Download
1 variant available
This checkpoint includes a config file, download and place it along side the checkpoint.
1150 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
(9)
Jun 15, 2026
ACE Audio

2.1K0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K
1.9K0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K

Version 4.0:
I've added two helpful notes:
The brown note is for you to copy the reference text for each voice you use.
I've noticed that if you don't use reference text with what the example voice says, it sometimes mispronounces the phrase you input.
For example, mine would start lisping, or pronounce the J's in an English way even when I told it I was speaking in Spanish. But when I added the reference text, this didn't happen.
The texts provided are from the Spanish voices I use from Elevenlabs. Use the ones for your languages.

In the blue note, I've included useful parameters for modulating the voice and adding emphasis to feelings or pronunciation.

Use my other workflow, the one for extracting subtitles from audio, to easily and quickly obtain the reference text.
https://civitai.red/models/2704860/workflow-for-generating-subtitles-for-a-video
https://civitai.red/articles/33166/extract-voices-from-elevenlabs-for-my-voice-clone-workflow
Versión 4.0:
He añadio 2 notas utiles:
La nota marrón para que copieis el texto de referencia de cada voz que pongais.
He visto que sino pones un texto de referencia con lo que dice la voz de ejemplo a veces pronuncia mal la frase que pongais.
A mí por ejemplo se me ponía a cecear, o pronunciaba las J al estilo inglés aunque le dijese que estaba hablando en español. Pero al ponerle el texto de referencia no ocurría esto.
Los textos puestos son los de las voces que uso yo de elevenlabs en español. Poned las de vuestros idiomas.
En la nota azul he puesto parametros utiles para modular la voz y añadirle enfasis en sentimientos o forma de pronunciar el texto.
Usad mi otro workflow, el de sacar subtitulos del audio para obtener el texto de referencia de una forma fácil y rápida.
https://civitai.red/models/2704860/workflow-for-generating-subtitles-for-a-video
https://civitai.red/articles/33166/extract-voices-from-elevenlabs-for-my-voice-clone-workflow
Version 3.0:
I've added a node to make loading .mp3 files easier, so they don't have to be in the input folder to load. Now they can be in any folder.
Versión 3.0:
He puesto un nodo para cargar más comodamente los .mp3 y no haya que ponerlos en la carpeta input para poder cargarlos. Ahora pueden estar en cualquier carpeta.
Version 2.0:
I've added a node to save it directly as an .mp3 file, eliminating the need to change the extension from the previous .flac file.
Versión 2.0:
He puesto un nodo para que lo guarde directamente en .mp3 y no haya que hacer el paso de cambiar la extensión de antes que era en .flac.
Hi!
This is a very simple workflow for cloning any voice. Simply put about 3 minutes of the voice you want, type a phrase, and click run.
It's best to do it phrase by phrase. I have an RTX 4090, and if I try to make a 10-minute audio file, my computer crashes, haha. In this tutorial, I'll show you how to create longer audio files.
https://civitai.red/articles/31334/how-to-clone-a-voice-and-create-large-audio-files-for-videos
P.S: Since Civitai takes a long time to review the images in the article explaining how to use it, here's the link to download the necessary model: download all the files and place them here: C:\Users\****\Documents\ComfyUI\models\fishaudioS2\s2-pro
https://huggingface.co/fishaudio/s2-pro/tree/main
¡Hola!
Este es un workflow muy simple para poder clonar cualquier voz. Simplemente pon el audio de unos 3 minutos de la voz que quieras, escribe una frase y dale a ejecutar.
Es mejor hacerlo frase a frase, yo tengo una rtx4090 y si intento hacer un audio de 10 minutos, el ordenador se me muere jeje. En este tutorial enseño cómo conseguir audios largos.
https://civitai.red/articles/31333/como-clonar-una-voz-y-crear-audios-de-gran-tamano-para-videos
P.D: Como civitai tarda mucho en revisar las imágenes del artículo que explica cómo usarlo, aquí os dejo el enlace para descargar el modelo necesario: baja todos los archivos y colocalos aquí C:\Users\****\Documents\ComfyUI\models\fishaudioS2\s2-pro
