Ace Step1.5 is a streamlined text-to-audio model designed for developers needing efficient, high-fidelity speech synthes...
ACE-Step
text-to-audio
63.4K
44
Ace Step1.5 XL DF11 is a specialized text-to-audio model optimized for the ComfyUI ecosystem. Unlike general-purpose aud...
mingyi456
text-to-audio
9.7K
0
FastSpeech 2 Conformer is a non-autoregressive text-to-speech (TTS) model designed for high-fidelity audio synthesis wit...
espnet
text-to-audio
9.6K
0
This model combines the FastSpeech 2 architecture with Conformer blocks and a HiFi-GAN vocoder to deliver high-fidelity,...
espnet
text-to-audio
8.8K
0
MiniMax Music3 is a specialized text-to-audio model designed for high-fidelity music generation. For developers, its pri...
MiniMaxAI
text-to-audio
8.6K
0
The acestep 5Hz lm 0.6B is a lightweight, text-to-audio model designed for low-latency synthesis. At 0.6B parameters, it...
ACE-Step
text-to-audio
5.4K
0
acestep v15 xl sft is a specialized text-to-audio model designed for high-fidelity sound synthesis. Unlike general-purpo...
ACE-Step
text-to-audio
4.5K
10
The acestep 5Hz lm 4B is a lightweight, open-source text-to-audio model designed for low-latency synthesis. With a 4-bil...
ACE-Step
text-to-audio
3.1K
2