This GGUF-quantized music generation model brings high-fidelity audio synthesis to local environments, eliminating the n...
mradermacher
audio-generation
393
0
The distilgpt2-abc-irish-music model is a specialized generative tool designed for developers working with symbolic musi...
ehcalabres
audio-generation
153
0
This music generation model by DancingIguana offers an open-source alternative for integrating high-fidelity audio synth...
DancingIguana
audio-generation
100
0
The Sinhala Audio to Text model provides a specialized speech-to-text pipeline for the Sinhala language, addressing a cr...
AqeelShafy7
audio-generation
41
0
LongCat AudioDiT 3.5B is a high-capacity text-to-speech model leveraging a Diffusion Transformer (DiT) architecture to a...
9r4n4y
audio-generation
26
0
LongCat AudioDiT 3.5B is a diffusion-transformer based text-to-speech model designed for high-fidelity audio synthesis. ...
orbitalhd
audio-generation
19
0
This model adapts the GPT-2 transformer architecture for symbolic music generation, treating musical notes and timing as...
hardikpatel
audio-generation
11
0
This music generation model provides a lightweight, MIT-licensed solution for developers needing programmatic audio synt...
metoonhathung
audio-generation
11
0
This audio generation model provides a flexible toolset for developers looking to integrate programmatic music synthesis...
nagayama0706
audio-generation
8
0
This Swahili ASR model provides a specialized pipeline for converting Swahili speech to text, filling a critical gap in ...
Peed911
audio-generation
8
0
The SAAD speech recognition model is a specialized audio-to-text tool designed specifically for the Hausa language. For ...
Baghdad99
audio-generation
8
0
This music generation model provides a flexible, MIT-licensed solution for developers looking to integrate programmatic ...
lopushanskyy
audio-generation
6
0
AudioSangraha is a specialized audio-to-text transcription model designed for developers needing reliable speech-to-text...
AqeelShafy7
audio-generation
6
0
The Sinhala Audio to Text CD model is a specialized speech-to-text tool designed for high-accuracy transcription of the ...
AqeelShafy7
audio-generation
6
0
LongCat AudioDiT 3.5B is a high-capacity text-to-speech model leveraging a Diffusion Transformer (DiT) architecture to a...
kkjipndy
audio-generation
6
0
The 'audio to texto pt br' model is a specialized speech-to-text engine optimized specifically for Brazilian Portuguese....
handsonbrain
audio-generation
5
0
Whisper is a robust speech-to-text model designed for high-accuracy transcription and translation across multiple langua...
AventIQ-AI
audio-generation
5
0