AI Tools Collection

One-stop directory of AI models, skills and MCP tools — search and find the AI that suits you.

AI Model · speech-enhancement

All (20)

sepformer wham16k enhancement

AI Model
The SepFormer WHAM!16k is a specialized speech enhancement model designed to isolate clean voice signals from
speechbrain speech-enhancement

sepformer dns4 16k enhancement

AI Model
SepFormer DNS4 16k is a specialized speech enhancement model designed to isolate clean voice signals from nois
speechbrain speech-enhancement

sepformer whamr enhancement

AI Model
SepFormer WHAMR is a specialized speech enhancement model designed to isolate clean speech from complex, noisy
speechbrain speech-enhancement

sepformer wham enhancement

AI Model
SepFormer WHAM is a specialized speech enhancement model designed to isolate target speech from complex, noisy
speechbrain speech-enhancement

speech enhancement sgmse

AI Model
SGMSE is a generative speech enhancement model designed to isolate clean voice signals from noisy environments
sp-uhh speech-enhancement

mpse smoe speech enhancement

AI Model
MPSE SMOE is a specialized speech enhancement model designed to isolate clean vocals from noisy environments.
AdityaRaikar speech-enhancement

unet speech enhancement

AI Model
The UNet Speech Enhancement model is a deep learning architecture designed to isolate target speech from compl
SSB258369 speech-enhancement

speech enhancement mask unet

AI Model
The Speech Enhancement Mask UNet is a deep learning architecture designed to isolate clean speech from noisy b
huseinzol05 speech-enhancement

AI Model · text-to-video

All (19)

Sulphur 2 base

AI Model
Sulphur 2 base is a foundational text-to-video model designed for developers building generative media pipelin
SulphurAI text-to-video

Wan2.2 TI2V 5B Diffusers

AI Model
Wan2.2 TI2V 5B is a specialized image-to-video diffusion model designed for high-fidelity temporal animation.
Wan-AI text-to-video

Wan2.1 T2V 1.3B Diffusers

AI Model
Wan2.1 T2V 1.3B is a lightweight text-to-video diffusion model optimized for efficiency and accessibility. Unl
Wan-AI text-to-video

MiniMax H3 Turbo Lora ComfyUI

AI Model
The MiniMax H3 Turbo LoRA for ComfyUI brings professional-grade text-to-video generation directly into a modul
drbaph text-to-video

Sulphur 2 base GGUF

AI Model
Sulphur 2 Base GGUF provides developers with a quantized, hardware-efficient implementation of the Sulphur 2 a
Abiray text-to-video

Wan2.2 T2V A14B Diffusers

AI Model
Wan2.2 T2V A14B is a high-capacity text-to-video diffusion model designed for developers seeking a balance bet
Wan-AI text-to-video

Wan2.2 T2V A14B GGUF

AI Model
Wan2.2 T2V A14B GGUF provides a quantized implementation of the 14-billion parameter text-to-video model, spec
QuantStack text-to-video

FastWan2.2 TI2V 5B FullAttn Diffusers

AI Model
FastWan2.2 TI2V 5B is a specialized Text-to-Image-to-Video (TI2V) diffusion model designed for high-fidelity c
FastVideo text-to-video

AI Model · text-generation

All (18)

Llama 3.1 405B

AI Model
Open source large language model
Meta text-generation

Mixtral 8x7B

AI Model
Sparse mixture of experts model
Mistral AI text-generation

Qwen3 0.6B

AI Model
Qwen3 0.6B is a highly compact language model designed for efficiency and low-latency deployment. At under one
Qwen text-generation

Qwen3 8B

AI Model
Qwen3 8B is a compact yet powerful language model designed for high-efficiency deployment without sacrificing
Qwen text-generation

tiny Qwen2ForCausalLM 2.5

AI Model
The tiny Qwen2ForCausalLM 2.5 is a compact, efficient language model designed for low-latency applications and
trl-internal-testing text-generation

gpt2

AI Model
GPT-2 is a foundational transformer-based language model that marked a shift toward zero-shot learning in NLP.
openai-community text-generation

Qwen2.5 7B Instruct

AI Model
Qwen2.5 7B Instruct is a dense decoder-only model designed for high-efficiency deployment without sacrificing
Qwen text-generation

Qwen3.6 35B A3B NVFP4

AI Model
The Qwen3.6 35B A3B NVFP4 is a specialized iteration of the Qwen series, optimized specifically for NVIDIA har
nvidia text-generation

AI Model · sentence-similarity

All (18)

all-MiniLM-L6-v2

AI Model
Fast sentence embeddings
Sentence Transformers sentence-similarity

paraphrase multilingual MiniLM L12 v2

AI Model
The paraphrase-multilingual-MiniLM-L12-v2 is a lightweight, high-performance transformer model optimized for g
sentence-transformers sentence-similarity

bge m3

AI Model
BGE-M3 is a versatile embedding model designed for high-performance retrieval across diverse linguistic landsc
BAAI sentence-similarity

all mpnet base v2

AI Model
all-mpnet-base-v2 is a high-performance sentence-transformer model optimized for mapping text to a dense vecto
sentence-transformers sentence-similarity

nomic embed text v1.5

AI Model
nomic-embed-text-v1.5 is a high-performance text embedding model designed for scalable retrieval and semantic
nomic-ai sentence-similarity

multilingual e5 small

AI Model
Multilingual E5 Small is a lightweight, high-efficiency embedding model designed for cross-lingual sentence si
intfloat sentence-similarity

paraphrase multilingual mpnet base v2

AI Model
The paraphrase-multilingual-mpnet-base-v2 is a robust sentence-embedding model designed for cross-lingual sema
sentence-transformers sentence-similarity

multilingual e5 base

AI Model
Multilingual-e5-base is a high-performance text embedding model designed for cross-lingual semantic search and
intfloat sentence-similarity

AI Model · question-answering

All (18)

electra large discriminator squad2 512

AI Model
The ELECTRA-Large Discriminator, fine-tuned on SQuAD 2.0, is a specialized encoder model designed for high-pre
ahotrod question-answering

roberta base squad2

AI Model
roberta-base-squad2 is a robust extractive question-answering model based on the RoBERTa architecture, specifi
deepset question-answering

koelectra base v3 finetuned korquad

AI Model
The koelectra-base-v3-korquad model is a specialized encoder-based transformer optimized for Korean extractive
monologg question-answering

distilbert base cased distilled squad

AI Model
DistilBERT-base-cased-distilled-SQuAD is a lightweight, transformer-based model optimized for extractive quest
distilbert question-answering

bert large uncased whole word masking squad2

AI Model
The BERT-Large-Uncased-WWM-SQuAD2 model is a specialized encoder designed for high-precision extractive questi
deepset question-answering

tinyroberta squad2

AI Model
TinyRoBERTa-SQuAD2 is a compact, distilled version of the RoBERTa architecture specifically fine-tuned for ext
deepset question-answering

koelectra small v2 distilled korquad 384

AI Model
KoELECTRA-Small v2 is a lightweight, distilled transformer model optimized specifically for Korean language un
monologg question-answering

xlm roberta base squad2

AI Model
xlm-roberta-base-squad2 is a cross-lingual transformer model fine-tuned specifically for extractive question a
deepset question-answering

AI Model · table-question-answering

All (18)

tapex base finetuned wikisql

AI Model
TaPEX (Table Pre-training Experiment) is a specialized encoder-decoder model designed specifically for table-b
microsoft table-question-answering

tapas large finetuned sqa

AI Model
The TAPAS Large model, fine-tuned on the SQuAD-style table question answering (SQuA) dataset, is designed spec
google table-question-answering

tapas base finetuned wtq

AI Model
The tapas-base-finetuned-wtq model is a specialized transformer architecture designed specifically for Table Q
google table-question-answering

tapas base finetuned sqa

AI Model
The TAPAS Base model, fine-tuned on the SQuAD-style table questioning dataset, is designed specifically for ta
google table-question-answering

tapas large finetuned wtq

AI Model
The TAPAS Large model, fine-tuned on the WikiTableQuestions (WtQ) dataset, is a specialized transformer design
google table-question-answering

tapas tiny finetuned sqa

AI Model
The Tapas Tiny SQA model is a specialized transformer designed for Table Question Answering (TableQA). Unlike
google table-question-answering

tiny tapas random wtq

AI Model
Tiny TAPAS (Random WTQ) is a specialized, lightweight model designed for table-based question answering. Unlik
lysandre table-question-answering

tiny tapas random sqa

AI Model
Tiny Tapas Random SQA is a lightweight, specialized model designed for table-based question answering. Unlike
lysandre table-question-answering

AI Model · image-to-text

All (18)

blip image captioning base

AI Model
BLIP (Bootstrapping Language-Image Pre-training) is a versatile vision-language model designed to bridge the g
Salesforce image-to-text

PP OCRv5 server det

AI Model
PP OCRv5 server det is a high-performance text detection model designed for industrial-scale OCR pipelines. Un
PaddlePaddle image-to-text

manga ocr base

AI Model
Manga OCR Base is a specialized image-to-text model engineered specifically for the complexities of Japanese m
kha-white image-to-text

blip image captioning large

AI Model
BLIP (Bootstrapping Language-Image Pre-training) Large is a versatile vision-language model designed for high-
Salesforce image-to-text

en PP OCRv5 mobile rec

AI Model
The PP-OCRv5 mobile recognition model is a lightweight, high-efficiency text recognition engine optimized for
PaddlePaddle image-to-text

UVDoc

AI Model
UVDoc is a specialized image-to-text model developed by PaddlePaddle, designed to handle complex document pars
PaddlePaddle image-to-text

trocr small handwritten

AI Model
TrOCR-small is a lightweight transformer-based model designed specifically for optical character recognition o
microsoft image-to-text

pix2text mfr

AI Model
pix2text mfr is a specialized image-to-text model designed for high-accuracy mathematical formula recognition.
breezedeus image-to-text

AI Model · image-text-to-text

All (18)

Qwen3.5 9B

AI Model
Qwen3.5 9B is a versatile multimodal model designed to bridge the gap between lightweight efficiency and high-
Qwen image-text-to-text

Qwen3.6 35B A3B FP8

AI Model
Qwen3.6 35B A3B FP8 is a multimodal model designed for efficient image-text processing. By utilizing FP8 quant
Qwen image-text-to-text

gemma 4 26B A4B it

AI Model
Gemma 4 26B A4B is a multimodal model from Google, designed for developers needing high-performance image-to-t
google image-text-to-text

gemma 4 31B it

AI Model
Gemma 4 31B IT is a mid-sized, instruction-tuned multimodal model designed for developers who need a balance b
google image-text-to-text

Qwen2.5 VL 7B Instruct

AI Model
Qwen2.5 VL 7B Instruct is a versatile vision-language model designed for high-precision image and video unders
Qwen image-text-to-text

Qwen3.6 27B FP8

AI Model
Qwen3.6 27B FP8 is a high-efficiency multimodal model designed for developers who need a balance between reaso
Qwen image-text-to-text

Qwen3.5 4B

AI Model
Qwen3.5 4B is a compact multimodal model designed for efficient image-and-text processing. Unlike larger LLMs
Qwen image-text-to-text

Qwen2.5 VL 3B Instruct

AI Model
Qwen2.5 VL 3B Instruct is a lightweight yet powerful vision-language model designed for efficient multimodal p
Qwen image-text-to-text

AI Model · object-detection

All (18)

table transformer structure recognition

AI Model
Table Transformer (TATR) is a specialized object detection model designed to automate the structural analysis
microsoft object-detection

table transformer detection

AI Model
Table Transformer (TATR) is a specialized object detection model designed to solve the complex problem of tabl
microsoft object-detection

yolos small

AI Model
YOLOS-Small is a streamlined object detection model that replaces traditional convolutional backbones with a V
hustvl object-detection

PP DocLayoutV3 safetensors

AI Model
PP DocLayoutV3 is a specialized object detection model optimized for document layout analysis. Unlike general-
PaddlePaddle object-detection

rtdetr r101vd coco o365

AI Model
The RT-DETR R101VD is a real-time end-to-end object detector that eliminates the need for non-maximum suppress
PekingU object-detection

rtdetr v2 r18vd

AI Model
RT-DETR v2 (R18VD) is a high-performance real-time object detector that bridges the gap between the speed of Y
PekingU object-detection

detr resnet 50

AI Model
DETR (Detection Transformer) with a ResNet-50 backbone represents a fundamental shift in object detection by r
facebook object-detection

table transformer structure recognition v1.1 all

AI Model
Table Transformer (TATR) v1.1 is a specialized object-detection model designed to automate the extraction of s
microsoft object-detection

AI Model · summarization

All (17)

BART Large CNN

AI Model
Summarization and text generation
Facebook summarization

distilbart cnn 12 6

AI Model
DistilBART-cnn-12-6 is a streamlined version of the BART architecture, specifically optimized for abstractive
sshleifer summarization

pegasus xsum

AI Model
Pegasus XSum is a specialized transformer model engineered specifically for extreme summarization. Unlike gene
google summarization

financial summarization pegasus

AI Model
Financial Summarization Pegasus is a domain-specific encoder-decoder model fine-tuned for the nuances of finan
human-centered-summarization summarization

bart large cnn samsum

AI Model
The bart-large-cnn-samsum model is a specialized encoder-decoder transformer fine-tuned specifically for conve
philschmid summarization

distilbart xsum 12 6

AI Model
DistilBART-xsum-12-6 is a compressed version of the BART architecture specifically fine-tuned for extreme summ
sshleifer summarization

text summarization

AI Model
This model, hosted by Falconsai, is a specialized tool designed for efficient text summarization tasks. Unlike
Falconsai summarization

MEETING SUMMARY

AI Model
The MEETING SUMMARY model is a specialized tool designed for distilling lengthy meeting transcripts into conci
knkarthick summarization

AI Model · text-classification

All (17)

bge reranker v2 m3

AI Model
The BGE Reranker v2 M3 is a cross-encoder model designed to refine the output of initial retrieval stages in R
BAAI text-classification

finbert

AI Model
FinBERT is a domain-specific adaptation of the BERT architecture, pre-trained on a massive corpus of financial
ProsusAI text-classification

Prompt Guard 86M

AI Model
Prompt Guard 86M is a lightweight, specialized classifier designed to secure LLM pipelines by detecting prompt
meta-llama text-classification

bge reranker base

AI Model
The bge-reranker-base is a cross-encoder model designed to refine the results of initial vector searches. Unli
BAAI text-classification

distilbert base uncased finetuned sst 2 english

AI Model
DistilBERT base uncased finetuned SST-2 is a lightweight, distilled version of BERT optimized for binary senti
distilbert text-classification

twitter roberta base sentiment latest

AI Model
The twitter-roberta-base-sentiment-latest model is a specialized text classifier fine-tuned on a massive corpu
cardiffnlp text-classification

koelectra small v3 nsmc

AI Model
KoELECTRA Small v3 (NSMC) is a lightweight, discriminative transformer model optimized for Korean sentiment an
daekeun-ml text-classification

tiny Qwen2ForSequenceClassification 2.5

AI Model
The tiny Qwen2ForSequenceClassification 2.5 is a lightweight, encoder-style adaptation of the Qwen2 architectu
trl-internal-testing text-classification

AI Model · token-classification

All (17)

indonesian roberta base posp tagger

AI Model
The indonesian-roberta-base-posp-tagger is a specialized token-classification model designed for Part-of-Speec
w11wo token-classification

stanford deidentifier base

AI Model
The Stanford Deidentifier Base is a specialized token-classification model designed to automate the removal of
StanfordAIMI token-classification

bert base NER

AI Model
The bert-base-NER model is a specialized token-classification transformer fine-tuned for Named Entity Recognit
dslim token-classification

bert large cased finetuned conll03 english

AI Model
This model is a BERT-Large architecture specifically fine-tuned on the CoNLL-03 dataset for English Named Enti
dbmdz token-classification

punctuate all

AI Model
Punctuate All is a specialized token-classification model designed to automate punctuation restoration in unst
kredor token-classification

ner english fast

AI Model
The 'ner-english-fast' model is a lightweight token-classification tool designed for high-throughput Named Ent
flair token-classification

sat 3l sm

AI Model
The sat 3l sm model is a specialized token-classification engine designed for precise text segmentation. Unlik
segment-any-text token-classification

bert portuguese ner

AI Model
The BERT Portuguese NER model is a specialized token-classification engine fine-tuned specifically for Named E
lfcc token-classification

AI Model · fill-mask

All (17)

bert base uncased

AI Model
BERT Base Uncased is a foundational transformer-based encoder designed for bidirectional representation learni
google-bert fill-mask

xlm roberta base

AI Model
XLM-RoBERTa Base is a multilingual transformer model designed for high-performance Natural Language Understand
FacebookAI fill-mask

roberta base

AI Model
RoBERTa-base is an optimized evolution of the BERT architecture, designed for developers who need high-perform
FacebookAI fill-mask

roberta large

AI Model
RoBERTa-large is a robustly optimized version of BERT, designed for developers who need high-performance natur
FacebookAI fill-mask

distilbert base uncased

AI Model
DistilBERT base uncased is a streamlined, lightweight version of the BERT architecture, designed for developer
distilbert fill-mask

xlm roberta large

AI Model
XLM-RoBERTa Large is a powerful multilingual transformer model designed for cross-lingual representation learn
FacebookAI fill-mask

ModernBERT base

AI Model
ModernBERT is a complete architectural refresh of the BERT encoder, designed to bridge the gap between legacy
answerdotai fill-mask

mdeberta v3 base

AI Model
mdeberta-v3-base is a refined encoder-only transformer that optimizes the DeBERTa architecture for superior pe
microsoft fill-mask

AI Model · feature-extraction

All (17)

bge small en v1.5

AI Model
The bge-small-en-v1.5 is a lightweight embedding model designed for high-efficiency vectorization of English t
BAAI feature-extraction

bge large en v1.5

AI Model
The bge-large-en-v1.5 is a high-performance embedding model designed for transforming text into dense vectors
BAAI feature-extraction

bge base en v1.5

AI Model
The bge-base-en-v1.5 is a high-performance embedding model designed for transforming text into dense vectors f
BAAI feature-extraction

Qwen3 Embedding 0.6B

AI Model
Qwen3 Embedding 0.6B is a compact, high-efficiency feature extraction model designed for developers building R
Qwen feature-extraction

multilingual e5 large

AI Model
Multilingual E5 Large is a high-performance text embedding model designed for cross-lingual semantic search an
intfloat feature-extraction

bge small zh v1.5

AI Model
The bge-small-zh-v1.5 is a lightweight embedding model optimized for Chinese language tasks. Designed for effi
BAAI feature-extraction

granite embedding small english r2

AI Model
The granite-embedding-small-english-r2 is a lightweight, English-centric feature extraction model designed for
ibm-granite feature-extraction

mxbai embed large v1

AI Model
mxbai-embed-large-v1 is a high-performance embedding model designed for developers building RAG pipelines and
mixedbread-ai feature-extraction

AI Model · voice-activity-detection

All (17)

segmentation 3.0

AI Model
Segmentation 3.0 by pyannote is a specialized voice activity detection (VAD) model designed to distinguish hum
pyannote voice-activity-detection

segmentation

AI Model
The pyannote segmentation model is a specialized tool for Voice Activity Detection (VAD) and speaker change de
pyannote voice-activity-detection

Namo Turn Detector v1 Korean

AI Model
Namo Turn Detector v1 is a specialized Voice Activity Detection (VAD) model optimized specifically for the Kor
videosdk-live voice-activity-detection

speaker diarization precision 2

AI Model
Speaker Diarization Precision 2, powered by pyannote, is a specialized voice-activity-detection tool designed
pyannote voice-activity-detection

Silero VAD v5 MLX

AI Model
Silero VAD v5 MLX is a specialized voice activity detection model optimized for Apple Silicon via the MLX fram
aufklarer voice-activity-detection

silero vad coreml

AI Model
Silero VAD CoreML is a specialized voice activity detection model optimized for Apple's hardware acceleration.
FluidInference voice-activity-detection

brouhaha

AI Model
Brouhaha is a specialized voice activity detection (VAD) model developed by pyannote, designed to distinguish
pyannote voice-activity-detection

Pyannote Segmentation MLX

AI Model
Pyannote Segmentation MLX is a specialized voice activity detection (VAD) and speaker segmentation model optim
aufklarer voice-activity-detection

AI Model · image-classification

All (17)

mobilenetv3 small 100.lamb in1k

AI Model
MobileNetV3-Small is a lightweight convolutional neural network optimized for low-latency inference on mobile
timm image-classification

vit base patch16 224

AI Model
The ViT-B/16 is a foundational Vision Transformer that replaces traditional convolutional layers with a pure t
google image-classification

nsfw image detection

AI Model
The nsfw image detection model by Falconsai is a specialized image classification tool designed to automate co
Falconsai image-classification

tf efficientnetv2 s.in21k ft in1k

AI Model
The EfficientNetV2-S (in21k-ft-in1k) is a high-performance convolutional neural network optimized for image cl
timm image-classification

fairface age image detection

AI Model
The FairFace age detection model is a specialized image classification tool designed to mitigate demographic b
dima806 image-classification

resnet50.a1 in1k

AI Model
The resnet50.a1_in1k is a refined implementation of the classic ResNet-50 architecture, optimized for image cl
timm image-classification

resnet18.a1 in1k

AI Model
The resnet18.a1_in1k is a lightweight convolutional neural network based on the ResNet-18 architecture, pre-tr
timm image-classification

resnet 50

AI Model
ResNet-50 is a foundational deep residual network that solved the vanishing gradient problem in deep architect
microsoft image-classification

AI Model · audio-generation

All (17)

music generation model GGUF

AI Model
This GGUF-quantized music generation model brings high-fidelity audio synthesis to local environments, elimina
mradermacher audio-generation

distilgpt2 abc irish music generation

AI Model
The distilgpt2-abc-irish-music model is a specialized generative tool designed for developers working with sym
ehcalabres audio-generation

music generation

AI Model
This music generation model by DancingIguana offers an open-source alternative for integrating high-fidelity a
DancingIguana audio-generation

GPT2 Music Generation Trained

AI Model
This model adapts the GPT-2 transformer architecture for symbolic music generation, treating musical notes and
hardikpatel audio-generation

music generation

AI Model
This music generation model provides a lightweight, MIT-licensed solution for developers needing programmatic
metoonhathung audio-generation

music generation model

AI Model
This audio generation model provides a flexible toolset for developers looking to integrate programmatic music
nagayama0706 audio-generation

music generation

AI Model
This music generation model provides a flexible, MIT-licensed solution for developers looking to integrate pro
lopushanskyy audio-generation

Sinhala Audio to Text

AI Model
The Sinhala Audio to Text model provides a specialized speech-to-text pipeline for the Sinhala language, addre
AqeelShafy7 audio-generation

AI Model · translation

All (16)

t5 small

AI Model
T5-Small is a lightweight, encoder-decoder transformer model designed for text-to-text tasks. Unlike decoder-o
google-t5 translation

t5 base

AI Model
T5 (Text-to-Text Transfer Transformer) Base is a versatile encoder-decoder model that treats every NLP task as
google-t5 translation

opus mt nl en

AI Model
The opus-mt-nl-en model is a specialized neural machine translation (NMT) tool designed specifically for Dutch
Helsinki-NLP translation

opus mt fr en

AI Model
The Opus MT FR-EN model is a specialized translation engine developed by Helsinki-NLP, designed specifically f
Helsinki-NLP translation

vntl llama3 8b v2 gguf

AI Model
The vntl-llama3-8b-v2 is a specialized GGUF quantization of Meta's Llama 3 8B, fine-tuned specifically for hig
lmg-anon translation

opus mt en ru

AI Model
Opus-MT EN-RU is a specialized neural machine translation model developed by Helsinki-NLP, designed specifical
Helsinki-NLP translation

opus mt en de

AI Model
The opus-mt-en-de model is a specialized neural machine translation (NMT) tool developed by Helsinki-NLP, desi
Helsinki-NLP translation

opus mt de en

AI Model
The opus-mt-de-en model is a specialized neural machine translation (NMT) engine developed by Helsinki-NLP, de
Helsinki-NLP translation

AI Model · visual-question-answering

All (16)

blip vqa base

AI Model
BLIP-VQA Base is a vision-language model designed for visual question answering, bridging the gap between imag
Salesforce visual-question-answering

vilt b32 finetuned vqa

AI Model
The ViLT b32 finetuned VQA model is a streamlined vision-and-language transformer designed for Visual Question
dandelin visual-question-answering

MiniCPM V 2

AI Model
MiniCPM-V 2 is a compact yet powerful vision-language model designed to bridge the gap between edge-device eff
openbmb visual-question-answering

deplot

AI Model
DePlot is a specialized visual-question-answering (VQA) model designed to bridge the gap between plotted data
google visual-question-answering

blip vqa capfilt large

AI Model
The BLIP VQA Capfilt Large is a specialized vision-language model designed for precise visual question answeri
Salesforce visual-question-answering

Qwen3.5 2B MedVL

AI Model
Qwen3.5 2B MedVL is a compact vision-language model specifically tuned for the medical domain. Designed for ef
OpenMed visual-question-answering

VideoScore2

AI Model
VideoScore2 is a specialized visual-question-answering (VQA) model designed to quantify video quality and sema
TIGER-Lab visual-question-answering

llava med v1.5 mistral 7b hf

AI Model
LLaVA-Med v1.5 Mistral 7B HF is a domain-specific multimodal model designed for biomedical visual question ans
chaoyinshe visual-question-answering

AI Model · automatic-speech-recognition

All (16)

whisperkit coreml

AI Model
WhisperKit CoreML brings OpenAI's Whisper speech-to-text capabilities directly to Apple silicon, optimizing in
argmaxinc automatic-speech-recognition

speaker diarization 3.1

AI Model
Speaker Diarization 3.1, powered by pyannote, is a specialized framework designed to solve the 'who spoke when
pyannote automatic-speech-recognition

whisper large v3 turbo

AI Model
Whisper large-v3-turbo is a streamlined version of OpenAI's state-of-the-art speech recognition model, enginee
openai automatic-speech-recognition

wav2vec2 large xlsr 53 japanese

AI Model
The wav2vec2-large-xlsr-53-japanese model is a robust automatic speech recognition (ASR) tool fine-tuned for J
jonatasgrosman automatic-speech-recognition

speaker diarization community 1

AI Model
Speaker Diarization Community 1, powered by pyannote, is a specialized tool for the 'who spoke when' problem i
pyannote automatic-speech-recognition

wav2vec2 large xlsr 53 portuguese

AI Model
The wav2vec2-large-xlsr-53-portuguese model is a robust Automatic Speech Recognition (ASR) tool fine-tuned spe
jonatasgrosman automatic-speech-recognition

voice activity detection

AI Model
Pyannote's Voice Activity Detection (VAD) is a specialized tool designed to distinguish human speech from sile
pyannote automatic-speech-recognition

Qwen3 ASR 1.7B

AI Model
Qwen3 ASR 1.7B is a compact, efficient automatic speech recognition model designed for low-latency transcripti
Qwen automatic-speech-recognition

AI Model · text-to-image

All (15)

Stable Diffusion XL

AI Model
High-quality image generation
Stability AI text-to-image

stable diffusion xl base 1.0

AI Model
Stable Diffusion XL (SDXL) 1.0 represents a significant architectural leap over previous versions, moving to a
stabilityai text-to-image

stable diffusion v1 5

AI Model
Stable Diffusion v1.5 remains a foundational pillar for open-source generative AI, offering a versatile text-t
stable-diffusion-v1-5 text-to-image

dreamshaper 7

AI Model
DreamShaper 7 is a refined Stable Diffusion checkpoint designed to bridge the gap between photorealism and dig
Lykon text-to-image

Z Image Turbo

AI Model
Z Image Turbo is a high-performance text-to-image model designed for developers who need a balance between gen
Tongyi-MAI text-to-image

sd turbo

AI Model
SD Turbo is a distilled version of Stable Diffusion designed specifically for real-time image synthesis. Unlik
stabilityai text-to-image

Realistic Vision V5.1 noVAE

AI Model
Realistic Vision V5.1 (noVAE) is a fine-tuned Stable Diffusion checkpoint optimized specifically for photoreal
SG161222 text-to-image

stable diffusion v1 4

AI Model
Stable Diffusion v1.4 is a latent diffusion model designed for high-efficiency text-to-image synthesis. Unlike
CompVis text-to-image

AI Model · image-to-image

All (14)

Qwen Image Edit 2509

AI Model
Qwen Image Edit 2509 is a specialized image-to-image model designed for precise visual manipulation. Unlike ge
Qwen image-to-image

FLUX.2 klein 4B

AI Model
FLUX.2 klein 4B is a streamlined image-to-image model designed for developers who need a balance between gener
black-forest-labs image-to-image

Qwen Image Edit 2511 Lightning

AI Model
Qwen Image Edit 2511 Lightning is a specialized image-to-image model designed for high-speed visual manipulati
lightx2v image-to-image

Qwen Edit 2509 Multiple angles

AI Model
Qwen Edit 2509 is a specialized image-to-image model designed for precise spatial and perspective manipulation
dx8152 image-to-image

Qwen Image Edit 2511 GGUF

AI Model
Qwen Image Edit 2511 GGUF is a specialized image-to-image model optimized for local deployment via the GGUF fo
unsloth image-to-image

Qwen Image Edit 2511

AI Model
Qwen Image Edit 2511 is a specialized image-to-image model designed for precise visual modifications. Unlike g
Qwen image-to-image

FLUX.2 small decoder

AI Model
FLUX.2 small decoder is a streamlined image-to-image component designed for developers who need efficient late
black-forest-labs image-to-image

FLUX.2 klein base 4B

AI Model
FLUX.2 klein base 4B is a compact, efficient image-to-image model designed for high-fidelity visual transforma
black-forest-labs image-to-image

AI Model · audio-to-audio

All (14)

bigvgan v2 22khz 80band 256x

AI Model
BigVGAN v2 is a high-fidelity neural vocoder designed for high-resolution audio synthesis. Operating at 22kHz
nvidia audio-to-audio

bigvgan v2 44khz 128band 512x

AI Model
BigVGAN v2 is a high-fidelity neural vocoder designed for professional-grade speech and audio synthesis. Opera
nvidia audio-to-audio

Qwen3 TTS Tokenizer 12Hz

AI Model
The Qwen3 TTS Tokenizer 12Hz is a specialized audio-to-audio component designed to bridge the gap between raw
Qwen audio-to-audio

neucodec

AI Model
Neucodec is an open-source audio-to-audio model designed for high-fidelity neural audio compression and recons
neuphonic audio-to-audio

TIGER DnR

AI Model
TIGER DnR is a specialized audio-to-audio model designed for high-fidelity denoising and restoration. Unlike g
JusperLee audio-to-audio

bigvgan v2 24khz 100band 256x

AI Model
BigVGAN v2 is a high-fidelity neural vocoder designed for high-resolution audio synthesis. Unlike traditional
nvidia audio-to-audio

MP SENet DNS

AI Model
MP SENet DNS is a specialized audio-to-audio model designed for high-fidelity deep noise suppression. Unlike g
JacobLinCool audio-to-audio

distill neucodec

AI Model
Distill NeuCodec is an efficient audio-to-audio model designed for high-fidelity signal processing and compres
neuphonic audio-to-audio

AI Model · zero-shot-classification

All (13)

bart large mnli

AI Model
BART-large-MNLI is a specialized encoder-decoder model fine-tuned on the Multi-Genre Natural Language Inferenc
facebook zero-shot-classification

distilbert base uncased mnli

AI Model
DistilBERT base uncased MNLI is a streamlined transformer model optimized for zero-shot text classification. B
typeform zero-shot-classification

mDeBERTa v3 base mnli xnli

AI Model
mDeBERTa v3 base mnli xnli is a multilingual encoder model optimized for zero-shot text classification. Built
MoritzLaurer zero-shot-classification

DeBERTa v3 base mnli fever anli

AI Model
DeBERTa v3 base mnli fever anli is a specialized NLI (Natural Language Inference) model optimized for zero-sho
MoritzLaurer zero-shot-classification

DeBERTa v3 large mnli fever anli ling wanli

AI Model
This model is a specialized DeBERTa v3 large variant fine-tuned on a comprehensive suite of Natural Language I
MoritzLaurer zero-shot-classification

deberta v3 base zeroshot v2.0

AI Model
DeBERTa-v3-base-zeroshot-v2.0 is a specialized encoder model optimized for zero-shot text classification. Unli
MoritzLaurer zero-shot-classification

deberta v3 large zeroshot v2.0

AI Model
DeBERTa v3 Large Zero-Shot v2.0 is a specialized encoder model optimized for Natural Language Inference (NLI)
MoritzLaurer zero-shot-classification

bge m3 zeroshot v2.0

AI Model
The bge-m3-zeroshot-v2.0 is a specialized classification model designed for high-efficiency zero-shot tasks, e
MoritzLaurer zero-shot-classification

AI Model · text-to-speech

All (13)

Kokoro 82M

AI Model
Kokoro 82M is a highly efficient, small-footprint text-to-speech model designed for developers who need high-q
hexgrad text-to-speech

Qwen3 TTS 12Hz 1.7B CustomVoice

AI Model
Qwen3 TTS 12Hz 1.7B CustomVoice is a lightweight, high-efficiency text-to-speech model designed for low-latenc
Qwen text-to-speech

chatterbox

AI Model
Chatterbox is a lightweight, MIT-licensed text-to-speech (TTS) engine developed by ResembleAI, designed for de
ResembleAI text-to-speech

Kokoro 82M v1.0 ONNX

AI Model
Kokoro 82M v1.0 is a lightweight, high-efficiency text-to-speech model optimized for local deployment via the
onnx-community text-to-speech

Qwen3 TTS 12Hz 0.6B CustomVoice

AI Model
Qwen3 TTS 12Hz 0.6B CustomVoice is a lightweight, high-efficiency text-to-speech model designed for low-latenc
Qwen text-to-speech

OmniVoice

AI Model
OmniVoice is an open-source text-to-speech (TTS) engine designed for developers needing high-fidelity voice sy
k2-fsa text-to-speech

VibeVoice Realtime 0.5B

AI Model
VibeVoice Realtime 0.5B is a lightweight, low-latency text-to-speech model designed for real-time applications
microsoft text-to-speech

Qwen3 TTS GGUF

AI Model
Qwen3 TTS GGUF brings high-fidelity text-to-speech capabilities to local environments via the GGUF format, sig
Serveurperso text-to-speech

AI Model · depth-estimation

All (13)

Depth Anything V2 Small hf

AI Model
Depth Anything V2 Small is a lightweight monocular depth estimation model designed for real-time spatial analy
depth-anything depth-estimation

DA3METRIC LARGE

AI Model
DA3METRIC LARGE is a specialized depth-estimation model designed to recover high-fidelity metric depth from si
depth-anything depth-estimation

dpt hybrid midas

AI Model
DPT Hybrid MIDAS is a depth-estimation model designed to bridge the gap between high-resolution spatial detail
Intel depth-estimation

depth anything large hf

AI Model
Depth Anything Large is a powerful monocular depth estimation model designed to extract high-resolution spatia
LiheYoung depth-estimation

Distill Any Depth Large hf

AI Model
Distill Any Depth Large is a specialized depth-estimation model designed for high-fidelity spatial mapping fro
xingyang1 depth-estimation

DA3MONO LARGE

AI Model
DA3MONO LARGE is a specialized depth-estimation model designed to extract high-fidelity monocular depth maps f
depth-anything depth-estimation

zoedepth nyu kitti

AI Model
ZoeDepth is a zero-shot metric depth estimation model designed to bridge the gap between relative and absolute
Intel depth-estimation

DA3 LARGE 1.1

AI Model
DA3 Large 1.1 is a specialized depth-estimation model designed to extract high-fidelity spatial information fr
depth-anything depth-estimation

AI Model · video-classification

All (13)

vjepa2 vitl fpc64 256

AI Model
V-JEPA (ViT-L/fpc64/256) is a self-supervised video representation model designed for efficient spatial-tempor
facebook video-classification

vjepa2 vitg fpc64 256

AI Model
V-JEPA (ViT-G FPC64 256) is a sophisticated video representation model designed for self-supervised learning.
facebook video-classification

xclip base patch32

AI Model
xclip base patch32 is a specialized vision-language model optimized for video classification tasks. Unlike sta
microsoft video-classification

vivit b 16x2 kinetics400

AI Model
ViViT-B/16x2 is a Video Vision Transformer designed specifically for high-accuracy video classification. Unlik
google video-classification

xclip base patch32 16 frames

AI Model
The xclip-base-patch32 model is a specialized video classification tool designed to extend CLIP's visual-langu
microsoft video-classification

vivit b 16x2

AI Model
ViViT-B/16x2 is a Video Vision Transformer designed for efficient video classification by treating video clips
google video-classification

ms eff gcvit deepfake b0 kodf

AI Model
The ms-eff-gcvit-deepfake-b0 is a specialized video classification model designed to detect synthetic media an
KoreaPeter video-classification

ms eff gcvit deepfake b5 kodf

AI Model
The ms-eff-gcvit-deepfake-b5-kodf is a specialized video classification model designed for high-accuracy deepf
KoreaPeter video-classification

AI Model · document-question-answering

All (12)

layoutlm document qa

AI Model
LayoutLM is a specialized transformer model designed for Document AI, bridging the gap between raw text and vi
impira document-question-answering

donut base finetuned docvqa

AI Model
The Donut-base model finetuned for DocVQA represents a shift toward OCR-free document understanding. Unlike tr
naver-clova-ix document-question-answering

tiny doc qa vision encoder decoder

AI Model
The tiny-doc-qa-vision-encoder-decoder is a lightweight, specialized model designed for efficient document que
fxmarty document-question-answering

donut base finetuned docvqa

AI Model
The Donut-based DocVQA model is an OCR-free transformer designed for document visual question answering. Unlik
Xenova document-question-answering

layoutlmv2 base uncased finetuned docvqa

AI Model
LayoutLMv2-base-uncased-finetuned-docvqa is a multimodal transformer designed for Document Visual Question Ans
tiennvcs document-question-answering

Llama 3.1 PersianQA

AI Model
Llama 3.1 PersianQA is a specialized adaptation of the Llama 3.1 architecture, fine-tuned specifically for doc
zpm document-question-answering

layoutlmv3 docvqa t11c5000

AI Model
LayoutLMv3 is a multimodal transformer designed for Document Visual Question Answering (DocVQA). Unlike tradit
xhyi document-question-answering

bart qg finetune squad

AI Model
The bart-qg-finetune-squad model is a specialized encoder-decoder architecture optimized for question generati
mghan3624 document-question-answering

AI Model · image-feature-extraction

All (11)

dinov2 small

AI Model
DINOv2-Small is a lightweight, self-supervised vision transformer designed for high-performance image feature
facebook image-feature-extraction

dinov2 base

AI Model
DINOv2 Base is a self-supervised vision transformer designed for high-performance image feature extraction wit
facebook image-feature-extraction

vit small patch14 dinov2.lvd142m

AI Model
The ViT-Small Patch14 (DINOv2) is a lightweight vision transformer pre-trained on a massive, curated dataset o
timm image-feature-extraction

dinov2 large

AI Model
DINOv2 Large is a self-supervised vision transformer designed for high-performance image feature extraction wi
facebook image-feature-extraction

vit base patch16 224 in21k

AI Model
The ViT-Base-Patch16-224 (in21k) is a Vision Transformer pre-trained on the massive ImageNet-21k dataset. Unli
google image-feature-extraction

dino vitb16

AI Model
DINO ViT-B/16 is a self-supervised vision transformer designed for high-performance image feature extraction.
facebook image-feature-extraction

vit base patch14 dinov2.lvd142m

AI Model
The ViT-Base Patch14 DINOv2 model is a high-performance vision transformer trained via self-supervised learnin
timm image-feature-extraction

vit large patch14 reg4 dinov2.lvd142m

AI Model
The ViT-L/14 DINOv2 model is a high-capacity vision transformer trained via self-supervised learning on a mass
timm image-feature-extraction

AI Model · text2text-generation

All (11)

flan t5 base

AI Model
Flan-T5 Base is an instruction-tuned version of the original T5 encoder-decoder framework, designed for develo
google text2text-generation

chronos t5 base

AI Model
Chronos-T5 Base is a specialized time-series forecasting model that treats numerical sequences as language. By
amazon text2text-generation

chronos t5 tiny

AI Model
Chronos-T5 Tiny is a lightweight, text-to-text model based on the T5 architecture, optimized for efficiency an
amazon text2text-generation

parrot paraphraser on T5

AI Model
The Parrot Paraphraser is a specialized text-to-text model built on the T5 architecture, designed specifically
prithivida text2text-generation

prot t5 xl uniref50

AI Model
ProtT5-XL-UniRef50 is a specialized encoder-decoder transformer trained on the UniRef50 protein database, desi
Rostlab text2text-generation

chronos t5 small

AI Model
Chronos-T5 Small is a specialized time-series forecasting model based on the T5 architecture, treating numeric
amazon text2text-generation

unifiedqa t5 small

AI Model
UnifiedQA T5-Small is a lightweight text-to-text model fine-tuned for a broad spectrum of question-answering t
allenai text2text-generation

flan t5 small

AI Model
Flan-T5 Small is a lightweight, encoder-decoder model designed for efficient text-to-text generation. Unlike t
google text2text-generation

AI Model · code-generation

All (10)

DeepSeek Coder 33B

AI Model
Code generation and understanding
DeepSeek code-generation

SWE BENCH generation claude reasoning llm correct swe gym 1500 plus critic qwen code 14b

AI Model
This specialized model is engineered for high-autonomy software engineering tasks, specifically optimized for
secmlr code-generation

Code Generation LLM LoRA Combined Model

AI Model
The Code Generation LLM LoRA Combined Model is a specialized adapter-based solution designed to enhance standa
Rabinovich code-generation

falcon rw 1b code generation llm task2 modelC

AI Model
Falcon RW 1B is a lightweight, specialized small language model (SLM) optimized for code generation tasks. Des
Katochh code-generation

falcon code generation llm

AI Model
Falcon Code is a specialized LLM optimized for software engineering tasks, designed to bridge the gap between
Katochh code-generation

falcon rw 1b code generation llm task2

AI Model
Falcon RW 1B is a compact, code-centric language model designed for efficient deployment in resource-constrain
Katochh code-generation

Code Generation LLM LoRA

AI Model
This LoRA adapter is specifically tuned for code generation tasks, designed to be layered atop a base LLM to e
Rabinovich code-generation

falcon code generation task llm

AI Model
Falcon Code is a specialized LLM optimized for programmatic tasks, designed to bridge the gap between general-
Katochh code-generation

AI Model · natural-language-inference

All (10)

mDeBERTa v3 base xnli multilingual nli 2mil7

AI Model
mDeBERTa-v3-base-xnli is a specialized multilingual model optimized for Natural Language Inference (NLI). Base
MoritzLaurer natural-language-inference

nli mpnet base v2

AI Model
The nli-mpnet-base-v2 is a robust sentence-transformer model optimized for generating high-quality semantic em
sentence-transformers natural-language-inference

nli deberta v3 small

AI Model
The NLI DeBERTa-v3-small is a lightweight cross-encoder optimized for Natural Language Inference (NLI). Unlike
cross-encoder natural-language-inference

bert base turkish cased mean nli stsb tr

AI Model
This model is a specialized BERT-base variant fine-tuned for Turkish Natural Language Inference (NLI) and Sema
emrecan natural-language-inference

nli deberta v3 base

AI Model
The NLI DeBERTa-v3-base is a high-performance cross-encoder optimized for Natural Language Inference tasks. Un
cross-encoder natural-language-inference

bert base nli mean tokens

AI Model
The bert-base-nli-mean-tokens model is a specialized encoder designed to transform text into dense vector embe
sentence-transformers natural-language-inference

nli MiniLM2 L6 H768

AI Model
The nli MiniLM2 L6 H768 is a lightweight cross-encoder optimized for Natural Language Inference (NLI) tasks. U
cross-encoder natural-language-inference

nli deberta v3 xsmall

AI Model
The NLI DeBERTa-v3-xsmall is a compact cross-encoder optimized for Natural Language Inference (NLI) tasks. Unl
cross-encoder natural-language-inference

AI Model · audio-classification

All (10)

clap htsat fused

AI Model
CLAP HTSAT Fused is a specialized audio representation model designed for high-accuracy audio classification a
laion audio-classification

ast finetuned audioset 10 10 0.4593

AI Model
This model is a fine-tuned implementation of the Audio Spectrogram Transformer (AST), specifically optimized f
MIT audio-classification

wav2vec2 large xlsr 53 gender recognition librispeech

AI Model
This model is a specialized fine-tuned version of Meta's wav2vec2-large-xlsr-53, optimized specifically for ge
alefiury audio-classification

ced gguf

AI Model
The ced gguf model is a specialized audio classification tool optimized for local deployment via the GGUF form
mudler audio-classification

audiobox aesthetics

AI Model
Audiobox Aesthetics is a specialized audio classification model from Meta designed to quantify the subjective
facebook audio-classification

gender cls svm ecapa voxceleb

AI Model
This model is a specialized gender classification tool leveraging ECAPA-TDNN embeddings trained on the VoxCele
griko audio-classification

voice gender classifier

AI Model
The Voice Gender Classifier is a specialized audio classification model designed to determine the perceived ge
JaesungHuh audio-classification

hubert large speech emotion recognition russian dusha finetuned

AI Model
The hubert-large-speech-emotion-recognition-russian-dusha is a specialized audio classification model fine-tun
xbgoose audio-classification

AI Model · image-text-retrieval

All (10)

clip vit base patch32

AI Model
CLIP ViT-B/32 is a versatile vision-language model designed for zero-shot image and text understanding. Unlike
openai image-text-retrieval

clip vit large patch14

AI Model
CLIP ViT-L/14 is a powerful vision-language model designed for zero-shot image and text understanding. Unlike
openai image-text-retrieval

clip vit large patch14 336

AI Model
CLIP ViT-L/14@336 is a high-resolution vision-language model designed for precise image-text alignment. By uti
openai image-text-retrieval

CLIP ViT B 32 laion2B s34B b79K

AI Model
CLIP ViT-B/32 (trained on LAION-2B) is a robust vision-language model designed for high-performance image-text
laion image-text-retrieval

fashion clip

AI Model
Fashion CLIP is a domain-specific adaptation of the CLIP architecture, fine-tuned specifically for the fashion
patrickjohncyh image-text-retrieval

CLIP convnext base w laion2B s13B b82K augreg

AI Model
This model is a high-performance vision-language encoder based on the ConvNeXt architecture, trained on the ma
laion image-text-retrieval

clip vit base patch16

AI Model
CLIP ViT-B/16 is a versatile vision-language model designed to map images and text into a shared embedding spa
openai image-text-retrieval

tiny clip text 2

AI Model
Tiny CLIP Text 2 is a lightweight image-text retrieval model designed for developers who need efficient embedd
peft-internal-testing image-text-retrieval

AI Model · image-generation

All (9)

NL Diffusion Image GGUF

AI Model
NL Diffusion Image GGUF brings high-quality image synthesis to local environments by leveraging the GGUF quant
realrebelai image-generation

diffusion models image

AI Model
This diffusion-based image generation model provides a flexible, open-source solution for developers needing h
f5aiteam image-generation

diffusion models image

AI Model
This image generation model leverages a diffusion-based architecture to transform text prompts into high-fidel
nguoidoncui image-generation

stable diffusion trained on yujiro hanma images baki anime fun project

AI Model
This specialized Stable Diffusion checkpoint is fine-tuned on a targeted dataset of Yujiro Hanma from the Baki
nicky007 image-generation

blip image2promt stable diffusion base

AI Model
The BLIP image2prompt model is a specialized vision-language tool designed to reverse-engineer text prompts fr
ifmain image-generation

Medical X ray image generation stable diffusion

AI Model
This Stable Diffusion fine-tune is engineered specifically for synthesizing medical X-ray imagery, bridging th
Osama03 image-generation

stable diffusion finetuned

AI Model
This fine-tuned iteration of Stable Diffusion is designed for developers needing higher precision and stylisti
ImageInception image-generation

stable diffusion base 2.0 text to image 04

AI Model
Stable Diffusion v2.0 is a latent diffusion model designed for high-fidelity text-to-image synthesis. Unlike i
mhbkb image-generation

AI Model · text-to-audio

All (8)

Ace Step1.5

AI Model
Ace Step1.5 is a streamlined text-to-audio model designed for developers needing efficient, high-fidelity spee
ACE-Step text-to-audio

Ace Step1.5 XL DF11 ComfyUI

AI Model
Ace Step1.5 XL DF11 is a specialized text-to-audio model optimized for the ComfyUI ecosystem. Unlike general-p
mingyi456 text-to-audio

fastspeech2 conformer

AI Model
FastSpeech 2 Conformer is a non-autoregressive text-to-speech (TTS) model designed for high-fidelity audio syn
espnet text-to-audio

acestep 5Hz lm 4B

AI Model
The acestep 5Hz lm 4B is a lightweight, open-source text-to-audio model designed for low-latency synthesis. Wi
ACE-Step text-to-audio

fastspeech2 conformer with hifigan

AI Model
This model combines the FastSpeech 2 architecture with Conformer blocks and a HiFi-GAN vocoder to deliver high
espnet text-to-audio

MiniMax Music3

AI Model
MiniMax Music3 is a specialized text-to-audio model designed for high-fidelity music generation. For developer
MiniMaxAI text-to-audio

acestep 5Hz lm 0.6B

AI Model
The acestep 5Hz lm 0.6B is a lightweight, text-to-audio model designed for low-latency synthesis. At 0.6B para
ACE-Step text-to-audio

acestep v15 xl sft

AI Model
acestep v15 xl sft is a specialized text-to-audio model designed for high-fidelity sound synthesis. Unlike gen
ACE-Step text-to-audio

AI Model · image-to-video

All (7)

Wan2.2 I2V A14B GGUF

AI Model
Wan2.2 I2V A14B GGUF is a quantized image-to-video generation model designed for developers seeking high-fidel
QuantStack image-to-video

Minimax h3 Turbo

AI Model
Minimax h3 Turbo is a specialized image-to-video generation model designed for developers needing high-fidelit
lightx2v image-to-video

Wan2.2 Distill Loras

AI Model
Wan2.2 Distill LoRAs are specialized adapters designed to optimize image-to-video generation by leveraging kno
lightx2v image-to-video

LTX2.3 10Eros

AI Model
LTX2.3 10Eros is a specialized image-to-video generation model designed for developers building dynamic visual
TenStrip image-to-video

Wan2.2 I2V A14B Diffusers

AI Model
Wan2.2 I2V A14B is a high-capacity image-to-video diffusion model designed for developers requiring cinematic
Wan-AI image-to-video

MiniMax H3 encoder GGUF

AI Model
The MiniMax H3 encoder, now available in GGUF format, provides a quantized implementation of the vision encodi
joeygambino image-to-video

Wan2.1 I2V 14B 480P gguf

AI Model
Wan2.1 I2V 14B 480P is a specialized image-to-video diffusion model optimized for developers seeking a balance
city96 image-to-video

AI Model · image-segmentation

All (7)

clipseg rd64 refined

AI Model
ClipSeg RD64 Refined is a specialized image segmentation model designed for zero-shot performance, leveraging
CIDAS image-segmentation

BiRefNet

AI Model
BiRefNet is a high-performance image segmentation model specifically engineered for precise binary reference-b
ZhengPeng7 image-segmentation

segformer b0 finetuned ade 512 512

AI Model
SegFormer-B0 (finetuned on ADE20K) is a lightweight, hierarchical Transformer-based model designed for efficie
Xenova image-segmentation

face parsing

AI Model
This face parsing model provides high-precision semantic segmentation of human facial features, mapping specif
jonathandinu image-segmentation

oneformer cityscapes swin large

AI Model
OneFormer Cityscapes Swin-Large is a high-performance image segmentation model designed for precise semantic,
shi-labs image-segmentation

coco panoptic eomt large 640

AI Model
The coco panoptic eomt large 640 is a specialized image segmentation model designed for high-precision panopti
tue-mps image-segmentation

modnet

AI Model
ModNet is a lightweight, real-time portrait matting model designed for high-fidelity alpha matte estimation. U
Xenova image-segmentation

AI Model · ocr

All (6)

chandra ocr 2

AI Model
Chandra OCR 2 is a specialized vision-language model optimized for high-accuracy text extraction and document
datalab-to ocr

DeepSeek OCR

AI Model
DeepSeek OCR is a specialized vision-language model designed to bridge the gap between raw image pixels and st
deepseek-ai ocr

surya ocr 2

AI Model
Surya OCR 2 is a high-performance vision-language model designed for precise document analysis and text extrac
datalab-to ocr

DeepSeek OCR 2

AI Model
DeepSeek OCR 2 is a specialized vision-language model engineered to bridge the gap between raw image data and
deepseek-ai ocr

Unlimited OCR AWQ

AI Model
Unlimited OCR AWQ is a specialized vision-language model optimized for high-accuracy optical character recogni
sahilchachra ocr

GOT OCR2 0

AI Model
GOT OCR2.0 is a specialized vision-language model designed to bridge the gap between raw image data and struct
stepfun-ai ocr

AI Model · sentiment-analysis

All (4)

sentiment analysis fine tuned model

AI Model
This fine-tuned sentiment analysis model is designed for developers needing a lightweight, specialized tool fo
alsgyu sentiment-analysis

twitter roberta base sentiment

AI Model
The twitter-roberta-base-sentiment model is a specialized encoder based on the RoBERTa architecture, fine-tune
cardiffnlp sentiment-analysis

Bangla twoclass Sentiment Analyzer

AI Model
The Bangla Two-Class Sentiment Analyzer is a specialized NLP model designed for binary sentiment classificatio
Arunavaonly sentiment-analysis

rubert base cased sentiment rusentiment

AI Model
RuBERT-base-cased-sentiment (RuSentiment) is a specialized transformer model optimized for sentiment analysis
blanchefort sentiment-analysis

AI Model · multimodal-representation

All (4)

ChestVision Fine Tuned Vision Language Model

AI Model
ChestVision is a specialized Vision Language Model (VLM) fine-tuned specifically for medical imaging, focusing
AhmedOkasha multimodal-representation

Touch Vision Language Models

AI Model
Touch Vision Language Models (TVLMs) extend traditional multimodal architectures by integrating tactile sensin
mlfu7 multimodal-representation

vision language garment model

AI Model
The Vision Language Garment model is a specialized multimodal representation tool designed to bridge the gap b
ackermannj multimodal-representation

vision language model

AI Model
This Vision Language Model (VLM) is a multimodal architecture designed to bridge the gap between visual percep
kirangowda3101 multimodal-representation

AI Model · image-quality-assessment

All (3)

generated image quality assessment

AI Model
This model provides a specialized solution for evaluating the fidelity and visual quality of synthetic images.
arman-chopikyan image-quality-assessment

face image quality assessment ediffiqa

AI Model
EdiffiQA is a specialized image quality assessment (IQA) model integrated within the OpenCV ecosystem, designe
opencv image-quality-assessment

AAL Plus Image Quality Assessment

AI Model
AAL Plus is a specialized Image Quality Assessment (IQA) model designed to provide objective, quantitative met
LoliRimuru image-quality-assessment

AI Model · text generation

All (3)

GPT-4

AI Model
Most capable GPT model for complex tasks
OpenAI text generation

Claude 3.5 Sonnet

AI Model
Advanced AI assistant with superior reasoning
Anthropic text generation

Gemini 1.5 Pro

AI Model
Multimodal AI with long context
Google text generation

AI Model · audio

All (1)

Whisper Large V3

AI Model
Robust speech recognition model
OpenAI audio

AI Skill · General

All (366)

Job Interviewer

AI Skill
I want you to act as an interviewer. I will be the candidate and you will ask me the interview questions for t
General f

Travel Guide

AI Skill
I want you to act as a travel guide. I will write you my location and you will suggest a place to visit near m
General koksalkapucuoglu

Plagiarism Checker

AI Skill
I want you to act as a plagiarism checker. I will write you sentences and you will only reply undetected in pl
General yetk1n

Character

AI Skill
I want you to act like {character} from {series}. I want you to respond and answer like {character} using the
General BRTZL

Advertiser

AI Skill
I want you to act as an advertiser. You will create a campaign to promote a product or service of your choice.
General devisasari

Storyteller

AI Skill
I want you to act as a storyteller. You will come up with entertaining stories that are engaging, imaginative
General devisasari

Football Commentator

AI Skill
I want you to act as a football commentator. I will give you descriptions of football matches in progress and
General devisasari

Stand-up Comedian

AI Skill
I want you to act as a stand-up comedian. I will provide you with some topics related to current events and yo
General devisasari

AI Skill · Coding

All (284)

Code Reviewer

AI Skill
Review code for bugs, security issues, and best practices
Coding

Ethereum Developer

AI Skill
Imagine you are an experienced Ethereum developer tasked with creating a smart contract for a blockchain messe
Coding ameya-2003

Linux Terminal

AI Skill
I want you to act as a linux terminal. I will type commands and you will reply with what the terminal should s
Coding f

JavaScript Console

AI Skill
I want you to act as a javascript console. I will type commands and you will reply with what the javascript co
Coding omerimzali

Excel Sheet

AI Skill
I want you to act as a text based excel. you'll only reply me the text-based 10 rows excel sheet with row numb
Coding f

UX/UI Developer

AI Skill
I want you to act as a UX/UI developer. I will provide some details about the design of an app, website or oth
Coding devisasari

Cyber Security Specialist

AI Skill
I want you to act as a cyber security specialist. I will provide some specific information about how data is s
Coding devisasari

Web Design Consultant

AI Skill
I want you to act as a web design consultant. I will provide you with details related to an organization needi
Coding devisasari

AI Skill · Data

All (53)

ad-campaign-analyzer

AI Skill
Analyze cross-channel campaign data, quantify uncertainty, and propose evidence-labeled budget tests without o
Data Agentic Awesome Skills 社区

alpha-vantage

AI Skill
Access 20+ years of global financial data: equities, options, forex, crypto, commodities, economic indicators,
Data Agentic Awesome Skills 社区

analytics-tracking

AI Skill
Set up, audit, and debug analytics tracking implementation — GA4, Google Tag Manager, event taxonomy, conversi
Data Alireza Rezvani

angular-ui-patterns

AI Skill
Modern Angular UI patterns for loading states, error handling, and data display. Use when building UI componen
Data Agentic Awesome Skills 社区

anti-reversing-techniques

AI Skill
AUTHORIZED USE ONLY: This skill contains dual-use security techniques. Before proceeding with any bypass or an
Data Agentic Awesome Skills 社区

apify-audience-analysis

AI Skill
Understand audience demographics, preferences, behavior patterns, and engagement quality across Facebook, Inst
Data Agentic Awesome Skills 社区

apify-ecommerce

AI Skill
Extract product data, prices, reviews, and seller information from any e-commerce platform using Apify's E-com
Data Agentic Awesome Skills 社区

apify-market-research

AI Skill
Analyze market conditions, geographic opportunities, pricing, consumer behavior, and product validation across
Data Agentic Awesome Skills 社区

AI Skill · Business

All (41)

Marketing Expert

AI Skill
Create marketing strategies and ad copy
Business

Recruiter

AI Skill
I want you to act as a recruiter. I will provide some information about job openings, and it will be your job
Business devisasari

Startup Idea Generator

AI Skill
Generate digital startup ideas based on the wish of the people. For example, when I say "I wish there's a big
Business buddylabsai

Salesperson

AI Skill
I want you to act as a salesperson. Try to market something to me, but make what you're trying to market look
Business biaksoy

Startup Tech Lawyer

AI Skill
I will ask of you to prepare a 1 page draft of a design partner agreement between a tech startup with IP and a
Business jonathandn

Product Manager

AI Skill
Please acknowledge my following request. Please respond to me as a product manager. I will ask for subject, an
Business orinachum

internal-comms

AI Skill
Use when a Head of People Ops, BizOps lead, or Internal Communications owner needs to draft and sequence an in
Business Alireza Rezvani

slack-gif-creator

AI Skill
Knowledge and utilities for creating animated GIFs optimized for Slack. Provides constraints, validation tools
Business Anthropic

AI Skill · Design

All (41)

Prompt Generator

AI Skill
I want you to act as a prompt generator. Firstly, I will give you a title like this: "Act as an English Pronun
Design iuzn

Digital Art Gallery Guide

AI Skill
I want you to act as a digital art gallery guide. You will be responsible for curating virtual exhibits, resea
Design devisasari

Midjourney Prompt Generator

AI Skill
I want you to act as a prompt generator for Midjourney's artificial intelligence program. Your job is to provi
Design iuzn

ChatGPT Prompt Generator

AI Skill
I want you to act as a ChatGPT prompt generator, I will send a topic, you have to generate a ChatGPT prompt ba
Design aitrainee

brand-guidelines

AI Skill
When the user wants to apply, document, or enforce brand guidelines for any product or company. Also use when
Design Alireza Rezvani

canvas-design

AI Skill
Create beautiful visual art in .png and .pdf documents using design philosophy. You should use this skill when
Design Anthropic

frontend-design

AI Skill
Guidance for distinctive, intentional visual design when building new UI or reshaping an existing one. Helps w
Design Anthropic

theme-factory

AI Skill
Toolkit for styling artifacts with a theme. These artifacts can be slides, docs, reportings, HTML landing page
Design Anthropic

AI Skill · Writing

All (30)

Screenwriter

AI Skill
I want you to act as a screenwriter. You will develop an engaging and creative script for either a feature len
Writing devisasari

Poet

AI Skill
I want you to act as a poet. You will create poems that evoke emotions and have the power to stir people's sou
Writing devisasari

Essay Writer

AI Skill
I want you to act as an essay writer. You will need to research a given topic, formulate a thesis statement, a
Writing devisasari

Journalist

AI Skill
I want you to act as a journalist. You will report on breaking news, write feature stories and opinion pieces,
Writing devisasari

doc-coauthoring

AI Skill
Guide users through a structured workflow for co-authoring documentation. Use when user wants to write documen
Writing Anthropic

docx

AI Skill
Use this skill whenever the user wants to create, read, edit, or manipulate Word documents (.docx files) or Wo
Writing Anthropic

adhx

AI Skill
Fetch any X/Twitter post as clean LLM-friendly JSON. Converts x.com, twitter.com, or adhx.com links into struc
Writing Agentic Awesome Skills 社区

article-illustrations

AI Skill
Generate hand-drawn 16:9 article illustrations with the Grav character IP, sparse annotations, and absurd but
Writing Agentic Awesome Skills 社区

AI Skill · Office

All (10)

pdf

AI Skill
Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text
Office Anthropic

pptx

AI Skill
Use this skill any time a .pptx or .potx file is involved in any way — as input, output, or both. This include
Office Anthropic

xlsx

AI Skill
Use this skill any time a spreadsheet file is the primary input or output. This means any task where the user
Office Anthropic

azure-ai-translation-document-py

AI Skill
Azure AI Document Translation SDK for batch translation of documents with format preservation. Use for transla
Office Agentic Awesome Skills 社区

board-prep

AI Skill
Board meeting preparation for the adversarial scenario, not the friendly one. Forces numbers-cold mastery, ant
Office Alireza Rezvani

contract-and-proposal-writer

AI Skill
Generate professional, jurisdiction-aware business documents: freelance contracts, project proposals, SOWs, ND
Office Alireza Rezvani

demo-video

AI Skill
Use when the user asks to create a demo video, product walkthrough, feature showcase, animated presentation, m
Office Alireza Rezvani

md-slides

AI Skill
Converts a markdown deck (slides separated by `---` HR boundaries or by `# ` H1 headings, with optional `<!--
Office Alireza Rezvani

AI Skill · Education

All (8)

Motivational Coach

AI Skill
I want you to act as a motivational coach. I will provide you with some information about someone's goals and
Education devisasari

Debate Coach

AI Skill
I want you to act as a debate coach. I will provide you with a team of debaters and the motion for their upcom
Education devisasari

Relationship Coach

AI Skill
I want you to act as a relationship coach. I will provide some details about the two people involved in a conf
Education devisasari

AI Writing Tutor

AI Skill
I want you to act as an AI writing tutor. I will provide you with a student who needs help improving their wri
Education devisasari

Life Coach

AI Skill
I want you to act as a life coach. I will provide some details about my current situation and goals, and it wi
Education vduchew

Instructor in a School

AI Skill
I want you to act as an instructor in a school, teaching algorithms to beginners. You will provide code exampl
Education omt66

Public Speaking Coach

AI Skill
I want you to act as a public speaking coach. You will develop clear communication strategies, provide profess
Education devisasari

Talent Coach

AI Skill
I want you to act as a Talent Coach for interviews. I will give you a job title and you'll suggest what should
Education guillaumefalourd

AI Skill · Translation

All (6)

English Translator and Improver

AI Skill
I want you to act as an English translator, spelling corrector and improver. I will speak to you in any langua
Translation f

English Pronunciation Helper

AI Skill
I want you to act as an English pronunciation assistant for ${Mother Language:Turkish} speaking people. I will
Translation f

Spoken English Teacher and Improver

AI Skill
I want you to act as a spoken English teacher and improver. I will speak to you in English and you will reply
Translation atx735

New Language Creator

AI Skill
I want you to translate the sentences I wrote into a new made up language. I will write the sentence, and you
Translation willfeldman

Language Detector

AI Skill
I want you act as a language detector. I will type a sentence in any language and you will answer me in which
Translation dogukandogru

Speech-Language Pathologist (SLP)

AI Skill
I want you to act as a speech-language pathologist (SLP) and come up with new speech patterns, communication s
Translation leonwangg1

AI Skill · Marketing

All (1)

SEO Optimizer

AI Skill
Optimize content for search engines
Marketing

MCP Tool · General

All (29)

jdubois/azure-cli-mcp

MCP Tool
A wrapper around the Azure CLI command line that allows you to talk directly to Azure
General Community

nwiizo/tfmcp

MCP Tool
🦀 🏠 - A Terraform MCP server allowing AI assistants to manage and operate Terraform environments, enabling r
General Community

automateyournetwork/pyATS_MCP

MCP Tool
Cisco pyATS server enabling structured, model-driven interaction with network devices.
General Community

maxim-saplin/mcp_safe_local_python_executor

MCP Tool
Safe Python interpreter based on HF Smolagents `LocalPythonExecutor`
General Community

sonirico/mcp-shell

MCP Tool
🏎️ 🏠 🍎 🪟 🐧 Give hands to AI. MCP server to run shell commands securely, auditably, and on demand on isola
General Community

tumf/mcp-shell-server

MCP Tool
A secure shell command execution server implementing the Model Context Protocol (MCP)
General Community

areweai/tsgram-mcp

MCP Tool
TSgram: Telegram + Claude with local workspace access on your phone in typescript. Read, write, and vibe code
General Community

teddyzxcv/ntfy-mcp

MCP Tool
The MCP server that keeps you informed by sending the notification on phone using ntfy
General Community

MCP Tool · Web

All (16)

thinkchainai/mcpbundles

MCP Tool
MCP Bundles: Create custom bundles of tools and connect providers with OAuth or API keys. Use one MCP server a
Web Community

microsoft/playwright-mcp

MCP Tool
Official Microsoft Playwright MCP server, enabling LLMs to interact with web pages through structured accessib
Web Community

hashicorp/terraform-mcp-server

MCP Tool
🎖️🏎️☁️ - The official Terraform MCP Server seamlessly integrates with the Terraform ecosystem, enabling prov
Web Community

sawa-zen/vrchat-mcp

MCP Tool
📇 🏠 This is an MCP server for interacting with the VRChat API. You can retrieve information about friends, w
Web Community

saurabhsharma2u/search-console-mcp

MCP Tool
An MCP server to interact with Google Search Console and Bing Webmasters.
Web Community

OpenZeppelin/contracts-wizard

MCP Tool
A Model Context Protocol (MCP) server that allows AI agents to generate secure smart contracts in multiples la
Web Community

apiarya/wemo-mcp-server

MCP Tool
[![wemo-mcp-server MCP server](https://glama.ai/mcp/servers/@apiarya/wemo-mcp-server/badges/score.svg)](https:
Web Community

hbg/mcp-paperswithcode

MCP Tool
🐍 ☁️ MCP to search through PapersWithCode API
Web Community

MCP Tool · Database

All (11)

mindsdb/mindsdb

MCP Tool
Connect and unify data across various platforms and databases with [MindsDB as a single MCP server](https://do
Database Community

Aiven-Open/mcp-aiven

MCP Tool
🐍 ☁️ 🎖️ - Navigate your [Aiven projects](https://go.aiven.io/mcp-server) and interact with the PostgreSQL®,
Database Community

alexanderzuev/supabase-mcp-server

MCP Tool
Supabase MCP Server with support for SQL query execution and database exploration tools
Database Community

bram2w/baserow

MCP Tool
Baserow database integration with table search, list, and row create, read, update, and delete capabilities.
Database Community

planetscale/mcp

MCP Tool
The PlanetScale CLI includes an MCP server that provides AI tools direct access to your PlanetScale databases.
Database Community

longportapp/openapi

MCP Tool
🐍 ☁️ - LongPort OpenAPI provides real-time stock market data, provides AI access analysis and trading capabil
Database Community

traceloop/opentelemetry-mcp-server

MCP Tool
🐍🏠 - An MCP server for connecting to any OpenTelemetry backend (Datadog, Grafana, Dynatrace, Traceloop, etc.
Database Community

thinkchainai/agentinterviews_mcp

MCP Tool
Conduct AI-powered qualitative research interviews and surveys at scale with [Agent Interviews](https://agenti
Database Community

MCP Tool · File System

All (6)

finmap-org/mcp-server

MCP Tool
[finmap.org](https://finmap.org/) MCP server provides comprehensive historical data from the US, UK, Russian a
File System Community

urlbox/urlbox-mcp-server

MCP Tool
📇 🏠 A reliable MCP server for generating and managing screenshots, PDFs, and videos, performing AI-powered s
File System Community

Azure/azure-mcp

MCP Tool
Official Microsoft MCP server for Azure services including Storage, Cosmos DB, and Azure Monitor.
File System Community

mediar-ai/screenpipe

MCP Tool
🎖️ 🦀 🏠 🍎 Local-first system capturing screen/audio with timestamped indexing, SQL/embedding storage, seman
File System Community

MonadsAG/capsulecrm-mcp

MCP Tool
📇 ☁️ Allows AI clients to manage contacts, opportunities and tasks in Capsule CRM including Claude Desktop re
File System Community

MCP Filesystem

MCP Tool
Secure file system access for Claude
File System Anthropic

MCP Tool · Search

All (3)

andybrandt/mcp-simple-arxiv

MCP Tool
🐍 ☁️ MCP for LLM to search and read papers from arXiv
Search Community

andybrandt/mcp-simple-pubmed

MCP Tool
🐍 ☁️ MCP to search and read medical / life sciences papers from PubMed.
Search Community

tan-yong-sheng/ai-vision-mcp

MCP Tool
📇 🏠 🍎 🪟 🐧 - Multimodal AI vision MCP server for image, video, and object detection analysis. Enables UI/U
Search Community

MCP Tool · Development

All (1)

TamarEngel/jira-github-mcp

MCP Tool
MCP server that integrates Jira and GitHub to automate end-to-end developer workflows, from issue tracking to
Development Community
Join our Telegram