Categories

AI Models Marketplace

Explore the latest and most popular AI models — LLM, vision, audio and more — to pick and integrate the right one.

tiny Qwen2ForCausalLM 2.5

Apache-2.0
The tiny Qwen2ForCausalLM 2.5 is a compact, efficient language model designed for low-latency applications and edge depl...
trl-internal-testing text-generation
15.6M 0

Qwen3.6 35B A3B NVFP4

apache-2.0
The Qwen3.6 35B A3B NVFP4 is a specialized iteration of the Qwen series, optimized specifically for NVIDIA hardware usin...
nvidia text-generation
12.2M 0

Llama 3.2 1B Instruct

llama3.2
Llama 3.2 1B Instruct is a lightweight, instruction-tuned model designed for high-efficiency deployment on edge devices ...
meta-llama text-generation
8.9M 0

Llama 3.1 405B

Llama 3.1
Open source large language model
Meta text-generation 405B
8.0M 9.5K

gpt oss 20b

apache-2.0
gpt oss 20b is a mid-sized open-source language model designed for developers who need a balance between local deploymen...
openai text-generation
8.0M 0

Llama 3.1 8B Instruct

llama3.1
Llama 3.1 8B Instruct is a dense decoder-only model designed for high-efficiency deployment without sacrificing complex ...
meta-llama text-generation
7.5M 0

Qwen3 8B

apache-2.0
Qwen3 8B is a compact yet powerful language model designed for high-efficiency deployment without sacrificing reasoning ...
Qwen text-generation
7.3M 325

Qwen2.5 7B Instruct

apache-2.0
Qwen2.5 7B Instruct is a dense decoder-only model designed for high-efficiency deployment without sacrificing reasoning ...
Qwen text-generation
7.2M 502

Qwen2.5 0.5B Instruct

apache-2.0
Qwen2.5-0.5B-Instruct is a lightweight, instruction-tuned LLM designed for environments where memory overhead and latenc...
Qwen text-generation
6.8M 0

Qwen3 0.6B

apache-2.0
Qwen3 0.6B is a highly compact language model designed for efficiency and low-latency deployment. At under one billion p...
Qwen text-generation
5.6M 248

OTel 2.0 LLM 31B IT

apache-2.0
OTel 2.0 LLM 31B IT is an instruction-tuned model designed for developers needing a balance between high-parameter reaso...
farbodtavakkoli text-generation
5.3M 0

Mixtral 8x7B

Apache 2.0
Sparse mixture of experts model
Mistral AI text-generation 47B
5.0M 7.2K

Qwen3 32B

apache-2.0
Qwen3 32B is a mid-sized LLM designed to balance high-reasoning performance with efficient deployment. For developers, t...
Qwen text-generation
4.4M 338

Qwen3 1.7B

apache-2.0
Qwen3 1.7B is a compact, high-efficiency language model designed for developers who need strong performance without the ...
Qwen text-generation
1.2M 79

Qwen2.5 1.5B Instruct

apache-2.0
Qwen2.5 1.5B Instruct is a lightweight, instruction-tuned LLM designed for high-efficiency deployment. Despite its small...
Qwen text-generation
931.4K 139

DeepSeek R1

mit
DeepSeek R1 is a reasoning-focused model designed to compete with high-end frontier LLMs by implementing advanced reinfo...
deepseek-ai text-generation
926.9K 1.4K

Qwen3 Coder 30B A3B Instruct GGUF

apache-2.0
Qwen3 Coder 30B A3B Instruct is a specialized Mixture-of-Experts (MoE) model optimized for high-performance programming ...
unsloth text-generation
108.1K 28

gpt2

mit
GPT-2 is a foundational transformer-based language model that marked a shift toward zero-shot learning in NLP. For devel...
openai-community text-generation
46.0K 11
Join our Telegram