The tiny Qwen2ForCausalLM 2.5 is a compact, efficient language model designed for low-latency applications and edge depl...
trl-internal-testing
text-generation
15.6M
0
The Qwen3.6 35B A3B NVFP4 is a specialized iteration of the Qwen series, optimized specifically for NVIDIA hardware usin...
nvidia
text-generation
12.2M
0
Llama 3.2 1B Instruct is a lightweight, instruction-tuned model designed for high-efficiency deployment on edge devices ...
meta-llama
text-generation
8.9M
0
Open source large language model
Meta
text-generation
405B
8.0M
9.5K
gpt oss 20b is a mid-sized open-source language model designed for developers who need a balance between local deploymen...
openai
text-generation
8.0M
0
Llama 3.1 8B Instruct is a dense decoder-only model designed for high-efficiency deployment without sacrificing complex ...
meta-llama
text-generation
7.5M
0
Qwen3 8B is a compact yet powerful language model designed for high-efficiency deployment without sacrificing reasoning ...
Qwen
text-generation
7.3M
325
Qwen2.5 7B Instruct is a dense decoder-only model designed for high-efficiency deployment without sacrificing reasoning ...
Qwen
text-generation
7.2M
502
Qwen2.5-0.5B-Instruct is a lightweight, instruction-tuned LLM designed for environments where memory overhead and latenc...
Qwen
text-generation
6.8M
0
Qwen3 0.6B is a highly compact language model designed for efficiency and low-latency deployment. At under one billion p...
Qwen
text-generation
5.6M
248
OTel 2.0 LLM 31B IT is an instruction-tuned model designed for developers needing a balance between high-parameter reaso...
farbodtavakkoli
text-generation
5.3M
0
Sparse mixture of experts model
Mistral AI
text-generation
47B
5.0M
7.2K
Qwen3 32B is a mid-sized LLM designed to balance high-reasoning performance with efficient deployment. For developers, t...
Qwen
text-generation
4.4M
338
Qwen3 1.7B is a compact, high-efficiency language model designed for developers who need strong performance without the ...
Qwen
text-generation
1.2M
79
Qwen2.5 1.5B Instruct is a lightweight, instruction-tuned LLM designed for high-efficiency deployment. Despite its small...
Qwen
text-generation
931.4K
139
DeepSeek R1 is a reasoning-focused model designed to compete with high-end frontier LLMs by implementing advanced reinfo...
deepseek-ai
text-generation
926.9K
1.4K
Qwen3 Coder 30B A3B Instruct is a specialized Mixture-of-Experts (MoE) model optimized for high-performance programming ...
unsloth
text-generation
108.1K
28
GPT-2 is a foundational transformer-based language model that marked a shift toward zero-shot learning in NLP. For devel...
openai-community
text-generation
46.0K
11