Distilled LLaMA by DeepSeek, fast and optimized for real-world tasks
1y
100K+
79
Google’s latest Gemma, in its QAT (quantization aware trained) variant
11m
100K+
23
Newest LLama 3 release with improved reasoning and generation quality
1y
100K+
21
Ministral 3: compact vision-enabled model with near-24B performance, optimized for local edge use
9m
100K+
4
Efficient 80B MoE coding model with 3B activated params, 256K context, and agentic capabilities
6m
50K+
3
DeepCoder-14B-Preview is a code reasoning LLM fine-tuned to scale up to long context lengths
1y
50K+
15
Gemma 4: multimodal open AI models by Google, optimized for reasoning, coding, and long context.
5m
50K+
SmolLM3 is a 3.1B model for efficient on-device use, with strong performance in chat
1y
50K+
8
Image generation model, uses a base latent diffusion model plus a refiner.
7m
50K+
7
Kimi K2 Thinking: open-source agent with deep reasoning, stable tool use, fast INT4, 256k context.
9m
50K+
4
DeepSeek-V3.2 boosts efficiency and reasoning with DSA, scalable RL, agentic data—IMO/IOI wins.
9m
50K+
11
Ministral 3: compact vision-enabled model with near-24B performance, optimized for local edge use
9m
50K+
4