GLM-5.1 is our next-generation flagship model for agentic engineering, with significantly stronger coding capabilities than its predecessor. It achieves state-of-the-art performance on SWE-Bench Pro and leads GLM-5 by a wide margin.
Araç Kullanımı
Akıl Yürütme
⬇ 2.3M indirme · 0 varyant
· 04.2026
MiniMax's M2-series model for coding, agentic workflows, and professional productivity.
Araç Kullanımı
Akıl Yürütme
⬇ 2.3M indirme · 0 varyant
· 03.2026
Cogito v1 Preview is a family of hybrid reasoning models by Deep Cogito that outperform the best available open models of the same size, including counterparts from LLaMA, DeepSeek, and Qwen across most standard benchmarks.
Araç Kullanımı
3b
8b
14b
32b
⬇ 2.1M indirme · 20 varyant
· 04.2025
Qwen3-Coder-Next is a coding-focused language model from Alibaba's Qwen team, optimized for agentic coding workflows and local development.
Araç Kullanımı
⬇ 2M indirme · 3 varyant
· 02.2026
Meta's latest collection of multimodal models.
Görüntü
Araç Kullanımı
16x17b
128x17b
⬇ 1.8M indirme · 11 varyant
· 06.2025
Hermes 3 is the latest version of the flagship Hermes series of LLMs by Nous Research
Araç Kullanımı
3b
8b
70b
405b
⬇ 1.6M indirme · 65 varyant
· 12.2024
Command R is a Large Language Model optimized for conversational interaction and long context tasks.
Araç Kullanımı
35b
⬇ 1.5M indirme · 32 varyant
· 08.2024
Phi-4-mini brings significant enhancements in multilingual support, reasoning, and mathematics, and now, the long-awaited function calling feature is finally supported.
Araç Kullanımı
3.8b
⬇ 1.4M indirme · 5 varyant
· 02.2025
The Ministral 3 family is designed for edge deployment, capable of running on a wide range of hardware.
Görüntü
Araç Kullanımı
3b
8b
14b
⬇ 1.4M indirme · 13 varyant
· 12.2025
Granite 4 features improved instruction following (IF) and tool-calling capabilities, making them more effective in enterprise applications.
Araç Kullanımı
350m
1b
3b
⬇ 1.4M indirme · 17 varyant
· 10.2025
As the strongest model in the 30B class, GLM-4.7-Flash offers a new option for lightweight deployment that balances performance and efficiency.
Araç Kullanımı
Akıl Yürütme
⬇ 1.4M indirme · 4 varyant
· 06.2026
Magistral is a small, efficient reasoning model with 24B parameters.
Araç Kullanımı
Akıl Yürütme
24b
⬇ 1.4M indirme · 5 varyant
· 06.2025
LFM2.5 is a new family of hybrid models designed for on-device deployment.
Araç Kullanımı
Akıl Yürütme
1.2b
⬇ 1.3M indirme · 5 varyant
· 01.2026
Mistral Large 2 is Mistral's new flagship model that is significantly more capable in code generation, mathematics, and reasoning with 128k context window and support for dozens of languages.
Araç Kullanımı
123b
⬇ 1.3M indirme · 32 varyant
· 11.2024
IBM Granite 2B and 8B models are 128K context length language models that have been fine-tuned for improved reasoning and instruction-following capabilities.
Araç Kullanımı
2b
8b
⬇ 1.1M indirme · 3 varyant
· 04.2025
LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient.
Araç Kullanımı
24b
⬇ 1.1M indirme · 6 varyant
· 02.2026
Devstral: the best open source model for coding agents
Araç Kullanımı
24b
⬇ 1M indirme · 5 varyant
· 07.2025
The IBM Granite 2B and 8B models are text-only dense LLMs trained on over 12 trillion tokens of data, demonstrated significant improvements over their predecessors in performance and speed in IBM’s initial testing.
Araç Kullanımı
2b
8b
⬇ 1M indirme · 33 varyant
· 01.2025
The IBM Granite 2B and 8B models are designed to support tool-based use cases and support for retrieval augmented generation (RAG), streamlining code generation, translation and bug fixing.
Araç Kullanımı
2b
8b
⬇ 1M indirme · 33 varyant
· 11.2024
Cohere For AI's language models trained to perform well across 23 different languages.
Araç Kullanımı
8b
32b
⬇ 992.2K indirme · 33 varyant
· 10.2024
A series of models from Groq that represent a significant advancement in open-source AI capabilities for tool use/function calling.
Araç Kullanımı
8b
70b
⬇ 979.3K indirme · 33 varyant
· 07.2024
A compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.
Görüntü
Araç Kullanımı
2b
⬇ 971.6K indirme · 5 varyant
· 02.2025
The IBM Granite 1B and 3B models are the first mixture of experts (MoE) Granite models from IBM designed for low latency usage.
Araç Kullanımı
1b
3b
⬇ 937.8K indirme · 33 varyant
· 11.2024
24B model that excels at using tools to explore codebases, editing multiple files and power software engineering agents.
Görüntü
Araç Kullanımı
24b
⬇ 935.1K indirme · 6 varyant
· 12.2025