Building upon Mistral Small 3, Mistral Small 3.1 (2503) adds state-of-the-art vision understanding and enhances long context capabilities up to 128k tokens without compromising text performance.
Görüntü
Araç Kullanımı
24b
⬇ 777.2K indirme · 5 varyant
· 04.2025
EXAONE Deep exhibits superior capabilities in various reasoning tasks including math and coding benchmarks, ranging from 2.4B to 32B parameters developed and released by LG AI Research.
2.4b
7.8b
32b
⬇ 758K indirme · 13 varyant
· 03.2025
DeepSeek-V3.1-Terminus is a hybrid model that supports both thinking mode and non-thinking mode.
Araç Kullanımı
Akıl Yürütme
671b
⬇ 720.5K indirme · 7 varyant
· 09.2025
An experimental 1.1B parameter model trained on the new Dolphin 2.8 dataset by Eric Hartford and based on TinyLlama.
1.1b
⬇ 714.7K indirme · 18 varyant
· 01.2024
nomic-embed-text-v2-moe is a multilingual MoE text embedding model that excels at multilingual retrieval.
Embedding
⬇ 703.8K indirme · 0 varyant
· 12.2025
A commercial-friendly small language model by NVIDIA optimized for roleplay, RAG QA, and function calling.
Araç Kullanımı
4b
⬇ 698.2K indirme · 17 varyant
· 09.2024
Nemotron-3-Nano is a new Standard for Efficient, Open, and Intelligent Agentic Models, now updated with a 4B parameter count model.
Araç Kullanımı
Akıl Yürütme
4b
30b
⬇ 686.6K indirme · 9 varyant
· 03.2026
A versatile model for AI software development scenarios, including code completion.
9b
⬇ 681K indirme · 17 varyant
· 07.2024
Mistral OpenOrca is a 7 billion parameter model, fine-tuned on top of the Mistral 7B model using the OpenOrca dataset.
7b
⬇ 671.1K indirme · 17 varyant
· 10.2023
Fine-tuned Llama 2 model to answer medical questions based on an open source medical dataset.
7b
⬇ 652.2K indirme · 17 varyant
· 10.2023
NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
Görüntü
Araç Kullanımı
Akıl Yürütme
33b
⬇ 642.6K indirme · 4 varyant
· 04.2026
Uncensored version of Wizard LM model
13b
⬇ 634.1K indirme · 18 varyant
· 10.2023
OpenCoder is an open and reproducible code LLM family which includes 1.5B and 8B models, supporting chat in English and Chinese languages.
1.5b
8b
⬇ 626.4K indirme · 9 varyant
· 11.2024
A high-performing model trained with a new technique called Reflection-tuning that teaches a LLM to detect mistakes in its reasoning and correct course.
70b
⬇ 608.8K indirme · 17 varyant
· 09.2024
Llama-3.1-Nemotron-70B-Instruct is a large language model customized by NVIDIA to improve the helpfulness of LLM generated responses to user queries.
Araç Kullanımı
70b
⬇ 606.3K indirme · 17 varyant
· 10.2024
Great code generation model based on Llama2.
13b
⬇ 585.6K indirme · 19 varyant
· 10.2023
The Nous Hermes 2 model from Nous Research, now trained over Mixtral.
8x7b
⬇ 581.8K indirme · 18 varyant
· 12.2024
Athene-V2 is a 72B parameter model which excels at code completion, mathematics, and log extraction tasks.
Araç Kullanımı
72b
⬇ 580.9K indirme · 17 varyant
· 11.2024
The first installment in the Qwen3-Next series with strong performance in terms of both parameter efficiency and inference speed.
Araç Kullanımı
Akıl Yürütme
80b
⬇ 579.2K indirme · 9 varyant
· 12.2025
MegaDolphin-2.2-120b is a transformation of Dolphin-2.2-70b created by interleaving the model with itself.
120b
⬇ 560K indirme · 19 varyant
· 01.2024
Uncensored Llama2 based model with support for a 16K context window.
13b
⬇ 556.6K indirme · 18 varyant
· 12.2023
Solar Pro Preview: an advanced large language model (LLM) with 22 billion parameters designed to fit into a single GPU
22b
⬇ 549.3K indirme · 18 varyant
· 09.2024
🎩 Magicoder is a family of 7B parameter models trained on 75K synthetic instruction data using OSS-Instruct, a novel approach to enlightening LLMs with open-source code snippets.
7b
⬇ 545.9K indirme · 18 varyant
· 12.2023
EXAONE 3.5 is a collection of instruction-tuned bilingual (English and Korean) generative models ranging from 2.4B to 32B parameters, developed and released by LG AI Research.
2.4b
7.8b
32b
⬇ 543K indirme · 13 varyant
· 12.2024