Braina Logo

AI Language Model Library

For information on how to run these models on your PC, please refer this guide : How to Run AI Language Models on Your PC Locally.


Qwen3.8

qwen3.8 📋 Copied!
Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks.
Category: Language
Downloads: 1.8M
Last Updated: 1 month ago
muse-glimmer 📋 Copied!
Meta's latest open model built for always-on local agents. 30B parameters, licensed under Apache 2.0 and runs on a single GPU — tuned for tool use, long tasks, and failure recovery.
Category: Language
Downloads: 200.1K
Last Updated: 3 weeks ago

Read more:
Muse Glimmer Evaluation Methodology
qwen3.8-flash-next 📋 Copied!
This experimental preview of the architecture that will underpin Qwen4.
Category: Language
Downloads: 96.3K
Last Updated: 3 weeks ago

Read more:
Qwen Blog
Hugging Face
License: Qwen Community License 1.0

Minicpm-v4.6

minicpm-v4.6 📋 Copied!
A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone
Category: Language
Downloads: 46.4K
Last Updated: 3 months ago

Gemma4

gemma4 📋 Copied!
Gemma 4 models are designed to deliver frontier-level performance at each size. They are well-suited for reasoning, agentic workflows, coding, and multimodal understanding.
Category: Language
Downloads: 24.7M
Last Updated: 3 weeks ago

Nomic-embed-text

nomic-embed-text 📋 Copied!
An efficient open embedding model featuring an extensive token context window.
Category: Tiny,embedding
Downloads: 85.4M
Last Updated: 2 years ago

Gemma3

gemma3 📋 Copied!
The current, most capable model that runs on a single GPU.
Category: Tiny,
Downloads: 40.3M
Last Updated: 1 year ago

Phi4-mini

phi4-mini 📋 Copied!
Phi-4-mini brings significant enhancements in multilingual support, reasoning, and mathematics, and now, the long-awaited function calling feature is finally supported.
Category: Language
Downloads: 1.4M
Last Updated: 1 year ago

Mxbai-embed-large

mxbai-embed-large 📋 Copied!
Cutting-edge large embedding model developed by mixedbread.ai.
Category: Tiny,embedding
Downloads: 14.5M
Last Updated: 2 years ago

Qwen3-coder-next

qwen3-coder-next 📋 Copied!
Qwen3-Coder-Next is a coding-focused language model from Alibaba's Qwen team, optimized for agentic coding workflows and local development.
Category: Language
Downloads: 2M
Last Updated: 7 months ago

Deepseek-v3

deepseek-v3 📋 Copied!
A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token.
Category: Language
Downloads: 3.8M
Last Updated: 1 year ago

Phi4

phi4 📋 Copied!
Phi-4 is a 14B parameter, state-of-the-art open model from Microsoft.
Category: Language
Downloads: 7.7M
Last Updated: 1 year ago

Nemotron

nemotron 📋 Copied!
NVIDIA has tailored the Llama-3.1-Nemotron-70B-Instruct model to enhance the quality of responses generated by large language models. This customization aims to increase the usefulness of LLM outputs for user inquiries.
Category: Language
Downloads: 619.1K
Last Updated: 1 year ago

Shieldgemma

shieldgemma 📋 Copied!
ShieldGemma consists of instructed models designed to assess the safety of text prompts and responses based on established safety guidelines.
Category: Language
Downloads: 935.9K
Last Updated: 1 year ago

Llama-guard3

llama-guard3 📋 Copied!
Llama Guard 3 consists of a set of models specifically refined for classifying the content safety of inputs and responses from large language models (LLMs). These models are designed to enhance the reliability of content interactions.
Category: Tiny,
Downloads: 1M
Last Updated: 1 year ago

All-MiniLM

all-minilm 📋 Copied!
Embedding models applied to extensive datasets at the sentence level.
Category: Tiny,embedding
Downloads: 3.5M
Last Updated: 2 years ago

Bge-large

bge-large 📋 Copied!
BAAI's embedding model converts texts into vectors.
Category: Tiny,embedding
Downloads: 284.5K
Last Updated: 2 years ago

Bge-m3

bge-m3 📋 Copied!
The BGE-M3, a recent model from BAAI, stands out for its versatility in multi-functionality, multilingual capabilities, and multi-granularity.
Category: Embedding
Downloads: 6.5M
Last Updated: 2 years ago

Snowflake

snowflake-arctic-embed 📋 Copied!
Snowflake offers a collection of performance-optimized text embedding models. These models are designed for enhanced efficiency.
Category: Tiny,embedding
Downloads: 3.1M
Last Updated: 2 years ago
mistral-large 📋 Copied!
Mistral Large 2 is the latest flagship model from Mistral, offering enhanced performance in code generation, mathematics, and reasoning, thanks to its 128k context window and support for numerous languages. Its capabilities are significantly improved compared to previous models.
Category: Language
Downloads: 1.3M
Last Updated: 1 year ago
Read more about: Mistral Large 2

Dolphin3

dolphin3 📋 Copied!
Dolphin 3.0 Llama 3.1 8B is the next generation of the Dolphin series of instruct-tuned models designed to be the ultimate general purpose local model, enabling coding, math, agentic, function calling, and general use cases.
Category: Uncensored
Downloads: 4.1M
Last Updated: 1 year ago

Smallthinker

smallthinker 📋 Copied!
A compact reasoning model has been fine-tuned from the Qwen 2.5 3B Instruct model.
Category: Language
Downloads: 259.1K
Last Updated: 1 year ago

Granite3.1-moe

granite3.1-moe 📋 Copied!
The IBM Granite 1B and 3B models are designed for low-latency applications and feature a long-context mixture of experts (MoE) architecture. These Granite models have been developed by IBM.
Category: Tiny,
Downloads: 3M
Last Updated: 1 year ago

Falcon3

falcon3 📋 Copied!
A group of effective AI models with fewer than 10 billion parameters excels in science, math, and coding thanks to advanced training methods. These models demonstrate high performance across these areas.
Category: Language
Downloads: 2.6M
Last Updated: 1 year ago

Granite-embedding

granite-embedding 📋 Copied!
The IBM Granite Embedding 30M and 278M models models are text-only dense biencoder embedding models, with 30M available in English only and 278M serving multilingual use cases.
Category: Tiny,embedding
Downloads: 366K
Last Updated: 1 year ago

Exaone3.5

exaone3.5 📋 Copied!
EXAONE 3.5 is a collection of instruction-tuned bilingual (English and Korean) generative models ranging from 2.4B to 32B parameters, developed and released by LG AI Research.
Category: Language
Downloads: 556.9K
Last Updated: 1 year ago

Llama3.3

llama3.3 📋 Copied!
New state of the art 70B model. Llama 3.3 70B offers similar performance compared to the Llama 3.1 405B model.
Category: Language
Downloads: 4.1M
Last Updated: 1 year ago

Snowflake-arctic-embed2

snowflake-arctic-embed2 📋 Copied!
Snowflake's frontier embedding model. Arctic Embed 2.0 adds multilingual support without sacrificing English performance or scalability.
Category: Embedding
Downloads: 450K
Last Updated: 1 year ago

Sailor2

sailor2 📋 Copied!
Sailor2 are multilingual language models made for South-East Asia. Available in 1B, 8B, and 20B parameter sizes.
Category: Tiny,
Downloads: 414.1K
Last Updated: 1 year ago

Qwq

qwq 📋 Copied!
QwQ is the reasoning model of the Qwen series.
Category: Language
Downloads: 2.3M
Last Updated: 1 year ago

Marco-o1

marco-o1 📋 Copied!
An open large reasoning model for real-world solutions by the Alibaba International Digital Commerce Group (AIDC-AI).
Category: Language
Downloads: 212.7K
Last Updated: 1 year ago

Tulu3

tulu3 📋 Copied!
Tülu 3 is a leading instruction following model family, offering fully open-source data, code, and recipes by the The Allen Institute for AI.
Category: Language
Downloads: 395.3K
Last Updated: 1 year ago

Athene-v2

athene-v2 📋 Copied!
Athene-V2 is a 72B parameter model which excels at code completion, mathematics, and log extraction tasks.
Category: Language
Downloads: 590.6K
Last Updated: 1 year ago

Opencoder

opencoder 📋 Copied!
OpenCoder is an open and reproducible code LLM family which includes 1.5B and 8B models, supporting chat in English and Chinese languages.
Category: Language
Downloads: 643.2K
Last Updated: 1 year ago

Ornith-1.5

ornith-1.5 📋 Copied!
Chirp Chirp! ? We are introducing Ornith-1.5, a major step toward building foundation models through end-to-end self-improvement.
Category: Language
Downloads: 305.7K
Last Updated: 1 month ago

Granite4.2

granite4.2 📋 Copied!
IBM Granite Models are a family of enterprise-ready, open foundation models that support multilingual capabilities, coding, retrieval-augmented generation (RAG), tool use, thinking and structured JSON output. Released under Apache 2.0 license.
Category: Language
Downloads: 49.9K
Last Updated: 1 month ago

Nemotron-3.5-lightning

nemotron-3.5-lightning 📋 Copied!
NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for always-on agents.
Category: Language
Downloads: 160.5K
Last Updated: 3 weeks ago

Laguna-s-2.1

laguna-s-2.1 📋 Copied!
Our most capable model to date, designed for long-horizon work. 70.2% on Terminal-Bench 2.1 at 118B-A8B.
Category: Language
Downloads: 135.7K
Last Updated: 3 weeks ago

Laguna-xs-2.1

laguna-xs-2.1 📋 Copied!
Laguna XS 2.1 is a 33B total parameter Mixture-of-Experts model with 3B activated parameters per token designed for agentic coding and long-horizon work on a local machine.
Category: Language
Downloads: 108.4K
Last Updated: 3 weeks ago

Ornith

ornith 📋 Copied!
A self-improving family of open-source models for agentic coding
Category: Language
Downloads: 479.3K
Last Updated: 2 months ago

North-mini-code-1.0

north-mini-code-1.0 📋 Copied!
North Mini Code is Cohere's first model for developers — a 30B Mixture-of-Experts model with 3B active parameters, built for agentic software engineering.
Category: Language
Downloads: 53K
Last Updated: 2 months ago

Lfm2.5

lfm2.5 📋 Copied!
LFM2.5-8B-A1B, an edge model built for fast, reliable tool calling on consumer hardware.
Category: Language
Downloads: 143.9K
Last Updated: 3 months ago

Minicpm-v4.5

minicpm-v4.5 📋 Copied!
A GPT-4o Level MLLM for Single Image, Multi Image and High-FPS Video Understanding on Your Phone
Category: Language
Downloads: 36.6K
Last Updated: 3 months ago

Granite4.1-guardian

granite4.1-guardian 📋 Copied!
Granite Guardian 4.1 is a specialized safety and judging model from IBM Research that evaluates whether LLM prompts and responses meet specified harm criteria.
Category: Language
Downloads: 15.4K
Last Updated: 3 months ago

Mistral-medium-3.5

mistral-medium-3.5 📋 Copied!
Mistral Medium 3.5 is the first flagship model of Mistral AI that merged instruction-following, reasoning, and coding in a single set of 128B weights.
Category: Language
Downloads: 385.1K
Last Updated: 4 months ago

Granite4.1

granite4.1 📋 Copied!
IBM Granite Models are a family of enterprise-ready, open foundation models that support multilingual capabilities, coding, retrieval-augmented generation (RAG), tool use, and structured JSON output. Released under Apache 2.0 license.
Category: Language
Downloads: 424.5K
Last Updated: 3 months ago

Nemotron3

nemotron3 📋 Copied!
NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows.
Category: Language
Downloads: 658.2K
Last Updated: 4 months ago

Deepseek-v4-flash

deepseek-v4-flash 📋 Copied!
DeepSeek-V4-Flash is the official release of DeepSeek-V4-Flash, built for efficient reasoning across a 1M-token context window, outperforming DeepSeek-V4-Pro (Preview).
Category: Language
Downloads: 446.7K
Last Updated: 1 month ago

Deepseek-v4-pro

deepseek-v4-pro 📋 Copied!
DeepSeek-V4-Pro is a frontier Mixture-of-Experts model with a large context window and three reasoning modes.
Category: Language
Downloads: 396.4K
Last Updated: 1 month ago

Laguna-xs.2

laguna-xs.2 📋 Copied!
Laguna XS.2 is a 33B total parameter Mixture-of-Experts model with 3B activated parameters per token designed for agentic coding and long-horizon work on a local machine.
Category: Language
Downloads: 29K
Last Updated: 1 month ago

Qwen3.6

qwen3.6 📋 Copied!
Qwen3.6 delivers substantial upgrades in agentic coding and thinking preservation than previous Qwen models.
Category: Language
Downloads: 6.5M
Last Updated: 3 weeks ago

Medgemma1.5

medgemma1.5 📋 Copied!
MedGemma 1.5 4B is an updated version of the MedGemma 4B model.
Category: Language
Downloads: 173.6K
Last Updated: 4 months ago

Medgemma

medgemma 📋 Copied!
MedGemma is a collection of Gemma 3 variants that are trained for performance on medical text and image comprehension.
Category: Language
Downloads: 368.9K
Last Updated: 4 months ago

Nemotron-cascade-2

nemotron-cascade-2 📋 Copied!
An open 30B MoE model from NVIDIA with 3B activated parameters that delivers strong reasoning and agentic capabilities.
Category: Language
Downloads: 146.6K
Last Updated: 5 months ago

Lfm2

lfm2 📋 Copied!
LFM2 is a family of hybrid models designed for on-device deployment. LFM2-24B-A2B is the largest model in the family, scaling the architecture to 24 billion parameters while keeping inference efficient.
Category: Language
Downloads: 1.1M
Last Updated: 6 months ago

Nemotron-3-super

nemotron-3-super 📋 Copied!
NVIDIA Nemotron 3 Super is a 120B open MoE model activating just 12B parameters to deliver maximum compute efficiency and accuracy for complex multi-agent applications.
Category: Language
Downloads: 2.9M
Last Updated: 6 months ago

Qwen3.5

qwen3.5 📋 Copied!
Qwen 3.5 is a family of open-source multimodal models that delivers exceptional utility and performance.
Category: Language
Downloads: 20.1M
Last Updated: 3 weeks ago

Glm-ocr

glm-ocr 📋 Copied!
GLM-OCR is a multimodal OCR model for complex document understanding, built on the GLM-V encoder–decoder architecture.
Category: Language
Downloads: 7.1M
Last Updated: 7 months ago

Lfm2.5-thinking

lfm2.5-thinking 📋 Copied!
LFM2.5 is a new family of hybrid models designed for on-device deployment.
Category: Tiny,
Downloads: 1.3M
Last Updated: 7 months ago

Glm-4.7-flash

glm-4.7-flash 📋 Copied!
As the strongest model in the 30B class, GLM-4.7-Flash offers a new option for lightweight deployment that balances performance and efficiency.
Category: Language
Downloads: 1.7M
Last Updated: 3 months ago

Translategemma

translategemma 📋 Copied!
A new collection of open translation models built on Gemma 3, helping people communicate across 55 languages.
Category: Language
Downloads: 2.3M
Last Updated: 7 months ago

Functiongemma

functiongemma 📋 Copied!
FunctionGemma is a specialized version of Google's Gemma 3 270M model fine-tuned explicitly for function calling.
Category: Tiny,
Downloads: 188.2K
Last Updated: 8 months ago

Olmo-3.1

olmo-3.1 📋 Copied!
Olmo is a series of Open language models designed to enable the science of language models. These models are pre-trained on the Dolma 3 dataset and post-trained on the Dolci datasets.
Category: Language
Downloads: 289.5K
Last Updated: 8 months ago

Nemotron-3-nano

nemotron-3-nano 📋 Copied!
Nemotron-3-Nano is a new Standard for Efficient, Open, and Intelligent Agentic Models, now updated with a 4B parameter count model.
Category: Language
Downloads: 822.6K
Last Updated: 5 months ago

Olmo-3

olmo-3 📋 Copied!
Olmo is a series of Open language models designed to enable the science of language models. These models are pre-trained on the Dolma 3 dataset and post-trained on the Dolci datasets.
Category: Language
Downloads: 456.6K
Last Updated: 8 months ago

Embeddinggemma

embeddinggemma 📋 Copied!
EmbeddingGemma is a 300M parameter embedding model from Google.
Category: Tiny,
Downloads: 2M
Last Updated: 1 year ago
gemma2 📋 Copied!
Google Gemma 2 is now offered in three sizes: 2B, 9B and 27B.
Category: Language
Downloads: 32.2M
Last Updated: 2 years ago
Read more about: Gemma 2

Deepseek-v3.1

deepseek-v3.1 📋 Copied!
DeepSeek-V3.1-Terminus is a hybrid model that supports both thinking mode and non-thinking mode.
Category: Language
Downloads: 726.7K
Last Updated: 11 months ago
llama3 📋 Copied!
Meta Llama 3 is the most advanced openly accessible language model available so far. It stands out as the leading choice among open-source LLMs.
Category: Language
Downloads: 25.2M
Last Updated: 2 years ago
Read more about: Llama 3

Gpt-oss

gpt-oss 📋 Copied!
OpenAI’s open-weight models designed for powerful reasoning, agentic tasks, and versatile developer use cases.
Category: Language
Downloads: 12.7M
Last Updated: 11 months ago
qwen2 📋 Copied!
Qwen2 is a newly launched series of extensive language models developed by the Alibaba Group.
Category: Tiny,
Downloads: 6.2M
Last Updated: 2 years ago
Read more about: Qwen2

Qwen3-coder

qwen3-coder 📋 Copied!
Alibaba's performant long context models for agentic and coding tasks.
Category: Language
Downloads: 9.1M
Last Updated: 11 months ago
deepseek-coder-v2 📋 Copied!
A code language model based on an open-source Mixture-of-Experts approach, it delivers performance similar to GPT-4 Turbo for code-related tasks.
Category: Coding
Downloads: 3.1M
Last Updated: 2 years ago
Read more about: DeepSeek-Coder-v2

Mistral-small3.2

mistral-small3.2 📋 Copied!
An update to Mistral Small that improves on function calling, instruction following, and less repetition errors.
Category: Language
Downloads: 2.5M
Last Updated: 1 year ago
phi3 📋 Copied!
Phi-3 consists of advanced, lightweight open models, including 3B (Mini) and 14B (Medium) versions, developed by Microsoft.
Category: Language
Downloads: 18.1M
Last Updated: 2 years ago
Read more about: Phi-3

Gemma3n

gemma3n 📋 Copied!
Gemma 3n models are designed for efficient execution on everyday devices such as laptops, tablets or phones.
Category: Language
Downloads: 2.2M
Last Updated: 1 year ago
aya 📋 Copied!
Cohere has launched Aya 23, a cutting-edge family of multilingual models that accommodate 23 different languages.
Category: Language
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: Aya 23

Magistral

magistral 📋 Copied!
Magistral is a small, efficient reasoning model with 24B parameters.
Category: Language
Downloads: 1.5M
Last Updated: 1 year ago
mistral 📋 Copied!
Mistral AI has released the 7B model, now updated to version 0.3.
Category: Language
Downloads: 33.4M
Last Updated: 1 year ago
Read more about: Mistral 7B

Devstral

devstral 📋 Copied!
Devstral: the best open source model for coding agents
Category: Language
Downloads: 1M
Last Updated: 1 year ago
mixtral 📋 Copied!
Mistral AI has released a set of Mixture of Experts (MoE) models featuring open weights, available in parameter sizes of 8x7b and 8x22b.
Category: Language
Downloads: 2.9M
Last Updated: 1 year ago
Read more about: Mistral 7B and 22B

Qwen2.5vl

qwen2.5vl 📋 Copied!
Flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL.
Category: Language
Downloads: 4.8M
Last Updated: 1 year ago
codegemma 📋 Copied!
CodeGemma comprises robust, lightweight models capable of various coding tasks, including code completion, generation, natural language understanding, mathematical reasoning, and instruction adherence. These models are designed to efficiently handle diverse programming challenges.
Category: Coding
Downloads: 3.2M
Last Updated: 2 years ago
Read more about: CodeGemma

Phi4-reasoning

phi4-reasoning 📋 Copied!
Phi 4 reasoning and reasoning plus are 14-billion parameter open-weight reasoning models that rival much larger models on complex reasoning tasks.
Category: Language
Downloads: 1.7M
Last Updated: 1 year ago
command-r 📋 Copied!
Command R is a Large Language Model designed for effective conversational interaction and handling extended context tasks. It is optimized to facilitate engaging dialogue and manage lengthy discussions.
Category: Language
Downloads: 1.5M
Last Updated: 2 years ago
Read more about: Command-R

Phi4-mini-reasoning

phi4-mini-reasoning 📋 Copied!
Phi 4 mini reasoning is a lightweight open model that balances efficiency with advanced reasoning ability.
Category: Language
Downloads: 309.9K
Last Updated: 1 year ago
command-r-plus 📋 Copied!
Command R+ is a robust and scalable large language model designed specifically for optimal performance in retrieval augmented generation (RAG) and real-world enterprise applications.
Category: Specialized
Downloads: 800.7K
Last Updated: 2 years ago
Read more about: Command R+

Qwen3

qwen3 📋 Copied!
Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models.
Category: Tiny,
Downloads: 36.8M
Last Updated: 11 months ago
llava 📋 Copied!
LLaVA is a novel large multimodal model that combines a vision encoder and Vicuna for general-purpose visual and language understanding. Updated to version 1.6.
Category: Multimodal
Downloads: 14.9M
Last Updated: 2 years ago
Read more about: LLaVA

Granite3.3

granite3.3 📋 Copied!
IBM Granite 2B and 8B models are 128K context length language models that have been fine-tuned for improved reasoning and instruction-following capabilities.
Category: Language
Downloads: 1.1M
Last Updated: 1 year ago
gemma 📋 Copied!
Gemma is a cutting-edge family of lightweight open models developed by Google DeepMind. It has recently been updated to version 1.1.
Category: Language
Downloads: 8.3M
Last Updated: 2 years ago
Read more about: Gemma

Deepcoder

deepcoder 📋 Copied!
DeepCoder is a fully open-Source 14B coder model at O3-mini level, with a 1.5B version also available.
Category: Language
Downloads: 950.8K
Last Updated: 1 year ago
qwen 📋 Copied!
Qwen 1.5 is a range of large language models developed by Alibaba Cloud, featuring parameter sizes from 0.5 billion to 110 billion.
Category: Tiny,
Downloads: 7.7M
Last Updated: 2 years ago
Read more about: Qwen 1.5

Mistral-small3.1

mistral-small3.1 📋 Copied!
Building upon Mistral Small 3, Mistral Small 3.1 (2503) adds state-of-the-art vision understanding and enhances long context capabilities up to 128k tokens without compromising text performance.
Category: Language
Downloads: 787.3K
Last Updated: 1 year ago
llama2 📋 Copied!
Llama 2 consists of a series of foundational language models with parameter sizes between 7 billion and 70 billion.
Category: Language
Downloads: 7.5M
Last Updated: 2 years ago
Read more about: Llama 2

Cogito

cogito 📋 Copied!
Cogito v1 Preview is a family of hybrid reasoning models by Deep Cogito that outperform the best available open models of the same size, including counterparts from LLaMA, DeepSeek, and Qwen across most standard benchmarks.
Category: Language
Downloads: 2.1M
Last Updated: 1 year ago
codellama 📋 Copied!
A powerful language model capable of generating and discussing code based on text prompts. It leverages natural language to facilitate coding tasks.
Category: Coding
Downloads: 6.1M
Last Updated: 2 years ago
Read more about: Codella

Llama4

llama4 📋 Copied!
Meta's latest collection of multimodal models.
Category: Language
Downloads: 1.8M
Last Updated: 1 year ago
dolphin-mixtral 📋 Copied!
Eric Hartford developed the uncensored 8x7b and 8x22b fine-tuned models using the Mixtral mixture of experts, which are particularly effective for coding tasks.
Category: Uncensored
Downloads: 1.9M
Last Updated: 1 year ago
Read more about: Dolphin Mixtral

Exaone-deep

exaone-deep 📋 Copied!
EXAONE Deep exhibits superior capabilities in various reasoning tasks including math and coding benchmarks, ranging from 2.4B to 32B parameters developed and released by LG AI Research.
Category: Language
Downloads: 766.5K
Last Updated: 1 year ago

Command-a

command-a 📋 Copied!
111 billion parameter model optimized for demanding enterprises that require fast, secure, and high-quality AI
Category: Language
Downloads: 227.8K
Last Updated: 1 year ago
llama2-uncensored 📋 Copied!
The Uncensored Llama 2 model was developed by George Sung and Jarrad Hope.
Category: Uncensored
Downloads: 2.7M
Last Updated: 2 years ago
Read more about: Llama 2 Uncensored
deepseek-coder 📋 Copied!
DeepSeek Coder is a powerful coding model that has been trained on two trillion tokens of code and natural language. Its extensive training allows it to perform well in various coding tasks.
Category: Tiny,coding
Downloads: 4.6M
Last Updated: 2 years ago
Read more about: DeepSeek-Coder

Command-r7b-arabic

command-r7b-arabic 📋 Copied!
A new state-of-the-art version of the lightweight Command R7B model that excels in advanced Arabic language capabilities for enterprises in the Middle East and Northern Africa.
Category: Language
Downloads: 203.8K
Last Updated: 1 year ago
phi 📋 Copied!
Phi-2 is a 2.7 billion parameter language model developed by Microsoft Research, showcasing exceptional reasoning and language comprehension skills.
Category: Language
Downloads: 1.5M
Last Updated: 2 years ago
Read more about: Phi-2

Qwen2.5-coder

qwen2.5-coder 📋 Copied!
The newest series of Code-Specific Qwen models features major enhancements in code generation, reasoning, and correction. These advancements lead to more effective coding solutions.
Category: Tiny,
Downloads: 21.3M
Last Updated: 1 year ago

Granite3.2-vision

granite3.2-vision 📋 Copied!
A compact and efficient vision-language model, specifically designed for visual document understanding, enabling automated content extraction from tables, charts, infographics, plots, diagrams, and more.
Category: Language
Downloads: 996.2K
Last Updated: 1 year ago
dolphin-mistral 📋 Copied!
The uncensored Dolphin model, which is built on Mistral, performs exceptionally well in coding tasks. It has been updated to version 2.8.
Category: Uncensored
Downloads: 1.7M
Last Updated: 2 years ago
Read more about: Dolphin Mistral

Solar-pro

solar-pro 📋 Copied!
Solar Pro Preview is an advanced large language model (LLM) featuring 22 billion parameters, optimized for deployment on a single GPU.
Category: Language
Downloads: 559.5K
Last Updated: 1 year ago
orca-mini 📋 Copied!
A versatile model with a parameter range of 3 billion to 70 billion, designed for use on entry-level hardware.
Category: Language
Downloads: 3M
Last Updated: 2 years ago
Read more about: Orca Mini

Nemotron-mini

nemotron-mini 📋 Copied!
NVIDIA has developed a small language model that is optimized for commercial use, particularly in roleplay, retrieval-augmented generation (RAG) question answering, and function calling. It is designed to be user-friendly for various applications.
Category: Language
Downloads: 711.4K
Last Updated: 1 year ago

Granite3.2

granite3.2 📋 Copied!
Granite-3.2 is a family of long-context AI models from IBM Granite fine-tuned for thinking capabilities.
Category: Language
Downloads: 456.3K
Last Updated: 1 year ago
dolphin-llama3 📋 Copied!
Dolphin 2.9, developed by Eric Hartford and built on Llama 3, comes in 8B and 70B sizes. It boasts a range of capabilities in instruction, conversation, and coding.
Category: Language
Downloads: 2.1M
Last Updated: 2 years ago
Read more about: Dolphin LLama3

Qwen2.5

qwen2.5 📋 Copied!
The Qwen2.5 models have been pretrained on Alibaba's latest extensive dataset, which includes up to 18 trillion tokens. This model can handle up to 128K tokens and offers multilingual support.
Category: Tiny,
Downloads: 39.9M
Last Updated: 1 year ago

R1-1776

r1-1776 📋 Copied!
A version of the DeepSeek-R1 model that has been post trained to provide unbiased, accurate, and factual information by Perplexity.
Category: Language
Downloads: 421.4K
Last Updated: 1 year ago

Bespoke-minicheck

bespoke-minicheck 📋 Copied!
Bespoke Labs has created an advanced fact-checking model. This state-of-the-art technology is at the forefront of verifying information.
Category: Language
Downloads: 530K
Last Updated: 1 year ago

Deepscaler

deepscaler 📋 Copied!
A fine-tuned version of Deepseek-R1-Distilled-Qwen-1.5B that surpasses the performance of OpenAI’s o1-preview with just 1.5B parameters on popular math evaluations.
Category: Language
Downloads: 1.3M
Last Updated: 1 year ago
mistral-openorca 📋 Copied!
Mistral OpenOrca is a 7 billion parameter model that has been fine-tuned from the Mistral 7B model with the OpenOrca dataset.
Category: Language
Downloads: 690.4K
Last Updated: 2 years ago
Read more about: Mistral OpenOrca

Mistral-small

mistral-small 📋 Copied!
Mistral Small is an efficient, lightweight model tailored for affordable applications such as translation and summarization. Its design prioritizes performance while minimizing costs.
Category: Language
Downloads: 3.1M
Last Updated: 1 year ago
starcoder2 📋 Copied!
StarCoder2 represents the next evolution of openly trained LLMs for code, available in three sizes: 3B, 7B, and 15B parameters.
Category: Coding
Downloads: 3M
Last Updated: 2 years ago
Read more about: StarCoder2

Reader-lm

reader-lm 📋 Copied!
A collection of models designed to transform HTML content into Markdown format, facilitating content conversion tasks. These models are beneficial for efficiently handling content changes.
Category: Tiny,
Downloads: 930.5K
Last Updated: 1 year ago

Openthinker

openthinker 📋 Copied!
A fully open-source family of reasoning models built using a dataset derived by distilling DeepSeek-R1.
Category: Language
Downloads: 1.2M
Last Updated: 1 year ago
zephyr 📋 Copied!
Zephyr comprises refined versions of the Mistral and Mixtral models, designed to serve as effective assistants. These models are specifically trained to enhance their helpfulness.
Category: Language
Downloads: 1.3M
Last Updated: 2 years ago
Read more about: Zephyr

MiniCPM-V

minicpm-v 📋 Copied!
A collection of multimodal large language models (MLLMs) created for understanding the relationship between vision and language. These models are specifically tailored for tasks that involve interpreting both visual and textual information.
Category: Multimodal
Downloads: 5.5M
Last Updated: 1 year ago

Deepseek-r1

deepseek-r1 📋 Copied!
DeepSeek-R1 is a family of open reasoning models with performance approaching that of leading models, such as O3 and Gemini 2.5 Pro.
Category: Language
Downloads: 92.6M
Last Updated: 1 year ago
yi 📋 Copied!
Yi 1.5 is an advanced bilingual language model that delivers exceptional performance. It excels in understanding and generating text in two languages.
Category: Language
Downloads: 1.4M
Last Updated: 2 years ago
Read more about: Yi 1.5

Deepseek-v2.5

deepseek-v2.5 📋 Copied!
An enhanced edition of DeekSeek-V2 that combines the general and coding capabilities of both DeepSeek-V2-Chat and DeepSeek-Coder-V2-Instruct. This upgraded version integrates features from both platforms.
Category: Language
Downloads: 289.1K
Last Updated: 1 year ago

Olmo2

olmo2 📋 Copied!
OLMo 2 is a new family of 7B and 13B models trained on up to 5T tokens. These models are on par with or better than equivalently sized fully open models, and competitive with open-weight models such as Llama 3.1 on English academic benchmarks.
Category: Language
Downloads: 3.8M
Last Updated: 1 year ago
llama2-chinese 📋 Copied!
A Llama 2-based model has been fine-tuned to enhance its ability to engage in Chinese dialogue. This adjustment aims to improve conversational performance in that language.
Category: Language
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: Llama 2 Chinese

Reflection

reflection 📋 Copied!
A highly effective model has been developed using a novel method known as Reflection-tuning, which enables a large language model to identify and rectify errors in its reasoning. This technique enhances the model's ability to improve its responses.
Category: Language
Downloads: 617.8K
Last Updated: 2 years ago

Command-r7b

command-r7b 📋 Copied!
The smallest model in Cohere's R series delivers top-tier speed, efficiency, and quality to build powerful AI applications on commodity GPUs and edge devices.
Category: Language
Downloads: 320.5K
Last Updated: 1 year ago
llava-llama3 📋 Copied!
A LLaVA model, fine-tuned from Llama 3 Instruct, has achieved improved scores across multiple benchmarks.
Category: Multimodal
Downloads: 2.3M
Last Updated: 2 years ago
Read more about: LLaVA Llama 3

Yi-coder

yi-coder 📋 Copied!
Yi-Coder is a collection of open-source programming language models that provide top-tier coding capabilities while using less than 10 billion parameters.
Category: Tiny,
Downloads: 1.1M
Last Updated: 1 year ago
vicuna 📋 Copied!
A general-purpose chat model built on Llama and Llama 2, offering context sizes ranging from 2K to 16K.
Category: Language
Downloads: 1.2M
Last Updated: 2 years ago
Read more about: Vicuna

Smollm2

smollm2 📋 Copied!
SmolLM2 is a series of compact language models offered in three sizes: 135M, 360M, and 1.7B parameters.
Category: Tiny,
Downloads: 4M
Last Updated: 1 year ago
nous-hermes2 📋 Copied!
Nous Research offers a robust family of models that excel in scientific discussions and coding tasks. These powerful models are designed to enhance productivity in various research applications.
Category: Language
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: Nous Hermes 2

Granite3-guardian

granite3-guardian 📋 Copied!
The IBM Granite Guardian 3.0 models 2B and 8B are created to identify risks in both prompts and responses. Their primary function is risk detection.
Category: Language
Downloads: 336.6K
Last Updated: 1 year ago
wizard-vicuna-uncensored 📋 Copied!
Wizard Vicuna Uncensored is a model with 7B, 13B, and 30B parameters, developed from Llama 2 by Eric Hartford. It is an uncensored version of the original model.
Category: Uncensored
Downloads: 1.3M
Last Updated: 2 years ago
Read more about: Wizard Vicuna Uncensored

Aya-expanse

aya-expanse 📋 Copied!
Cohere for AI has developed language models that excel in 23 diverse languages. These models are designed to deliver strong performance across all supported languages.
Category: Language
Downloads: 1M
Last Updated: 1 year ago
tinyllama 📋 Copied!
The TinyLlama SLM (small language model) is a compact AI language model with 1.1B parameters, designed for local on-premise inference on consumer grade hardware.
Category: Tiny,
Downloads: 5.6M
Last Updated: 2 years ago
Read more about: TinyLlama

Granite3-moe

granite3-moe 📋 Copied!
The IBM Granite 1B and 3B models are IBM's inaugural mixture of experts (MoE) Granite models, specifically created for low latency applications.
Category: Tiny,
Downloads: 955.7K
Last Updated: 1 year ago
codestral 📋 Copied!
Codestral is the inaugural code model from Mistral AI, specifically developed for tasks involving code generation.
Category: Coding
Downloads: 1.3M
Last Updated: 2 years ago
Read more about: Codestral

Granite3-dense

granite3-dense 📋 Copied!
The IBM Granite 2B and 8B models are tailored for tool-based applications, facilitating retrieval augmented generation (RAG) to enhance code generation, translation, and bug fixing. These features help streamline various programming tasks.
Category: Language
Downloads: 1M
Last Updated: 1 year ago
starcoder 📋 Copied!
StarCoder is a model designed for code generation, trained in over 80 programming languages.
Category: Tiny,coding
Downloads: 1.3M
Last Updated: 2 years ago
Read more about: StarCoder
wizardlm2 📋 Copied!
Microsoft AI has developed a cutting-edge large language model that enhances performance in complex chat, multilingual interactions, reasoning, and agent applications.
Category: Language
Downloads: 1.2M
Last Updated: 2 years ago
Read more about: WizardLM-2
openchat 📋 Copied!
An open-source model family trained on diverse data sets has outperformed ChatGPT on multiple benchmarks. It has been updated to version 3.5-0106.
Category: Language
Downloads: 1.3M
Last Updated: 2 years ago
Read more about: OpenChat
tinydolphin 📋 Copied!
Eric Hartford developed an experimental model with 1.1 billion parameters, utilizing the new Dolphin 2.8 dataset and based on TinyLlama.
Category: Tiny,
Downloads: 726.5K
Last Updated: 2 years ago
Read more about: TinyDolphin
openhermes 📋 Copied!
OpenHermes 2.5 is a 7B model that has been fine-tuned by Teknium using Mistral and entirely open datasets.
Category: Language
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: OpenHermes 2.5

Llama3.2

llama3.2 📋 Copied!
Meta's Llama 3.2 introduces smaller models with 1 billion and 3 billion parameters. These compact versions aim to enhance performance and accessibility.
Category: Tiny,
Downloads: 83M
Last Updated: 1 year ago
wizardcoder 📋 Copied!
Advanced code generation model.
Category: Coding
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: Wizard Coder
stable-code 📋 Copied!
Stable Code 3B is a coding model that offers instruct and code completion options comparable to larger models like Code Llama 7B, which is 2.5 times its size.
Category: Coding
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: Stable Code 3B
codeqwen 📋 Copied!
CodeQwen1.5 is a substantial language model that has been pretrained using extensive code datasets.
Category: Coding
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: CodeQwen 1.5
neural-chat 📋 Copied!
A well-tuned Mistral model that effectively covers both domain and language.
Category: Language
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: NeuralChat
wizard-math 📋 Copied!
The model specializes in solving mathematical and logical challenges. It emphasizes analytical reasoning and problem-solving skills.
Category: Specialized
Downloads: 1M
Last Updated: 2 years ago
Read more about: WizardMath
stablelm2 📋 Copied!
Stable LM 2 is an advanced language model with 1.6 billion and 12 billion parameters, designed using multilingual data in several languages including English, Spanish, German, Italian, French, Portuguese, and Dutch.
Category: Tiny,
Downloads: 1.1M
Last Updated: 2 years ago
Read more about: Stable LM 2
phind-codellama 📋 Copied!
A code generation model utilizing Code Llama.
Category: Coding
Downloads: 978.6K
Last Updated: 2 years ago
Read more about: Phind CodeLlama
granite-code 📋 Copied!
IBM has developed a family of open foundation models designed for code intelligence.
Category: Coding
Downloads: 1.5M
Last Updated: 2 years ago
Read more about: Granite Code
dolphincoder 📋 Copied!
A 7B and 15B uncensored version of the Dolphin model family, which is highly effective at coding, is derived from StarCoder2.
Category: Uncensored,coding
Downloads: 1M
Last Updated: 2 years ago
Read more about: Dolphin Coder
nous-hermes 📋 Copied!
Nous Research has developed general use models based on Llama and Llama 2.
Category: Language
Downloads: 1.2M
Last Updated: 2 years ago
Read more about: Nous Hermes
sqlcoder 📋 Copied!
SQLCoder is a code completion model specifically optimized for SQL generation tasks, built upon StarCoder. It offers enhanced capabilities for coding assistance in SQL.
Category: Coding,specialized
Downloads: 1.7M
Last Updated: 2 years ago
Read more about: SQLCoder
llama3-gradient 📋 Copied!
This model enhances the context length of LLama-3 8B from 8,000 to more than 1,000,000 tokens.
Category: Language
Downloads: 1M
Last Updated: 2 years ago
Read more about: Llama 3 Gradient
starling-lm 📋 Copied!
Starling is a substantial language model developed through reinforcement learning based on feedback from AI, aimed at enhancing the helpfulness of chatbots.
Category: Language
Downloads: 986.3K
Last Updated: 2 years ago
Read more about: Starling
deepseek-llm 📋 Copied!
A sophisticated language model built with 2 trillion bilingual tokens.
Category: Language
Downloads: 1.2M
Last Updated: 2 years ago
Read more about: DeepSeek
yarn-llama2 📋 Copied!
A version of Llama 2 that allows for a context of up to 128k tokens. This extension enhances its capacity for processing larger amounts of information.
Category: Language
Downloads: 983.5K
Last Updated: 2 years ago
Read more about: Yarn Llama 2

Phi3.5

phi3.5 📋 Copied!
A lightweight AI model consisting of 3.8 billion parameters surpasses the performance of similarly sized and larger models.
Category: Language
Downloads: 1M
Last Updated: 2 years ago
xwinlm 📋 Copied!
A conversational model built on Llama 2 demonstrates strong performance across multiple benchmarks. It competes effectively with others in the field.
Category: Language
Downloads: 1M
Last Updated: 2 years ago
Read more about: Xwin-LM

Smollm

smollm 📋 Copied!
? A family of small models with 135M, 360M, and 1.7B parameters, trained on a new high-quality dataset.
Category: Tiny,
Downloads: 2.1M
Last Updated: 2 years ago
llama3-chatqa 📋 Copied!
NVIDIA has developed a model leveraging Llama 3, which is particularly effective at conversational question answering (QA) and retrieval-augmented generation (RAG). It excels in these tasks, making it a powerful tool for generating responses and retrieving information.
Category: Specialized
Downloads: 1M
Last Updated: 2 years ago
Read more about: LLama 3 ChatQA-1.5

Paraphrase-multilingual

paraphrase-multilingual 📋 Copied!
The sentence-transformers model is suitable for tasks such as clustering and semantic search. It can efficiently process and understand textual data for these applications.
Category: Tiny,
Downloads: 940.1K
Last Updated: 2 years ago
orca2 📋 Copied!
Orca 2, developed by Microsoft Research, is an optimized version of Meta's Llama 2 models. It is specifically crafted to excel in reasoning tasks.
Category: Language
Downloads: 938.5K
Last Updated: 2 years ago
Read more about: Orca 2
wizardlm 📋 Copied!
A general-purpose model that utilizes Llama 2.
Category: Language
Downloads: 914.8K
Last Updated: 2 years ago
Read more about: WizardLM
samantha-mistral 📋 Copied!
An assistant specializing in philosophy, psychology, and personal relationships. Built on the Mistral model.
Category: Specialized
Downloads: 1M
Last Updated: 2 years ago
Read more about: Samantha Mistral
dolphin-phi 📋 Copied!
The uncensored Dolphin model, developed by Eric Hartford, is a 2.7 billion parameter model based on Microsoft's Phi language model.
Category: Uncensored
Downloads: 1.6M
Last Updated: 2 years ago
Read more about: Dolphin Phi
stable-beluga 📋 Copied!
A Llama 2-based model has been fine-tuned using an Orca-style dataset, originally named Free Willy.
Category: Language
Downloads: 945.4K
Last Updated: 2 years ago
Read more about: Stable Beluga
bakllava 📋 Copied!
BakLLaVA is a multimodal model that combines the Mistral 7B base model with the LLaVA architecture.
Category: Multimodal
Downloads: 887.7K
Last Updated: 2 years ago
Read more about: BakLLaVA
wizardlm-uncensored 📋 Copied!
Unedited version of the Wizard LM model.
Category: Uncensored
Downloads: 645.5K
Last Updated: 2 years ago
Read more about: WizardLM Uncensored
medllama2 📋 Copied!
The Llama 2 model has been fine-tuned to respond to medical queries using an open-source medical dataset. This adaptation enhances its effectiveness in addressing health-related inquiries.
Category: Specialized
Downloads: 702.7K
Last Updated: 2 years ago
Read more about: MedLlama2
yarn-mistral 📋 Copied!
A version of Mistral has been expanded to accommodate context windows of 64K or 128K.
Category: Language
Downloads: 903.6K
Last Updated: 2 years ago
Read more about: Yarn Mistral
nous-hermes2-mixtral 📋 Copied!
The Nous Hermes 2 model by Nous Research has now been trained on Mixtral.
Category: Language
Downloads: 591.6K
Last Updated: 1 year ago
Read more about: Nous Hermes 2 Mixtral
llama-pro 📋 Copied!
An extended version of Llama 2 designed to combine general language comprehension with specialized knowledge, especially in programming and mathematics.
Category: Language
Downloads: 913.3K
Last Updated: 2 years ago
Read more about: LLaMa-Pro
deepseek-v2 📋 Copied!
An efficient and cost-effective Mixture-of-Experts language model that demonstrates strength in performance.
Category: Language
Downloads: 1.2M
Last Updated: 2 years ago
Read more about: DeepSeek-V2
meditron 📋 Copied!
An open-source medical large language model, tailored from Llama 2 for healthcare applications.
Category: Specialized
Downloads: 869.8K
Last Updated: 2 years ago
Read more about: Meditron
codeup 📋 Copied!
An excellent code generation model built on Llama2.
Category: Coding
Downloads: 596K
Last Updated: 2 years ago
Read more about: CodeUp
nexusraven 📋 Copied!
Nexus Raven is a 13B model specifically designed for function calling tasks. It has been fine-tuned with a focus on instruction execution.
Category: Specialized
Downloads: 898.5K
Last Updated: 2 years ago
Read more about: Nexus Raven 13B
everythinglm 📋 Copied!
An uncensored model based on Llama2 that offers support for a 16K context window.
Category: Uncensored
Downloads: 566.7K
Last Updated: 2 years ago
Read more about: Everything LM
llava-phi3 📋 Copied!
A compact LLaVA model has been fine-tuned based on Phi 3 Mini.
Category: Multimodal
Downloads: 333.8K
Last Updated: 2 years ago
Read more about: LLaVa Phi-3
magicoder 📋 Copied!
? Magicoder is a family of 7B parameter models trained on 75K synthetic instruction data using OSS-Instruct, a novel approach to enlightening LLMs with open-source code snippets.
Category: Language
Downloads: 555.8K
Last Updated: 2 years ago
Read more about: Magicoder
codebooga 📋 Copied!
A highly effective code instruction model has been developed by combining two existing code models.
Category: Coding
Downloads: 501K
Last Updated: 2 years ago
Read more about: Codebooga
mistrallite 📋 Copied!
MistralLite is a refined model derived from Mistral, designed to improve the handling of lengthy contexts. Its advanced capabilities enable better processing for extended information.
Category: Language
Downloads: 524.4K
Last Updated: 2 years ago
Read more about: Mistrallite
wizard-vicuna 📋 Copied!
Wizard Vicuna is a 13 billion parameter model built on Llama 2, developed by MelodysDreamj.
Category: Language
Downloads: 518.4K
Last Updated: 2 years ago
Read more about: Wizard Vicuna
glm4 📋 Copied!
An impressive multi-lingual general language model that competes well with Llama 3. Its performance is notably strong.
Category: Language
Downloads: 1.2M
Last Updated: 2 years ago
Read more about: GLM4
duckdb-nsql 📋 Copied!
The 7B parameter text-to-SQL model was developed by MotherDuck and Numbers Station.
Category: Specialized
Downloads: 528.2K
Last Updated: 2 years ago
Read more about: DuckDB-NSQL
falcon2 📋 Copied!
Falcon2 is a decoder-only model with 11 billion parameters, developed by TII and trained on 5 trillion tokens.
Category: Language
Downloads: 541.2K
Last Updated: 2 years ago
Read more about: Falcon2
dbrx 📋 Copied!
DBRX is a versatile, open-source large language model developed by Databricks. It is designed for a wide range of applications.
Category: Language
Downloads: 366.9K
Last Updated: 2 years ago
Read more about: DBRX
codegeex4 📋 Copied!
A flexible model for AI software development applications, such as code completion. It adapts to various use cases effectively.
Category: Coding
Downloads: 691.5K
Last Updated: 2 years ago
Read more about: Codegeex4
internlm2 📋 Copied!
InternLM2.5 is a 7B parameter model designed for real-world applications, showcasing exceptional reasoning abilities.
Category: Language
Downloads: 1M
Last Updated: 2 years ago
Read more about: InternLM2.5
mathstral 📋 Copied!
MathΣtral is a 7 billion parameter model developed by Mistral AI, aimed at enhancing mathematical reasoning and facilitating scientific discoveries.
Category: Specialized
Downloads: 550.1K
Last Updated: 2 years ago
Read more about: Mathstral
llama3-groq-tool-use 📋 Copied!
Groq has developed a series of models that significantly enhance open-source AI capabilities for tool usage and function calling. These advancements represent a major step forward in the field.
Category: Specialized
Downloads: 999.8K
Last Updated: 2 years ago
Read more about: Llama3-groq-tool-use
firefunction-v2 📋 Copied!
A function-calling model utilizing open weights, built on Llama 3, offers capabilities that are competitive with GPT-4's function calling.
Category: Language
Downloads: 524.6K
Last Updated: 2 years ago
Read more about: Firefunction-v2
nuextract 📋 Copied!
A 3.8 billion parameter model has been fine-tuned on a private, high-quality synthetic dataset for the purpose of information extraction, utilizing Phi-3 as its foundation.
Category: Language
Downloads: 547K
Last Updated: 2 years ago
Read more about: NuExtract
mistral-nemo 📋 Copied!
A cutting-edge 12B model featuring a 128k context length, developed by Mistral AI in partnership with NVIDIA.
Category: Language
Downloads: 5.7M
Last Updated: 1 year ago
Read more about: Mistral NeMo
llama3.1 📋 Copied!
Meta has released Llama 3.1, a cutting-edge model offered in parameter sizes of 8B, 70B, and 405B.
Category: Language
Downloads: 119.3M
Last Updated: 1 year ago
Read more about: Llama 3.1

Qwen2-math

qwen2-math 📋 Copied!
Qwen2 Math consists of a collection of specialized math language models based on the Qwen2 LLMs. These models greatly surpass the mathematical performance of both open-source and even some closed-source models, such as GPT4o.
Category: Tiny,
Downloads: 1.1M
Last Updated: 2 years ago

Hermes3

hermes3 📋 Copied!
Hermes 3 is the newest iteration of Nous Research's flagship series of LLMs. It continues the legacy of the Hermes series.
Category: Language
Downloads: 1.6M
Last Updated: 1 year ago

Granite3.1-dense

granite3.1-dense 📋 Copied!
The IBM Granite 2B and 8B models are dense LLMs that focus solely on text and have been trained on more than 12 trillion tokens. In initial testing by IBM, they showed notable enhancements in both performance and speed compared to earlier models.
Category: Tiny,
Downloads: 1M
Last Updated: 1 year ago

Llama3.2-vision

llama3.2-vision 📋 Copied!
Llama 3.2 Vision is a collection of instruction-tuned image reasoning generative models in 11B and 90B sizes.
Category: Language
Downloads: 5.2M
Last Updated: 1 year ago

Nomic-embed-text-v2-moe

nomic-embed-text-v2-moe 📋 Copied!
nomic-embed-text-v2-moe is a multilingual MoE text embedding model that excels at multilingual retrieval.
Category: Tiny,
Downloads: 877.8K
Last Updated: 9 months ago

Rnj-1

rnj-1 📋 Copied!
Rnj-1 is a family of 8B parameter open-weight, dense models trained from scratch by Essential AI, optimized for code and STEM with capabilities on par with SOTA open-weight models.
Category: Language
Downloads: 507.4K
Last Updated: 9 months ago

Devstral-small-2

devstral-small-2 📋 Copied!
24B model that excels at using tools to explore codebases, editing multiple files and power software engineering agents.
Category: Language
Downloads: 1.1M
Last Updated: 9 months ago

Devstral-2

devstral-2 📋 Copied!
123B model that excels at using tools to explore codebases, editing multiple files and power software engineering agents.
Category: Language
Downloads: 355.1K
Last Updated: 9 months ago

Ministral-3

ministral-3 📋 Copied!
The Ministral 3 family is designed for edge deployment, capable of running on a wide range of hardware.
Category: Language
Downloads: 1.4M
Last Updated: 9 months ago

Mistral-large-3

mistral-large-3 📋 Copied!
A general-purpose multimodal mixture-of-experts model for production-grade tasks and enterprise workloads.
Category: Language
Downloads: 106.6K
Last Updated: 9 months ago

Deepseek-ocr

deepseek-ocr 📋 Copied!
DeepSeek-OCR is a vision-language model that can perform token-efficient OCR.
Category: Language
Downloads: 530.3K
Last Updated: 9 months ago

Cogito-2.1

cogito-2.1 📋 Copied!
The Cogito v2.1 LLMs are instruction tuned generative models. All models are released under MIT license for commercial use.
Category: Language
Downloads: 222.5K
Last Updated: 9 months ago

Gpt-oss-safeguard

gpt-oss-safeguard 📋 Copied!
gpt-oss-safeguard-20b and gpt-oss-safeguard-120b are safety reasoning models built-upon gpt-oss
Category: Language
Downloads: 155.4K
Last Updated: 10 months ago

Granite4

granite4 📋 Copied!
Granite 4 features improved instruction following (IF) and tool-calling capabilities, making them more effective in enterprise applications.
Category: Tiny,
Downloads: 1.5M
Last Updated: 10 months ago

Qwen3-embedding

qwen3-embedding 📋 Copied!
Building upon the foundational models of the Qwen3 series, Qwen3 Embedding provides a comprehensive range of text embeddings models in various sizes
Category: Tiny,
Downloads: 3.9M
Last Updated: 11 months ago

Qwen3-vl

qwen3-vl 📋 Copied!
The most powerful vision-language model in the Qwen model family to date.
Category: Language
Downloads: 6M
Last Updated: 10 months ago

Qwen3-next

qwen3-next 📋 Copied!
The first installment in the Qwen3-Next series with strong performance in terms of both parameter efficiency and inference speed.
Category: Language
Downloads: 587.3K
Last Updated: 9 months ago