The open-source model observatory

50 curated models. Live stats. Honest cards. No hype.

Modality
License
Hardware

50 models

all-MiniLM-L6-v2

Sentence-Transformers · MiniLM

Permissive
22MEmbedding

The default lightweight embedding baseline for years: tiny, extremely fast, and good enough for most semantic search prototypes.

— downloads·— likes

via Hugging Face

Bark

Suno · Bark

Permissive
~1B (combined sub-models)Audio

A fully generative text-to-audio model that can produce speech, music, and sound effects, but with less voice-cloning control than dedicated TTS models.

— downloads·— likes

via Hugging Face

BGE-M3

BAAI · BGE

Permissive
568MEmbedding

A versatile embedding model that unifies dense, sparse, and multi-vector (ColBERT-style) retrieval in one checkpoint across 100+ languages.

— downloads·— likes

via Hugging Face

Command R (08-2024)

Cohere · Command R

Restricted
32BText

A retrieval-augmented-generation-focused model with strong citation and tool-use behavior, but licensed for research and non-commercial use only.

— downloads·— likes

via Hugging Face

Command R+ (08-2024)

Cohere · Command R

Restricted
104BText

Cohere's larger RAG/tool-use flagship, strong at grounded enterprise-style workflows but explicitly non-commercial without a separate license.

— downloads·— likes

via Hugging Face

DeepSeek-R1

DeepSeek · DeepSeek

Permissive
671B total, 37B active (MoE)Text

The reasoning-focused sibling of DeepSeek-V3, released under a fully permissive MIT license, that triggered a wave of open reasoning-model distillation.

— downloads·— likes

via Hugging Face

DeepSeek-R1-Distill-Qwen-32B

DeepSeek · DeepSeek

Permissive
32BText

A Qwen2.5-32B base fine-tuned on DeepSeek-R1 reasoning traces, bringing most of R1's reasoning gains to a size that fits on a single GPU.

— downloads·— likes

via Hugging Face

DeepSeek-V3

DeepSeek · DeepSeek

Conditional
671B total, 37B active (MoE)Text

A massive sparse MoE that matched or beat leading closed models on many benchmarks at a fraction of the reported training cost.

— downloads·— likes

via Hugging Face

E5-Mistral-7B-Instruct

Microsoft · E5

Permissive
7BEmbedding

A large decoder-based embedding model that topped the MTEB leaderboard at release, trading model size for retrieval quality.

— downloads·— likes

via Hugging Face

FLUX.1 [dev]

Black Forest Labs · FLUX.1

Restricted
12BImage

Widely considered the best open-weight text-to-image model for prompt adherence and detail as of its release, but explicitly non-commercial.

— downloads·— likes

via Hugging Face

FLUX.1 [schnell]

Black Forest Labs · FLUX.1

Permissive
12BImage

A distilled, few-step FLUX.1 variant that is both fast enough for near-real-time generation and fully Apache-licensed for commercial use.

— downloads·— likes

via Hugging Face

Gemma 2 27B

Google · Gemma

Conditional
27BText

The largest Gemma 2 release, competitive with models roughly twice its size on several benchmarks at launch.

— downloads·— likes

via Hugging Face

Gemma 2 2B

Google · Gemma

Conditional
2BText

A very small Google model distilled from larger Gemma teachers, aimed at on-device use where every gigabyte of VRAM matters.

— downloads·— likes

via Hugging Face

Gemma 2 9B

Google · Gemma

Conditional
9BText

A well-rounded mid-size Gemma release that was among the strongest open 9B models at launch, distilled from a larger internal teacher.

— downloads·— likes

via Hugging Face

Idefics3 8B Llama3

Hugging Face · Idefics

Permissive
8BMultimodal

A fully open vision-language model built on Llama 3 that improved substantially on document/OCR tasks over its Idefics2 predecessor.

— downloads·— likes

via Hugging Face

Conditional
405BText

Meta's first frontier-scale open-weight model, competitive with closed frontier models at release but out of reach for anyone without a GPU cluster.

— downloads·— likes

via Hugging Face

Conditional
70BText

A mid-2024 workhorse dense model that still holds up for general reasoning and coding when 405B-class compute is not available.

— downloads·— likes

via Hugging Face

Conditional
8BText

The default pick for a local 8B chat model: strong instruction following for its size, easy to quantize and run on a single consumer GPU.

— downloads·— likes

via Hugging Face

Conditional
3BText

A small edge/on-device model that trades raw capability for speed and a tiny memory footprint, fine for summarization and simple assistants.

— downloads·— likes

via Hugging Face

Conditional
70BText

A late-2024 refresh that gets close to Llama 3.1 405B quality on many benchmarks at a fraction of the compute cost.

— downloads·— likes

via Hugging Face

LLaVA-1.6 Mistral 7B

LLaVA community · LLaVA

Permissive
7BMultimodal

One of the earliest and still widely-used open vision-language models, combining a CLIP vision encoder with a Mistral 7B backbone.

— downloads·— likes

via Hugging Face

Mistral 7B Instruct v0.3

Mistral AI · Mistral

Permissive
7BText

The classic fully-permissive 7B baseline; not the strongest 7B anymore but has no usage restrictions at all.

— downloads·— likes

via Hugging Face

Mistral Nemo 12B Instruct 2407

Mistral AI / NVIDIA · Mistral

Permissive
12BText

A joint Mistral/NVIDIA release that packs a 128K context window and a new Tekken tokenizer into a single-consumer-GPU-friendly 12B.

— downloads·— likes

via Hugging Face

Mistral Small 24B Instruct 2501

Mistral AI · Mistral

Permissive
24BText

A dense 24B tuned to fit on a single high-end GPU, aimed squarely at teams that want Apache-licensed weights without MoE serving complexity.

— downloads·— likes

via Hugging Face

Mixtral 8x22B Instruct v0.1

Mistral AI · Mistral

Permissive
8x22B (141B total, ~39B active)Text

A large permissively-licensed MoE that competes with dense 70B+ models while activating a fraction of its total parameters per token.

— downloads·— likes

via Hugging Face

Mixtral 8x7B Instruct v0.1

Mistral AI · Mistral

Permissive
8x7B (46.7B total, ~12.9B active)Text

The model that put sparse mixture-of-experts on the open-weight map: 70B-class quality at 13B-class inference cost, no license strings attached.

— downloads·— likes

via Hugging Face

MusicGen Large

Meta · MusicGen

Restricted
3.3BAudio

A strong text/melody-to-music generator from Meta's AudioCraft project, but licensed for research and non-commercial use only.

— downloads·— likes

via Hugging Face

Nomic Embed Text v1.5

Nomic AI · Nomic Embed

Permissive
137MEmbedding

A fully open embedding model, including training data and code, with Matryoshka representation learning for variable-size, resizable embeddings.

— downloads·— likes

via Hugging Face

OLMo 2 13B Instruct

Allen Institute for AI · OLMo

Permissive
13BText

The larger fully-open OLMo 2 release, giving researchers a reproducible 13B baseline with published data and training code.

— downloads·— likes

via Hugging Face

OLMo 2 7B Instruct

Allen Institute for AI · OLMo

Permissive
7BText

A fully open model in the truest sense: weights, training data, and training code are all published, not just the checkpoint.

— downloads·— likes

via Hugging Face

Phi-3.5 Mini Instruct

Microsoft · Phi

Permissive
3.8BText

A tiny MIT-licensed model built for on-device and edge use, with a 128K context window that is unusually large for its size class.

— downloads·— likes

via Hugging Face

Phi-4

Microsoft · Phi

Permissive
14BText

A data-quality-first small model that punches far above its 14B size on math and reasoning benchmarks, fully MIT licensed.

— downloads·— likes

via Hugging Face

PixArt-Sigma XL 2 1024MS

PixArt / Huawei Noah's Ark Lab · PixArt

Conditional
0.6BImage

A remarkably small diffusion transformer that reaches 4K-capable text-to-image quality at a training cost far below SDXL-class models.

— downloads·— likes

via Hugging Face

Qwen2-VL 72B Instruct

Alibaba · Qwen

Conditional
72BMultimodal

The flagship Qwen2-VL tier, competitive with closed multimodal models on visual reasoning, document understanding, and video benchmarks.

— downloads·— likes

via Hugging Face

Qwen2-VL 7B Instruct

Alibaba · Qwen

Permissive
7BMultimodal

A vision-language model that handles arbitrary image resolutions and understands video, strong on document/OCR-style tasks for its size.

— downloads·— likes

via Hugging Face

Qwen2.5 14B Instruct

Alibaba · Qwen

Permissive
14BText

A sweet-spot size for local deployment: noticeably stronger reasoning than 7B while still quantizing down to a single consumer GPU.

— downloads·— likes

via Hugging Face

Qwen2.5 32B Instruct

Alibaba · Qwen

Permissive
32BText

A strong permissively-licensed mid-size model that often trades blows with much larger dense models on reasoning benchmarks.

— downloads·— likes

via Hugging Face

Qwen2.5 72B Instruct

Alibaba · Qwen

Conditional
72BText

Alibaba's flagship dense open model at 72B, competitive with Llama 3.1 405B on many benchmarks despite far fewer parameters.

— downloads·— likes

via Hugging Face

Qwen2.5 7B Instruct

Alibaba · Qwen

Permissive
7BText

One of the strongest fully-permissive 7B models available, with unusually good math and coding for its size.

— downloads·— likes

via Hugging Face

Permissive
32BText

Widely regarded as the strongest fully open code model at its size, often cited as close to GPT-4-class coding performance.

— downloads·— likes

via Hugging Face

Permissive
7BText

A code-specialized 7B that punches well above its weight on code generation and repair benchmarks, and is fully permissive.

— downloads·— likes

via Hugging Face

QwQ 32B

Alibaba · Qwen

Permissive
32BText

A reasoning-tuned model that produces long chain-of-thought traces before answering, aimed at math and logic tasks rather than general chat.

— downloads·— likes

via Hugging Face

SmolLM2 1.7B Instruct

Hugging Face · SmolLM2

Permissive
1.7BText

A carefully data-curated small model that outperforms most other sub-2B models on reasoning and knowledge benchmarks.

— downloads·— likes

via Hugging Face

SmolLM2 360M Instruct

Hugging Face · SmolLM2

Permissive
360MText

A sub-billion-parameter model small enough to run in a browser via WebGPU, useful for constrained on-device tasks rather than general chat.

— downloads·— likes

via Hugging Face

Stable Diffusion 3.5 Large

Stability AI · Stable Diffusion

Conditional
8.1BImage

Stability's 2024 flagship diffusion transformer, a major quality jump over SDXL with much better text rendering and prompt adherence.

— downloads·— likes

via Hugging Face

Stable Diffusion 3.5 Medium

Stability AI · Stable Diffusion

Conditional
2.5BImage

A smaller SD3.5 variant tuned to run on more modest consumer hardware while keeping most of the architecture's quality gains.

— downloads·— likes

via Hugging Face

Stable Diffusion XL Base 1.0

Stability AI · Stable Diffusion

Conditional
3.5BImage

Still the most widely-deployed open text-to-image base model, with by far the largest ecosystem of fine-tunes, LoRAs, and ControlNets.

— downloads·— likes

via Hugging Face

Whisper Large v3

OpenAI · Whisper

Permissive
1.55BAudio

The de facto open standard for speech recognition, with strong multilingual accuracy and no usage restrictions.

— downloads·— likes

via Hugging Face

Whisper Large v3 Turbo

OpenAI · Whisper

Permissive
809MAudio

A pruned-decoder version of Whisper large-v3 that runs several times faster with only a small accuracy tradeoff, good for real-time use.

— downloads·— likes

via Hugging Face

XTTS v2

Coqui · XTTS

Restricted
466MAudio

A widely-used voice-cloning TTS model that can clone a voice from just a few seconds of audio across 17 languages, but non-commercial by default.

— downloads·— likes

via Hugging Face