LLM Releases

All tracked models

Last updated Jun 18, 2026

LLM model catalog

A searchable catalog of large language model releases, with source links, lifecycle dates, access details, model size, context window, and modality filters.

229
Models
40
Labs
160
Open
4
Recent

229 models

Kimi K2.7 Code

Available
Moonshot AIFrontierOpen weights

Moonshot's open coding-focused agentic model built on K2.6, with native vision/video input, forced thinking mode, and stronger long-horizon software-engineering performance.

MoE1T262K ctxJun 18, 2026

GLM-5.2

Available
Z.ai (Zhipu AI)FrontierOpen source

Z.ai's latest open flagship for long-horizon coding, agentic engineering, and million-token workflows, adding IndexShare sparse-attention reuse over GLM-5.1.

MoE753B1M ctxJun 17, 2026

MiniMax-M3

Available
MiniMaxFrontierOpen weights

Native multimodal MiniMax model with a one-million-token context, sparse attention, and agentic coding/cowork positioning.

MoE428B1M ctxJun 16, 2026

GPT-5.6

Preview
OpenAIFrontierProprietary

OpenAI's mid-2026 flagship, headlined by an industry-leading 1.5M-token context window and long-horizon agentic tool use.

MoEUndisc.1.5M ctxJun 9, 2026

Claude Fable 5

Withdrawn
AnthropicFrontierProprietary

The public, guardrailed sibling of Mythos and Anthropic's most capable widely-released model, built for long-horizon agentic work. Launched June 9, 2026 across the Claude API, AWS, and Microsoft Foundry — then pulled three days later under a US government export-control directive barring access by foreign nationals.

Undisc. ctxJun 9, 2026

Nemotron 3 Ultra 550B-A55B

Available
NVIDIAFrontierOpen weights

NVIDIA's largest Nemotron 3 open-weight hybrid Mamba-Transformer MoE, tuned for agentic reasoning, coding, planning, and tool calling.

Hybrid550B1M ctxJun 4, 2026

Claude Opus 4.8

Available
AnthropicFrontierProprietary

Anthropic's most capable model, with strengthened agentic and long-running task performance.

Undisc.500K ctxMay 28, 2026

MiniMax-M2.7

Available
MiniMaxFrontierOpen weights

Open-weight agentic model from MiniMax focused on real-world software engineering, office tasks, tool use, and self-improving training workflows.

MoE229.9B ctxMay 26, 2026

Gemini 3.5 Pro

Preview
Google DeepMindFrontierProprietary

Announced at Google I/O 2026; emphasizes deep multimodal reasoning over a 2M-token context.

MoEUndisc.2M ctxMay 19, 2026

Qwen3.6-27B

Available
Alibaba (Qwen)Open source

Dense 27B that punches far above its weight on agentic coding — easy to self-host on a single GPU node.

Dense27B256K ctxMay 12, 2026

Grok 4.3

Available
xAIFrontierProprietary

xAI's agentic flagship with a 1M-token context and aggressive API pricing.

MoEUndisc.1M ctxMay 6, 2026

DeepSeek V4-Flash

Preview
DeepSeekOpen source

Efficient V4 companion model with 284B total / 13B active parameters and the same one-million-token context window.

MoE284B1M ctxApr 24, 2026

DeepSeek V4-Pro

Preview
DeepSeekFrontierOpen source

Preview-series sparse MoE flagship with a one-million-token context window and 1.6T total / 49B active parameters.

MoE1.6T1M ctxApr 24, 2026

Hunyuan-A13B-Instruct

Available
Tencent HunyuanOpen weights

Tencent Hunyuan open-weight fine-grained MoE model with 80B total parameters and 13B active parameters, optimized for agentic tool use.

MoE80B ctxApr 22, 2026

GLM-5.1

Available
Z.ai (Zhipu AI)FrontierOpen source

Z.ai agentic-engineering follow-up to GLM-5, with stronger coding performance and better long-horizon tool-use behavior.

MoE754B ctxApr 8, 2026

Claude Mythos

Preview
AnthropicFrontierProprietary

A frontier model Anthropic disclosed on April 7, 2026 but declined to release publicly, citing security risk. Shipped only via 'Project Glasswing' to ~50 defensive-security partners, then suspended on June 12, 2026 under a US government directive.

Undisc. ctxApr 7, 2026

Gemma 4 31B

Available
Google DeepMindOpen source

Google DeepMind's Gemma 4 advanced-reasoning open model for personal computers, part of the April 2026 Gemma 4 family.

Dense31B ctxApr 2, 2026

Kimi K2.6

Available
Moonshot AIFrontierOpen weights

Moonshot's open native multimodal agentic model for long-horizon coding, visual interface generation, and autonomous tool orchestration.

MoE1T256K ctxMar 30, 2026

Mistral Medium 3.5

Available
Mistral AIOpen weights

Dense 128B open-weight model with a 256k context and strong coding performance for its size.

Dense128B256K ctxMar 18, 2026

Nemotron 3 Super 120B-A12B

Available
NVIDIAFrontierOpen weights

Open-weight hybrid Mamba-Transformer MoE designed for collaborative agents and high-volume enterprise workflows.

Hybrid120B1M ctxMar 16, 2026

Step-3.5-Flash

Available
StepFunOpen source

StepFun's Apache-licensed sparse MoE model for fast agentic execution, coding, math, browsing, and tool-use workflows.

MoE196B256K ctxMar 14, 2026

Sarvam-105B

Available
Sarvam AIOpen source

Apache-licensed Indian-context MoE from Sarvam AI, optimized for reasoning, coding, agentic tasks, and 22 Indian languages.

MoE105B128K ctxMar 6, 2026

GPT-5.4

Available
OpenAIFrontierProprietary

Workhorse GPT-5 release with a dedicated Thinking mode; widely deployed across ChatGPT and the API.

MoEUndisc.400K ctxMar 5, 2026

Qwen3.5-397B

Available
Alibaba (Qwen)FrontierOpen source

Native vision-language MoE supporting 201 languages with a 1M-token context.

MoE397B1M ctxFeb 20, 2026

Gemini 3.1 Pro

Available
Google DeepMindFrontierProprietary

Generally available multimodal flagship with native tool use and a 2M-token context.

MoEUndisc.2M ctxFeb 19, 2026

GLM-5

Available
Z.ai (Zhipu AI)FrontierOpen source

Z.ai flagship for complex systems engineering and long-horizon agentic tasks, scaling the GLM line to 744B total / 40B active parameters.

MoE744B ctxFeb 11, 2026

Claude Opus 4.6

Available
AnthropicFrontierProprietary

Introduced genuinely autonomous multi-file coding and stronger computer use.

Undisc.200K ctxFeb 5, 2026

Qwen3-Coder-Next

Available
Alibaba (Qwen)Open source

Apache-licensed Qwen3-Next coding-agent model with 80B total / 3B active parameters, 256K context, and long-horizon tool-use training.

Hybrid80B262K ctxFeb 3, 2026

Kimi K2.5

Available
Moonshot AIFrontierOpen weights

Open multimodal Kimi model that adds native visual agentic intelligence, instant and thinking modes, and agent-swarm workflows on top of the K2 base.

MoE1T256K ctxJan 27, 2026

GLM-4.7

Available
Z.ai (Zhipu AI)FrontierOpen source

Coding-focused GLM release with improved multilingual agentic coding, terminal tasks, tool use, and interface generation.

MoE358B ctxJan 8, 2026

OLMo 3 Think 32B

Available
Allen Institute for AI (Ai2)Open source

Ai2's fully open thinking model with public weights, code, data, checkpoints, and training details across the OLMo 3 pipeline.

Dense32B ctxDec 15, 2025

Nemotron 3 Nano 30B-A3B

Available
NVIDIAOpen weights

Efficient Nemotron 3 MoE checkpoint for agentic reasoning and coding, activating about 3B parameters while supporting 1M-token contexts.

Hybrid30B1M ctxDec 15, 2025

GLM-4.6V

Available
Z.ai (Zhipu AI)Open source

Open 106B-class vision-language model with native multimodal function calling for visual agents.

MoE106B128K ctxDec 8, 2025

Mistral Large 3

Available
Mistral AIFrontierOpen weights

Mistral's largest open-weight MoE, aimed at frontier reasoning while remaining self-hostable.

MoE675B256K ctxDec 2, 2025

DeepSeek-V3.2

Available
DeepSeekFrontierOpen source

Reasoning-first agent model that adds DeepSeek Sparse Attention and thinking directly inside tool-use workflows.

MoE685B128K ctxDec 1, 2025

DeepSeek-V3.2-Speciale

Available
DeepSeekFrontierOpen source

High-compute reasoning variant of V3.2, positioned for olympiad-level math, programming, and other deep reasoning tasks.

MoE685B128K ctxDec 1, 2025

LFM2 1.2B

Available
Liquid AIOpen weights

Liquid AI hybrid model for efficient CPU/GPU/NPU local deployment, using short convolutions plus attention blocks.

Hybrid1.17B33K ctxNov 28, 2025

Kimi K2 Thinking

Available
Moonshot AIFrontierOpen weights

Open K2 reasoning-agent variant that interleaves step-by-step thinking with tool calls and supports stable 200-300 step tool-use trajectories.

MoE1T256K ctxNov 6, 2025

Kimi-Linear-48B-A3B-Instruct

Available
Moonshot AIOpen source

MIT-licensed hybrid linear-attention model using Kimi Delta Attention, built for million-token contexts with much lower KV-cache usage.

Hybrid48B1.0M ctxOct 31, 2025

GLM-4.6

Available
Z.ai (Zhipu AI)FrontierOpen source

Agentic reasoning and coding upgrade over GLM-4.5, expanding the text context window from 128K to 200K tokens.

MoE357B200K ctxSep 30, 2025

DeepSeek-V3.2-Exp

Preview
DeepSeekOpen source

Experimental checkpoint that introduced DeepSeek Sparse Attention as an efficiency bridge between V3.1-Terminus and V3.2.

MoE685B128K ctxSep 29, 2025

DeepSeek-V3.1-Terminus

Available
DeepSeekOpen source

Stability update to V3.1 focused on language consistency, code-agent reliability, and search-agent behavior.

MoE685B128K ctxSep 22, 2025

Kimi K2 Instruct 0905

Available
Moonshot AIFrontierOpen weights

September 2025 K2 update with stronger agentic coding, better frontend generation, and a doubled 256K context window.

MoE1T256K ctxSep 5, 2025

Gemma 3 27B

Available
Google DeepMindOpen weights

Google's open multimodal model: 128k context, 140+ languages, runs on a single GPU.

Dense27B128K ctxSep 4, 2025

DeepSeek-V3.1

Available
DeepSeekOpen source

Hybrid thinking/non-thinking release that upgraded tool calling, long-context training, and agent task performance.

MoE671B128K ctxAug 21, 2025

Seed-OSS-36B-Instruct

Available
ByteDance SeedOpen source

ByteDance Seed's Apache-licensed long-context reasoning and agent model, with controllable thinking budgets and a native 512K context.

Dense36B512K ctxAug 20, 2025

GLM-4.5V

Available
Z.ai (Zhipu AI)Open source

Vision-language GLM based on GLM-4.5-Air, covering image, video, document, grounding, and GUI-agent tasks.

MoE106B ctxAug 11, 2025

gpt-oss-20b

Available
OpenAIOpen source

Smaller gpt-oss reasoning model optimized for local inference on systems with about 16GB of memory.

MoE21B128K ctxAug 5, 2025

gpt-oss-120b

Available
OpenAIOpen source

OpenAI's larger open-weight reasoning model, a 117B-total / 5.1B-active MoE with 128K context for local and self-hosted deployment.

MoE117B128K ctxAug 5, 2025

Falcon-H1 34B

Available
Technology Innovation InstituteOpen weights

A hybrid attention + state-space-model (SSM) design that matches 70B-class models with fewer parameters.

Hybrid34B256K ctxJul 31, 2025

GLM-4.5

Available
Z.ai (Zhipu AI)FrontierOpen source

Open agentic, reasoning, and coding foundation model that marked Z.ai international rebrand and MIT-licensed GLM push.

MoE355B128K ctxJul 28, 2025

GLM-4.5-Air

Available
Z.ai (Zhipu AI)Open source

Compact GLM-4.5 companion with 106B total / 12B active parameters for efficient agentic reasoning and coding.

MoE106B128K ctxJul 28, 2025

EXAONE 4.0 32B

Available
LG AI ResearchOpen weights

LG AI Research's unified model with non-reasoning and reasoning modes, agentic tool use, and English, Korean, and Spanish support.

Dense32B ctxJul 15, 2025

Kimi K2 Instruct

Available
Moonshot AIFrontierOpen weights

Original open K2 post-trained model: a 1T-parameter MoE optimized for coding, reasoning, and tool-using agentic workflows.

MoE1T128K ctxJul 11, 2025

Grok 4

Deprecated
xAIProprietary

xAI's fourth-generation Grok line, preceding the later 4.x API updates already tracked in the catalog.

Undisc. ctxJul 9, 2025

SmolLM3 3B

Available
Hugging FaceOpen source

Hugging Face's fully open 3B multilingual long-context model with optional reasoning mode and 128K context.

Dense3B128K ctxJul 8, 2025

ERNIE-4.5-300B-A47B

Available
BaiduOpen source

Baidu's open ERNIE 4.5 language MoE, part of a 10-variant Apache-licensed model family built with heterogeneous multimodal MoE training.

MoE300B128K ctxJun 30, 2025

ERNIE-4.5-VL-424B-A47B

Available
BaiduOpen source

Baidu's largest ERNIE 4.5 vision-language MoE, supporting text, image, and video inputs with thinking and non-thinking modes.

MoE424B128K ctxJun 30, 2025

Kimi-VL-A3B-Thinking-2506

Available
Moonshot AIOpen source

Updated MIT-licensed Kimi-VL reasoning model with better multimodal reasoning, video understanding, high-resolution perception, and lower thinking-token use.

MoE16B128K ctxJun 21, 2025

Kimi-Dev-72B

Available
Moonshot AIOpen source

MIT-licensed coding LLM trained with repository-level reinforcement learning for software issue resolution.

Dense73B ctxJun 17, 2025

MiniMax-M1-80k

Available
MiniMaxFrontierOpen source

Open Apache-licensed hybrid-attention reasoning model with 456B total / 45.9B active parameters and a native 1M-token context.

Hybrid456B1M ctxJun 16, 2025

Magistral Medium

Available
Mistral AIProprietary

Mistral's first dedicated reasoning model family, released in Small open-weight and Medium enterprise/API tiers.

Undisc. ctxJun 10, 2025

Magistral Small

Available
Mistral AIOpen weights

Open-weight 24B reasoning model from Mistral's Magistral family, popular for local reasoning experiments.

Dense24B40K ctxJun 10, 2025

DeepSeek-R1-0528

Available
DeepSeekFrontierOpen source

Major R1 reasoning update with stronger math, programming, general logic, function calling, and reduced hallucinations.

MoE671B128K ctxMay 28, 2025

Claude Opus 4

Deprecated
AnthropicProprietary

First Claude 4 Opus model, positioned for long-running agentic and coding work before the 4.x point releases.

Undisc.200K ctxMay 22, 2025

Seed Thinking v1.5

Available
ByteDance SeedProprietary

ByteDance Seed reasoning model focused on long-horizon thinking and problem solving.

Undisc. ctxMay 22, 2025

Sarvam-M

Available
Sarvam AIOpen weights

Sarvam's medium-scale open model for multilingual Indian-language chat, reasoning, and translation tasks.

DenseUndisc. ctxMay 21, 2025

Phi-4 Reasoning

Available
MicrosoftOpen weights

Phi-4 reasoning-specialized model family for math, science, and chain-of-thought style tasks.

Dense14B ctxApr 30, 2025

Granite 3.3 8B

Available
IBMOpen source

Granite 3.3 text update for enterprise chat, RAG, and instruction-following workflows.

Dense8B128K ctxApr 30, 2025

Qwen3-235B-A22B

Available
Alibaba (Qwen)Open source

Largest open Qwen3 MoE, introducing hybrid thinking/non-thinking modes and 119-language coverage.

MoE235B128K ctxApr 28, 2025

Kimi-Audio-7B-Instruct

Available
Moonshot AIOpen source

Open audio foundation model for audio understanding, generation, speech recognition, audio QA, captioning, and speech conversation.

Hybrid10B ctxApr 25, 2025

Kimi-VL-A3B-Instruct

Available
Moonshot AIOpen source

Efficient MIT-licensed vision-language MoE for OCR, image/video understanding, long documents, and OS-style agent tasks.

MoE16B128K ctxApr 17, 2025

OpenAI o3

Available
OpenAIProprietary

Reasoning model released alongside o4-mini with tool use, image reasoning, and stronger agentic problem solving.

Undisc. ctxApr 16, 2025

GPT-4.1

Deprecated
OpenAIProprietary

API model family focused on coding, instruction following, and one-million-token long-context work.

Undisc.1M ctxApr 14, 2025

Llama 4 Maverick

Available
Meta AIFrontierOpen weights

Meta's flagship open-weight MoE; highest MMLU among open models at release.

MoE400B1M ctxApr 5, 2025

Llama 4 Scout

Available
Meta AIOpen weights

Efficient open-weight MoE designed for very long context on modest hardware.

MoE109B10M ctxApr 5, 2025

Llama-3.3-Nemotron-Super-49B

Available
NVIDIAOpen weights

Open Llama Nemotron reasoning model from NVIDIA's 2025 Nemotron family.

Dense49B128K ctxApr 2, 2025

Qwen2.5-Omni-7B

Available
Alibaba (Qwen)Open weights

Local omni-modal Qwen model that supports text, image, audio, video, and speech generation in a 7B package.

Dense7B ctxMar 26, 2025

DeepSeek-V3-0324

Available
DeepSeekOpen source

Post-R1 V3 update with improved reasoning, front-end coding, Chinese writing, search, and function calling.

MoE671B128K ctxMar 25, 2025

Gemini 2.5 Pro

Deprecated
Google DeepMindProprietary

Reasoning-focused Gemini 2.5 model that made thinking a core part of Google's flagship model line.

Undisc.1M ctxMar 25, 2025

Mistral Small 3.1

Available
Mistral AIOpen source

Apache-licensed Small update adding vision and a 128K context window to the efficient 24B line.

Dense24B128K ctxMar 17, 2025

ERNIE X1

Available
BaiduProprietary

Baidu's reasoning model released alongside ERNIE 4.5 before the open ERNIE 4.5 weights.

Undisc. ctxMar 16, 2025

OLMo 2 32B

Available
Allen Institute for AI (Ai2)Open source

A fully open model — weights, data, and training code all public — and the first such to beat GPT-3.5 / GPT-4o mini.

Dense32B4K ctxMar 13, 2025

Command A

Available
CohereOpen weights

Enterprise-grade model tuned for RAG, tool use, and multilingual business workloads.

Dense111B256K ctxMar 13, 2025

Granite 3.2 8B

Available
IBMOpen source

Granite 3.2 update with reasoning controls and multimodal/document-oriented Granite variants.

Dense8B128K ctxFeb 26, 2025

Claude 3.7 Sonnet

Retired
AnthropicProprietary

Anthropic's first hybrid-reasoning Sonnet. Shut down May 11, 2026 as the 4.x line matured.

Undisc.200K ctxFeb 24, 2025

Moonlight-16B-A3B-Instruct

Available
Moonshot AIOpen source

MIT-licensed 16B/3B-active MoE trained with Moonshot's scalable Muon optimizer experiments.

MoE16B8K ctxFeb 24, 2025

DeepHermes 3 Llama 3 8B

Available
Nous ResearchOpen weights

Nous reasoning-oriented Hermes model trained to combine concise answers with optional deep reasoning traces.

Dense8B8K ctxFeb 18, 2025

Grok 3

Deprecated
xAIProprietary

xAI's third-generation model family, introduced with stronger reasoning, search, and coding modes.

Undisc. ctxFeb 17, 2025

Dolphin 3.0 Llama 3.1 8B

Available
Cognitive ComputationsOpen weights

Popular local assistant model tuned for coding, math, function calling, and agentic workflows.

Dense8B128K ctxFeb 2, 2025

Mistral Small 3

Available
Mistral AIOpen source

A latency-optimized 24B dense model under Apache-2.0 — a popular local-deployment workhorse.

Dense24B32K ctxJan 30, 2025

Qwen2.5-Max

Available
Alibaba (Qwen)Proprietary

Proprietary MoE flagship for the Qwen2.5 generation, released through Qwen Chat and Alibaba Cloud APIs.

MoEUndisc. ctxJan 29, 2025

Qwen2.5-VL-72B

Available
Alibaba (Qwen)Open weights

Vision-language Qwen2.5 model for image, document, video, and agentic visual grounding tasks.

Dense72B128K ctxJan 26, 2025

Doubao-1.5-pro

Available
ByteDance SeedProprietary

Doubao 1.5 Pro update positioned for stronger multimodal, reasoning, and agentic work in Volcano Engine.

Undisc. ctxJan 22, 2025

DeepSeek-R1

Available
DeepSeekFrontierOpen source

Breakout open reasoning model trained with large-scale reinforcement learning and released with weights under MIT.

MoE671B128K ctxJan 20, 2025

Kimi k1.5

Available
Moonshot AIProprietary

Moonshot's multimodal reinforcement-learning reasoning model, reported as matching OpenAI o1 on math, coding, and multimodal reasoning.

Undisc. ctxJan 20, 2025

MiniMax-01

Available
MiniMaxOpen weights

Open MiniMax generation with MiniMax-Text-01 and MiniMax-VL-01 long-context models.

Hybrid456B4M ctxJan 15, 2025

DeepSeek-V3

Available
DeepSeekOpen source

The 671B/37B-active MoE release that made DeepSeek a central open-model lab before the R1 breakthrough.

MoE671B128K ctxDec 26, 2024

Step-2

Available
StepFunProprietary

Second-generation StepFun foundation model line with larger-scale multimodal and reasoning ambitions.

Undisc. ctxDec 23, 2024

Granite 3.1 8B

Available
IBMOpen source

IBM's enterprise-focused open model with a 128k context, Apache-2.0 licensed.

Dense8B128K ctxDec 18, 2024

Falcon 3 10B

Available
Technology Innovation InstituteOpen weights

UAE's TII open model designed to run on light infrastructure, including laptops.

Dense10B32K ctxDec 17, 2024

Command R7B

Available
CohereOpen weights

Cohere's smallest, fastest R-series model, tuned for RAG and tool use on modest hardware.

Dense8B128K ctxDec 13, 2024

Phi-4

Available
MicrosoftOpen source

A 14B dense model that rivals far larger ones on math and reasoning, under a permissive MIT license.

Dense14B16K ctxDec 12, 2024

Gemini 2.0 Flash

Deprecated
Google DeepMindProprietary

First Gemini 2.0 release, built for native multimodal input/output, tool use, and agentic product integrations.

Undisc.1M ctxDec 11, 2024

EXAONE 3.5 32B

Available
LG AI ResearchOpen weights

EXAONE 3.5 32B open-weight model for bilingual reasoning, coding, and long-context tasks.

Dense32B32K ctxDec 9, 2024

Llama 3.3 70B

Available
Meta AIOpen weights

Late-2024 70B Llama update delivering much of the 405B instruction-following quality at lower serving cost.

Dense70B128K ctxDec 6, 2024

OpenAI o1

Deprecated
OpenAIProprietary

General release of OpenAI's o1 reasoning model with stronger deliberative reasoning and multimodal ChatGPT integration.

Undisc. ctxDec 5, 2024

Amazon Nova Pro

Available
AmazonProprietary

AWS-native multimodal model with a 300k context; size and architecture undisclosed.

Undisc.300K ctxDec 3, 2024

Amazon Nova Lite

Available
AmazonProprietary

Lower-cost multimodal Nova understanding model for text, image, and video inputs.

Undisc.300K ctxDec 3, 2024

QwQ-32B-Preview

Available
Alibaba (Qwen)Open source

Qwen's first public reasoning-preview model, aimed at math, coding, and deliberate problem solving.

Dense32B32K ctxNov 28, 2024

Tulu 3 405B

Available
Allen Institute for AI (Ai2)Open weights

Ai2's post-trained open instruction model line, scaling the Tulu recipe to Llama 3.1 405B.

Dense405B128K ctxNov 21, 2024

DeepSeek-R1-Lite-Preview

Retired
DeepSeekProprietary

Reasoning-preview model exposed in DeepSeek Chat ahead of the open DeepSeek-R1 release.

Undisc. ctxNov 20, 2024

Qwen2.5-Coder-32B

Available
Alibaba (Qwen)Open source

Code-specialized Qwen2.5 model family, with the 32B checkpoint as the flagship open coding model.

Dense32B128K ctxNov 12, 2024

Hunyuan-Large

Available
Tencent HunyuanOpen weights

Tencent's 389B total / 52B active open-weight Transformer MoE, released with a 256K pretraining context and 128K instruct context.

MoE389B128K ctxNov 4, 2024

SmolLM2 1.7B

Available
Hugging FaceOpen source

Compact on-device model family trained on 11T tokens, popular for lightweight local chat and experimentation.

Dense1.7B ctxNov 4, 2024

Claude 3.5 Haiku

Deprecated
AnthropicProprietary

Fast, lower-cost Claude 3.5 model for latency-sensitive coding, tool-use, and customer-facing workloads.

Undisc.200K ctxOct 22, 2024

Sarvam-1

Available
Sarvam AIOpen weights

Sarvam's 2B open model trained for ten major Indian languages.

Dense2B ctxOct 22, 2024

Granite 3.0 8B

Available
IBMOpen source

Apache-licensed Granite 3.0 text model, part of IBM's push toward enterprise-friendly open models.

Dense8B4K ctxOct 21, 2024

Yi-Lightning

Available
01.AIProprietary

01.AI's MoE API model that reached the global top-10 on Chatbot Arena, strong in Chinese, math, and coding.

MoEUndisc. ctxOct 16, 2024

Ministral 8B

Available
Mistral AIProprietary

Small Mistral model line optimized for edge and low-latency workloads.

Dense8B128K ctxOct 16, 2024

Llama-3.1-Nemotron-70B

Available
NVIDIAOpen weights

NVIDIA-tuned Llama 3.1 70B instruction model optimized with Nemotron reward and alignment recipes.

Dense70B128K ctxOct 15, 2024

Llama 3.2 90B Vision

Available
Meta AIOpen weights

First Llama family release with native vision models, alongside smaller edge-oriented 1B and 3B text models.

Dense90B128K ctxSep 25, 2024

Molmo 72B

Available
Allen Institute for AI (Ai2)Open weights

Open multimodal model family trained for strong image understanding, pointing, and visual grounding.

Dense72B ctxSep 25, 2024

Qwen2.5-72B

Available
Alibaba (Qwen)Open weights

Broad Qwen2.5 foundation-model update spanning general, coding, math, and multimodal descendants.

Dense72B128K ctxSep 19, 2024

Pixtral 12B

Available
Mistral AIOpen source

Mistral's first open multimodal model, adding image understanding to a Mistral text backbone.

Dense12B128K ctxSep 17, 2024

OpenAI o1-preview

Retired
OpenAIProprietary

OpenAI's first public reasoning-model preview, optimized to spend more inference time on hard math, coding, and science tasks.

Undisc. ctxSep 12, 2024

Yi-Coder-9B

Available
01.AIOpen weights

01.AI's compact code model trained for repository-scale programming and code completion tasks.

Dense9B128K ctxSep 5, 2024

DeepSeek-V2.5

Available
DeepSeekOpen source

Unified DeepSeek V2 generation combining general-chat and coding strengths before the V3 series.

MoE236B128K ctxSep 5, 2024

Hunyuan Turbo

Available
Tencent HunyuanProprietary

Tencent's faster, lower-cost Hunyuan update before the open Hunyuan-Large model card.

Undisc. ctxSep 5, 2024

OLMoE 1B-7B

Available
Allen Institute for AI (Ai2)Open source

Fully open sparse MoE model with 7B total and about 1B active parameters.

MoE7B ctxSep 3, 2024

Jamba 1.5 Large

Available
AI21 LabsOpen weights

Israel's AI21 hybrid Mamba-Transformer MoE, with a 256k context and strong long-document throughput.

Hybrid398B256K ctxAug 22, 2024

Phi-3.5 MoE

Available
MicrosoftOpen weights

Phi-3.5 mixture-of-experts model, scaling Microsoft's small-model line while preserving efficient active parameters.

MoE42B128K ctxAug 20, 2024

Hermes 3 Llama 3.1 405B

Available
Nous ResearchOpen weights

Large Hermes 3 instruction-tuned model built on Meta's Llama 3.1 405B.

Dense405B128K ctxAug 15, 2024

Grok-2

Retired
xAIProprietary

Second-generation Grok release with Grok-2 and Grok-2 mini for chat, coding, reasoning, and image-enabled product experiences.

Undisc. ctxAug 13, 2024

EXAONE 3.0 7.8B

Available
LG AI ResearchOpen weights

LG's first open-weight EXAONE model, a compact bilingual instruction model for Korean and English.

Dense7.8B ctxAug 7, 2024

MiniCPM-V 2.6

Available
OpenBMBOpen weights

8B vision-language model for local image, multi-image, OCR, and video understanding, with llama.cpp and Ollama support.

Dense8B ctxAug 2, 2024

Llama 3.1 405B

Available
Meta AIOpen weights

Meta's first frontier-scale open Llama model, with 405B parameters, 128K context, multilingual support, and tool-use improvements.

Dense405B128K ctxJul 23, 2024

Mistral NeMo

Available
Mistral AIOpen source

Apache-licensed 12B model co-developed with NVIDIA, including a 128K context window and strong multilingual tokenization.

Dense12B128K ctxJul 18, 2024

Gemma 2 27B

Available
Google DeepMindOpen weights

Second-generation Gemma model, improving open-weight quality and efficiency at 9B and 27B sizes.

Dense27B8K ctxJun 27, 2024

Claude 3.5 Sonnet

Retired
AnthropicProprietary

Major Sonnet upgrade that became Anthropic's default high-intelligence workhorse for coding, writing, and visual reasoning.

Undisc.200K ctxJun 20, 2024

DeepSeek-Coder-V2

Available
DeepSeekOpen source

Open code-focused MoE built from DeepSeek-V2, expanding programming-language coverage and coding benchmark performance.

MoE236B128K ctxJun 17, 2024

Nemotron-4 340B

Available
NVIDIAOpen weights

NVIDIA's large open model family for synthetic data generation and reward modeling.

Dense340B4K ctxJun 14, 2024

Qwen2-72B

Available
Alibaba (Qwen)Open weights

Qwen2's largest dense model, introducing stronger multilingual support, coding/math gains, and long-context variants.

Dense72B128K ctxJun 7, 2024

GLM-4-9B

Available
Z.ai (Zhipu AI)Open weights

Open GLM-4 9B model family, covering chat, long-context, and code-oriented variants.

Dense9B128K ctxJun 5, 2024

Codestral 22B

Available
Mistral AIOpen weights

Mistral's first code-specialized model, trained for code generation, fill-in-the-middle, and multi-language programming tasks.

Dense22B32K ctxMay 29, 2024

Aya 23 35B

Available
CohereOpen weights

Open multilingual research model covering 23 languages, released by Cohere For AI.

Dense35B ctxMay 23, 2024

Doubao-pro

Available
ByteDance SeedProprietary

ByteDance's commercial Doubao foundation model line for text, code, and assistant workloads.

Undisc. ctxMay 15, 2024

GPT-4o

Retired
OpenAIProprietary

The 2024 omni-modal model that defined a generation of assistants. Deprecated in Feb 2026 and fully retired across ChatGPT on April 3, 2026.

Undisc.128K ctxMay 13, 2024

Yi-1.5-34B

Available
01.AIOpen weights

Yi 1.5 update with stronger instruction following, coding, math, and multilingual performance.

Dense34B4K ctxMay 13, 2024

Falcon 2 11B

Available
Technology Innovation InstituteOpen weights

Falcon 2 generation, including text and vision-language 11B models under a permissive TII license.

Dense11B8K ctxMay 13, 2024

DeepSeek-V2

Available
DeepSeekOpen source

DeepSeek's first major MoE general model with Multi-head Latent Attention and low-cost API positioning.

MoE236B128K ctxMay 7, 2024

Granite Code 34B

Available
IBMOpen source

Apache-2.0 code model from IBM's Granite Code family, used for local code generation and enterprise coding assistants.

Dense34B8K ctxMay 6, 2024

Amazon Titan Text Premier

Available
AmazonProprietary

Larger Titan text model for enterprise RAG, summarization, and agent workflows in Amazon Bedrock.

Undisc. ctxApr 30, 2024

Snowflake Arctic

Available
Snowflake AI ResearchOpen source

Apache-2.0 enterprise LLM with 480B total / 17B active parameters, optimized for SQL, code, and instruction following.

MoE480B ctxApr 24, 2024

Phi-3 Mini

Available
MicrosoftOpen weights

3.8B-parameter Phi-3 model released as a phone-capable small model with 4K and 128K variants.

Dense3.8B128K ctxApr 23, 2024

Llama 3 70B

Available
Meta AIOpen weights

First Llama 3 release, with 8B and 70B open models and a stronger tokenizer, data mix, and post-training stack.

Dense70B8K ctxApr 18, 2024

Mixtral 8x22B

Available
Mistral AIOpen source

Larger open Mixtral sparse MoE with 141B total and 39B active parameters, released under Apache-2.0.

MoE141B64K ctxApr 17, 2024

abab6.5

Available
MiniMaxProprietary

MiniMax's commercial long-context abab model generation before the open MiniMax-01 and M series.

Undisc.1M ctxApr 17, 2024

WizardLM-2 8x22B

Available
MicrosoftOpen weights

Microsoft's WizardLM-2 MoE chat model, widely mirrored and run locally after its model-card release.

MoE141B66K ctxApr 15, 2024

Step-1V

Available
StepFunProprietary

StepFun's first major vision-language model, released after the Step-1 language model.

Undisc. ctxApr 12, 2024

CodeGemma 7B

Available
Google DeepMindOpen weights

Open code-specialized Gemma model for local code completion, generation, and instruction-following.

Dense7B8K ctxApr 9, 2024

Command R+

Deprecated
CohereProprietary

Higher-capability RAG and tool-use model in Cohere's Command R family.

Undisc.128K ctxApr 4, 2024

Grok-1.5

Retired
xAIProprietary

Grok update with stronger reasoning and a 128K context window.

Undisc.128K ctxMar 28, 2024

Jamba

Available
AI21 LabsOpen weights

First Jamba hybrid Transformer-Mamba MoE model with open weights and a 256K context length.

Hybrid52B256K ctxMar 28, 2024

DBRX Instruct

Available
Databricks / MosaicMLOpen weights

Databricks' 132B-total / 36B-active open MoE model for code, math, RAG, and enterprise self-hosted workloads.

MoE132B32K ctxMar 27, 2024

Step-1

Available
StepFunProprietary

StepFun's first public foundation model generation, introduced as a trillion-parameter Chinese model line.

Undisc. ctxMar 23, 2024

Kimi 1M

Available
Moonshot AIProprietary

Long-context Kimi upgrade advertised with support for million-character document and conversation contexts.

Undisc. ctxMar 18, 2024

Command R

Deprecated
CohereProprietary

Enterprise RAG-focused model with tool use, citations, multilingual retrieval, and long-context support.

Undisc.128K ctxMar 11, 2024

Claude 3 Opus

Deprecated
AnthropicProprietary

Highest-capability Claude 3 model, launched with Sonnet and Haiku and Anthropic's first major vision-capable Claude family.

Undisc.200K ctxMar 4, 2024

StarCoder2 15B

Available
BigCodeOpen weights

Next-generation BigCode code model trained on 4T+ tokens and 600+ programming languages, with 16K context.

Dense16B16K ctxFeb 28, 2024

Mistral Large

Deprecated
Mistral AIProprietary

Mistral's first proprietary flagship API model, introduced alongside Le Chat and stronger multilingual/coding performance.

Undisc.32K ctxFeb 26, 2024

Gemma 7B

Available
Google DeepMindOpen weights

First Gemma open-weight text model family, derived from the same research lineage as Gemini.

Dense7B8K ctxFeb 21, 2024

Gemini 1.5 Pro

Deprecated
Google DeepMindProprietary

Gemini generation that introduced production-scale long context, eventually expanding to a two-million-token window.

MoEUndisc.2M ctxFeb 15, 2024

Qwen1.5-110B

Available
Alibaba (Qwen)Open weights

Largest Qwen1.5 model, released as the bridge from the original Qwen line to Qwen2.

Dense110B32K ctxFeb 5, 2024

OLMo 7B

Available
Allen Institute for AI (Ai2)Open source

Ai2's first fully open language model release, including weights, training data, code, logs, and intermediate checkpoints.

Dense7B4K ctxFeb 1, 2024

Stable LM 2 1.6B

Available
Stability AIOpen weights

Small multilingual Stable LM release built for low hardware barriers and local experimentation.

Dense1.6B ctxJan 19, 2024

GLM-4

Available
Z.ai (Zhipu AI)Proprietary

Zhipu's GLM-4 flagship generation, launched as the successor to ChatGLM3 with stronger tool use and multimodal variants.

Undisc.128K ctxJan 16, 2024

DeepSeekMoE 16B

Available
DeepSeekOpen source

Early DeepSeek sparse MoE research model that foreshadowed the later V2/V3 architecture direction.

MoE16B4K ctxJan 11, 2024

Nous Hermes 2 Mixtral

Available
Nous ResearchOpen source

Nous instruction-tuned Mixtral model with strong open-chat and tool-use adoption.

MoE47B32K ctxJan 11, 2024

OpenChat 3.5

Available
OpenChatOpen source

Compact Mistral-based local chat model trained with C-RLFT, popular in early 2024 local leaderboards.

Dense7B ctxJan 6, 2024

TinyLlama 1.1B Chat

Available
TinyLlamaOpen source

Compact Llama-style 1.1B chat model trained for local experimentation and low-memory deployments.

Dense1.1B ctxJan 1, 2024

Phi-2

Available
MicrosoftOpen weights

2.7B-parameter Phi model showing strong reasoning and language understanding at small scale.

Dense2.7B ctxDec 12, 2023

OpenHathi-7B

Available
Sarvam AIOpen weights

Sarvam AI's first open Indic language model, adapted from Llama 2 for Hindi and Indian-language work.

Dense7B ctxDec 12, 2023

Mixtral 8x7B

Available
Mistral AIOpen source

The open sparse Mixture-of-Experts that brought MoE efficiency to the open ecosystem.

MoE47B32K ctxDec 11, 2023

Gemini 1.0 Ultra

Deprecated
Google DeepMindProprietary

Google's first natively multimodal Gemini flagship, since superseded by the 1.5/2/3 lines.

Undisc.32K ctxDec 6, 2023

Qwen-72B

Available
Alibaba (Qwen)Open weights

Alibaba's first major open Qwen model and the start of a prolific open-weight line.

Dense72B32K ctxNov 30, 2023

DeepSeek LLM 67B

Available
DeepSeekOpen source

First general DeepSeek language model family, with 7B and 67B base/chat checkpoints.

Dense67B4K ctxNov 29, 2023

Claude 2.1

Retired
AnthropicProprietary

Claude update with a 200K context window, lower hallucination rates, and improved tool-use beta support.

Undisc.200K ctxNov 21, 2023

Yi-34B

Available
01.AIOpen weights

01.AI's strong bilingual open model, with a 200k-context variant.

Dense34B200K ctxNov 6, 2023

GPT-4 Turbo

Deprecated
OpenAIProprietary

Lower-cost GPT-4 generation with a 128K context window, introduced at OpenAI DevDay.

Undisc.128K ctxNov 6, 2023

Grok-1

Available
xAIOpen source

xAI's first Grok model, later released as open weights with a 314B-parameter MoE checkpoint.

MoE314B ctxNov 4, 2023

DeepSeek Coder 33B

Available
DeepSeekOpen source

DeepSeek's first public code-model family, released before the general DeepSeek LLM line.

Dense33B16K ctxNov 2, 2023

ERNIE 4.0

Available
BaiduProprietary

Baidu's fourth-generation ERNIE flagship, announced with stronger understanding, generation, reasoning, and memory.

Undisc. ctxOct 17, 2023

Kimi Chat

Available
Moonshot AIProprietary

Moonshot's first Kimi assistant release, establishing the long-context product line before the open Kimi model cards.

Undisc. ctxOct 9, 2023

LLaVA 1.5 13B

Available
LLaVAOpen weights

Open vision-language assistant and one of the most widely run early local multimodal models.

Hybrid13B ctxSep 30, 2023

Amazon Titan Text Express

Available
AmazonProprietary

Amazon's first-party Titan text generation model exposed through Bedrock, initially alongside embeddings and image models.

Undisc. ctxSep 28, 2023

Mistral 7B

Available
Mistral AIOpen source

The 7B that punched far above its weight and put Mistral on the map.

Dense7B8K ctxSep 27, 2023

Qwen-14B

Available
Alibaba (Qwen)Open weights

Second open Qwen size, expanding the first-generation Qwen language-model lineup.

Dense14B8K ctxSep 25, 2023

Granite 13B

Available
IBMOpen weights

IBM's early Granite foundation model family for enterprise language and code tasks.

Dense13B ctxSep 7, 2023

Hunyuan

Available
Tencent HunyuanProprietary

Tencent's first Hunyuan foundation model release, introduced as a general-purpose Chinese enterprise model.

Undisc. ctxSep 7, 2023

Falcon 180B

Available
Technology Innovation InstituteOpen weights

At launch the largest openly available model, from the UAE's TII.

Dense180B2K ctxSep 6, 2023

Code Llama 34B

Available
Meta AIOpen weights

Meta's first code-specialized Llama model family, released in base, Python, and instruction-tuned variants.

Dense34B16K ctxAug 24, 2023

Qwen-7B

Available
Alibaba (Qwen)Open weights

Alibaba's first open Qwen checkpoint and the start of the Qwen open-model line.

Dense7B32K ctxAug 3, 2023

Nous-Hermes-Llama2-13B

Available
Nous ResearchOpen weights

Early Nous Hermes instruction model on Llama 2, widely used in the open-model fine-tuning ecosystem.

Dense13B4K ctxJul 24, 2023

EXAONE 2.0

Retired
LG AI ResearchProprietary

Second EXAONE generation, improving bilingual Korean-English performance and enterprise deployment options.

Undisc. ctxJul 19, 2023

Llama 2 70B

Available
Meta AIOpen weights

The release that made capable open-weight models genuinely usable for production.

Dense70B4K ctxJul 18, 2023

Claude 2

Retired
AnthropicProprietary

Anthropic's first widely-available Claude, notable for an early 100k-token context window.

Undisc.100K ctxJul 11, 2023

ChatGLM2-6B

Available
Z.ai (Zhipu AI)Open weights

Second open ChatGLM generation, improving long context, inference efficiency, and bilingual chat quality.

Dense6B32K ctxJun 25, 2023

Phi-1

Available
MicrosoftOpen weights

Microsoft's first Phi small-language-model release, demonstrating strong code performance from textbook-quality synthetic data.

Dense1.3B ctxJun 21, 2023

Falcon 40B

Available
Technology Innovation InstituteOpen weights

TII's breakout open Falcon model, released before Falcon 180B and trained on the RefinedWeb corpus.

Dense40B2K ctxMay 25, 2023

PaLM 2

Retired
Google DeepMindProprietary

Google's improved multilingual, reasoning, and coding foundation model family introduced at I/O 2023.

DenseUndisc. ctxMay 10, 2023

MPT-7B

Available
Databricks / MosaicMLOpen source

MosaicML's permissively licensed 7B model, an early favorite for commercial local fine-tuning and long-context variants.

Dense7B2K ctxMay 5, 2023

Vicuna 13B

Available
LMSYS / SkyLabOpen weights

LMSYS instruction-tuned LLaMA model that became a landmark early local ChatGPT-style assistant.

Dense13B ctxMar 30, 2023

ERNIE Bot

Available
BaiduProprietary

Baidu's public chat assistant launch, built on the ERNIE foundation-model line.

Undisc. ctxMar 16, 2023

GPT-4

Deprecated
OpenAIProprietary

The model that brought reliable multi-step reasoning to the mainstream; size never disclosed.

Undisc.8K ctxMar 14, 2023

ChatGLM-6B

Available
Z.ai (Zhipu AI)Open weights

Zhipu AI and Tsinghua KEG's first widely used open bilingual ChatGLM checkpoint.

Dense6B2K ctxMar 14, 2023

Claude 1

Retired
AnthropicProprietary

Anthropic's first broadly announced Claude assistant model, launched through an API and select product partners.

Undisc. ctxMar 14, 2023

Jurassic-2 Ultra

Deprecated
AI21 LabsProprietary

Second-generation Jurassic model with better multilingual support, lower latency, and instruction following.

Undisc. ctxMar 9, 2023

GPT-3.5 Turbo

Retired
OpenAIProprietary

OpenAI's first ChatGPT API model, bringing the GPT-3.5 chat-tuned line to developers at much lower cost than text-davinci-003.

Undisc.4K ctxMar 1, 2023

LLaMA

Available
Meta AIOpen weights

Meta's first LLaMA, released to researchers; its leak catalyzed the open-weight movement.

Dense65B2K ctxFeb 24, 2023

Galactica

Withdrawn
Meta AIOpen weights

A science-focused model whose public demo was withdrawn after just three days over confidently wrong outputs — an early, instructive retraction.

Dense120B2K ctxNov 15, 2022

BLOOM

Available
BigScienceOpen weights

An open, multilingual 176B model (46 languages) from a global research collaboration.

Dense176B2K ctxJul 12, 2022

PaLM

Retired
Google DeepMindProprietary

Google's 540B Pathways model; the API was later deprecated in favor of Gemini.

Dense540B ctxApr 4, 2022

EXAONE 1.0

Retired
LG AI ResearchProprietary

LG AI Research's first EXAONE foundation model generation, introduced as a large multimodal expert AI.

Undisc. ctxDec 14, 2021

ERNIE 3.0 Titan

Retired
BaiduProprietary

Baidu's 260B-parameter ERNIE 3.0 Titan model, an early Chinese frontier-scale language model.

Dense260B ctxDec 8, 2021

Jurassic-1 Jumbo

Retired
AI21 LabsProprietary

AI21's first major API language model, launched through AI21 Studio.

Dense178B ctxAug 11, 2021

GPT-3

Retired
OpenAIProprietary

The 175B model that proved in-context learning at scale; its base API models were retired in 2024.

Dense175B2K ctxJun 11, 2020

GPT-2

Available
OpenAIOpen source

Initially withheld over misuse fears, then fully released in Nov 2019 — an early 'limited release' debate.

Dense1.5B1K ctxNov 5, 2019

BERT

Available
Google DeepMindOpen source

The bidirectional encoder that reshaped NLP and seeded the transformer era.

Dense0.34B512 ctxOct 11, 2018