02 / MODEL INDEX

AI model index

Browse first release dates, organizations, families, architecture, context length, and weight availability for every model in the archive.

182 models
GLM-5.3-Flash

The first natively multimodal GLM-5 Flash model: 320B total, 18B active parameters with hybrid sparse and linear attention.

Z.aiGLM
HYBRID320B1M ctxOpen
Qwen3.8-Flash-Next

A Qwen4 architecture preview: 125B main model plus 51B N-gram embeddings, 6B active per token, extendable to 1M context.

Alibaba CloudQwen
HYBRID176B262K ctxOpen
DeepSeek-V4-Flash-Vision-Exp

An experimental V4 Flash API model adding image input.

DeepSeekDeepSeek V / R
MOE284B1M ctx
GLM-5.3

Kept the GLM-5.2 base and improved complex coding, long-horizon tasks, and cyber capability entirely through scaled post-training.

Z.aiGLM
MOE
Qwen3.8-27B

The 27B dense open Qwen3.8 model for local deployment, vision, coding, and tool use.

Alibaba CloudQwen
DENSE27B262K ctxOpen
Gemini 3.7 Flash

The current Gemini Flash workhorse for software engineering, web development, multi-step agents, and design fidelity.

Google DeepMindGemini
1M ctx
DeepSeek-V4-Pro-0813

The production V4-Pro snapshot with stronger agents, low/high/max reasoning effort, and Responses API support.

DeepSeekDeepSeek V / R
MOE1.6T1M ctx
Grok 4.6

Built on Grok 4.5 with stronger long-running agents, complex-codebase work, and interactive visual projects.

SpaceXAIGrok
500K ctx
Muse Glimmer 30B

A 30B open model for always-on local agents, designed to run on one consumer GPU or a Mac or PC.

MetaMuse
DENSE30B128K ctxOpen
Wan 3.0 Video

The current preview all-in-one video model, generating up to 30 seconds with as many as 20 multimodal references.

Alibaba CloudWan
Muse Spark 1.2

The current Muse flagship, with stronger coding and the multi-agent Muse Code tool.

MetaMuse
1M ctx
Qwen3.8-Max

The Qwen3.8 flagship with 2.4T total, 95B active parameters, text-image-video input, and 1M context.

Alibaba CloudQwen
MOE2.4T1M ctxOpen
DeepSeek-V4-Flash-0731

The production API version, re-post-trained on the same architecture as V4-Flash-Preview with stronger agent capability.

DeepSeekDeepSeek V / R
MOE284B1M ctx
Seedance 2.5

Generates up to 30 seconds with multi-round extension, 30 image/10 video/10 audio references, and timestamp editing.

ByteDance SeedSeedance
MiniMax H3

An open omni-modal generation system producing up to 15-second, 2K, 24fps video with native stereo audio.

MiniMaxHailuo / H3
Open
Claude Opus 5

The new Opus flagship with a default 1M context and adaptive thinking for long-running agents, research, and enterprise work.

AnthropicClaude
1M ctx
Midjourney V8.2

The current default, focused on aesthetics, image quality, personalization, and fewer low-quality results.

MidjourneyMidjourney
Gemini 3.5 Flash-Lite

A low-latency 3.5 tier designed for high-volume automation and subagents.

Google DeepMindGemini
1M ctx
Gemini 3.6 Flash

Improved token efficiency, coding, and agent planning over 3.5 Flash while reducing output verbosity.

Google DeepMindGemini
1M ctx
Qwen-Image 3.0

Targets directly usable visual content with 4.5K-token prompts, 10px text, multilingual rendering, and complex layouts.

Alibaba CloudQwen-Image
Kimi K3

A 2.8T open MoE using KDA and Attention Residuals with native vision and a 1M context window.

Moonshot AIKimi
MOE2.8T1M ctxOpen
Grok 4.5

A fast flagship for coding, agents, and knowledge work, available in Grok Build and the API.

SpaceXAIGrok
Muse Spark 1.1

Improved tools, computer use, coding, and multimodal understanding with a 1M-context Meta Model API preview.

MetaMuse
1M ctx
Seedream 5.0 Pro

The professional-design flagship for information visualization, pixel-level interactive editing, realism, and multilingual generation.

ByteDance SeedSeedream
Muse Image

An agentic image model that uses search and code, self-refines, and supports single- and multi-image references.

MetaMuse
Muse Video Preview

A preview Muse video model emphasizing fidelity, temporal consistency, and native audio; not yet generally available.

MetaMuse
Hy3

The final Hy3 used product feedback to improve agents, coding, and output reliability.

Tencent HyHunyuan / Hy
MOE295B256K ctxOpen
Claude Sonnet 5

A new Sonnet generation with a default 1M context, adaptive thinking, and stronger agent capability closer to Opus.

AnthropicClaude
1M ctx
Nano Banana 2 Lite

The fastest and lowest-cost high-throughput model in the Nano Banana family.

Google DeepMindNano Banana
Gemini Omni Flash

Google’s Gemini-native model for cost-efficient video generation and conversational editing.

Google DeepMindVeo / Gemini Video
GPT-5.6

A three-tier series—Sol, Terra, and Luna—covering flagship, balanced, and cost-efficient capability levels.

OpenAIGPT
UNKNOWN
Seed2.1

The current Seed line, improving long-running general agents, end-to-end coding, and scientific workflows.

ByteDance SeedSeed
GLM-5.2

Extended context to 1M and optimized for large codebases, long documents, and project-scale agent work.

Z.aiGLM
MOE1M ctxOpen
DiffusionGemma 26B

An experimental Gemma 4 text-diffusion model exploring lower latency through parallel block generation.

Google DeepMindGemma
MOE26BOpen
Claude Fable 5

A Mythos-class model made generally available with conservative safeguards for software engineering, research, and complex knowledge work.

AnthropicClaude
Ray3.2

The current Ray model for fuller creative control and production API workflows.

Luma AIRay
Gemma 4 12B

A unified encoder-free multimodal model for laptops, filling the gap between E4B and 26B.

Google DeepMindGemma
DENSE12B128K ctxOpen
MiniMax-M3

MiniMax’s current open flagship with 1M context, native image/video understanding, coding, and computer use.

MiniMaxMiniMax M
1M ctxOpen
Claude Opus 4.8

Improved agent judgment, tool efficiency, computer use, and long-session collaboration over Opus 4.7.

AnthropicClaude
1M ctx
Gemini 3.5 Flash

The first public model in Google’s 3.5 generation, combining frontier intelligence with low-latency action in the Flash tier.

Google DeepMindGemini
1M ctx
Gemini 3.1 Flash-Lite

The low-latency, high-throughput, cost-efficient tier in the Gemini 3 family.

Google DeepMindGemini
1M ctx
GPT-5.5 Instant

The everyday fast tier of GPT-5.5, improving accuracy, concision, and use of personalized context.

OpenAIGPT
DeepSeek-V4-Pro-Preview

The high-capability DeepSeek V4 preview: a 1.6T-total, 49B-active MoE with a 1M context window.

DeepSeekDeepSeek V / R
MOE1.6T1M ctxOpen
DeepSeek-V4-Flash-Preview

The efficient V4 preview with 284B total, 13B active parameters, and a 1M context window.

DeepSeekDeepSeek V / R
MOE284B1M ctxOpen
Hy3 Preview

The first Hy3 preview after Tencent rebuilt its training stack, combining fast/slow reasoning with sparse MoE.

Tencent HyHunyuan / Hy
MOE295B256K ctxOpen
GPT-5.5

A 1M-context frontier model for complex knowledge work, computer use, coding, and scientific research.

OpenAIGPT
1M ctx
MiMo-V2.5

MiMo’s current general model combining native full-modal understanding, 1M context, and agent execution.

Xiaomi MiMoMiMo
1M ctxOpen
MiMo-V2.5-Pro

The high-intensity MiMo-V2.5 agent tier for long-range reasoning, coding, and complex-task efficiency.

Xiaomi MiMoMiMo
MOE1T1M ctxOpen
Qwen3.6-27B

The 27B dense Qwen3.6 tier for local deployment and general multimodal agents.

Alibaba CloudQwen
DENSE27B262K ctxOpen
GPT Image 2

The current flagship for 2K production assets, complex layouts, multilingual text, and high-fidelity editing.

OpenAIGPT Image
Kimi K2.6

The open successor to K2.5 with stronger long-horizon coding, general agents, vision, and Agent Swarm.

Moonshot AIKimi
MOE256K ctxOpen
Claude Opus 4.7

Focused on hard software engineering, long-run consistency, higher-resolution vision, and self-verification.

AnthropicClaude
1M ctx
Qwen3.6-35B-A3B

The efficient Qwen3.6 MoE tier, using 3B active parameters for multimodal, coding, and tool tasks.

Alibaba CloudQwen
HYBRID35B262K ctxOpen
Midjourney V8.1

Added native 2K HD, roughly 4–5× faster standard generation, and stronger text and detail retention.

MidjourneyMidjourney
Muse Spark

The first Muse model, with native multimodal reasoning, tools, visual chain of thought, and multi-agent orchestration.

MetaMuse
GLM-5.1

A GLM-5 update for long-horizon work, designed to execute autonomously for hours within one task.

Z.aiGLM
MOE200K ctxOpen
Wan 2.7

Unified audio text-to-video, first/last frames, continuation, references, and editing in a hosted 1080p line.

Alibaba CloudWan
Gemma 4 31B

The high-quality dense Gemma 4 flagship for local reasoning, coding, and agent workflows.

Google DeepMindGemma
DENSE31B256K ctxOpen
Gemma 4 26B MoE

The Gemma 4 MoE activating only 3.8B parameters per step for local speed and intelligence density.

Google DeepMindGemma
MOE26B256K ctxOpen
Wan 2.7 Image

Wan’s image model for image sequences, multi-reference, interactive editing, and output up to 4K.

Alibaba CloudWan
MiniMax-M2.7

Strengthened agent teams, self-evolution, and end-to-end software productivity.

MiniMaxMiniMax M
Open
MiMo-V2-Pro

A trillion-parameter, 42B-active agent flagship with a 1M context window.

Xiaomi MiMoMiMo
MOE1T1M ctx
MiMo-V2-Omni

Unified text, vision, and speech perception with tools and GUI control in one agent foundation.

Xiaomi MiMoMiMo
256K ctx
GPT-5.4 mini

A fast small model for coding and subagents with text, image, tools, and a 400K context window.

OpenAIGPT
400K ctx
GPT-5.4 nano

The smallest GPT-5.4 tier for high-throughput extraction, ranking, and simple coding subtasks.

OpenAIGPT
GPT-5.4

Unified GPT-5.3-Codex coding advances with general reasoning, tools, and professional document work.

OpenAIGPT
GPT-5.3 Instant

A GPT-5.3 update for everyday conversation, focused on fewer unnecessary refusals, better search accuracy, and smoother writing.

OpenAIGPT
Nano Banana 2

Brings 0.5K–4K generation, thinking, and image-search grounding at Flash speed and price.

Google DeepMindNano Banana
Gemini 3.1 Pro Preview

The successor preview to Gemini 3 Pro, with a separate endpoint tuned to prioritize custom tools.

Google DeepMindGemini
1M ctx
Claude Sonnet 4.6

A full Sonnet 4.5 upgrade across coding, computer use, long context, agent planning, and design.

AnthropicClaude
1M ctx
Qwen3.5-397B-A17B

The first Qwen3.5 flagship, extending the Qwen3-Next hybrid architecture into a native multimodal agent model.

Alibaba CloudQwen
HYBRID397B262K ctxOpen
Seed2.0

Systematically improved multimodal understanding, foundation reasoning, and production agent performance.

ByteDance SeedSeed
Seedream 5.0 Lite

Brought deeper reasoning and live search into lightweight image generation and editing.

ByteDance SeedSeedream
GLM-5

Scaled to 744B total parameters with DSA, repositioning the model from code generation toward complex systems engineering.

Z.aiGLM
MOE744BOpen
GPT-5.3-Codex-Spark

A small, ultra-fast Codex research preview designed for real-time interactive coding.

OpenAIGPT
128K ctx
MiniMax-M2.5

A high-efficiency update for coding, tools, and workplace productivity.

MiniMaxMiniMax M
Open
Seedance 2.0

Unified text, image, audio, and video references and editing with 15-second multi-shot stereo output.

ByteDance SeedSeedance
Qwen-Image 2.0

Unified generation and editing with 1K-token complex prompts, native 2K, and more realistic detail.

Alibaba CloudQwen-Image
GPT-5.3-Codex

A Codex model extending agentic coding into research, tool use, and end-to-end computer work.

OpenAIGPT
400K ctx
Claude Opus 4.6

Brought a 1M context beta to Opus with stronger long-running agents, code review, and adaptive thinking.

AnthropicClaude
1M ctx
Kling AI 3.0

The native multimodal 3.0 line unifies image, video, and audio understanding, generation, and editing in narratives up to 15 seconds.

KuaishouKling
Kimi K2.5

A natively multimodal continuation of K2 with visual coding and an Agent Swarm of up to 100 subagents.

Moonshot AIKimi
MOE256K ctxOpen
Ray3.14

Brought Ray3 to native 1080p while substantially reducing latency and cost.

Luma AIRay
FLUX.2 [klein] 4B

A sub-second generation and editing model for consumer GPUs.

Black Forest LabsFLUX
HYBRID4BOpen
MiniMax-M2.1

Improved real-world coding, tool use, and agent reliability.

MiniMaxMiniMax M
MOE230BOpen
GLM-4.7

The open successor to GLM-4.5 with continued gains in coding, reasoning, and agent tasks.

Z.aiGLM
MOE200K ctxOpen
GPT-5.2-Codex

A GPT-5.2 variant optimized for long-horizon software engineering, large code changes, Windows, and context compaction.

OpenAIGPT
Seed1.8

A general agent model for search, code, GUI interaction, and complex workflows.

ByteDance SeedSeed
Gemini 3 Flash Preview

The low-latency Gemini 3 preview with stronger visual-spatial reasoning and agentic coding.

Google DeepMindGemini
1M ctx
MiMo-V2-Flash

Used hybrid sliding-window attention and MTP for faster long-context coding and agents.

Xiaomi MiMoMiMo
MOE309B256K ctxOpen
GPT Image 1.5

Improved precise editing, detail preservation, typography, and generation speed by roughly four times.

OpenAIGPT Image
Seedance 1.5 Pro

Added joint audio-video generation, multilingual dialogue, camera control, and fuller narrative expression.

ByteDance SeedSeedance
GPT-5.2

A GPT-5 series upgrade for professional work and long-running agents, offered in Instant, Thinking, and Pro tiers.

OpenAIGPT
DeepSeek-V3.2

A general model unifying non-thinking and thinking modes with sparse attention and everyday agent capability.

DeepSeekDeepSeek V / R
MOE671B128K ctxOpen
DeepSeek-V3.2-Speciale

A high-reasoning V3.2 variant for math, competitive programming, and longer deliberation.

DeepSeekDeepSeek V / R
MOE671B128K ctxOpen
Runway Gen-4.5

Runway’s current flagship for cinematic fidelity, complex sequenced prompts, and professional HDR output.

RunwayRunway Gen
Kling 2.6

Unified video generation, references, and editing across Kling O1/2.6 with native audio.

KuaishouKling
FLUX.2 [dev]

The second-generation visual family with multi-reference editing, precise color control, and stronger realism.

Black Forest LabsFLUX
HYBRID32BOpen
Claude Opus 4.5

Introduced effort control with stronger coding, agents, computer use, and multi-agent coordination.

AnthropicClaude
200K ctx
Nano Banana Pro

A professional Gemini 3 Pro image model with reasoning, real-world knowledge, multilingual text, and 4K control.

Google DeepMindNano Banana
Grok 4.1 Fast

A 2M-context fast agent model for real-world tool use and deep research, offered in reasoning and non-reasoning variants.

SpaceXAIGrok
2M ctx
Gemini 3 Pro Preview

The first Gemini 3 Pro preview for complex reasoning, coding, tools, and multimodal work.

Google DeepMindGemini
1M ctx
Grok 4.1

Used large-scale reinforcement learning to improve style, intent understanding, collaboration, and factuality in everyday conversation.

SpaceXAIGrok
GPT-5.1

An adaptive-reasoning API update that reduced latency on easy tasks and added apply_patch and shell tools.

OpenAIGPT
Kimi K2 Thinking

A K2-based thinking agent trained to reason while using tools natively.

Moonshot AIKimi
MOE1T256K ctxOpen
Hailuo 2.3

Improved body motion, micro-expressions, stylization, and motion prompts with a lower-cost Fast variant.

MiniMaxHailuo / H3
MiniMax-M2

Shifted MiniMax’s main line toward efficient coding and agent execution.

MiniMaxMiniMax M
MOE230BOpen
Claude Haiku 4.5

A low-latency Haiku model offering near-Sonnet-4 coding and computer-use capability at lower cost.

AnthropicClaude
200K ctx
Veo 3.1

Improved narrative control, reference consistency, and editing, adding 1080p, 4K, and vertical output.

Google DeepMindVeo / Gemini Video
Sora 2

Improved physics, control, and multi-shot state persistence with native dialogue and sound; the product ended on 2026-04-26.

OpenAISora
Claude Sonnet 4.5

Strengthened long-running coding, computer use, and complex agents alongside the Claude Agent SDK.

AnthropicClaude
200K ctx
HunyuanImage 3.0

A native multimodal MoE image model with 80B total and 13B active parameters.

Tencent HyHunyuan Image
MOE80BOpen
Qwen3-Max

A trillion-plus-parameter Qwen3 flagship for coding, agents, and million-token professional work.

Alibaba CloudQwen
MOE1T1M ctx
Grok 4 Fast

An efficient Grok 4 variant unifying reasoning and non-reasoning modes with 2M context and live search.

SpaceXAIGrok
2M ctx
Ray3

Added visual reasoning, self-evaluation, native HDR EXR, and draft-to-4K workflows.

Luma AIRay
Qwen3-Next-80B-A3B

An 80B-total, 3B-active architecture preview combining hybrid linear attention and ultra-sparse MoE.

Alibaba CloudQwen
HYBRID80B262K ctxOpen
Seedream 4.0

Unified text-to-image and editing, using a vision-language model for world knowledge and supporting up to 4K.

ByteDance SeedSeedream
Nano Banana

Applied Gemini 2.5 Flash knowledge and conversation to fast image generation and editing.

Google DeepMindNano Banana
GPT-5

GPT-5 unified fast responses, deeper reasoning, and automatic routing, with flagship, mini, and nano variants in the API.

OpenAIGPT
GPT-5 mini

The lower-cost GPT-5 tier for workloads balancing reasoning quality, throughput, and price.

OpenAIGPT
GPT-5 nano

The smallest GPT-5 tier for classification, extraction, ranking, and lightweight subtasks.

OpenAIGPT
Qwen-Image

A 20B MMDiT image foundation model focused on complex Chinese/English text rendering and precise editing.

Alibaba CloudQwen-Image
HYBRID20BOpen
GLM-4.5

A 355B-total, 32B-active hybrid-reasoning MoE unifying reasoning, coding, and agent capabilities.

Z.aiGLM
MOE355B128K ctxOpen
Wan 2.2

Introduced dual-expert MoE video diffusion with open 720p text, image, speech, and animation variants.

Alibaba CloudWan
MOE27BOpen
Kimi K2

A 1T-total, 32B-active open MoE focused on agentic tool use and coding.

Moonshot AIKimi
MOE1T128K ctxOpen
Grok 4

Scaled reinforcement learning beyond Grok 3 Reasoning with native tools, web search, and X search.

SpaceXAIGrok
Gemma 3n E4B

Used MatFormer and Per-Layer Embeddings to bring multimodal understanding to mobile devices.

Google DeepMindGemma
HYBRID8B32K ctxOpen
Seed1.6

Introduced adaptive reasoning on a 230B MoE foundation unifying text and vision.

ByteDance SeedSeed
MOE230B256K ctx
Midjourney Video V1

Animates one image into five seconds, extendable to 21 seconds with end frames, loops, and motion modes.

MidjourneyMidjourney
Hailuo 02

Advanced Hailuo with native 1080p, complex motion, and stronger physics.

MiniMaxHailuo / H3
MiniMax-M1

An open reasoning model with 1M context using hybrid linear and softmax attention.

MiniMaxMiniMax M
HYBRID456B1M ctxOpen
Seedance 1.0

Established Seedance with native multi-shot storytelling, 1080p, and stable complex motion.

ByteDance SeedSeedance
FLUX.1 Kontext [dev]

Unified generation and contextual editing with multi-turn changes, character consistency, and local control.

Black Forest LabsFLUX
HYBRID12BOpen
Claude Opus 4

The Claude 4 flagship, focused on long-running coding, complex agents, and sustained multi-hour execution.

AnthropicClaude
200K ctx
Claude Sonnet 4

The balanced Claude 4 tier with stronger coding, reasoning, instruction following, and parallel tool use.

AnthropicClaude
200K ctx
Imagen 4

Supports up to 2K and multiple aspect ratios with stronger detail, spelling, and typography.

Google DeepMindImagen
Veo 3

Added native dialogue, ambience, and sound effects to Veo.

Google DeepMindVeo / Gemini Video
MiMo-7B-RL

Xiaomi’s first open reasoning model, exploring the full pretraining-to-RL path at 7B scale.

Xiaomi MiMoMiMo
DENSE7BOpen
Qwen3-235B-A22B

The flagship Qwen3 MoE, combining thinking and non-thinking modes in one open model.

Alibaba CloudQwen
MOE235B128K ctxOpen
GPT Image 1

Brought GPT-4o-native image generation to the API with world knowledge, style adherence, and typography.

OpenAIGPT Image
Llama 4 Scout

A natively multimodal MoE with 17B active parameters, 16 experts, and up to 10M context.

MetaLlama
MOE109B10M ctxOpen
Llama 4 Maverick

The higher-capability Llama 4 open model with 400B total parameters and 128 experts.

MetaLlama
MOE400B1M ctxOpen
Midjourney V7

Added default personalization, Omni Reference, 10× draft mode, and conversational creation.

MidjourneyMidjourney
Runway Gen-4

Improved subject, location, and style consistency across scenes with more direct camera control.

RunwayRunway Gen
Gemini 2.5 Pro

Google built “thinking” directly into the general Gemini generation, with emphasis on complex reasoning and coding.

Google DeepMindGemini
1M ctx
Gemma 3 27B

Added image input, a 128K context window, and support for more than 140 languages.

Google DeepMindGemma
DENSE27B128K ctxOpen
Wan 2.1

Released open 1.3B and 14B text/image-to-video weights plus VACE general video editing.

Alibaba CloudWan
HYBRID14BOpen
Claude 3.7 Sonnet

Combined near-instant responses and controllable extended thinking in one model, alongside Claude Code.

AnthropicClaude
200K ctx
Grok 3

xAI’s reasoning-agent generation, extending complex-task capability through Think and DeepSearch modes.

SpaceXAIGrok
Gemini 2.0 Flash

A low-latency multimodal model that introduced native tool use and the Flash Thinking experimental path.

Google DeepMindGemini
1M ctx
DeepSeek-R1

Used large-scale reinforcement learning to shape reasoning and released both full and distilled model weights.

DeepSeekDeepSeek V / R
MOE671B128K ctxOpen
Kimi K1.5

Moonshot’s multimodal reasoning work exploring long-context reinforcement learning and test-time reasoning scale.

Moonshot AIKimi
Ray2

Luma’s second-generation large video model with stronger natural motion, physics, and camera expression.

Luma AIRay
DeepSeek-V3

A 671B-total, 37B-active MoE model that extended the efficient-training and open-release path.

DeepSeekDeepSeek V / R
MOE671B128K ctxOpen
Veo 2

Advanced Google’s video line with stronger physics, cinematic language, and research output up to 4K.

Google DeepMindVeo / Gemini Video
Sora Turbo

Sora’s first production release with up to 20 seconds, 1080p, extension, blending, and storyboards.

OpenAISora
Llama 3.3 70B

Brought near-Llama-3.1-405B instruction quality to a much smaller 70B model.

MetaLlama
DENSE70B128K ctxOpen
Hunyuan Large

Tencent’s first large open MoE with 389B total and 52B active parameters.

Tencent HyHunyuan / Hy
MOE389BOpen
Stable Diffusion 3.5 Large

A customizable open image tier spanning 8.1B Large, Turbo, and 2.5B Medium variants.

Stability AIStable Diffusion
HYBRID8.1BOpen
Llama 3.2 Vision

Added 11B and 90B vision models plus 1B and 3B edge-oriented variants.

MetaLlama
128K ctxOpen
Imagen 3

A major Google text-to-image generation with better detail, lighting, and prompt adherence.

Google DeepMindImagen
FLUX.1 [dev]

BFL’s first flow-matching image family across open developer, fast, and hosted professional variants.

Black Forest LabsFLUX
HYBRID12BOpen
Llama 3.1 405B

Llama’s first 405B open flagship, extending the family to 128K context.

MetaLlama
DENSE405B128K ctxOpen
Gemma 2 27B

Raised reasoning quality in 9B and 27B sizes designed for accessible accelerators.

Google DeepMindGemma
DENSE27B8K ctxOpen
Stable Diffusion 3 Medium

Introduced MMDiT in a 2B model aimed at consumer GPUs.

Stability AIStable Diffusion
HYBRID2BOpen
Qwen2-72B

The largest dense Qwen2 model, using GQA across the family and extending support to 27 additional languages.

Alibaba CloudQwen
DENSE72.7B128K ctxOpen
GPT-4o

An “omni” model natively spanning text, vision, and audio with substantially lower interaction latency.

OpenAIGPT
UNKNOWN128K ctx
DeepSeek-V2

Combined Multi-head Latent Attention with fine-grained MoE routing to reduce training and inference cost.

DeepSeekDeepSeek V / R
MOE236B128K ctxOpen
Llama 3

Opened Meta’s new open-model generation in 8B and 70B sizes.

MetaLlama
DENSE70B8K ctxOpen
Claude 3 Opus

The flagship Claude 3 model, combining visual input with a 200K context window.

AnthropicClaude
200K ctx
Gemma

Google’s first lightweight open models, released in 2B and 7B sizes.

Google DeepMindGemma
DENSE7B8K ctxOpen
Gemini 1.5 Pro

Used an MoE architecture to extend usable context to 1M tokens while retaining cross-modal retrieval.

Google DeepMindGemini
MOE1M ctx
Midjourney V6

Significantly improved long-prompt understanding, coherence, knowledge, and image prompting.

MidjourneyMidjourney
Gemini 1.0 Ultra

The first Gemini flagship, designed from training onward for native text, image, audio, and video understanding.

Google DeepMindGemini
DENSE
Qwen-7B

The starting point of Qwen’s open-model path, emphasizing Chinese-English, multilingual, and tool-use capabilities.

Alibaba CloudQwen
DENSE7.7B32K ctxOpen
SDXL 1.0

Used a base-plus-refiner pipeline to bring open generation to native 1024 resolution.

Stability AIStable Diffusion
HYBRID6.6BOpen
Claude 2

Expanded to a 100K context window and reached broader users through the API and claude.ai.

AnthropicClaude
100K ctx
GPT-4

A large multimodal model for complex professional tasks; core scale and training details were not disclosed.

OpenAIGPT
UNKNOWN8K ctx
Claude 1

Anthropic’s first public Claude model, centered on its Constitutional AI training approach.

AnthropicClaude
Stable Diffusion 1.4

Brought strong text-to-image weights to local GPUs and catalyzed a broad tooling and fine-tuning ecosystem.

Stability AIStable Diffusion
HYBRIDOpen
GPT-3

A 175B autoregressive model that systematically demonstrated few-shot and in-context learning without fine-tuning.

OpenAIGPT
DENSE175B2K ctx
GPT-2

Demonstrated broad task transfer from unsupervised language modeling and introduced a staged-weight release.

OpenAIGPT
DENSE1.5B1K ctxOpen