SYS://REGISTRY|

The chronological release log of major foundation model checkpoints and research breakthroughs.

September 2026[4 releases]

Google DeepMind/Gemini 3.8 Flash
Sep 2, 2026

Released Sept 2, 2026; Google DeepMind's premier frontier workhorse powering advanced agentic coding loops, recursive self-correction, and 1M token real-time multimodal streaming.

1,048,576 tokensNEWEST FLASH SOTA
INSPECT SPEC
Anthropic/Claude Fable 5.1
Sep 1, 2026

Released Sept 1, 2026; Anthropic's reigning flagship for agentic coding and knowledge work. Features a 75% reduction in cache read costs ($0.25/M tokens) and 52.6% on Terminal-Bench-Science.

1,000,000 tokensFRONTIER FLAGSHIP #1
INSPECT SPEC
OpenAI/OpenAI Astra
Sep 1, 2026

Announced Sept 1, 2026; the first LLM to meet OpenAI's 'Critical' cybersecurity threshold under its Preparedness Framework, capable of identifying zero-day exploits.

1,050,000 tokensCRITICAL CYBERSECURITY
INSPECT SPEC
Meta AI/Muse Voice Transcribe
Sep 1, 2026

Announced Sept 1, 2026; streaming speech-to-text foundation model supporting 70+ languages, 20+ voice diarization, and multilingual code-switching.

128,000 tokensNEW AUDIO CHECKPOINT
INSPECT SPEC

August 2026[6 releases]

DeepSeek/DeepSeek V4 Flash Vision Exp
Aug 30, 2026

Open-sourced under MIT license on August 30, 2026. 305B parameter multimodal vision-language model with native document understanding.

262,144 tokensOPEN WEIGHTS (MIT)
INSPECT SPEC
Alibaba Cloud (Qwen)/Qwen3.8 Flash
Aug 26, 2026

Released August 26, 2026; combines visual document understanding, fast agentic workflows, and 1M context with ultra-cheap $0.15/$0.47 pricing.

1,000,000 tokens1M MULTIMODAL VALUE
INSPECT SPEC
DeepSeek/DeepSeek V4-Pro (0813)
Aug 13, 2026

DeepSeek's primary 1.6T parameter powerhouse with 49B activated per token, configurable thinking budget, and 1M context.

1,048,576 tokens1.6T HOSTED SOTA
INSPECT SPEC
xAI/Grok 4.6
Aug 12, 2026

Released August 12, 2026; xAI's smartest model with frontier performance in coding and autonomous agents, integrated natively into Cursor and Grok Build.

500,000 tokens1.5T SUPERCOMPUTE
INSPECT SPEC
Alibaba Cloud (Qwen)/Qwen3.8 2.4T A95B
Aug 12, 2026

The largest open-weight MoE model in existence. 2.4 Trillion parameters with 95B activated per token and 1M context, available on HuggingFace and ModelScope.

1,000,000 tokens2.4T OPEN TITAN
INSPECT SPEC
OpenAI/GPT-5.6 Sol
Aug 5, 2026

OpenAI's primary general-purpose flagship model. Combines native fast execution, multi-agent orchestration, and 1.05M context for production software development.

1,050,000 tokensPRIMARY FLAGSHIP
INSPECT SPEC

July 2026[2 releases]

OpenAI/GPT-5.6 Luna
Jul 28, 2026

High-speed developer model designed for streaming code completions, real-time function calling, and high-frequency queries at $0.20/1M tokens.

256,000 tokensHIGH THROUGHPUT
INSPECT SPEC
Anthropic/Claude Opus 5
Jul 15, 2026

Deep multi-hour research, complex code synthesis, and long-horizon workflow orchestration.

1,000,000 tokensHEAVYWEIGHT AGENTIC
INSPECT SPEC

June 2026[1 release]

Cohere/North Mini Code
Jun 17, 2026

30B MoE agentic coding model specifically tuned for multi-file workspace inspection and test suites, offered free for developers.

256,000 tokensFREE AGENTIC CODE
INSPECT SPEC

May 2026[1 release]

Meta AI/Llama 4 Scout (16E)
May 10, 2026

Extended-context open-weights model capable of ingesting 1.31 million tokens in a single prompt for codebase-wide document analysis.

1,310,720 tokensLONG CONTEXT MOE
INSPECT SPEC

April 2026[2 releases]

Mistral AI/Mistral Medium 3.5
Apr 30, 2026

Dense 128B multimodal instruction-following model with native text and image understanding, tuned for European enterprise compliance.

262,144 tokensEUROPEAN FLAGSHIP
INSPECT SPEC
Meta AI/Llama 4 Maverick (128E)
Apr 18, 2026

Meta's flagship open-weights foundation model. 128-expert MoE architecture with 1M token context, powering enterprise on-premise deployments and community fine-tuning.

1,048,576 tokensOPEN WEIGHTS SOTA
INSPECT SPEC

March 2026[1 release]

Cohere/Command A (111B)
Mar 20, 2026

Cohere's primary flagship model optimized for business agentic search, long-document question answering, and multi-step tool use.

256,000 tokensENTERPRISE FLAGSHIP
INSPECT SPEC