0
SIGNAL + NOISE / INDEX 013

Model Signal

Model releases, translated into operating decisions.

A recurring Signal + Noise series on frontier, open, multimodal, and specialist models — what changed, where each model fits, what breaks, and what teams should do next.

TRACKED101
LATESTGLM-5.3
UPDATEDAugust 14, 2026
LATEST MODEL SIGNAL·August 14, 2026·NEW

GLM-5.3

GLM-5.3 is a post-training update from Z.ai focused on long-horizon coding and advanced cybersecurity tasks.

Z.AICODE1M CONTEXT
AUGUST 2026 · 14 OF 101 TRACKED
Z.ai·August 14, 2026·NEW

GLM-5.3

GLM-5.3 is a post-training update from Z.ai focused on long-horizon coding and advanced cybersecurity tasks.

Code·1M
Qwen·August 13, 2026·NEW

Qwen3.8 27B

Qwen3.8-27B is a dense 27-billion-parameter native vision-language model that understands images and videos, with flexible thinking control, designed to carry complex, multi-step tasks through to completion with greater reliability.

Multimodal·262k
Qwen·August 13, 2026·NEW

Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95B is a sparse Mixture-of-Experts causal language model from Qwen with 2.4 trillion total parameters, 95 billion activated parameters per token, and open-weight availability.

Multimodal·1010000
Google DeepMind·August 13, 2026·NEW

Gemini 3.7 Flash

Gemini 3.7 Flash is a lightweight, high-speed model from Google DeepMind optimized for coding and agentic workflows at a lower price point.

Multimodal·1,048,576
Writer·August 13, 2026·NEW

Palmyra X6

Writer released Palmyra X6, an enterprise-focused general model optimized for agentic workflows and reduced operating costs.

General·1M tokens
DeepSeek·August 13, 2026·NEW

DeepSeek V4 Pro 0813

DeepSeek V4 Pro 0813 is the official GA release of DeepSeek V4 Pro, superseding the preview version and available on DeepSeek’s app, web, and API.

Multimodal·1M tokens
Pathway·August 11, 2026·NEW

BDH-CQ

Pathway's BDH-CQ is a 150-million-parameter reasoning model in a Post-Transformer architecture that scored 29.5% pass@2 on the public ARC-AGI-1 evaluation set at a computed inference cost of $0.0007 per task.

Reasoning
Meta·August 10, 2026·NEW

Muse Glimmer 30B

Muse Glimmer is a ~29.6B-parameter dense multimodal causal language model with a dedicated perception encoder, distilled from Muse Spark and purpose-built for autonomous agentic tasks and local agent workflows on consumer hardware, supporting multimodal understanding, tool use, and failure recovery without requiring cloud infrastructure or network access.

Multimodal·131,072+ tokens
Meta·August 10, 2026·NEW

Muse Glimmer

Muse Glimmer is a 30-billion-parameter open-weight model by Meta designed specifically for local agentic workflows.

General·30B parameters
Meta·August 5, 2026·NEW

Muse Spark 1.2

Muse Spark 1.2 is a natively multimodal reasoning model from Meta optimized for complex agentic tasks and coding workflows.

Multimodal·1M tokens
Meta·August 5, 2026·NEW

Muse Code

Meta has released a beta version of Muse Code, a terminal-based AI coding agent designed to autonomously navigate and modify large codebases.

Code
Liquid AI·August 4, 2026·NEW

LFM2.5-2.6B

A 2.6B dense model trained for agentic workloads, built for on-device deployment with native tool calling.

General·131072
Alibaba Qwen·August 3, 2026

Qwen3.8-Max

Qwen3.8-Max is Alibaba's frontier multimodal foundation model, offering a 1-million-token context window for complex reasoning and large-scale visual understanding tasks.

Multimodal·1,000,000 tokens
OpenAI·August 1, 2026

Astra

OpenAI has internally previewed Astra, a next major model family oriented toward deep, long‑horizon reasoning, which it used to generate advances on ten long‑standing problems in mathematics and theoretical computer science.

Reasoning
JULY 2026 · 24 OF 101 TRACKED
MiniMax·July 31, 2026

H3

MiniMax H3 is a general-purpose multimodal video-generation model that jointly understands text, images, video, and audio and generates up to 15‑second 2K video clips with native stereo (dual‑channel) audio.

Multimodal·multimodal (text, images, video, audio)
Google·July 30, 2026

Gemini Spark

Gemini Spark is an agentic AI model integrated directly into Google Chrome to autonomously execute long-running background web tasks.

General
Google DeepMind·July 30, 2026

Gemini Robotics 2

A vision-language-action model designed to provide whole-body control and physical reasoning capabilities for humanoid and general-purpose robots.

Multimodal
Thinking Machines·July 30, 2026

Inkling Small

An open-weights, 276-billion-parameter sparse Mixture-of-Experts multimodal transformer (276B total / 12B active) designed to deliver much of the flagship Inkling model’s reasoning, coding, and multimodal performance at lower compute cost and latency.

General·1000000
Google DeepMind·July 30, 2026

Gemini Robotics ER 2

Gemini Robotics ER 2 is Google DeepMind’s most capable embodied reasoning vision-language model, acting as a high-level brain for robots to communicate with humans, understand the physical world, plan multi-step tasks, and orchestrate tools and action models, including multi-robot collaboration.

Multimodal·128k
Qwen·July 27, 2026

Qwen3.7 Flash

Qwen3.7 Flash is a high-efficiency vision-language model from Alibaba featuring a 1M-token context window, optimized for multimodal agents, visual coding, and spatial understanding.

Multimodal·1,000,000 tokens
Anthropic·July 24, 2026

Claude Opus 5

Claude Opus 5 is Anthropic's step-change improvement over Claude Opus 4.8, designed for deep reasoning, agentic and long-horizon tasks, and test-time compute scaling.

Multimodal·1000000
Google·July 21, 2026

Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency multimodal model from Google optimized for coding, agentic workflows, and large-scale knowledge tasks.

Multimodal·1,048,576 tokens
Google·July 21, 2026

Gemini 3.6 Flash

Gemini 3.6 Flash is a high-efficiency, natively multimodal model optimized for fast execution, coding, and agentic workflows across large context lengths.

Multimodal·1000000
Poolside·July 21, 2026

Laguna S 2.1

Poolside has released Laguna S 2.1, a 118-billion-parameter open-weight Mixture-of-Experts model optimized for agentic coding and software development tasks.

Code·1M tokens
Google·July 21, 2026

Gemini 3.5 Flash Cyber

Gemini 3.5 Flash Cyber is a lightweight, domain-specific model fine-tuned from the Flash architecture to identify, validate, and patch software vulnerabilities.

Code
Google·July 21, 2026

Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-throughput, low-cost execution for subagent tasks and document parsing.

Multimodal·1,048,576 input tokens
Google·July 21, 2026

Gemini 3.6 Flash

Gemini 3.6 Flash is a workhorse, frontier‑level multimodal model that delivers better coding, knowledge work, and token‑efficient performance than Gemini 3.5 Flash, optimized for real‑world agentic and everyday tasks at higher speed and lower cost.

General·Up to 1,000,000 input tokens and up to 64,000 output tokens
Alibaba·July 17, 2026

Qwen3.8-Max

Alibaba's 2.4-trillion-parameter flagship model, released as Qwen3.8-Max-Preview.

General
Moonshot AI·July 16, 2026

Kimi K3

Moonshot AI has released Kimi K3, a 2.8-trillion parameter open-weight model that achieves frontier-level coding capabilities and outperformed Claude Fable 5 on the Frontend Code Arena.

General·1000000
OpenAI·July 9, 2026

GPT-5.6 Sol

GPT-5.6 Sol is OpenAI's flagship frontier model in the 5.6 series, optimized for complex reasoning, multi-step coding, and agentic workflows.

Multimodal·1050000
OpenAI·July 9, 2026

GPT-5.6 Luna

GPT-5.6 Luna is OpenAI’s fastest and most affordable GPT-5.6 tier, positioned for low-latency, cost-sensitive use cases.

Multimodal·1M tokens
OpenAI·July 9, 2026

ChatGPT Work

ChatGPT Work is an agent inside ChatGPT, powered by GPT‑5.6 and Codex, that can gather context across your apps, files, and workflows on web, mobile, and desktop to turn high‑level goals into finished work artifacts such as documents, spreadsheets, presentations, reports, and web applications, staying with complex multi‑step projects for hours.

General·GPT‑5.6 Sol, Terra, and Luna are available in different ways across plans: Free and Go users access GPT‑5.6 Terra in Cha
Meta·July 9, 2026

Muse Spark 1.1

Meta's Muse Spark 1.1 is an API-accessible, agentic coding model featuring a 1-million-token context window and optimized for automated tool and computer use.

Code
OpenAI·July 9, 2026

GPT-5.6 Sol

GPT-5.6 Sol is OpenAI's flagship frontier model, offering top-tier capabilities in complex coding, scientific reasoning, and cybersecurity.

General
OpenAI·July 9, 2026

GPT-5.6

OpenAI's GPT-5.6 is a major frontier model family featuring three tiers—Sol, Terra, and Luna—designed to deliver improved intelligence and cost efficiency across enterprise workloads.

General
xAI·July 8, 2026

Grok 4.5

Grok 4.5 is xAI's flagship multimodal model, featuring a 500,000-token context window and advanced capabilities in coding, STEM, and general knowledge work.

Multimodal·500000
OpenAI·July 8, 2026

GPT-Live

OpenAI's GPT-Live is a family of full-duplex voice models designed for low-latency, real-time human-AI interaction.

Multimodal
Anthropic·July 1, 2026

Claude Sonnet 5

Anthropic's Claude Sonnet 5 is a Sonnet-tier model with a 1M-token context window, 128k max output tokens, adaptive thinking, and standard pricing of $3 per million input tokens and $15 per million output tokens; it launched with introductory pricing of $2/$10 through August 31, 2026 and became the default model for Free and Pro plans.

General·1M tokens
JUNE 2026 · 17 OF 101 TRACKED
Google·June 30, 2026

TabFM

Google's TabFM is a zero-shot foundation model capable of performing classification and regression on tabular data without requiring per-dataset training.

General·N/A
Google·June 30, 2026

Nano Banana 2 Lite

A fast, cost-efficient Gemini Image model from Google (Gemini 3.1 Flash-Lite Image / Nano Banana 2 Lite) designed for high-throughput text-to-image generation and editing with latency of roughly four seconds per image.

Image
Base44·June 29, 2026

Base 1

Base 1 is Base44’s first proprietary large language model, fine-tuned on an open‑source base and trained on tens of millions of real user interactions, purpose‑built to create and edit web applications inside the Base44 vibe‑coding platform.

Code
OpenAI·June 26, 2026

GPT-5.6 Terra

GPT-5.6 Terra is the balanced, mid‑tier model in OpenAI’s GPT‑5.6 family, designed to provide strong reasoning and tool‑use capabilities with lower cost and high scalability, and it supports up to a 1M‑token context window for large‑scale applications.

Multimodal·1000000
OpenAI·June 26, 2026

GPT-5.6 Sol

GPT-5.6 Sol is OpenAI's flagship frontier model in the GPT-5.6 family, built for complex professional work such as deep reasoning, advanced coding, and long‑horizon agentic workflows, with multimodal (text + vision) support.

Multimodal·1050000
OpenAI·June 26, 2026

GPT-5.6 Terra

GPT-5.6 Terra is OpenAI's mid-tier multimodal model, positioned between the flagship Sol and efficient Luna tiers to provide a balance of reasoning performance and cost-efficiency.

Multimodal·~1.05M tokens (often rounded as a 1M-token context window in OpenAI product messaging)
Liquid AI·June 25, 2026

LFM2.5-230M

Liquid AI has released LFM2.5-230M, a 230-million-parameter open-weight language model built to run anywhere (CPUs, GPUs, NPUs) for on-device and edge AI workloads with strong performance on tool use and structured data extraction.

General·32000
Mistral·June 23, 2026

OCR 4

A specialized document extraction model providing structured outputs with layout analysis and confidence scoring across 170 languages.

Multimodal
Sakana AI·June 21, 2026

Sakana Fugu

Sakana Fugu is an orchestration model that functions as a multi-agent system by dynamically routing tasks across a swappable pool of underlying large language models.

General
WeiboAI·June 17, 2026

VibeThinker-3B

VibeThinker-3B is a 3-billion-parameter open-weights reasoning model developed by WeiboAI.

Reasoning
Z.ai·June 13, 2026

GLM-5.2

GLM-5.2 is a 753-billion parameter open-weights large language model from Z.ai that is specifically optimized for long-horizon coding tasks.

Code·1000000
Avataar AI·June 11, 2026

Varya

An open-weight video generation model optimized for Indian cultural contexts and highly cost-efficient inference.

Multimodal
Anthropic·June 9, 2026

Claude Fable 5

Claude Fable 5 is Anthropic’s most capable generally available Mythos‑class model, designed for long‑horizon agentic knowledge work and coding, with multimodal inputs and a 1‑million‑token context window.

Multimodal·1000000
Anthropic·June 9, 2026

Claude Mythos 5

Claude Mythos 5 is the same model as Claude Fable 5 but with cyber safeguards lifted, and it initially became available to existing Claude Mythos Preview users such as cybersecurity partners in Project Glasswing. Anthropic suspended access to both models on June 12, 2026 after a U.S. government export-control directive, then restored access to Mythos 5 for a set of U.S. organizations after government approval.

Reasoning
Anthropic·June 9, 2026

Claude Fable 5

Anthropic’s most capable publicly available model and its first Mythos‑class model made safe for general use, released for global access across the Claude API, Claude Platform, Claude Code, Claude Cowork, and major cloud partners.

Reasoning·1M
NVIDIA·June 4, 2026

Nemotron-3-Ultra-550B-A55B-NVFP4

A frontier-scale 550-billion parameter hybrid model from NVIDIA featuring a 1-million token context window and NVFP4 precision.

General·1000000
Alibaba Qwen·June 1, 2026

Qwen3.7-Plus

Qwen3.7-Plus is a mid-tier, multimodal model from Alibaba featuring a one-million token context window, optimized for cost-effective reasoning and agentic workflows.

Multimodal·1000000 tokens
MAY 2026 · 7 OF 101 TRACKED
Anthropic·May 28, 2026

Claude Opus 4.8

Claude Opus 4.8 is Anthropic's generally available Opus model, announced and released on May 28, 2026.

Multimodal·1000000
Google·May 28, 2026

Gemini 3.1 Flash Image

Gemini 3.1 Flash Image (Nano Banana 2) is a high-efficiency image generation and conversational editing model from Google, optimized for speed, low latency, and high-volume developer use cases.

Image·131072
Anthropic·May 28, 2026

Claude Opus 4.8

Builds on Opus 4.7 with stronger honesty and reliability, improved long-horizon agentic coding and professional work, dynamic workflows in Claude Code that orchestrate hundreds of parallel subagents, support for mid-conversation system messages, and a 2.5x fast mode at unchanged base pricing but much cheaper than prior fast-mode offerings.

Reasoning·1M
Alibaba·May 22, 2026

Qwen3.7 Max

A frontier-discount agent substrate from Alibaba — strong agentic index, 1M context, and prompt-cache economics that change what becomes economical to automate.

General·1M
xAI·May 20, 2026

Grok Build 0.1

Grok Build 0.1 is xAI’s fast coding model trained specifically for agentic software engineering workflows and interactive coding agents, supporting text and image inputs with text output and a 256K-token context window.

Code·256K tokens
Cohere·May 20, 2026

Command A+

Cohere’s open-weight sparse Mixture-of-Experts model built for enterprise agentic workloads, combining text and vision inputs, multilingual support, tool use, and complex reasoning within a 128K context for sovereign, privately deployable AI.

General·128K input, 64K output
Google·May 19, 2026

Gemini 3.5 Flash

Gemini 3.5 Flash is Google's high-efficiency multimodal model optimized for fast, cost-effective coding and parallel agentic execution.

Multimodal·1048576
APRIL 2026 · 10 OF 101 TRACKED
DeepSeek·April 24, 2026

DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is a specific revision of DeepSeek's efficient mixture-of-experts model, updated via re-post-training for enhanced coding, reasoning, and agent workflows.

General·1,000,000 tokens
DeepSeek·April 24, 2026

DeepSeek V4 Flash

Efficiency-focused Mixture-of-Experts variant in the DeepSeek V4 series, optimized for fast, cost-efficient inference while preserving strong reasoning and coding performance with a 1M-token context window.

Reasoning·1,000,000
OpenAI·April 23, 2026

GPT-5.5

OpenAI's latest flagship with noticeably stronger reasoning and autonomy than GPT-5.4. Available in standard and Pro variants. 1M context window, $5/$30 per 1M tokens for standard.

Reasoning·1M
Qwen·April 22, 2026

Qwen3.6-27B

A 27-billion-parameter dense multimodal (image-text-to-text) model with native 262K token context, released as the first open-weight Qwen3.6 variant and designed for strong coding and agentic/repository-level reasoning.

Multimodal·262144
Qwen·April 16, 2026

Qwen3.6-35B-A3B

An open-weight sparse Mixture-of-Experts multimodal model from Qwen with 35 billion total parameters and about 3 billion active per token, combining a causal language model with a vision encoder and optimized for long-context, agentic coding and reasoning.

Multimodal·262144
Z.ai·April 7, 2026

GLM-5.1

Open-weights model from Z.ai (Zhipu AI) with strong web-dev coding performance. Ranks above Kimi K2.6 on Code Arena WebDev leaderboard (1,534 Elo). Competitive on multilingual tasks.

Code·200K
Google·April 3, 2026

gemma-4-26B-A4B-it

A 26-billion-parameter, instruction-tuned Mixture-of-Experts model from Google supporting multimodal inputs and a 256K-token context window.

Multimodal·256K
Google·April 2, 2026

gemma-4-31B-it

An instruction-tuned ~31-billion parameter dense open-weights model from Google's Gemma 4 (Gemma 4 31B IT) family, offering multimodal input support for text, image, and video (via frames) and a 256K-token context window.

Multimodal·256K
Moonshot AI·April 1, 2026

Kimi K2.6

1 trillion-parameter vision-language model from Moonshot AI. Highest-ranked open-weights model on Artificial Analysis leaderboard (score 54). Designed for long-horizon agentic coding with plan-write-test-debug loops lasting days.

Code·128K
DeepSeek·April 1, 2026

DeepSeek V4

DeepSeek's latest open-source flagship. Rivals closed frontier models on coding and reasoning benchmarks while maintaining extremely low inference cost. MoE architecture.

Code·128K
MARCH 2026 · 4 OF 101 TRACKED
FEBRUARY 2026 · 5 OF 101 TRACKED
DECEMBER 2025 · 1 OF 101 TRACKED
SEPTEMBER 2025 · 1 OF 101 TRACKED
AUGUST 2025 · 2 OF 101 TRACKED
JULY 2025 · 3 OF 101 TRACKED
MAY 2025 · 2 OF 101 TRACKED
APRIL 2025 · 5 OF 101 TRACKED
FEBRUARY 2025 · 2 OF 101 TRACKED
OCTOBER 2024 · 1 OF 101 TRACKED
SEPTEMBER 2024 · 1 OF 101 TRACKED
DATE PENDING · 2 OF 101 TRACKED
About Model Signal

Model Signal tracks AI model releases across every provider. Each model has a compounding Model Signal report — a synthesized brief that auto-updates as the Signal + Noise wire mentions it. It covers releases through an operator lens: who should care, where each model fits, what changes, and what does not.