Capital & Compute
Release record· Updated July 29, 2026· 20 releases

AI model releases, month by month

A dated record of every AI model release we can verify against the lab that shipped it, grouped by the month it happened. 20 releases so far, each with the price it launched at, the context window, whether the weights are open, and a link to the announcement. For what these models cost today and what is expected next, see themodel tracker; for how they rank on independent benchmarks, thevalue leaderboard.

How many AI models were released in July 2026?

20 distinct model releases in July 2026 are verifiable against a dated source, from 13 labs. 4 shipped with open weights. The flagship launches were Claude Opus 5, Kimi K3, GPT-5.6 Sol, Grok 4.5. That number is smaller than the counts on auto-ingested catalogs, which list every fine-tune and re-hosted variant; counted here is a distinct model a lab announced as a release.

What is the most recent AI model release?

Claude Opus 5 from Anthropic, released Jul 24, 2026. Took the top spot on the Artificial Analysis Intelligence Index at 61, and did it at half the per-token price of Claude Fable 5. It launched at $5 / $25 per million tokens.

20
Releases in July
from 13 labs
5
Busiest single day
July 21
$0.1
Cheapest launch, input
Laguna S 2.1 per Mtok
61
Top intelligence score
Claude Opus 5
AI model releases in July 2026, by dayA stacked column chart of the 20 model releases in July 2026, one cell per release, positioned on the day it shipped and coloured by whether it was a flagship, notable or incremental launch. The busiest day was July 21 with 5 releases.012345July 7: Cohere Transcribe ArabicJuly 8: Grok 4.5July 8: SWE-1.7July 9: GPT-5.6 SolJuly 9: GPT-5.6 LunaJuly 9: GPT-5.6 TerraJuly 9: Muse Spark 1.1July 15: InklingJuly 16: Kimi K3July 19: Qwen3.8-Max-PreviewJuly 20: Qwen-Audio-3.0-TTS PlusJuly 20: Qwen-Audio-3.0-TTS FlashJuly 21: Gemini 3.6 FlashJuly 21: Laguna S 2.1July 21: Gemini 3.5 Flash CyberJuly 21: Gemini 3.5 Flash-LiteJuly 21: Qwen-Image-3.0July 23: FLUX 3July 23: Ling-3.0-flashJuly 24: Claude Opus 517891516192021232431FlagshipNotableIncremental
AI model releases in July 2026, by day
Day of July 2026ReleasesModels
71Cohere Transcribe Arabic
82Grok 4.5, SWE-1.7
94GPT-5.6 Sol, GPT-5.6 Luna, GPT-5.6 Terra, Muse Spark 1.1
151Inkling
161Kimi K3
191Qwen3.8-Max-Preview
202Qwen-Audio-3.0-TTS Plus, Qwen-Audio-3.0-TTS Flash
215Gemini 3.6 Flash, Laguna S 2.1, Gemini 3.5 Flash Cyber, Gemini 3.5 Flash-Lite, Qwen-Image-3.0
232FLUX 3, Ling-3.0-flash
241Claude Opus 5
Model releases in July 2026, by the day they shippedSource: Capital & Compute, from lab announcements. Verified July 2026.

Releases do not arrive evenly. July 2026 produced 20 of them across just 10 separate days, and 5 of those landed on July 21 alone. The pattern is competitive rather than coincidental: labs hold a launch until a rival moves, then answer within hours, which is why the calendar above is a run of spikes separated by quiet weeks rather than a steady drip.

Launch price per million output tokens, July 2026 releasesA ranked ladder of the 10 July 2026 model releases that shipped with a published per-token rate card, cheapest output price first, from $0.2 to $30 per million output tokens. Flagship launches are highlighted.$0$10$20$30Laguna S 2.1$0.2Gemini 3.5 Flash-Lite$2.5Inkling$4.68GPT-5.6 Luna$6Grok 4.5$6Gemini 3.6 Flash$7.5Kimi K3$15GPT-5.6 Terra$15Claude Opus 5$25GPT-5.6 Sol$30
Launch price per million output tokens, July 2026 releases
ItemValue
Laguna S 2.1$0.2
Gemini 3.5 Flash-Lite$2.5
Inkling$4.68
GPT-5.6 Luna$6
Grok 4.5$6
Gemini 3.6 Flash$7.5
Kimi K3$15
GPT-5.6 Terra$15
Claude Opus 5$25
GPT-5.6 Sol$30
What each July model cost on the day it launched, output tokensSource: Capital & Compute, from each lab's published rate card. Verified July 2026.

The spread inside one month is about 150x, from Laguna S 2.1 at $0.2 per million output tokens to GPT-5.6 Sol at $30. Read that as a range of positions, not a quality ranking: the cheapest per-token model is frequently not the cheapest way to finish a task, because a model that reasons for four times as many tokens erases a four-times-lower rate. The cost-per-task ranking works the figure the other way round, from a completed task backwards.

Why do AI release trackers disagree on dates?

Because most of them ingest an API catalog instead of reading an announcement, so they record when a model appeared in a listing rather than when the lab shipped it. Two live examples from July 2026 alone. Grok STT 1.0 is dated July 23, 2026 on one major timeline; xAI actually released it on March 16, 2026, and July 23 is the day it was added to OpenRouter. Kimi K3 is dated July 14 by one tracker and July 16 by two others; Moonshot unveiled it on July 16. Every date below is read off the lab announcement, and the 4 rows that rest on reporting are marked as such.

AI models released in July 2026

20 releases from 13 labs.

DateModelLaunch pricein / out per MtokContextSource
Jul 24Claude Opus 5AnthropicTook the top spot on the Artificial Analysis Intelligence Index at 61, and did it at half the per-token price of Claude Fable 5.$5 / $251MOfficial
Jul 23FLUX 3Black Forest LabsOne network spanning image, video, audio and robot action prediction. Video runs in early access; the open-weight Dev build is promised later in 2026.Official
Jul 23Ling-3.0-flashAnt GroupA 124B mixture-of-experts firing only 5.1B parameters per token, which Ant says matches its own 1T flagship. Announced as open weight, but shipped API-only with the weights still unpublished.256KOfficial
Jul 21Gemini 3.6 FlashGoogleGoogle's workhorse tier, cut to $7.50 output from the $9.00 that Gemini 3.5 Flash charged, on a claimed 17% drop in output tokens per task.$1.5 / $7.51.05MOfficial
Jul 21Laguna S 2.1poolside · open weight (OpenMDW-1.1)118B total and 8B active, released under OpenMDW-1.1 at $0.10 in and $0.20 out, the cheapest agentic coding model to ship in July.$0.1 / $0.21MOfficial
Jul 21Gemini 3.5 Flash CyberGoogleTuned to find and patch software vulnerabilities, and restricted to governments and trusted partners through the CodeMender pilot.Official
Jul 21Gemini 3.5 Flash-LiteGoogleThe cheapest launch price of any July model at $0.30 in, aimed at classification, extraction and routing rather than reasoning.$0.3 / $2.51.05MOfficial
Jul 21Qwen-Image-3.0AlibabaClosed and invite-only at launch, with no model card, weights or published benchmarks to check the multi-panel layout claims against.Reporting
Jul 20Qwen-Audio-3.0-TTS PlusAlibabaTook first place on the Artificial Analysis text-to-speech arena. Billed per character, not per token, at roughly a third of what ElevenLabs charges.Reporting
Jul 20Qwen-Audio-3.0-TTS FlashAlibabaThe real-time tier of the same text-to-speech release, covering 16 languages and 20 Chinese dialect regions, hosted only on Alibaba Cloud Model Studio.Reporting
Jul 19Qwen3.8-Max-PreviewAlibabaA 2.4T-parameter preview shown at the World AI Conference with no model card, no license and no independent benchmarks published alongside it.1MReporting
Jul 16Kimi K3Moonshot · open weight2.8T parameters and the highest-scoring open-weight model of the month. Weights followed on July 26, a day inside Moonshot's own deadline.$3 / $151.05MOfficial
Jul 15InklingThinking Machines · open weight (Apache 2.0)975B total and 41B active under Apache 2.0, the most permissive license anyone has attached to a model at that parameter scale.$1.87 / $4.681MOfficial
Jul 9GPT-5.6 SolOpenAIThe reasoning tier of the GPT-5.6 line and the most expensive output token of any July release at $30 per million.$5 / $301.05MOfficial
Jul 9GPT-5.6 LunaOpenAIThe high-volume tier at $1 in and $6 out, and the best intelligence-per-dollar of anything OpenAI shipped in July.$1 / $61.05MOfficial
Jul 9GPT-5.6 TerraOpenAIThe middle tier, priced at half of Sol while scoring within four points of it on the Artificial Analysis Intelligence Index.$2.5 / $151.05MOfficial
Jul 9Muse Spark 1.1MetaMeta Superintelligence Labs shipped its first paid model, breaking the open-weight-only posture Meta had held since Llama.1MOfficial
Jul 8Grok 4.5xAITrained on real Cursor session data and priced at $2 in and $6 out, well under half of what Opus 4.8 and GPT-5.5 charged at the time.$2 / $6500KOfficial
Jul 8SWE-1.7CognitionPost-trained on top of Moonshot's already RL-heavy Kimi K2.7 Code, and served inside Devin only. Not sold as a standalone API.Official
Jul 7Cohere Transcribe ArabicCohere · open weight (Apache 2.0)A 2B open-weight speech recognition model built for Arabic dialect variation and Arabic-English code-switching.Official

Filter every tracked release

Search by model, lab or description, then narrow by category or lab. Price is the rate the model launched at, not necessarily what it costs today.

20 of 20
Sort by
Filter
ReleasedModelLaunch pricein / out per MtokAA IndexContextSource
Jul 24, 2026Claude Opus 5 Anthropic Took the top spot on the Artificial Analysis Intelligence Index at 61, and did it at half the per-token price of Claude Fable 5.$5 / $25611MOfficial
Jul 23, 2026FLUX 3 Black Forest Labs One network spanning image, video, audio and robot action prediction. Video runs in early access; the open-weight Dev build is promised later in 2026.Official
Jul 23, 2026Ling-3.0-flash Ant Group A 124B mixture-of-experts firing only 5.1B parameters per token, which Ant says matches its own 1T flagship. Announced as open weight, but shipped API-only with the weights still unpublished.256KOfficial
Jul 21, 2026Gemini 3.6 Flash Google Google's workhorse tier, cut to $7.50 output from the $9.00 that Gemini 3.5 Flash charged, on a claimed 17% drop in output tokens per task.$1.5 / $7.51.05MOfficial
Jul 21, 2026Laguna S 2.1 poolside· open weight(OpenMDW-1.1) 118B total and 8B active, released under OpenMDW-1.1 at $0.10 in and $0.20 out, the cheapest agentic coding model to ship in July.$0.1 / $0.21MOfficial
Jul 21, 2026Gemini 3.5 Flash Cyber Google Tuned to find and patch software vulnerabilities, and restricted to governments and trusted partners through the CodeMender pilot.Official
Jul 21, 2026Gemini 3.5 Flash-Lite Google The cheapest launch price of any July model at $0.30 in, aimed at classification, extraction and routing rather than reasoning.$0.3 / $2.51.05MOfficial
Jul 21, 2026Qwen-Image-3.0 Alibaba Closed and invite-only at launch, with no model card, weights or published benchmarks to check the multi-panel layout claims against.Reporting
Jul 20, 2026Qwen-Audio-3.0-TTS Plus Alibaba Took first place on the Artificial Analysis text-to-speech arena. Billed per character, not per token, at roughly a third of what ElevenLabs charges.Reporting
Jul 20, 2026Qwen-Audio-3.0-TTS Flash Alibaba The real-time tier of the same text-to-speech release, covering 16 languages and 20 Chinese dialect regions, hosted only on Alibaba Cloud Model Studio.Reporting
Jul 19, 2026Qwen3.8-Max-Preview Alibaba A 2.4T-parameter preview shown at the World AI Conference with no model card, no license and no independent benchmarks published alongside it.1MReporting
Jul 16, 2026Kimi K3 Moonshot· open weight 2.8T parameters and the highest-scoring open-weight model of the month. Weights followed on July 26, a day inside Moonshot's own deadline.$3 / $15571.05MOfficial
Jul 15, 2026Inkling Thinking Machines· open weight(Apache 2.0) 975B total and 41B active under Apache 2.0, the most permissive license anyone has attached to a model at that parameter scale.$1.87 / $4.68411MOfficial
Jul 9, 2026GPT-5.6 Sol OpenAI The reasoning tier of the GPT-5.6 line and the most expensive output token of any July release at $30 per million.$5 / $30591.05MOfficial
Jul 9, 2026GPT-5.6 Luna OpenAI The high-volume tier at $1 in and $6 out, and the best intelligence-per-dollar of anything OpenAI shipped in July.$1 / $6511.05MOfficial
Jul 9, 2026GPT-5.6 Terra OpenAI The middle tier, priced at half of Sol while scoring within four points of it on the Artificial Analysis Intelligence Index.$2.5 / $15551.05MOfficial
Jul 9, 2026Muse Spark 1.1 Meta Meta Superintelligence Labs shipped its first paid model, breaking the open-weight-only posture Meta had held since Llama.1MOfficial
Jul 8, 2026Grok 4.5 xAI Trained on real Cursor session data and priced at $2 in and $6 out, well under half of what Opus 4.8 and GPT-5.5 charged at the time.$2 / $654500KOfficial
Jul 8, 2026SWE-1.7 Cognition Post-trained on top of Moonshot's already RL-heavy Kimi K2.7 Code, and served inside Devin only. Not sold as a standalone API.Official
Jul 7, 2026Cohere Transcribe Arabic Cohere· open weight(Apache 2.0) A 2B open-weight speech recognition model built for Arabic dialect variation and Arabic-English code-switching.Official

Launch price is the standard non-batch rate the lab published on release day, in USD per million tokens. A dash means the model has no per-token rate card: speech and image models are billed per character or per generation, and some models ship only inside a product. AA Index is the Artificial Analysis Intelligence Index at or near launch. Source links to the lab announcement, or is marked Reporting where no primary source exists.

What the July record actually shows

Three things stand out once the month is laid out chronologically rather than as a leaderboard.

The open-weight gap closed at the top, not the bottom. 4 of 20 releases shipped with open weights, but the notable one is Kimi K3: an open-weight model scoring 57 on the Artificial Analysis Intelligence Index, within four points of the best proprietary model released the same month. A year ago the open-weight frontier trailed by a wide margin. It now trails by a rounding error at the top and wins outright on price.

Price moved down while capability moved up. Claude Opus 5 took the top intelligence score at 61 while charging $5 and $25 per million tokens, half of what Claude Fable 5 charges. Google cut its workhorse output rate from $9.00 to $7.50 and claims a 17% reduction in output tokens per task on top, which compounds into a real bill reduction rather than a headline one. See the model comparison tool to put any two of these against the same task.

Most of the month was not frontier models. Speech, image, video and security-tuned models made up a large share of the releases. That is the shape of a maturing market: the frontier gets a few launches a month, and the volume moves to specialised models that do one job cheaply. Several of them cannot be priced per token at all, which is why those rows carry a dash rather than a number.

For the narrative version of this month, with the benchmark detail and what it means for what you pay, read the July 2026 model roundup. For which of these you can call at no cost, see free AI models, and for what the benchmark names in each announcement actually measure, the AI benchmark directory.

Frequently asked questions

How many AI models were released in July 2026?
20 model releases in July 2026 are verifiable against a dated source, from 13 different labs. That count is deliberately narrower than the auto-ingested catalogs, which list every fine-tune, quantization and re-hosted variant and run to dozens of rows a month. Counted here are distinct models a lab announced as a release: Claude Opus 5, Kimi K3, GPT-5.6 Sol, Grok 4.5 were the flagship launches, and 4 of the 20 shipped with open weights.
What was the most recent AI model release?
Claude Opus 5 from Anthropic, released Jul 24, 2026. Took the top spot on the Artificial Analysis Intelligence Index at 61, and did it at half the per-token price of Claude Fable 5.
Why do AI model release trackers disagree on dates?
Because most of them ingest a catalog rather than read an announcement, so they record the date a model appeared in an API listing rather than the date the lab shipped it. Two live examples: Grok STT 1.0 is dated July 23, 2026 by one major timeline, but xAI released it on March 16, 2026 and July 23 is only when it was added to OpenRouter. Kimi K3 is dated July 14 by one tracker and July 16 by two others; Moonshot unveiled it on July 16. Every date on this page is read off the lab announcement, and the handful that are not are labelled as reporting.
What is a launch price, and why track it separately?
A launch price is the standard per-token rate the lab published on release day. It is worth freezing because per-token rates move: a model repriced downward six months later looks cheap in a current price table, which hides how the market actually shifted. Holding the launch price next to the current rate is how you see the drift. No other release tracker records it.
Which lab shipped the most models in July 2026?
Alibaba, with 4. The full split: Alibaba 4, Google 3, OpenAI 3, Ant Group 1, Anthropic 1, Black Forest Labs 1, Cognition 1, Cohere 1, Meta 1, Moonshot 1, poolside 1, Thinking Machines 1, xAI 1. Volume is not quality: a three-model drop in one announcement is one product decision, not three independent launches.
How much did AI model prices vary at launch?
By about 150x within a single month. In July 2026 the cheapest launch output rate was $0.2 per million tokens (Laguna S 2.1) and the most expensive was $30 (GPT-5.6 Sol). Per-token price is a poor guide to what work costs, though: token consumption varies more between models than the sticker rate does.
Where does this release data come from?
Each row is read off the lab's own announcement, documentation or press release and stamped with the date it was checked. 16 of 20 rows are sourced that way. 4 rest on reporting because the lab published no announcement page, and those are marked as such in the table and in the JSON. Nothing is taken from memory or from another tracker.

Get each breakdown before it makes the rounds

You get one email when a new source-backed analysis goes live: what AI agents actually cost, which models are worth running, and what the benchmarks really mean. No hype.

No spam. Unsubscribe anytime.

Sources

One primary source per release, with the date it was checked. Rows markedreporting had no citable announcement from the lab.

  • Anthropic (2026). Claude Opus 5. Announcement. Released Jul 24, verified Jul 29.
  • Black Forest Labs (2026). FLUX 3. Announcement. Released Jul 23, verified Jul 29.
  • Ant Group (2026). Ling-3.0-flash. Announcement. Released Jul 23, verified Jul 29.
  • Google (2026). Gemini 3.6 Flash. Announcement. Released Jul 21, verified Jul 29.
  • poolside (2026). Laguna S 2.1. Announcement. Released Jul 21, verified Jul 29.
  • Google (2026). Gemini 3.5 Flash Cyber. Announcement. Released Jul 21, verified Jul 29.
  • Google (2026). Gemini 3.5 Flash-Lite. Announcement. Released Jul 21, verified Jul 29.
  • Alibaba (2026). Qwen-Image-3.0. Reporting. Released Jul 21, verified Jul 29.
  • Alibaba (2026). Qwen-Audio-3.0-TTS Plus. Reporting. Released Jul 20, verified Jul 29.
  • Alibaba (2026). Qwen-Audio-3.0-TTS Flash. Reporting. Released Jul 20, verified Jul 29.
  • Alibaba (2026). Qwen3.8-Max-Preview. Reporting. Released Jul 19, verified Jul 29.
  • Moonshot (2026). Kimi K3. Announcement. Released Jul 16, verified Jul 29.
  • Thinking Machines (2026). Inkling. Announcement. Released Jul 15, verified Jul 29.
  • OpenAI (2026). GPT-5.6 Sol. Announcement. Released Jul 9, verified Jul 29.
  • OpenAI (2026). GPT-5.6 Luna. Announcement. Released Jul 9, verified Jul 29.
  • OpenAI (2026). GPT-5.6 Terra. Announcement. Released Jul 9, verified Jul 29.
  • Meta (2026). Muse Spark 1.1. Announcement. Released Jul 9, verified Jul 29.
  • xAI (2026). Grok 4.5. Announcement. Released Jul 8, verified Jul 29.
  • Cognition (2026). SWE-1.7. Announcement. Released Jul 8, verified Jul 29.
  • Cohere (2026). Cohere Transcribe Arabic. Announcement. Released Jul 7, verified Jul 29.
  • Artificial Analysis (2026). Intelligence Index and model pages. artificialanalysis.ai/models. Independent benchmark composite, used for the AA Index column.

← All tools & trackers