Mistral prices Large 4 at a twelfth of OpenAI's GPT-6 Astra rate as a LangChain test cut coding costs 64%

Mistral listed Large 4 at a twelfth of GPT-6 Astra's output price, LangChain's router cut coding cost 64% with no quality loss, and Atlassian took OpenAI money while refusing to make Astra its default. The evidence points one way: the tollbooth is moving from the model to the door.

Vincent JiangVincent Jiang · 3 min read
Share
Mistral AI chief executive Arthur Mensch meets Indian Prime Minister Narendra Modi in Paris
1 / 7Slide 1 of 7
Mistral AI chief executive Arthur Mensch meets Indian Prime Minister Narendra Modi in Paris, June 2026. Mensch's company launched Large 4 this week at a fraction of frontier rivals' token prices.

Rippling runs its production AI agents on a LangChain architecture with one supervisor agent coordinating five to seven specialized sub-agents, built on Deep Agents and LangSmith, per a LangChain case study disclosed this week 5. None of those sub-agents needs to be expensive. That is the week's trade, hiding inside a price list and an experiment.

The model layer just priced itself like a commodity

Mistral launched Large 4 on October 6, a 1-trillion-parameter mixture-of-experts model, at $1.36 per million input tokens and $4.18 per million output tokens 4. GPT-6 Astra charges $10 and $50 for the same units 4. Claude Opus 5.5 charges $4 and $20 4. On output, Large 4 runs at roughly a fifth of Opus 5.5's price, and Decrypt's own comparison puts it at "roughly a seventh and a twelfth" of Astra's on input and output 4. The open-weights release lands by the end of October, which lets buyers test the quality claims themselves rather than trust the vendor's benchmark card 4.

Mistral's Large 4 lists at a twelfth of Astra's output price

  • Input $/M tokens
  • Output $/M tokens
0204060GPT-6 AstraClaude Opus 5.5Mistral Large 41.36
Data
Input $/M tokensOutput $/M tokens
GPT-6 Astra1050
Claude Opus 5.5420
Mistral Large 41.364.18
Published list prices in US dollars per million tokens, input and output, as of October 6, 2026. Source: Decrypt reporting on the Mistral Large 4 launch.4

The router got receipts the same week

LangChain published a 973-thread experiment on October 1 in which its model router cut median AI coding cost by 64%, with no statistically significant difference in code-merge outcomes 1. Read that twice: the router committed to one model before each conversation unfolded, and the merged code did not get worse 1. Atlassian spent the same week deepening its OpenAI partnership, which an OpenAI representative described as "effectively a spend commitment," while telling VentureBeat that GPT-6 Astra is not Rovo's default; its internal AI gateway routes dynamically across providers per task 3. Atlassian's rebuilt MCP server handles more than 15 million calls a day across 220-plus tools, at up to 25% fewer tokens than the prior version 3. A survey-reported figure rounds the point off: only 2% of enterprises said the model was the main reason they chose an agent platform 2.

The routing layer is absorbing everything beneath it

LangSmith Engine v2 shipped October 6 6. AWS published agentic retrieval inside the langchain-aws package, with its managed knowledge base explicitly owning the storage layer; at five retrieved results, one embedding covered just 4 of 6 sub-intents in a comparative question, which is why AWS itself recommends routing by query shape rather than picking one retrieval path 7. CoreWeave's new Partner Network, announced October 6, lists LanceDB in its data-services segment, survival by being named in someone else's distribution 8. The pattern is consistent: whoever owns the customer owns the router, and the router treats models as interchangeable parts.

What this means for the $2 trillion book

Anthropic is reportedly targeting an IPO at around a $2 trillion valuation, raising roughly $60 billion to $100 billion, against 2025 net losses of nearly $42 billion and an annualized run rate that passed $65 billion by the end of July 2026 9. That ask dates from September 13, three weeks old, and the roadshow it sets up lands in October 9. The multiple only holds if frontier pricing holds. The same week it is being marketed, the routing layer's own published data says the model is interchangeable and the cheapest competent one wins the workload 14. A router optimized for cost will have every reason to pick Large 4 once its open weights are downloadable by month's end 4.

The counterpoint and the open door

The strongest case for paying the mark: frontier capability still commands a premium on the hardest tasks, and Surge AI's blind human evaluation had Claude Opus 5 at 4.22 against Large 4's 3.74 in coding quality 4. But routing exists precisely to buy the premium only where it pays. Watch the October roadshow, and watch whether Atlassian-style gateways keep Astra out of the default slot once the weights land. The tollbooth is moving from the model to the door.

Deepdive

AI-generated from this story and its cited sources. Not investment advice.

Reader comments

0 comments

    Sign up

    Get your curated digest

    After email confirmation, you will receive a daily digest of the most relevant news that matter to your portfolio