AI Releases & Pricing

July 24, 2026

DeepSeek's deepseek-chat and deepseek-reasoner aliases retire today at 15:59 UTC

Subscribe
Listen

The legacy model names stop resolving at 8:59 AM Pacific while V4 general availability still has no entry in the official changelog; Microsoft shipped priced in-house image and voice models, and the Claude Opus 5 rumor gained its first real-name sighting.

The DeepSeek migration deadline lands today, but V4 GA still has not

The deadline this feed has tracked for a week arrives this morning. DeepSeek's two legacy model aliases, deepseek-chat and deepseek-reasoner, are deprecated today, July 24, 2026, at 15:59 UTC (8:59 AM Pacific). After that, any call passing either name returns an error. The fix is a one-string rename: deepseek-chat to deepseek-v4-flash (non-thinking) or deepseek-v4-pro, and deepseek-reasoner to deepseek-v4-flash (thinking mode) or deepseek-v4-pro. Base URL, API keys, and billing are unchanged.

One detail worth re-checking before you cut over: until 15:59 UTC, the aliases route to V4-Flash, not Pro. If your app called deepseek-reasoner assuming it hit the strongest model, you have been on Flash the whole time. Moving to deepseek-v4-pro is an upgrade, so re-test your prompts against Pro rather than assuming parity.

What has not arrived is the general-availability launch that deadline was meant to accompany. DeepSeek's official news page still lists the April 24 V4 Preview as its newest entry, with zero July posts. The "official V4 version" that DeepSeek emailed API users about on June 29 and that Chinese press reported for "as early as Monday" on July 19 has slipped past its mid-July window, with the delay attributed to bundling a first-party coding harness. The deadline ships on time even when the launch does not.

That matters for your bill. The announced peak/off-peak structure, the first time-of-day pricing on a frontier API, doubles rates during Beijing business hours (09:00 to 12:00 and 14:00 to 18:00 Beijing, which is 01:00 to 04:00 and 06:00 to 10:00 UTC). Off-peak stays at the current baseline: V4 Pro $0.435 in / $0.87 out, V4 Flash $0.14 / $0.28 per million tokens. Peak doubles those to $0.87 / $1.74 and $0.28 / $0.56. As of this morning I could not confirm the peak tier is live on the docs, so budget for it but verify the rate card before assuming the surcharge is active.

DeepSeek V4 peak pricing doubles output cost during Beijing business hours
Output $ per 1M tokens, off-peak vs announced peak. Peak windows: 09-12 and 14-18 Beijing (01-04 and 06-10 UTC). Source: DeepSeek API docs + TechNode; off-peak live, peak pending official GA.

A reminder on sourcing: a post at deepseek.ai/blog claiming "V4 GA July 24" is a fan site ("Deep Seek Fan Hub"), not DeepSeek's official channel. DeepSeek's own preview post tells readers to "rely only on our official accounts." Treat any GA claim as false unless a July entry appears in the official changelog.

Microsoft ships priced in-house image and voice models, claims up to 89% cost cut versus OpenAI

Microsoft AI pushed two purpose-built models into public preview on July 23, and the pricing is concrete enough to model against current spend.

MAI-Image-2.5-Pro, billed as Microsoft's highest-fidelity image generator, is priced at $5 per 1M text input tokens, $8 per 1M image input tokens, and $106 per 1M image output tokens. That more than doubles the image-output price of the base MAI-Image-2.5 ($47), targeting the premium hero-imagery and in-image-text-rendering tier rather than volume generation. It is available in Microsoft Foundry and the MAI Playground.

MAI-Voice-2-Flash, first shown at Build, is a low-latency text-to-speech model at $15 per 1M characters, 2x faster and 32% cheaper than MAI-Voice-2 (which runs $22 per 1M characters). It is aimed at high-volume call-center and voice-agent workloads.

The deployment list is the pricing-pressure signal. Microsoft says the models already run in Bing, PowerPoint, OneDrive, Dynamics 365, Excel, GitHub Copilot, and Azure, and reports vendor-supplied production figures: up to 84% GPU cost reduction in PowerPoint versus OpenAI's GPT-Image-2, and up to 89% in Dynamics 365 Contact Center (VentureBeat, July 23). Those are Microsoft's own numbers on Microsoft's own workloads, so read them as a vendor claim, not an independent benchmark. But the framing, "Microsoft products, powered by Microsoft models," is the clearest sign yet that OpenAI's largest backer is building a first-party stack that competes with OpenAI's frontier models on cost for specific modalities.

Claude Opus 5 rumor: the codename era may be over, but Anthropic's page still says 4.8

RUMOR. The trail on Claude Opus 5 just produced its most specific artifact yet, and it is still not an announcement.

For weeks every leaked trace used the "Honeycomb" codename. That changed on July 23 to 24: screenshots circulating on X show a Cursor error dialog naming the model outright, "The model claude-opus-5-thinking-high requires Max Mode to be enabled." That is the first sighting of the literal claude-opus-5 string in a shipping product, and the -thinking-high suffix matches how Cursor labels reasoning variants, which reads like launch plumbing rather than an experiment. Deployment tracker @M1Astra posted that Opus 5 had begun rolling out across providers, with some users reportedly served the new model under the "Opus 4.8" label before a full switchover. A separate screenshot, attributed to someone posing as an Anthropic employee, showed a Fable 5 guardrail routing to Opus 5.

Against all of that stands the primary record. Anthropic's official models overview still lists Claude Opus 4.8 (May 28, 2026) as the newest Opus, with no Opus 5 model card, pricing page, or API model ID. A forensic audit published July 24 found no Opus 5 entry in Anthropic's, Google Cloud's, or Cursor's public model catalogs. The community's favorite Thursday target, July 23, came and went with no Anthropic launch post. Until anthropic.com/claude/opus or the API docs ship a named model, the public Opus flagship remains Opus 4.8 at $5 in / $25 out per million tokens. Build on that, watch the rest.

Tracking

Current prices

Output and input dollars per 1M tokens, each linked to the official pricing or docs page, verified July 24, 2026. Off-peak unless noted.

The output-price spread across this list is roughly 178x, from DeepSeek V4 Flash at $0.28 to Claude Fable 5 at $50 per million tokens.

That’s the reading for this issue.