Home/Archive/24 August 2026
Weekend digest · 21–24 August 2026

DeepSeek makes vision free at the text rate

Price was the release this weekend, not capability. DeepSeek bolted image understanding onto its budget Flash model and billed it at the same rate as text, which quietly ends multimodality as a premium feature. OpenAI took a third off GPT-5.6 Sol's output price for three months, and Meta opened a contributor tier for Muse Spark 1.2 at ten cents per million tokens, both aimed at buying developer habit rather than topping a leaderboard. The other pattern was distribution: Grok 4.6 reached Google's Vertex AI after Bedrock and GitHub Copilot, and xAI bundled its Grok Bot agent into Cursor's paid tiers. The one capability story, and the most interesting for builders, was NVIDIA's AVO scoring a perfect 100 per cent on ARC-AGI-3 using scaffolding around a model that manages 30 per cent alone. If you have not re-run your cost model in a month, it is now out of date.

DeepSeek

DeepSeek puts image understanding into the cheap tier at no premium

This is the most consequential release in the window, and the pricing is the reason. DeepSeek has bolted vision onto its budget Flash model and is billing images at the same token rate as text, with each image costing up to 384 tokens after automatic resizing, so document parsing, chart reading and screenshot-driven agent loops stop being a premium feature. On text it matches the existing V4-Flash for agents, reasoning and world knowledge, and DeepSeek's own multimodal agent benchmarks put it close to Opus 4.8, a bold claim for a sparse mixture-of-experts model running 13 billion active parameters out of 284 billion.

It is reachable through the same endpoint you already use, with model deepseek-v4-flash-vision-exp, and DeepSeek Harness 0.1.1 shipped the same day with support built in. The label says experimental, so treat it as something to benchmark against your own data rather than move production traffic to this week.

Sources
DeepSeek on X, 21 Aug [Direct] DeepSeek API docs, 21 Aug OpenRouter, 21 Aug
OpenAI

GPT-5.6 Sol gets more than 20 per cent cheaper, but only for three months

If any part of your stack calls Sol, your unit economics just changed. Input drops from 5 dollars to 4 dollars per million tokens and output from 30 dollars to 20 dollars, a 33 per cent cut on the output side where most agentic and coding workloads actually spend their money. The discount reaches the API, ChatGPT Work credits and Codex credits, while Pro, Plus and Business subscriptions are untouched, so this is squarely aimed at developers rather than seat holders.

The catch is the clock: OpenAI has framed it as a promotion running roughly to 21 November, which makes it a window to migrate volume or renegotiate a budget rather than a permanent repricing. It is also OpenAI's second pricing move in under a month after July's cuts to Terra and Luna, which tells you the frontier tier is now competing on cost as hard as on capability.

Sources
OpenAI on X, 21 Aug [Direct] Business Standard, 22 Aug OpenAI Community, 21 Aug
Meta

Muse Spark 1.2 opens a contributor tier at ten cents per million tokens

Meta has put a stripped-back access tier for Muse Spark 1.2 on OpenRouter at 0.10 dollars per million input and 0.20 dollars per million output, roughly a twelfth of the standard rate and cheap enough to change what prototyping costs. You keep the 1M-token context window and multimodal input covering text, images, video, audio and PDFs.

The trade is explicit and worth reading twice before you use it: prompts and outputs on the contributor tier may be used to improve Meta's products, so it is fine for experimentation, evaluation harnesses and learning, and wrong for anything touching client data or anything you would not want in a training set. Meta is clearly buying developer habit rather than benchmark headlines, which fits its wider pattern of undercutting rather than out-scoring Anthropic and OpenAI.

Sources
OpenRouter, 21 Aug Price Per Token, 21 Aug
xAI

Grok 4.6 arrives on Google Cloud Vertex AI

Grok 4.6 is now in preview through Vertex AI Model Garden, which matters mainly if your organisation already runs its procurement, billing and data controls through Google Cloud and could not previously touch xAI models. Pricing starts at 2 dollars per million input tokens and 6 dollars per million output, with cached input at roughly a quarter of the standard rate, and you get the 500,000-token context window plus four selectable reasoning-effort levels.

Coming a fortnight after Grok 4.6 landed on Amazon Bedrock and in GitHub Copilot, this looks less like a product update and more like xAI accepting that enterprise adoption runs through the hyperscalers. For anyone comparing long-context coding models, it means Grok can now sit in the same bake-off as Gemini and Claude without a separate vendor agreement.

Sources
xAI on X, 21 Aug [Direct] xAI news, 21 Aug
Also noted
Grok Bot, xAI's always-on agent with its own machine and tool access, is now included with SuperGrok Plus, Cursor Pro+ and every Cursor Teams plan rather than sold separately. Bundling an autonomous agent into an existing coding subscription is an aggressive distribution play given how new the product is: if you already pay for Cursor at those tiers, you have an agent sitting there whether you asked for one or not, which is worth checking against your policies on what may run unattended against a repository. (Source: xAI news, 21 Aug)
NVIDIA

NVIDIA AVO scores 100 per cent on ARC-AGI-3, and the scaffolding did the work

NVIDIA's general-purpose coding agent completed all 183 levels across all 25 public ARC-AGI-3 environments with no instructions, stated rules or goals, using 12 per cent fewer actions than the previous best system. The detail that should interest anyone building agents is what sits underneath: AVO runs on Anthropic's Claude Opus 5, which manages around 30 per cent on the same benchmark unaided. The gap between 30 and 100 is architecture, not model weights, a useful counterweight to the assumption that better results always require a better base model.

Two caveats keep this from being more than a strong signal: the result covers only the public environment set, with ARC Prize's held-out sets untested, and AVO is a research demonstration rather than something you can buy or integrate today.

Sources
NVIDIA AI on X, 21 Aug [Direct] NVIDIA blog, 21 Aug
Also noted
NVIDIA's AI safety and security teams published a deep dive on where security belongs in the agent stack, drawing a line between what the harness guides an agent to attempt and what the infrastructure actually permits. Useful framing if you are writing agent policy rather than shipping one. (Source: NVIDIA AI on X, 21 Aug)
Google DeepMind

A research partnership aimed at the four things agents are still bad at

DeepMind has announced a research partnership with Fenris Creations to use a persistent, populated virtual world as a testbed, and the stated target list reads like an honest inventory of what current agents cannot do: continual learning without catastrophic forgetting, memory systems that reach well past today's context windows, planning over weeks or months rather than minutes, and multi-agent dynamics covering cooperation, negotiation and emergent economic behaviour.

The team's framing is that SIMA taught agents to understand 3D space, but understanding human social dynamics needs a world that keeps running when the agent stops looking. For anyone building long-running agents commercially, this is a signal about where the next round of capability gains is expected to come from, and none of the four items on that list will arrive this quarter.

Sources
Google DeepMind on X, 21 Aug [Direct] Google DeepMind blog, Aug
Quiet in the last 72 hours
Anthropic Nothing significant in products or releases. The AnthropicAI account's most recent post is 18 August, though Claude Code 2.1.239 surfaced on release trackers on 22 August, an incremental build covering cost estimates, cloud sessions, Bedrock and Alpine support, with no dated announcement alongside it.
Microsoft · Mistral · Perplexity Nothing significant. Microsoft Research last posted on 18 August, Mistral on 11 August, and Perplexity on 18 August.
Hugging Face · ElevenLabs · Others Nothing qualifying from Hugging Face itself, ElevenLabs, IBM, Snowflake, Cohere, Runway, Stability AI or Scale AI in the window.

Industry themes

Price is now the release. Three of the five substantive items this window are about what a token costs or where you can buy it, not what a model can do: OpenAI taking a third off Sol's output price, Meta opening a ten-cents-per-million contributor tier, and DeepSeek giving away vision at the text rate. If you have not re-run your cost model in a month, it is out of date.

Multimodality has finished being a premium feature. DeepSeek charging nothing extra for image input on its budget tier sets a reference point that every other provider's vision pricing now gets compared against, and it makes document and screenshot workflows viable at volumes that did not add up a fortnight ago.

Distribution beats exclusivity. Grok 4.6 landing on Vertex AI after Bedrock and GitHub Copilot, and Meta pushing a tier onto OpenRouter, both point the same way: models are placed where developers already have billing set up rather than behind their own front doors. NVIDIA's AVO is the odd one out and the most interesting for builders, because it shows a 30 per cent base model reaching 100 per cent through architecture alone. One process note: the OpenAI price cut surfaced on the company's own X account and developer forum and had still not reached openai.com/news by this morning, so without the social scan the single most commercially relevant item here would have been missed.

← 21 August 2026 All issues →

Every day I sift through the noise, cut out the hype, and serve up the AI updates that actually matter for your business. Straight to your inbox before 8am. No fluff, no jargon, no faff.

👩👨👩👨+
Join 2,400+ business professionals already subscribed

You're in! First issue lands tomorrow morning.