Skip to content
Larnaca, Cyprus
BINA CYINNOVATION HUBLarnaca · est. 2026
Aerial long-exposure photograph of a highway interchange at night, with blue and amber light trails from traffic converging at a toll plaza checkpoint
AIAI14 August 20265 min read

AI Brief — August 14, 2026

DeepSeek V4-Pro goes GA with a price shock; Cerebras makes GPT-5.6 Sol 14× faster; Anthropic ships Claude Code auto mode; China unwinds Manus deal.

By BINA Editorial

Four stories dominate the AI wire on August 14: OpenAI and Cerebras clock a 14× inference leap on GPT-5.6 Sol; Anthropic flips Claude Code's auto mode on by default for paid users; DeepSeek graduates V4-Pro to general availability — and follows the announcement with a price increase that catches developers off guard; and China's regulators complete the forced unwind of Meta's $2 billion acquisition of Manus, with Tencent waiting in the wings.

OpenAI and Cerebras Unlock 14× Faster GPT-5.6 Sol

OpenAI and Cerebras announced an "Ultrafast Mode" for GPT-5.6 Sol on Thursday morning, clocking 750 output tokens per second — roughly 14 times the ~53 tokens per second available in Standard Mode. The partnership leverages Cerebras' wafer-scale chip architecture, which bypasses the memory bandwidth bottleneck that slows conventional GPU clusters during the model's decoding phase.

Ultrafast Mode enters limited preview for API customers today. OpenAI describes GPT-5.6 Sol as its strongest model for long-form professional work — legal briefs, financial models, engineering reports — tasks where a 14× speed multiplier translates directly into faster iteration cycles and lower wall-clock latency in production pipelines.

The commercial arrangement marks Cerebras, which trades on the NASDAQ as CBRS, as an inference infrastructure partner rather than a competitor to OpenAI. It continues a pattern visible throughout 2026: frontier labs outsourcing specialized inference to chip startups rather than building or acquiring the hardware themselves.

No change to pricing or output quality accompanies the new tier. OpenAI has not said when Ultrafast Mode will exit limited preview or reach enterprise and consumer tiers.

Anthropic Makes Claude Code's Auto Mode the Default

Starting today, new Claude Code sessions on Pro, Max, and Team plans run in auto mode by default. Rather than pausing to request human approval before each file edit, shell command, or network call, Claude Code proceeds unless its built-in classifier judges an action "irreversible, destructive, or aimed outside your environment."

The rationale for the switch comes from Anthropic's own controlled study: human reviewers caught a deliberately planted dangerous command 13.6% of the time; the classifier caught it 89% of the time. On that evidence, AI-supervised review outperforms human spot-checking for routine command oversight.

Users who previously set a custom permission mode keep it. Those who have pinned a preference see no change. Anthropic has also removed the extra token charge for the classifier on paid plans — a cost that had previously made some teams hesitant to enable auto mode.

Enterprise, API, and cloud platform users can expect the same default within the next 30 days. The move is the clearest signal yet that Anthropic treats AI-supervised agentic coding as more reliable than the per-command human-in-the-loop workflow it launched with. Whether the broader developer community agrees will become clear as feedback comes in from today's rollout.

DeepSeek V4-Pro Goes GA — and Brings a Price Increase Up to 12×

DeepSeek graduated V4-Pro from preview to general availability on Wednesday, designating the build as DeepSeek-V4-Pro-0813. The model has been in limited preview since April; Wednesday marked the first time it opened for unrestricted API access. The timing was immediately overshadowed by a pricing announcement: a new peak and off-peak billing structure takes effect at 16:00 UTC on August 16.

On benchmarks focused on autonomous coding and tool use, V4-Pro scored 87.9 on Terminal Bench 2.1, 62.7 on DeepSWE, and 61.5 on NL2Repo — numbers that put it in competitive range with frontier Western models on agent-specific tasks. The context window reaches 1 million tokens, with a maximum output length of 384,000 tokens. Thinking and non-thinking modes are both supported, and the API now exposes three effort levels — low, high, and max — for the reasoning budget.

V4-Pro is also fully compatible with the OpenAI Responses API format, meaning developers can point existing client code at DeepSeek's endpoint without rewrites. That deliberate interoperability lowers the cost of switching workloads over — or switching back, now that pricing has changed.

The pricing restructuring is the headline detail. At peak hours (01:00–04:00 and 06:00–10:00 UTC), cache-miss input rises to $1.32 per million tokens, up from $0.435. Output tokens jump to $3.96 per million at peak, up from $0.87. Off-peak rates run at half the peak level. Depending on when workloads run, effective cost increases range from roughly 1.9× on input tokens to 12× on cached reads. DeepSeek frames the change as a demand-shaping mechanism intended to shift traffic toward less congested periods, but introducing the change on the same day as the GA release gives developers only 72 hours to adjust before costs change.

China Unwinds Meta's Manus Deal — Tencent Circles

AI agent startup Manus announced on August 11 that it will return to operating as an independent company after Chinese regulators ordered Meta to reverse its acquisition. Meta had completed the deal on December 29, 2025, for approximately $2 billion. China's National Development and Reform Commission issued its withdrawal order in April 2026, citing violations related to technology exports and foreign investment rules.

The separation has immediate, practical consequences for users. Data created on or after December 29, 2025 — the date Meta's ownership began — will be deleted between August 23 and August 24 in the jurisdictions subject to the regulatory order. Affected users have a backup window through 7:59 p.m. EDT on August 22; restoration access opens from August 25.

During the unwinding review, Manus co-founders CEO Xiao Hong and chief scientist Ji Yichao were reportedly barred from leaving China while authorities examined the transaction. Reuters and the Financial Times reported in July that Tencent is in talks to become Manus' largest shareholder — an outcome that would return the company to Chinese domestic ownership without reopening a Western acquisition path.

The forced reversal is the most direct example to date of Chinese regulators actively unwinding a completed outbound AI acquisition. It establishes a precedent that will factor into due diligence for any future deal involving a Chinese AI company and a foreign acquirer. For Manus' users, who built workflows around its autonomous agent assistant, the episode surfaces both a near-term data continuity problem and an open question about what the product roadmap looks like under new ownership.