News

OpenAI Daybreak & Claude Code Auto Mode: AI News Aug 11

OpenAI Daybreak & Claude Code Auto Mode: AI News Aug 11

I. Today's Headlines

OpenAI is putting its frontier cyber models into the hands of the security industry. On August 10 the company announced an expansion of its Daybreak Cyber Partner Program, giving approved partners access to two controlled-release models — Daybreak Blue for broad defensive workflows, and Daybreak Red for tightly governed offensive work such as red teaming and penetration testing. The partner list is unusually heavyweight: services firms Accenture, IBM, Capgemini, Cognizant, EY, KPMG, PwC, NCC Group and SpecterOps, alongside technology partners Palo Alto Networks' Unit 42, CrowdStrike, Cisco, Sophos, Akamai, Fortinet and Cloudflare. Crucially, access stays with the partner: the models are not handed to the end customer, and engagements carry identity verification, defined testing scopes, logging, monitoring and human oversight. OpenAI's framing is that finding a vulnerability is the easy part, and that the real gap is triage and remediation at machine speed. (Source: OpenAI, August 10)

Anthropic is removing the per-command approval prompt from Claude Code. Starting August 14, auto mode becomes the default for Pro, Max and Team plans. Instead of asking the user to sign off on each file write and shell command, every tool call is routed through a classifier trained to block actions that are irreversible, destructive, or aimed outside the user's environment. The justification is a controlled study with 1,053 paid testers: humans caught 13.6% of deliberately planted dangerous commands, while auto mode blocked 89% of the same set. Human performance decayed with time on task — roughly 17% early in a session, falling to about 5% after fifty prompts — and users approved 97% of all permission prompts they were shown, which suggests the checkpoint had largely become muscle memory. Three consecutive blocks, or twenty across a session, drop the session back to manual approval. Anthropic has also stopped charging Pro, Max and Team users for the tokens the classifier consumes. Enterprise, API, Amazon Bedrock, Google Cloud's Agent Platform and Microsoft Foundry stay opt-in for now, with the default expected to follow within the month. (Sources: Anthropic announcement August 7; industry coverage August 10; subject to vendor confirmation)

II. Model Releases and Product Updates

Anthropic published prompt-injection numbers alongside the auto mode change. In a third-party evaluation by Trajectory Labs covering 72 scenarios run ten times each — 720 attack attempts in total — zero indirect prompt injections succeeded against Fable 5, Opus 5 or Sonnet 5 running auto mode. The same evaluation reported a 5.83% attack success rate against GPT-5.6 Sol running Codex's auto-review mode. Anthropic also cites production data showing teams on auto mode ship roughly 25% more pull requests, with Adobe, Nuro, Gusto and Garner Health named as production adopters. Treat vendor-run comparisons with the usual caution, but the direction of travel is clear. (Sources: Anthropic; Trajectory Labs evaluation as reported August 10; subject to vendor confirmation)

ChatGPT Business is getting premium seats. OpenAI announced the new tier on August 10, part of a broader push to package higher-capability access for business workspaces rather than leaving it to per-seat API budgets. (Source: OpenAI, August 10)

ChatGPT Atlas stopped working on August 9. OpenAI's standalone agentic browser, launched on macOS in October 2025, has been retired after less than ten months. Browser-based agentic capability moves into the ChatGPT desktop app — multi-tab support, downloads, improved navigation and account login — with a Chrome extension and sidebar covering users who stay on Chrome. Bookmarks, open tabs and history did not transfer automatically; OpenAI told users to export them before the cutoff. The lesson is not that agentic browsing failed, but that a whole browser was the wrong container for it. (Source: OpenAI Help Center; coverage August 9–10)

MiniMax H3's open weights are now running on consumer hardware. A week after the August 3 open-source release, community reports describe H3-Base running locally on a single RTX 5070 Ti with 32GB of system RAM — roughly 160 seconds for a five-second 480p clip, and around 560 seconds at 720p, with community repackaging bringing a ten-second generation down to about 318 seconds. H3 is an omni-modal system that takes text, images, video and audio as context and emits up to fifteen seconds of video at up to 2K with native stereo audio; it currently sits first on Artificial Analysis' audio-video editing leaderboard with an Elo of 1,130. Sixteen chip vendors and platforms completed day-zero adaptation, including Huawei Ascend, Moore Threads, Metax, Hygon, Kunlunxin, Iluvatar CoreX, Biren, AMD and Intel. (Sources: MiniMax official release; community benchmarks August 10; subject to vendor confirmation)

xAI shipped Imagine Image 2.0 on August 7, positioned around precise image generation and editing for production creative work rather than one-shot novelty. (Source: xAI, August 7)

III. Industry and Capital

Harvey is reportedly raising at least $500 million at a $15.5 billion valuation. The legal AI company closed $200 million at $11 billion in March; five months later the reported mark is $4.5 billion higher, on annualised revenue that has grown from $190 million in January to more than $350 million. That works out to roughly 44 times revenue. The strategically interesting detail is not the multiple but the roadmap: Harvey said in June it intends to build its own legal foundation models, which would put it in direct competition with Anthropic and OpenAI, the vendors currently powering its platform. (Sources: The Information via SiliconANGLE, August 7; subject to confirmation)

Cloudflare says machines now outnumber humans on the web. On its second-quarter earnings call the company reported that 60.4% of HTTP requests returning HTML across its network came from bots, against 39.6% from humans, with agent traffic up sharply year over year. Whatever the exact figure, the composition of web traffic has crossed a line that changes the economics of publishing, rate limiting and bot policy. (Sources: Cloudflare Q2 earnings call, August 9; subject to vendor confirmation)

The cost side is starting to bite. SAP has reportedly paused most travel and hiring, citing AI compute costs, at the same time as data circulating this week put roughly 70% of AI revenue in the hands of OpenAI and Anthropic. Both data points point the same way: the spending is broad, the revenue is narrow, and companies that are neither model labs nor chip vendors are absorbing the difference. (Sources: industry reporting, August 10; subject to confirmation)

IV. China in Focus

ByteDance is reportedly pre-training a model with up to ten trillion parameters. According to the Financial Times, cited by Chinese tech outlets, the run is in its early stages with a full cycle expected to take three to six months. If accurate, that would be several times larger than Kimi K3, currently the largest model shipped by a Chinese lab. It follows ByteDance elevating AI to its first core business track and founder Zhang Yiming reportedly rejecting model distillation proposals three times inside the Seed team — a position that reads as expensive but deliberate. (Sources: Financial Times via PConline, August 10; subject to vendor confirmation)

Chinese models took four of the top five slots on OpenRouter last week. Platform token volume reached 69 trillion for August 3–9, up 21.48% week over week, with DeepSeek V4 Flash 0731, Hy3, DeepSeek V4 Flash 0423 and MiMo-V2.5 ahead of GPT-5.6 Luna. Whatever the benchmark leaderboards say, developer routing behaviour is now a separate signal — and on that signal the price-performance tier is where the volume lives. (Sources: Guoyuan Securities industry weekly, August 10; OpenRouter data; subject to confirmation)

Unitree opened its STAR Market IPO subscription on August 10, making it the first humanoid robotics company to list on the board. Embodied AI is moving from demo reels to disclosure documents, which is a different and more revealing genre. (Sources: Chinese financial media, August 10; subject to confirmation)

V. Today's Observation

Two of today's biggest stories are the same story told from opposite ends.

OpenAI's argument for Daybreak is that attackers now move at machine speed, so defenders need frontier models inside the workflows they already run — and it wraps that access in identity checks, scoped engagements and human oversight, precisely because the capability is dangerous. Anthropic's argument for auto mode is that the human in the loop was never really in the loop: a checkpoint approved 97% of the time is not review, it is a click. Both companies looked at a place where human judgment was supposed to be the safeguard, concluded it was not working, and reached for a model.

The number worth holding onto is not 89%. It is the 11% the classifier misses. Under the old default those actions still had a person in front of them, catching roughly one in seven. Under the new default they have nobody. That may still be the better trade — 89% of a large number beats 13.6% of it — but it is a trade, not a free upgrade, and the industry should be honest that the party proposing to remove the human is also the party reporting the low human score.

The quieter lesson today comes from Atlas. OpenAI built a whole browser to host agentic browsing, ran it for under ten months, and concluded the container was wrong while the capability was right. In a year where every lab is shipping agents into new surfaces, that is a useful reminder: the feature usually survives; the product around it often does not.


Compiled from vendor announcements and public reporting. Where a claim originates from media or community sources rather than a vendor's own publication, it is marked subject to vendor confirmation.

Comments (0)

Share:XHatena

Post a Comment

Loading...