Caught Up · Today

One edition and you’re caught up on AI.

today's briefing

Two Modes of AGI: DeepMind Ships, Anthropic Negotiates

Google DeepMind flooded the week with agents, translation and a safety roadmap; Anthropic won back an export license, lost an Alibaba account, and watched Claude help a researcher forge festival tickets.

2026-07-053 min readgoogle-deepmindanthropicai-safety

This week is not about a single model. It is a study in contrast between the two largest Western labs and how each is trying to survive an era in which capability is outrunning regulation. Google DeepMind ran a saturation campaign: new products, new safety program, new money for outside researchers. Anthropic spent its week on politics, incidents and export law — and did so on the front page.

Saturation from Mountain View

DeepMind unveiled computer use in Gemini 3.5 Flash, giving the cheaper, faster Gemini variant the ability to drive a browser and desktop applications as an agent. At the same time it shipped Gemini 3.5 Live Translate, bringing near real-time natural voice translation to Google AI Studio, Translate and Meet. What is interesting is less any single release than the density of them. DeepMind has clearly decided that distribution — folding agents directly into Meet, Chrome and Workspace — matters more than waiting for the perfect model.

To prevent all of this from reading as pure product velocity, the lab published an AI Control Roadmap for securing agent systems, pairing traditional safeguards with real-time behavioural monitoring, and opened a $10M funding call for multi-agent safety research. The timing is not coincidental: if you are pushing an agent into the world that will click wherever you tell it, you also need to be seen thinking about what happens when someone other than the account holder does the telling.

Anthropic between Washington and Front Gate

Anthropic is living a very different week. After the White House ordered the company to suspend foreign-national access to its most advanced models, the Trump administration has now lifted export controls on Fable 5 and Mythos 5. The re-entry ticket was not free: according to Wired, Anthropic had to add a new security measure to get back into the administration's good graces. The specifics remain private, but the precedent is not: the safety architecture of frontier models is now a bargaining chip between labs and the executive branch.

As if to underline the risk, Wired also reported that a researcher used Claude Opus 4.7 to break into Front Gate, the ticketing platform behind essentially every major US festival from Lollapalooza to Bonnaroo, and could issue any ticket he wanted. The model did nothing malicious in the abstract; it was simply good enough to find a hole in someone else's code faster than the vendor's own security team. That is exactly the flavour of incident that keeps policy leads awake, because it is very hard to explain to a senator.

And then a third beat: Alibaba has reportedly banned employees from using Claude Code, classifying it as high-risk software. In a single week Anthropic won an export reprieve in Washington, an internal ban in Hangzhou, and a security embarrassment in Austin. It is hard to imagine a cleaner illustration that model safety is no longer a technical category but a geopolitical one.

Second-order effects

The overlap is more interesting than either story alone. DeepMind is betting that trust comes from combining distribution with visible safety investment: a published roadmap, a public grant, agents embedded in products people already use. Anthropic is betting on a tighter, deeper relationship with the state: fewer product headlines, more meetings behind closed doors. Both strategies have a price. DeepMind risks agent incidents arriving before its monitoring stack matures. Anthropic risks becoming a vendor whose safety roadmap is effectively written in the White House.

The second-order effect worth watching is the coupling between safety and access. If the emerging pattern is that labs pay for export licenses in the currency of new security layers, we end up in a world where the most capable models have the tightest political tether to the government that let them out. For US customers that may be acceptable. For European buyers weighing Gemini against Claude, it starts to look less like a pricing question and more like a sovereignty one.

What to watch next

Three things. First, whether DeepMind publishes any measurable data from computer use pilots — error rates, prompt-hijack incidents, kill-switch actions. Without numbers the AI Control Roadmap is just a document. Second, what exactly Anthropic's "new security measure" for the administration turns out to be; if it involves something like mandatory logging of foreign queries, that reshapes the enterprise Claude market in Europe overnight. And third, how Front Gate and Anthropic explain the ticketing incident, together or separately. If Claude Opus 4.7 can reliably surface vulnerabilities in production sites, agent misuse stops being a hypothetical and becomes a line item on cyber insurance forms.

Source ledger

You’re caught up.

Recent issues

archive