ZeroNoise Logo zeronoise
Post
AI labs are selling governed domain workflows, not just models
4 min read
1097 docs
OpenAI, Anthropic, and Google DeepMind are packaging frontier capability for law, biology, and genomics, while Anthropic puts numbers around AI-led R&D and agent oversight. The common thread is a shift from general-purpose models toward domain-specific infrastructure, permissions, and verification.

Domain stacks move into high-stakes work

OpenAI turns GPT-6 Astra into a legal stack

OpenAI launched Astra for Law as a foundation for law firms and legal-technology companies, combining GPT-6 Astra with legal-analysis and writing instructions, tailored settings, tools, context, and a legal search index. The index covers U.S. case law, statutes, regulations, court rules, and administrative decisions across more than 230 million URLs, with sources added daily.

OpenAI reports that the complete setup passed the overall correctness check on 54.0% of 200 questions in the private Vals AI Legal Research Bench, versus 38.7% for GPT-6 Astra using web search alone—a 40% relative improvement. The result is vendor-reported and comes from a private validation set, but it shows the product’s intended advantage: retrieval and legal context are being packaged alongside the frontier model rather than left to users to assemble.

The initial rollout is limited to selected firms through Trusted Access, with zero data retention on the API and default exclusion of ChatGPT Enterprise usage from human review; OpenAI is also working with Latham & Watkins on permissions, ethical walls, client instructions, and oversight. The launch adds 26 partner-built plugins and 47 adaptable community skills, reinforcing a strategy built around firm-specific workflows and an ecosystem of legal tools rather than a standalone chatbot.

Anthropic makes biology access both more permissive and more controlled

Anthropic opened applications for a beta Life Sciences Verification Program that gives verified teams access to Mythos, Opus, and Sonnet with safeguards more permissive for biology work than those on its generally available models. Applicants are reviewed for research credentials, security standards, and ethical oversight; standard grants cover broad team workflows, while high-risk grants are project-specific, require additional vetting, and renew every six months.

The program ties access to an organization’s stated use cases and continuously monitors traffic for activity outside that scope. Anthropic says it is shifting some enforcement from real-time blocking to offline behavioral monitoring, retaining data associated with flagged activity for 30 days while keeping it compartmentalized and out of model training.

Alongside the access program, Anthropic reports that Claude optimized more than 30 open-source biomolecular models, producing roughly fourfold average speedups with minimal precision loss and nearly twofold speedups with identical outputs; the optimization code is open-sourced. A new Adaptyv Bio competition will experimentally validate more than 5,000 community designs, backed by up to $1 million in Claude credits and $250,000 in Modal compute credits. The combination points to a practical biology strategy: expand access where users can be verified, while lowering the compute and experimental cost of the work itself.

Research infrastructure becomes a shared utility

AlphaGenome Atlas precomputes a map of human genetic variation

Google DeepMind introduced AlphaGenome Atlas as a free academic resource containing predictions for all 9 billion possible single-nucleotide variants in the human genome. The company describes it as a roughly 1-petabyte dataset with thousands of molecular-effect predictions per variant, an AlphaGenome Variant Impact score, and more than 2,500 recurring DNA motifs across hundreds of cell types and tissues.

DeepMind says collaborators used the atlas to prioritize a DNM1 variant in unsolved rare-disease research, with experimental screens validating the predicted mechanism. In a separate analysis of more than 54,000 UK Biobank participants, the University of Exeter team reported 22% more detectable non-coding genetic associations; Stowers researchers used the atlas’s motifs to classify regulatory activity.

The important shift is from a model researchers query one case at a time to a precomputed, searchable research layer that exposes model predictions at genome scale. DeepMind’s results are company-reported, but the release pairs the infrastructure with a concrete validation example and makes the resource available through a portal, API, and Google Antigravity.

Transparency gets more precise, not yet comparable

Anthropic puts numbers around AI-led R&D and oversight

Anthropic published three internal measures of frontier development: how much AI R&D is performed by AI systems, how well agents are overseen, and how compute is allocated. It argues that other frontier developers could publish comparable measures and that third parties could verify them.

Its August snapshot says Claude “led” 26% of Anthropic’s AI R&D work, performed at or above the “AI collaborates” level for more than 90% of the work, and was not fully autonomous for any measured subset. On the company’s most-used internal research platform, roughly 30,000 agents had 100% of actions pass through online or offline monitoring; online monitors blocked 0.002% of more than a billion decisions. Anthropic also reports that 6% of AI-R&D compute, and 12% of compute going to AI-driven AI R&D, was allocated to safety in a July 13–20 snapshot.

Anthropic labels the automation index a prototype, limits the oversight figures to one internal platform, and says cross-lab comparisons need a common methodology and independent checks because the lab is using its own models as judges. Nathan Lambert’s reaction was similar: he called the disclosure a step in the right direction but said the 26% figure does not define what counts as AI R&D. The verification problem is therefore part of the development story itself; Geoffrey Hinton separately called independent verification organizations a good start, arguing that reliance on whistleblowers is not enough.

AI labs are selling governed domain workflows, not just models
Summary
Coverage start
4 days ago
Coverage end
3 days ago
Frequency
Daily
Published
3 days ago
Reading time
4 min
Research time
15 hrs 12 min
Documents scanned
1097
Documents used
8
Citations
20
Sources monitored
116 / 116
Insights
Skipped contexts
148
View
Source details
Source Docs Insights Status
Cohere 3 0
Shane Legg 0 0
AI at Meta 2 1
Prof. Anima Anandkumar 0 0
Ian Goodfellow 0 0
Chip Huyen 0 0
Oriol Vinyals 0 0
Nathan Lambert 6 1
Ashish Vaswani 0 0
Sherjil Ozair 0 0
Raquel Urtasun 0 0
Greg Brockman 2 1
Sebastian Raschka 0 0
Thomas Wolf 7 2
Jeremy Howard 0 0
LocalLLM 904 4
The Cognitive Revolution 2 2
Richard Socher 0 0
John Carmack 0 0
Mustafa Suleyman 0 0
Emad 12 2
Tim Dettmers 0 0
Geoffrey Hinton 0 0
Machine Learning Street Talk 0 0
hardmaru 0 0
swyx 0 0
Lukas Biewald 0 0
a16z 0 0
Nando de Freitas 0 0
Pieter Abbeel 0 0
Rowan Cheung 0 0
Percy Liang 0 0
Logan Kilpatrick 0 0
Interconnects 0 0
Latent.Space 1 1
ChinAI Newsletter 0 0
Big Technology 0 0
Machine Learning 22 0
Import AI 0 0
Latent Space 0 0
Gradient 0 0
Lex Fridman 0 0
Arxiv Insights 0 0
Aleksa Gordić - The AI Epiphany 0 0
Matt Wolfe 0 0
sarah guo 7 1
martin_casado 21 2
Marc Andreessen 🇺🇸 0 0
Elad Gil 0 0
François Chollet 1 0
Vinod Khosla 4 2
Yann LeCun 5 1
Fei-Fei Li 0 0
Ilya Sutskever 0 0
Jeff Dean 0 0
clem 🤗 3 1
Jim Fan 0 0
Sara Hooker 4 0
Soumith Chintala 0 0
Sebastian Ruder @ ACL 0 0
Dario Amodei 0 0
Google DeepMind 0 0
Demis Hassabis 0 0
Satya Nadella 0 0
Sam Altman 0 0
Elon Musk 11 0
Arthur Mensch 0 0
Aravind Srinivas 2 1
Aidan Gomez 3 1
OpenAI 0 0
Yannic Kilcher 0 0
Andrej Karpathy 0 0
Andrew Ng 0 0
Sundar Pichai 0 0
Kate Crawford 0 0
Gary Marcus 46 15
NVIDIA Blog 1 0
Jay Alammar 0 0
inFERENCe 0 0
arg min 1 0
Jack Clark 0 0
Two Minute Papers 0 0
Anthropic 5 3
OpenAI 5 1
Google DeepMind 5 1
Gary Marcus 0 0
Guillaume Lample @ NeurIPS 2024 0 0
Sarah Guo 0 0
Christopher Manning 0 0
Pieter Abbeel 0 0
Ilya Sutskever 0 0
Jerry Liu 0 0
Sebastian Raschka 2 2
Harrison Chase 0 0
Elad Gil 0 0
Sara Hooker 0 0
Aidan Gomez 0 0
Oriol Vinyals 0 0
Jeremy Howard 0 0
Simon Willison 0 0
Arthur Mensch 0 0
Andrej Karpathy 0 0
Sam Altman 0 0
Yoshua Bengio 0 0
Geoffrey Hinton 3 3
Andrew Ng 3 3
Demis Hassabis 2 2
Jeff Dean 0 0
Nathan Benaich 0 0
Clément Delangue 0 0
Fei-Fei Li 0 0
Ben Thompson 1 1
Dario Amodei 0 0
Yann LeCun 1 1
Percy Liang 0 0
Jack Clark 0 0