Anthropic opens safety reviews to outside evaluators

Dario Amodei also wants slower capability gains, but shared standards still require other labs and governments to opt in.

In partnership with

SPD-BEEHIIV:486868b9-607e-42f0-a24b-53e2b733bc03:R11:8a1ce6965503ed0b0f37b62b
Dario Amodei also wants slower capability gains, but shared standards still require other labs and governments to opt in.
Superpower DailyRead online/Account
Weekly digest / The Weekly DigestSunday, September 13, 2026
Our toolsSuperpower ChatGPT/WFH.team/Snipman

This week's briefing

What happened this week

This week, AI safety shifted from a statement of principle to a question of operating rules: Anthropic opened its practices to outside evaluation, OpenAI took a 2026 IPO off the table, and Washington’s role remained unsettled. At the same time, builders got new agent guidance while deployment, distribution, and capacity decisions kept moving.

Inside this week's digest
01Anthropic says independent evaluators will get employee-level access and can publish findings without company editorial control.
02OpenAI says it will not pursue an IPO in 2026 while safety, alignment, business, and social-readiness work remains unfinished.
03The Pentagon says roughly 90% of its classified AI workloads have moved off Anthropic, with the remainder due by September’s end.
04OpenAI’s Codex guidance favors leaner prompts, clearer completion states, and approval gates tied to concrete risks.
Listen to this newsletterAudio edition / About 3 min
Anthropic’s Dario Amodei Calls for Slower AI Progress and Outside Safety Reviews

Lead story / security risk

Anthropic gives outside reviewers access to its safety work

Anthropic CEO Dario Amodei is calling on frontier AI developers to slow capability gains while putting an immediate oversight mechanism in place at his own company. Anthropic says independent safety evaluators will receive employee-level access, matching the access available to its internal risk-assessment teams.

Those evaluators can examine safety practices and report incidents, and may publish findings without Anthropic’s editorial control. Amodei says the arrangement takes effect immediately, creating a standing review mechanism at Anthropic rather than an industry-wide requirement.

The broader agenda is a request for alignment beyond the company. Amodei proposes permanent independent evaluation across frontier AI companies, common safety standards among democratic countries, and coordination between democratic governments and authoritarian states on shared interests, including a ban on using AI to develop biological weapons.

Amodei’s case for slower progress rests on the possibility that increasingly capable systems could help create their successors, accelerating development faster than people can understand or control it. Anthropic has made its own access commitment; whether slower capability gains and shared standards follow depends on other companies and governments choosing to participate.

Read full story  ↗

Your take

Should AI labs accept shared limits on capability gains?

Join the discussion  →
 

How Jennifer Aniston’s LolaVie brand grew sales 40% with CTV ads

For its first CTV campaign, Jennifer Aniston’s DTC haircare brand LolaVie had a few non-negotiables. The campaign had to be simple. It had to demonstrate measurable impact. And it had to be full-funnel.

LolaVie used Roku Ads Manager to test and optimize creatives — reaching millions of potential customers at all stages of their purchase journeys. Roku Ads Manager helped the brand convey LolaVie’s playful voice while helping drive omnichannel sales across both ecommerce and retail touchpoints.

The campaign included an Action Ad overlay that let viewers shop directly from their TVs by clicking OK on their Roku remote. This guided them to the website to buy LolaVie products.

Discover how Roku Ads Manager helped LolaVie drive big sales and customer growth with self-serve TV ads.

The DTC beauty category is crowded. To break through, Jennifer Aniston’s brand LolaVie, worked with Roku Ads Manager to easily set up, test, and optimize CTV ad creatives. The campaign helped drive a big lift in sales and customer growth, helping LolaVie break through in the crowded beauty category.

 
WFH.team logoA tool for your workflowWFH.teamA focused feed of carefully selected remote roles and practical work-from-home resources.Remote work, without the noisy job-board scroll
 
Browse remote roles  ↗
OpenAI Says It Will Skip a 2026 IPO Over AI Safety Concerns

business

OpenAI rules out a 2026 IPO over safety concerns

Sam Altman says OpenAI will not pursue an IPO in 2026, calling it an ill-advised moment while safety, alignment, business readiness, and society’s response to more capable AI remain unresolved. He left later timing open. OpenAI has discussed capability pauses and coordination, but announced no measurable readiness test, specific safeguards, or development freeze.

Continue reading  ↗
Arm CEO Says AI Could Help Cure Cancer Within His Lifetime

business

Arm CEO says AI could help cure cancer, but current systems fall short

Arm CEO Rene Haas says increasingly capable AI and computing could eventually help cure cancer, but acknowledges that current systems cannot model how the disease affects a DNA marker. The nearer-term evidence is narrower: the NHS says AI-powered X-ray tools helped more than 4 million patients receive faster lung diagnoses earlier this year. Haas also says chip shortages are constraining humanoid-robot deployment.

Continue reading  ↗
OpenAI Publishes a Leaner Prompting Playbook for Codex Agents

tools

OpenAI tells developers to use leaner Codex prompts

OpenAI’s guidance for GPT-6 Astra asks teams to trim broad skill descriptions, turn AGENTS.md files into conditional maps, and define exactly what completion means. It retains approval gates for production access, destructive migrations, external effects, and credentials while permitting safe, reversible local work. OpenAI says Astra can stop early when broad workflows lack a clear end state.

Continue reading  ↗
Pentagon Says It Has Moved 90% of Classified AI Work Off Anthropic

government action

Pentagon moves 90% of classified AI work off Anthropic

The Defense Department says roughly 90% of its classified AI workloads have moved away from Anthropic, with the remainder expected to transition by the end of September. The break followed a contract dispute over Anthropic’s proposed restrictions on mass surveillance and fully autonomous lethal weapons. The Pentagon is building a multi-vendor classified AI stack using OpenAI, Google, and xAI while Anthropic’s legal challenge remains pending.

Continue reading  ↗
 

Understand how AI is used—wherever your workforce runs it.

Knowing which AI tools employees use is only the start.

Harmonic Security reveals why they’re used and the value they create, classifying tasks by business use case across tools, shadow AI, assistants, and agents.

See what drives productivity, where adoption grows, spend overlaps, and sensitive data moves. Then scale with confidence.

 
The Weekly Digest themed section header

Who tests powerful models, whether competitors can coordinate on safety, and what a voluntary slowdown would require.

Policy file 01
 
 
 
 
 
 
 
regulation
Senate AI Bill Hits Dispute Over Who Tests Powerful ModelsRead story ↗
Policy file 02
 
 
 
 
 
 
 
government action
OpenAI Seeks Congress’s View on a Coordinated AI SlowdownRead story ↗
Policy file 03
 
 
 
 
 
 
 
policy
Sam Altman Reportedly Told Staff OpenAI Could Slow AI-Agent Development With Other LabsRead story ↗
Missed Wednesday's Workbench? Five standout AI tools live in their own visual edition, keeping this digest focused. Open the Workbench ↗
 

Weekend tool drop

5 AI tools worth knowing this weekend

Selected for fit, not rank
QApilot MCP for Android logoQApilot MCP for AndroidRuns plain-English Android tests and saves passes as replayable Gherkin cases.Best for / Android teams using coding agentsOpen ↗
Youkti logoYoukti Tracks accounts and deals to recommend the next sales action from signals and history.Best for / AEs, RevOps, and outbound teamsOpen ↗
TIM PG logoTIM PGMasks clipboard and document data locally before sending text to LLMs.Best for / Privacy-conscious Windows AI usersOpen ↗
Suno v6 logoSuno v6Creates and edits music from text, audio, images, or video inputs.Best for / Creators iterating on AI musicOpen ↗
Live Captions by Subanana logoLive Captions by SubananaDelivers live translated captions or audio to audience phones, screens, and broadcast feeds.Best for / Multilingual events and presentationsOpen ↗
 
Quick reads
Meta’s Muse Reaches No. 2 on U.S. App Store, Sensor Tower EstimatesMeta’s Muse reaches No. 2 in the U.S. App Store ↗launch
Roblox Adds AI Game Tools and Plans Browser and Standalone Game AccessRoblox expands AI game tools and plans wider distribution ↗launch
OpenAI Halts New $200 Pro Sign-Ups as Astra Demand Strains CapacityOpenAI pauses new $200 Pro sign-ups over Astra demand ↗platform shift
Apple Introduces iPhone Duo With an AI-Assisted HingeApple uses AI in iPhone Duo hinge production ↗launch
 

The Internet Had a Point

From the timelineOpenAI’s Trojan Horse of Free Models
Screenshot of a social post by u/userusertion titled “OpenAi: ‘Just being helpful’,” labeled “Funny.” It shows a cartoon Trojan horse entering an ancient city while people emerge from it. An overlaid OpenAI card reads: “We’re giving scientists, mathematicians, and engineers free access to our frontier models—starting with 10,000 researchers expanding to 100,000 through 2027….”
 
From our network. Tools built for the way you work. Useful products from the team behind Superpower Daily.
WFH.team logo
Reader tool / From our teamWFH.teamFind carefully selected remote roles and practical resources for distributed work.Remote work, without the noisy job-board scroll
Browse remote roles  ↗
WFH.team product preview

Reader check-in

Help shape tomorrow's briefing

One click tells us what to keep, improve, or tighten.

01Useful02Interesting03Too long
Superpower Daily tracks the companies, models, products, tools, policy decisions, and cultural shifts moving AI.Follow Superpower Daily
X
LinkedIn
Discord
YouTube
RSS
Follow on Google News
Sponsor Superpower DailyEmail preferences