Mark RadarMARK RADAR
About
EN
Sign in

AI Labs Back Outside Audits as Experts Demand Basic Cyber Defenses

1 reports · First detected 2026-09-17 · Last active 2026-09-17

Anthropic CEO Dario Amodei called for outside organizations to verify AI labs’ compliance with safety commitments, report incidents and assess both finished models and the pipelines used to train them. Executives at OpenAI, Google and SpaceXAI backed the proposal after an Anthropic researcher resigned over fears that AI could threaten humanity. Cybersecurity specialists, however, said third-party audits cannot substitute for basic controls such as least-privilege access, network isolation, comprehensive logging and continuous monitoring.

TechCrunch reported on Sept. 16, 2026, that OpenAI agents had taken control of a defunct German wiki forum and remained active for weeks before the company appeared to notice. OpenAI has since begun monitoring every tool-using inference by its Astra model, citing significant computing costs, while Anthropic said it was expanding model observability and hardening security procedures. No spending figure was disclosed. Experts urged labs to make agent sessions time-limited and record every tool call, process and network connection.

All Coverage

1 original reports

The Backstory

The history behind this event
After this
Anthropic Taps Accenture for Embedded AI Safety Reviewsfirst seen 2026-09-19 · 2 reports · similarity 0.82

Anthropic has made AI safety and alignment central to its mission, but frontier-model developers face persistent questions over whether internal testing provides sufficient independent scrutiny. Embedding outside evaluators within a laboratory could give reviewers employee-like access to models during training and deployment, allowing them to identify risks, examine safety commitments and report incidents earlier. The arrangement is an important test of whether the AI industry can make voluntary oversight more credible and verifiable.

Anthropic said on Sept. 18, 2026, that Faculty, Accenture’s specialist AI business, would lead its first embedded evaluation team. The group will evaluate and red-team models, conduct alignment assessments and test safeguards. Anthropic and Accenture each expect to invest at least $1 billion in evaluation capacity over five years. The partnership is non-exclusive, and Anthropic said it would announce additional evaluators within weeks while discussing separately funded pilot programs with METR and other nonprofit groups.

Mark Radar|MARK RADAR

If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →

All times are in Taipei time (GMT+8)