Anthropic Opens Near-Internal Access to External Security Teams, CEO Says Model Capability Development Should Be Slowed Down

Anthropic Opens Near-Internal Access to External Security Teams, CEO Says Model Capability Development Should Be Slowed Down

2026-09-12 15:53View Original

According to reports, Anthropic has announced a long-term on-site presence for third-party security teams, providing them with tools and access nearly equivalent to those of internal risk assessment teams. External reviewers can inspect training pipelines, safety protocols, and incidents, and independently publish their findings; Anthropic may only remove information related to security, legal compliance, or customer privacy, but cannot delete content due to unfavorable conclusions. CEO Dario stated that this approach goes further than any current AI company.

He also shifted his stance on "slowing down," noting that AI is increasingly involved in the development of next-generation AI systems, accelerating capability advancements—making pure safety research insufficient. The model's own capabilities must also be slowed down. He cited OpenAI’s Hugging Face incident as a warning, cautioning that within 6 to 12 months, similar agent swarms could establish botnets spanning large-scale internet infrastructures, potentially causing trillions in losses. He urged leading AI companies in democratic nations like the U.S. to unify safety standards and R&D speed, then attempt global coordination with China on AI development pace, while continuing to restrict advanced chip access and model distillation to maintain U.S. technological leadership.

Disclaimer: Contains third-party opinions, does not constitute financial advice

Recommended Reading

Florida and Texas Halt Flock License Plate Camera System Over Privacy Concerns

11 days ago
Florida and Texas Halt Flock License Plate Camera System Over Privacy Concerns

Caterpillar leverages its automated mining expertise for AI deployment, launching the Cat AI voice assistant and planning a $100 million investment over five years to train employees

13 days ago
Caterpillar leverages its automated mining expertise for AI deployment, launching the Cat AI voice assistant and planning a $100 million investment over five years to train employees

Flock Safety CEO Calls for Compromise on Privacy and Security, Company Faces Accusations of Surveillance Abuse

20 days ago
Flock Safety CEO Calls for Compromise on Privacy and Security, Company Faces Accusations of Surveillance Abuse

FTC Urged to Investigate AI Firm's Practice of Purchasing and Destroying Books for Training Data

21 days ago
FTC Urged to Investigate AI Firm's Practice of Purchasing and Destroying Books for Training Data

OpenAI to Launch Astra Within Weeks: Employees Already Testing New Version in Codex

23 days ago
OpenAI to Launch Astra Within Weeks: Employees Already Testing New Version in Codex

U.S. legal AI company Harvey launches its proprietary legal model Tenet, powered by the Moonshot AI Kimi K3 foundation model

24 days ago
U.S. legal AI company Harvey launches its proprietary legal model Tenet, powered by the Moonshot AI Kimi K3 foundation model

OpenAI Launches New Security Measures Following Hugging Face Incident: Enhanced Monitoring, Network Isolation, and 30-Minute Alert System

24 days ago
OpenAI Launches New Security Measures Following Hugging Face Incident: Enhanced Monitoring, Network Isolation, and 30-Minute Alert System