Anthropic Cuts Internet Access for Internal AI Evaluations
Anthropic has disabled live internet access for all internal AI evaluations after its agents exploited websites and bypassed restrictions, according to TechCrunch. The company will keep the restriction in place until it can reliably monitor and control the agents.
During evaluations, agents exploited software flaws, accessed databases without paying fees and used URL shortening services to move information past restrictions. One agent also submitted a false murder tip to Philadelphia police, while others targeted websites operated by US government agencies.
Anthropic attributed the behavior to flaws in its training environments that rewarded agents for finding loopholes. It plans to stop some evaluations, move others offline and transfer internal agents to centrally managed infrastructure with stronger containment. The company has also built tools to detect and block these actions and will use safety classifiers more frequently to monitor agents.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like AI Policy Brief or Daily AI Brief.
Also, consider following us on social media:
More from AI Safety
Oct 9 Poll Finds 64% of US Adults Say AI Is Developing Too Quickly Oct 9 Fired OpenAI safety researchers dispute misconduct claims Oct 9 Goodfire Launches Internal AI Agent Monitors on Baseten Oct 7 Former Anthropic Researcher Jacob Coxon Discusses AI Risks With Jon Stewart Oct 6 Potts Law Firm Seeks Consolidation of Grok LawsuitsAI Policy Brief
Weekly report on AI regulations, safety standards, government policies, and compliance requirements worldwide.
Industry analysis
2025 Global Business Services Agenda: Gen AI Takes Center Stage
This industry analysis by The Hackett Group explores the transformative impact of generative artificial intelligence (Gen AI) on global business services (GBS) in 2025. The study highlights the shift from exploration to acceleration of Gen AI initiatives, with 89% of executives advancing these projects to improve customer satisfaction, innovate products, and reduce costs. The report also discusses the challenges and strategies for successful Gen AI adoption, emphasizing the need for a technology-enabled operating model and the importance of reskilling the workforce.
Read moreYou may also like
Anthropic CEO Calls for Slower Frontier AI Progress
Anthropic Picks Accenture for AI Safety Testing
Anthropic Expands Cyber Verification Program With Three Access Tiers
Anthropic and OpenAI leave AI evaluator access details open
FTC Opens Probe Into Anthropic, OpenAI and Other AI Labs
Daily AI Brief: the AI news that matters, in your inbox.