Goodfire Launches Internal AI Agent Monitors on Baseten
Goodfire AI launched internal activation monitors for AI agents through Baseten, according to TechCrunch. The monitors inspect signals inside a model rather than using a separate model to review every output.
Customers can monitor risks including offensive hacking, chemical and biological weapons misuse, and reward hacking. They can configure the system to log flagged activity, send it for human review, or refuse a request.
In tests using Kimi K3, Goodfire said monitoring about 1 million exchanges cost roughly $185. The probes detected 93 percent of malicious hacking sessions and flagged 5.5 percent of harmless sessions for further review. Running four probes added less than 2 percent to the time before the model began responding.
We hope you enjoyed this article
Consider subscribing to one of our newsletters like Cybersecurity AI Weekly, AI Policy Brief or Daily AI Brief.
Also, consider following us on social media:
More from Cybersecurity
Oct 8 NewCore Wins Identity Management Platform of the Year Award Oct 8 Noma Security Wins 2026 SINET Innovator Award Oct 8 Rein Security raises $25 million for AI agent security Oct 8 Cogent Security Launches Attack Path Analysis for AI Agent Threats Oct 8 CUJO AI and F-Secure Partner on Home and Mobile SecurityMore from AI Safety
Oct 9 Fired OpenAI safety researchers dispute misconduct claims Oct 7 Former Anthropic Researcher Jacob Coxon Discusses AI Risks With Jon Stewart Oct 6 Potts Law Firm Seeks Consolidation of Grok Lawsuits Oct 6 Anthropic Alerts Florida Police Over Alleged Threat in Claude Chat Oct 6 Former Anthropic Researcher Jacob Coxon to Testify at New York City AI HearingCybersecurity AI Weekly
Weekly newsletter about AI in Cybersecurity.
Market report
2025 Generative AI in Professional Services Report
Thomson Reuters
This report by Thomson Reuters explores the integration and impact of generative AI technologies, such as ChatGPT and Microsoft Copilot, within the professional services sector. It highlights the growing adoption of GenAI tools across industries like legal, tax, accounting, and government, and discusses the challenges and opportunities these technologies present. The report also examines professionals' perceptions of GenAI and the need for strategic integration to maximize its value.
Read moreYou may also like
US Agencies Accuse Chinese AI Firms of Industrial Scale Model Distillation
Chinese AI Agents Deceived Evaluators in Controlled Tests
PRE Security Launches AgentGuard for AI Agent Monitoring
Researchers Track Chinese AI Agent Fleet Querying Alibaba's Amap
ScamAdviser Introduces AgentLooker.ai to Protect AI Agents From Scams
Daily AI Brief: the AI news that matters, in your inbox.