Business only (hides war / politics / culture / sports)
Wednesday, Sep 16, 2026
1 Baseten and Hugging Face Launch Open AI Safety Tools After OpenAI Swarm AttackPVT:BSTNPVT:HGFCPVT:OPAI 🤖 AI Sep 16, 2:11 PM EDT 6/5
1
Baseten and Hugging Face Launch Open AI Safety Tools After OpenAI Swarm AttackPVT:BSTNPVT:HGFCPVT:OPAI
🤖 AI Sep 16, 2:11 PM EDT 6/5
Base Labs is developing open safety research and monitoring methods to be integrated into Baseten's inference infrastructure. The initiative, created with collaborators Hugging Face and Goodfire AI, provides training for models to follow explicit policies and includes runtime failure detection with linked intervention controls to create a transparent standard for how open AI models are deployed.
Baseten will offer the safety infrastructure as a managed service to its customers to establish security standards for open-source models. The effort follows an agent swarm attack by OpenAI on Hugging Face, which Baseten describes as the type of failure providers must prepare for as open models reach parity with closed frontier labs.
You're reading an older version of the story.
Goodfire Probes Cut AI Monitoring Costs 90% to Detect Reward Hacking 24 tweets • 14 sources