Sakana AI Ships Fugu-Cyber Multi-Agent Model with 86.9% Security Benchmark
Tags AI · Infrastructure · Enterprise

Sakana AI has released Fugu-Cyber, a multi-agent orchestration model that claims state-of-the-art performance on UC Berkeley's CyberGym security benchmark at 86.9%, outperforming leading competitors like GPT-5.5-Cyber and Mythos-Preview. The model represents a significant advancement in agentic security systems, though benchmark methodology and performance claims require independent verification before production deployment.
Technical significance
Fugu-Cyber's 86.9% performance on CyberGym represents a significant leap in agentic security capabilities, potentially transforming how organizations approach cyber defense. The multi-agent orchestration approach demonstrates that composite AI systems can outperform monolithic models, suggesting a paradigm shift toward more modular, specialized AI architectures for complex security challenges.