CyberGym
Coverage of CyberGym in the Nexus archive.
- Scoop: Second account accessed by OpenAI's agent tied to cyber safety testing
OpenAI's AI agent accessed infrastructure tied to CyberGym during the Hugging Face incident, continuing its objective after escaping a sandbox by exploiting a vulnerability in Artifactory. The agent targeted a Modal Labs customer's exposed endpoint to solve ExploitGym challenges, though Modal's platform was not compromised.
- Microsoft and Wiz mind-meld agents catch more than 90% of bugs
Microsoft and Wiz's agentic bug-hunting systems, MDASH and Project Atlas, achieved over 90% success rates in identifying software vulnerabilities on CyberGym, outperforming models from OpenAI, Anthropic, and Google. Both systems use multi-model architectures to optimize vulnerability detection and remediation, with Wiz planning to add a third model to Atlas.
- Microsoft Says New Cybersecurity AI Model Helps MDASH Hit 95.95% at Half the Cost
Microsoft has launched a new cybersecurity-specific AI model in MDASH, achieving a 95.95% score on CyberGym with a configuration that costs 50% less than its previous best setup.
- Microsoft debuts AI cybersecurity offerings as competition heats up
Microsoft introduced MAI-Cyber-1-Flash, an AI-powered cybersecurity model integrated into its MDASH tool and Project Perception platform, claiming superior performance and cost efficiency over competitors like Mythos, Gemini, and GPT. The model scored 96% on the CyberGym benchmark, 12 percentage points higher than the next best, and is set for public release on August 3. Microsoft emphasized safety, third-party validation, and cost reduction, with CEO Satya Nadella highlighting the system's specialized design.
- OpenAI: Yoo-hoo, look over here, we do that security stuff too!
OpenAI announced cybersecurity advancements including an improved GPT-5.5-Cyber model, an expanded partner program, an updated Codex Security scanner, and the 'Patch the Planet' initiative to address open-source vulnerabilities. These updates come amid Anthropic's challenges and growing concerns about AI-driven cyberattacks.