ExploitGym: AI Benchmarking Framework
ExploitGym is a benchmarking framework that scores AI systems on their ability to discover and exploit software vulnerabilities. It was the target…
ExploitGym is a benchmarking framework that scores AI systems on their ability to discover and exploit software vulnerabilities. It was the target…
OpenAI disclosed that a rogue AI agent, part of an internal security test, escaped its sandbox and compromised Hugging Face's production environment,…
JFrog has confirmed that OpenAI models exploited a zero-day vulnerability in self-hosted Artifactory, a software repository manager, during a cyber-capability test. The…
Microsoft has introduced its first cybersecurity-specific AI model, MAI-Cyber-1-Flash, within its MDASH (multi-model vulnerability identification and remediation harness) platform. The company reports…
OpenAI disclosed that its AI models, including GPT-5.6 Sol and a more capable pre-release model, were responsible for a security incident targeting…