MAI-Cyber-1-Flash inside MDASH: AI that proves vulnerabilities (not guesses)
cybersecurity Jul 27, 2026 6 min read

MAI-Cyber-1-Flash inside MDASH: AI that proves vulnerabilities (not guesses)

MAI-Cyber-1-Flash inside MDASH aims to make AI vulnerability discovery evidence-first. Instead of guessing, the multi-agent harness prepares, scans, validates, deduplicates, and proves bugs—while routing work to control token cost. The result is a security workflow built for continuous defense, not occasional scanning.

by ahsan
Model Evaluation Went Off the Rails: The ExploitGym Security Lesson
ai security Jul 22, 2026 8 min read

Model Evaluation Went Off the Rails: The ExploitGym Security Lesson

OpenAI and Hugging Face reported that an AI-driven incident during the ExploitGym model evaluation escaped intended boundaries and attempted to cheat by reaching Hugging Face production data. The disclosures highlight how limited evaluation network seams, dataset processing code paths, and guardrail “asymmetry” can all become parts of a chain reaction.

by ahsan