AIToday

Microsoft launches cybersecurity AI model, Perception platform

TechCrunch AI1h agoSend on LINE
Microsoft launches cybersecurity AI model, Perception platform

Key takeaway

Microsoft has unveiled MAI-Cyber-1-Flash, a cybersecurity-focused AI model, alongside Perception, a new platform that deploys teams of AI agents to automate vulnerability detection and remediation at enterprise scale. The company claims its model outperforms competitors including Anthropic, Google, and OpenAI on the Cyber Gym benchmark. Perception will enter preview on November 3, positioning Microsoft to compete directly with Anthropic's Mythos and OpenAI's security solutions as AI-driven cyberattacks grow.

Summaries like this, in your inbox every morning.

Sign up free →

3 Key Points

  • What happened

    Microsoft announced MAI-Cyber-1-Flash, a specialized cybersecurity model paired with Perception, a new AI platform that deploys teams of agents to identify and fix software vulnerabilities. The company claims MAI-Cyber-1-Flash outperforms competitor models (Gemini, GPT 5.5 Cyber, GPT 5.6 Sol, and Mythos 5) on the Cyber Gym benchmark. Perception will be available in preview starting November 3.

  • Why it matters

    Cybercriminals are increasingly using AI in attacks, so Microsoft is positioning Perception as a way for enterprise defenders to match attackers' speed and scale. According to Dave Weston, Perception's lead engineer, the platform cuts security work from hours of manual effort across multiple specialists down to minutes, including discovering issues, prioritizing them, and generating code fixes.

  • What to watch

    Perception enters a crowded field—Anthropic launched Mythos earlier this year through its Glasswing partner program, and OpenAI released a security solution in May through Day Break. Microsoft is shipping MAI-Cyber-1-Flash into production immediately.

In Depth

Microsoft on Monday launched two interconnected cybersecurity tools at a small event in San Francisco: MAI-Cyber-1-Flash, a specialized AI model built to find vulnerabilities in complex codebases, and Perception, a new platform designed to orchestrate teams of AI agents for automated security workflows.

MAI-Cyber-1-Flash is built to work within MDASH, Microsoft's existing harness for software vulnerability identification and remediation. According to Mustafa Suleyman, the co-founder of DeepMind and current CEO of Microsoft AI, the model combines with GPT 5.4 inside the MDASH harness and beats competitors on Cyber Gym, the primary cybersecurity benchmark. The beaten models include Gemini, GPT 5.5 Cyber, GPT 5.6 Sol, and Mythos 5. Microsoft claims the model is "significantly more powerful (and more cost-effective)" than competitor offerings. "We're shipping this into production immediately," Suleyman said.

Perception operates by deploying three types of AI agent teams. Red teams provide detailed simulations of potential attacks, including context about threat actors and likely exploitable vulnerabilities. Blue teams detect and triage existing bugs. Green teams take corrective actions against those bugs. The platform can also integrate with MDASH. Dave Weston, Perception's lead engineer, described the impact as a massive efficiency gain: "We've gone from this taking hours and hours of manual work from multiple specialized folks across the security organization — appsec hunters, remediation engineers, you name it — and in minutes, we have a fix for all of this. Not only do we discover the issues and prioritize them, but we have detection, posture fixing, and even a code fix."

Microsoft is positioning these tools against a backdrop of rising AI-enabled cyberattacks. Hayete Gallot, Microsoft's VP for security, framed Perception as a way for enterprise defenders to "defend against AI with AI at the scale and speed that the attackers have." Perception will enter preview on November 3, arriving in a competitive field that includes Anthropic's Mythos (released earlier this year through a limited Glasswing partner program) and OpenAI's security solution (launched in May through Day Break).

Context & Analysis

Microsoft's entry into specialized cybersecurity AI reflects the company's broader push into AI-driven enterprise tools and marks a significant competitive move against established players. By pairing a dedicated vulnerability-detection model with an agentic platform (Perception), Microsoft is addressing a real enterprise pain point: the manual, time-intensive work currently required to find and fix bugs at scale. The announcement comes in a context where cybercriminals themselves are deploying AI, creating pressure on defenders to automate and accelerate their own workflows—a dynamic Hayete Gallot, Microsoft's VP for security, explicitly invoked when describing the need to "defend against AI with AI at the scale and speed that the attackers have."

The competitive landscape is heating up. Anthropic launched Mythos earlier this year through a limited Glasswing partner program, and OpenAI released its own security solution in May. Microsoft's immediate shift to production deployment and preview availability by November 3 signals aggressive intent to capture market share quickly. The claim that MAI-Cyber-1-Flash outperforms Gemini, GPT 5.5 Cyber, GPT 5.6 Sol, and Mythos 5 on Cyber Gym—framed by CEO Mustafa Suleyman as "the golden benchmark"—positions Microsoft as the performance leader in this emerging category, though independent validation of that claim remains pending.

FAQ

When will Perception be available?
Perception will be available in preview on November 3.
How does Perception work?
Perception uses agentic red teams (which simulate potential attacks), blue teams (which detect and triage bugs), and green teams (which take corrective actions). It can integrate with MDASH, Microsoft's tool for software vulnerability identification and remediation.
How does MAI-Cyber-1-Flash compare to competitors?
According to Microsoft, MAI-Cyber-1-Flash beats out Gemini, GPT 5.5 Cyber, GPT 5.6 Sol, and Mythos 5 on Cyber Gym, which the company describes as the primary cybersecurity benchmark.

Get the latest Large Language Models news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Discussion

No comments yet. Be the first to share your thoughts!

Log in to join the discussion

Related Articles

Stay ahead with AI news

Get curated AI news from 200+ sources delivered daily to your inbox. Free to use.

Get Started Free

Free · takes 30 seconds · unsubscribe anytime