Artificial Intelligence

Microsoft Launches its First Cybersecurity AI Model, MAI-Cyber-1-Flash

San Francisco: Microsoft unveiled its first cybersecurity-focused AI model, MAI-Cyber-1-Flash at an event in San Francisco. The launch also included a new agentic security platform called Project Perception.

The model does not work as a standalone product. It is built to power MDASH, Microsoft’s harness dedicated to identifying and fixing software vulnerabilities. Access is limited to select customers through an Azure AI Foundry private preview, and requires extra approval, since tools that find vulnerabilities can be misused by attackers too.

Strong benchmark results

Microsoft says the results are notable. In testing on the CyberGym cybersecurity evaluation framework, MAI-Cyber-1-Flash paired with MDASH and GPT-5.4 outperformed Google’s newly launched 3.5 Flash Cyber, OpenAI’s GPT-5.6 Sol, and Anthropic’s Mythos 5 in finding vulnerabilities. The MDASH system scored 95.95 percent on CyberGym using this combination.

Mustafa Suleyman, CEO of Microsoft AI, called the results a major milestone. He said the combined system “beats out Gemini, GPT 5.5 Cyber, GPT 5.6 Sol, and Mythos 5 on Cyber Gym, which is the primary benchmark.”

Built for cost savings, not just performance

A key part of the pitch is affordability. Microsoft says MAI-Cyber-1-Flash is designed to handle up to 90 percent of MDASH’s tasks, while GPT-5.4 is reserved only for the hardest 10 percent. This routing approach is claimed to cost 50 percent less than the company’s current best MDASH setup, which uses a combination of GPT-5.4, GPT-5.4 mini, and GPT-5.3 Codex.

Technically, the new model is a sparse mixture-of-experts fine-tune of MAI-Code-1-Flash, carrying 137 billion total parameters with 5 billion active at a time, and supports a 256k context window. It was trained only for defensive work such as patching bugs, not offensive tasks, which is why it scores zero by design on ExploitGym, a benchmark for writing exploits.

Independent scrutiny

Some outlets have flagged caveats. The 95.95 percent figure wasn’t reflected on CyberGym’s public leaderboard as of July 28, which still showed lower scores for competing systems. Details like token usage, call volume, and compute allocation behind the cost claim haven’t been disclosed, making the comparison hard to independently verify.

Part of a bigger industry trend

The launch comes just days after Google Cloud rolled out its own tool, CodeMender, with AWS, Nvidia, and others also racing to defend against AI-driven cyberthreats. Analysts note the emerging theme is that cybersecurity increasingly needs a multi-model approach, and Microsoft’s move also reflects its broader push to build its own AI model family and reduce reliance on costlier frontier models from partners like OpenAI.

Microsoft also signed on to Nvidia’s newly launched Open Secure AI Alliance, aimed at sharing security tools across the industry.

Shobhit Kalra

Shobhit Kalra is the Chief Sub Editor at Tea4Tech, with over 12 years of experience across digital media, digital marketing, and health technology. He is responsible for editorial review, content structuring, and quality control of articles covering software, SaaS products, and developments across the technology ecosystem. || At Tea4Tech, Shobhit oversees content accuracy, clarity, and adherence to editorial standards, ensuring published stories meet the newsroom’s guidelines for originality, sourcing, and consistency.

Recent Posts

Facebook Blocks PM Modi’s NEET Video, Meta Calls It an Error

New Delhi: The Indian government was caught off guard when social media giant Meta blocked…

1 day ago

Moonshot’s Kimi K3 Becomes Largest Open-Weight AI Model Ever

BEIJING: Moonshot AI releases Kimi K3, a 2.8 trillion-parameter model that instantly becomes the largest…

1 day ago

Fireworks AI Hits $17.5B as Enterprises Flee Frontier API Prices

SAN MATEO, Calif.: Fireworks AI closes a $1.505 billion Series D at a $17.5 billion…

2 days ago

AegisAI Lands $36M Series A to Secure the Agentic Enterprise

SAN FRANCISCO: AegisAI raises $36 million in Series A funding to secure enterprises against threats…

2 days ago

Paper Raises $34M as AI Coding Agents Redraw Design’s Borders

SAN FRANCISCO: Paper raises $34 million in Series A funding led by Accel and ICONIQ…

2 days ago

Anthropic Launches Claude Opus 5, Says it Costs Less and Does More

New Delhi: Anthropic has released a new AI model called Claude Opus 5. The company…

2 days ago