Microsoft has developed an autonomous security system featuring red, blue, and green team AI agents, which will enter public preview on August 3.

Microsoft has developed an autonomous security system featuring red, blue, and green team AI agents, which will enter public preview on August 3.

      TL;DR: Microsoft unveiled Project Perception, an autonomous security system featuring red/blue/green AI agents. Its MAI-Cyber-1-Flash model achieves a 96% score on CyberGym, outperforming Mythos by 12 points, at 50% reduced costs. It will enter public preview on August 3.

      On Monday, Microsoft introduced Project Perception, an autonomous security system that utilizes three types of AI agents in a continuous feedback loop: red team agents to identify vulnerabilities ahead of potential attackers, blue team agents to assess the significance of identified risks, and green team agents to strengthen defenses throughout the environment. The system is set for public preview on August 3 and is characterized by Microsoft as a “new Cyber Stack” designed for a landscape where AI-driven attacks evolve faster than human defenses can respond.

      The initial performance metric is Microsoft’s proprietary MAI-Cyber-1-Flash model, integrated within its MDASH software vulnerability management tool, which achieved a 96% score on CyberGym, an industry benchmark for assessing vulnerabilities. This score exceeds Anthropic’s Mythos, which is recognized as the leading frontier model in cybersecurity, by 12 points. Additionally, Microsoft asserts that this configuration offers close to 50% savings compared to the existing MDASH production setup by utilizing appropriate models for specific tasks instead of processing everything through a single costly frontier model.

      Project Perception employs a multi-model architecture instead of depending on a single model for every function. Frontier models are used for complex reasoning tasks while specialized cybersecurity models are implemented for high-volume, low-latency operations. The system leverages Microsoft’s extensive insights into identities, endpoints, applications, data, clouds, and AI systems, enabling it to act across these domains rather than merely providing alerts. Microsoft’s AI has already detected a record number of vulnerabilities in its own software, and Project Perception aims to extend these capabilities to customers’ environments with agents operating continuously, rather than merely during monthly updates.

      The competitive jab at Anthropic appears intentional. Mythos became the go-to cybersecurity model after gaining significant attention due to a temporary White House ban, highlighting its offensive capabilities. Now, Microsoft claims its specialized model surpasses Mythos in performance at half the cost, citing its extensive training data from years of defending enterprise environments—an advantage Anthropic lacks. This month, the White House launched Gold Eagle to unify AI-driven cyber defense, and Project Perception positions itself as Microsoft’s platform to lead in defense operations. The public preview on August 3 will allow enterprises to evaluate it, raising the question of whether the 96% benchmark can withstand real-world attacks as opposed to synthetic assessments.

Other articles

Amazon requested the FCC's approval for 5,105 satellites intended for a network that transmits internet directly to your mobile device. Amazon requested the FCC's approval for 5,105 satellites intended for a network that transmits internet directly to your mobile device. Amazon submitted a proposal to deploy 5,105 satellites for direct-to-device connectivity. This initiative merges Kuiper and Globalstar resources, with a deployment goal set for 2028. SpaceX is already providing direct-to-device services in partnership with T-Mobile. AI dramas require performers, leading Chinese platforms to convert human faces into stock assets. AI dramas require performers, leading Chinese platforms to convert human faces into stock assets. Chinese platforms are compensating individuals for licensing their faces for AI-generated dramas, but the contracts may provide minimal protection once those digital likenesses are used outside the marketplace. Angola secures $321 million in its largest IPO to date with the Unitel launch. Angola divested a 15% share in Unitel for approximately $321 million, marking its largest initial public offering to date. The shares originate from a seizure in 2022 involving Isabel dos Santos. Dopl Technologies secures $6.3 million to introduce remote robotic ultrasound services in rural hospitals. Dopl Technologies secured $6.3 million to develop a telerobotic ultrasound system. Sonographers will control it remotely using haptic controllers. The next step is obtaining FDA clearance. The system promises a 90% reduction in wait times. AI providers are shifting from subscription models to consumption-based pricing. AI PCs serve as a safeguard. AI providers are shifting from subscription models to consumption-based pricing. AI PCs serve as a safeguard. Leading software companies are moving away from per-seat AI pricing in favor of token-based consumption. AI PCs that operate models locally provide businesses with an option to control expenses as cloud costs increase. Anthropic's open-weight models remain quiet following its own ban on Fable 5. Anthropic's open-weight models remain quiet following its own ban on Fable 5. Anthropic is the sole major AI laboratory that has not endorsed the open-weights letter. In June, it made a similar argument when Washington withdrew Fable 5.

Microsoft has developed an autonomous security system featuring red, blue, and green team AI agents, which will enter public preview on August 3.

Project Perception employs AI agents that continuously attack, investigate, and resolve security vulnerabilities. Microsoft's MAI-Cyber-1-Flash achieved a score of 96% on CyberGym, surpassing Mythos by 12 points.