HackMyIP
← Back to News
2026-09-02 The Hacker News

Google, Anthropic, OpenAI Unveil Cyber AI Models and Safeguards

AI SecurityVulnerabilityThreat Intel

Major AI labs are doubling down on defensive cybersecurity, with Google, Anthropic, and OpenAI rolling out specialized models, access programs, and safeguards aimed at giving defenders an edge over attackers. Google on Wednesday announced Gemini 3.8 Flash Cyber, its most capable cybersecurity model to date, and made it available through a new initiative called the Fairwind Program. The program grants early access to high-priority defenders including governments, healthcare providers, and telecommunications services, and is already supported by over 650 partners such as CrowdStrike, Datadog, Menlo Security, Palo Alto Networks, and Snowflake. According to Tulsee Doshi, senior director of product management, and Raluca Ada Popa, Gemini Security Lead at Google DeepMind, the model demonstrates frontier-level performance in autonomous vulnerability discovery and surpasses larger models from rivals, including Anthropic's Mythos 5 and OpenAI's GPT-5.6 Sol and GPT-5.5-Cyber. Google noted it prioritized defensive capabilities like vulnerability fixing over offensive capabilities such as exploitation, marking a deliberate ethical stance in model design.

Anthropic simultaneously debuted Claude Fable 5.1 and Claude Mythos 5.1, both featuring tiered safeguards. While Fable 5.1 is approved for identifying software vulnerabilities, Mythos 5.1 is restricted to trusted access programs and supports cybersecurity and life sciences research. Anthropic confirmed that more sensitive tasks such as penetration testing, exploit generation, and binary-based vulnerability scanning will continue to be routed to its Opus models. The company reported that Mythos 5.1 refused malicious agentic coding and computer-use requests at rates comparable to Mythos 5, Sonnet 5, and Opus 5, and ranked as its most robust model to date on external prompt injection benchmarks. Alongside the model releases, Anthropic introduced Enterprise Frontier Safeguards (EFS), a solution that combines zero data retention (ZDR) privacy with state-of-the-art misuse detection, giving enterprises full control over how their data is reviewed and stored — a critical consideration for teams responsible for managing SSL/TLS configurations across enterprise environments.

The coordinated push from the three largest AI providers signals a maturing consensus that frontier models must ship with robust guardrails on both ends of the security spectrum. With critical infrastructure operators, healthcare networks, and government agencies facing relentless threats, the early-access models offered through Google's Fairwind Program and Anthropic's tiered release strategy could reshape how defenders evaluate and deploy AI tooling. Organizations seeking to harden their own environments before integrating AI-driven defenses should run a port scanner to identify exposed services and review any credentials against an email breach checker to catch compromised accounts before they become initial access vectors.

Source: The Hacker News →

Related Tools

Check whether this kind of story affects you — free, no signup:

My IP →IP Lookup →Privacy Checkup →

Related Guides

Learn the background behind this story:

What is my IP and why it matters →IP address security →How to stop being tracked online →