Radar / AI security / OpenAI, Anthropic and Google DeepMind jointly…
OpenAI, Anthropic and Google DeepMind jointly unveil cyber-focused safety models and safeguards
The three rival labs disclosed weeks of behind-the-scenes coordination and jointly released a set of cyber-focused AI models, safeguards, and access programs — a first for the fiercely competitive labs.
Why it matters
Unprecedented joint safety release from competing frontier labs, aimed squarely at the dual-use cyber capability problem (models that defend and attack).
What you could build with it
A cybersecurity vendor could align its agent red-teaming and guardrail products with the labs' shared cyber safeguards, selling joint-standard-compliant AI security reviews to regulated enterprises.
Does it hold up?
Too early to judge — a coordination announcement with few public details on the actual models, safeguards, or how joint maintenance will work going forward.
Built with OpenAI, Anthropic and Google DeepMind jointly…
- Radar Digital: coordinated slowdown debate, external evals adopted by Anthropic and OpenAIarticle · Companion reporting on the labs' coordinated safety positioning, external evaluation adoption, and the NYC AI hearing context.
Learn more
First spotted on article: source.
More AI security
Cloudflare security-audit-skill: multi-phase security audits for coding agentsA coding-agent skill that runs multi-phase security audits with independently verified, machine-checkable findings.security · JEV 0.63LiveNerf: a pre-registered benchmark for post-release model driftOpen-source, pre-registered 30-day benchmark that detects whether Claude Opus 5.5 quietly gets worse after launch.security · JEV 0.61Sandlock 0.8.8: deferred commit for agent sandboxesThe process-based Linux AI sandbox (no container, no VM) ships deferred commit: every run returns a changeset of what…security · JEV 0.58OpenAI disrupts Moonshot-linked 'adversarial distillation' campaignOpenAI disclosed it shut down a coordinated campaign by 15,000+ users to extract hidden model reasoning, attributing a…security · JEV 0.58ClawSecure: free security scanner for OpenClaw AI agent skillsFree scanner that audits OpenClaw agent skills for vulnerabilities; the maker's audit of 2,890+ OpenClaw skills found…security · JEV 0.5Google DeepMind SynthID BioWatermarking technology that embeds an imperceptible, verifiable signature into AI-designed protein sequences and…security · JEV 0.49
Get the week's best AI launches, plus 3 ideas worth building
One email every Saturday. Ranked by traction, not hype. Free.