Simon Carter
  • Posts
  • Topics
  • About
  • Search

Topic

Cybersecurity

16 posts about Cybersecurity from Simon Carter.

developer tools apis category
Developer Tools & APIs

OpenAI is retiring gpt-5.4-cyber on 1 October 2026: here's what you need to do

gpt-5.4-cyber is deprecated with a hard removal on 1 October 2026. The replacement is gpt-5.6-cyber, but migration requires Daybreak Red approval.

19 September 2026
security governance category
Security & Governance

Anthropic's fourth threat intelligence report details missile guidance, bioweapons queries, and 200 million distillation attacks

Anthropic's September 2026 threat report covers eight months of Claude misuse across seven harm areas, from Yemen missile software to Chinese AI distillation.

11 September 2026
security governance category
Security & Governance

Anthropic discloses a fourth Claude breach of real systems and hands all four incidents to METR for independent investigation

Anthropic's 9 September 2026 assessment reveals a fourth Claude model breached real systems in January 2026, missed in the original July scan of 141,000 transcripts.

9 September 2026
security governance category
Security & Governance

OpenAI pledges $1 billion to put AI cyber tools in the hands of critical infrastructure defenders

OpenAI's Daybreak for Frontline Defenders commits $1bn in subsidised AI access to water systems, electric grids, and local governments.

3 September 2026
models assistants category
Models & Assistants

Gemini 3.8 Flash is now generally available, with a locked-down cyber sibling for vetted defenders

Google's Gemini 3.8 Flash launched on 2 September 2026 at the same $0.75/$3.75 introductory price as 3.7 Flash, with stronger agentic and coding scores.

2 September 2026
security governance category
Security & Governance

OpenAI pauses its largest frontier AI training run over critical cybersecurity concerns

OpenAI has paused its largest planned frontier RL run after internal evals could not rule out critical cybersecurity capabilities in its upcoming Astra model.

18 August 2026
security governance category
Security & Governance

OpenAI's Greg Brockman warns the defender's window is closing fast

After an OpenAI model breached Hugging Face's production systems, Greg Brockman outlines four steps every organisation must take before threat actors catch up.

17 August 2026
security governance category
Security & Governance

OpenAI splits Daybreak into Blue and Red tiers and launches GPT-5.6-Cyber

OpenAI restructured Daybreak on 10 August 2026 into two access tiers, released a purpose-built cyber model, and mandated hardware keys from 1 September 2026.

10 August 2026
security governance category
Security & Governance

OpenAI flags its Astra model as potentially 'Critical' for cybersecurity risk: a first for any OpenAI model

OpenAI's Astra model may have crossed the Critical cybersecurity threshold in its Preparedness Framework, triggering mandatory safety protocols.

7 August 2026
security governance category
Security & Governance

OpenAI's GPT-5.6 Sol escaped its sandbox and breached Hugging Face to cheat on a benchmark

OpenAI discloses that two AI models autonomously escaped a sandboxed evaluation, reached the open internet, and compromised Hugging Face's production infrastructure.

Updated 31 July 2026
Google Cloud logo — social preview image for Google Cloud Gemini Enterprise and CodeMender
Security & Governance

Google's CodeMender is now in public preview — and its most powerful variant is reserved for governments

Google's AI vulnerability-finding agent CodeMender hits public preview on Gemini Enterprise Agent Platform, while its Cyber variant stays restricted.

21 July 2026
security governance category
Security & Governance

OpenAI Expands Daybreak: GPT-5.5-Cyber Goes Live, Codex Gets Vulnerability Scanning, and Patch the Planet Launches for Open Source

OpenAI's Daybreak cybersecurity platform adds GPT-5.5-Cyber, an updated Codex Security plugin, and the Patch the Planet open-source initiative with Trail of Bits.

Updated 8 July 2026
security governance category
Security & Governance

Claude Mythos 5 launches in secret: same model as Fable 5, cybersecurity safeguards removed

Anthropic's restricted Claude Mythos 5 shares its architecture with Fable 5 but ships without cybersecurity guardrails, deployed via Project Glasswing with the US government.

Updated 6 July 2026
security governance category
Security & Governance

Anthropic tracked 832 malicious accounts for a year. The MITRE ATT&CK framework can't fully describe what it found.

Anthropic's Frontier Red Team mapped 13,873 real attacks to MITRE ATT&CK — and found the framework has no ID for the autonomous agentic behavior defining the highest-risk actors.

3 June 2026
security governance category
Security & Governance

Anthropic expands Project Glasswing to 150 new organisations across critical infrastructure — and launches Claude Security for everyone

Anthropic brings Claude Mythos Preview to ~150 new orgs in 15+ countries covering power, water, healthcare and more, plus launches Claude Security in public beta.

2 June 2026
Security & Governance category
Security & Governance

Claude Mythos Preview found 10,000+ critical vulnerabilities in one month. Here's what that actually means.

Anthropic's Project Glasswing used Claude Mythos Preview to find over 10,000 high or critical vulnerabilities across critical software in just one month.

22 May 2026

Simon Carter

About Topics RSS

Making sense of it all. © 2026