cyber
22 stories and discussions about cyber, aggregated from every source we track.
GLM-5.3 can autonomously build end-to-end cyber exploits, but unlike other frontier models, it was released without meaningful safeguards to limit misuse.
We conducted cyber evaluations of Anthropic’s Claude Mythos Preview and found continued improvement in capture-the-flag (CTF) challenges and significant improvement on multi-step cyber-attack simulations.
Now that DARPA’s AI Cyber Challenge (AIxCC) has officially ended, we can finally make Buttercup, our CRS (Cyber Reasoning System), open source!
This guide answers the ten questions practitioners, legal teams, and buyers are actually asking — each chapter stands on its own.
Five people whose work supported missions at U.S Cyber Command and the National Security Agency at Fort George G. Meade died by suicide within a month of each other this summer. Lawmakers are asking questions and…
The Cyber Resilience Act (CRA) aims to make sure all digital products are safe from cyber threats. This rulebook requires that devices and software are designed, updated, and maintained to protect users in our…
Google claims Gemini 4 Argon stacks up against OpenAI and Anthropic’s models.
Contribute to usnistgov/caisi-cyber-evals development by creating an account on GitHub.
The PRC-based company Z.ai (formerly known as Zhipu AI) released a new AI model, GLM-5.3, on August 14, 2026.
Traditional virtual machines are inadequate for isolating cyber-capable autonomous agents. Tests using GPT-5.6-Cyber indicated multiple escape attempts due to kernel flaws. While Firecracker provided some containment,…
The Artificial Analysis Cyber Index Alliance brings together industry partners to set a new standard for evaluating how AI models perform on enterprise cyber defense tasks. The Alliance launches alongside the…
Gemini 3.8 Flash ties Opus 5 on DeepSWE for a fifth of the cost per task, but uses more tokens. Flash Cyber is gated; Muse Spark is cheap if Meta trains on you.
Strengthen your cybersecurity strategy with Andersen cyber risk management services. Identify, assess, and mitigate risks to improve resilience, compliance, and business security.
Google today revealed its next AI frontier model, which it's calling Gemini 4 Argon. The new model delivers "frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense," according to chief AI architect and Google DeepMind SVP Koray Kavukcuoglu. But the company is limiting access at first to a "set of trusted cyber defenders," and Kavukcuoglu says that Google is "actively engaged in the U.S. government's voluntary process for pre-release model access while we gradually expand access." Kavukcuoglu says Gemini 4 Argon is already powering Google's … Read the full story at The Verge.