OpenAI and Anthropic models went on a hacking spree when tested by the UK's AI research i…
Browse
AI News
931 items — filtered, classified, deduplicated
OK, Well, Rogue AI Agents Are Hacking Again
AI fuels more than half of cybercrime in Africa as scams surge – Interpol
OpenAI, Anthropic AI Models Involved in More Security Incidents
Nvidia doesn’t mess around: A week after open AI industry group formed, it’s already show…
Data Centers Exposed US Telecoms to China Hacks, US Panel to Say
Hackers Steal Bitcoin, Visa Outlines Stablecoin Plans | Bloomberg Crypto 8/4/2026
Third-party cyber evaluations involving OpenAI models
Apple caps bug bounty program due to deluge of AI submissions
Apple Asks Judge to Bar OpenAI From Using Alleged Trade Secrets
AI Now Fuels Over Half of Africa’s Cybercrime, Study Finds
Apple says more ex-employees may have taken confidential data to OpenAI
VLAGuard: A Framework for Evaluating and Mitigating Physical Attention Hijacking in Visio…
Adversarial Attacks in Multi-Agent LLM Pipelines: Unveiling Structural Vulnerabilities in…
No Single Neuron of Failure: Distributed Safety Alignment Against White-Box Attacks
From Chasing Ghosts to Missed Attacks: Perspectives and Perceptions of SOC Practitioners …
FL-OA: A Byzantine-Robust Federated Learning Framework with Outsourced Auditing for Intel…
Crushing the Evidence: A Dual-Penalty Evasion Framework for Fooling White-Box Explainable…
Exposed by Design: A Dynamic Security Assessment of Internet-Facing MCP Servers at Scale
Salami Attack: Stealthy Collusive Memory Poisoning against OpenClaw
MineGrad: Gradient Inversion Attacks on LoRA Fine-Tuning
Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based L…
When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems
Securing Agentic AI: From Per-Action Checks to Trajectory Assurance
Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models
Decoy Images Amplify Caption-Mediated Defenses Against Encoded Jailbreaks
Robust Watermarks Meet Backdoored Models: Evading Diffusion Semantic Watermarks via Steal…
Caliber: Cross-Architecture Extraction-Cost Control for Score-Returning APIs
AI Security Leaderboard: Methodology, Results and Minimal Standard
OpenAI Hack Could Have Been 'Way Worse,' Hugging Face CEO Says