Steganalysis of Adaptive Covert Collusion in Tool-Using Agent Populations: A Black-Box, C…
Explorar
Noticias de IA
1057 elementos — filtrados, clasificados y sin duplicados
Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC…
MAFIA: Query-Only Memory Attacks via Probing and Factual Injection against Audited LLM Ag…
Privacy-Preserving AI Verification via Minimal Information Disclosure
Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic)
OK, Well, Rogue AI Agents Are Hacking Again
AI fuels more than half of cybercrime in Africa as scams surge – Interpol
OpenAI, Anthropic AI Models Involved in More Security Incidents
Nvidia doesn’t mess around: A week after open AI industry group formed, it’s already show…
Data Centers Exposed US Telecoms to China Hacks, US Panel to Say
Hackers Steal Bitcoin, Visa Outlines Stablecoin Plans | Bloomberg Crypto 8/4/2026
Third-party cyber evaluations involving OpenAI models
Apple caps bug bounty program due to deluge of AI submissions
Apple Asks Judge to Bar OpenAI From Using Alleged Trade Secrets
AI Now Fuels Over Half of Africa’s Cybercrime, Study Finds
Apple says more ex-employees may have taken confidential data to OpenAI
ICO: Enhancing Semantic-Shift Jailbreaks via Iterative Context Optimization
When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems
From Chasing Ghosts to Missed Attacks: Perspectives and Perceptions of SOC Practitioners …
No Single Neuron of Failure: Distributed Safety Alignment Against White-Box Attacks
Salami Attack: Stealthy Collusive Memory Poisoning against OpenClaw
Adversarial Attacks in Multi-Agent LLM Pipelines: Unveiling Structural Vulnerabilities in…
MineGrad: Gradient Inversion Attacks on LoRA Fine-Tuning
Decoy Images Amplify Caption-Mediated Defenses Against Encoded Jailbreaks
Why Formal Monitors Fail: Attack Distribution Entropy as a Coverage Bound for LTL-Based L…
Exposed by Design: A Dynamic Security Assessment of Internet-Facing MCP Servers at Scale
Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models
Securing Agentic AI: From Per-Action Checks to Trajectory Assurance
Robust Watermarks Meet Backdoored Models: Evading Diffusion Semantic Watermarks via Steal…
Caliber: Cross-Architecture Extraction-Cost Control for Score-Returning APIs