Corrupt Plans, Clean Traces: Evading Chain-of-Thought Monitoring with Plan Injection
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
Automating Attack Graph Construction for Agentic Pentesting. Towards Neuro-Symbolic Vulne…
HazardAuditor: From Executable Threats to Safer Computer-Use Agents
Measuring and Exploiting Contextual Bias in LLM-Assisted Security Code Review
An AI Agent Execution Environment to Safeguard User Data
ActGuard: Pre-execution Action Auditing against Indirect Prompt Injection in LLM Agents
SynGhost: Invisible and Universal Task-agnostic Backdoor Attack via Syntactic Transfer
Data Security in Large Language Models: Risks, Defense, and Directions
Trustworthy Agentic AI: A Comprehensive Cybersecurity and Systems Survey on Threat Landsc…
When Malicious Instructions Persist: Persistent Memory Poisoning Attack on Harness-Based …
Vulnerability Localization Benchmark: Measuring Agentic Security Analysis at Repository S…
DSS: Dynamic Semantic Steering for Robust Concept Erasure in Diffusion Models
Exploring Automated Vulnerability Identification in JavaScript Code Using Large Language …
Divide, Consult, Conquer: Capability Laundering Through Aligned LLMs
Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges
IntraGuard: Committee-Side Defenses Against Review Outsourcing to Commercial Chatbots
PIDS-Bench: Evaluating Prompt-Injection Detectors Under Over-Defense, Obfuscation, and Di…
SENTINEL: A Multi-Pathway Architecture for Detecting Living-Off-the-Land APT Attacks on W…
When the World Lies: Backdoor Attacks on Latent World Models for Downstream Control
Hugging Face Scientist on Safety Issues With Agentic AI
Google DeepMind Staffer Says AI May ‘Kill Us All’ in Exit Post
Gebru: AI Security & Safety Is About Human Control
Cloudflare CEO: Good Guys Have More Tools Than Bad Guys
AI bots "Timmy," "Ren," and "Jackie" are flooding social media with slop
New York Seizes a Dozen Celebrity Deepfake Websites
Ex-Google DeepMind Insider: Why We MUST Slow Down AI Now
Adversarial Fashion Makes a Statement on AI Panopticon
What a time to be alive – rouge AI agents attack RubyGems.org
Sexually Explicit Deepfake Sites Target 100-Plus Politicians in Europe
OpenAI President on Doing Business in the Wake of Hugging Face