An AI Agent Execution Environment to Safeguard User Data
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
IntraGuard: Committee-Side Defenses Against Review Outsourcing to Commercial Chatbots
Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges
Corrupt Plans, Clean Traces: Evading Chain-of-Thought Monitoring with Plan Injection
When the World Lies: Backdoor Attacks on Latent World Models for Downstream Control
Automating Attack Graph Construction for Agentic Pentesting. Towards Neuro-Symbolic Vulne…
Data Security in Large Language Models: Risks, Defense, and Directions
When Malicious Instructions Persist: Persistent Memory Poisoning Attack on Harness-Based …
SynGhost: Invisible and Universal Task-agnostic Backdoor Attack via Syntactic Transfer
HazardAuditor: From Executable Threats to Safer Computer-Use Agents
SENTINEL: A Multi-Pathway Architecture for Detecting Living-Off-the-Land APT Attacks on W…
Exploring Automated Vulnerability Identification in JavaScript Code Using Large Language …
Vulnerability Localization Benchmark: Measuring Agentic Security Analysis at Repository S…
Divide, Consult, Conquer: Capability Laundering Through Aligned LLMs
ActGuard: Pre-execution Action Auditing against Indirect Prompt Injection in LLM Agents
DSS: Dynamic Semantic Steering for Robust Concept Erasure in Diffusion Models
Trustworthy Agentic AI: A Comprehensive Cybersecurity and Systems Survey on Threat Landsc…
Measuring and Exploiting Contextual Bias in LLM-Assisted Security Code Review
PIDS-Bench: Evaluating Prompt-Injection Detectors Under Over-Defense, Obfuscation, and Di…
Hugging Face Scientist on Safety Issues With Agentic AI
Google DeepMind Staffer Says AI May ‘Kill Us All’ in Exit Post
Gebru: AI Security & Safety Is About Human Control
Cloudflare CEO: Good Guys Have More Tools Than Bad Guys
AI bots "Timmy," "Ren," and "Jackie" are flooding social media with slop
New York Seizes a Dozen Celebrity Deepfake Websites
Ex-Google DeepMind Insider: Why We MUST Slow Down AI Now
Adversarial Fashion Makes a Statement on AI Panopticon
What a time to be alive – rouge AI agents attack RubyGems.org
Sexually Explicit Deepfake Sites Target 100-Plus Politicians in Europe
OpenAI President on Doing Business in the Wake of Hugging Face