Security and Privacy in Agentic AI: Grand Challenges and Future Directions
Explorar
Noticias de IA
1057 elementos — filtrados, clasificados y sin duplicados
When Agents Remember Too Much: Memory Poisoning Attacks on Large Language Model Agents
From Beats to Breaches:How Offensive AI Infers Sensitive User Information from Playlists
Large Language Models (LLMs) and Generative AI in Cybersecurity and Privacy: A Survey of …
Suspecting AI cheating, Ivy League prof ordered an in-person final; scores fell 50%
Google’s deepfake detector system used to debunk McConnell hoax pic
Lawsuit: Man used Grok to make 7K sex images of stepdaughter, then shot himself
Discord confirms AI moderators have banned thousands over harmless images
Hackers can use 9 of the most popular AI tools to assemble massive botnets
GitLost: We Tricked GitHub's AI Agent into Leaking Private Repos
The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access …
Evaluating calibrated refusal and safe usefulness in dual-use biology settings
Discord admits AI moderation bug wrongfully banned users over harmless images
AI Meets Cryptography 1: What AI Found in Cloudflare's Circl
ECB Asks Banks for Plans to Address AI Cybersecurity Threats
Taylor Swifts wedding photos are going viral. Theres just one problem.
Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies
Measuring the Robustness of Audio Deepfake Detection under Real-World Corruption
Don't Make Models Guess Security and Safety: Symbolic Guardrails for Domain-Specific AI A…
Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Mo…
When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Age…
Your Agent's Memories Are Not Its Own: Forged Reasoning Attacks on LLM Agent Memory and D…
Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
Beyond Static Rules: Automated Discovery of Latent Vulnerabilities in Text-to-SQL
DualView: Preventing Indirect Prompt Injection in Personal AI Agents
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Age…
Agent Data Injection Attacks are Realistic Threats to AI Agents
Agentic SABRE: An Uncertainty-Aware Neuro-Symbolic Multi-Agent Framework for Adaptive Ran…
The ‘first’ AI-run ransomware attack still needed a human