When Poison Fails After Retrieval: Revisiting Corpus Poisoning under Chunking and Reranki…
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization
Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models
Goal-Autopilot: A Verifiable Anti-Fabrication Firewall for Unattended Long-Horizon Agents
Learning to Inject: Automated Prompt Injection via Reinforcement Learning
South Korea Fines Coupang Record $409 Million for Data Leak
AI agent runs amok in Fedora and elsewhere
OpenAI Says China-Linked Accounts Aim to Fuel US Data Center Pushback
Cybersecurity researchers aren’t happy about the guardrails on Anthropic’s Fable
Anthropic CEO Doesn’t Know If Claude Used in Iran School Strike
EC2’s formally verified “isolation engine” provides mathematical assura…
Wrongful Arrest Exposes Failures in One of the Oldest Police Face-Recognition Tools in th…
A €0.01 bank transfer could compromise a banking AI agent
PRC-linked influence operations are targeting AI debates in the US
Dynamics of Adversarial Attacks on Large Language Model-Based Search Engines
Understanding and mitigating the risks of OpenClaw for non-technical users: A practical g…
Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation
BadRobot: Jailbreaking Embodied LLM Agents in the Physical World
The Distributed Detectability Band Against Marginal-Preserving Attacks
GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines
Attacks on Machine-Text Detectors Retain Stylistic Fingerprints
A Source Domain is All You Need: Source-Only Cross-OS Transfer Learning for APT Anomaly D…
Assessing Automated Prompt Injection Attacks in Agentic Environments
Advancing the State-of-the-Art in Empirical Privacy Auditing
A Bayesian Network Approach for Enhancing Security-Focused Decision Support Systems
'Sloppenheimer:' Amazon Employees Mock the Company's AI on Slack
The French government's internal messaging service was compromised in a security breach
Conan OBrien, deepfake master, wants to stop you from getting pwned
Autonomous AI Data Loss in DevOps: Building Efficient Defenses
Microsoft's open source tools were hacked to steal passwords of AI developers