not much happened today
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
PlanFlip: Attacking Multi-Agent LLM Systems via Planning-Phase Prompt Injection
STAC: When Innocent Tools Form Dangerous Chains for LLM Agents
Adaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent Security
How Do You Choose Your AI Component? An Interview Study of Secure AI Integration in Pract…
Do Speech Tokens Leak Voiceprints? Speaker Inversion Attacks Against End-to-End Speech La…
Between Safe Boundaries: Exploiting Temporal Consistency for Jailbreaking Text-To-Video G…
Broken Gates: Re-evaluating Web Bot Defenses in the Age of LLM Agents
Hardware Mechanisms to Dynamically Throttle AI Performance
An Early Warning of Emerging Biosecurity Risks in Frontier LLMs
Self-State Attacks on Self-Hosted AI Agents: How Far Can OS Defenses Go?
Latent Fusion Jailbreak: Blending Harmful and Harmless Representations to Elicit Unsafe L…
Poison to Detect: Detection of Targeted Overfitting in Federated Learning
FLINT: Fingerprinting Federated Learning Architectures from 5G PHY-Layer Side Channels
Your Period Tracker Is (Probably) Spying on You
Prompt Injection Attacks Are Thwarting AI Hacking Agents
The Zoom hack that says, ‘Don’t record me’
IPhone Hacking Firm Sues Ex-Worker Over Alleged Theft of Secrets
Zoox issues software recall because its robotaxis may be confused by smoke
AI Meets Cryptography 2: What AI Found in OpenVM's ZkVM
Most smart appliances collect data: Here's how to find out what yours tracks
The risk of weather data sabotage is rising
xAI can’t deny Grok makes CSAM anymore. So it’s suing users.
The agent security gap: 54% of enterprises have already had an AI agent incident, and mos…
NTSB investigators confirm Tesla driver overrode Full Self-Driving system in fatal crash
xAI sues Grok user for generating nonconsensual sexualized deepfakes
Adversarial Prompting Framework for AI Safety Assessment
Safeguard-Conditioned Uplift: Measuring Utility-Risk Frontiers for Dual-Use Biology Assis…
Rethinking Penetration Testing for AI-Enabled Systems: From Resource Compromise to Behavi…
PersGuard: Preventing Malicious Personalization in Text-to-Image Diffusion Models via Mod…