Are we on fire, dude? When a Waymo ran over a firework
Explorar
Noticias de IA
1057 elementos — filtrados, clasificados y sin duplicados
Secret Claude tracker shocks users after Anthropic’s anti-surveillance stance
AI Poses Biggest Security Challenge of Decade, UK’s Cooper Warns
A sociotechnical threat model for AI-driven smart home devices
DualView: Preventing Indirect Prompt Injection in Personal AI Agents
Greek Politician Investigating Spyware Had Mobile Phone Hacked
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
DRL-CLBA: A Clean Label Backdoor Attack for Speech Classification via DDPG Reinforcement …
Breaking Safety at the Token Boundary: How BPE Tokenization Creates Exploitable Gaps in L…
How Amazon Bedrock catches AI-generated phishing
Threads' ubiquitous Mr Beast spam is part of a massive crypto scam network
Pmeta-TLA: Backdoor Attacks for Speech Classification Models via Meta-Learning with Timbr…
Beyond Gradient-Based Attacks: Adversarial Robustness and Explainability Stability in Cyb…
SoK: Attack and Defense Landscape of Mobile On-device AI Systems
Recovering Input Text from Hidden States: Study of Gradient-Based Inversion of Decoder-On…
Beyond the Prompt: Jailbreaking Function-Calling LLMs via Simulated Moderation Traces
Hey, That's My Model! Introducing Chain & Hash, An LLM Fingerprinting Technique
Apple's Hide My Email may not be hiding anything
Accelerating the quantum-safe timeline
You Can Now Sound the Alarm on AI Behaving Badly
Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Em…
Claude Helped a Hacker Find a Way to Issue Tickets to Almost Every US Music Festival
Detecting Audio Deepfakes on the Edge:Lightweight SSL-Based Detection in a Browser Plugin
Curvature-Guided Module Localization for Low-Rank Detoxification of Backdoored Large Lang…
From Similarity to Vulnerability: Key Collision Attack on LLM Semantic Caching
Containment Verification: AI Safety Guarantees Independent of Alignment
An AI-Based Solution for Secure Service Provisioning in IoT
CVE-TTP KG: Knowledge Graph Linking Software Vulnerabilities to Attack Behaviors
Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense
AI-Generated PowerShell Malware: An Experimental Framework and Dataset