HauntAttack: When Attack Follows Reasoning as a Shadow
Explorar
Noticias de IA
1057 elementos — filtrados, clasificados y sin duplicados
Prompt Injection in Automated R\'esum\'e Screening with Large Language Models: Single and…
From Celebrities to Anyone: Characterizing AI Nudification Content, Technology, and Commu…
Fortress and Gatekeeper: Theorizing Transitive Trust in Third-Party Cybersecurity Risk Go…
The Inattentional Gap: Task-Conditioned Language and Vision Models Omit the Safety-Critic…
Chai: Agentic Discovery of Cryptographic Misuse Vulnerabilities
Inherited Circuits, Learned Semantics: How Fine-Tuning Creates Evasion Vulnerabilities In…
What happened after 2k people tried to hack my AI assistant
Anthropic says Alibaba must be punished for largest Claude cloning attack
Russia allegedly used a forensics platform to hack an activist's phone, despite having it…
A Marketplace for AI-Generated Adult Content and Deepfakes
What Does It Mean to Break a Distillation Defense?
The Unfireable Safety Kernel: Execution-Time AI Alignment for AI Agents and Other Escapab…
SoK: AI Secure Code Generation: Progress, Pitfalls, and Paths Forward
Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelin…
Epistemic Bias Injection: Manipulating LLM Opinion via Selective Context Retrieval
Privacy Vulnerabilities of Attention Layers in Tabular Foundation Models and Protection o…
Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study
What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics
PVF:Understanding AI Vulnerability Against SDCs
A Hybrid CNN-LSTM Intrusion Detection Framework for Cybersecurity in Smart Renewable Ener…
Color Matters: Trigger Color Affects Success in Federated Backdoor Attacks
Scaling cybercrime disruption through innovation and AI
Anthropic says Alibaba illicitly extracted Claude AI model capabilities
Anthropic Accuses Alibaba of ‘Illicitly’ Accessing AI Models
AI Snitches Get Glitches: Towards Evading Agentic Surveillance
Microsoft Says Copilot AI Helped Knock Down Cybercrime Tools
Fight Against Child Predators Gets Short Shrift While Cases Explode in AI Era
Scammers Are Using AI to Create Fake Auto Loan Documents
VPA-Guard: Defending and Benchmarking Image-to-Video Generation Against Visual Prompt Att…