PiMRef: Deducing Ever-evolving Spear-phishing Emails with Knowledge Base Invariants
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
Towards a Resilience-Theoretic Foundation for Adversarial Robustness in Industrial Contro…
Unveiling Hidden Threats: Using Fractal Triggers to Boost Stealthiness of Distributed Bac…
Bait-and-Recover: Poisoning Internal Refusal Signals to Defend LLMs against White-Box Edi…
Multimodal Resource-Exhaustion Attacks on Vision-Language Models via Joint Pixel-Prompt O…
Style Over Substance: Content-Invariant Wrappers Flip LLM Safety-Judge Verdicts
Revoked but Still Authoritative: An Empirical Study of Revocation Enforcement in Agent-Me…
Versioned Transitive Dependency-Closure Binding and Operation-Time Effect Governance for …
A TTP by TTP Approach: Precise Malware Detection via Malicious TTP Recognition
How to Backdoor Image Knowledge Distillation
Cisco President on Defending Against AI Attacks
Why this month's Microsoft patch release is a doozy
Hackers are stealing Claude tokens from subscribers
“This is the AI men actually use”: Meta ads pushed apps nudifying real teens
Chrome is now shipping updates every 2 weeks as AI changes the security landscape
Meta Ran Hundreds of Ads Showing AI Child Sexual Abuse, NGO Says
Boston Scientific Warns of Financial Hit from Cyber Attack
OpenAI is figuring out how to tell people when its agents go rogue
AlcaTRAz - Anchored Tree-Rule Defense Against Jailbreaks
CONTINUITY: Security-Context Contracts for Composable LLM Agent Controls
Repeat-After-Me: Black-Box Adaptive Visual Prompt Injection
When Seeing Overrides Knowing: Visual Dominance and Deferral-Based Method for Personalize…
Uncensored Open-weight Models: Redistribution as the Persistence Layer
Rethinking Indirect Prompt Injection as a Test-Time Search Problem
The Struggle Between Continuation and Refusal: A Mechanistic Analysis of the Continuation…
OpenAI responds after report exposed another incident in which its AI agents went rogue
Rogue AI agents commandeered German website and used it as a messaging board
OpenAI admits to German wiki ‘incident’
OpenAI Agents Hacked Another Website
[AINews] Collusion.wiki: A second undisclosed OpenAI agent swarm incident...