CEAR: Certified Ensemble Adversarial Robustness in DNNs
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
Benchmarking Security Risk Detection and Verification in Open Agentic Skill Ecosystems
Digital-to-Physical Transfer of Adversarial Patches for Aerial Vehicle Detection
PrivacyPeek: Auditing What LLM-Based Agents Acquire, Not Just What They Say
Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Mode…
Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research
Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs
Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem
"Do Not Mention This to the User": Detecting and Understanding Malicious Agent Skills
The Alignment Curse: Modality Alignment Supercharges Audio Attacks via Text Transfer
SilentDrift: Exploiting Action Chunking for Stealthy Backdoor Attacks on Vision-Language-…
Ethical Hyper-Velocity (EHV): A Hardware-Rooted Zero-Trust Runtime Enforcement Architectu…
Safety Must Precede the Deployment of Open-Ended AI
SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents
Cross-modal linkage risk in clinical vision-language models
AgentRedBench: Dynamic Redteaming and Integration-Aware Defense for LLM Agents over SaaS …
Jailbreaking Multimodal Large Language Models using Multi-Clip Video
THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Languag…
Defenses & Enablers For Skill Injection Attacks on Terminal Based Agents
DiscourseFlip: An Oblique Discourse-Level Opinion Manipulation Attack against Black-box R…
ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree
Hackers duped Meta AI support chatbot to steal celebrity Instagram accounts
Meta's AI support chatbot made it ridiculously easy for hackers to take over Instagram ac…
Meta’s own AI was exploited to hijack Instagram accounts
Hackers say that Meta AI helped them compromise big Instagram accounts
Allegedly trashing Airbnbs to test robots puts startup in legal trouble
Monitoring Agentic Systems Before They're Reliable
SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems
When LLMs Learn to Be Consistently Wrong: A Multi-Model Study of Linear Representations o…
The Surface You Test Is Not the Surface That Breaks