Explorar

Noticias de IA

1057 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Security & Safety
TrustErase: Auditable Instant Machine Unlearning with Passport-Embedded Representations
arXiv cs.AI Security & Safety
Security and Privacy Prompts in the Wild: What Users Ask LLMs and How LLMs Respond
arXiv cs.AI Security & Safety
Detecting and Mitigating DDoS Attacks with AI: A Survey
arXiv cs.AI Security & Safety
Breaking the Code: Security Assessment of AI Code Agents Through Systematic Jailbreaking …
arXiv cs.AI Security & Safety
Adversarial Attacks Leverage Interference Between Features in Superposition
arXiv cs.AI Security & Safety
BadScientist: Can a Research Agent Write Convincing but Unsound Papers that Fool LLM Revi…
arXiv cs.AI Security & Safety
Membership Inference Attacks against Large Audio Language Models
arXiv cs.AI Security & Safety
Like a Hammer, It Can Build, It Can Break: Large Language Model Uses, Perceptions, and Ad…
arXiv cs.AI Security & Safety
Learning Red Agent Policy from Observations for Neurosymbolic Autonomous Cyber Agents
arXiv cs.AI Security & Safety
Graph neural networks at war: integrating cybersecurity and drone intelligence in the Isr…
arXiv cs.AI Security & Safety
SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents
arXiv cs.AI Security & Safety
An AI Security Agent for Banking: Multi-Vector Fraud and AML Detection Across Retail and …
Mashable Security & Safety
This Copilot vulnerability could expose emails, 2FA codes, and other sensitive data
Google DeepMind Security & Safety
Securing the future of AI agents
Ars Technica AI Security & Safety
Critical Copilot vulnerability allowed hackers to seal 2FA code from users
AI News Security & Safety
AI Red Teaming Explained: What It Is and Why You Need It
arXiv cs.AI Security & Safety
Red-Teaming Agent Execution Contexts: Open-World Security Evaluation on OpenClaw
arXiv cs.AI Security & Safety
Are Neuro-Inspired Multi-Modal Vision-Language Models Resilient to Membership Inference P…
arXiv cs.AI Security & Safety
Can We Stop Malicious AI? KILLBENCH: A Benchmark for External AI Kill Switch Feasibility
arXiv cs.AI Security & Safety
A Survey on Agentic Security: Applications, Threats and Defenses
arXiv cs.AI Security & Safety
Discrete optimal transport is a strong audio adversarial attack
arXiv cs.AI Security & Safety
Automated jailbreak attack targeting multiple defense strategies
arXiv cs.AI Security & Safety
GAS-Leak-LLM: Genetic Algorithm-Based Suffix Optimization for Black-Box LLM Jailbreaking
arXiv cs.AI Security & Safety
Snyk VulnBench JS 1.0: Can LLMs Find the Same Bugs Twice?
arXiv cs.AI Security & Safety
InstantForget: Update-Free Backdoor Unlearning with Inference-Time Feature Reset
arXiv cs.AI Security & Safety
Defending against Adaptive Prompt Injection Attacks via Reasoning-enabled Task Alignment
arXiv cs.AI Security & Safety
FragFuse: Bypassing Access Control of Large Language Model Agents via Memory-Based Query …
arXiv cs.AI Security & Safety
SkillVetBench: LLM-as-Judge for Multi-Dimensional Security Risk Evaluation in Open-Source…
arXiv cs.AI Security & Safety
MASCOT-Android: A Curated Dataset and Automated Collection Pipeline for Android Malware S…
arXiv cs.AI Security & Safety
Dual-Granularity Orthogonal Disentanglement for Generalizable Audio Deepfake Detection