Explorar

Noticias de IA

1359 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Security & Safety
Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence
arXiv cs.AI Security & Safety
Jailbreak to Protect: Buffering and Reinforcing via Temporary Jailbreaking for Safe Fine-…
arXiv cs.AI Security & Safety
Hidden-State Privacy Has an Empty Middle
arXiv cs.AI Security & Safety
LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injection
arXiv cs.AI Security & Safety
Membership Inference Attacks on Tokenizers of Large Language Models
arXiv cs.AI Security & Safety
Explainable Attention-Guided Stacked Graph Neural Networks for Malware Detection
arXiv cs.AI Security & Safety
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models
arXiv cs.AI Security & Safety
Enhancing Reliability in LLM-Based Secure Code Generation
arXiv cs.AI Security & Safety
Attested Tool-Server Admission: A Security Extension to the Model Context Protocol
arXiv cs.AI Security & Safety
When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Age…
arXiv cs.AI Security & Safety
Security of OpenClaw Agents: Fundamentals, Attacks, and Countermeasures
arXiv cs.AI Security & Safety
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
Mashable Security & Safety
Megalodon cyberattack infects 5,500 GitHub open-source repositories with malware, researc…
Wired AI Security & Safety
The AI Era Is Creating a Bug Hunting Arms Race
arXiv cs.AI Security & Safety
Codec-Robust Attacks on Audio LLMs
arXiv cs.AI Security & Safety
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
arXiv cs.AI Security & Safety
BarrierSteer: LLM Safety via Learning Barrier Steering
arXiv cs.AI Security & Safety
AI Security Research Should Better Incentivize Defense Research
arXiv cs.AI Security & Safety
PoisonForge: Task-Level Targeted Poisoning Benchmark for Instruction-Tuned LLMs
arXiv cs.AI Security & Safety
Security of LLM-generated Code: A Comparative Analysis
arXiv cs.AI Security & Safety
MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structura…
arXiv cs.AI Security & Safety
RAG-Pull: Turning Retrieval into a Code-Injection Channel via Invisible Unicode Perturbat…
arXiv cs.AI Security & Safety
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
arXiv cs.AI Security & Safety
The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Syst…
arXiv cs.AI Security & Safety
Are Frontier LLMs Ready for Cybersecurity? Evidence for Vertical Foundation Models from D…
arXiv cs.AI Security & Safety
Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Feat…
arXiv cs.AI Security & Safety
TCAP: Tri-Component Attention Profiling for Unsupervised Backdoor Detection in MLLM Fine-…
arXiv cs.AI Security & Safety
Adversarial Vulnerability Under Temporal Concept Drift: A Longitudinal Study of Android M…
arXiv cs.AI Security & Safety
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
arXiv cs.AI Security & Safety
GenAI-Driven Threat Detection with Microsoft Security Copilot