Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
Jailbreak to Protect: Buffering and Reinforcing via Temporary Jailbreaking for Safe Fine-…
Hidden-State Privacy Has an Empty Middle
LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injection
Membership Inference Attacks on Tokenizers of Large Language Models
Explainable Attention-Guided Stacked Graph Neural Networks for Malware Detection
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models
Enhancing Reliability in LLM-Based Secure Code Generation
Attested Tool-Server Admission: A Security Extension to the Model Context Protocol
When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Age…
Security of OpenClaw Agents: Fundamentals, Attacks, and Countermeasures
Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses
Megalodon cyberattack infects 5,500 GitHub open-source repositories with malware, researc…
The AI Era Is Creating a Bug Hunting Arms Race
Codec-Robust Attacks on Audio LLMs
Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG
BarrierSteer: LLM Safety via Learning Barrier Steering
AI Security Research Should Better Incentivize Defense Research
PoisonForge: Task-Level Targeted Poisoning Benchmark for Instruction-Tuned LLMs
Security of LLM-generated Code: A Comparative Analysis
MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structura…
RAG-Pull: Turning Retrieval into a Code-Injection Channel via Invisible Unicode Perturbat…
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Syst…
Are Frontier LLMs Ready for Cybersecurity? Evidence for Vertical Foundation Models from D…
Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Feat…
TCAP: Tri-Component Attention Profiling for Unsupervised Backdoor Detection in MLLM Fine-…
Adversarial Vulnerability Under Temporal Concept Drift: A Longitudinal Study of Android M…
GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents
GenAI-Driven Threat Detection with Microsoft Security Copilot