GPT-5.5 Bio Bug Bounty
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
Large Language Models (LLMs) and Generative AI in Cybersecurity and Privacy: A Survey of …
When Agents Remember Too Much: Memory Poisoning Attacks on Large Language Model Agents
From Beats to Breaches:How Offensive AI Infers Sensitive User Information from Playlists
Security and Privacy in Agentic AI: Grand Challenges and Future Directions
Suspecting AI cheating, Ivy League prof ordered an in-person final; scores fell 50%
Google’s deepfake detector system used to debunk McConnell hoax pic
Lawsuit: Man used Grok to make 7K sex images of stepdaughter, then shot himself
Discord confirms AI moderators have banned thousands over harmless images
Hackers can use 9 of the most popular AI tools to assemble massive botnets
GitLost: We Tricked GitHub's AI Agent into Leaking Private Repos
The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access …
Evaluating calibrated refusal and safe usefulness in dual-use biology settings
Discord admits AI moderation bug wrongfully banned users over harmless images
AI Meets Cryptography 1: What AI Found in Cloudflare's Circl
ECB Asks Banks for Plans to Address AI Cybersecurity Threats
Taylor Swifts wedding photos are going viral. Theres just one problem.
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Age…
Beyond Static Rules: Automated Discovery of Latent Vulnerabilities in Text-to-SQL
Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies
Agent Data Injection Attacks are Realistic Threats to AI Agents
Undetectable Backdoors in Model Parameters: Hiding Sparse Secrets in High Dimensions
Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives
Don't Make Models Guess Security and Safety: Symbolic Guardrails for Domain-Specific AI A…
Measuring the Robustness of Audio Deepfake Detection under Real-World Corruption
Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Mo…
When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Age…
Your Agent's Memories Are Not Its Own: Forged Reasoning Attacks on LLM Agent Memory and D…
DualView: Preventing Indirect Prompt Injection in Personal AI Agents
Agentic SABRE: An Uncertainty-Aware Neuro-Symbolic Multi-Agent Framework for Adaptive Ran…