Meta Deletes Face-Recognition System From Its Smart Glasses App After WIRED Report
Explorar
Noticias de IA
1057 elementos — filtrados, clasificados y sin duplicados
Reddit ads pose as news stories to promote AI investment scams
Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabiliti…
Context-Fractured Decomposition Attacks on Tool-Using LLM Agents: Exploiting Artifact Pro…
It's a TRAP! Task-Redirecting Agent Persuasion Benchmark for Web Agents
MalTree: Tracing Malware Evolution from Embeddings at Scale
Latent-space Attacks for Refusal Evasion in Language Models
Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in L…
EVA: Evolving Semantic Adversaries for Red-Teaming GUI Agents Against Environmental Injec…
RAVEN: Retrieval-Augmented Vulnerability Exploration Network for Memory Corruption Analys…
Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software
Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety
OpenAI unveils Lockdown Mode to protect sensitive data from prompt injection attacks
Meta confirms 1000s of Instagram accounts were hacked by abusing its AI chatbot
Crypto-Funded Chinese Peptide Labs Are Booming
Bitcoin drops below $60,000, despite Trumps embrace
OpenAI rolls out a Lockdown Mode for extra protection against prompt injection attacks
Rubrik CEO's Big AI Warning
The Meta hack shows there’s more to AI security than Mythos
Explainable AI-Driven Cyber Risk Analytics and Model Reliability Assessment for Intellige…
AttackPathGNN: Cross-function vulnerability detection in smart contracts using state inte…
Safety Paradox: How Enhanced Safety Awareness Leaves LLMs Vulnerable to Posterior Attack
Domain-Conditioned Safety in Frontier Computer-Using Agents: A 793-Episode Browser Benchm…
Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage?
CaMeLs Can Use Computers Too: System-level Security for Computer Use Agents
Data Flow Control: Data Safety Policies for AI Agents
Cognitive Threat Intelligence and Explainable Federated Security Analytics for distribute…
TinyML-Driven Cybersecurity for Autonomous Spacecraft: Latency-Accuracy Analysis for SPAR…
IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures
SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks