Explorar

Noticias de IA

1057 elementos — filtrados, clasificados y sin duplicados

Wired AI Security & Safety
Meta Deletes Face-Recognition System From Its Smart Glasses App After WIRED Report
Mashable Security & Safety
Reddit ads pose as news stories to promote AI investment scams
Hugging Face Daily Papers Security & Safety
Unveiling Privacy Risks in Multi-modal Large Language Models: Task-specific Vulnerabiliti…
Hugging Face Daily Papers Security & Safety
Context-Fractured Decomposition Attacks on Tool-Using LLM Agents: Exploiting Artifact Pro…
arXiv cs.AI Security & Safety
It's a TRAP! Task-Redirecting Agent Persuasion Benchmark for Web Agents
arXiv cs.AI Security & Safety
MalTree: Tracing Malware Evolution from Embeddings at Scale
arXiv cs.AI Security & Safety
Latent-space Attacks for Refusal Evasion in Language Models
arXiv cs.AI Security & Safety
Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in L…
arXiv cs.AI Security & Safety
EVA: Evolving Semantic Adversaries for Red-Teaming GUI Agents Against Environmental Injec…
arXiv cs.AI Security & Safety
RAVEN: Retrieval-Augmented Vulnerability Exploration Network for Memory Corruption Analys…
arXiv cs.AI Security & Safety
Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software
arXiv cs.AI Security & Safety
Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety
TechCrunch AI Security & Safety
OpenAI unveils Lockdown Mode to protect sensitive data from prompt injection attacks
Hacker News (AI filter) Security & Safety
Meta confirms 1000s of Instagram accounts were hacked by abusing its AI chatbot
Wired AI Security & Safety
Crypto-Funded Chinese Peptide Labs Are Booming
Mashable Security & Safety
Bitcoin drops below $60,000, despite Trumps embrace
Engadget Security & Safety
OpenAI rolls out a Lockdown Mode for extra protection against prompt injection attacks
Bloomberg Technology Security & Safety
Rubrik CEO's Big AI Warning
MIT Technology Review AI Security & Safety
The Meta hack shows there’s more to AI security than Mythos
arXiv cs.AI Security & Safety
Explainable AI-Driven Cyber Risk Analytics and Model Reliability Assessment for Intellige…
arXiv cs.AI Security & Safety
AttackPathGNN: Cross-function vulnerability detection in smart contracts using state inte…
arXiv cs.AI Security & Safety
Safety Paradox: How Enhanced Safety Awareness Leaves LLMs Vulnerable to Posterior Attack
arXiv cs.AI Security & Safety
Domain-Conditioned Safety in Frontier Computer-Using Agents: A 793-Episode Browser Benchm…
arXiv cs.AI Security & Safety
Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage?
arXiv cs.AI Security & Safety
CaMeLs Can Use Computers Too: System-level Security for Computer Use Agents
arXiv cs.AI Security & Safety
Data Flow Control: Data Safety Policies for AI Agents
arXiv cs.AI Security & Safety
Cognitive Threat Intelligence and Explainable Federated Security Analytics for distribute…
arXiv cs.AI Security & Safety
TinyML-Driven Cybersecurity for Autonomous Spacecraft: Latency-Accuracy Analysis for SPAR…
arXiv cs.AI Security & Safety
IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures
arXiv cs.AI Security & Safety
SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks