Explorar

Noticias de IA

1359 elementos — filtrados, clasificados y sin duplicados

Bloomberg Technology Security & Safety
OpenAI Makes AI Safety Changes in Wake of Hugging Face Breach
Bloomberg Technology Security & Safety
French Tax Office to Use AI to Probe Vulnerabilities After Hack
AI News Security & Safety
OpenAI president urges enterprises to hasten AI security defences
Ars Technica AI Security & Safety
Microsoft Copilot reveals secret input that allowed it to be hacked
OpenAI News Security & Safety
Pacing model development in an era of cyber-critical capabilities
AINews / smol.ai Security & Safety
not much happened today
Hugging Face Daily Papers Security & Safety
Fair ASR: Re-Evaluating Black-Box Jailbreaks under Shared Target-Call Budgets
arXiv cs.AI Security & Safety
Ventor-QTest: Threat-Model-Driven Verification of Vendor-Hosted LLM APIs
arXiv cs.AI Security & Safety
Propaganda Forensics: Recovering the Generation Pipeline of an AI-Driven Influence Campai…
arXiv cs.AI Security & Safety
SMA: Who Said That? Auditing Membership Leakage in Semi-Black-box RAG Controlling
arXiv cs.AI Security & Safety
Towards Risk-free AI Agent Deployment
arXiv cs.AI Security & Safety
Steering the Flow: Inverting Face Recognition Models via Gradient-Guided Flow Matching
arXiv cs.AI Security & Safety
When State Becomes an Attack Surface: State-Semantic Injection in LLM-Driven Embodied Age…
arXiv cs.AI Security & Safety
Workspace Topology as an Attack Vector in Agentic Coding Assistants
arXiv cs.AI Security & Safety
Individual Disempowerment through an Advice Channel: Control Loss when Influence is Endog…
arXiv cs.AI Security & Safety
Synchronized Logit Steering: Real-world Steganography
Hacker News (AI filter) Security & Safety
Israel creates fake think tank in likely attempt to dupe AI chatbots
Hugging Face Daily Papers Security & Safety
Digital Twin-Based Intrusion Detection for Vehicle Powertrain CAN Bus Systems
Mashable Security & Safety
Grok CSAM lawsuit expands as more step forward
Hugging Face Daily Papers Security & Safety
Steering the Flow: Inverting Face Recognition Models via Gradient-Guided Flow Matching
Bloomberg Technology Security & Safety
Odd Lots: Is There An AI Kill Switch If Things Go Wrong?
Hacker News (AI filter) Security & Safety
AI-Generated GitHub Copilot "Autofix" Allowed Compromise of Snowflake's Jira
Hugging Face Daily Papers Security & Safety
DSPrompt: Dynamic Soft Prompt Defense Against M-RAG Corruption
Bloomberg Technology Security & Safety
What the OpenAI/Hugging Face Hack Really Tells Us About AI Danger
Engadget Security & Safety
Another woman joins lawsuit accusing Grok of generating CSAM
OpenAI News Security & Safety
The Defender’s Window
arXiv cs.AI Security & Safety
Tripwire: Triggering Aligned Refusal via Statistically Certified Safety Neurons
arXiv cs.AI Security & Safety
Mandato: Protocol-Level Enforcement of Digitally Signed Mandates on AI Agent Actions with…
TechCrunch AI Security & Safety
Woman claims her stepfather used Grok to transform childhood photo into explicit imagery
TechCrunch AI Security & Safety
How to tell if your AI platforms’ accounts have been hacked