Explorar

Noticias de IA

1359 elementos — filtrados, clasificados y sin duplicados

Import AI (Jack Clark) Security & Safety
Import AI 441: My agents are working. Are yours?
Hugging Face Blog Security & Safety
AprielGuard: A Guardrail for Safety and Adversarial Robustness in Modern LLM Systems
OpenAI News Security & Safety
Continuously hardening ChatGPT Atlas against prompt injection
OpenAI News Security & Safety
Strengthening cyber resilience as AI capabilities advance
OpenAI News Security & Safety
Mixpanel security incident: what OpenAI users need to know
OpenAI News Security & Safety
Strengthening our safety ecosystem with external testing
OpenAI News Security & Safety
Understanding prompt injections: a frontier security challenge
OpenAI News Security & Safety
Doppel’s AI defense system stops attacks before they spread
Google DeepMind Security & Safety
Strengthening our Frontier Safety Framework
Google DeepMind Security & Safety
Introducing CodeMender: an AI agent for code security
Hugging Face Blog Security & Safety
Hugging Face and VirusTotal collaborate to strengthen AI security
OpenAI News Security & Safety
Disrupting malicious uses of AI: October 2025
OpenAI News Security & Safety
PRC-linked abuse: Surveillance and influence activity
OpenAI News Security & Safety
Cyber Operation: Russian-speaking malware tooling
OpenAI News Security & Safety
Scam operations: Online fraud networks
OpenAI News Security & Safety
Cyber Operation: Korean-language malware support
OpenAI News Security & Safety
Operation “Stop News”: Recidivist influence activity
OpenAI News Security & Safety
Cyber Operation: Phishing and scripting support
OpenAI News Security & Safety
Operation “Nine–emdash Line”: Regional influence activity
OpenAI News Security & Safety
Launching Sora responsibly
Hugging Face Blog Security & Safety
Democratizing AI Safety with RiskRubric.ai
OpenAI News Security & Safety
OpenAI and Anthropic share findings from a joint safety evaluation
OpenAI News Security & Safety
Helping people when they need it most
OpenAI News Security & Safety
Preparing for future AI risks in biology
OpenAI News Security & Safety
Scaling security with responsible disclosure
OpenAI News Security & Safety
Disrupting malicious uses of AI: June 2025
OpenAI News Security & Safety
Operation “Wrong Number”: AI-assisted task scam
OpenAI News Security & Safety
Operation “Sneer Review”: China-origin influence activity
OpenAI News Security & Safety
Operation “Uncle Spam”: US polarization influence activity
OpenAI News Security & Safety
Operation “Helgoland Bite”: German-language influence activity