Explorar

Noticias de IA

1057 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Security & Safety
Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirica…
arXiv cs.AI Security & Safety
Runtime Skill Audit: Targeted Runtime Probing for Agent Skill Security
arXiv cs.AI Security & Safety
When Poison Fails After Retrieval: Revisiting Corpus Poisoning under Chunking and Reranki…
arXiv cs.AI Security & Safety
JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization
Bloomberg Technology Security & Safety
South Korea Fines Coupang Record $409 Million for Data Leak
Hacker News (AI filter) Security & Safety
AI agent runs amok in Fedora and elsewhere
Bloomberg Technology Security & Safety
OpenAI Says China-Linked Accounts Aim to Fuel US Data Center Pushback
TechCrunch AI Security & Safety
Cybersecurity researchers aren’t happy about the guardrails on Anthropic’s Fable
Bloomberg Technology Security & Safety
Anthropic CEO Doesn’t Know If Claude Used in Iran School Strike
Wired AI Security & Safety
Wrongful Arrest Exposes Failures in One of the Oldest Police Face-Recognition Tools in th…
Hacker News (AI filter) Security & Safety
A €0.01 bank transfer could compromise a banking AI agent
OpenAI News Security & Safety
PRC-linked influence operations are targeting AI debates in the US
arXiv cs.AI Security & Safety
Advancing the State-of-the-Art in Empirical Privacy Auditing
arXiv cs.AI Security & Safety
Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation
arXiv cs.AI Security & Safety
GitInject: Real-World Prompt Injection Attacks in AI-Powered CI/CD Pipelines
arXiv cs.AI Security & Safety
A Bayesian Network Approach for Enhancing Security-Focused Decision Support Systems
arXiv cs.AI Security & Safety
BadRobot: Jailbreaking Embodied LLM Agents in the Physical World
arXiv cs.AI Security & Safety
A Source Domain is All You Need: Source-Only Cross-OS Transfer Learning for APT Anomaly D…
arXiv cs.AI Security & Safety
Attacks on Machine-Text Detectors Retain Stylistic Fingerprints
arXiv cs.AI Security & Safety
Understanding and mitigating the risks of OpenClaw for non-technical users: A practical g…
arXiv cs.AI Security & Safety
Dynamics of Adversarial Attacks on Large Language Model-Based Search Engines
arXiv cs.AI Security & Safety
Assessing Automated Prompt Injection Attacks in Agentic Environments
arXiv cs.AI Security & Safety
The Distributed Detectability Band Against Marginal-Preserving Attacks
Hacker News (AI filter) Security & Safety
'Sloppenheimer:' Amazon Employees Mock the Company's AI on Slack
Engadget Security & Safety
The French government's internal messaging service was compromised in a security breach
Mashable Security & Safety
Conan OBrien, deepfake master, wants to stop you from getting pwned
AI News Security & Safety
Autonomous AI Data Loss in DevOps: Building Efficient Defenses
Hacker News (AI filter) Security & Safety
Microsoft's open source tools were hacked to steal passwords of AI developers
arXiv cs.AI Security & Safety
Instrumental convergence and power-seeking
arXiv cs.AI Security & Safety
RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-di…