Learning Intrusion Response Strategies for OT Systems
Explorar
Noticias de IA
1359 elementos — filtrados, clasificados y sin duplicados
Arbitrary Cipher Attacks Against Large Language Models Do Not Require Fine-Tuning
How Fragile Is Safety Alignment at Frontier Scale? A Single-Direction Attack on a 320B MoE
An Experimental Evaluation of Multimodal Prompt Injection Attacks on Agentic AI Frameworks
AgentHijack: Visual Patch Attacks on Multimodal Computer-Use Agents
Cyber-Financial Contagion: Modeling the Propagation of an AI Vendor Compromise Through th…
CS-Guard: Benchmarking LLM Guardrails for Code Generation Security
Trust Me, I'm Your Developer: Self-Issued Authentication in Large Language Models
Black-Box Red Teaming of Agentic AI: A Taxonomy-Driven Framework for Automated Risk Disco…
Six Chinese AI firms accused of aggressively copying US frontier models
I Let an AI Agent Hack All My Gadgets—and I’d Do It Again
Anthropic Worker Resigns, Warns of AI Risks to Humanity
Ledger Hires New Security Chief as Crypto Hacks Hit $1.4 Billion
Man told ChatGPT he was feeling delusional. ChatGPT insisted he was Jesus.
Worried Anthropic researchers warn that AI ‘could kill all humans’
Anthropic researcher quits, says AI could kill us all by the end of the decade
not much happened today
Versioned Transitive Dependency-Closure Binding and Operation-Time Effect Governance for …
Evaluating Deep-Search Agents under Hierarchical Web Evidence Poisoning
Style Over Substance: Content-Invariant Wrappers Flip LLM Safety-Judge Verdicts
Bait-and-Recover: Poisoning Internal Refusal Signals to Defend LLMs against White-Box Edi…
SWE-Test: Benchmarking LLM Vulnerability Discovery via Input Prediction
Detokenization Leaks: Reconstructing Local LLM Outputs From Cache Traces
Hardware Trojan Threats to Multi-Chiplet Photonic Neural Network Accelerators
From Review to Authorization: Key-Isolated Threshold Signing for LLM Agents
PiMRef: Deducing Ever-evolving Spear-phishing Emails with Knowledge Base Invariants
Towards a Resilience-Theoretic Foundation for Adversarial Robustness in Industrial Contro…
How to Backdoor Image Knowledge Distillation
FATS: A Prompt Injection Attack Utilizing Feign Security Agents with Deceptive Few-shots …
Revoked but Still Authoritative: An Empirical Study of Revocation Enforcement in Agent-Me…