CrackedPDFs: A Controlled Benchmark for Hidden Prompt Injection in PDFs
Explorar
Noticias de IA
30711 elementos — filtrados, clasificados y sin duplicados
ITPEval: Benchmarking Formal Translation Across Interactive Theorem Provers
The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception …
Knowledge-Centric Self-Improvement
An Intelligent-Cloud Edge Multimodal Interaction System for Robots
OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System At…
NEXUS: Structured Runtime Safety for Tool-Using LLM Agents
Measuring LLM Trust Allocation Across Conflicting Software Artifacts
Internal Knowledge Without External Expression: Probing the Generalization Boundary of a …
Information Aggregation with AI Agents
LoRA-Tuned Large Language Models for Dementia Detection via Multi-View Speech-Derived Fea…
FineServe: A Fine-Grained Dataset and Characterization of Global LLM Serving Workloads
Toward Adaptable Multi-Agent Reinforcement Learning: An Assumption-Aware Review
Prompt Programming for Cultural Bias and Alignment of Large Language Models
In-Run Data Shapley for Adam Optimizer
Edge Intelligence in Civil Aviation: Paradigms, Techniques, and Applications
Sidewalk Moments: Are Richer Representations Always More Human-Aligned? Evidence from Cit…
Anti-Goal Reasoning: Rethinking the Theory of Goal Reasoning in Non-Axiomatic Logic
DINO-VPT: Hierarchical Visual Prompt Tuning for Joint Physical-Digital Face Anti-Spoofing
Engine-Native Editable 3D World Reconstruction with Objects and Lighting
Webly Supervised Multi-Label Recognition: Evaluation Benchmark and Dual-Branch Multi-Labe…
ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes?
Probabilistic Residual Learning for Online Recommendations
Code Monitor Red Teaming for Public-Test-Passing Code
Offline RL with Hierarchical Action Chunking
REFACT: Adaptive Fact Restatement for Compact and Faithful Chain-of-Thought Reasoning
Khosla Ventures in Talks to Raise $5.5 Billion in New Funds
New Complexity-Theoretic Frontiers of Tractability for Neural Network Training
Open AI Models Spent Hours on Hack That Usually Takes Weeks
Profiling Lightweight Large Language Models