Managing Uncertainty in LLM-Generated Procedural Knowledge for Virtual Laboratory Planning
Explorar
Noticias de IA
30294 elementos — filtrados, clasificados y sin duplicados
Certified Purity for Cognitive Workflow Executors: From Static Analysis to Cryptographic …
Jailbreak susceptibility prediction and mitigation via the behavioral geometry of models
Annotator Positionality as Signal: Psychometric Weighting for Anti-Autistic Ableism Detec…
Reconstructing Multi-Scale Physical Fields from Extremely Sparse Measurements with an Aut…
Linear and Neural Dueling Bandits with Delayed Feedback
Certified Causal Attribution for Real-Time Attack Forensics in 6G Network Slicing
Prospective evaluation of multimodal respiratory failure prediction: Do chest X-rays impr…
The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence
FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives
Tool Calling is Linearly Readable and Steerable in Language Models
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs
RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations
When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of…
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
Recursive Flow Matching
Composition Collapse: Stable Factual Knowledge Does Not Imply Compositional Reasoning
Where Hindsight Credit Can Reside: A Signed-Capacity View of Token Updates in RLVR
Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Langua…
Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments
The ATOM Report: Measuring the Open Language Model Ecosystem
Governed Capability Evolution: Lifecycle-Time Compatibility Checking and Rollback for AI-…
SenBen: Sensitive Scene Graphs for Explainable Content Moderation
Unified Panoramic Geometry Estimation via Multi-View Foundation Models
Understanding the Challenges in Iterative Generative Optimization with LLMs
From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Ques…
Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argum…
Why LLMs Hallucinate on Structured Knowledge: A Mechanistic Analysis of Reasoning over Li…
Detached Skip-Links and $R$-Probe: Decoupling Feature Aggregation from Gradient Propagati…