ModelEquivBench: Certifying Multi-Relational Evaluation of LLM-Generated Optimization Mod…
Explorar
Noticias de IA
30294 elementos — filtrados, clasificados y sin duplicados
Combining Large Language Models and Symbolic Reasoning for Multi-Robot Temporal Planning …
Latent Actions from Factorized Transition Effects under Agent Ambiguity
FBFM: A Training-Free Asynchronous Feedback Mechanism for Flow-Matching in World-Action M…
The Capability Convergence Hypothesis: Capability from Access Structure, Not Scale
DynaResize: Runtime GPU Reallocation for Disaggregated LLM Post-Training
SREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios
Fast Feature Field ($\text{F}^3$): A Predictive Representation of Events
On the Expressive Power of Sparse Geometric MPNNs
Revisiting Multi-Permutation Equivariance through the Lens of Irreducible Representations
SEDR-Seq2P: A Lightweight Dilated Residual Sequence-to-Point Network for Multi-Task Indus…
DungeonBench: A Benchmark for Rules-Rich Tactical Reasoning in Dungeons & Dragons Combat
A Human-Centered Validation of the Explainability-Performance Coefficient
Dual-Force: Enhanced Offline Diversity Maximization under Imitation Constraints
ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction
Predicting Steel Fatigue Life from Micrographs Using Physics-Informed Deep Learning
Scaffolding Critical Engagement with GenAI: Transforming Ethnic Minority Preparatory Stud…
MOT-SR: Multi-Objective Tool-Augmented Scientific Equation Discovery with Large Language …
Reproducing Human Individual Motor Signatures: A Data-Driven Approach for Repetitive Moti…
Epistemic-aware Vision-Language Foundation Model for Fetal Ultrasound Interpretation
RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems
Topology-Aware Data Movement for Disaggregated GPU Inference
The Asymmetric Effects of Knowledge Distillation on Bias in Small Language Models
Seeing Differently: Modeling Interpretive Perspectives in Computational Creativity using …
Monotone and Separable Set Functions: Characterizations and Neural Models
Shall We Play a Game? Language Models for Open-ended Wargames
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feed…
Sensitivity Analysis of GRU, LSTM and Transformer Encoder in Classification of Automated …
Guarantees on Dynamical System Distinguishability for LLM Token Generation
Dual-Dimensional Consistency: Balancing Budget and Quality in Adaptive Inference-Time Sca…