A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agen…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Effective Reasoning Chains Reduce Intrinsic Dimensionality
Breaking the Simplification Bottleneck in Amortized Neural Symbolic Regression
Reliable Self-Improvement Training by Verifying Reasoning, Not Just Answers
Linear Ordering Problem: Time for a Change
Automatically Attacking Software Reverse Engineering AI Agents
Vector Linking via Cross-Model Local Isometric Consistency
Beyond Memorization: Assessing Semantic Generalization in Large Language Models Using Phr…
Stop the Flip-Flop: Context-Preserving Verification for Fast Revocable Diffusion Decoding
UniScale: Adaptive Unified Inference Scaling via Online Joint Optimization of Model Routi…
Structure-Induced Information for Rerooting Levin Tree Search
Routing on the Stiefel Manifold: When Does Adaptive Subspace Selection Help for Cross-Dom…
An Odd Estimator for Shapley Values
Mixture of Concept Bottleneck Experts
Generating Reports or Repeating Templates? Measuring and Mitigating Template Collapse in …
Annealed Softmax Greedy in Many-Armed Bayesian Bandits
ParalESN: Enabling parallel information processing in Reservoir Computing
Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Au…
DEM: A Distilled Explanation Model for Interpretable Anomaly Detection in Physiological S…
TraceGraph: Shared Decision Landscapes for Diagnosing and Improving Agent Trajectories
Multi-Agent Teams Hold Experts Back
SKETCH: Semantic Key-Point Conditioning for Long-Horizon Vessel Trajectory Prediction
ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and …
Reading Between the Citations: A Typed Claim Network for Scientific Literature
Do Large Language Models Encode Institutional Experience? Evidence from Cross-Linguistic …
A Persona-Based Evaluation Framework for Pluralistic Alignment in Generative AI
HypoAgent: An Agentic Framework for Interactive Abductive Hypothesis Generation over Know…
What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucinati…
BlueFin: Benchmarking LLM Agents on Financial Spreadsheets
Inverse Reinforcement Learning without an Optimal Demonstrator: A Feasible Reward Set App…