Functorial Neural Architectures from Higher Inductive Types
Explorar
Noticias de IA
30239 elementos — filtrados, clasificados y sin duplicados
REAL: Regression-Aware Reinforcement Learning for LLM-as-a-Judge
Empirical Characterization of Inference-Time Elicited Probability Transformations in Larg…
SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs
Variational Routing: A Scalable Bayesian Framework for Calibrated Mixture-of-Experts Tran…
Prompt Injection as Role Confusion
LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric R…
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM U…
NGDBench: Towards Neural Graph Data Management
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
The Global Landscape of Environmental AI Regulation: From the Cost of Reasoning to a Righ…
Position: Evaluation of ECG Representations Must Be Fixed
DTBench: A Synthetic Benchmark for Document-to-Table Extraction
HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Lang…
A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agen…
Controllable Lung Nodule Synthesis via Histogram-Regularized Latent Diffusion Models
Effective Reasoning Chains Reduce Intrinsic Dimensionality
Breaking the Simplification Bottleneck in Amortized Neural Symbolic Regression
Stop the Flip-Flop: Context-Preserving Verification for Fast Revocable Diffusion Decoding
An Odd Estimator for Shapley Values
Mixture of Concept Bottleneck Experts
ParalESN: Enabling parallel information processing in Reservoir Computing
Multi-Agent Teams Hold Experts Back
PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation
SKETCH: Semantic Key-Point Conditioning for Long-Horizon Vessel Trajectory Prediction
FEM-Bench: A Structured Scientific Reasoning Benchmark for Evaluating Code-Generating LLMs
Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments
Mixture of Horizons in Action Chunking
Reasoning-Aware Multimodal Fusion for Hateful Video Detection
Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language M…