Explorar

Noticias de IA

30239 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Functorial Neural Architectures from Higher Inductive Types
arXiv cs.AI Research & Papers
REAL: Regression-Aware Reinforcement Learning for LLM-as-a-Judge
arXiv cs.AI Research & Papers
Empirical Characterization of Inference-Time Elicited Probability Transformations in Larg…
arXiv cs.AI Research & Papers
SimulCost: A Cost-Aware Benchmark and Toolkit for Automating Physics Simulations with LLMs
arXiv cs.AI Research & Papers
Variational Routing: A Scalable Bayesian Framework for Calibrated Mixture-of-Experts Tran…
arXiv cs.AI Security & Safety
Prompt Injection as Role Confusion
arXiv cs.AI Research & Papers
LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric R…
arXiv cs.AI Research & Papers
Language Models Learn Constructional Semantics, Not To Mention Syntax: Investigating LM U…
arXiv cs.AI Research & Papers
NGDBench: Towards Neural Graph Data Management
arXiv cs.AI Security & Safety
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
arXiv cs.AI Policy & Regulation
The Global Landscape of Environmental AI Regulation: From the Cost of Reasoning to a Righ…
arXiv cs.AI Research & Papers
Position: Evaluation of ECG Representations Must Be Fixed
arXiv cs.AI Research & Papers
DTBench: A Synthetic Benchmark for Document-to-Table Extraction
arXiv cs.AI Research & Papers
HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Lang…
arXiv cs.AI Research & Papers
A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agen…
arXiv cs.AI Research & Papers
Controllable Lung Nodule Synthesis via Histogram-Regularized Latent Diffusion Models
arXiv cs.AI Research & Papers
Effective Reasoning Chains Reduce Intrinsic Dimensionality
arXiv cs.AI Research & Papers
Breaking the Simplification Bottleneck in Amortized Neural Symbolic Regression
arXiv cs.AI Research & Papers
Stop the Flip-Flop: Context-Preserving Verification for Fast Revocable Diffusion Decoding
arXiv cs.AI Research & Papers
An Odd Estimator for Shapley Values
arXiv cs.AI Research & Papers
Mixture of Concept Bottleneck Experts
arXiv cs.AI Research & Papers
ParalESN: Enabling parallel information processing in Reservoir Computing
arXiv cs.AI Research & Papers
Multi-Agent Teams Hold Experts Back
arXiv cs.AI Policy & Regulation
PASTA: A Scalable Framework for Multi-Policy AI Compliance Evaluation
arXiv cs.AI Research & Papers
SKETCH: Semantic Key-Point Conditioning for Long-Horizon Vessel Trajectory Prediction
arXiv cs.AI Research & Papers
FEM-Bench: A Structured Scientific Reasoning Benchmark for Evaluating Code-Generating LLMs
arXiv cs.AI Research & Papers
Flow Equivariant World Models: Memory for Partially Observed Dynamic Environments
arXiv cs.AI Research & Papers
Mixture of Horizons in Action Chunking
arXiv cs.AI Research & Papers
Reasoning-Aware Multimodal Fusion for Hateful Video Detection
arXiv cs.AI Research & Papers
Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language M…