Latent Reasoning in TRMs is Secretly a Policy Improvement Operator
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers
CARES: Context-Aware Resolution Selector for VLMs
Characterizing Web Search in The Age of Generative AI
Rethinking RL Evaluation: Can Benchmarks Truly Reveal Failures of RL Methods?
Make a Video Call with LLM: A Measurement Campaign over Six Mainstream Apps
End-to-End Deep Learning for Predicting Metric Space-Valued Outputs
A Methodological Framework for Explicit Control of the Speed-Accuracy Trade-off in Brain-…
Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice v…
Beyond Model Base Retrieval: Weaving Knowledge to Master Fine-grained Neural Network Desi…
DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?
A Lightweight Context-Driven Training-Free Network for Scene Text Segmentation and Recogn…
EuroBERT: Scaling Multilingual Encoders for European Languages
Efficient Weighted Sampling via Score-based Generative Models
Visual-Noise Guided In-Context Distillation for Multimodal Large Language Model Unlearning
CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated…
Score Function Gradient Estimation to Widen the Applicability of Decision-Focused Learning
Coding Agent Is Good As World Simulator
Agricultural Landscape Understanding At Country-Scale
Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigat…
KnowledgeBerg: Evaluating Systematic Knowledge Coverage and Compositional Reasoning in La…
RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography
Diffusion Image Generation with Explicit Modeling of Data Manifold Geometry
Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory
Model Parallelism With Subnetwork Data Parallelism
Aligning Cellular Sheaves with Classifier Attention for Interpretable Weakly-Supervised P…
On Information Self-Locking in Reinforcement Learning for Active Reasoning of LLM agents
LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs?
DLLM-JEPA: Joint Embedding Predictive Architectures for Masked Diffusion Language Models
SentimentLens: Reconciling Sentiment and Ratings via Dual-Modality in the Hospitality Sec…