MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop
Explorar
Noticias de IA
21861 elementos — filtrados, clasificados y sin duplicados
LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs?
On Information Self-Locking in Reinforcement Learning for Active Reasoning of LLM agents
Model Parallelism With Subnetwork Data Parallelism
Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory
RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography
KnowledgeBerg: Evaluating Systematic Knowledge Coverage and Compositional Reasoning in La…
Agricultural Landscape Understanding At Country-Scale
Coding Agent Is Good As World Simulator
Score Function Gradient Estimation to Widen the Applicability of Decision-Focused Learning
Efficient Weighted Sampling via Score-based Generative Models
A Lightweight Context-Driven Training-Free Network for Scene Text Segmentation and Recogn…
Beyond Model Base Retrieval: Weaving Knowledge to Master Fine-grained Neural Network Desi…
Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice v…
End-to-End Deep Learning for Predicting Metric Space-Valued Outputs
Make a Video Call with LLM: A Measurement Campaign over Six Mainstream Apps
Rethinking RL Evaluation: Can Benchmarks Truly Reveal Failures of RL Methods?
Characterizing Web Search in The Age of Generative AI
CARES: Context-Aware Resolution Selector for VLMs
Optimal Bayesian Stopping for Efficient Inference of Consistent LLM Answers
Latent Reasoning in TRMs is Secretly a Policy Improvement Operator
ShelfAware: Real-Time Semantic Localization in Quasi-Static Environments with Low-Cost Se…
VocSim: A Training-free Benchmark for Zero-shot Content Identity in Single-source Audio
Calibrating Uncertainty for Zero-Shot Adversarial CLIP
Dynamic Entropy Tuning in Reinforcement Learning Low-Level Quadcopter Control: Stochastic…
Paradoxical noise preference in RNNs
Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics
Physics-Encoded Inverse Modeling for Arctic Snow Depth Prediction
Med-Scout: Curing MLLMs' Geometric Blindness in Medical Perception via Geometry-Aware RL …
Global Geometry Is Not Enough for Vision Representations