ICR-RL: Deep Reinforcement Learning via In-Context Regression
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Agentic Artificial Intelligence for Multistage Physics Experiments at a Large-Scale User …
Polarity Detection of Sustainable Development Goals in News Text
Restricted Bernoulli Matrix Factorization: Balancing the trade-off between prediction acc…
Adaptive Margin RLHF via Preference over Preferences
OctoPipe: Reducing Pipeline Bubbles for Heterogeneous Models via Co-Optimizing Partitioni…
Learning to Visually Connect Actions and their Effects
MAD-PINN: A Decentralized Physics-Informed Machine Learning Framework for Safe and Optima…
Graph Unitary Message Passing
VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models
Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation…
Robust Counterfactual Explanations under Model Multiplicity Using Multi-Objective Optimiz…
The Three Regimes of Offline-to-Online Reinforcement Learning
Evaluating LLM-Based Regression Test Generation
Verifier-free Test-Time Sampling for Vision-Language-Action Models
TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Releva…
SilvaScenes: Tree Detection and Species Classification from Under-Canopy Images in Natura…
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of A…
Leveraging Natural Language Processing to Unravel the Mystery of Life: A Review of NLP Ap…
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models
A Model Can Help Itself: Reward-Free Self-Training for LLM Reasoning
ELBO-T2IAlign: A Generic ELBO-Based Method for Calibrating Pixel-level Text-Image Alignme…
Last Layer Hamiltonian Monte Carlo
Extending Foundational Monocular Depth Estimators to Fisheye Cameras with Calibration Tok…
VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models
Empirical Computation: Prompting versus Programming
Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters
Large Language Models Develop Novel Social Biases Through Adaptive Exploration
SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards
Exploring Context-aware and LLM-driven Locomotion for Immersive Virtual Reality