Beyond Reproducibility: Towards Security-Aware Evaluation of Research Artifacts
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
A Low-Cost, Open Platform for End-to-End Autonomous Driving on a Miniature Ackermann Vehi…
Artificial Intelligence for Energy Optimization in Data Centers
Transfiver: Human-AI Co-Inference through a Shared Editable State
Semantic Bayesian World Models
Refusal geometry reflects refusal training: diverse refusal prefixes can raise stable ran…
Making Every Tool Call Count: Necessary Tool-Evidence Path Rewards for Agentic Vision-Lan…
A computable representation of the physical laboratory enables verifiable workflows
Dalek: A Constructive Agent Machine
Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Tran…
Identifying AI Web Scrapers Using Canary Tokens
LLM Evaluation as Tensor Completion: Low Rank Structure and Semiparametric Efficiency
CASCADE: A Component Ablation and Corpus Audit of a Layered Local Defense for MCP-Based S…
LRConv-NeRV: Low Rank Convolution for Efficient Neural Video Compression
One Model to Translate Them All? A Journey to Mount Doom for Multilingual Model Merging
DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training
PeroMAS: A Multi-agent System of Perovskite Material Discovery
HOMURA: Taming the Sand-Glass for Time-Constrained LLM Translation via Reinforcement Lear…
Mixed Data Clustering Survey and Challenges
VoxPrivacy: A Benchmark for Evaluating Interactional Privacy of Speech Language Models
User Perceptions vs. Proxy LLM Judges: Privacy and Helpfulness in LLM Responses to Privac…
Evolving Excellence: Automated Optimization of LLM-based Agents
Short-Window Sliding Learning for Real-Time Violence Detection via LLM-based Auto-Labeling
AnyBox: Efficient Zero-Shot 9DoF Pose Estimation of Boxes for Robotic Manipulation
EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering
Decentralized Vision-Based Autonomous Aerial Wildlife Monitoring
Medical Reasoning in the Era of LLMs: A Systematic Review of Enhancement Techniques and A…
ScoreMix: Synthetic Data Generation by Score Composition in Diffusion Models Improves Rec…
LDC: Learning to Generate Research Idea with Dynamic Control
AgentRM: Enhancing Agent Generalization with Reward Modeling