Beyond Reproducibility: Towards Security-Aware Evaluation of Research Artifacts
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Identifying AI Web Scrapers Using Canary Tokens
PeroMAS: A Multi-agent System of Perovskite Material Discovery
LRConv-NeRV: Low Rank Convolution for Efficient Neural Video Compression
FedPS: Federated Preprocessing for structured data via aggregated Statistics
LLM Evaluation as Tensor Completion: Low Rank Structure and Semiparametric Efficiency
VoxPrivacy: A Benchmark for Evaluating Interactional Privacy of Speech Language Models
CASCADE: A Component Ablation and Corpus Audit of a Layered Local Defense for MCP-Based S…
HOMURA: Taming the Sand-Glass for Time-Constrained LLM Translation via Reinforcement Lear…
Temperature Scaling Attack Disrupting Model Confidence in Federated Learning
AnyBox: Efficient Zero-Shot 9DoF Pose Estimation of Boxes for Robotic Manipulation
User Perceptions vs. Proxy LLM Judges: Privacy and Helpfulness in LLM Responses to Privac…
Short-Window Sliding Learning for Real-Time Violence Detection via LLM-based Auto-Labeling
Mixed Data Clustering Survey and Challenges
Decentralized Vision-Based Autonomous Aerial Wildlife Monitoring
EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering
Evolving Excellence: Automated Optimization of LLM-based Agents
One Model to Translate Them All? A Journey to Mount Doom for Multilingual Model Merging
LightEMMA: A Longitudinal Evaluation of Vision-Language Models for Autonomous Driving
LDC: Learning to Generate Research Idea with Dynamic Control
ScoreMix: Synthetic Data Generation by Score Composition in Diffusion Models Improves Rec…
AgentRM: Enhancing Agent Generalization with Reward Modeling
Medical Reasoning in the Era of LLMs: A Systematic Review of Enhancement Techniques and A…
Not All Preferences Deserve Gradients: Understanding Gradient Utility in Offline Reasonin…
WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquit…
PaperScout: An Autonomous Agent for Academic Paper Search with Process-Aware Sequence-Lev…
ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize
Grammar-Aligned Decoding
RECAST: Expanding the Boundaries of LLMs' Complex Instruction Following with Multi-Constr…
Complete Identification of Deep ReLU Networks through {\L}ukasiewicz Logic