What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend
Explorar
Noticias de IA
21272 elementos — filtrados, clasificados y sin duplicados
Towards a satellite image manipulation and deepfake localization benchmark dataset
Masked diffusion enables coherent beat tracking
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models
A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination
SimMOF: AI agent for Automated MOF Simulations
Easy to Complete, Hard to Choose: Investigating LLM Performance on the ProverbIT Benchmark
Representational separation between unitary and channel quantum generative models via sha…
AudioScape-TTA: A Structured Soundscape Benchmark for Fine-Grained Text-to-Audio Evaluati…
CARVE: Cross-Slice Anisotropic Reallocation of Visual Evidence for Efficient 3D Medical V…
EuroExec: Frontier Language Models Fall Short of Expert Judgment on European Executive De…
When Absence Is Evidence: Evaluating Completeness-Sensitive Negative Reasoning in Large L…
TwinIR: Coordinated Invisible Dual-Point Attacks on Online HD Map Construction
Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Medi…
Tropical Algebraic Geometry for Neuronal Representations: An Arakelov-Green Measure Based…
Generative Optimization for Incentivized Advertising with Global Level Constraints
ExeCRE: Execution-Consistency Guided Reliability Estimation for Self-Correcting Code Gene…
EndoVLM: An Endoscopy Vision-Language Pre-training Model via Anatomy-Guided Sparsity and …
Rethinking Reservoir Pruning: A Dynamical Perspective for Echo State Networks
Approximate Multi-Objective Search Under Rulebooks
Towards Trustworthy Hypergraph Neural Networks under Label Noise
Training-Free Hashing-Based Attention via Binary Principal Components
MESH: Memory-Efficient Sinkhorn Optimization for Mixture-of-Experts Training
Not Every Divergence Should Be Suppressed: Counterfactual Recoverability in On-Policy Dis…
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation
HyPASE: Hyperbolic Geometry for Parameter-Efficient Speech Emotion Fine-Tuning Framework …
Combating Knowledge Corruption in Agent Systems: A Byzantine-Tolerant Secure Collaborativ…
The Order Is the Guarantee: Verifier-Budgeted Code Deletion with Static-First Learned Pro…
FinReportBench: Measuring and Improving Institution-Grade Financial Report Generation
Robust and Efficient Motion Reasoning for Privacy-Aware Classroom Incident Recognition