Best-of-Evidence: Best-of-N Selection under Partial Verification
Explorar
Noticias de IA
22116 elementos — filtrados, clasificados y sin duplicados
not much happened today
Representing Entity Importance in AI Knowledge Systems: A Dual-Signal Framework of Audien…
MagicMakeup: A Region-Controllable Diffusion Transformer for High-Fidelity Makeup-Transfer
Three-Pronged Spectral Control for Federated Parameter Efficient Fine Tuning
Multi-turn RL with Structural and Performance Aware Rewards for CUDA Kernel Generation
Active Inference as a Convex Markov Decision Process
Hybrid LSTM-Graph Neural Framework for Robust Financial Fraud Detection and Adversarial R…
Benchmarking Confidential GPU Inference on NVIDIA H100 under Intel TDX
FormulaSPIN: Self-Play Fine-Tuning for Natural Language to Spreadsheet Formula Generation
Information Discernment in Large Language Models
LISA: Linear-Indexed Sparse Attention for Efficient Long-Context Reasoning
Lifted Representation Hypothesis in Language Models
AdaRoPE: Not All Attention Heads Should Rotate and Scale Equally
Statistically Grounded Sparse-Feature Interventions for Activation-Space Control in Large…
Logic-Guided Data Extraction with Answer Set Programming and Large Language Models
Rethinking Uncertainty Evaluation in Large Language Models
Spectral-LSH: Sub-Quadratic Prompt Compression via Krylov-Projected Locality-Sensitive Ha…
Beyond Tracking or Shortcut: Composition-Bounded Predictive States in Poker Autoregressiv…
Mitigating Scaffolding Collapse in Socratic Tutors via Representation Alignment
Euclean: Automated Geometry Problem Formalization with Unified Verification in Lean
HyGRL: Adaptive Hybrid Graph Reasoning for Multi-Entity Questions
FORCE-Bench: A Benchmark, Dataset, and Evaluation Harness for Agentic AI in Enterprise Fi…
Sophisticated Policies from Epistemic Priors
Edge Intelligence in Civil Aviation: Paradigms, Techniques, and Applications
Rewarding Better Thinking for LLM Preference Alignment
Symbol and Footprint Database for Electronic Components by Agentic Recognition and Genera…
Long-Term Sequential Decision Making under Risk
MOF-Sleuth: Tool-Grounded Reward Alignment for Explainable Fine-Grained MOF CIF Auditing
SenWorld: A Digital-Twin Simulation for Generating Context-Rich Evaluation Data