WhiFlash: Accelerating Speculative Decoding with Token-Level Cross-Paradigm Routing
Explorar
Noticias de IA
22104 elementos — filtrados, clasificados y sin duplicados
Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward …
SearchSwarm: Towards Delegation Intelligence in Agentic LLMs for Long-Horizon Deep Resear…
Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting
BRAIN: Bayesian Reasoning via Active Inference for Agentic and Embodied Intelligence in M…
From `May' to `Is': Certainty Distortion in Language Model Rewriting
Bidirectional Small-Granularity Search between Code and Text
ABLE: Representing and Mapping LLMs via Attribution-Based Large-model Embedding
Post-training is (Massive) Supervised Learning
mllm-shap: A Shapley Value Explainability Platform for Text-Audio Multimodal Large Langua…
Principled Agent Debate: Adversarial Arbitration for Sycophancy Reduction in Large Langua…
Item Response Scaling Laws: A Measurement Theory Approach for Efficient and Generalizable…
Query Lens: Interpreting Sparse Key-Value Features with Indirect Effects
RAPID: Layer-Wise Redundancy-Aware Pruning and Importance-Driven Token Merging for Effici…
Beyond Accuracy: Interpreting Topic Representation in Suicide Ideation Detection Models
Quantum-Enhanced Similarity Measures for Polarimetric Materials Classification
Decoupling Semantics and Logic: A Training-Free Coarse-to-Fine Pipeline for Video Retriev…
Summarization is Not Dead Yet
Human-Centered Benchmarking of Driver Monitoring Models
Few-shot Class-variable Incremental Audio Classification via Prototype Adaptation and Pse…
Constrained Paraphrase Consistency for LLM Hallucination Detection
CLASP: Language-Driven Robot Skill Selection and Composition using Task-Parameterized Lea…
GlobeAudio: A Multilingual Multicultural Benchmark for Naturalistic Evaluation of Large A…
How Deep Are Deep GPs, Really? A Sharp Threshold and a Non-Gaussian Limit for Composition…
More Yap Less Meaning: Uncovering Self-Improvement Behavior in SLMs
TRIAGE: Dialectical Reasoning for Explainable Risk Prediction on Irregularly Sampled Medi…
Failure-Aware Refinement of Vision-Language Model for Lithography Defect Detection
PAI: Preserving Amplitude Information in Representation-Based Time-Series Anomaly Detecti…
PACT: Learning Diverse Diagnostic Strategies via Privileged Synthesis and Branch Consensus
NutriMLLM: Multimodal Large Language Models for Dietary Micronutrient Analysis