Wiring Beats Blending: What Transfers Between Transformer Sizes -- and What Doesn't
Explorar
Noticias de IA
29349 elementos — filtrados, clasificados y sin duplicados
AFD-Ledger: Deployment Provisioning for Attention--FFN Disaggregation
GeoReward: Mitigating Contextual Variable Overestimation in Vision-Language Models for Cr…
IMFACT: Counterfactual Explanations for Time Series via Intrinsic Mode Function Substitut…
A Long-Run Persistence Theory for AI Systems under the Redundancy-Adjusted Artificial Age…
CheMLFlow: An Open-Source Platform for Cheminformatics and Materials Informatics Applicat…
TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale
ORCA-bench: How Ready Are Language Model Agents for Oncall?
CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research
iStructTab: Structured Feature Sequencing for Multimodal Learning of Image and Tabular Da…
Monte Carlo Tree Search for Table-to-Multimodal Report Generation
A Systematic Review and Taxonomy of Reinforcement Learning-Model Predictive Control Integ…
Just Repair: A Minimal Denoising Network for Time Series Anomaly Detection
Image Classification Using CNN-QNN Hybrid Model with Optimized Correlated Features
From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs
Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Mode…
DeepImagine: Clinical Trial Outcome Prediction via Stepwise Local Counterfactual Imaginat…
Place-it-R1: Unlocking Environment-aware Reasoning Potential of MLLM for Video Object Ins…
Formal Analysis and Supply Chain Security for Agentic AI Skills
MOON3.0: Reasoning-aware Multimodal Representation Learning for E-commerce Product Unders…
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
A Survey of Agent Memory in the Second Half: Towards Self-Evolving and Long-Horizon Agents
Beyond the Dirac Delta: Mitigating Diversity Collapse in Reinforcement Fine-Tuning for Ve…
RooflineBench: A Benchmarking Framework for On-Device LLMs via Roofline Analysis
Generative Experiences for Digital Mental Health Interventions: Evidence from a Randomize…
AdaCorrection: Adaptive Offset Cache Correction for Accurate Diffusion Transformers
D$^2$F-ReAG: Dynamic Decomposition and Filtering for Multi-Hop Reasoning-Augmented Genera…
When Large Language Models Know the Table: A Framework for Assessing Data Contamination i…
Neural Diversity Regularizes Hallucinations in Language Models
Bi-Level Reinforcement Learning Pathway for Sim-to-Real Optimality