LegoLM: Structured Weight Sharing for Large Language Models
Explorar
Noticias de IA
21270 elementos — filtrados, clasificados y sin duplicados
RippleKV: Cross-Layer KV Cache Allocation via Perturbation Propagation
CuteTTS: Efficient and High-Quality Speech Synthesis via Autoregressive Modeling of Conti…
Dynamic gain neuromodulation attenuates the stability gap under joint training
TomaMMU: A Comprehensive Multimodal Understanding Benchmark for Tomato Leaf Diseases
DoGMA: A Central-Dogma-Guided Foundation Model for Multi-Omics Alignment and Multi-Task L…
Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution
How to Ask the AI: A User Perspective Survey for Large Language Model Prompting
Verication-driven closed-loop multi-agent large language modelframework for code-complian…
Distilling CT Foundation Models into Editable Concept Bottlenecks for Lung Nodule Maligna…
MMArch: Benchmarking Multimodal Reasoning Grounded in Architectural Evidence
DA-NBV: A Direction-Aware Next-Best-View Planner for Efficient 3D Reconstruction of Ships…
COMEX: A Composition-Grounded Benchmark and Learning Framework for Explainable Aesthetic …
Spectral Outliers Reveal Dominant Learned Structure in Transformer Attention
FitAQA: A Benchmark of Fitness Action Quality Assessment for Multimodal Large Language Mo…
Persistent Semantic Entities in Tool-Augmented LLM Systems
RAG-Based Auto-Configuration for Industrial Fieldbus Devices
Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Pro…
How Far Do Foundation Models Transfer to Infant Signals? A Cross-Dataset Transfer Audit w…
NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs
Dramarrator: Object-Based Audio Editing for Audio Drama Production from Books
Deployable Per-Instance Multi-Layer Activation Steering for Large Language Models
Detecting Clear Contact Lenses for Iris Recognition: A Two-Stage Mask-Guided Attention Ap…
Continuous Interaction Diffusion: A Diffusion-Native Runtime for Asynchronous Tool-Augmen…
Conversational versus Dashboard Explainable AI for UAV Intrusion Detection: An Empirical …
Do Time-Series Forecasters Use the Right History: Recoverability, Recovery, and Functiona…
GeoSeg-OV: Bridging Geospatial Gaps with Structural Guidance for Open-Vocabulary Remote S…
VisEditBench: Can Vision-Language Models Edit Visualization Code from Multimodal Feedback?
Automatic Field-of-View Adjustment for a View-Expansive Microscope via LSTM-Based Gaze an…
FormStruct-Bench:A Hierarchical and Diagnostic Benchmark for Table-Form Document Structur…