Governed Capability Evolution: Lifecycle-Time Compatibility Checking and Rollback for AI-…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
The ATOM Report: Measuring the Open Language Model Ecosystem
Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Langua…
Where Hindsight Credit Can Reside: A Signed-Capacity View of Token Updates in RLVR
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of…
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
Tool Calling is Linearly Readable and Steerable in Language Models
FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives
Rotation-Invariant Spherical Watermarking via Third-Order SO(3) Representation Coupling
L2Rec: Towards Dual-View Understanding of LLMs for Personalized Recommendation
Towards Generalization-Oriented Models for Vehicle Routing Problems with Mixture-of-Exper…
SeDT: Sentence-Transformer Decision-Transformer Conditioning for Multi-Turn Conversation …
HTMLCure: Turning Browser Experience into State Guided Repair for Interactive HTML
The Kalman Evolve: Closing the Gap in Kalman Filtering via Interpretable Algorithm Discov…
Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Models
Strategies for Guiding LLMs to Use Software Design Patterns: A Case of Singleton
Tournament-GRPO: Group-Wise Tournament Rewards for Reinforcement Learning in Open-Ended L…
Tracing Computation Density in LLMs
QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Soci…
LitSeg: Narrative-Aware Document Segmentation for Literary RAG
An investigation of AI integration in sound designer workflows and experiences
TWIST: Closed-Loop token Synchronization for Application-Aware Wireless Digital Twins
LUCoS: Latent Unsupervised Context Selection for Tabular Foundation Models
Pair-In, Pair-Out: Latent Multi-Token Prediction for Efficient LLMs
Does RAG Know When Retrieval Is Wrong? Diagnosing Context Compliance under Knowledge Conf…
It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertain…
EdgeFlow: Edge-Map Augmented VLM-Based Flowchart Processing for Industrial Requirements E…
LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding