Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actuall…
Explorar
Noticias de IA
21270 elementos — filtrados, clasificados y sin duplicados
People Are Not Just Their Countries. Disentangling Social Determinants of LLM Value Align…
CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems
SkillEval: Decomposing Agent Skill Quality into Interpretable Signals
Deterministic Preprocessing and Interpretable Fuzzy Banding for Cost-per-Student Reportin…
Fast LapSum: Exact Differentiable Top-k at Million Scale
Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence
MemWM: Memory-Augmented Text-Based World Model
Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory
CASA: Classification Augmented with Safety Attention for Robust Multimodal Alignment
Surg-UniWorld: A Unified Surgical World Model with Multimodal Control Experts
From Cheap Fakes to Pure Synthesis: Addressing the New Era of T2V Fake News Videos
Mind the Gap: A Dual Knowledge Graph Framework for Unified Multi-task User Intent Inferen…
Automated item evaluation: Predicting item acceptance and rejection using LLM-generated c…
CellWorld: From Gene-Level Reconstruction to Latent Cell Prediction in Spatial Transcript…
MolBioKG: Grounding Out-of-Graph Molecules in Biomedical Knowledge Graphs via Multi-Resol…
Can MLLMs Decode the Creative Leap? Introducing C4 for Cross-Concept Understanding
TaskSense: Focusing on What Matters in World Models
MemPrism: Task-Conditioned Relational Memory Views for Long-Horizon Agents
PULSE: Agentic Investigation with Passive Sensing for Proactive Affective Intervention in…
Towards Multi-Label Graph Foundation Models: from Single-Vector Representation Learning t…
Beyond Routing Weights: Faithful Response-Level Interpretation of Mixture-of-Experts Rewa…
Critical Acclaim Orientation in Large Language Models: Evidence from Film Preference Elic…
ADIAS: Automated Design of Interactive Agentic Systems
Ask-E: An Environment for Calibrated Question Generation
Playing Games with My Heart: An Evaluation of AI Companion Apps
On Seeding Watermarks to Detect Verbatim LLM Copy-Paste Responses
Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low…
Cluster Attention for Graph Machine Learning
From Plan to Action: How Well Do Agents Follow the Plan?