The First EgoCross Challenge at EgoVis 2026: Cross-Domain Egocentric Video Question Answe…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Revealed Rationality: Label-Free Evaluation and Regularization from Representation Theore…
ORACLE: A Multi-Objective Reinforcement Learning-Based Analog Circuit Design Optimizer wi…
When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LL…
Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First
COMPAS: Difficulty-Aware Joint Search for Optimizing Code Generation
Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits
Compass: Continuously Aligning Social Media Feeds via In-Situ Reflections
A Unified Model for Cross-Domain Clone Detection via Model Merging
Visualizing Graph-to-Answer Mechanism Recovery in Materials-Science Hypothesis Generation
InsightEmb: Learning Action-Intent Embeddings for Agentic Insight Retrieval
Design Choices That Matter: A Functional ANOVA Analysis for Remote Sensing Multi-Label Cl…
Active-SWE: Benchmarking Coding Agents for Proactive Bug Fixing without Issue Reports
ABSeeker: Training Long-Horizon Search Agents via Answer-Backtracked Credit Assignment
NodeJEPA: Structure-Conditioned Latent Prediction for Node-Level Graph Self-Supervised Le…
A/B Agent: A Self-Evolving Agent for Strategy Iteration in Industrial A/B Testing
Leak-Resistant Unlearning: A New Benchmark for Evaluating Multi-Hop Reasoning Consistency…
Architectural Implications of Agentic AI Workflows
NeuMoSync: End-to-End Neuromodulatory Control for Plasticity and Adaptability in Continua…
AgentForge: An Immersive Role-Playing Platform for Learning Agentic Software Engineering
Perception Before Reasoning: Dynamic Latent Reasoning for Video Understanding and Questio…
Out-Of-The-Loop Multi-Fidelity Bayesian Optimization
Recurrent Residual Quantization: A Progressive Multi-Precision Representation for LLMs
FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents
FinProBench: Evaluating Financial AI Agents with Role-Grounded Rubrics Derived from Profe…
Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings
PhysMind: From Video to Executable Worlds for Training-Free Physical Reasoning
A Model Merging Approach for Continual MLLM Unlearning
GUARD: Grounding Uncertainty and Ablation-Based Risk Detection for Diffusion-Based VLAs
Beyond Linear Dynamics: Neural Bilinear Dynamical Models for Time Series Forecasting