Trends in AI and Human-AI Interaction in Clinical Trials -- A Hybrid Human-AI Exploration
Explorar
Noticias de IA
30321 elementos — filtrados, clasificados y sin duplicados
The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic D…
OmniRetrieval: Unified Retrieval across Heterogeneous Knowledge Sources
BlockBatch: Multi-Scale Consensus Decoding for Efficient Diffusion Language Model Inferen…
Toward Ethical Facial Age Estimation: A Generalized Zero-Shot Benchmark Without Training …
Compute Allocation in Evolutionary Search: From Depth-Breadth to Multi-Armed Bandits
Causal Label Recovery in Payment Networks
LoopFM: Learning frOm HistOrical RePresentations of Foundation Model for Recommendation
Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Mul…
A Minimal Bifurcation Model of Load Imbalance in a Softmax Mixture-of-Experts Router
When and How Long? The Readout-Mediator Angle in Temporal Reasoning
When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodi…
Composing Non-Conjugate Factor Graphs with Closed-Form Variational Inference
How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignmen…
Test Time Training for Supervised Causal Learning
Do Proactive Agents Really Need an LLM to Decide When to Wake and What to Anchor?
AgentLens: Revealing The Lucky Pass Problem in SWE-Agent Evaluation
Many-Shot CoT-ICL: Making In-Context Learning Truly Learn
MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing
Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
AttenA+: Rectifying Action Inequality in Robotic Foundation Models
TRACER: Persistent Regularization for Robust Multimodal Finetuning
On the Geometry of Games and their Solvers
The Good, the Bad, and the Ugly of Markov Boundary for Tabular Prediction
Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Po…
iLoRA: Bayesian Low-Rank Adaptation with Latent Interaction Graphs for Microbiome Diagnos…
ProtoMedAgent: Multimodal Clinical Interpretability via Privacy-Aware Agentic Workflows
EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents
Hilbert-Geo: Solving Solid Geometric Problems by Neural-Symbolic Reasoning
Unveiling Multi-regime Patterns in SciML: Distinct Failure Modes and Regime-specific Opti…