Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
DeCo-MIL: Debiased Counterfactual Reasoning for Long-Tailed Whole Slide Image Analysis
arXiv cs.AI Research & Papers
SMOPD: Selective Token-Entropy Masking for Dirty-History Multi-Turn On-Policy Self-Distil…
arXiv cs.AI Research & Papers
iFuzz-Meta: An Interpretable Fuzzy Learning Framework Bridging Top-Down and Bottom-Up Kno…
arXiv cs.AI Research & Papers
Measuring Reward Hacking and Reasoning-Answer Decoupling Under Position-Confounded Optimi…
arXiv cs.AI Research & Papers
When Do Cheap Probes Predict Expensive Training? Probing 3D-CT Encoders for Text Generati…
arXiv cs.AI Research & Papers
Which Question Is Your Attention Metric Answering? Attention Rows as Compositional Data
arXiv cs.AI Research & Papers
Beyond Correctness: Toward Automated Novelty Verification with Lean 4
arXiv cs.AI Research & Papers
From Generalist to Specialist: A Context-Fusion Framework for Endoscopic Polyp Reporting …
arXiv cs.AI Research & Papers
Admission Without Answers: Label-Free Certification and Experience Learning for LLM-Based…
arXiv cs.AI Research & Papers
Unraveling the Size Determination Mechanism of Nanocrystal Synthesis via Interpretable Ne…
arXiv cs.AI Research & Papers
PolyComp: A Polycube-based Benchmark for Compositional 3D Spatial Reasoning in Multimodal…
arXiv cs.AI Research & Papers
JarvisBench: Always-on Intelligence Between Humans and Agents
arXiv cs.AI Research & Papers
When Entropy Is Not Enough: Reclaiming Lost Semantics in LLM Output Length Prediction
arXiv cs.AI Research & Papers
Do Personalized Skills Help Coding Agents? An Empirical Study of Developer Interaction Hi…
arXiv cs.AI Research & Papers
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-…
arXiv cs.AI Research & Papers
Rotation-Invariant Multi-IMU Activity Recognition under Independent Per-Location Orientat…
arXiv cs.AI Research & Papers
TRACE: Trajectory Aware Reasoning for Multi-Turn Adversarial Conversation Evaluation
arXiv cs.AI Research & Papers
Agent Gym: A Framework for Continuous Evaluation and Evolution of LLM Agents Through Huma…
arXiv cs.AI Research & Papers
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Mo…
arXiv cs.AI Research & Papers
AgentMV: A State-Guided Multi-Agent Framework for Budget-Aware Music Video Generation
arXiv cs.AI Research & Papers
Improving the matrix multiplication exponent with modern optimization and AlphaEvolve
arXiv cs.AI Research & Papers
Afterlife Delegation Protocol: Speculative Design of Self-Sovereign Agents that Outlive T…
arXiv cs.AI Opinion & Editorial
Position: AI Agents in Scientific Teams Should Be Studied as Human-Agent Systems
arXiv cs.AI Research & Papers
From Contexts to Values: Context-Dependent Defeat in Abstract Argumentation
arXiv cs.AI Research & Papers
Invariant Pretraining for Robust Code Representations
arXiv cs.AI Research & Papers
Path2ST: Hierarchical Cell-Tissue Grounded Cross-Modal Translation for Spatial Transcript…
arXiv cs.AI Research & Papers
ER-KANs: Efficient and Robust Kolmogorov-Arnold Networks for Data-Scarce Scientific Machi…
arXiv cs.AI Research & Papers
Beyond Visual CoT: Internalized Visual Thinking for Proactive Video Reasoning
arXiv cs.AI Research & Papers
KV-Rescue: Recovering Reasoning Language Model KV Eviction Loss via Stepwise Interleaving
arXiv cs.AI Research & Papers
The Authority Resolution Framework: A Five-Domain Ontology for Governing Who and What Dec…