Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Toward Fine-Grained Forgetting:Attribute Unlearning for Multimodal Large Language Models
arXiv cs.AI Research & Papers
From AI Technical Debt to Agentic Technical Debt: A Systematic Mapping of Root Causes and…
arXiv cs.AI Research & Papers
Judging Is Not Enumerating: Silent Omissions in LLM-Authored Acceptable Sets
arXiv cs.AI Research & Papers
Auditing Discovery Claims: A Two-Sided Criterion for Agentic Science, with the Negative S…
arXiv cs.AI Research & Papers
Passing Coarse Marginal Checks Can Be Cheap: Persona Mixtures and Imprecise Treatment-Res…
arXiv cs.AI Research & Papers
PROGRESS: Coverage-guided RL to Train Search-augmented LLM Agent
arXiv cs.AI Research & Papers
TrajWiki: Source-Grounded Memory Trajectories for Long-Horizon Dialogue Agents
arXiv cs.AI Research & Papers
Modeling Social Dynamics with an LLM-Enabled Agent Based Network-Dynamic (LAND) Model
arXiv cs.AI Research & Papers
CADIR: A Cross-Backend Editable Intermediate Representation for Agentic CAD Generation
arXiv cs.AI Research & Papers
Neuro-Evolved Heuristics for Variable Gapped Common Subsequence Identification
arXiv cs.AI Research & Papers
Assuming You Knew: Fixing an Epistemic Semantics for Flow Policies Using Agentic AI
arXiv cs.AI Research & Papers
The Scaling Paradox in Human-AI Collaboration
arXiv cs.AI Research & Papers
AgentSLABench: Evaluating and Benchmarking Agentic Systems Under Resource Constraints
arXiv cs.AI Research & Papers
Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation
arXiv cs.AI Research & Papers
AI-Based Thesis Assessment: An Empirical Study of Human Evaluation Priorities and Their I…
arXiv cs.AI Research & Papers
Tracing the Cascade: A Topology-Aware Evaluation Framework for Scientific Agent Hallucina…
arXiv cs.AI Research & Papers
Multi-Dimensional Assessment for AI Cognition (MAAC): A Theoretical Framework for Process…
arXiv cs.AI Research & Papers
HetGPS: Scalable Graph Multi-Agent Reinforcement Learning with Physics-Anchored Adaptive …
arXiv cs.AI Research & Papers
DASH: Decoupled Adaptive Surrogate - Acquisition Harness for Automated Bayesian Optimizat…
arXiv cs.AI Research & Papers
Slides2MindMap: Reconstructing Cognitively Efficient Knowledge Hierarchies from Lecture S…
arXiv cs.AI Research & Papers
Why Does the Future Branch? Identifiable Closure Tests for Stochastic Physical World Mode…
arXiv cs.AI Research & Papers
Through the LENS: Local Geometric Decomposition of Vision-Language Model Representations
arXiv cs.AI Research & Papers
CURE: Local Uncertainty Repair for Block-Parallel Speculative Decoding
arXiv cs.AI Research & Papers
BayesSeg: A Bayesian Optimization Framework for State Segmentation of Electricity Consump…
arXiv cs.AI Research & Papers
The Bayesian Reflex: A Predictive Coding Engine for Artificial Intelligence
arXiv cs.AI Research & Papers
F-WANDA: Fisher-Reweighted Post-Training Pruning for Sustainable Deployment of Large Lang…
arXiv cs.AI Research & Papers
Diagnose Before You Compress: Prediction-Independent Bottleneck Witness Refinement for LL…
arXiv cs.AI Research & Papers
TrAC: Trace-Conditioned Answer Consistency for Efficient Uncertainty Quantification in LL…
arXiv cs.AI Research & Papers
SymboUQ: Symbolic Uncertainty Quantification for Spatial Reasoning in LLMs
arXiv cs.AI Research & Papers
Where did the ambiguity go? Examining how multimodal models interpret polysemous words