Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
MARLA: A Conceptual Scaffold for Regulatory Learning under the EU AI Act
arXiv cs.AI Research & Papers
SQL-Zero: Self-Evolving Text-to-SQL
arXiv cs.AI Research & Papers
Better Understanding, Better Fixes? A Study of Hallucination in LLM-based Automated Progr…
arXiv cs.AI Research & Papers
SciDocBench: A Workflow-Centered Benchmark and Data Pipeline for Scientific Document Unde…
arXiv cs.AI Research & Papers
Adaptive Multi-Granularity Temporal Modeling for Weakly Supervised Video Anomaly Detection
arXiv cs.AI Research & Papers
Compact Bellman-Grounded Cognitive Maps for Cost-Aware Navigation
arXiv cs.AI Research & Papers
The Mirror Agent Model: a Bayesian Architecture for Interpretable Agent Behavior
arXiv cs.AI Research & Papers
Comparables XAI: Faithful Example-based AI Explanations with Counterfactual Trace Adjustm…
arXiv cs.AI Research & Papers
TIER: Threat Implicitness Benchmark for Evaluating LLM Safety Behaviors
arXiv cs.AI Research & Papers
ARGOS: Who, Where, and When in Agentic Multi-Camera Person Search
arXiv cs.AI Research & Papers
ProCA: Progressive Contrastive Alignment for Robust EEG Visual Decoding
arXiv cs.AI Research & Papers
PerfReasoning: How Well Do LLMs Reason on Hardware Performance?
arXiv cs.AI Research & Papers
Large Language Models for HVAC Operations in Building Energy Systems: A Critical Review o…
arXiv cs.AI Research & Papers
GUT: Quantifying and Optimizing the Reasoning Uncertainty of LLMs via Graph Complexity
arXiv cs.AI Research & Papers
Beyond Aggregate Scores: Behavioral Correctness Assumptions for Assessing Reference-Based…
arXiv cs.AI Research & Papers
RISE: Recursive Improvement via Self-Extrapolating Policy Distillation
arXiv cs.AI Research & Papers
LLM-Driven Algorithm Design for Quantum Circuit Synthesis based on Binary Decision Diagra…
arXiv cs.AI Research & Papers
MePo++: Unifying Representation Refinement and Reconciliation for General Continual Learn…
arXiv cs.AI Research & Papers
TruthInsightBench: An Evidence-Grounded Benchmark for Automated Evaluation of Open-Ended …
arXiv cs.AI Research & Papers
Multi-Modal Time Series Prediction via Mixture of Modulated Experts
arXiv cs.AI Research & Papers
AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
arXiv cs.AI Research & Papers
Constructing and Evaluating Clinical Reasoning Trajectories for Medical Agent
arXiv cs.AI Research & Papers
Leveraging Low-Level Symbolic Competences for Unsupervised Grounding in Hallucination Det…
arXiv cs.AI Research & Papers
A Schema Bounded Language Model for Refining Robot Policies Without Destabilizing Local L…
arXiv cs.AI Research & Papers
Direction for Detection: A Survey of Automated Vulnerability Detection and all of its Pai…
arXiv cs.AI Research & Papers
Graph Foundation Models for Recommendation: A Comprehensive Survey
arXiv cs.AI Research & Papers
LLM-Guided Program Evolution for Circle Packing: Breaking 10 Packomania Records for $28
arXiv cs.AI Research & Papers
Octopus Protocol: One-Shot Hardware Discovery and Control for AI Agents via Infrastructur…
arXiv cs.AI Research & Papers
Technical Manual for a Toolkit for Measuring Contextual Individuation in Transformer Lang…
arXiv cs.AI Research & Papers
Towards a universal language of concepts: A survey