MARLA: A Conceptual Scaffold for Regulatory Learning under the EU AI Act
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
SQL-Zero: Self-Evolving Text-to-SQL
Better Understanding, Better Fixes? A Study of Hallucination in LLM-based Automated Progr…
SciDocBench: A Workflow-Centered Benchmark and Data Pipeline for Scientific Document Unde…
Adaptive Multi-Granularity Temporal Modeling for Weakly Supervised Video Anomaly Detection
Compact Bellman-Grounded Cognitive Maps for Cost-Aware Navigation
The Mirror Agent Model: a Bayesian Architecture for Interpretable Agent Behavior
Comparables XAI: Faithful Example-based AI Explanations with Counterfactual Trace Adjustm…
TIER: Threat Implicitness Benchmark for Evaluating LLM Safety Behaviors
ARGOS: Who, Where, and When in Agentic Multi-Camera Person Search
ProCA: Progressive Contrastive Alignment for Robust EEG Visual Decoding
PerfReasoning: How Well Do LLMs Reason on Hardware Performance?
Large Language Models for HVAC Operations in Building Energy Systems: A Critical Review o…
GUT: Quantifying and Optimizing the Reasoning Uncertainty of LLMs via Graph Complexity
Beyond Aggregate Scores: Behavioral Correctness Assumptions for Assessing Reference-Based…
RISE: Recursive Improvement via Self-Extrapolating Policy Distillation
LLM-Driven Algorithm Design for Quantum Circuit Synthesis based on Binary Decision Diagra…
MePo++: Unifying Representation Refinement and Reconciliation for General Continual Learn…
TruthInsightBench: An Evidence-Grounded Benchmark for Automated Evaluation of Open-Ended …
Multi-Modal Time Series Prediction via Mixture of Modulated Experts
AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
Constructing and Evaluating Clinical Reasoning Trajectories for Medical Agent
Leveraging Low-Level Symbolic Competences for Unsupervised Grounding in Hallucination Det…
A Schema Bounded Language Model for Refining Robot Policies Without Destabilizing Local L…
Direction for Detection: A Survey of Automated Vulnerability Detection and all of its Pai…
Graph Foundation Models for Recommendation: A Comprehensive Survey
LLM-Guided Program Evolution for Circle Packing: Breaking 10 Packomania Records for $28
Octopus Protocol: One-Shot Hardware Discovery and Control for AI Agents via Infrastructur…
Technical Manual for a Toolkit for Measuring Contextual Individuation in Transformer Lang…
Towards a universal language of concepts: A survey