Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Quipu: A Governed Bitemporal Knowledge Graph Store
arXiv cs.AI Research & Papers
Policy Iteration with Human Feedback: Bringing Post-Training RL to In-context Learning
arXiv cs.AI Research & Papers
What Do Compliance Detectors Read? An Audit of Activation Probes and Guard Models
arXiv cs.AI Research & Papers
A Temporal Reasoning Benchmarking Framework for LRMs via Difficulty-controlled and Dynami…
arXiv cs.AI Research & Papers
WARA: Toward Automated Wireless Optimization Research with Closed-Loop LLM Agents
arXiv cs.AI Research & Papers
From Reactive to Autonomous: Evolution of AI Operations in Cloud Network Infrastructure
arXiv cs.AI Research & Papers
Would this change your answer? Evaluating Explanations of LLM Behavior In The Wild with C…
arXiv cs.AI Research & Papers
TAHB: A Comprehensive Benchmark for Text-Attributed Hypergraph Learning
arXiv cs.AI Research & Papers
SCOPE: Score-Isolated Agentic Optimization for Video World Models
arXiv cs.AI Research & Papers
LLM-Based Hierarchical Coordinated Control with Continuation-Aware Policy Learning
arXiv cs.AI Research & Papers
SKILL: Self-correcting Knowledge-guided Iterative Large Language Model Agent for Logic Op…
arXiv cs.AI Research & Papers
ClawGym II: Exploring Black-Box RL on Agent Harness
arXiv cs.AI Research & Papers
Neurosymbolic Embodied Agents
arXiv cs.AI Research & Papers
Logical Embeddings for Argument Analysis
arXiv cs.AI Research & Papers
RETRACE: Resilience-Guided Trait-Conditioned Craving Estimation from Wearable Physiology …
arXiv cs.AI Research & Papers
CaliBench: Are the Stochastic Dynamics of Video World Models Physically Calibrated?
arXiv cs.AI Research & Papers
ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language…
Hugging Face Daily Papers Research & Papers
LLMs for Medical Consultation Are Evaluated Too Late: The Preformulation Gap
Hugging Face Daily Papers Research & Papers
SignalReasoner: Assessing the Upper Bound of 3B Models for Signal Mathematical Reasoning
Hugging Face Daily Papers Research & Papers
LiveHouse-TS: An Open-world Living Benchmark for Time Series Foundation Models
Hugging Face Daily Papers Research & Papers
Beyond MSE: Rethinking the Evaluation Metric and Benchmarking for Irregular Time Series F…
Hugging Face Daily Papers Research & Papers
Do LLMs Know a Good Hypothesis When They See One? Logit-Based Energy Scoring Outperforms …
Hugging Face Daily Papers Research & Papers
Information fusion and machine learning for sensitivity analysis using physics knowledge …
Hugging Face Daily Papers Research & Papers
Delta2Gamma: Band-Wise Adaptive Contrastive Learning of EEG for Alzheimer's Disease Detec…
Hugging Face Daily Papers Research & Papers
Probing Association Instability with Track-State Perturbations for Clip-Level Active Lear…
Hugging Face Blog Research & Papers
Multi-Vector (Late Interaction) Embedding Models with Sentence Transformers
Hugging Face Daily Papers Research & Papers
SCENARIODIFF: A Scenario-level Guidance Framework for Multimodal Time Series Forecasting-…
Hugging Face Daily Papers Research & Papers
Causal Discovery in Equal Variance Linear Gaussian DAGs via SURE-Tuned Ridge Regression
Hugging Face Daily Papers Research & Papers
Deep Learning for Cross-Border Electricity Price Forecasting: A Comparative Study
Ars Technica AI Research & Papers
Hidden Airtag reveals Amazon is trashing rare books to train AI