Explorar

Noticias de IA

21271 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Simulator-Grounded Large Language Models for Industrial Causal Reasoning: Tool-Use, Struc…
arXiv cs.AI Research & Papers
Test-Time Scaling in Reasoning Models Is Not Effective for Knowledge-Intensive Tasks Yet
arXiv cs.AI Research & Papers
What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implication…
arXiv cs.AI Research & Papers
Does FLAIR super-resolution erase or hallucinate small white-matter lesions?
arXiv cs.AI Research & Papers
Measuring and Detecting Harmful AI Sycophancy
arXiv cs.AI Research & Papers
Stochasticity Is Not the Hard Part: Reduction and Complexity in Instructional Sequencing …
arXiv cs.AI Research & Papers
Bias Analysis of L2 Speaking Assessment Systems Using Concept Activation Vectors
arXiv cs.AI Research & Papers
When Agentic AI Meets Integrated Sensing and Communication
arXiv cs.AI Research & Papers
When Self-Evolution Backfires: Pre-Commit Gating against Skill Contamination in LLM Agents
arXiv cs.AI Research & Papers
Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Pers…
arXiv cs.AI Research & Papers
Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay
arXiv cs.AI Research & Papers
Hyper-ES: Effective Evolution Strategies for LLM Reasoning via Descent Direction Merging
arXiv cs.AI Research & Papers
BaKron: Efficient Quantization with Kronecker-Factored Hessians
arXiv cs.AI Research & Papers
Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for …
arXiv cs.AI Research & Papers
Unified Agent: Managing Interactions across Devices
arXiv cs.AI Research & Papers
FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial…
arXiv cs.AI Research & Papers
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Mode…
arXiv cs.AI Research & Papers
Gender-Based Heterogeneity in Youth Privacy-Protective Behavior for Smart Voice Assistant…
arXiv cs.AI Research & Papers
ProDVI: Programmatic Dynamics Priors for Value Network Initialization
arXiv cs.AI Research & Papers
SCP-NL2TL: Selective Conformal Prediction with Semantic Verification for Natural Language…
arXiv cs.AI Research & Papers
MM-ISTS: Cooperating Irregularly Sampled Time Series Forecasting with Multimodal Vision-T…
arXiv cs.AI Research & Papers
Stochastic Parrots or Singing in Harmony? Testing Five Leading LLMs for their Ability to …
arXiv cs.AI Research & Papers
ECG-LENS: Lead-Aware Clinical Context Enriched ECG Report Generation and Evaluation
arXiv cs.AI Research & Papers
TriQua: Reconciling Granularity and Context in Factuality Evaluation
arXiv cs.AI Research & Papers
From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models
arXiv cs.AI Research & Papers
Personalized Deep Research Query Refinement with Graph-Scaffolded Evidence Grounding
arXiv cs.AI Research & Papers
GSBF: Gaussian Splatting for Environment-Aware Beamforming
arXiv cs.AI Research & Papers
CourseGraph: Finding overlaps and differences in Computer Science courses across universi…
arXiv cs.AI Research & Papers
Grounded Well-Condition Anomaly Detection on the Volve Field: Constructed Labels, a Basel…
arXiv cs.AI Research & Papers
VLMs for Videogame Data Annotation