Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Automated Construction of FAIR Digital Object Knowledge Graphs from Flat Cultural Heritag…
arXiv cs.AI Research & Papers
Is Next-Chunk Reasoning RL Really Better than SFT? Revisiting Training Strategies under n…
arXiv cs.AI Research & Papers
Why Does Robustness Reduce Superposition?
arXiv cs.AI Research & Papers
What LLMs explain is not what they believe: Evaluating explanation sufficiency under mode…
arXiv cs.AI Research & Papers
RIACT: A Responsible AI System for Personalized Study Habit Tracking and Early Burnout Si…
arXiv cs.AI Research & Papers
CONTRAMEM: Learning Self-Evolving Procedural Memory from Contrasting Multi-Model Trajecto…
arXiv cs.AI Research & Papers
Towards Comprehensive Basketball Understanding
arXiv cs.AI Research & Papers
CausalCache: Conditional High-Fidelity Restoration for Long-Horizon GUI Agents
arXiv cs.AI Research & Papers
What's the Catch? Evaluating Temporal Consistency in Vision-Language Models
arXiv cs.AI Research & Papers
Procedural Knowledge Extraction from Industrial Troubleshooting Guides Using Vision Langu…
arXiv cs.AI Research & Papers
From Solver Feedback to Faithful Plans: Multi-Role Reinforcement Learning for Symbolic Pl…
arXiv cs.AI Research & Papers
Beyond RGB: Benchmarking and Enhancing MLLMs for Hyperspectral Image Understanding via Tr…
arXiv cs.AI Research & Papers
Machine Learning Assisted Inverse Design of Pixelated mmWave Patch Antennas
arXiv cs.AI Research & Papers
ClawProBench: Trace-Aware Evaluation of AI Agents with Runtime Coverage and Frozen Workpl…
arXiv cs.AI Research & Papers
From SQL Generation to Tool Selection: A Domain-Oriented Pattern for MCP Servers
arXiv cs.AI Research & Papers
Artificial Empathy: Towards a Framework for Unsupervised Agency Detection and Policy Reco…
arXiv cs.AI Research & Papers
Runtime Action Interference for AI Control of AlphaStar in StarCraft II
arXiv cs.AI Research & Papers
AdaR: A Framework for Equipping LLMs with Adaptive Reasoning
arXiv cs.AI Research & Papers
LitReview Arena: Evaluating Literature Review Agents with Battle-Style Peer Review Platfo…
arXiv cs.AI Research & Papers
ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Ga…
arXiv cs.AI Research & Papers
Deep-Learning-Based Pixelated Microwave Filter Design and Characterization using Electro-…
arXiv cs.AI Research & Papers
NoTB: Oracle-Free Triage of LLM-Generated RTL via Cross-Model Formal Consensus
arXiv cs.AI Research & Papers
Concepts for Securing Agentic AI Coding and the Terok Environment
arXiv cs.AI Security & Safety
PsychJail: Exploring Psychological Jailbreaks via Multi-Turn Persuasion of LLM Policies
arXiv cs.AI Research & Papers
MACD: Multi-Agent Clinical Diagnosis with Self-Learned Knowledge for LLM
arXiv cs.AI Research & Papers
GIM: Evaluating models via tasks that integrate multiple cognitive domains
arXiv cs.AI Research & Papers
Learning from the Test: Self-Referential Differential Testing for Deep RL Agents
arXiv cs.AI Research & Papers
Proxy reliance in large language model decisions is uncalibrated to predictive evidence
arXiv cs.AI Research & Papers
CDEG: Learning Decision-Critical Evidence for Long-Horizon Diagnostic Agents
arXiv cs.AI Research & Papers
On the Role of Citations in Preference Data