Explorar

Noticias de IA

21270 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
How Chain-of-Thought Works? Tracing Information Flow from Decoding, Projection, and Activ…
arXiv cs.AI Research & Papers
LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop P…
arXiv cs.AI Research & Papers
Governed Capability Evolution: Lifecycle-Time Compatibility Checking and Rollback for AI-…
arXiv cs.AI Research & Papers
GENESIS: Harnessing AI Agents for Autonomous 6G RAN Synthesis, Research, and Testing
arXiv cs.AI Research & Papers
When Eyes Betray AI: Social Gaze Consistency as a Semantic Cue for AI-Generated Image Det…
arXiv cs.AI Research & Papers
The ATOM Report: Measuring the Open Language Model Ecosystem
arXiv cs.AI Research & Papers
Risk Averse Alert Prioritization for IDS Using Subnormal Gaussian Fuzzy Models
arXiv cs.AI Research & Papers
Lost in Sampling: Assessing Lexical Reachability in LLMs via the Word Coverage Score (WCS)
arXiv cs.AI Research & Papers
Governed Evolution of Agent Runtimes through Executable Operational Cognition
arXiv cs.AI Research & Papers
Many Logics, One Methodology: A Plea for Logical Pluralism in Formalised Reasoning (prepr…
arXiv cs.AI Research & Papers
FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
arXiv cs.AI Research & Papers
Chat2Workflow: A Benchmark for Generating Executable Visual Workflows with Natural Langua…
arXiv cs.AI Research & Papers
Algebraic Semantics of Governed Execution: Monoidal Categories, Effect Algebras, and Cote…
arXiv cs.AI Research & Papers
Mechanized Foundations of Structural Governance: Machine-Checked Proofs for Governed Inte…
arXiv cs.AI Research & Papers
Where Hindsight Credit Can Reside: A Signed-Capacity View of Token Updates in RLVR
arXiv cs.AI Research & Papers
Mind the Tool Failures: Achieving Synergistic Tool Gains for Medical Agents
arXiv cs.AI Research & Papers
OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling
arXiv cs.AI Research & Papers
Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation
arXiv cs.AI Research & Papers
BrainDINO: A Brain MRI Foundation Model for Generalizable Clinical Representation Learning
arXiv cs.AI Research & Papers
When VLMs 'Fix' Students: Identifying and Penalizing Over-Correction in the Evaluation of…
arXiv cs.AI Research & Papers
Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Dat…
arXiv cs.AI Research & Papers
Bridging Classification and Reconstruction: Cooperative Time Series Anomaly Detection
arXiv cs.AI Research & Papers
From Detection to Recovery: Operational Analysis on LLM Pre-training with 504 GPUs
arXiv cs.AI Research & Papers
VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning
arXiv cs.AI Research & Papers
FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentat…
arXiv cs.AI Research & Papers
Yes, Q-learning Helps Offline In-Context RL
arXiv cs.AI Research & Papers
Self-Cascaded Diffusion Models for Arbitrary-Scale Image Super-Resolution
arXiv cs.AI Research & Papers
Generative Animations: A Multi-Model Pipeline for Prompt-Driven Motion Synthesis
arXiv cs.AI Research & Papers
Tool Calling is Linearly Readable and Steerable in Language Models
arXiv cs.AI Research & Papers
Beyond Questions: Evaluating What Large Language Models (Actually) Know