Explorar

Noticias de IA

21270 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Governed Capability Evolution: Lifecycle-Time Compatibility Checking and Rollback for AI-…
arXiv cs.AI Research & Papers
Qiskit QuantumKatas: Adapting Microsoft's Quantum Computing exercises for LLM evaluation
arXiv cs.AI Research & Papers
Governed Metaprogramming for Intelligent Systems: Reclassifying Eval as a Governed Effect
arXiv cs.AI Research & Papers
Post-training makes large language models less human-like
arXiv cs.AI Research & Papers
JobBench: Aligning Agent Work With Human Will
arXiv cs.AI Research & Papers
FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
arXiv cs.AI Research & Papers
ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence
arXiv cs.AI Research & Papers
Reasoning, Code, or Both? How Large Language Models Handle Variations in Math Questions
arXiv cs.AI Research & Papers
Governed Evolution of Agent Runtimes through Executable Operational Cognition
arXiv cs.AI Research & Papers
Which Changes Matter? Towards Trustworthy Legal AI via Relevance-Sensitive Evaluation and…
arXiv cs.AI Research & Papers
When Eyes Betray AI: Social Gaze Consistency as a Semantic Cue for AI-Generated Image Det…
arXiv cs.AI Research & Papers
AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compressio…
arXiv cs.AI Research & Papers
MemFail: Stress-Testing Failure Modes of LLM Memory Systems
arXiv cs.AI Research & Papers
GENESIS: Harnessing AI Agents for Autonomous 6G RAN Synthesis, Research, and Testing
arXiv cs.AI Research & Papers
How to Square Tensor Networks and Circuits Without Squaring Them
arXiv cs.AI Research & Papers
A Dataset of Robot-Patient and Doctor-Patient Medical Dialogues for Spoken Language Proce…
arXiv cs.AI Research & Papers
What Makes Chain-of-Thought Work at Probe Time? Local Co-occurrence Rather Than Global De…
arXiv cs.AI Research & Papers
Multi-Stakeholder LLM Alignment: Decomposing Estimation from Aggregation
arXiv cs.AI Research & Papers
TADDLE: A Tool-Augmented Agent for Detecting Deficient LLM-Generated Peer Reviews
arXiv cs.AI Research & Papers
LELA: An End-to-end LLM-based Entity Linking Framework with Zero-shot Domain Adaptation
arXiv cs.AI Research & Papers
AI-Driven Contribution Evaluation and Conflict Resolution: A Framework & Design for Group…
arXiv cs.AI Research & Papers
PaTAS: A Framework for Trust Propagation in Neural Networks Using Subjective Logic
arXiv cs.AI Research & Papers
BatteryMFormer: Multi-level Learning for Battery Degradation Trajectory Forecasting
arXiv cs.AI Research & Papers
Can Broad Biomedical Knowledge be Contextualized into Scenario-Grounded Propositions?
arXiv cs.AI Research & Papers
A Sharper Picture of Generalization in Transformers
arXiv cs.AI Research & Papers
Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving
arXiv cs.AI Research & Papers
CFG-OEC: Classifier Free Guidance with Orthogonal Error Correction
arXiv cs.AI Research & Papers
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic
arXiv cs.AI Research & Papers
The Compressive Knowledge Graph Hypothesis: Which Graph Facts Matter for Scientific Hypot…
arXiv cs.AI Research & Papers
Maat: The Agentic Legal Research Assistant for Competition Protection