Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Distributed Optimization with Streaming Data: A Temporal Weighting Perspective
arXiv cs.AI Research & Papers
ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models
arXiv cs.AI Research & Papers
PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering
arXiv cs.AI Research & Papers
JaleesBench: Are AI Assistants Good Spiritual Company?
arXiv cs.AI Research & Papers
Multi-objective Evolutionary Merging Enables Efficient Reasoning Models
arXiv cs.AI Research & Papers
Fairness is Not Flat: Geometric Phase Transitions Against Shortcut Learning
arXiv cs.AI Research & Papers
Controllable Video Object Insertion via Multi-View Priors
arXiv cs.AI Research & Papers
Hybrid Policy Distillation for LLMs
arXiv cs.AI Research & Papers
From Local to Cluster: A Unified Framework for Causal Discovery with Latent Variables
arXiv cs.AI Research & Papers
Culturally Situated AI Safety for Youth: Saudi Arabian Perspectives of Youth, Parents and…
arXiv cs.AI Research & Papers
Path-Lock Expert: Separating Reasoning Mode in Hybrid Thinking via Architecture-Level Sep…
arXiv cs.AI Research & Papers
Controlled Memory Interference in Continual LLM Agents
arXiv cs.AI Research & Papers
From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Or…
arXiv cs.AI Research & Papers
Agent-MD: Selective LLM Intervention with Event-Driven Escalation for Stateful GCMC--MD C…
arXiv cs.AI Research & Papers
Contextual Value Alignment via Multilayer Combinatorial Fusion
arXiv cs.AI Research & Papers
An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Gl…
arXiv cs.AI Research & Papers
Towards Researcher Agents for Knowledge-Graph Question Answering
arXiv cs.AI Research & Papers
Protecting patient privacy in clinical foundation models: Technical and legal perspectives
arXiv cs.AI Research & Papers
Adaptive Two-Level Allocation of a Conserved Capacity Budget Across Locations and Service…
arXiv cs.AI Research & Papers
Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation
arXiv cs.AI Research & Papers
AndroidReality: How Far Are Mobile Agents from the Real World?
arXiv cs.AI Research & Papers
The Capability Ladder: A Curriculum-Modernization Framework for Workforce Readiness in th…
arXiv cs.AI Research & Papers
Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space
arXiv cs.AI Research & Papers
CausalNav: Reliability-Certified Causal World Models for Control under Physical-Parameter…
arXiv cs.AI Research & Papers
When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-…
arXiv cs.AI Research & Papers
Counterfactual Benchmarking and Training for Factuality Consistency and Order-Robust Grou…
arXiv cs.AI Research & Papers
Back to the Future: A workbook time machine for spread sheet creation benchmarks
arXiv cs.AI Research & Papers
GRACE: LLM-Grounded Semantic Metric Spaces for Scalable Mixed-Data Clustering
arXiv cs.AI Research & Papers
Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills
arXiv cs.AI Research & Papers
TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?