Distributed Optimization with Streaming Data: A Temporal Weighting Perspective
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models
PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering
JaleesBench: Are AI Assistants Good Spiritual Company?
Multi-objective Evolutionary Merging Enables Efficient Reasoning Models
Fairness is Not Flat: Geometric Phase Transitions Against Shortcut Learning
Controllable Video Object Insertion via Multi-View Priors
Hybrid Policy Distillation for LLMs
From Local to Cluster: A Unified Framework for Causal Discovery with Latent Variables
Culturally Situated AI Safety for Youth: Saudi Arabian Perspectives of Youth, Parents and…
Path-Lock Expert: Separating Reasoning Mode in Hybrid Thinking via Architecture-Level Sep…
Controlled Memory Interference in Continual LLM Agents
From Single Chatbots to Governed Agent Ecosystems: An Agentic AI Pattern Catalogue and Or…
Agent-MD: Selective LLM Intervention with Event-Driven Escalation for Stateful GCMC--MD C…
Contextual Value Alignment via Multilayer Combinatorial Fusion
An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Gl…
Towards Researcher Agents for Knowledge-Graph Question Answering
Protecting patient privacy in clinical foundation models: Technical and legal perspectives
Adaptive Two-Level Allocation of a Conserved Capacity Budget Across Locations and Service…
Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation
AndroidReality: How Far Are Mobile Agents from the Real World?
The Capability Ladder: A Curriculum-Modernization Framework for Workforce Readiness in th…
Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space
CausalNav: Reliability-Certified Causal World Models for Control under Physical-Parameter…
When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-…
Counterfactual Benchmarking and Training for Factuality Consistency and Order-Robust Grou…
Back to the Future: A workbook time machine for spread sheet creation benchmarks
GRACE: LLM-Grounded Semantic Metric Spaces for Scalable Mixed-Data Clustering
Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills
TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?