Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Earth System World Model for What-If Simulations: A Case Study for Terrestrial Ecosystems
arXiv cs.AI Research & Papers
SQLMorph: Query Mutation and Fine-Grained Metrics for Text-to-SQL Evaluation
arXiv cs.AI Research & Papers
Personalizing LLM Agent Memory Using Biometrics
arXiv cs.AI Research & Papers
LEBGen: An LLM-Enhanced Bayesian Network Framework for Few-Shot Travel Survey Data Genera…
arXiv cs.AI Research & Papers
Feyospace-v1: How the Cyber Mercury Seven Trained Frontier Cyber Models
arXiv cs.AI Research & Papers
AgentGrad: Intervention-guided Prompt Optimization for Multi Agent Systems
arXiv cs.AI Research & Papers
When Can One Obtain Certificates of Optimality Using Positivstellensaetze?
arXiv cs.AI Research & Papers
Hi-FLoop: Hierarchical State-Feedback Loops for Multi-Timescale World Modeling
arXiv cs.AI Research & Papers
Optimal Experiments for Partial Causal Effect Identification
arXiv cs.AI Research & Papers
zScore-N: A Neural Network for On-Chain Wallet Reputation Scoring
arXiv cs.AI Research & Papers
Agentic ML Exploration (A-MLE) for Ads Ranking
arXiv cs.AI Research & Papers
CircuTutor: Transforming Static Circuit Problems into Intelligent and Dynamic Tutoring
arXiv cs.AI Research & Papers
Evidence-Aligned Entity Verification for Hallucination Detection in Retrieval-Augmented G…
arXiv cs.AI Research & Papers
Hyperparameter Scaling Laws Across MoE Sparsity
arXiv cs.AI Research & Papers
Qiushi Engine on AstaBench E2E-Bench-Hard
arXiv cs.AI Research & Papers
MemForest: Efficient Agent Memory Management via EventTree Partitioning and Progressive M…
arXiv cs.AI Research & Papers
SE-GoS: Self-Evolving Graph-of-Skills for Skill Library at Scale
arXiv cs.AI Research & Papers
Closing the Consistency Gap: Self-Evolving Agents That Learn to Stay on Course
arXiv cs.AI Research & Papers
Silent Revision: Measuring Undisclosed Change in the Safety Frameworks of Frontier AI Dev…
arXiv cs.AI Research & Papers
API Benchmark Scores Do Not Reliably Transfer to Chatbot Interfaces
arXiv cs.AI Research & Papers
WorldAgen: Unified State-Action Prediction with Test-Time World Model Training
arXiv cs.AI Research & Papers
Key Path Identification for Resolving Knowledge Conflicts via SAE-based Steering
arXiv cs.AI Research & Papers
Inference-Time Nash Alignment
arXiv cs.AI Research & Papers
RevalExo: A Functional Daily-Activity Benchmark for Inertial and Visual Locomotion Mode R…
arXiv cs.AI Research & Papers
OntologyBench: Can Dense Retrieval Satisfy Structured Biomedical Constraints?
arXiv cs.AI Research & Papers
SciLitBench: Benchmark and Design Principles for LLM-Powered Systematic Literature Reviews
arXiv cs.AI Research & Papers
Damage-Aware Bandit Pruning for Vision and Language Transformers
arXiv cs.AI Research & Papers
Compiling VGDL into Causal Models
arXiv cs.AI Research & Papers
Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collaps…
arXiv cs.AI Research & Papers
MoEMB: Scaling Universal Multimodal Embeddings with Efficient Mixture-of-Experts Models