Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Auditable by Construction: An Ontology-Driven Framework for Trustworthy LLM Analytics in …
arXiv cs.AI Research & Papers
Beyond Effectiveness: A Multi-Criteria Framework for Comparing Practical Socio-Technical …
arXiv cs.AI Research & Papers
Investigating Target Class Influence on Neural Network Compressibility for Energy-Autonom…
arXiv cs.AI Research & Papers
Foundation Models for Partial Causal Identification
arXiv cs.AI Research & Papers
Dual-Cache Latent Space Communication between Heterogeneous Language Models
arXiv cs.AI Research & Papers
Evaluating Skills, Not Just Agents: Agentic Continuous Evaluation of Skills
arXiv cs.AI Research & Papers
FL-MAESTRO: Multi-Agent LLM Orchestration for Resource-Constrained Federated Learning
arXiv cs.AI Research & Papers
Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol
arXiv cs.AI Research & Papers
Volumetric Radiology AI in the Era of Multimodal Large Language Models
arXiv cs.AI Research & Papers
FlavourBench: Ranking Frontier Language Models with Executable Culinary Ground Truth
arXiv cs.AI Research & Papers
ACQ: A Deployed Two-Stage Framework for Automated Creative Quota Allocation in Large-Scal…
arXiv cs.AI Research & Papers
Difficulty-Aware Semantic-ID Optimization for Generative Recommendation
arXiv cs.AI Research & Papers
BC-Bench: Evaluating Agentic Engineering in a Domain-Specific Language for ERP
arXiv cs.AI Research & Papers
Curriculum-Aware Interpolate-then-Refine: Learned Physiological Time-Series Imputation un…
arXiv cs.AI Research & Papers
An Automated Pipeline for Few-Shot Bird Call Classification: A Case Study with the Tooth-…
arXiv cs.AI Research & Papers
Specification Portability Across LLM Development Agents: Cross-Agent Compatibility in Spe…
arXiv cs.AI Research & Papers
Open-Weight Masked Introspection: Measuring What Language Models Can Report About Their O…
arXiv cs.AI Research & Papers
Efficient Exploration at Scale
arXiv cs.AI Research & Papers
Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlene…
arXiv cs.AI Research & Papers
Lost in Translation: How Universal Ethical Values Fail to Translate Across Global Contexts
arXiv cs.AI Research & Papers
StateSight: Benchmarking Latent Spatial-State Reconstruction in Vision-Language Models
arXiv cs.AI Research & Papers
SCOPE: A Generative Approach for LLM Prompt Compression
arXiv cs.AI Research & Papers
Significant Other AI: Identity, Memory, and Emotional Regulation as Long-Term Relational …
arXiv cs.AI Research & Papers
GeoExplain: Multimodal Reasoning based on Hierarchy of Visual Information in Street View
arXiv cs.AI Research & Papers
MeltwaterBench: Deep learning for spatiotemporal downscaling of surface meltwater
arXiv cs.AI Research & Papers
WeedNet: A Foundation Model-Based Global-to-Local AI Approach for Real-Time Weed Species …
arXiv cs.AI Research & Papers
WorldLines: Benchmarking and Modeling Long-Horizon Stateful Embodied Agents
arXiv cs.AI Research & Papers
Neural-Primitive: An Efficient End-to-end Local Planner with Primitive-based Imitation Le…
arXiv cs.AI Research & Papers
Mind the Style: Impact of Communication Style on Human-Chatbot Interaction
arXiv cs.AI Research & Papers
Interpretable Multimodal Classification with Linear Discriminant Tree Ensembles