Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
VideoNorms: Benchmarking Cultural Awareness of Video Language Models
arXiv cs.AI Research & Papers
Partner Capability Estimation for Task-Agnostic Adaptation in Ad-Hoc Teamwork
arXiv cs.AI Research & Papers
A Methodology for Designing Knowledge-Driven Missions for Robots
arXiv cs.AI Research & Papers
Predict before you train: Scaling Laws for particle physics foundation models
arXiv cs.AI Research & Papers
A Graph-Native Bitemporal Memory Store for Conversational AI Agents
arXiv cs.AI Research & Papers
Large-Scale ChatBot Validation Through Customer Digital Twin Simulations
arXiv cs.AI Research & Papers
Do Methods Support the Claims? Intra-Paper Verification for Peer Review
arXiv cs.AI Research & Papers
Zero-Fi: Zero-Shot Wi-Fi-Based Human Activity Recognition via Contrastive Signal-Language…
arXiv cs.AI Research & Papers
Structurally Separated Uncertainty in Supervised Latent Variable Models
arXiv cs.AI Research & Papers
ForgetBench: Benchmarking Forgetting Dynamics of Long-Term Parametric Memory in Language …
arXiv cs.AI Research & Papers
SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search
arXiv cs.AI Research & Papers
IFCMemoryBench: Evaluating Long-Term Memory of LLM-Based Agents in BIM Information Retrie…
arXiv cs.AI Research & Papers
FinCacheServe: Dependency-Consistent Answer Reuse for Cost-Efficient RAG Serving over Mut…
arXiv cs.AI Research & Papers
A Reference-Free Score for Detecting Silent Reasoning Failures in Large Language Models
arXiv cs.AI Research & Papers
Audio-Anchored Fusion of Multi-Ratio DiT Reconstruction Residuals for Cross-Domain Audio …
arXiv cs.AI Security & Safety
FakeIDet3-DB: Refining Digital Attacks and Patch Extraction for Secure ID Benchmarking
arXiv cs.AI Security & Safety
Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak Defen…
arXiv cs.AI Research & Papers
Linguistic Monoculture in LLM-Assisted Language Use
arXiv cs.AI Research & Papers
AgentMap: Joint Equivalence and Subsumption Discovery for Ontology Matching
arXiv cs.AI Research & Papers
Setoka: A Benchmark for Hierarchical User Understanding in Personalized Agents over Heter…
arXiv cs.AI Research & Papers
Belief-Guided Decision Making with Uncertainty Gating in the Game of Go
arXiv cs.AI Research & Papers
Property-driven Causal Abstractions for Markov Decision Processes
arXiv cs.AI Research & Papers
UrbanDS: A Graph-Guided LLM Multi-Agent System for Data-Intensive Urban Tasks
arXiv cs.AI Research & Papers
AlphaSchema: Exploring the Space of Trading Semantics for LLM-Based Alpha Mining
arXiv cs.AI Research & Papers
Eco3S: Complex Socio-Economic System Simulation via Agent-Based Models
arXiv cs.AI Research & Papers
Shared Symbolic Backbones for Physically Consistent Multi-Output Symbolic Regression
arXiv cs.AI Research & Papers
Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical B…
arXiv cs.AI Research & Papers
Post-Training at the Edge of Detectability: A Game-Theoretic Approach to Fine-Tuning
arXiv cs.AI Research & Papers
Multi-Agent Debate Strategies: Survey, Taxonomy, and Challenges
arXiv cs.AI Research & Papers
Harnessing X-ray Absorption Spectroscopy Data through Multimodal Mining of Battery Litera…