Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
IndexTTS 2.5 Technical Report
arXiv cs.AI Research & Papers
ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Cou…
arXiv cs.AI Research & Papers
Structure-Enhanced Features and Quality-Aware Dynamic Anchor Scoring for Robust Lane Dete…
arXiv cs.AI Research & Papers
Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression
arXiv cs.AI Research & Papers
Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset…
arXiv cs.AI Research & Papers
RangeFactory: Scalable Construction of Multi-Hop Cyber Ranges
arXiv cs.AI Research & Papers
Not All Visual Tokens Are Equally Safe to Remove:Consequence-Sensitive Visual Token Compr…
arXiv cs.AI Research & Papers
Can Open-Weight Models Compete on Financial Text Comprehension?
arXiv cs.AI Research & Papers
Distributed Optimization with Streaming Data: A Temporal Weighting Perspective
arXiv cs.AI Research & Papers
Decoding Phenotypes: A Framework for Fusing Genomic Language Models and Neuroimaging
arXiv cs.AI Research & Papers
SemPIC: Learning Semantic Position-Independent KV Caches
arXiv cs.AI Research & Papers
Signature-Guided Capacity Occupancy for Dense Expert Merging
arXiv cs.AI Research & Papers
Full-bandwidth transformer
arXiv cs.AI Research & Papers
AquiLLM: An Architecture for Supporting Tacit Knowledge Capture in Research Groups
arXiv cs.AI Research & Papers
Can LLMs Rank? A Tale of Triads and Triage
arXiv cs.AI Research & Papers
SafeSceneReason: A Multimodal Reasoning Benchmark Connecting Industrial Hazards with Acci…
arXiv cs.AI Research & Papers
VCU-Bridge: Hierarchical Visual Connotation Understanding via Semantic Bridging
arXiv cs.AI Research & Papers
LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns
arXiv cs.AI Research & Papers
CRUISE: Vision-Language Model-Guided Uncertainty-Aware Cross-Modal Sensor Fusion for Robu…
arXiv cs.AI Research & Papers
SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery
arXiv cs.AI Research & Papers
From Manuals to Maintenance: Fine-Tuning MedGemma for Multi-Modal Imaging System Support …
arXiv cs.AI Research & Papers
TeXFix-Bench: An Empirically Grounded Multi-Format Benchmark for LLM-Based Document Sourc…
arXiv cs.AI Research & Papers
RAG-3DSG: Enhancing 3D Scene Graphs with Re-Shot Guided Retrieval-Augmented Generation
arXiv cs.AI Research & Papers
TGIF: Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs
arXiv cs.AI Research & Papers
Protecting patient privacy in clinical foundation models: Technical and legal perspectives
arXiv cs.AI Research & Papers
Intelligence Foundation Model: A New Perspective to Approach Artificial General Intellige…
arXiv cs.AI Research & Papers
The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI T…
arXiv cs.AI Research & Papers
RareLens: Towards End-to-End Rare Disease Care via Aligning Divergent Large Language Mode…
arXiv cs.AI Research & Papers
Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation
arXiv cs.AI Research & Papers
Rethinking Factor Sharing in Federated LoRA: A Rank-Aware Adaptive Approach