IndexTTS 2.5 Technical Report
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Cou…
Structure-Enhanced Features and Quality-Aware Dynamic Anchor Scoring for Robust Lane Dete…
Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression
Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset…
RangeFactory: Scalable Construction of Multi-Hop Cyber Ranges
Not All Visual Tokens Are Equally Safe to Remove:Consequence-Sensitive Visual Token Compr…
Can Open-Weight Models Compete on Financial Text Comprehension?
Distributed Optimization with Streaming Data: A Temporal Weighting Perspective
Decoding Phenotypes: A Framework for Fusing Genomic Language Models and Neuroimaging
SemPIC: Learning Semantic Position-Independent KV Caches
Signature-Guided Capacity Occupancy for Dense Expert Merging
Full-bandwidth transformer
AquiLLM: An Architecture for Supporting Tacit Knowledge Capture in Research Groups
Can LLMs Rank? A Tale of Triads and Triage
SafeSceneReason: A Multimodal Reasoning Benchmark Connecting Industrial Hazards with Acci…
VCU-Bridge: Hierarchical Visual Connotation Understanding via Semantic Bridging
LAUDE: LLM-Assisted Unit Test Generation and Debugging of Hardware DEsigns
CRUISE: Vision-Language Model-Guided Uncertainty-Aware Cross-Modal Sensor Fusion for Robu…
SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery
From Manuals to Maintenance: Fine-Tuning MedGemma for Multi-Modal Imaging System Support …
TeXFix-Bench: An Empirically Grounded Multi-Format Benchmark for LLM-Based Document Sourc…
RAG-3DSG: Enhancing 3D Scene Graphs with Re-Shot Guided Retrieval-Augmented Generation
TGIF: Text-Guided Layer Fusion Mitigates Hallucination in Multimodal LLMs
Protecting patient privacy in clinical foundation models: Technical and legal perspectives
Intelligence Foundation Model: A New Perspective to Approach Artificial General Intellige…
The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI T…
RareLens: Towards End-to-End Rare Disease Care via Aligning Divergent Large Language Mode…
Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation
Rethinking Factor Sharing in Federated LoRA: A Rank-Aware Adaptive Approach