Test-Time Scaling for Scientific Equation Discovery
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation
A Large-scale Evaluation of Text-guided Models for Facial Editing
Toward Scalable Audio Description Quality Control: A Workflow for Evaluating Human and VL…
Can Large Language Models Identify Meaningful Touchpoints in Conversion Attribution?
Beyond Dense States: Sparse Transcoders as Causally Testable Operators for LLM Latent Rea…
Cross Lingual Transfer in Tulu Legal Comprehension: Script-Dependent Improvement and RAG-…
PromptKWS: A Novel Prompt-Guided Open-Vocabulary Keyword Spotting Framework
FAMPWQ: Fisher Information-based Adaptive Mixed Precision Weight Quantization for Effecti…
PAUSE: Editable Strategy Artifacts for Long-Form Cultural Story Adaptation
FigMirror: Ground It, Code It, Plot It
Intelligent Identification and Repair of Design Defects in BIM via Domain-Specific Large …
Asymmetric Within-Document Predictive Learning for Scientific Document Representation
The Race between Agentic AI Capabilities and Data Quality Control in Online Surveys
PUFFER: Incremental Fuzzy Deduplication for Continuously Evolving Corpora
Gurukul AI: An Interactive AI-Driven Educational Platform for Indian Education System
Parametric Multimodal User Memory: Storing What Captions Cannot Carry
NLP-Driven Knowledge Extraction and Thematic Classification of Translated Ancient Indian …
Looking Again: Measuring Sycophancy in the Reasoning Chains of Multimodal Models Under Pr…
MA-RAG: Multi-Agent Retrieval-Augmented Generation for Query-Driven Summarization of Long…
PowerSlider: Exploiting Phase Asymmetry for LLM Serving under Demand Response
OntoAligner-Ensemble: Voting-Based Fusion across Heterogeneous Ontology Alignment Techniq…
Interpretable Predictability-Based AI Text Detection: A Replication Study
MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative…
BLOOM-WILT: Logit Tilting for Behaviour Elicitation in Automated LLM Auditing
"Act Like a 5th Grader" is Not Enough: Bounding Knowledge in LLM-Based User Simulators
Let Prompts Bridge Defense Knowledge: Transferable Graph Purification via Vulnerability-A…
Token-Efficient Data Reasoning Agents via Adaptive Structuring of Unstructured Data
The Signal in the Noise: An Auditable Reliability Layer for Biomedical Text Classification
Measure Before You Manage: Evaluating Agent Working Memory in Coding Agents