OncoTriad-QA: A Patient-Level Radiology-Pathology-Genomics Benchmark for Pan-Cancer Reaso…
Explorar
Noticias de IA
29627 elementos — filtrados, clasificados y sin duplicados
Reversing Arrows in Large Language Models
State Propagation Also Satisfies: A Complex-Valued State-Space Model for Deterministic St…
A neural operator framework for data-driven discovery of stability and receptivity in phy…
Large Language Models provide support for the parallelogram theory of analogy
The production of meaning in the processing of natural language
TransVLM: A Vision-Language Framework and Benchmark for Detecting Any Shot Transitions
A Deployment-Friendly Foundational Framework for Efficient Computational Pathology
Solver-Aware Decompositions for Programming-by-Example: When Dividing Requires Knowing ho…
Failure-Informed Image Self-Augmentation for Multimodal Large Language Model Self-Improve…
AutoSND: From Execution Evidence to Structural Policies for Automated Network Dismantling…
Leveraging System-Level Observations to Inform Bayesian Learning of Model Parameters for …
Unequal Verdicts: Investigating Gender Bias in LLM-Based Fake News Detection
Oilbird: Training-Free Speculative Decoding with Keys the Verifier Already Computes
Quantifying Hallucinations in Language Language Models on Medical Textbooks
Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling
Rethinking Modality Reliability in Multimodal Sentiment Analysis with Incomplete Observat…
Distilled Roads: Generalisable Road Network Extraction Across Sensors, Resolutions, and R…
CARE-Bench: Benchmarking Patient-Facing LLM Triage
Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation
ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning
ChiEngMixBench: Evaluating Large Language Models on Expert-Style Chinese-English Terminol…
DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning
AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detecti…
SKILL-KD: Contrastive Skill Distillation for LLM Agents
CT-HEG: A Bidirectional, Timestamp-Attributed Event Graph for ICU In-Hospital Mortality P…
AI Alignment and Fiduciary Obligation
PRIVEE: Privacy-Preserving Vertical Federated Learning Against Feature Inference Attacks
Verifier-Guided Model Discovery for Physical Dynamical Systems with Pretrained Symbolic T…
A Unified Framework for Human AI Collaboration in Security Operations Centers with Truste…