Supervision versus Demonstration-Based In-Context Learning for Multiword Expression Class…
Explorar
Noticias de IA
22065 elementos — filtrados, clasificados y sin duplicados
PaperFlow: Profiling, Recommending, and Adapting Across Daily Paper Streams
TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Impr…
The Masked Advantage: Uncovering Local-Language Access to Cultural Knowledge in LLMs
OGA-AID: Clinician-in-the-loop AI Report Drafting Assistant for Multimodal Observational …
Evaluating AI-based Scientific Knowledge Synthesis with Epidemiological Systematic Reviews
VeriHGN: Heterogeneous Graph-Based Congestion Prediction for Chip Layout Verification
TRUE: A Trustworthy Unified Explanation Framework for Large Language Model Reasoning
ScenicRules: An Autonomous Driving Benchmark with Multi-Objective Specifications and Abst…
MACS: Modality-Aware Capacity Scaling for Efficient Multimodal MoE Inference
Automatic Causal Fairness Analysis with LLM-Generated Reporting
Self-Consistency from Only Two Samples: CoT-PoT Ensembling for Efficient LLM Reasoning
ViVa: A Video-Generative Value Model for Robot Reinforcement Learning
When is 3D Worth It? A Resource-Performance Frontier for CNNs and Transformers in Lung CT
DataEvolver: Automatic Data Preparation for Large Language Models through Multi-Level Sel…
Never Seen Before: Benchmarking Genuine Zero-Shot Composed Image Retrieval with Consisten…
TRACE: Trajectory Reasoning through Adaptive Cross-Step Evidence Aggregation for LLM Agen…
Coordinated optimization of departure sequencing and section-track allocation in railway …
FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantizat…
Multi-Scale Feature Attention Network for Polymer Classification using THz Dual-Comb Spec…
MacArena: Benchmarking Computer Use Agents on an Online macOS Environment
WAV: Multi-Resolution Block Residual Routing for Deep Decoder-Only Transformers
NTILC: Neural Tool Invocation via Learned Compression
ChinaHeritaQA: A Culturally-Grounded Visual Question Answering Dataset for World Heritage…
From Hazard Functions to Language Space: Cox-Supervised Distillation of Survival Risk int…
Order Matters: Unveiling the Hidden Impact of Macro Placement Sequences via Proxy-Guided …
Synthetic but Not Realistic: The Evaluation Challenge in Generative Modelling for Structu…
FAME: Forecastability-Aware Mixture of Experts for Heterogeneous Time Series Forecasting
Cheap Reward Hacking Detection
The Open Source Community is backing OpenEnv for Agentic RL