Do Language Models Reason Across Languages?
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
From Location Phrases to Geographic Entities: Task-Adapted Retrieval for People Search
Enhancing Low-Resource Language Reasoning via High-Resource Language Feature Transfer
Error Detection for PET/CT Radiology Reports: Domain-Specific vs Large Language Models
Online Estimation of Dynamic Origin-Destination Matrices Using Reinforcement Learning wit…
Higher-Dimensional Rotary Position Embedding
A Composition-Aware Pretraining Framework for Geospatial Foundation Models
LCoT-GV: Graph Attention Networks for Verifying Long Reasoning Chains in Large Language M…
Graph4BiLO: Graph Neural Network Approximation for Bilevel Mixed-Integer Linear Optimizat…
Does Latent Planning Survive Point Clouds? Action-Conditioned JEPA World Models for Geome…
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
Repair or Resample? Rethinking Failure Debugging in LLM Multi-Agent Systems
LLMODE: Aligning ODEs with LLMs via Gated Token Injection for Irregular Spatio-Temporal F…
Adaptive Multi-Branching for Shallow Decision Tree Induction
Masked Distillation: Internalizing the Chain-of-Thought in Language Models
SUB-PLAY: Adversarial Policies against Partially Observed Multi-Agent Reinforcement Learn…
AgenticRag-R1: Agentic Reinforcement Learning with Stack Memory for Multi-Step Reasoning,…
QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation
FRAC-MAS: A Safe and Explainable Multi-Agent System for Fracture Diagnosis
Toward Scalable Audio Description Quality Control: A Workflow for Evaluating Human and VL…
Enhancing SAE-based Steering via Neighbor Integrated Feature Selection
Scoring, Reasoning, and Selecting the Best! Ensembling Large Language Models via a Peer-R…
Pro-Router: Token-Aware Progressive Model Routing with Adaptive Edge-Cloud Collaboration …
Capability-Stratified Degradation in Ternary Language Models
Standardizing Longitudinal Radiology Report Evaluation via Large Language Model Annotation
How Language Models Choose Sides: Internal Representations of Instruction Hierarchy
Conducting Stylistic Analysis of Paintings through an Art-History Agent
Rate-Coding Bundle Memory: A Unified Model of Memory and Control for Symbolic Computation…
When Do Larger Batches Help Scale LLM Reinforcement Learning?
AgentLogs: A Dataset for Opening the Black Box of GitHub's Cloud Agent