Credit Assignment with Resets in Language Model Reasoning
Explorar
Noticias de IA
29675 elementos — filtrados, clasificados y sin duplicados
From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language M…
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
Spectral Retrieval: Multi-Scale Sinc Convolution over Token Embeddings for Localized Retr…
Who judges the judges? Governance from metrics: a runtime framework for continuous LLM co…
The Path Matters: Learning a Token-Commitment Policy for Diffusion Language Models
VaaWIT: Visual-Aware Adaptation of Large Language Models for Multilingual Web Image Trans…
CyBOKClaw: Human-in-the-Loop CyBOK Mapping for Cybersecurity Curriculum
ADMFormer: An Adaptive-Decomposition Transformer with Time-Varying Masked Spatial Attenti…
Demystifying the Mythos or Disrupting Bugonomics? From Zero-Day Asymmetry to Defender Rem…
Detecting Unfaithful Chain-of-Thought via Circuit-Guided Internal-External Discrepancy
AgentHijack: Benchmarking Computer Use Agent Robustness to Common Environment Corruptions
Agent-Centric Social Trajectory Prediction: A Free Energy Principle Perspective
PILOT: Policy-Informed Learned Optimization for Adaptive Deep Network Training
SemanticZip: A Pilot Framework for Lossy Text Compression with LLMs as Semantic Decompres…
Is Decentralized AI Governable? From Regulative Policy to Constitutive Protocol
FoodMonitor: Benchmarking MLLMs for Explainable Compliance Analysis
MuCRASP: Multimodal Chain-of-thought Reasoning aware Structured Pruning
LECTOR: Joint Optimization of Scientific Reasoning Graphs and Introduction Generation
Neural Scalable Symbolic Search Framework for Complex Logical Queries with Multiple Free …
L2IR: Revealing Latent Intent in Graph Fraud Detection
TriVAL: A Tri-Validation Framework for Faithful Automatic Optimization Modeling
Raon-Speech Technical Report
Authority Signals in Claude AI Health Citations: A Descriptive Analysis Using the Authori…
Coarse-to-Fine Domain Incremental Learning with Attentive Distillation for Mining Footpri…
AI-Driven Controlled Environment Agriculture as Resilient Infrastructure for U.S. Fresh-P…
Code2UML: Agentic LLMs with context engineering for scalable software visualization
{\Phi}-Noise: Training-Free Temporal Video Conditioning via Phase-Based Noise Manipulation
Treatment Effect Estimation with Differentiated Networked Effect on Graph Data
A World Model of Radiologist Reading for Medical Image Representation Learning