Reward-free Alignment for Conflicting Objectives
Browse
AI News
30934 items — filtered, classified, deduplicated
Agent Primitives: Reusable Latent Building Blocks for Multi-Agent Systems
BEAR: Towards Beam-Search-Aware Optimization for Recommendation with Large Language Models
RecGOAT: Graph Optimal Adaptive Transport for LLM-Enhanced Multimodal Recommendation with…
QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis
Future-KL Regularized GRPO: Process-Level Credit Assignment from $f$-Divergence Regulariz…
Pixelwise Uncertainty Quantification of Accelerated MRI Reconstruction
Extreme-value forest fire prediction A study of the Loss Function in an Ordinality Scheme
NeoAMT: Neologism-Aware Agentic Machine Translation with Reinforcement Learning
Routing by Analogy: kNN-Augmented Expert Assignment for Mixture-of-Experts
$M^3-Verse$: A "Spot the Difference" Challenge for Large Multimodal Models
Multimodal Functional Maximum Correlation for Emotion Recognition
Simply Stabilizing the Loop via Fully Looped Transformer
DynaPURLS: Dynamic Refinement of Part-Aware Representations for Skeleton-Based Zero-Shot …
Bridging the Semantic-Action Gap in Visual Token Pruning for Efficient VLA Inference
Understanding, Accelerating, and Improving MeanFlow Training
Membership Inference Attacks on Tokenizers of Large Language Models
INSIGHT: INference-time Sequence Introspection for Generating Help Triggers in Vision-Lan…
Equip Pre-ranking with Target Attention by Residual Quantization
What Happens Next? Anticipating Future Motion by Generating Point Trajectories
LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries
Metacognition Should Be the Scientific Framework for Bounded and Effective Self-Governanc…
Explainable Attention-Guided Stacked Graph Neural Networks for Malware Detection
Plan for Speed: Dilated Scheduling for Masked Diffusion Language Models
FLoRIST: Singular Value Thresholding for Efficient and Accurate Federated Fine-Tuning of …
SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models
MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email De…
PhySense: Sensor Placement Optimization for Accurate Physics Sensing
Pragmatic Reasoning improves LLM Code Generation
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model