Reward-Conditioned Attention: How Reward Design Shapes What Autonomous Driving Agents See
Explorar
Noticias de IA
30685 elementos — filtrados, clasificados y sin duplicados
PVF:Understanding AI Vulnerability Against SDCs
An Approach for a Supporting Multi-LLM System for Automated Certification Based on the Ge…
TokenMinds: Pretrained User Tokens and Embeddings for User Understanding in Large Recomme…
Safe Learning Control with Optimality and Stability Guarantees
Silent Failures in Physics-Informed Neural Networks: Parameter Poisoning and the Limits o…
ATMA: Length-Invariant Language Modeling via Polar Attention and Gated-Delta Compression …
Do vision-language models search like humans? Reasoning tokens as a reaction-time analog …
Phoneme-Level Mispronunciation Screening in Polish-Speaking Children with an Explainable …
ACT-JEPA: Novel Joint-Embedding Predictive Architecture for Efficient Policy Representati…
OmegAMP: Targeted AMP Discovery via Biologically Informed Generation
Towards Understanding The Calibration Benefits of Sharpness-Aware Minimization
HiT-JEPA: A Hierarchical Self-supervised Trajectory Embedding Framework for Similarity Co…
Position: Reasoning After Perception Means Reasoning Without Vision
Steering Vision-Language Models with Joint Sparse Autoencoders
Bias Fitting to Mitigate Length Bias of Reward Model in RLHF
Agentic Software Engineering: Foundational Pillars and a Research Roadmap
Discovering New Theorems via LLMs with In-Context Proof Learning in Lean
Reinforcement Learning Improves Traversal of Parametric Knowledge in LLMs
Introduction to Automated Negotiation
Weight Space Representation Learning via Neural Field Adaptation
Auto-exploration for online reinforcement learning
CustomX: Unified Character, Action, and Scene Customization in Video World Models
A Marketplace for AI-Generated Adult Content and Deepfakes
Membox: Weaving Topic Continuity into Long-Range Memory for LLM Agents
TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control
Streaming-dLLM: Accelerating Diffusion LLMs via Suffix Pruning and Dynamic Decoding
Latent Space Analysis for Interpretable Uncertainty in Melanoma Classification
Verifiable Manifest Signing and Transparency Enforcement for Secure MCP-Based LLM Pipelin…
ReaDy-Go: Real-to-Sim Dynamic 3D Gaussian Splatting Simulation for Environment-Specific V…