DYNA : Dynamic Episodic Memory Networks for Augmenting Large Language Models with Tempora…
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Infant Spontaneous Movement Noise Improves Exploration in Deep RL
Proximal Policy Optimization for Amortized Discrete Sampling
SACE: Concept Erasure at the Semantic Singularity in Visual Autoregressive Models
Free Energy Heuristics: Fast-And-Frugal Cognition as Active Inference Under Uncertain Pre…
HoloRec: Holistic Encoding and Interleaved Reasoning for Generative Recommendation
ArtNet: A JEPA-Like Articulatory Predictive Framework for Robust Zero-Shot Phoneme Recogn…
Edu-Theater: A Data-Efficient Agent Framework for Scalable Learner Behavior Simulation th…
Enabling Real-Time Point-of-Care Ultrasound Segmentation: A GPU-Free Deployment in Resour…
MimicIK: Real-Time Generative Inverse Kinematics from Teleoperation with FK Consistency
PACUTE: Phonology-, Affix-, and Character-level Understanding of Tokens for Filipino
Beyond Scalar Distances: Semantic Attribute Gradients from Frozen MLLMs for Visual Embedd…
EyeMVP: OCT-Informed Fundus Representation Learning via Paired CFP--OCT Pretraining
Teacher-Student Structure for Domain Adaptation in Ensemble Audio-Visual Video Deepfake D…
Bridging Geographic Bias in Urban Streetscape Inference via Lifelong Learning with Visual…
PANDA: An LLM-Enhanced Performance-Driven Analog Design Framework Bridging Design Intent …
Multi-Modal Attention for Automated Disaster Damage Assessment Using Remote Sensing Image…
Separable Neural Architectures as Physical World Models: from Mathematical Theory to Appl…
An Ensemble Deep Learning Approach for Reliable and Scalable Lemon Leaf Disease Classific…
Knowledge-Based Zero-Replay Debugging of Multi-Agent LLM Traces
XFlow: An Executable Protocol Programming System for Reliable Multi-Agent Workflows
MatchLM2Lite: A Scalable MLLM-to-Lite Framework for Reproduced Content Identification
JoyAI-VL-Interaction: Real-Time Vision-Language Interaction Intelligence
ScoutVLA: UAV-Centric Active Perception via a Dual-Expert VLA Model for Open-World Embodi…
Where Does Texture Evidence Live in SAM? Features, Proposal Masks, and Texture Segmentati…
Sub-Semantic Image Segmentation
Pixel-TTS: Image based Text Rendering for Robust Text-to-Speech
Wasserstein Equilibrium Decoding for Reliable Medical Visual Question Answering
Improved Baselines with Representation Autoencoders
Decoupling Semantics from Distortions: Multi-Scale Two-Stream Vision-Language Alignment f…