Grammar-Aware Literate Generative Mathematical Programming with Compiler-in-the-Loop
Explorar
Noticias de IA
30334 elementos — filtrados, clasificados y sin duplicados
COMET: Concept Space Dissection of the Modality Gap in Audio-Text Multimodal Contrastive …
Architecture-Induced Recoverability Bias in Differentiable Symbolic Regression
SERC: LDPC-Inspired Semantic Error Correction for Retrieval-Augmented Generation
Thoughts-as-Planning: Latent World Models for Chain-of-Thoughts Optimization via Reinforc…
Honest Lying: Understanding Memory Confabulation in Reflexive Agents
KLAS: Using Similarity to Stitch Neural Networks for Improved Accuracy-Efficiency Tradeof…
Wait! There's a Way Out: A Decision Mechanism for Forecasting Conversational Derailment
Relevance as a Vulnerability: How Web Retrieval Degrades Safety Alignment in LLM Agents
GPF-LiveNews: A Streaming Evaluation Protocol for Group-Conditioned Framing in Large Lang…
Self-Play Reinforcement Learning under Imperfect Information in Big 2
KBF: Knowledge Boundary as Fingerprint for Language Model and Black-Box API Auditing
Emergent Semantic Representations in World Models through Physical Interaction without Li…
Balancing Multimodal Learning through Label Space Reshaping
Bridging the Sim-to-Real Gap in Reinforcement Learning-Based Industrial Dispatching throu…
TaxDistill: Improving Metagenomic Taxonomic Annotation via Distilled Genomic Foundation M…
Context Distillation as Latent Memory Management
Quantum-Enhanced Adversarial Robustness in Artificial Intelligence
Mind Your Tone: Does Tone Alter LLM Performance?
The Hamilton-Jacobi Theory of Deep Learning
HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous S…
When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis
VFEAgent: A Multimodal Agent Framework for End-to-End Automated Finite Element Analysis
Frontier LLM-based agents can overcome the ontology curation bottleneck for natural pheno…
Review Arcade: On the Human Alignment and Gameability of LLM Reviews
Learn from A Rationalist: Distilling Intermediate Interpretable Rationales
Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning
Behavior-Induced Mirror-Prox Temporal-Difference Learning for Faster Off-Policy Prediction
BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices
Loong: A Human-Like Long Document Translation Agent with Observe-and-Act Adaptive Context…