LLM4LLM: Bridging Kernel Benchmarks and Real Deployment via Closed-Loop Agentic Optimizat…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Learning with Boolean threshold functions
Physics-Knowledge-Guided Hybrid Neural Learning for Arctic Sea Ice Concentration Evolutio…
Consistency Is Not Coherence: Orientation Search for Certified Alignments Between 4D Defe…
The Retriever Should Remember: Experience-Amortized Reranking for Long-Term Agent Memory
Meta-Ctrl: Guaranteed Plan Generation by Decoupling Syntactic and Semantic Constraints
Correcting a learned physical invariant improves world-model rollouts
Reasoning Shortcuts and Value Symmetries: What Symmetry Permits, Architecture Realizes, a…
AI University: An LLM-Powered Learning Assistant for Engineering---A Finite Element Metho…
ParallelWorld: Test-Time Scaling for Embodied Reasoning
Redteaming Leading Arabic LLMs with ASAS
Hack-Verifiable Terminal Bench: Evaluating Reward Hacking in Terminal Tasks
ConvergeFlow: Language Flow with Provable Convergence to Token Embeddings
An Empirical Study on Preference Tuning Generalization and Diversity Under Domain Shift
ATHENA: Knowledge-guided agentic neural architecture search for AutoFormer-based electron…
ReasonEdit: Editing Vision-Language Models using Human Reasoning
Balancing Safety and Optimality in Robot Path Planning: Algorithm and Metric
Aligning Human Sense: Calibrated Distributional Reward Learning for Video Generation
Mamba-based Selective State Space Modeling Improves the Accuracy-Complexity Tradeoff of S…
Text-ADBench: Text Anomaly Detection Benchmark Based on LLM Embeddings
Safety Hacking in Constrained Best-of-$N$ Inference-time Scaling
KONTOGRAPH: Verified Point-in-Time Feature Consistency and Amortised Explanation for Real…
TO-Agents: A Multi-Agent AI Framework for Subjective Preference-Guided Topology Optimizat…
Reinforcing the World's Edge: A Continual Learning Problem in the Multi-Agent-World Bound…
Beyond Observed Auxiliary Relations: Environment-Conditioned Modeling for Multi-Behavior …
Agentic AI for Safety-critical Multi-drone Systems: Challenges and Opportunities
Power-Performance Characterization of TinyML Systems
MemGuard: Persisting Verifier Signals for LLM-Agent Memory Governance
SEAM: Shot Entity-Attribute Memory for Consistent Short-Drama Generation at Scale
CyrillicQA: The Influence of Phonetically Encoded Secret Language on LLM Performance