HyperGuide: Hyperbolic Guidance for Efficient Multi-Step Reasoning in Large Language Mode…
Browse
AI News
30934 items — filtered, classified, deduplicated
Go witheFlow: Real-time Emotion Driven Audio Effects Modulation
ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference
SkillEvolBench: Benchmarking the Evolution from Episodic Experience to Procedural Skills
Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic U…
LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward …
BackWeak: Backdooring Knowledge Distillation Simply with Weak Triggers and Fine-tuning
Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform
E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomo…
LGMT: Logic-Grounded Metamorphic Testing for Evaluating the Reasoning Reliability of LLMs
Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs
LECTOR: Joint Optimization of Scientific Reasoning Graphs and Introduction Generation
JT-SAFE-V2: Safety-by-Design Foundation Model with World-Context Data
SPACE: Unifying Symmetric and Asymmetric Routing Problems for Generalist Neural Solver
AI-generated podcasts: Synthetic Intimacy and Cultural Mistranslation in NotebookLM's Aud…
A Two-Dimensional Framework for AI Agent Design Patterns: Cognitive Function and Executio…
AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning
The Many Faces of On-Policy Distillation: Pitfalls, Mechanisms, and Fixes
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents
Learning Preference-Based Objectives from Clinical Narratives for Dynamic Sepsis Treatment
OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search
FLINGO -- Instilling ASP Expressiveness into Linear Integer Constraints
Decoding ML Decision: An Agentic Reasoning Framework for Large-Scale Ranking System
Reasoning as an Attack Surface: Adaptive Evolutionary CoT Jailbreaks for LLMs
Architecting Agentic Communities using Design Patterns
Hypothesis Generation and Inductive Inference in Children and Language Models
Sensing Intelligence as a Trainable Metamaterial Property
IPR-1: Interactive Physical Reasoner
Rewarding Structural Conformance of Reasoning using Process Mining