Efficient Test-time Inference for Generative Planning Models
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment
AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Re…
An Abstract Worlds Semantic Framework for Belief Change Operators
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief
MOSAIC: Modular Orchestration for Structured Agentic Intelligence and Composition
LLM-Driven Co-Evolutionary Automated Heuristic Design for Bi-Component Coupled Combinator…
AI Sovereignty as National Learning Capacity: A Human-Centered Learning Mechanics Viewpoi…
Interaction-Centered Intelligence: Toward Interaction as the Primary Unit of Analysis in …
DarkVesselNet: Multi-Modal Remote Sensing and Trajectory Reasoning for Dark Vessel Detect…
Adversarial Feeds Steer LLM Agent Decisions Against Their Defaults
Decoupled Behavioral Cloning for Scalable Inductive Generalization in RL from Specificati…
Large Language Models in Transportation Systems Management and Operations: From Text Reas…
Towards Understanding Modality Interaction in Multimodal Language Models via Partial Info…
Can AI Review Improve Paper Drafting? An Empirical Study on 20 Computer Architecture Subm…
TriLens: Per-Layer Logit-Lens Entropy for White-Box Hallucination Detection
DAG-MoE: From Simple Mixture to Structural Aggregation in Mixture-of-Experts
Before the Model Learns the Bug:Fuzzing RLVR Verifiers
SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision
Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches
"Skill issues'': data-centric optimization of lakehouse agents
The Case for Model Science: Verify, Explore, Steer, Refine
The Shape of Wisdom: Decision Trajectories in Language Models
Brain-Atlas-Guided Generative Counterfactual Attention for Explainable Cognitive Decline …
SIRIUS-SQL: Anchoring Multi-Candidate Text-to-SQL in Execution Feedback
SkillSmith: Co-Evolving Skills and Tools for Self-Improving Agent Systems
Science Earth: Towards A Planet-Scale Operating System for AI-Native Scientific Discovery
FlowTime: Towards Continuous Generative Watch Time Prediction via Flow-based Personalized…
GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning
Don't Ask the LLM to Track Freshness: A Deterministic Recipe for Memory Conflict Resoluti…