Self-Evolving Cognitive Framework via Causal World Modeling for Embodied Scientific Intel…
Explorar
Noticias de IA
30588 elementos — filtrados, clasificados y sin duplicados
Efficient Multimodal Clinical Question Answering for Pulmonary Embolism Risk Assessment
Code Isn't Memory: A Structural Codebase Index Inside a Coding Agent
PlanBench-XL: Evaluating Long-Horizon Planning of LLM Tool-Use Agents in Large-Scale Tool…
Reference-Free Assessment of Physical Consistency in World Model-based Video Generation
Hypothesis-Driven Skill Optimization for LLM Agents
Nous: A Predictive World Model for Long-Term Agent Memory
Zhinong AI: A Design-Science Study of an AI-Enabled Agricultural Decision-Support Platfor…
Using Biometrics to Understand AI-Assisted Coding Performance and its Perception
AOHP: An Open-Source OS-Level Agent Harness for Personalized, Efficient and Secure Intera…
Cross-Architectural Mixture-of-Experts with Adaptive Soft Routing for Plant Leaf Disease …
Digital Humanism and Evolutionary Design
Abstract representational geometry supports inference in large language models
GIF: Locally Sound Geometric Information Flow Control for LLMs
HOLMES: Evaluating Higher-Order Logical Reasoning in LLMs
SPADE: Structure-Prior Adaptive Decision Estimation
Capable but Careless: Do Computer-Use Agents Follow Contextual Integrity?
DART: Draft-Agreement Routing for Training-Free Adaptive Thinking Budgets in Hybrid Reaso…
Decomposing Financial Market Dynamics via Mechanism Analysis in an Evolutionary Multi-Age…
Managing Procedural Memory in LLM Agents: Control, Adaptation, and Evaluation
TTFT-Aware Graph Chain-of-Thought:Distance-Indexed Neural A* for Low-Hallucination Multi-…
Cognitive Digital Twins: Ethical Risks and Governance for AI Systems That Model the Mind
Some Results about the Expressivity of Preference-Incomplete Structured Argumentation Fra…
IPO Finance Agent: Evaluation of LLM Financial Analysts beyond Finance Agent v2, with Aut…
From numerical proportions to analogical proportions between probabilities
A Stackelberg Framework for Resource-Aware LLM Agents: Learning, Repair, and Conditional …
When Preferences Fail to Become Incentives: A Utility-Behavior Gap in Large Language Mode…
The Impact of VAE Design on Latent Pose Representations for Diffusion-based Sign Language…
ENVS: Environment-Native Verified Search for Long-Horizon GUI Agents
When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents