Towards Understanding the Cognitive Habits of Large Reasoning Models
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Towards Embodied Cognition in Robots via Spatially Grounded Synthetic Worlds
A context-adaptive policy framework for robust and reactive robotic manipulation via unce…
Visual prompt engineering for video models
Less is More: Modality-Decoupling for General AIGC Audio-Video Detection
Beyond Self-Knowledge: Propagating Uncertainty Across Reasoning and Retrieval in LLMs
Neural Network Learning of One-Bit Protocols for Qubit Measurement Simulation
What Gets Lost When Memory Becomes Media? Evaluating AI-Generated Oral History Visualizat…
From Idea to Classroom in Days: Using "Vibe Coding" to Create a Programming Process Visua…
CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models
From Naive RAG to Deep Agentic Retrieval: An Evolving Context Engineering Pipeline for Re…
AI-Assisted Knowledge Access for Legacy Enterprise Asset Management in Energy Operations:…
Reading Without a Reader: Large Language Models Collapse Reading and Writing into a Singl…
When Thinking Before Retrieval Hurts: TraceBound Diagnostics for Adaptive Knowledge-Graph…
Diffusion Model-based Parameter Estimation in Dynamic Power Systems
Preliminary Guidelines for Using and Evaluating GenAI Tools to Support Systematic Literat…
GLIDE: Guided Layerwise Hybrid Attention for Efficient LLM Inference
RoCo-ACE: Rollout-Conditioned Online Distillation for Retention-Aware Knowledge Injection
PATHFinder Agent for Tailored Prenatal Care
Agent Skills Matter: Inferring Proprietary Skills from Execution Trajectories
From Training to Deployment: Post-Hoc Causal Feature Identification via Sensitivity Ratios
Are the High-weight Neurons the Important Ones in Image Classification Neural Networks?
HANDBOOK.md: A Benchmark for Long-Context Agentic Instruction Following
Penelope: Localized Latent Recurrence for Efficient Structured Reasoning
Cyber-Capable AI Agents: Vulnerabilities, Evaluation Containment, and Defensive Response
ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning
Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessme…
From Cellular Responses to Pharmacological Domains: Multimodal Zero-Shot Drug Representat…
Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inv…
Instruction-Tuned Language Models Cannot Sample from Distributions They Can Describe