Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games
Explorar
Noticias de IA
30313 elementos — filtrados, clasificados y sin duplicados
From Brewing to Resolution: Tracing the Internal Lifecycle of Code Reasoning in LLMs
Beyond Domains: Reusing Web Skills via Transferable Interaction Patterns
First Proof Second Batch
IsabeLLM: Automated Theorem Proving Applied to Formally Verifying Consensus
EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navig…
Fixed-Point Reasoners: Stable and Adaptive Deep Looped Transformers
Shattering the Autoregressive Curse: Dynamic Epistemic Entropy Orchestrated Erasable Rein…
The Stanford EDGAR Filings Dataset: Reconstructing U.S. Corporate and Financial Disclosur…
DRFLOW: A Deep Research Benchmark for Personalized Workflow Prediction
Enhanced Evolutionary Multi-Objective Deep Reinforcement Learning for Reliable and Effici…
When AI Says "I have been in similar situations": Synthetic Lived Experience in Peer-Like…
PIVOT: Bridging Black-Scholes Implied-Volatility and Price Objectives via Differentiable …
Towards Distributed Inference of LLMs on a P2P Network
Curiosity-Critic: Cumulative Prediction Error Improvement as a Tractable Intrinsic Reward…
Towards Understanding and Measuring COGNITIVE ATROPHY in LLM Behaviour
OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipul…
Explicit Context-Driven Neural Acoustic Modeling for High-Fidelity RIR Generation
High-Fidelity 3D Geometric Reconstruction of Pelvic Organs from MRI: A Hybrid Deep Learni…
Looped World Models
The Measurement Gap in the Automation of EU Law: Benchmarking Doctrinal Legal Reasoning u…
RLRC: Reinforcement Learning-based Recovery for Compressed Vision-Language-Action Models
Non-negative Elastic Net Decoding for Information Retrieval
Surrogate Assisted Pedestrian Protection Design via a Foundation Model Orchestrated Workf…
Can LLMs Be CEOs? Benchmarking Strategic Resource Reallocation with Multi-Role Agent Simu…
Dissecting model behavior through agent trajectories
LLM-as-Judge in Education: A Curriculum-Grounded Marking Pipeline
EmoFSM: A Finite State Machine for Emotional Support Conversation
Top-Theta Attention: Sparsifying Transformers by Compensated Thresholding
Agentic AI-based Framework for Mitigating Premature Diagnostic Handoff and Silent Halluci…