Behavior and Representation in Open-Weight Large Language Models for Combinatorial Optimi…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Towards Human Motion World Models via Executable Behaviour Representations
Tools as Continuous Flow for Evolving Agentic Reasoning
ComBodied Agents: a New Paradigm of Human-Centric Agentic AI
Deep Activity Model: A Generative Approach for Human Mobility Pattern Synthesis
Explainability in Practice: A Survey of Explainable NLP Across Various Domains
COLORA: Efficient Fine-Tuning for Convolutional Models with a Study Case on Optical Coher…
Glance, Scrutinize, and Think: Advancing Video Anomaly Detection from Training-Free to Ag…
Adaptive Hybrid Particle Swarm Optimization with Gradient Descent
Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonom…
TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy F…
Detecting a Route Flip Is Easier Than Knowing Whether to Fix It: Causal Route-Mediated Da…
AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research
Harnessing agent memory to build lifelong AI partners for materials scientists
Towards Sustainable Learning in Online Education: A Reinforcement Learning Approach
Identity from the Outside: A Conceptual Framework and Research Program for AI Personality…
Cutting AI Datacenter Energy with Reinforcement Learning: Measured Power Control of LLM T…
Pretraining large language models with MXFP4 on Native FP4 Hardware
Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurement
Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing Long-Hori…
Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Tex…
From Numbers to Judgment: Specialist LLM Agents and Reinforcement Learning for European L…
From Prompting to Behavioral Alignment: Personalized LLM Judges for Recommendation Evalua…
When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problem…
Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diag…
A Modular Agentic Framework for Synthetically Constrained Multi-Objective Hit-to-Lead Opt…
Localizing Safety Alignment: MLP Layers and Mid-Network Blocks Encode Refusal Behavior in…
CoAdapt-GUI: Joint Workflow Context and Policy Adaptation for Unseen GUI Applications
Making AI-Generated Feedback Matter: From Provision to Student Enactment
MBA: Multimodal Benchmark and Agents for Real-World Business Ideation