How memory can affect collective and cooperative behaviors in an LLM-Based Social Particl…
Explorar
Noticias de IA
21863 elementos — filtrados, clasificados y sin duplicados
AdaMARP: An Adaptive Multi-Agent Interaction Framework for General Immersive Role-Playing
EvoPINN: Agentic Discovery of Executable Algorithms for Physics-Informed Neural Networks
TraceCoder: Explainable and Auditable Code Generation with Position-Key Snippet Versioning
Exploring Structures in Physics Problems: Can AI Agents Discover Statistical Mechanical M…
CG-World: A Large-Scale World-State Dataset and Protocol for World Models
FilmBench: A Film-Grade Benchmark for Cinematic Video Generation
OmegaUse-OfficeVal: Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Econo…
Structurally Separated Uncertainty in Supervised Latent Variable Models
Where Is the Cost of Third-Party API Routers in Agentic Software Development?
Towards simultaneous decoding of kinetic and kinematic movement parameters during grasp a…
Compressed Video Aggregator: Content-driven Module for Efficient Micro-Video Recommendati…
Harnessing X-ray Absorption Spectroscopy Data through Multimodal Mining of Battery Litera…
SARC-DQ: Runtime Data-Quality Gating for Agentic AI: Silent Evidence Defects, the Incompe…
An Unofficial FastLAS Tutorial: A Programmer's Guide
Contextualized Counterspeech Can Be More Persuasive Than Generic Counterspeech
Train Large, Deploy Compact: Structured Compression for Compact Low-Rank Adaptation
Plans Work in Mysterious Ways: Evaluating a Plan Mode for Spreadsheet Agents
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
Shot-based quantum encoding: a data-loading paradigm for quantum neural networks
BioHiCL: Hierarchical Multi-Label Contrastive Learning for Biomedical Retrieval with MeSH…
MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Suppor…
Top-$k$ Pareto Bandits: Hypervolume Regret for Multi-Objective Slate Selection
Calibrate Globally, Measure Everywhere: Scaling LLM-Based Prevalence Measurement Across A…
Deep Expert Injection for Anchoring Retinal VLMs with Domain-Specific Knowledge
State-Dependent Safety Failures in Multi-Turn Language Model Interaction
$\texttt{AMEND++}$: Benchmarking Eligibility Criteria Amendments in Clinical Trials
Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection The…
Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities
VideoNorms: Benchmarking Cultural Awareness of Video Language Models