Information Specialization and Constrained Synthesis in Multi-Agent LLM Forecasting: A Pr…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Reproducing and Evaluating the Generalizability of Subliminal Learning in Open-Weight Mod…
Learning to adapt GR(1) specifications through degradation
EvoRS: On-Policy Self-Evolution of Reward Systems for Open-Ended Reinforcement Learning
Accelerating battery research with an interoperable interface between FINALES and Kadi4Mat
Hierarchical Belief Modeling for Zero-Shot Opponent Adaptation in Partially Observable Mu…
LifeFuse-Mem: Lifecycle-Aware State Fusion Against Temporary Overwriting for Long-Term Me…
TripPattern: A Pattern-based Text Watermarking Method for Large Language Models
SCQ: Stabilizing Conservative Q-Learning with Sigmoid-Bounded Entropy
MPT: Missing Prototype Tracking via Barycentric Reconstruction in Vehicular Federated Lea…
K-Bench: A Benchmark for LLM Unlearning in Agentic Deployments
Skill Issue: Lessons from Optimizing Repository SKILLs for Coding Agents
Diffusion Models and Concept Formation
MedRoundsQA: A Persona and Difficulty Aware Evaluation for Multi-Turn Medical Consultatio…
Class-wise Contribution Estimation via Logit Maximization for Federated Learning
I Am AdMan: A Pipeline for Automatic Generation of Personalized Advertising Imagery
Implicit Personality Representations in Humans and LLMs
Embodied-BenchForge: A Closed-Loop Agentic Workflow for Embodied Benchmark Construction
Representation Before Training: A Practical Benchmark for Generative Medical Event Model …
PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corre…
El Agente Quntur: A research collaborator agent for quantum chemistry
Chopthin-Consensus Power Sampling: A Diversity-Preserving Approach to LLM Decoding
Meddies-PII: A Multilingual Framework for Personally Identifiable Information Extraction …
DU-NO: A Parameter-Efficient Double U-Shaped Neural Operator for Phase-Resolving Wave Mod…
DenseTRF: Texture-Aware Unsupervised Representation Adaptation for Surgical Scene Dense P…
Can LLMs in Draft-Verify-Revise Pipelines Resolve Deictic Ambiguity?
Measuring Pragmatic Influence in Large Language Model Instructions
The Vienna 4G/5G Drive-Test Dataset
A Dataset and Benchmarks for Atrial Fibrillation Detection from Electrocardiograms of Int…
Unified Text-Image Generation with Weakness-Targeted Post-Training