Information Specialization and Constrained Synthesis in Multi-Agent LLM Forecasting: A Pr…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
TripPattern: A Pattern-based Text Watermarking Method for Large Language Models
Toward Robust Personalized Alignment for LLMs: Mitigating Persona Drift in Multi-Turn Dia…
Reproducing and Evaluating the Generalizability of Subliminal Learning in Open-Weight Mod…
Hierarchical Belief Modeling for Zero-Shot Opponent Adaptation in Partially Observable Mu…
LifeFuse-Mem: Lifecycle-Aware State Fusion Against Temporary Overwriting for Long-Term Me…
Is Gaussian Splatting Becoming Neural Again? A Taxonomy and Controlled Study of Learned P…
GTA: Graph Theory Agent and Benchmark for Algorithmic Graph Reasoning with LLMs
EvoRS: On-Policy Self-Evolution of Reward Systems for Open-Ended Reinforcement Learning
Skill Issue: Lessons from Optimizing Repository SKILLs for Coding Agents
SCQ: Stabilizing Conservative Q-Learning with Sigmoid-Bounded Entropy
MPT: Missing Prototype Tracking via Barycentric Reconstruction in Vehicular Federated Lea…
I Am AdMan: A Pipeline for Automatic Generation of Personalized Advertising Imagery
Implicit Personality Representations in Humans and LLMs
K-Bench: A Benchmark for LLM Unlearning in Agentic Deployments
PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corre…
MedRoundsQA: A Persona and Difficulty Aware Evaluation for Multi-Turn Medical Consultatio…
DenseTRF: Texture-Aware Unsupervised Representation Adaptation for Surgical Scene Dense P…
Class-wise Contribution Estimation via Logit Maximization for Federated Learning
Countdown-Code: A Testbed for Studying The Emergence and Generalization of Reward Hacking…
The Vienna 4G/5G Drive-Test Dataset
OA-NBV: Occlusion-Aware Next-Best-View Planning for Human-Centered Active Perception on M…
Measuring Pragmatic Influence in Large Language Model Instructions
FEAT: A Linear-Complexity Foundation Model for Extremely Large Structured Data
Dead Weights, Live Signals: Feedforward Graphs of Frozen Language Models
LLM Compression by Block Removal with Constrained Binary Optimization
A Dataset and Benchmarks for Atrial Fibrillation Detection from Electrocardiograms of Int…
Diffusion Models and Concept Formation
Unified Text-Image Generation with Weakness-Targeted Post-Training
Representation Before Training: A Practical Benchmark for Generative Medical Event Model …