Hacking Generative Perplexity: Why Unconditional Text Evaluation Needs Distributional Met…
Explorar
Noticias de IA
22115 elementos — filtrados, clasificados y sin duplicados
PACT: Self-Evolving Physical Safety Alignment for Diffusion Policies in Embodied Manipula…
TimpaTeks: Automatic In-place Text Sequence Modification via Diffusion Language Model Ste…
SceneConductor: 3D Scene Generation from Single Image with Multi-Agent Orchestration
STAR-KV: Low-Rank KV Cache Compression via Soft Thresholding for Adaptive Rank Control
Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy
Pre-Intervention Prediction of Sparse Autoencoder Steering Side Effects
Learning to Attack and Defend: Adaptive Red Teaming of Language Models via GRPO
Tyan-WP: A Wind Power Foundation Model for Ultra-Short-Term Probabilistic Forecasting
Trait-space Monitoring for Emergent Misalignment During Supervised Finetuning
Active Learning with Foundation Model Priors: Efficient Learning under Class Imbalance
HASA: Subnet Allocation for Compute-Constrained Model-Heterogeneous Federated Learning
ScaleSweep: Accurate NVFP4 Post-Training Quantization of LLMs via Block Scale Initializat…
No Free Lunch for Synthetic Images under Data Scarcity Conditions
NeuroAlign: Hierarchical Multimodal Fusion of Dynamic and Structural Neuroimaging for MCI…
Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Syste…
Reconstructing and forecasting disease trajectories of patients with Alzheimer's disease …
Syll: Open-Source Personal Automation with Cross-Surface Execution
SmartMixed: A Two-Phase Training Strategy for Adaptive Activation Function Learning in Ne…
DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback
FormalASR: End-to-End Spoken Chinese to Formal Text
LEAP: Learnable End-to-End Adaptive Pruning of Large Language Models
TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models
CLPO: Curriculum Learning meets Policy Optimization for LLM Reasoning
Sound and Complete Neurosymbolic Reasoning with LLM-Grounded Interpretations
IDEQ -- Improving Diffusion Models for the Traveling Salesman Problem (TSP) by Leveraging…
A Survey on Large Language Model-Based Game Agents
APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Mu…
Learning Behavioral Signals from Encrypted Smartphone Network Traffic
DynamicPO: Dynamic Preference Optimization for Recommendation