LCSHBench: A Multilingual, Consensus-Grounded Benchmark for Library of Congress Subject H…
Explorar
Noticias de IA
29346 elementos — filtrados, clasificados y sin duplicados
Strabo: Declarative Specification and Implementation of Agentic Interaction Protocols
Fog of Love: Engineering Virtuous Agent Behavior with Affinity-based Reinforcement Learni…
Characterizing initial human-AI proof formalization workflows
Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation
Cascading Hallucination in Agentic RAG: The CHARM Framework for Detection and Mitigation
AgentJet: A Flexible Swarm Training Framework for Agentic Reinforcement Learning
DEFLECT: Temporal Counterfactual Preference Learning for Delay-Robust Asynchronous VLAs
Extending Fair Null-Space Projections for Continuous Attributes to Kernel Methods
Retrieval and competition: how a protein foundation model starts a protein
On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers
Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis
Platonic Transformers: A Solid Choice For Equivariance
SAM 3D: 3Dfy Anything in Images
Invariant Gradient Alignment for Robust Reasoning Distillation
From Prompt to Process: a Process Taxonomy and Comparative Assessment of Frameworks Suppo…
OA-CutMix: Correcting the Label Bias of CutMix
AdaKoop: Efficient Modeling of Nonlinear Dynamics from Nonstationary Data Streams with Ko…
Contextual Multi-Task Reinforcement Learning for Autonomous Reef Monitoring
ZeroUnlearn: Few-Shot Knowledge Unlearning in Large Language Models
Quantum entanglement provides a competitive advantage in adversarial games
TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering
Spatial Transcriptomics as Images for Large-Scale Pretraining
FinTradeBench: A Financial Reasoning Benchmark for LLMs
PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Rol…
SUSD: Structured Unsupervised Skill Discovery through State Factorization
Does Order Matter : Connecting The Law of Robustness to Robust Generalization
A Unified Framework for Locality in Scalable MARL
Tuning the Implicit Regularizer of Masked Diffusion Language Models: Enhancing Generaliza…
L$^3$: Large Lookup Layers