S2MAM: Semi-supervised Meta Additive Model for Robust Estimation and Variable Selection
Explorar
Noticias de IA
21272 elementos — filtrados, clasificados y sin duplicados
DynaSchedBench: Calibrated Dynamic Scheduling Benchmarks and Observability Paradox in LLM…
On the Origin of Synthetic Information by Means of Steganographic Inheritance
Anomaly as Non-Conformity via Training-Free Graph Laplacian Energy Minimization
Reward Bias Substitution: Single-Axis Bias Mitigations Redirect Optimization Pressure
PetroBench: A Benchmark for Large Language Models in Petroleum Engineering
MolLingo: Molecule-Native Representations for LLM-Powered Scientific Agents
LACUNA: Safe Agents as Recursive Program Holes
EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents
SPARC: Spatial-Aware Path Planning via Attentive Agent Communication
A Fixed-Budget, Cluster-Aware Standard for LLM-as-a-Judge Evaluation: A Multi-Hop RAG Str…
A Unified Framework for the Evaluation of LLM Agentic Capabilities
A Query Engine for the Agents
Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resoluti…
Relational Semantic Reasoning on 3D Scene Graphs for Open World Interactive Object Search
Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning f…
MAVEN A Multi-Agent Framework for Multicultural Text-to-Video Generation
ArcVQ-VAE: A Spherical Vector Quantization Framework with ArcCosine Additive Margin
HO-SFL: Hybrid-Order Split Federated Learning with Backprop-Free Clients and Dimension-Fr…
PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience …
Sense Representations Are Inducible Interfaces
Identifying and Understanding Human Values in Text: A Tailorable LLM-based Architecture
Behavioural Analysis of Alignment Faking
DeepSciVerify: Verifying Scientific Claim--Citation Alignment via LLM-Driven Evidence Esc…
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
RULER: Representation-Level Verification of Machine Unlearning
Cyberbullying Governance on Social Media: A Unified Framework from Content Identification…
Hybrid Neural World Models
SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces
A Policy-Driven Runtime Layer for Agentic LLM Serving