Exploration and Online Transfer with Behavioral Foundation Models
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Beyond Triplet Plausibility: Relation Set Completion in Knowledge Graphs
A swap-adversarial framework for improving domain generalization in electrocorticography-…
A Tutorial on Autonomous Fault-Tolerant Control Using Knowledge-Grounded LLM Agents
Containment Verification: AI Safety Guarantees Independent of Alignment
LiteResearcher: A Scalable Agentic RL Training Framework for Deep Research Agent
GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug…
ECHO: Prune to act, trace to learn with selective turn memory in agentic RL
Sparsity-Inducing Divergence Losses for Biometric Verification
ComplianceGate: Classifier-Gated Multi-Tier LLM Routing for Inference in Regulated Indust…
WorldRoamBench: An Open-World Benchmark for Long-Horizon Stability of Interactive World M…
ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Scien…
When to Truncate a Feature Ranking: A Residual-Overlap Stopping Rule for Subset Selection
Improving LLM Reasoning with Homophily-aware Structural and Semantic Text-Attributed Grap…
ShopX: A Foundation Model for Intent-to-Item Fulfillment in Agentic Shopping
Seeing Is Not Sharing: Some Vision-Language Models Overestimate Common Ground in Asymmetr…
STEB: Style Text Embedding Benchmark
LLM-Empowered Agentic MAC Protocols: A Dynamic Stackelberg Game Approach
FedXDS: Leveraging Model Attribution Methods to counteract Data Heterogeneity in Federate…
Look But Don't Touch with Sparse Autoencoders for Unlearning in Diffusion Models
The HydroGym Reinforcement Learning Platform for Fluid Dynamics
A Technical Typology of AI Systems in Public Administration
Geometry-Preserving Orthonormal Initialization for Low-Rank Adaptation in RLVR
Disentangling Reasoning Logic to Resolve Explicit Knowledge Conflicts
Introspective Coupling: Self-Explanation Training Tracks Behavioral Change Despite Fixed …
Real-Time Source-Free Object Detection
OpenLife: Toward Open-World Artificial Life with Autonomous LLM Agents
Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expressio…
Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models
Belief Contraction in Dynamic Epistemic Logic