Safety Alignment of LMs via Non-cooperative Games
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Adaptive data selection improves wearable prediction under low baseline performance
AI-PROPELLER: Warehouse-Scale Interprocedural Code Layout Optimization with AlphaEvolve
Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop
Structured Visual Evidence Decomposition for Evidence-Grounded Multimodal Screening of Ob…
Beyond Text and Tables: Vision-Language Model Integration in ComProScanner for Extracting…
Physics-Informed Neural Networks for Radial Consolidation of Combined Electroosmotic, Vac…
LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs?
On Information Self-Locking in Reinforcement Learning for Active Reasoning of LLM agents
Model Parallelism With Subnetwork Data Parallelism
Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory
TCAR-Gen: Temporal Graph Retrieval with Evidence Fusion for Knowledge-Grounded Generation
RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography
Understanding Stigmatizing Language in Clinical Documentation: A Paired Comparison of Amb…
Toward Robust In-Context Learning: Leveraging Out-of-distribution Proxies for Target Inac…
KnowledgeBerg: Evaluating Systematic Knowledge Coverage and Compositional Reasoning in La…
Bridging the Last Mile of Time Series Forecasting with LLM Agents
Beyond One-shot: AI Agents for Learning in Field Experiments
Agricultural Landscape Understanding At Country-Scale
Coding Agent Is Good As World Simulator
Forget Attention: Importance-Aware Attention Is All You Need
From Capability Models to Automated Planning: An AAS-Native Approach for Automatic PDDL G…
Score Function Gradient Estimation to Widen the Applicability of Decision-Focused Learning
Explainable Data-driven Deep Reinforcement Learning Methods for Optimal Energy Management…
Learning to Construct Practical Agentic Systems
Efficient Weighted Sampling via Score-based Generative Models
A Lightweight Context-Driven Training-Free Network for Scene Text Segmentation and Recogn…
Subliminal Learning Is Steering Vector Distillation
When Single Answer Is Not Enough: Rethinking Single-Step Retrosynthesis Benchmarks for LL…