Predicting Consequences and Reinforcing Navigation Policies with Latent World Models
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Benchmarking AI Agents for Hardware Design Automation via MCP Tool Calling
Can You Say This for Me? Speaking Up by Proxy in Co-Located Discussion
TutorTrace: A Dataset and Taxonomy for Classifying Learner Behavioral States during AI-As…
Multi2AV-Safety: Benchmarking Safety in Multimodal-to-Audio-Video Generation
Feature Transformation Enhanced Jacobi Polynomial Graph Filtering for Graph Anomaly Detec…
Invocation-Level Reliability of Tool-Using Agents
Naive Prompt Optimization: Rethinking the Need for Complex Prompt Search
When Tool Outputs Become Commands: Separating Action Induction from Runtime Authorization…
Beyond Accuracy: A Qualitative Analysis of Vision-Language Models for Hate Speech Detecti…
ADeptS-Bench: Measuring the Trustworthiness of Computer Use Agents Across Devices
A Task-Centric Ontology and Deterministic Domain Rules as a Verifiable Core for AI-Assist…
Refusal Is Not Robustness: Auditing Confident Fabrication in Large Language Models on a P…
Is Your Neighborhood Safe? Place-based Stigma in Large Language Models' Urban Safety Judg…
MemToC: Benchmarking Memory-Tool Conflict Resolution in Large Language Models
What Makes Good Agentic Data? An ACE Lens on Data Generation for LLM Agents
Active Diffusion-Based Inference for Ill-Posed Inverse Problems under Incomplete Priors
A Safety-Gated Multimodal AI Backend for Mental-Health Support: Hierarchical State Repres…
LLMs Can Design Near-Optimal OR Algorithms
EEG-to-Report: An Annotation and Feature-Text Framework for Training Language Models on C…
Methodological and Conceptual Framework for 5D Multi-Table Analysis: A Unified Approach f…
Explainable Artificial Intelligence for Customer Churn Prediction in Telecommunications: …
Selection Bias Correction in Retail Intelligence
CIFQA: A Deterministic Tool-Grounded Multi-Agent LLM Framework for Financial Query Answer…
The Artificial Experimentalist: Discovery and Control of Self-Organizing Phenomena with A…
Co-Evolving Structured Knowledge and Reasoning in Language Models
Artificial Intelligence Models Can Predict and Collaboratively Modulate Human Memory Sear…
PICasso: An AI-Enabled Design Framework for Autonomous Optimization of Silicon Photonic D…
KnockGS:interaction-Grounded Calibrationof Physical Gaussian Representations
Large Models for Battery Prognostics and Health Management: A Review and Future Roadmap