Complete, Scalable, and Robust Prioritized Planning for Multi-Robot Ordered Storage and R…
Explorar
Noticias de IA
29336 elementos — filtrados, clasificados y sin duplicados
A Grounded and Decomposed Framework for Relation-Level Hallucination Evaluation in Abstra…
Second Order Drifting Models
NL2SHACL-Bench: A Benchmark Suite for Natural Language to SHACL Translation
JaleesBench: Are AI Assistants Good Spiritual Company?
Hidden Language Consistency Phenomena in Reasoning LLMs
HugSelect: An Explainable Multi-Criteria Decision-Support Framework for foundation-model …
Multi-Branch Policy Optimization for Multimodal Large Language Models
Motif 3: Technical Report
LLM Reasoning for Subjective Tasks: Failure Modes, Mitigation, and Dynamic Reasoning Rout…
Scaling Inherently Interpretable Language Models
Adversarial Attacks on Deep OCR Systems
Enhancing Knowledge Tracing through Leakage-Free and Recency-Aware Embeddings
COMEX: A Composition-Grounded Benchmark and Learning Framework for Explainable Aesthetic …
Geometry Beats Estimated Depth: RGB-Only Multi-Camera 3D Tracking under Sim2Real
LoRSA: Toward Generalizable Parameter-Efficient Fine-Tuning for Biomedical Downstream Tas…
Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control
Generalizing deep reinforcement learning across cable-driven parallel robot configuration…
Learning an Interior Layout Policy in a Domain Specific Language Action Space
GLocFM: A Geometry-Aware Foundation Model for 3D Indoor Wireless Localization
MasDrift: Benchmarking Authorization Preservation Across Multi-Agent Architectures
Beyond "I Can't Help With That": How Child Safety Experts Evaluate AI Chatbot Safety
Search-G1: Grounded Search Agents via Representation-Based Intrinsic Rewards
How sensitive do we want AI to be? Socio-communicative competencies of large language mod…
EMMR: Emotion-Mediated Multimodal Reasoning for Personality Assessment in Asynchronous Vi…
SAKE: Structured Agentic Knowledge Extrapolation for Complex LLM Reasoning via Reinforcem…
MCIF: Multimodal Crosslingual Instruction-Following Benchmark from Scientific Talks
Multilingual Agent-Based World Modeling for Social Science
Adaptive Sequential Test Planning for Multi-Mechanism Reliability Qualification via Bayes…
MOSAIC: Adversarial Co-evolution of Specialist Heuristics and Problem Instances for LLM-b…