Topological Attribution Distance (TAD): Revealing Segment-Level RAG Influence on LLM Outp…
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
OGX: An Open-Source, Vendor-Neutral Generative AI Application Server
Toward Safe LLM Agents: A Survey of Specification, Verification, and Enforcement
Euclid-Omni : A Unified Neuro-Symbolic Framework for Plane Geometry
An Agentic Framework Using Rules and LLMs for Embedding and Annotating Descriptive Docume…
The Hallucination Snowball: Modeling Error Propagation as State Transitions in Multi-Agen…
When Do LLMs Apply the Wrong Law? Diagnosing LLM Failures in Temporal Legal Reasoning
Do LLM Agents Negotiate Rationally? A Mechanism-Design Framework for Verifiable Multi-Age…
Large Language Models and their Awareness of Mechanics and Spatial Geometry
A Human-Centred Approach to Benchmarking LLMs for Parenting Advice
Learning Agent Execution for KV-Cache Management in Agentic Serving
Task- and Session-Level Model Routing: A Common-Interface Hybrid Evaluation of Four Open-…
Accuracy and Reliability of Large Language Models in Cosmetic Chemistry and Skin Health: …
Evaluating Multimodal LLMs across Text and Audio Modalities for Accessible Disaster Assis…
When Uncertainty Isn't Enough: An Empirical Study of Self-Correction in Code Generation
Cross-Domain Industrial Fault Detection by Causal Mechanism Monitoring
Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting
Path2ST: Hierarchical Cell-Tissue Grounded Cross-Modal Translation for Spatial Transcript…
Which Question Is Your Attention Metric Answering? Attention Rows as Compositional Data
DeCo-MIL: Debiased Counterfactual Reasoning for Long-Tailed Whole Slide Image Analysis
CoupVisor: Strategy Optimization by Round and Challenge Decision Support
Dear Algo: A Precision-First Agentic Intent Layer for Unified Search and Recommendation
Unified Pedestrian Path Prediction Using Inverse Reinforcement Learning
UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations
Solvable Sokoban Without a Solver via Diffusion
ALPS: Measuring Valid Creativity in Large Language Models with Mathematical Construction
MUPA$^{2}$E: Multimodal Unified Perception with Asymmetric Attention for Emotion Assessme…
Prior Audit-Repair Context Shifts LLM Verifier Thresholds Toward Leniency
Governance at the Boundary: How Agent Decomposition Degrades Policy Compliance
Eigenanalysis framework for autoregressive neural emulators of multi-scale chaotic dynami…