Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
How Good Are Frontier Models at Physics? Expert Re-Grading Reveals Broken Evaluations and…
arXiv cs.AI Research & Papers
Autonomous Research for Open-Ended Problems: A Case Study on Telecom Ticket Retrieval
arXiv cs.AI Research & Papers
Scaling Clinical Judgment to Evaluate Medical AI
arXiv cs.AI Research & Papers
Tracing and Coordinating Cross-Layer Influence for Multimodal Model Merging
arXiv cs.AI Research & Papers
When Agent Metrics Measure Different Things: An Evidence-Grounded Audit of the Praxa AI P…
arXiv cs.AI Research & Papers
Enabling and Understanding Personalization in AI-Generated Advertising Imagery
arXiv cs.AI Research & Papers
When Rubrics Fail: Hallucinations Reveal Blind Spots in Medical AI Evaluation
arXiv cs.AI Research & Papers
What Drives Recovery in Agentic Text-to-Cypher? LAST-CQ: An LLM Agent Self-Refinement Fra…
arXiv cs.AI Research & Papers
Generative AI Use Cases In Real Estate Marketing: Adoption and Constraints in Germany
arXiv cs.AI Research & Papers
Residual Vector-based Reconstruction as Long-Context Recall Regardless of Context Window …
arXiv cs.AI Research & Papers
SteerDuplex: Steerable Duplex Speech Dialogue Models
arXiv cs.AI Research & Papers
Assisted Spatial Cognition Through Vision-Language Models
arXiv cs.AI Research & Papers
Beyond Vector Similarity: Hierarchical Context-Aware Graph RAG vs Standard RAG in Enterpr…
arXiv cs.AI Research & Papers
When Does AI Augment Work? A Workflow-Level Framework for Human-Agent Collaboration
arXiv cs.AI Research & Papers
Generative AI Assisted Workflows in Architectural Conceptual Design: Performance, Creativ…
arXiv cs.AI Research & Papers
MAD: Modality-Adaptive Decoding for Mitigating Cross-Modal Hallucinations in Multimodal L…
arXiv cs.AI Research & Papers
From Collaboration to Capability: Internalizing Routed LLM Experts into Compact Reasoners
arXiv cs.AI Security & Safety
SoK: Rethinking Jailbreaking in the Era of Agentic AI: Attacks, Defenses, and Practical C…
arXiv cs.AI Research & Papers
Unified Agentic Video Editing Across Levels of Complexity and Creativity
arXiv cs.AI Research & Papers
Do LLMs Trust the Accuser or the Accusation? Measuring Belief Shifts in Werewolf
arXiv cs.AI Research & Papers
Beyond Generation and Accuracy: Diagnosing and Enhancing Visual Chain-of-Thought for Geom…
arXiv cs.AI Research & Papers
BlueLM-GUI Technical Report: A Real-Device-Centric Flywheel for Self-Improving Mobile GUI…
arXiv cs.AI Research & Papers
Decentralized Evolution of Hexapod Gaits with Independent Leg Controllers
arXiv cs.AI Research & Papers
Niching Agents in The Core
arXiv cs.AI Research & Papers
A Multi-Vehicle Dataset with Camera, LiDAR, and Radar Sensors and Scanned 3D Models for C…
arXiv cs.AI Research & Papers
Hybrid Physics-AI Framework of Body Center of Mass Dynamics from Wrist-Worn Sensors
arXiv cs.AI Research & Papers
Do Influence-Derived Data Perturbations Enable Machine Unlearning? A Controlled Study of …
arXiv cs.AI Research & Papers
Interpreting the predictions of neural network classification based on a Taylor Coefficie…
arXiv cs.AI Research & Papers
LettuceVisSim: A Simulator That Generates Lettuce Image Time-series for Vision-Based Rein…
arXiv cs.AI Research & Papers
Beyond ID Embeddings: Process-Grounded Language Modeling for Cognitive Diagnosis