Explorar

Noticias de IA

22318 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
C3-Bench: A Context-Aware Change Captioning Benchmark
arXiv cs.AI Research & Papers
A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation
arXiv cs.AI Research & Papers
HG-Bench: A Benchmark for Multi-Page Handwritten Answer-Region Grounding in Automated Hom…
arXiv cs.AI Research & Papers
The impact of artificial intelligence on enterprise software user roles
arXiv cs.AI Research & Papers
Evaluating LLMs on Real-World Software Performance Optimization
arXiv cs.AI Research & Papers
An Approach for a Supporting Multi-LLM System for Automated Certification Based on the Ge…
arXiv cs.AI Research & Papers
Probabilistic Agents in Deterministic Audits: Evaluating Multi-Agent Systems for Automate…
arXiv cs.AI Research & Papers
TL++: Accuracy and Privacy Preserving Traversal Learning for Distributed Intelligent Syst…
arXiv cs.AI Research & Papers
Staying In Character: Perspective-Bounded Memory For Book-Based Role-Playing Agents
arXiv cs.AI Research & Papers
Taxonomy of Risks on Automated Fact-Checking Systems Considering its Propagation
arXiv cs.AI Research & Papers
Is GraphRAG Needed? From Basic RAG to Graph-/Agentic Solutions with Context Optimization
arXiv cs.AI Research & Papers
Steering Vision-Language Models with Joint Sparse Autoencoders
arXiv cs.AI Research & Papers
Gradient-based inverse lithography for EUV masks via the waveguide method and a physics-i…
arXiv cs.AI Research & Papers
Uncertainty Quantification for Computer-Use Agents: A Benchmark across Vision-Language Mo…
arXiv cs.AI Research & Papers
Do Encoders Suffice? A Systematic Comparison of Encoder and Decoder Safety Judges for LLM…
arXiv cs.AI Research & Papers
MiniOpt: Reasoning to Model and Solve General Optimization Problems with Limited Resources
arXiv cs.AI Research & Papers
Edges Before Embeddings: A Confidence-Aware Blur Gate for Vision-Language Pipelines
arXiv cs.AI Research & Papers
Semantic Consistency Policy Optimization for Reinforcement Learning of LLM Agents
arXiv cs.AI Research & Papers
AutoRelAnnotator: Calibrated Model Cascades for Cost-Efficient Relevance Evaluation in Sp…
arXiv cs.AI Research & Papers
Pulmonary Embolism Risk Stratification from CTPA and Medical Records: Vascular Graphs Are…
arXiv cs.AI Research & Papers
Enhancing Brain MRI Anomaly Detection and Reasoning with ROI Rethink and Synthetic Data
arXiv cs.AI Research & Papers
Overview of HIPE-2026: Person-Place Relation Extraction from Multilingual Historical Texts
arXiv cs.AI Research & Papers
Multi-Agent Goal Recognition with Team- and Goal-Conditioned Reinforcement Learning and F…
arXiv cs.AI Research & Papers
Weave of Formal Thought
arXiv cs.AI Research & Papers
SpeechEQ: Benchmarking Emotional Intelligence Quotient in Socially Aware Voice Conversati…
arXiv cs.AI Research & Papers
FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Dist…
arXiv cs.AI Research & Papers
Can Trustless Agents Be Trusted? An Empirical Study of the ERC-8004 Decentralized AI Agen…
arXiv cs.AI Research & Papers
Natural Ungrokking: Asymmetric Control of Which Rules Survive Pretraining
arXiv cs.AI Research & Papers
Model Forensics: Investigating Whether Concerning Behavior Reflects Misalignment
arXiv cs.AI Research & Papers
On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity