Explorar

Noticias de IA

21271 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
IMMENSE: Inductive Multi-perspective User Classification in Social Networks
arXiv cs.AI Research & Papers
Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization u…
arXiv cs.AI Research & Papers
Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New S…
arXiv cs.AI Research & Papers
SafeDivertor: Faithful Divertor Heat Flux Reconstruction from Macroscopic Plasma State Si…
arXiv cs.AI Research & Papers
Once a Response, Always a Response: Detecting LLM-generated Text via Latent Prompt Restor…
arXiv cs.AI Research & Papers
Stochastic Parrots or Singing in Harmony? Testing Five Leading LLMs for their Ability to …
arXiv cs.AI Research & Papers
Is Self-Pretraining really useful to improve diagnosis in medical Time Series?
arXiv cs.AI Research & Papers
Gender-Based Heterogeneity in Youth Privacy-Protective Behavior for Smart Voice Assistant…
arXiv cs.AI Research & Papers
CREBench: Evaluating Large Language Models in Cryptographic Binary Reverse Engineering
arXiv cs.AI Research & Papers
PhaseCoder: Microphone Geometry-Agnostic Spatial Audio Understanding for Multimodal LLMs
arXiv cs.AI Research & Papers
Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for …
arXiv cs.AI Research & Papers
Beyond Feature Importance: A Comparative Analysis of Pattern Detection Methods in Cluster…
arXiv cs.AI Research & Papers
Mind the Gaps: Mixture-of-Minds for Human Simulation
arXiv cs.AI Research & Papers
Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning
arXiv cs.AI Research & Papers
Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Pers…
arXiv cs.AI Research & Papers
Epistemic Trustworthiness in Generative AI: A Normative Framework for Warranted Reliance …
arXiv cs.AI Research & Papers
Poli-Bias: Understanding and Measuring Large Language Model Biases in International Polit…
arXiv cs.AI Research & Papers
Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models
arXiv cs.AI Research & Papers
Layer-wise Positional Bias in Short-Context Language Modeling
arXiv cs.AI Research & Papers
Autonomous Learning From Success and Failure: Goal-Conditioned Supervised Learning with N…
arXiv cs.AI Research & Papers
ASAT: Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distributio…
arXiv cs.AI Research & Papers
SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution
arXiv cs.AI Research & Papers
Look Twice: Training-Free Evidence Highlighting for Knowledge-based Visual Question Answe…
arXiv cs.AI Research & Papers
Stochasticity Is Not the Hard Part: Reduction and Complexity in Instructional Sequencing …
arXiv cs.AI Research & Papers
The Impossibility Triangle of Long-Context Modeling
arXiv cs.AI Research & Papers
DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinic…
arXiv cs.AI Research & Papers
EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents
arXiv cs.AI Research & Papers
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Mode…
arXiv cs.AI Research & Papers
Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large La…
arXiv cs.AI Research & Papers
MACRO: Markov Chain Routing of Transformer Layers