Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
How to Ask the AI: A User Perspective Survey for Large Language Model Prompting
arXiv cs.AI Research & Papers
PAST: Privileged Adaptation from Complete Student Trajectories for On-Policy Self-Distill…
arXiv cs.AI Research & Papers
Persuasive and Compliant Tendencies Predict Group Decision-Making in Humans and Language …
arXiv cs.AI Research & Papers
Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States
arXiv cs.AI Research & Papers
Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safe…
arXiv cs.AI Research & Papers
DialectS2S: End-to-End Speech Dialogue Modeling for Low-Resource Chinese Dialects
arXiv cs.AI Research & Papers
Effect of Abstractions and Prompting Strategies on LLM-Guided High-Performance Optimizati…
arXiv cs.AI Research & Papers
ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Cou…
arXiv cs.AI Research & Papers
TLDChoiceNet: Quantitatively Choosing a Transfer Learning Dataset
arXiv cs.AI Research & Papers
Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Pol…
arXiv cs.AI Research & Papers
Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Pr…
arXiv cs.AI Research & Papers
Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Pr…
arXiv cs.AI Research & Papers
Evaluation of Motivational Interviewing Counsellors with Task-Aware Multi-Stage LLM-Based…
arXiv cs.AI Research & Papers
PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling
arXiv cs.AI Research & Papers
NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs
arXiv cs.AI Research & Papers
LLM Reasoning for Subjective Tasks: Failure Modes, Mitigation, and Dynamic Reasoning Rout…
arXiv cs.AI Research & Papers
Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinfo…
arXiv cs.AI Research & Papers
360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents
arXiv cs.AI Research & Papers
Deployable Per-Instance Multi-Layer Activation Steering for Large Language Models
arXiv cs.AI Research & Papers
Compositional Cross-Modality Translation via Whole-Volume Multitask Latent Flow Matching
arXiv cs.AI Research & Papers
AI Evaluation Should Measure Verification Cost, Not Correctness Alone
arXiv cs.AI Research & Papers
DoGMA: A Central-Dogma-Guided Foundation Model for Multi-Omics Alignment and Multi-Task L…
arXiv cs.AI Research & Papers
Agentic Anomaly Detection with ORCA-Style Dynamic Inductive Bias Adaptation in Multimodal…
arXiv cs.AI Research & Papers
FemWear: A Specialized Wearable Foundation Model for Women's Health
arXiv cs.AI Research & Papers
Illusion of Alignment: Detecting Hidden Disagreement in Collaborative Dialogue
arXiv cs.AI Research & Papers
Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalig…
arXiv cs.AI Research & Papers
A New Approach to Characterising Optimisation Problems Using Programmatic Representation …
arXiv cs.AI Research & Papers
From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Deco…
arXiv cs.AI Research & Papers
REVEAL: A Rubric-Guided Agent for Explicit Evidence Sufficiency Verificationin Long-Video…
arXiv cs.AI Research & Papers
Improving Generalization Robustness of Multimodal RLVR