How to Ask the AI: A User Perspective Survey for Large Language Model Prompting
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
PAST: Privileged Adaptation from Complete Student Trajectories for On-Policy Self-Distill…
Persuasive and Compliant Tendencies Predict Group Decision-Making in Humans and Language …
Prompt Embedding Probes (PEP): Hallucination Detection in LLMs from Hidden States
Do All LLMs Know When They're Being Harmful? A Reproducibility Study of Latent-Space Safe…
DialectS2S: End-to-End Speech Dialogue Modeling for Low-Resource Chinese Dialects
Effect of Abstractions and Prompting Strategies on LLM-Guided High-Performance Optimizati…
ComboShoppingBench: Evaluating LLM Agents for Budget-Constrained Basket Shopping with Cou…
TLDChoiceNet: Quantitatively Choosing a Transfer Learning Dataset
Privileged Likelihood Is Not Automatically Value: Three Checks for Token Credit in On-Pol…
Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Pr…
Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Pr…
Evaluation of Motivational Interviewing Counsellors with Task-Aware Multi-Stage LLM-Based…
PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling
NeuPAT: Neuron-aware Plasticity Allocation Tuning for Language-Preserving MLLMs
LLM Reasoning for Subjective Tasks: Failure Modes, Mitigation, and Dynamic Reasoning Rout…
Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinfo…
360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents
Deployable Per-Instance Multi-Layer Activation Steering for Large Language Models
Compositional Cross-Modality Translation via Whole-Volume Multitask Latent Flow Matching
AI Evaluation Should Measure Verification Cost, Not Correctness Alone
DoGMA: A Central-Dogma-Guided Foundation Model for Multi-Omics Alignment and Multi-Task L…
Agentic Anomaly Detection with ORCA-Style Dynamic Inductive Bias Adaptation in Multimodal…
FemWear: A Specialized Wearable Foundation Model for Women's Health
Illusion of Alignment: Detecting Hidden Disagreement in Collaborative Dialogue
Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalig…
A New Approach to Characterising Optimisation Problems Using Programmatic Representation …
From Recovery to Drop-off: How Action Post-training Reduces a VLM's Late-Layer Depth Deco…
REVEAL: A Rubric-Guided Agent for Explicit Evidence Sufficiency Verificationin Long-Video…
Improving Generalization Robustness of Multimodal RLVR