IMMENSE: Inductive Multi-perspective User Classification in Social Networks
Explorar
Noticias de IA
21271 elementos — filtrados, clasificados y sin duplicados
Reducing Hallucination in Vision-Language Models via Stage-wise Preference Optimization u…
Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New S…
SafeDivertor: Faithful Divertor Heat Flux Reconstruction from Macroscopic Plasma State Si…
Once a Response, Always a Response: Detecting LLM-generated Text via Latent Prompt Restor…
Stochastic Parrots or Singing in Harmony? Testing Five Leading LLMs for their Ability to …
Is Self-Pretraining really useful to improve diagnosis in medical Time Series?
Gender-Based Heterogeneity in Youth Privacy-Protective Behavior for Smart Voice Assistant…
CREBench: Evaluating Large Language Models in Cryptographic Binary Reverse Engineering
PhaseCoder: Microphone Geometry-Agnostic Spatial Audio Understanding for Multimodal LLMs
Beyond Sentiment: Comparing Traditional NLP and LLM-Based Multi-Dimensional Analysis for …
Beyond Feature Importance: A Comparative Analysis of Pattern Detection Methods in Cluster…
Mind the Gaps: Mixture-of-Minds for Human Simulation
Refining Over Resampling: Test-Time Self-Correction for LLM Reasoning
Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Pers…
Epistemic Trustworthiness in Generative AI: A Normative Framework for Warranted Reliance …
Poli-Bias: Understanding and Measuring Large Language Model Biases in International Polit…
Woodpecker Distillation: Weak Models Diagnose Reasoning Bugs in Strong Models
Layer-wise Positional Bias in Short-Context Language Modeling
Autonomous Learning From Success and Failure: Goal-Conditioned Supervised Learning with N…
ASAT: Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distributio…
SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution
Look Twice: Training-Free Evidence Highlighting for Knowledge-based Visual Question Answe…
Stochasticity Is Not the Hard Part: Reduction and Complexity in Instructional Sequencing …
The Impossibility Triangle of Long-Context Modeling
DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinic…
EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Mode…
Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large La…
MACRO: Markov Chain Routing of Transformer Layers