Prototype Transformer: Towards Language Model Architectures Interpretable by Design
Explorar
Noticias de IA
21861 elementos — filtrados, clasificados y sin duplicados
CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsio…
Why Self-Inconsistency Arises in GNN Explanations and How to Exploit It
ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts
AutoEval Done Right: Using Synthetic Data for Model Evaluation
The Role of Ambiguity in Error Prediction via Uncertainty Quantification
When Single Answer Is Not Enough: Rethinking Single-Step Retrosynthesis Benchmarks for LL…
MidSteer: Optimal Affine Framework for Steering Generative Models
naPINN: Noise-Adaptive Physics-Informed Neural Networks for Recovering Physics from Corru…
Deep Interest Mining for Intent-Enriched Semantic IDs in Multimodal Generative Recommenda…
Just Type It in Isabelle! AI Agents Drafting, Mechanizing, and Generalizing from Human Hi…
Vibe-driven model-based engineering
U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster
Understand and Accelerate Memory Processing Pipeline for Large Language Model Inference
Beyond String Matching: Semantic Evaluation of PDF Table Extraction
scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Pat…
Concept Heterogeneity-aware Representation Steering
IDLM: Inverse-distilled Diffusion Language Models
Mitigating Reward Hacking in RLHF via Bayesian Non-negative Reward Modeling
Learning to Remember, Learn, and Forget in Attention-Based Models
Collaborative and Efficient Fine-tuning: Leveraging Task Similarity
Perturbation Effects on Accuracy and Fairness among Similar Individuals
Coupling Language Models with Physics-based Simulation for Synthesis of Inorganic Materia…
Evaluating the Performance of Deep Learning Models in Whole-body Dynamic 3D Posture Predi…
A Foundation Model for Wearable Movement Data in Mental Health Research
Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation
MINTS: Minimalist Thompson Sampling
GC-MoE: Genomics-Guided Cell-Type-Specific Mixture of Experts for Histology-Based Single-…
Boosting Multimodal Federated Learning via Chained Modality Optimization
Logit Distillation on Manifolds: Mapping by Learning