Agentic Transformers Provably Learn to Search via Reinforcement Learning
Explorar
Noticias de IA
21861 elementos — filtrados, clasificados y sin duplicados
The New Social Image: How AI Competency and AI Proactivity Influence Self- and Peer-Perce…
Uncovering Competency Gaps in Large Language Models and Their Benchmarks
Beyond Augmentation: Score-Guided Pathological Prior for EEG-based Depression Detection
ChurnNet: A Optimized Modern AI for Churn Prediction
A Foundation Model for Wearable Movement Data in Mental Health Research
Coupling Language Models with Physics-based Simulation for Synthesis of Inorganic Materia…
Perturbation Effects on Accuracy and Fairness among Similar Individuals
Collaborative and Efficient Fine-tuning: Leveraging Task Similarity
Learning to Remember, Learn, and Forget in Attention-Based Models
Mitigating Reward Hacking in RLHF via Bayesian Non-negative Reward Modeling
IDLM: Inverse-distilled Diffusion Language Models
Concept Heterogeneity-aware Representation Steering
scicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated Pat…
Beyond String Matching: Semantic Evaluation of PDF Table Extraction
Understand and Accelerate Memory Processing Pipeline for Large Language Model Inference
U-Cast: A Surprisingly Simple and Efficient Frontier Probabilistic AI Weather Forecaster
Vibe-driven model-based engineering
Just Type It in Isabelle! AI Agents Drafting, Mechanizing, and Generalizing from Human Hi…
Deep Interest Mining for Intent-Enriched Semantic IDs in Multimodal Generative Recommenda…
MidSteer: Optimal Affine Framework for Steering Generative Models
Why Self-Inconsistency Arises in GNN Explanations and How to Exploit It
CLIP Tricks You: Training-free Token Pruning for Efficient Pixel Grounding in Large VIsio…
Avatar Forcing: Real-Time Interactive Head Avatar Generation for Natural Conversation
Improving IoT Intrusion Detection Through SMOTE-Based Oversampling and Extended Multi-Mod…
DataShield: Safety-degrading Data Filtering for LLM Benign Instruction Fine-Tuning
A physics-informed foundation model for quantitative diffusion MRI
A Protocol-Language Model for Network Intrusion (Without Deep Packet Inspection)
DiffCrossGait: Trajectory-Level Alignment for 2D-3D Cross-Modal Gait Recognition via Late…
Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying