Spectral Rewiring for Exploration, Purification, and Model Merging
Explorar
Noticias de IA
22318 elementos — filtrados, clasificados y sin duplicados
Don't Wait to Reply: Towards Responsive yet Thoughtful Dialogue through Proactive Thinking
S-EMBER: A Large-Scale Benchmark for Streaming Egocentric Memory Retrieval
Not All Refusals Are Equal: How Safety Alignment Fails Cybersecurity at Scale
SMOCS: A Streaming Framework for Simplified Deployment, Monitoring, and Optimization of M…
Safe Inference-Time Alignment via Lagrangian Reward Augmentation
Test-Time Adaptation for EEG Foundation Models: A Systematic Study under Real-World Distr…
Seduced by the Narrative: Assessing Rule Adherence in Semi-Open Textual Sandboxes
Teaming Up with AI: Coordination and Cooperation
Parallelized Autoregressive Decoding for Omni-Modal Dense Video Captioning
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling
CONFLUX: A Latent Diusion Model for 3D Chest-CT Synthesis with RL Post-Training
Back to Basics: Improving Molecular Understanding in LLMs via SMILES-Graph Translation
HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Dif…
Is Agentic Code Review Helpful? Mining Developers' Feedback to CodeRabbit Reviews in the …
From Judgments to Issues: Structured Extraction of Legal Reasoning with Citation-Hallucin…
FedAvg for HAR: Exploring the Tradeoff Between Personalized and Generalization Accuracy
Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation
PPE-Bench: A Benchmark for Evaluating MLLM Unlearning under Private-Public Entanglement
Learning Taxonomic Trees with Hierarchical Representation Regularization for Large Multim…
Bootstrap Flow-Map Tree Sampling Enables Online Feedback Driven Search
Full Glyph Images Beat Token Embeddings: A Controlled Study for Transformers
Separating Representation from Reconstruction Enables Scalable Text Encoders
Finite Reliability Representations: Noise-Calibrated Belief-Space Covers for Reliable Dec…
A Unified Algebraic Framework for Classification Performance Evaluation
OmniOpt: Taxonomy, Geometry, and Benchmarking of Modern Optimizers
The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling
Telescope: Improving Zero Shot Detection of LLM Generated Content By Measuring Token Repe…
BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources