Benchmarking Trustworthiness of SLMs: Pre-trained vs. Compressed
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operat…
OEIS Open: How many conjectures can language models turn into theorems?
Distillation of Foundation Models for Time-dependent PDEs
LazyTrain: Limited-resource Allocation toward Zero-waste Yield Optimization in Large Lang…
A Factor Graph Approach to Scalable Multi-Output Gaussian Process Regression
Do You See What You Draw? A Semantic Closed-Loop Framework for Holistic Evaluation of Uni…
Policy-as-logic for robust reasoning over rules
Agent Skills Can Be Harmful: An Empirical Study of Skill-Induced Failures in LLM Agents
CoQui: A Coordinate-Conditioned Quantum Implicit Generative Adversarial Network for End-t…
Forward and Inverse Virtual Metrology for Phototransistor Gain: A Hierarchical, Uncertain…
Small-Scale Experiments: Are We There Yet?
LookBack: Where and How to Score LVLM Responses via Visual Reference Usage
Air Quality Station Simulation via LSTM and Attention-Based Modelling
GeoBridge: Decoupled Semantic Conditioning for Generative Image Geolocalization
Towards Understanding On-Policy Distillation through the Lens of Test-Time Scaling
Located but Not Releasable: Silent Gate Inversion and Bounded Linear Release
CoDiR: Confidence-Guided Diffusion Refinement for Semi-Supervised Histopathology Segmenta…
Hybrid Gated Attention
GRPO for Financial Advice Generation: Outperforming Commercial LLMs under CATE Evaluation
Language-Conditional Dequantization: Recovering What Quantization Steals from Non-English…
TradingMoE: Routing the Right Experts in Evolving Markets
VOLA: Improving Open-World Driving by VLM-Based Semantic Attribute Prediction
The Sleeping Agent: What Gist-Based Context Compression Loses and Why
Diagnosis Before Recovery: Turning Agent Failures into Selective Self-Correction
HyperANFIS: Enhancing Rule Representation and Interpretability in Adaptive Neuro-Fuzzy Sy…
Tight Nonasymptotic Local Convergence of Sinkhorn-Knopp
Ripple-Pivot Search: Active Parallel Decoding for Diffusion Large Language Models
JieZi: A Large-Scale Expert-Audited Dataset and Benchmark for Ancient Chinese Character E…
Locating and Controlling Implicit Personalization in Large Language Models