Formal Verification of Romanov's Triplet Logic: A Verified Filter for Sliding-window 3-CN…
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
TTSD-FAR: Test-Time Self-Distillation with Fisher-Anchored Restoration for Missing-Modali…
Mechanistic Interpretability of Structure-Aware Numerical Reasoning in LLaMA 3.1 8B
Selection, Recombination, or a Fresh Solve? A Candidate-Free Control for Single-Pass Test…
Efficient Adaptation of LLMs for Hate Speech Detection in Low-Resource Languages: A Compa…
Coupled-cluster molecular properties across the main group that extrapolate beyond traini…
One Gate Is Not Enough: Composing Stateful Pre-Action Controls for Agentic AI
FairGlucose: A CGM Fairness Benchmark Reveals Subgroup Disparities Hidden in Population-L…
Low-Power, Neuromorphic, Acoustic Anomaly Detection for Persistent Machine Monitoring
From Inference to Adaptation: A Unified Optimal Transport View of Vision Language Model
Solving Is Not Drawing: A Benchmark for Diagrammatic Reasoning in Olympiad Geometry
Which Negatives Matter? Ask Your Text Encoder: Adaptive Similarity Margins for Dense-Capt…
FinRCA-Bench: Benchmarking Evidence Retrieval and Reasoning for Financial AI Systems
Visual-Prompt Guided Wildlife Instance-Level Recognition
GigaBrain-WBC-0.5: A Behavior World Model for Robust Whole-Body Control with Environment …
When Do LLMs Actually Help? Evaluating LLMs as Data Quality Annotators
A systematic review of machine learning techniques to address diagnosis and treatment of …
Are LLMs Safe Beyond Text: Do Emojis Expose Gaps in Safety Evaluation
How Quantum Is the Advantage? A Fair, Calibration- and Noise-Aware Benchmark and Attribut…
Accurate Decoding of Natural Sentences from Non-Invasive Brain Recordings
The Deontic Gap: Large Language Models and the Modal Language of Obligation
Global Index on Responsible AI 2026 : Conceptual Framework and Methodology
TokenPowerSandbox: Evidence-Gated CPU-First Screening for Energy-Aware LLM Serving
Different Facets of Verbalised Overconfidence: an Interpretability Study
StocksTalk: A Voice-Enabled Conversational Agent for Structured Query Generation over Web…
DeepTCM1.0: A Multi-Expert AI Agent for Deciphering Mechanisms of Chinese Herbal Formulae…
Same Facts, Different Updates: Inference Setup Shapes LLM Behavior in Medical Allocation
Computational Orientalism: Measuring Structural Discourse Bias in Large Language Models U…
Self- and Other-Labels Induce Bidirectional Bias in LLM Judges
NE-BERT: A Multilingual Language Model for Nine Northeast Indian Languages