Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting
Explorar
Noticias de IA
30313 elementos — filtrados, clasificados y sin duplicados
DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers
LLM Watermark Evasion via Bias Inversion
InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Tra…
Not All Pixels Are Equal: Pixel-wise Meta-Learning for Medical Segmentation with Noisy La…
QuITE: Query-Based Irregular Time Series Embedding
Whose Name Comes Up? III: Persona Prompting Effects in LLM-Based Scholar Recommendation
Can Quantum Federated Learning Withstand Circuit-Level Backdoors?
Transferable Reinforcement Learning via Probabilistic Latent Embeddings and Dynamic Polic…
Evaluation of AI Ethics Tools in Language Models: A Developers' Perspective Case Study
The Optimal Sample Complexity of Linear Contracts
Pruning and Distilling Mixture-of-Experts into Dense Language Models
RAGe: A Retrieval-Augmented Generation Evaluation Framework
Generic Interpretation Approach for Transformer Models Incorporating Heterogenous Attenti…
Adapting, Fast and Slow: On Few-Shot Transportability of Compositions
SAME: Stabilized Mixture-of-Experts for Multimodal Continual Instruction Tuning
Bridging the Stability-Expressivity Gap: Synthetic Data Scaling and Preference Alignment …
Memory-Based vs. Context-Only Conditioning Produces Distinct Behavioral Patterns in State…
On the Intrinsic Limits of Transformer Image Embeddings in Non-Solvable Spatial Reasoning
Sense Representations Are Inducible Interfaces
Identifying and Understanding Human Values in Text: A Tailorable LLM-based Architecture
Behavioural Analysis of Alignment Faking
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
RULER: Representation-Level Verification of Machine Unlearning
Cyberbullying Governance on Social Media: A Unified Framework from Content Identification…
A Policy-Driven Runtime Layer for Agentic LLM Serving
Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems
Revealing Algorithmic Deductive Circuits for Logical Reasoning
When Context Flips, Safety Breaks: Diagnosing Brittle Safety in Aligned Language Models
AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models