GPTKB 2.0: Browsing, Querying, and Auditing a Disambiguated LLM-Derived Knowledge Base
Explorar
Noticias de IA
29336 elementos — filtrados, clasificados y sin duplicados
FinRank: An Evidence-Grounded Benchmark for Financial Question Answering and Retrieval ov…
MEDLEY-BENCH: Benchmarking Behavioural Metacognition and Belief Revision Under Social Pre…
Curriculum as Code: An AI-Assisted Architecture for Instructional Design in STEM Education
A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs Whil…
SkillEval: Decomposing Agent Skill Quality into Interpretable Signals
Fluid-DiT: Graph-Free Diffusion Transformers for Fluid Flow Simulations Learning
Towards a Theoretical Understanding of Two Tower Recommendation Models
Not All Problems Are Best Modeled as MILP: A DSL-Centric Framework for Flexible and Accur…
Characterizing the Quality Profile of AI-Generated C++ in Production
Optimizing Spectral Prediction in MXene-Based Metasurfaces Through Multi-Channel Spectral…
Taxonomy-Driven Analysis of Open-Source AI Risk Mitigation Tools
MaskFlow: Precise, Consistent and Seamless Regional Image Editing
Measurements Automatically Extracted from Zero Echo Time MRI Using Deep Learning Image Se…
Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotatio…
Aftab: A Comprehensive Benchmark of CNN Encoders and Advanced Value Functions in Parallel…
Interaction Creates Dynamical AI Behavior Absent in Isolation
Unsupervised Adaptation of PDE Foundation Models
How Much AI Is in This Track? Quantifying the Proportion of AI-Generated Stems in Hybrid …
Accounting Graph Transformer for Short-History Multi-KPI Forecasting in Small Businesses
TRACE: A Multi-Layer Benchmark for Human AI Controller Coordination Under Drift and Failu…
AutoIntervene: Calibrated Intervention for Action-Chunking Imitation Learning Policies
Alignment has a Fantasia Problem
Multi-Level Modeling of Large Language Model Inference Latency and Energy via Hybrid Anal…
Scalable High-Fidelity Macromolecular Docking for GPU-Accelerated Supercomputers
Automated item evaluation: Predicting item acceptance and rejection using LLM-generated c…
TOFD: Target-Oriented Feature Decoupling against Poisoning Attacks in Split Federated Lea…
Mind the Gap: A Dual Knowledge Graph Framework for Unified Multi-task User Intent Inferen…
Debias in Text, Believe Your Eyes: Text-Anchored Cross-Modal Transfer for Visual Counter-…
EliSeg: Verified Target Construction for Report-Grounded Abnormality Segmentation