TerraNova: A Foundation Model for the Anthropocene
Explorar
Noticias de IA
30294 elementos — filtrados, clasificados y sin duplicados
ARB: A Matched Authorship-Rewriting Benchmark Dataset for AI-Text Detector Evaluation
TraceViT: Grounded Trace Supervision for Visual Abstract Reasoning
SERUM: State Extraction and Refinement for User Modeling
Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation
The Theoretical Foundation of Socratic Tests: Dynamic, Multimodal, Conversational Examina…
Federated Foundation Models Fine-Tuning with Heterogeneous Compressed Clients
GeoRA: Geometry-Aware Low-Rank Adaptation for RLVR
Harnessing the Wisdom of LLM Crowds through Complementarity-Driven Iterative Collaboration
Knowledge Restoration-driven Prompt Optimization: Unlocking LLM Potential for Open-Domain…
WebCoderBench: Benchmarking Web Application Generation with Comprehensive and Interpretab…
Towards the Holographic Characteristic of LLMs for Efficient Short-text Generation
AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-La…
Evidence-Grounded Constraint Checking in Construction Documents
Learning Lookahead Lemmas for Neural Network Verification
Improving scDiffusion with Sparsity-Biased Classifier-Free Guidance
Auto-JEPA: A Latent World Model of Continuous Intent for End-to-End Autonomous Driving
Adjudicated Captioning: Multi-Agent Alignment Scoring and Consensus-Distilled Beam Arbitr…
MMShopBench: A Real-Log Benchmark for Multimodal, Multi-Turn Shopping Agents
MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations
What Makes a Sale? Simulating End-to-End Seller--Buyer Retail Dynamics with LLM Agents
Retrieval-Driven Training-Free AI-Generated Video Attribution
DiffAttack: Evasion Attacks Against Face Recognition via Latent Diffusion Models
Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowle…
Step-Level Visual Grounding Faithfulness Predicts Out-of-Distribution Generalization in L…
Wrong Code, Right Structure: Learning Netlist Representations from Imperfect LLM-Generate…
Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Genera…
MAGA: Multi-Platform Self-Fusion of GUI Agents via Structured Action Distillation
StaQ: a Finite Memory Approach to Discrete Action Policy Mirror Descent
Agentic Harness for Real-World Compilers