LivingArena: Do LLMs Know What Other LLMs Don't? Peer-Probing as Scalable Evaluation
Explorar
Noticias de IA
30353 elementos — filtrados, clasificados y sin duplicados
HOBA: Hierarchical On-Policy Bidding Agents for Adaptive Online Advertising
Dual-Domain Manifold Modeling for Hyperspectral Image Fusion
OPERA: Offline Policy-guided Expert Routing and Adaptation for Universal Biomedical Image…
Spectral Truncation in Synthetic Control
Physics-Grounded Fluid Video Generation with a Simulation Dataset and Dual-Stream Optical…
Lantern: Conflict-Aware Gradient Blending for Physics-Guided Diffusion Models in Calorime…
Authoring Agent Skills: A Software-Engineering Approach
Steering topology distributions for unified generative design of architected metamaterials
Latent Stability Analysis of Malware Representations Under Feature-Space Perturbations
LLM as Forecasting Planner: Training-Free Text Conditioning for Time-Series Foundation Mo…
Calibrated Partial Resets: Preventing Policy Collapse in Continual Reinforcement Learning
GAUGE: Grading Agent-Built Financial Models Without a Golden Answer
Atmospheric Diffusion-Guided Spatio-Temporal Transformer for Nuclear Radiation Forecasting
Right-sizing Recommendations (RSR): Cloud Workload Conformal Prediction for Virtual Machi…
GraphRareBench: An Auditable Graph-Evidence Benchmark for Phenotype-Driven Rare-Disease D…
Beyond Epistemia: Epistemic Schizologia and Large Language Models as Techno-Semiotic Mach…
ContractHIL-HLS: Contract-Aligned Multi-Agent Workflow with Hardware-in-the-Loop Feedback…
Many-body Tipping Dynamics of ChatGPT-like AIs
RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Eval…
ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Lo…
Human Preference aligned Tabular Similarity
CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG F…
Tokens are All You Need: Dual-purpose Semantic IDs for Achieving LLM-Level I/O Efficiency…
LLM Scheming Inversely Scales with Pretraining Language Coverage
A Cost-Effective Multimodal LLM Reasoning Framework for Question Answering over Irregular…
How Small Can You Go? A Controlled Study of LoRA Rank, Target Modules, and Quantization T…
Matrix-Free Photoacoustic Image Reconstruction via Sensor-Token Self-Attention
Untrusted Authors, Trusted Answers: A Calculus of Fidelity-Graded Translations
HVM-GraphRAG: Holistic-View Multimodal Graph Retrieval-Augmented Generation on Complex Do…