StakeBench: Evaluating Language Understanding Grounded in Market Commitment
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
When Gradients Collide: Failure Modes of Multi-Objective Prompt Optimization for LLM Judg…
Aes3D: Aesthetic Assessment in 3D Gaussian Splatting
Efficient Preference Poisoning Attack on Offline RLHF
Leveraging Spreading Activation for Improved Document Retrieval in Knowledge-Graph-Based …
Chain-of-Thought Hijacking
Optimizing Sensor Placement for Flow Reconstruction in Urban Drainage Networks: A Digital…
BC Protocol: Structured Dual-Expert Dialogue for Eliciting High-Quality Chain-of-Thought …
PennySynth: RAG-Driven Data Synthesis for Automated Quantum Code Generation
A Controlled Synthetic Benchmark for Educational Aspect-Based Sentiment Analysis
A Multi-Agent LLM Framework for Rating the Quality of Surgical Feedback
AI Content Moderation in Therapy Conversations
Autoregression-Free Neural Operators for Time-Dependent PDEs
A Token/KV-Cache Communication Media Selection and Resource Allocation Strategy for Multi…
Weakly Supervised Camouflaged Object Detection Based on the SAM Model and Mask Guidance
Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoiet…
Performance Comparison of Classical and Neural Sampling Algorithms for Robotic Navigation
D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation
Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion
Scaling up Energy-Aware Multi-Agent Reinforcement Learning for Mission-Oriented Drone Net…
Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulne…
OSDTW: Optimal Shared Depth and Task Weighting for Long-Tailed Recognition
RealBench: Benchmarking Data-Driven Numerical Weather Forecasting Under Operational Condi…
On the Impact of Class Imbalance on the Learning Dynamics of Deep Neural Networks:An Intu…
Explainable Multi-Task Retinal Imaging Reveals Microvascular Signals for Systemic Risk St…
DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection
How Many Tools Should an LLM Agent See? A Chance-Corrected Answer
Beyond Generative Priors: Minority Sampling with JEPA-Guided Diffusion
Correcting Visual Blur Induced by Attention Distraction to Reduce Hallucinations: Algorit…
Catching MRI outliers: unsupervised detection and localization of MRI artefacts and clini…