Agentic Real2Sim: Physics-based World Modeling with Vision-Language Agents
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
MIRA-Ev:A Benchmark for Granular Evidence Detection and Relational Reasoning in Clinical …
Inference-Time Steering for Cross-Lingual Factual Consistency in LLMs
Prompt Design at Scale: How Format, Instruction Count, and Context Length Shape Instructi…
Toward Auditable Fraud Detection: Combining Graph Features, Model Explanations, and Agent…
PathAgentBench: Benchmarking Evidence-Seeking Vision-Language Models on Whole-Slide Patho…
They'll Verify. They Just Won't Act. How Authority Framing and Laundered Code Turn a Trus…
Planning as Emergent Behavior in Reinforcement Learning with Relational Hidden States
AutoIndex: Learning Representation Programs for Retrieval
Automated Data Engineering and Feature Selection for the Case Study of Warpage Detection …
From Trajectories to Instructions: Language-Conditioned Meta-Reinforcement Learning
Bounding Boxes to Improve Small Language Model Performance on Vision-Based Grading Tasks
Mitigating Matthew Effect: Multi-Hypergraph Boosted Multi-Interest Self-Supervised Learni…
LatentMT: Machine Translation with Latent Reasoning
What the Waveform Knows: Transparent-first Speech and Audio Intelligence with Caption Stu…
RRPO: Reference-Relative Policy Optimization with Stratified Conditional Rollouts
Competitive and Complementary Tools
ABOPD: Antibody CDR Design via On-Policy Distillation
ChainMark: Model-Free LLM Watermarking with Closed-Form Calibration
Regime-Aware Physics-Guided Early Warning of Lithium-Ion Battery Thermal Runaway Using Th…
Relay-Bench: Evaluating LLMs on Multi-Domain Reasoning Chains
RAMP: Recognition parametrisation by Amortised Message Passing
Multi-layer MIMO Relay as Deep Physical Neural Networks: Power Amplifiers as Activation F…
MambaLSTM: A Spatio-Temporal Framework for Enhanced Traffic Accident Risk Prediction
Circuit Claims Depend on What Is Extracted and How It Is Compared
AutoJourn: Multi-Perspective Summarisation, Bias Detection and Bias Neutralisation for LL…
An Analysis of Residual-Stream Geometry Across Transformer Depth
Addressing Limited Data in Auditory Attention Decoding with Diffusion Generative Models
Computational Humor with Multimodal LLMs: Methods, Datasets, Evaluation, and Challenges