Learning to Reason with Insight for Informal Theorem Proving
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy
From Weak Cues to Real Identities: Evaluating Inference-Driven De-Anonymization in LLM Ag…
From Out-of-Distribution Detection to Hallucination Detection: A Geometric View
Discovering Differences in Strategic Behavior Between Humans and LLMs
Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory
ConSensus: Multi-Agent Collaboration for Multimodal Sensing
DTop-p MoE: Sparsity-Controlled Dynamic Top-p MoE for Foundation Model Pre-training
HERMES: Towards Efficient and Verifiable Mathematical Reasoning in LLMs
Post-Training LLMs as Better Decision-Making Agents: A Regret-Minimization Approach
OLG++: A Semantic Extension of Obligation Logic Graph
ProofWala: A Framework for Multilingual Proof Data Synthesis and Theorem-Proving
Unifying and Optimizing Data Values for Selection via Sequential Decision-Making
Stateful Online Monitoring Catches Distributed Agent Attacks
TunerDiT: Training-free Progressive Steering of Diffusion Transformer for Multi-Event Vid…
Positional versus Symbolic Attention Heads: Learning Dynamics, RoPE Geometry, and Length …
RayDer: Scalable Self-Supervised Novel View Synthesis from Real-World Video
Separating Secrets from Placeholders: A Hybrid CNN-CodeBERT Framework for Three-Class Cre…
Controllable Lung Nodule Synthesis via Histogram-Regularized Latent Diffusion Models
On Efficient Scaling of GNNs via IO-Aware Layers Implementations
Fine-grained Verification via Diagnostic Reasoning Supervision for Aspect Sentiment Tripl…
Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Info…
DOA: Training-Free Decoder-Only Attention Policy for Long-Form Simultaneous Translation w…
The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of La…
Target-Side Paraphrase Augmentation for Sign Language Translation with Large Language Mod…
Scaling Higher-Order Graph Learning with Maximal Clique Complexes
dashi: A Python library for Dataset Shift Characterization to Support Trustworthy AI Deve…
Appropriateness of Empathy in AI: A Signal-Cost Perspective
Inconsistency-Aware Minimization: Improving Generalization with Unlabeled Data
Neither Replacement nor Panacea: Comparing LLM-Based Conversational and Graphical Decisio…