Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web …
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
CityGen: Structure-Guided City-Style Synthesis for Cross-City Autonomous Driving
Internal Representation, Not Clinical Knowledge: Where Apparent LLM Triage Failures Origi…
Selection Hyper-heuristics Can Automatically Adjust the Learning Period to Optimally Solv…
Evolutionary Dynamics of Cooperation in Next-Generation LLM Agent Systems: A Cross-Provid…
Mitigating Hallucination in Vision-Language Models through Barrier-Regulated Adaptive Clo…
HARP: Hadamard-Preconditioned Adaptive Rotation Processor for Extreme LLM Quantization
Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems
Inferring Code Correctness from Specification
Multi-Legal-Bench: Evaluating LLMs on Legal Reasoning Across Jurisdictions, Languages, an…
Energy-Aware NECO for Single-Pass Pixel-wise Out-of-Distribution Detection in Semantic Se…
OccamToken: Efficient VLM Inference with Training-Free and Budget-Adaptive Token Pruning
EviLink: Multi-Path Schema Linking with Uncertainty-Guided Evidence Acquisition for Large…
Entity-Collision: A Stratified Protocol for Attributing Retrieval Lift in Agent Memory
The Sample Complexity of Multiclass and Sparse Contextual Bandits
Brain-IT-VQA: From Brain Signals to Answers
Temporal Motif-aware Graph Test-time Adaptation for OOD Blockchain Anomaly Detection
GUITestScape: Towards Open-set Evaluation on Exploratory GUI Testing
SCOPE: A Lightweight-training LLM Framework for Air Traffic Control Readback Monitoring
Quotient DAGs for Off-Policy Evaluation:Forward-Flow Importance Sampling and Exact Slate …
The New Pro Se: Generative AI and the Surge in Federal Civil Self-Representation
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles
SciIntBench: Measuring LLM Compliance with Research Integrity Norms Under Adversarial Fra…
Adaptive Interviewing for Persona Simulation in LLMs: Evidence-Grounded Reasoning Improve…
Forget Less, Generalize More: Unifying Temporal and Structural Adaptation for Dynamic Gra…
DELOS: Detecting Shallow Transits in Kepler Photometry Using a Contrastive-Learning Frame…
Beyond Bilingual Transfer: Multilingual Code-Switching in Instruction Tuning
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language …
Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabular…