Do Latent Channels Actually Communicate? A Causal Audit of Latent Multi-Agent LLM
Explorar
Noticias de IA
30334 elementos — filtrados, clasificados y sin duplicados
FedTopo: Relation-Level Topology Sharing for Model-Heterogeneous Federated Learning
Cost-Sensitive Conformal Prediction and Human-in-the-Loop Abstention for Imbalanced High-…
Recover, Decode, Reguard: Guard-Agnostic Defense Amplification againstEncoded VLM Jailbre…
A Methodology for Designing Knowledge-Driven Missions for Robots
Partner Capability Estimation for Task-Agnostic Adaptation in Ad-Hoc Teamwork
IDP AutoOpt: Agent-Driven Optimization of Document Processing Pipeline Configurations
ServerlessT2I: Efficient Text-to-Image Workflow Serving on a Serverless Platform
From Passive Video to Editable Experience: Physically Grounded Experience Synthesis for E…
Predict before you train: Scaling Laws for particle physics foundation models
MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Suppor…
Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adap…
TRACE-CTI: Auditable Post-Extraction Governance of TTP Claims with Knowledge Graphs
MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning
Property-driven Causal Abstractions for Markov Decision Processes
Position: Evaluation Scores Are Perishable Knowledge Claims
CaM-Wolf: Causal-Aware Multimodal Agents for Social Deduction Games
GoGoTB: Agentic RTL Verification with Specification-Grounded Coverage Closure
FARI: Robust One-Step Inversion for Watermarking in Diffusion Models
PatchDenoiser: Parameter-efficient multi-scale patch learning and fusion denoiser for Low…
Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak Defen…
Evidence-Ledger Adjudication for Claim-Evidence Traceability
FakeIDet3-DB: Refining Digital Attacks and Patch Extraction for Secure ID Benchmarking
GuideSkill: Evolving Executable LLM Agent Skills for Guideline-Grounded Clinical Reasoning
ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Sc…
APEX-Accounting
When benchmark inferences do not compose: Projectibility in AI evaluation
ReCo: Reweighting GRPO Against Distributional Concentration
Living-Harness Is an Interactive-Agent Evolver
Improving Item Discoverability in e-Commerce Search via Related Intent Generation