From Prompt Injection to Persistent Control: Defending Agentic Harness Against Trojan Bac…
Explorar
Noticias de IA
29670 elementos — filtrados, clasificados y sin duplicados
STEP: Learning STructured Embeddings for Progressive Time Series
LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis
KnowledgeGain: Evaluating and Optimizing Science News Generation for Reader Learning
A Unified Framework for Gradient Aggregation in Multi-Objective Optimization
Graph-Conditioned Mixture of Graph Neural Network Experts for Traffic Forecasting
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
Active Timepoint Selection for Learning Measure-Valued Trajectories
Score Broadcast and Decorrelation: A General Framework for Broadcast-Based Credit Assignm…
COFT: Counterfactual-Conformal Decoding for Fair Chain-of-Thought Reasoning in Large Lang…
HypoAgent: An Agentic Framework for Interactive Abductive Hypothesis Generation over Know…
Developing an AI-Powered UX Research Point of View for Digital Health in A Regulatory Con…
Rethinking Multimodal Few-Shot 3D Point Cloud Segmentation: From Fused Refinement to Deco…
Performance and Complexity Trade-off Optimization of Speech Models During Training
LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric R…
PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Traini…
GPU Forecasters: Language Models as Selective Surrogates for Kernel Runtime Optimization
Inferring Events from Time Series using Language Models
The Gaussian-Head OFL Family: One-Shot Federated Learning from Client Global Statistics
Chain-of-Thought Reasoning In The Wild Is Not Always Faithful
Inverting Data Transformations via Diffusion Sampling
SCOPE: Selective Conformal Optimized Pairwise LLM Judging
PAC-Bayesian Reinforcement Learning Trains Generalizable Policies
Vision-Language Models Suppress Female Representations Under Ambiguous Input
DynaTree: Dynamic Agentic Retrieval Tree for Time-Sensitive News Retrieval
FBHM: Functional Benchmarking and Steering of VLMs for Hateful Meme Detection
Social welfare optimisation under institutional reward and punishment
Latent Space Disentanglement via Activation Steering for Interpretable Attribute Control …
Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs
Simulation of collision avoidance behavior in crowd movement by data-driven approach