Browse

AI News

26304 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
TrajWiki: Source-Grounded Memory Trajectories for Long-Horizon Dialogue Agents
arXiv cs.AI Research & Papers
SearchMaster: Grounded and Regulated Self-Play for Search Agents
arXiv cs.AI Research & Papers
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning
arXiv cs.AI Security & Safety
Exposed by Design: A Dynamic Security Assessment of Internet-Facing MCP Servers at Scale
arXiv cs.AI Research & Papers
Ranking Image Fusion the Way Humans Do: A Learned Pairwise Preference Metric for Infrared…
arXiv cs.AI Research & Papers
Automatic Annotation of Ancient Greek Vowel Length
arXiv cs.AI Research & Papers
MemArbiter: Decision-Time Memory Arbitration for Long-Horizon LLM Agents
arXiv cs.AI Research & Papers
When Memory Updates but Behavior Does Not: Repairing Implicit Stale Dependencies in Perso…
arXiv cs.AI Research & Papers
Deep Learning CNN and Recurrence Analysis for Alpha Gamma EEG Biomarkers in Fragile X Syn…
arXiv cs.AI Research & Papers
Long-Horizon Autonomous Architecture Research with a Language-Model Agent: A Behavioural …
arXiv cs.AI Research & Papers
Deep Agentic Search for Repository-Level Code Question Answering: An Empirical Study
arXiv cs.AI Research & Papers
A Fortran General-Purpose Transpiler: Proof of Concept
arXiv cs.AI Research & Papers
AgentSLABench: Evaluating and Benchmarking Agentic Systems Under Resource Constraints
arXiv cs.AI Research & Papers
AtumAI: A Principled Framework for Agentic Generation of Datacenter Control-Plane Policies
arXiv cs.AI Research & Papers
Beyond the Mean: Multi-Moment Policy Optimization for LLM Reasoning
arXiv cs.AI Research & Papers
TextNCA: Neural Cellular Automata for Language Modeling via Hierarchical Local Attention
arXiv cs.AI Research & Papers
G-ReAct: Graph-Guided Deep Search via Structure-State Co-Evolution
arXiv cs.AI Research & Papers
Request-Level Energy Attribution for Batched LLM Serving
arXiv cs.AI Security & Safety
Robust Watermarks Meet Backdoored Models: Evading Diffusion Semantic Watermarks via Steal…
arXiv cs.AI Security & Safety
When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems
arXiv cs.AI Research & Papers
Passing Coarse Marginal Checks Can Be Cheap: Persona Mixtures and Imprecise Treatment-Res…
arXiv cs.AI Research & Papers
AgentMemBench: A Systematic Benchmark for Evaluating Long-Term Memory Management Strategi…
arXiv cs.AI Research & Papers
When LLM Essays Outscore Student Essays: What a Korean Writing Rubric Rewards and Where R…
arXiv cs.AI Research & Papers
MAPLE: Metadata Augmented Private Language Evolution
arXiv cs.AI Research & Papers
Cross-Task Dissociation in Frontier Vision-Language Model Theory of Mind
arXiv cs.AI Research & Papers
Hierarchical Pre-Training of Vision Encoders with Large Language Model
arXiv cs.AI Security & Safety
Salami Attack: Stealthy Collusive Memory Poisoning against OpenClaw
arXiv cs.AI Research & Papers
Physics-Informed Neural Networks for Complex Eigenfrequency Identification and Mode Struc…
arXiv cs.AI Research & Papers
CRAFTS: Collaborative Role-Adaptive Fine-Tuning of LLM Agents for Chemical Process Simula…
arXiv cs.AI Research & Papers
DGA$_2$D: Directed Graph-Guided Automated Algorithm Design with Large Language Models