Browse

AI News

30934 items — filtered, classified, deduplicated

arXiv cs.AI Research & Papers
ScientistOne: Towards Human-Level Autonomous Research via Chain-of-Evidence
arXiv cs.AI Research & Papers
JobBench: Aligning Agent Work With Human Will
arXiv cs.AI Research & Papers
Post-training makes large language models less human-like
arXiv cs.AI Research & Papers
Governed Metaprogramming for Intelligent Systems: Reclassifying Eval as a Governed Effect
arXiv cs.AI Research & Papers
The Strongest Teacher Is Not Always the Best Teacher: Student-Centric Answer Selection
arXiv cs.AI Research & Papers
Persistent AI Agents in Academic Research: A Single-Investigator Implementation Case Study
arXiv cs.AI Research & Papers
Innovation: An Almost Characterization of Hallucination
arXiv cs.AI Research & Papers
Symmetry-Compatible Principle for Optimizer Design: Embeddings, LM Heads, SwiGLU MLPs, an…
arXiv cs.AI Research & Papers
Traceable Knowledge Graph Reasoning Enables LLM-Assisted Decision Support for Industrial …
arXiv cs.AI Research & Papers
ICCU: In-Context Continual Unlearning via Pattern-Induced Refusal Rules
arXiv cs.AI Research & Papers
VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation
arXiv cs.AI Opinion & Editorial
Generative artificial intelligence and the marginalization of minoritized knowledges in h…
arXiv cs.AI Security & Safety
Tracing the Dynamics of Refusal: Exploiting Latent Refusal Trajectories for Robust Jailbr…
arXiv cs.AI Research & Papers
CyberEvolver: Structured Self-Evolution for Cybersecurity Agents On the Fly
arXiv cs.AI Research & Papers
Measuring Prediction Uncertainty in Neural Cellular Automata
arXiv cs.AI Research & Papers
Beyond Trajectory-Level Attribution: Graph-Based Credit Assignment for Agentic Reinforcem…
arXiv cs.AI Research & Papers
Rethinking the Trust Region in LLM Reinforcement Learning
arXiv cs.AI Research & Papers
RulePlanner: All-in-One Reinforcement Learner for Unifying Design Rules in 3D Floorplanni…
arXiv cs.AI Research & Papers
DynFrame: Adaptive Reasoning-Driven Multimodal Framework with Dynamic Frame Augmentation …
arXiv cs.AI Research & Papers
More Expressive Feedforward Layers: Part I. Token-Adaptive Mixing of Activations
arXiv cs.AI Research & Papers
Spend Your Rollouts Where It Counts: Rollout Allocation for Group-Based RL Post-Training
arXiv cs.AI Policy & Regulation
Examining the Challenges of Intellectual Property in AI-Generated Productions
arXiv cs.AI Research & Papers
SWAP: Towards Copyright Auditing of Soft Prompts via Sequential Watermarking
arXiv cs.AI Research & Papers
Monte Carlo Permutation Search
arXiv cs.AI Research & Papers
When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation
arXiv cs.AI Research & Papers
A Physics-Informed Hierarchical Neural Network for Microwave Scattering Analysis of 3D PE…
arXiv cs.AI Research & Papers
Adaptive Multi-prompt Contrastive Network for Few-shot Out-of-distribution Detection
arXiv cs.AI Research & Papers
Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interac…
arXiv cs.AI Research & Papers
Can LLMs Introspect? A Reality Check
arXiv cs.AI Research & Papers
Bridging Control with Neural Network Verifier alpha-beta-CROWN: A Tutorial