Agent Learning via Early Experience
Browse
AI News
30934 items — filtered, classified, deduplicated
False Fixed Points: Kantian Feedback, Stable Miscalibration, and Representational Compres…
Voting with the Graph: Stable RLAIF via Topological Consistency Maximization
Efficient and Scalable Neural Symbolic Search for Knowledge Graph Complex Query Answering
FloorplanQA: A Benchmark for Spatial Reasoning in LLMs using Structured Representations
Hide-and-Shill: A Reinforcement Learning Framework for Market Manipulation Detection in S…
From Multi-Agent Systems and the Semantic Web to Agentic AI: A Unified Narrative of the W…
Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning
Squeezing Capacity from Multimodal Large Language Models for Subject-driven Generation
Beyond Summaries: Structure-Aware Labeling of Code Changes with Large Language Models
Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark
OrpQuant: Geometric Orthogonal Residual Projection for Multiplier-Free Power-of-Two Trans…
DRScaffold: Boosting Dense-Scene Reasoning in Lightweight Vision Language Models
Confidence and Calibration of Activation Oracles for Reliable Interpretation of Language …
AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models
Retrieval-Augmented Detection of Potentially Abusive Clauses in Chilean Terms of Service
A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot…
Step-TP: A Grounded, Step-Level Dataset with Chain-of-Thought Reasoning for LLM-Guided Te…
From Latent Space to Training Data: Explainable Specialization in Minimal MLPs
Explaining Too Much? Understanding How Large Language Model Reasoning Traces Influence Pe…
TTPrint: Evidence-Grounded TTP Extraction via Diverge-then-Converge Verification
Geometric Evolution Maps: Extracting Stable Concept Probes from Transformer Residual Stre…
Adaptive Graph Refinement and Label Propagation with LLMs for Cost-Effective Entity Resol…
NPSolver: Neural Poisson Solver with Iterative Physics Supervision
How Should LLMs Consume High-Quality Data? Optimal Data Scheduling via Quality-Aware Func…
DeGRe: Dense-supervised Generative Reranking for Recommendation
Benchmarking Pathology Foundation Models for Spatial Domain Understanding
Don't Retrain, Just Reuse: Recovering Dual-Target Molecules from Single-Target Diffusion …
Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversari…
Fine-Tuning and Serving Gemma 4 31B on Google Cloud TPU: A Technical Comparison with GPU …