GISAgentBench: A Practitioner-Sourced Benchmark for Evaluating LLM Agents on GIS Tasks
Browse
AI News
25887 items — filtered, classified, deduplicated
LongCat Sparse Attention: Taming the Lightning via Streaming-aware Hierarchical Cross-Lay…
RL-Lock: Reinforcement Learning for Generating Interlocking Assemblies
Latent Thought Credit: Multi-Answer Credit Assignment for Latent Reasoning
Temperature-driven inversion and nonlinear dynamics in ChatGPT-like AIs
KING: Embodiment-Aware Kinematic Graph Neural Network for Unified Motion Representation o…
FL-OA: A Byzantine-Robust Federated Learning Framework with Outsourced Auditing for Intel…
TabDPT-Turbo: Efficient In-Context Learning for Tabular Prediction
Sweet Little Lies: Strategic Deception in AI Emotional Support Chatbots
Emergence Invariance: From Symbolized Thought to Interface Refinement
Is More Privileged Information Better? From Solution Traces to Problem-Solving Structure …
Securing Agentic AI: From Per-Action Checks to Trajectory Assurance
Does the Competitive Component of Adversarial Self-Play Improve Legal Reasoning? A Contro…
Fairness Auditing: Lower Bounds on Company Manipulation
PolymerGPT: Multi-property Optimization with a Decoder-Based GPT Model for Generative Pol…
Loud or Silent? A Reusable Framework for Per-Modality Failure Analysis in Multimodal Clin…
Computing with Agentic Oracles
Training nGPT
KoVRE: Training an Efficient Embedding Model for Korean Visual Document Retrieval
Reusing Rollouts under Policy Lag: Prefix-Normalized Policy Optimization for LLM Reinforc…
CraftAlign: Feature-Grounded Evaluation and Revision Guidance for AI Stories
From Chasing Ghosts to Missed Attacks: Perspectives and Perceptions of SOC Practitioners …
EduPluginBench: Executable Assurance for AI-Generated Educational Plugins
Agentic Stage-One Stellarator Optimization: Autonomous Multi-Objective Search for Finite-…
G-ReAct: Graph-Guided Deep Search via Structure-State Co-Evolution
CRAFTS: Collaborative Role-Adaptive Fine-Tuning of LLM Agents for Chemical Process Simula…
Cognitive Demand Steering for Adaptive Meta-Reasoning in Large Language Models
CT-PrepAgent: Bounded Policy and Controlled Execution for Adaptive CT Data Preparation
MRAFnd: Multimodal Retrieval-Augmented Framework for Zero-Shot Fake News Detection
Leveraging AI for fine-grained food safety risk forecasting in sparse data conditions