Meta-Programming for Linear-time Temporal Answer Set Programming
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
It`s All About Speed: AI`s Impact on Workflow in Music Production
Redundant or Necessary? A Benchmark for Detecting Redundant Steps in Agent Trajectories
TRACE: Toulmin-based Reasoning Assessment through Constructive Elements for LLM CoT Evalu…
OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster…
PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing
SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Secu…
DynSess: Dynamic Session-Level Evaluation and Optimization Framework for Role-Playing Age…
BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation
Differentiable Belief-based Opponent Shaping
Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete…
Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search…
Benchmarking Positional Encoding Strategies for Transformer-Based EEG Foundation Models
Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Arti…
TIMEGATE: Sustainable Time-Boxed Promotion Gates for Continual ML Adaptation Under Resour…
FinVerBench: Benchmark Validity and Calibration in Large Language Model Financial Stateme…
Sustainable Metal-Organic Framework Water Harvesters in the Artificial Intelligence Era
UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents
The Chain Holds, the Answer Folds: Trace-Answer Dissociation in Reasoning Models Under Ad…
GRASP: Gated Regression-Aware Skill Proposer for Self-Improving LLM Agents
Think Fast, Talk Smart: Partitioning Deterministic and Neural Computation for Structured …
From Prompts to Context: An Ontology-Driven Framework for Human-Generative AI Collaborati…
DLM-SWAI: Steering Diffusion Language Models Before They Unmask
VikingMem: A Memory Base Management System for Stateful LLM-based Applications
Beyond Attack Success Rate: Temporal Logit Observability for LLM Safety Failures
Rethinking Literature Search Evaluation: Deep Research Helps, and Human Citation Lists Ar…
ReasonOps: Operator Segmentation for LLM Reasoning Traces
Rooted Absorbed Prefix Trajectory Balance with Submodular Replay for GFlowNet Training
Governing Technical Debt in Agentic AI Systems