Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
A Survey on Recent Advances in Conversational Data Generation
arXiv cs.AI Research & Papers
A comparative study of transformer-based embeddings for topic coherence
arXiv cs.AI Research & Papers
Physics Is All You Need? A Case Study in Physicist-Supervised AI Development of Scientifi…
arXiv cs.AI Research & Papers
MIRA: Mid-training Rubric Anchoring for Source-Aware Data Selection
arXiv cs.AI Research & Papers
Persona Conditioning of Brand Recommendations in Retrieval-Augmented Commercial Chat: A P…
arXiv cs.AI Research & Papers
Meta-Cognitive Memory Policy Optimization for Long-Horizon LLM Agents
arXiv cs.AI Research & Papers
AgentSchool: An LLM-Powered Multi-Agent Simulation for Education
arXiv cs.AI Research & Papers
Enhancing Multi-Agent Communication through Attention Steering with Context Relevance
arXiv cs.AI Research & Papers
PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers
arXiv cs.AI Research & Papers
Conformal Certification of Reasoning Trace Prefixes
arXiv cs.AI Research & Papers
Reliable Reasoning with Large Language Models via Preference-Based Maximum Satisfiability
arXiv cs.AI Research & Papers
Robust and Generalizable Safety Steering for Text-to-Image Diffusion Transformers
arXiv cs.AI Research & Papers
From GPS Points to Travel Patterns: Flexible and Semantic Trajectory Generation with LLMs
arXiv cs.AI Research & Papers
Cookie-Bench: Continuous On-screen Key Interaction Evaluation for Web Generation
arXiv cs.AI Research & Papers
Compass: Navigating Global Marine Lead Data Integration through Expert-Guided LLM Agent
arXiv cs.AI Research & Papers
Meta-Programming for Linear-time Temporal Answer Set Programming
arXiv cs.AI Research & Papers
It`s All About Speed: AI`s Impact on Workflow in Music Production
arXiv cs.AI Research & Papers
Redundant or Necessary? A Benchmark for Detecting Redundant Steps in Agent Trajectories
arXiv cs.AI Research & Papers
TRACE: Toulmin-based Reasoning Assessment through Constructive Elements for LLM CoT Evalu…
arXiv cs.AI Research & Papers
OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster…
arXiv cs.AI Research & Papers
PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing
arXiv cs.AI Research & Papers
SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow
arXiv cs.AI Research & Papers
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Secu…
arXiv cs.AI Research & Papers
Model Fusion via Retrofitting
arXiv cs.AI Research & Papers
DynSess: Dynamic Session-Level Evaluation and Optimization Framework for Role-Playing Age…
arXiv cs.AI Research & Papers
BEAMS: Benchmarking and Evaluating AI for Modeling and Simulation
arXiv cs.AI Research & Papers
Differentiable Belief-based Opponent Shaping
arXiv cs.AI Research & Papers
Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete…
arXiv cs.AI Research & Papers
Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search…
arXiv cs.AI Research & Papers
Benchmarking Positional Encoding Strategies for Transformer-Based EEG Foundation Models