Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Coding Agent Is Good As World Simulator
arXiv cs.AI Research & Papers
Agricultural Landscape Understanding At Country-Scale
arXiv cs.AI Research & Papers
PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning
arXiv cs.AI Research & Papers
CodeCytos: AI-assisted spatial molecular imaging analysis via code-augmented agent action…
arXiv cs.AI Research & Papers
KnowledgeBerg: Evaluating Systematic Knowledge Coverage and Compositional Reasoning in La…
arXiv cs.AI Research & Papers
RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography
arXiv cs.AI Research & Papers
Interpretable Policy Distillation for Power Grid Topology Control
arXiv cs.AI Research & Papers
Improving Visual Representation Alignment Generation with GRPO
arXiv cs.AI Research & Papers
SPADER: Step-wise Peer Advantage with Diversity-Aware Exploration Rewards for Multi-Answe…
arXiv cs.AI Research & Papers
Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory
arXiv cs.AI Research & Papers
Model Parallelism With Subnetwork Data Parallelism
arXiv cs.AI Research & Papers
LinguIUTics at PsyDefDetect: Iterative Imbalance-Aware Fine-tuning of Qwen3-8B for Psycho…
arXiv cs.AI Research & Papers
Demystifying the Optimal Fair Classifier in Multi-Class Classification
arXiv cs.AI Research & Papers
The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Sh…
arXiv cs.AI Research & Papers
On Information Self-Locking in Reinforcement Learning for Active Reasoning of LLM agents
arXiv cs.AI Research & Papers
LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs?
arXiv cs.AI Research & Papers
Information-Theoretic Lower Bounds for Bit-Constrained Stochastic Optimization via a Redu…
arXiv cs.AI Research & Papers
SORA: Free Second-Order Attacks in Fast Adversarial Training
arXiv cs.AI Research & Papers
Behavior-Invariant Task Representation Learning with Transformer-based World Models for O…
arXiv cs.AI Research & Papers
Extending Causal Metamodeling to a non-Markovian Queue
arXiv cs.AI Research & Papers
MulFeRL: Enhancing Reinforcement Learning with Verbal Feedback in a Multi-turn Loop
arXiv cs.AI Research & Papers
Safety Alignment of LMs via Non-cooperative Games
arXiv cs.AI Research & Papers
MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents
arXiv cs.AI Research & Papers
A Unified Evaluation-Instructed Framework for Query-Dependent Prompt Optimization
arXiv cs.AI Research & Papers
ACON: Optimizing Context Compression for Long-horizon LLM Agents
arXiv cs.AI Research & Papers
Benchmarks for Vision-Language Models in Urban Perception Should Be Reliability-Aware and…
arXiv cs.AI Research & Papers
Dive into Waves: Morlet Spectral Transformer for Cross-Subject Emotion Decoding from EEG
arXiv cs.AI Security & Safety
Benchmarking Security Risk Detection and Verification in Open Agentic Skill Ecosystems
arXiv cs.AI Research & Papers
Taming System Complexity: Demystifying Software Engineering Agents in Diagnosing Linux Ke…
arXiv cs.AI Research & Papers
Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-Tuning and Can Be Miti…