Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets
arXiv cs.AI Research & Papers
Verified Misguidance: Measuring Structural Citation Failures in Search-Augmented LLMs
arXiv cs.AI Security & Safety
Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem
arXiv cs.AI Research & Papers
Evaluating the Realism of LLM-powered Social Agents: A Case Study of Reactions to Spanish…
arXiv cs.AI Research & Papers
Measuring Form and Function in Language Models
arXiv cs.AI Security & Safety
Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking
arXiv cs.AI Research & Papers
The Attentional White Bear Effect in Transformer Language Models
arXiv cs.AI Research & Papers
A Fresh Look at Lamarckian Evolution and the Baldwin Effect
arXiv cs.AI Research & Papers
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
arXiv cs.AI Research & Papers
BIRDNet: Mining and Encoding Boolean Implication Knowledge Graphs as Interpretable Deep N…
arXiv cs.AI Research & Papers
IPO-Mine: A Toolkit and Dataset for Section-Structured Analysis of Long, Multimodal IPO D…
arXiv cs.AI Research & Papers
Rethinking Memory as Continuously Evolving Connectivity
arXiv cs.AI Research & Papers
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
arXiv cs.AI Research & Papers
OmniVerifier-M1: Multimodal Meta-Verifier with Explicit Structured Recalibration
arXiv cs.AI Models & Releases
Apple Intelligence Foundation Language Models
arXiv cs.AI Research & Papers
The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models
arXiv cs.AI Research & Papers
A Comparative Study of Rule-Based and Data-Driven Approaches in Industrial Monitoring
arXiv cs.AI Research & Papers
RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs
arXiv cs.AI Research & Papers
CircuitLM: A Multi-Agent LLM-Aided Design Framework for Generating Circuit Schematics fro…
arXiv cs.AI Research & Papers
Can I Have Your Order? Monte-Carlo Tree Search for Slot Filling Ordering in Diffusion Lan…
arXiv cs.AI Research & Papers
COOP$^2$: Defining, Observing, and Repairing Cooperation in LLM Multi-Agent Systems
arXiv cs.AI Research & Papers
FactReview: Evidence-Grounded Peer Review with Execution-Based Claim Verification
arXiv cs.AI Research & Papers
Graph-of-Skills: Dependency-Aware Structural Retrieval for Massive Agent Skills
arXiv cs.AI Research & Papers
DSSE: a drone swarm search environment
arXiv cs.AI Research & Papers
Escher-Loop: Mutual Evolution by Closed-Loop Self-Referential Optimization
arXiv cs.AI Research & Papers
Verifiable Process Rewards for Agentic Reasoning
arXiv cs.AI Research & Papers
Delay-Aware Reinforcement Learning for Highway On-Ramp Merging under Stochastic Communica…
arXiv cs.AI Research & Papers
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design
arXiv cs.AI Research & Papers
FLUID: From Ephemeral IDs to Multimodal Semantic Codes for Industrial-Scale Livestreaming…
arXiv cs.AI Research & Papers
Revisiting Graph Autoencoders as Implicit Contrastive Learners