Explorar

Noticias de IA

30675 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
EvoSQL: Memory-Augmented Critic-Generator Co-Evolution for Text-to-SQL
arXiv cs.AI Research & Papers
Autonomous Topology Mutation: Safe Runtime Restructuring for Multi-Agent LLM Systems with…
arXiv cs.AI Research & Papers
The Devil is in the Spectrum: Mitigating Representation Collapse in LLMs via Topologicall…
arXiv cs.AI Research & Papers
Tractable Hierarchical Control of Autoregressive Language Models
arXiv cs.AI Research & Papers
PersonaTrail: Benchmarking Personalized Web Agents through Browsing Trails
arXiv cs.AI Research & Papers
Enabling Scalable Topology Inference in Distribution Systems via Constrained Multi-Source…
arXiv cs.AI Research & Papers
Beyond Liars' Bench: The Impact of Lie Typology, Depth, and Sparsity on Deception Detecti…
arXiv cs.AI Research & Papers
Semi-Supervised Text-Attributed Graph Distillation
arXiv cs.AI Research & Papers
Benchmarking the Personalization Capabilities of Large Language Models
arXiv cs.AI Security & Safety
Robust Critics: Defending LLMs Against Multi-Turn Attacks
arXiv cs.AI Research & Papers
PlanE: Meta Planning of Data, Tuning, and Inference for Extractive-based LLMs
arXiv cs.AI Research & Papers
DecodeShare: Tracing the Shared Subspace of LLM Decode-Time Decisions
arXiv cs.AI Research & Papers
InferenceBench: A Benchmark for Open-Ended LLM Inference Optimization by AI Agents
arXiv cs.AI Research & Papers
DC-Leap: Training-Free Acceleration of dLLMs via Draft-Guided Contiguous Leaping Decoding
arXiv cs.AI Research & Papers
JAXBench: Benchmarking Autonomous TPU Kernel Optimization
arXiv cs.AI Research & Papers
Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts
arXiv cs.AI Research & Papers
AINTMA: Agentic AI Architecture for Autonomous Test Management with Generative Intelligen…
arXiv cs.AI Research & Papers
RUMBA: Russian User Memory Benchmark
arXiv cs.AI Research & Papers
Student-Centered Distillation Narrows the Agentic Gap Between Small and Large LLMs
arXiv cs.AI Research & Papers
Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-su…
arXiv cs.AI Research & Papers
Refusal-Gated Decoding: Preserving Refusal Behavior Under High-Temperature Sampling
arXiv cs.AI Research & Papers
OPOD: On-Policy Omni Distillation
arXiv cs.AI Research & Papers
Detecting LLM-Generated Tokens in Human--LLM Coauthored Text
arXiv cs.AI Research & Papers
DynaMark: A Reinforcement Learning Framework for Dynamic Watermarking in Industrial Machi…
arXiv cs.AI Research & Papers
Drive As You Like: Multi-Head Diffusion with Reinforcement Learning for Personalized Driv…
arXiv cs.AI Research & Papers
Loss-Complexity Landscape and Model Structure Functions
arXiv cs.AI Research & Papers
Generative AI and Agency in Education: A Critical Scoping Review and Thematic Analysis
arXiv cs.AI Research & Papers
Representative Sets in Propositional Abduction
arXiv cs.AI Research & Papers
Synthetic minority data is redundant or invalid: a data-dependent validity theory and a d…
arXiv cs.AI Research & Papers
Autonomous disproofs of the sum-product conjecture over $\mathbb R$ with GPT-5.5 Pro