Explorar

Noticias de IA

29396 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Security & Safety
Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software
arXiv cs.AI Research & Papers
Endogenous Resistance to Activation Steering in Language Models
arXiv cs.AI Research & Papers
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
arXiv cs.AI Research & Papers
Position: A Dynamical Systems Perspective is Needed to Advance Time Series Modeling
arXiv cs.AI Models & Releases
TokaMind: A Multi-Modal Transformer Foundation Model for Tokamak Plasma Dynamics
arXiv cs.AI Research & Papers
Forecasting as Rendering: A 2D Gaussian Splatting Framework for Time Series Forecasting
arXiv cs.AI Research & Papers
Training for Technology: Adoption and Productive Use of Generative AI in Legal Analysis
arXiv cs.AI Research & Papers
Benchmarking Language Modeling for Lossless Compression of Full-Fidelity Audio
arXiv cs.AI Research & Papers
EvoClaw: Evaluating AI Agents on Continuous Software Evolution
arXiv cs.AI Research & Papers
Chameleon: Control-Indexed Prospective Memory for Visuomotor Manipulation
arXiv cs.AI Research & Papers
Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry
arXiv cs.AI Research & Papers
CountsDiff: A Diffusion Model on the Natural Numbers for Generation and Imputation of Cou…
arXiv cs.AI Research & Papers
NTILC: Neural Tool Invocation via Learned Compression
arXiv cs.AI Research & Papers
WAV: Multi-Resolution Block Residual Routing for Deep Decoder-Only Transformers
arXiv cs.AI Research & Papers
SW-$A^2$-Bench: Benchmarking Autonomous Software Agent Generation for Agentic Web
arXiv cs.AI Research & Papers
MacArena: Benchmarking Computer Use Agents on an Online macOS Environment
arXiv cs.AI Research & Papers
MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Ret…
arXiv cs.AI Research & Papers
More Capable, Less Cooperative? When LLMs Fail At Zero-Cost Collaboration
arXiv cs.AI Research & Papers
Multi-Scale Feature Attention Network for Polymer Classification using THz Dual-Comb Spec…
arXiv cs.AI Security & Safety
RAVEN: Retrieval-Augmented Vulnerability Exploration Network for Memory Corruption Analys…
arXiv cs.AI Research & Papers
InvEvolve: Evolving White-Box Inventory Policies via Large Language Models with Performan…
arXiv cs.AI Research & Papers
Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval
arXiv cs.AI Developer Tooling
Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Ag…
arXiv cs.AI Research & Papers
FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantizat…
arXiv cs.AI Research & Papers
Coordinated optimization of departure sequencing and section-track allocation in railway …
arXiv cs.AI Research & Papers
CrowdMath: A Dataset of Crowdsourced Mathematical Research Discussions
arXiv cs.AI Research & Papers
CHoE: Cross-Domain Heterogeneous Graph Prompt Learning via Structure-Conditioned Experts
arXiv cs.AI Developer Tooling
Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review
arXiv cs.AI Research & Papers
Never Seen Before: Benchmarking Genuine Zero-Shot Composed Image Retrieval with Consisten…
arXiv cs.AI Research & Papers
DataEvolver: Automatic Data Preparation for Large Language Models through Multi-Level Sel…