Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Security & Safety
Technical Report: Exploring the Emerging Threats of the Agent Skill Ecosystem
arXiv cs.AI Research & Papers
InfiMed-ORBIT: Aligning LLMs on Open-Ended Complex Tasks via Rubric-Based Incremental Tra…
arXiv cs.AI Research & Papers
LLM Watermark Evasion via Bias Inversion
arXiv cs.AI Research & Papers
Evaluating the Realism of LLM-powered Social Agents: A Case Study of Reactions to Spanish…
arXiv cs.AI Research & Papers
The Principles of Diffusion Models
arXiv cs.AI Research & Papers
EAGer: Entropy-Aware GEneRation for Adaptive Inference-Time Scaling
arXiv cs.AI Research & Papers
Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting
arXiv cs.AI Research & Papers
Measuring Form and Function in Language Models
arXiv cs.AI Security & Safety
Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking
arXiv cs.AI Research & Papers
The Attentional White Bear Effect in Transformer Language Models
arXiv cs.AI Research & Papers
Visualizing Latent Phase Structures in Locomotion Policies: A Multi-Environment Study wit…
arXiv cs.AI Research & Papers
Smaller, Younger, and More Impactful: How AI-Assisted Writing Transforms Research Teams
arXiv cs.AI Research & Papers
Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empiri…
arXiv cs.AI Research & Papers
A Fresh Look at Lamarckian Evolution and the Baldwin Effect
arXiv cs.AI Research & Papers
Beyond Interpretability: When, Why, and How Sparse Autoencoders Enable Label-Free Visual …
arXiv cs.AI Research & Papers
MMTABREAL: Real-World Benchmark for Multimodal Table Understanding
arXiv cs.AI Research & Papers
LiDDA: Data Driven Attribution at LinkedIn
arXiv cs.AI Research & Papers
MemTrace: Tracing and Attributing Errors in Large Language Model Memory Systems
arXiv cs.AI Research & Papers
BIRDNet: Mining and Encoding Boolean Implication Knowledge Graphs as Interpretable Deep N…
arXiv cs.AI Research & Papers
IPO-Mine: A Toolkit and Dataset for Section-Structured Analysis of Long, Multimodal IPO D…
arXiv cs.AI Research & Papers
The Point, the Vision and the Text: Does Point Cloud Boost Spatial Reasoning of Large Lan…
arXiv cs.AI Research & Papers
Rethinking Memory as Continuously Evolving Connectivity
arXiv cs.AI Research & Papers
MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
arXiv cs.AI Research & Papers
Skill-Conditioned Gated Self-Distillation for LLM Reasoning
arXiv cs.AI Research & Papers
OmniVerifier-M1: Multimodal Meta-Verifier with Explicit Structured Recalibration
arXiv cs.AI Models & Releases
Apple Intelligence Foundation Language Models
arXiv cs.AI Research & Papers
The Shape of Reasoning: Topological Analysis of Reasoning Traces in Large Language Models
arXiv cs.AI Research & Papers
A Comparative Study of Rule-Based and Data-Driven Approaches in Industrial Monitoring
arXiv cs.AI Research & Papers
RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs
arXiv cs.AI Research & Papers
CircuitLM: A Multi-Agent LLM-Aided Design Framework for Generating Circuit Schematics fro…