Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk
arXiv cs.AI Research & Papers
GUIDE: Governed Unified Intelligence for Document-to-Artifact Generation in Enterprise Se…
arXiv cs.AI Research & Papers
How to Spend Your Oracle Budget: Practical Guidance for Protein Structure Prediction Mode…
arXiv cs.AI Research & Papers
Retry, Switch, or Abstain? Learning Strategy-Aware Tool-Use Policies via Controlled Error…
arXiv cs.AI Security & Safety
Evaluating LLM Generated Detection Rules in Cybersecurity
arXiv cs.AI Research & Papers
ExRole: From Team Trajectories to Executable Roles in Multi-Agent Language Models
arXiv cs.AI Research & Papers
The Sleeping Agent: What Gist-Based Context Compression Loses and Why
arXiv cs.AI Research & Papers
Towards Query-Agnostic RAG Evaluation via Query Coverage and Claim Verifiability
arXiv cs.AI Research & Papers
BEST-KAG: Enhancing Question Answering of Building Engineering Standards with Multimodal …
arXiv cs.AI Research & Papers
Glance, Scrutinize, and Think: Advancing Video Anomaly Detection from Training-Free to Ag…
arXiv cs.AI Research & Papers
Deployment Decision Reliability: A Generalizability-Theory Framework for Sizing Long-Hori…
arXiv cs.AI Research & Papers
AgonAlpha: Autonomous Alpha Discovery via Prompt Economy and Scalable Agentic Search
arXiv cs.AI Research & Papers
HyperANFIS: Enhancing Rule Representation and Interpretability in Adaptive Neuro-Fuzzy Sy…
arXiv cs.AI Research & Papers
Harnessing agent memory to build lifelong AI partners for materials scientists
arXiv cs.AI Research & Papers
AgenticTwin: An Agentic LLM Framework Integrated with Digital Twin for Anomaly Detection
arXiv cs.AI Research & Papers
Geometry-aware Incremental Neural Operator for Long-Horizon PDE prediction
arXiv cs.AI Research & Papers
Beyond Trial-and-Error: Agentic Optimization for Image-to-Video Adherence
arXiv cs.AI Research & Papers
When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problem…
arXiv cs.AI Research & Papers
VQ-bench: A Composable Vector Quantization Framework
arXiv cs.AI Research & Papers
FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance A…
arXiv cs.AI Research & Papers
TRACE Bench: Task-driven Roleplay Agentic Checklist Evaluation
arXiv cs.AI Research & Papers
TailBooster: A Dual-Layer Generative Framework for Extreme Value Augmentation with Operat…
arXiv cs.AI Research & Papers
Benchmark-Based Comparative Assessment of Publicly Benchmarked Indian Foundation Models: …
arXiv cs.AI Research & Papers
Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Con…
arXiv cs.AI Research & Papers
Causal inference for group-contaminated structured outcomes: observable quotients, lossle…
arXiv cs.AI Research & Papers
Adaptive Hybrid Particle Swarm Optimization with Gradient Descent
arXiv cs.AI Research & Papers
LookBack: Where and How to Score LVLM Responses via Visual Reference Usage
arXiv cs.AI Research & Papers
CoQui: A Coordinate-Conditioned Quantum Implicit Generative Adversarial Network for End-t…
arXiv cs.AI Research & Papers
Hamilton-Zero: A Neural Tensor-Network Foundation Model for Ground States of Arbitrary Qu…
arXiv cs.AI Research & Papers
Instruction Alignment for Binary Code Representation Learning