Explorar

Noticias de IA

30724 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Odyssey: Constructing Verifiable Local Truth-Preserving Foundation Models
arXiv cs.AI Research & Papers
When Does Personality Composition Matter for Multi-Agent LLM Teams?
arXiv cs.AI Research & Papers
Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health…
arXiv cs.AI Security & Safety
Just Ask: Curious Code Agents Reveal System Prompts in Frontier LLMs
arXiv cs.AI Research & Papers
CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching
arXiv cs.AI Research & Papers
Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Lea…
arXiv cs.AI Research & Papers
Algorithms for Deciding the Safety of States in Fully Observable Non-deterministic Proble…
arXiv cs.AI Research & Papers
AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization
arXiv cs.AI Research & Papers
Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook
arXiv cs.AI Research & Papers
Derivation of effective gradient flow equations and dynamical truncation of training data…
arXiv cs.AI Research & Papers
The Minimal Search Space for Conditional Causal Bandits
arXiv cs.AI Research & Papers
DMind Benchmark: Toward a Holistic Assessment of LLM Capabilities across the Web3 Domain
arXiv cs.AI Security & Safety
Seven Security Challenges That Must be Solved in Cross-domain Multi-agent LLM Systems
arXiv cs.AI Security & Safety
PRISON: Unmasking the Criminal Potential of Large Language Models
arXiv cs.AI Research & Papers
OSOR: One-Step Diffusion Inpainting for Effect-Aware Object Removal
arXiv cs.AI Research & Papers
ToolPrivacyBench: Benchmarking Purpose-Bound Privacy in Tool-Using LLM Agents
arXiv cs.AI Research & Papers
Freshness and the Limits of Heuristic Trend Detection in Temporal RAG
arXiv cs.AI Research & Papers
Unbiased Binning for Fairness-aware Attribute Representation
arXiv cs.AI Research & Papers
Ranking Before Serving: Low-Latency LLM Serving via Pairwise Learning-to-Rank
arXiv cs.AI Security & Safety
MetaBreak: Jailbreaking Online LLM Services via Special Token Manipulation
arXiv cs.AI Research & Papers
LieSolver: PDE-Constrained Learning for IBVPs via Lie Symmetries
arXiv cs.AI Research & Papers
Hybrid Fact-Checking that Integrates Knowledge Graphs, Large Language Models, and Search-…
arXiv cs.AI Research & Papers
DG^VoiC: Speaker Clustering for Fraud Investigation under Real Call-Centre Conditions
arXiv cs.AI Research & Papers
Parallel Rollout Approximation for Pixel-Space Autoregressive Image Generation
arXiv cs.AI Research & Papers
Trust Region Masking for Long-Horizon LLM Reinforcement Learning
arXiv cs.AI Research & Papers
Psychometric Comparability of LLM-Based Digital Twins
arXiv cs.AI Research & Papers
An Interpretable, Controllable Time-Varying IIR Denoiser for On-Device Assistive Hearing
arXiv cs.AI Research & Papers
Dismantling Pathological Shortcuts: A Causal Framework for Faithful LVLM Decoding
arXiv cs.AI Research & Papers
DDSA: Dual-Domain Strategic Attack for Spatial-Temporal Efficiency in Adversarial Robustn…
arXiv cs.AI Research & Papers
A Primer on SO(3) Action Representations in Deep Reinforcement Learning