Odyssey: Constructing Verifiable Local Truth-Preserving Foundation Models
Explorar
Noticias de IA
30724 elementos — filtrados, clasificados y sin duplicados
When Does Personality Composition Matter for Multi-Agent LLM Teams?
Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health…
Just Ask: Curious Code Agents Reveal System Prompts in Frontier LLMs
CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching
Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Lea…
Algorithms for Deciding the Safety of States in Fully Observable Non-deterministic Proble…
AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization
Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook
Derivation of effective gradient flow equations and dynamical truncation of training data…
The Minimal Search Space for Conditional Causal Bandits
DMind Benchmark: Toward a Holistic Assessment of LLM Capabilities across the Web3 Domain
Seven Security Challenges That Must be Solved in Cross-domain Multi-agent LLM Systems
PRISON: Unmasking the Criminal Potential of Large Language Models
OSOR: One-Step Diffusion Inpainting for Effect-Aware Object Removal
ToolPrivacyBench: Benchmarking Purpose-Bound Privacy in Tool-Using LLM Agents
Freshness and the Limits of Heuristic Trend Detection in Temporal RAG
Unbiased Binning for Fairness-aware Attribute Representation
Ranking Before Serving: Low-Latency LLM Serving via Pairwise Learning-to-Rank
MetaBreak: Jailbreaking Online LLM Services via Special Token Manipulation
LieSolver: PDE-Constrained Learning for IBVPs via Lie Symmetries
Hybrid Fact-Checking that Integrates Knowledge Graphs, Large Language Models, and Search-…
DG^VoiC: Speaker Clustering for Fraud Investigation under Real Call-Centre Conditions
Parallel Rollout Approximation for Pixel-Space Autoregressive Image Generation
Trust Region Masking for Long-Horizon LLM Reinforcement Learning
Psychometric Comparability of LLM-Based Digital Twins
An Interpretable, Controllable Time-Varying IIR Denoiser for On-Device Assistive Hearing
Dismantling Pathological Shortcuts: A Causal Framework for Faithful LVLM Decoding
DDSA: Dual-Domain Strategic Attack for Spatial-Temporal Efficiency in Adversarial Robustn…
A Primer on SO(3) Action Representations in Deep Reinforcement Learning