AI Persuasion as a Threat to Human Control
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
AcquireBound: Runtime Authorization for Resources Acquired by AI Agents
Self-Orchestrating Language Models: Leveraging Semantic Dependence for Efficient Inference
One Model, Two Physical Stories: Auditing Misalignment in Multi-Modal World Modeling
Four Ledgers, Not One Score: Responsible Communication of LLM-Judge Calibration in Biomed…
Domain Generalization for Smartphone-Based Human Activity Recognition: A Systematic Analy…
Externalizing Requirement-to-Repair Artifacts as Observable Traces for LLM-Based Program …
Geometric Flow enhanced Graph Coarsening
Converting Sequenced Fuzzy Cognitive Maps to Causal Virtual Worlds with Large Video Gener…
Shallow Beliefs: Synthetic document finetuning does not inoculate against emergent misali…
Enabling Creative Exploration for Vibe Design Agents
OpenAl4S: Code as Action, Science as Sessions
Attention Is All You Need (to Avoid Spurious Oscillations)
STHMoE: Hypergraph-Enhanced Heterogeneous Dependency Coordination for LLM-Based Urban Tra…
ProIQA: A Process-Based Framework for Fine-Grained Math Item Quality Assessment
Empirical Evaluation of Open-Source Large Language Models for Retrieval-Augmented Generat…
Parameter-Efficient Adaptation of Pretrained Language Models for Time-Series Forecasting
Reason What Matters: Retrieval-Grounded Reasoning for Universal Multimodal Embeddings
Who Teaches Which Token? Verifier-Gated Multi-Expert On-Policy Distillation for Scientifi…
Option-Aware Retrieval and Task-Specific VLM Adaptation for Medical VQA
Can AI systems have free will?
HISPO: Hierarchical Importance-Sampling Policy Optimization with Entropy-Derived Segments
Confuse the Model, Control the Flow: Understanding and Mitigating Privacy Leakage from LL…
NoteVQA: Benchmarking VLMs on Real-Life Questions from Human Communities
Data storytelling meets interpretable machine learning: Decoding AI decisions for non-exp…
EEG-Xplain: Decoding Neural Black-Boxes of EEG Foundation Models
Stellar Colosseum: A Many-Agent Harness for Long-Horizon Research in Mathematics and Theo…
vla-eval: A Unified Evaluation Harness for Vision-Language-Action Models
Atria Dawn: The Dawn of Agentic Superintelligence
Corrupt Plans, Clean Traces: Evading Chain-of-Thought Monitoring with Plan Injection