UTS at ELOQUENT 2026 Voight-Kampff: structural shifts in AI writing bypass state-of-the-a…
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Compaction as Epistemic Failure: How Agentic LLM Tools Fabricate Confirmed Results from K…
Verifying formulas for interventional distributions
DeepLoop: Depth Scaling for Looped Transformers
DevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments
GeoAnchor: Collaborative Reasoning via Latent Decomposition for 3D Spatial Understanding
Adversarial Prompting Framework for AI Safety Assessment
Symbiosis-Inspired Knowledge Distillation for Incremental Object Detection
Learning Physics-Guided Residual Dynamics for Deformable Object Simulation
Partially Correlated Verifier Cascades in LLM Harnesses: Concave Log-Odds, Polynomial Rel…
ScanFocus: A Coarse-to-Fine Framework for Spatio-Temporal Video Grounding
Can We Steer the Black-Box? Towards Controllability-Centric Evaluation of Recommender Sys…
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware …
Price of Fairness in Bandits: A Tight Minimax Characterization
From Observation to Insight: Mechanistic World Models and the Quest for Autonomous Discov…
Learning Engagement Assistant (LEA): Cross-Course Scalability and Classroom Evaluation of…
Traffic-Aware Randomized Smoothing for LLM-Based Network Intrusion Detection
The Refusal Residue: When Probes Catch Alignment Faking and When They Don't
Efficient Text-to-Audio Generation via Pruning
Privacy Preserving Recommender Systems Balancing Personalization with Privacy
Adapting Generalist Vehicle Models for High-Speed MPC Across Terrains
Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchm…
Early Adoption of Agentic Coding Tools by GitHub Projects
Faithful Autoformalization of Natural Language Assertions
Knowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Deci…
AI-Augmented Human Resource Management? Insights from German companies
HRO: Hierarchical Room-to-Object Framework for Zero-Shot Object Goal Navigation with Larg…
How Agents Ask for Permission: User Permissions for AI Agents, from Interfaces to Enforce…
Reassessing Muon for Matrix Factorization
Interventional Grounding Audits: Black-Box Premise-Dependency Tests for LLM Chain-of-Thou…