Forecasting Side Effects of Activation Steering
Explorar
Noticias de IA
29343 elementos — filtrados, clasificados y sin duplicados
CORA-Diff: Confidence-Oriented Residual Acceptance for Efficient Diffusion Language Model…
Is Per-Agent Policy Composition Safe? Rethinking Successor-Feature Transfer in Cooperativ…
Co-constructing sociotechnical AI governance: participatory system mapping using algorith…
Two-Stage Deformable-Convolutional Inverse Design of Nanophotonic Absorbers from Optical …
Who Thinks Best Depends on How Long You Let Them: Budget-Dependent Rankings in LLM Evalua…
How to Spend Your Oracle Budget: Practical Guidance for Protein Structure Prediction Mode…
Towards the Harness of Embodied Agents
Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing
Dion3: Full-Stack Orthogonal Updates
Evaluating LLM Generated Detection Rules in Cybersecurity
GUIDE: Governed Unified Intelligence for Document-to-Artifact Generation in Enterprise Se…
Dynamic Governance of Multi-LLM Agent Systems for Collaborative Conversational Outcomes
User-Assisted Collaborative Distributed Inference for Efficient QoS-Aware Autoscaling
Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration
Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop
VLM2Rec: Resolving Modality Collapse in Vision-Language Model Embedders for Multimodal Se…
Low-Interaction-Rank Learning: Unifying Multiplicative Dual-Encoder Heads
MOON: Multi-Objective OrthoNormalized Updates for Multitask Learning
From Monolithic to Modular: Segment-level Automatic Prompt Optimization
DORA Explorer: Improving the Exploration Ability of LLMs Without Training
Federated Learning for Distributed CNC Tool Wear Prediction
LLMs in Process Diagram Engineering: From Optimal PFDs to Validated P&IDs
Probably Approximately Correct Maximum A Posteriori Inference
Conflict and Congruency Effects in Large Language Models: In-Weight and In-Context Compet…
Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discove…
Identity from the Outside: A Conceptual Framework and Research Program for AI Personality…
VICBench: A Multi-Language Benchmark for Code Vulnerability Detection
Governing Agentic AI in FinTech
Semantic Lenia: Emergence of Homeostatic Solitons within the Semantic Space of Large Lang…