Position: We Need Practical AI Alignment Methods to Mirror Human Reasoning
Explorar
Noticias de IA
29336 elementos — filtrados, clasificados y sin duplicados
Beyond Outcome Rewards: Step-Level Self-Distilled Policy Optimization for Deep Search Age…
Mr3D-VL: A generalist vision language foundation model for Multiparametric 3D Magnetic Re…
Erase but Preserve: Controllable Removal of Copyrighted Animation Characters via Optimize…
SynAct: A Reasoning-Acting Large Language Model Agent for Adaptive Synthesis Optimization
ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval
How Do VLMs Behave When Blind or Misled? Behavioral Evaluation of VLMs on Scientific Figu…
HumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmark
Heterogeneous Vision-Language Ensemble with Disagreement-Aware Reranking for Text-Based P…
Interpretable Causal Discovery via Causal-Effect Constraints
BoardroomAI: Dependency-Aware Human-Steerable Multi-Agent Deliberation through Evolving D…
A Compositional Theory of Curvature in Probabilistic Circuits
Falsehood and Impossibility Are Different Directions in an AI's Representation of Language
AI and Consumer Rights in India Working Paper
CityRiSE: Reasoning Urban Socio-Economic Status in Large Vision-Language Models via Reinf…
ReflectFact: Self-Reflective Agents for Improving Comprehension and Reasoning in Multi-Ho…
AQuA: Recursively Self-Improving Quantitative Trading Research Agents
FSGR: Mitigating Token Frequency Bias for Fair SID-Based Generative Recommendation
Teach the Magnitude, Not the Direction: Verifier-Bounded Credit Assignment for Multi-Turn…
Towards Context-Aware Clinical Motion Understanding in Daily Living at Home: Freezing of …
Better Decomposition, Free Aggregation: A Synthesizer-Folding Framework for Multilingual …
Practice Makes Unsafe: Skill Misevolution in Self-Improving LLM Agents
Position: The Alignment Community is Unintentionally Building a Censor's Toolkit
Don't Want Your LLM to Recommend Nuclear Strike? Try Asking It in Japanese
NARU: A Benchmark for NARrative Evolution and Cultural Nuance Understanding in Japanese E…
GEM: A Generative Embedding Model Bridging Reasoning and Retrieval
Waymo receives permission to offer rides in Sacramento and San Diego
Nvidia Has $21 Billion SpaceX Stake, $30 Billion in Intel Shares
Anthropic Revenue Ahead of IPO Surges Over 14-Fold in Second Quarter
The Next Big Influencer Is This 4-Foot-Tall Robot From China