Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Thre…
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
Reproducing and Stress-Testing Two Approaches to LLM Reasoning Reliability: Test-Time Pro…
Fusion Training for Mathematical Generalization in Large Language Models
Designing for Ethical AI: HCI Feature Considerations to Improve Fairness and User Experie…
The Politician, the Liar, and the Obedient Worker: Emerging Behavior of LLM Agents in Hie…
CoRCi: Cross-Reconstruction of Coherent Interests Modeling in Cross-Domain Sequential Rec…
Second-Order Muon Done Right: A Principled Marriage of Spectral Geometry and Curvature
RAG-Audio: Retrieval-Augmented Generation for Faithful Brain-to-Audio Reconstruction
From Values to Benchmarks: Evaluating Large Language Models for Governmental Use in Dutch
CARD: Controlled Agentic Reddit Discussions for Credit Card Simulation
TSPORec: Token Selection via Preference Optimization for LLM-Based Sequential Recommendat…
Biologically Informed Representation Learning for Robust Cross-Center Generalization of M…
Coordinated incentives in AI-generated misinformation governance
STEMMA: An Adversarial Multi-Agent Framework for Evaluating Self-Identity Consistency in …
TongGuOCR: A Layout-Aware and Token-Augmented OCR Framework for Chinese Historical Docume…
One Adapter Pair per Model: A Universal Activation Interface for Language Models
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models
Capability Is Not Propensity: Measuring Pressure-Robust Cooperative Behavior in Civic LLM…
Cross-Model Humor Preference Modeling with Cards Against Humanity
Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discove…
The Scaffolding Matters More Than the Interface: A Controlled Comparison of MCP and CLI T…
Intelligence Foundation Model: A New Perspective to Approach Artificial General Intellige…
Vision-Language Grounding as Bidirectional Concept Correspondence
Listen, See and Track: Spatio-Temporal Audio-Visual Sound Event Reasoning for Omni-Modal …
CEAA: A Cognitive Embodied Agents Architecture for Interactive Computing Systems
UniDFKD: A Unified Semantic Prior Framework for Architecture-Agnostic Data-Free Knowledge…
Matryoshka Language Model Suites
PIVOT: Preference-based Intervention Vectors for Pedagogical Tutor Steering
MCIF: Multimodal Crosslingual Instruction-Following Benchmark from Scientific Talks
MateInfoUB: A Real-World Benchmark for Testing LLMs in Competitive, Multilingual, and Mul…