Can Small Language Models Reliably Resist Jailbreak Attacks? A Comprehensive Evaluation
Explorar
Noticias de IA
29655 elementos — filtrados, clasificados y sin duplicados
A Contractualist Argumentation Framework for Moral Decision-Making
Cost-Based Semantics for Querying Inconsistent Weighted Knowledge Bases
SKT: Skill-Use Training at Scale via Verified Synthetic Data Generation
Shared Organizational Memory for Enterprise Coding Agents: System Design and Deployment S…
Unifying biomedical knowledge in a modern multimodal graph
HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering
OntoTKGE: Ontology-Enhanced Temporal Knowledge Graph Extrapolation
High-Stakes Decisions with Language Models: Insights from Emergency Triage
GeoMind: An Agentic Workflow for Lithology Classification with Reasoned Tool Invocation
When to Vote, When to Rewrite: Disagreement-Guided Strategy Routing for Test-Time Scaling
Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models
When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Auditing
Group Selection as a Safeguard Against AI Substitution
FinToolBench: Evaluating LLM Agents for Real-World Financial Tool Use
CORE: Collaborative Reasoning via Cross Teaching
Mind the Sim2Real Gap in User Simulation for Agentic Tasks
TeachArena: Are Language Agents Ready for Realistic Teaching Work?
Unleashing the Potential of Large Language Models: A Blueprint for Real-Time, Enterprise-…
ICU-Bench:Benchmarking Continual Unlearning in Multimodal Large Language Models
Bounded Normative Equivalence in Human-AI Cooperation: Group Behaviour, Not Partner Label…
Trust or Check? Understanding the (Evolutionary) Dynamics of User Trust in AI Systems
LADY: Linear Attention for Autonomous Driving Efficiency without Transformers
SIEVE: Selective Integrity Verification and Escalation for Defending LLM Agents against I…
Panning for Gold: Expanding Domain-Specific Knowledge Graphs with General Knowledge
Knowledge Graph Augmented Large Language Models for Disease Prediction
AIC-VDS: Attention-Based In-Context Learning for Joint Velocity Control and Data Collecti…
Efficiency vs. Alignment: Investigating Safety and Fairness Risks in Parameter-Efficient …
ProbGuard: Proactive Runtime Monitoring for LLM Agent Safety via Probabilistic Prediction
DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys