Brian A. Zaboski, Gregory N Muller
The rapid integration of artificial intelligence (AI) into mental healthcare presents opportunities and ethical challenges, particularly for complex conditions like obsessive–compulsive disorder (OCD). In this perspective, we argue for a Dual Imperative: establishing safety architectures for AI-powered therapeutic tools to prevent algorithmic sycophancy (symptom accommodation), while mandating explainable AI (XAI) in prognostic models to ensure clinical auditability. In therapeutics, we propose a Guardian Angel architecture that utilizes patient-specific fear hierarchies and linguistic stance detection to distinguish compulsive reassurance-seeking from legitimate patient questions. This approach transforms potential therapeutic ruptures into opportunities for distress tolerance via the Digital Ulysses Pact, a patient-authorized, algorithmically enforced response prevention protocol. In diagnostics, we address the black box problem in precision psychiatry. We argue that as AI evolves from detection to high-stakes treatment selection, safety and accountability become a prerequisite for clinical application. Although distinct in implementation, these architectures form an integrated framework for aligning therapeutic and diagnostic AI. These architectures are not parallel tracks but a unified ecosystem: A patient’s XAI-audited profile can inform the Guardian Angel’s configuration, while the longitudinal data gathered during therapy enriches diagnostic precision. Grounded in ethical principles and best practices in OCD, this suggests a path toward AI that is auditable in its diagnostic logic, firm in its therapeutic boundaries, and enforceable through emerging regulatory frameworks.