Evaluating the Effectiveness and Acceptability of a GPT-4o and RAG-Based Voice Chatbot for Depression Screening Using PHQ-9
试验速览
- 阶段
- 不适用
- 状态
- Enrolling By Invitation
- 入组人数
- 100
- 试验地点
- 1
- 主要终点
- Feasibility and Acceptability of the GPT-4o and RAG Voice Chatbot
研究概览
简要总结
This study aims to assess the feasibility and acceptability of a voice-based chatbot, powered by GPT-4o and Retrieval-Augmented Generation (RAG), for conducting depression screening using the Patient Health Questionnaire-9 (PHQ-9). The PHQ-9 is a validated self-report instrument widely used to screen, diagnose, and monitor the severity of depression. It consists of nine questions that correspond to the Diagnostic and Statistical Manual of Mental Disorders (DSM-5) criteria for major depressive disorder. Respondents rate the frequency of symptoms experienced over the past two weeks on a scale from 0 ("not at all") to 3 ("nearly every day"). The total score (ranging from 0 to 27) indicates the severity of depressive symptoms, categorized into minimal, mild, moderate, moderately severe, or severe depression. The PHQ-9 is also used to assess functional impairment and guide treatment decisions in clinical and research settings.
The voice-based chatbot integrates GPT-4o, with RAG to enhance its ability to provide informed and contextualized responses during interactions. GPT-4o serves as the conversational engine, capable of generating human-like, empathetic, and contextually appropriate dialogue. RAG, on the other hand, enables the chatbot to retrieve and incorporate external, up-to-date knowledge from a curated database or knowledge repository, ensuring the accuracy and reliability of its responses.
详细描述
Depression is a prevalent mental health challenge with significant personal, social, and economic costs. Traditional mental health resources face barriers such as stigma, limited availability, and long wait times. Technology, particularly AI-powered tools, provides an opportunity to bridge these gaps. This study utilizes GPT-4o and RAG to create a voice-interactive chatbot capable of conversational engagement, administering the PHQ-9 questionnaire, and delivering personalized feedback.
Participants will fill in the PHQ-9 for self-testing before interacting with the chatbot (the results will not be disclosed to the public and will only be used for accuracy comparisons), and the results of their self-tests will be compared with the results given by the chatbot in terms of accuracy.
The chatbot interaction comprises three phases:
- Warm-up conversations for rapport-building and general support.
- The chatbot initiates casual, empathetic dialogues to build rapport with users, helping them feel comfortable and at ease before transitioning to the PHQ-9 screening.
- Users can ask general questions related to mental health, and the chatbot provides informed and supportive responses.
- Administration of the PHQ-9 questionnaire for depression screening.
研究设计
- 研究类型
- Observational
- 观察模型
- Other
- 时间视角
- Cross Sectional
入排标准
- 年龄范围
- 18 Years 至 65 Years(Adult, Older Adult)
- 性别
- All
- 接受健康志愿者
- 是
入选标准
- •Adults aged 18-65 years.
- •Fluent in English.
- •Access to a device capable of voice interaction and stable internet connection.
- •Willing to participate in chatbot interaction and a follow-up interview.
排除标准
- •Current severe psychiatric diagnoses (e.g., psychosis, bipolar disorder).
- •Participants undergoing active treatment for depression with a psychiatrist.
- •Discomfort with voice-based technology or inability to provide informed consent.
结局指标
主要结局
Feasibility and Acceptability of the GPT-4o and RAG Voice Chatbot
时间窗: Interviews are conducted immediately following the chatbot interaction.
Participants' perceptions of the chatbot's feasibility, acceptability, and effectiveness are assessed through semi-structured interviews conducted after the interaction session. These interviews explore themes such as the chatbot's empathy, usability, and the overall user experience.
次要结局
- Accuracy of PHQ-9 Scoring by the Chatbot(Measured immediately after the interaction session, once the chatbot has generated PHQ-9 scores)
