此页面是自动翻译的,不保证翻译的准确性。请参阅 英文版 对于源文本。

Patient AI Trust Dynamics Before and After Orthopedic Consultation (ORTHO-OP-GPT) (ORTHO-OP-GPT)

2026年8月26日 更新者:Utku Gürhan

Longitudinal Pre-Post Patient AI Trust Dynamics in Orthopedic Outpatients: A Mixed-Methods Observational Study With Matched Physician-Patient Dyads

Patients increasingly consult artificial intelligence (AI) chatbots such as ChatGPT for health information before clinical visits, yet the impact of an actual orthopedic consultation on patient trust in AI-derived information remains unknown. This prospective longitudinal observational study quantifies how a single orthopedic outpatient consultation modifies patient trust in AI chatbots, the concordance between AI-derived and physician-delivered information, and patient anxiety, using a paired pre-post survey design supplemented by a matched physician-side assessment. Adult patients (18 years and older) presenting to two orthopedic outpatient clinics in Cyprus complete a brief pre-consultation questionnaire (T0) capturing demographics, AI use patterns, prior AI consultation regarding the current complaint, baseline trust, expectations, and anxiety. Immediately after their consultation they complete a second questionnaire (T1) assessing concordance with physician advice, trust change, consultation facilitation, post-consultation anxiety, and future intention. The consulting physician completes a brief 30-second post-visit form capturing whether AI was discussed, the medical accuracy of AI-derived information conveyed by the patient, and the effect of the AI discussion on consultation duration. The primary outcomes are the paired within-patient change in AI trust between T0 and T1 and physician-patient concordance on AI versus physician advice. Target enrollment was 180 to obtain 150 paired completed assessments; 350 participants were enrolled.

研究概览

详细说明

Background and Rationale: Cross-sectional surveys have documented increasing patient use of AI chatbots for health information seeking. However, no published study has assessed how an actual physician consultation modifies patient trust in AI in a paired pre/post design, nor has any study captured the physician perspective on the same encounter in a matched dyad. Routine clinical encounters may be the primary mechanism by which patients calibrate their trust in AI-derived medical information.

Setting and Population: Two university-affiliated orthopedic outpatient clinics in North Cyprus.

Procedures:

  • T0 (pre-consultation, waiting room, approximately 5 minutes): 14-item self-report questionnaire.
  • Consultation: usual care.
  • T1 (post-consultation, departure, approximately 5 minutes): 10-item self-report questionnaire.
  • Physician form (post-consultation, approximately 30 seconds): 5-item brief assessment.
  • Patient and physician forms are linked by an anonymous Participant ID.

Statistical Analysis Plan: Paired t-tests or Wilcoxon signed-rank tests for paired continuous outcomes; McNemar test or Stuart-Maxwell for paired categorical outcomes; Cohen's kappa for inter-rater agreement (AI versus physician); multinomial logistic regression for predictors of trust shift. All analyses two-sided, alpha equals 0.05. SPSS version 28.

Data Management: Anonymous CSV stored locally, encrypted, retained for 5 years per institutional policy. De-identified participant-level data available upon reasonable request after publication.

No formal pilot study is conducted. Instead, the first 20 participants will be prospectively monitored for protocol feasibility (mean completion time, drop-out rate, item-level missing data) as an embedded running pilot.

研究类型

观察性的

注册 (实际的)

350

联系人和位置

本节提供了进行研究的人员的详细联系信息,以及有关进行该研究的地点的信息。

学习地点

      • Kyrenia、塞浦路斯
        • University of Kyrenia, Dr. Suat Gunsel Hospital - Orthopedic Outpatient Clinic
      • Nicosia、塞浦路斯
        • Near East University Hospital - Orthopedic Outpatient Clinic

参与标准

研究人员寻找符合特定描述的人,称为资格标准。这些标准的一些例子是一个人的一般健康状况或先前的治疗。

资格标准

适合学习的年龄

  • 成人
  • 年长者

接受健康志愿者

不

取样方法

非概率样本

研究人群

Consecutive adult patients (18 years and older) presenting to the participating orthopedic outpatient clinics during the recruitment window who consent to participate. No specific orthopedic diagnosis is required.

描述

Inclusion Criteria:

  • Age 18 years or older
  • Presenting to an orthopedic outpatient clinic for any consultation
  • Able to read and respond to a Turkish-language questionnaire
  • Provides informed consent

Exclusion Criteria:

  • Inability to complete a self-report questionnaire (e.g., severe cognitive impairment, language barrier)
  • Re-presentation within the same recruitment window (each patient is enrolled only once)
  • Refusal of consent for either T0 or T1

学习计划

本节提供研究计划的详细信息,包括研究的设计方式和研究的衡量标准。

研究是如何设计的?

设计细节

研究衡量的是什么?

主要结果指标

结果测量
措施说明
大体时间
Mean within-patient change in self-reported trust in artificial intelligence-derived health information, measured by a study-specific 5-point Likert item (T0.11) and a study-specific 3-level categorical change item (T1.4).
大体时间:Baseline (within 15 minutes pre-consultation in the orthopaedic outpatient waiting room) and immediately after the consultation (within 15 minutes of consultation exit, same-day index visit).
Trust in AI-derived health information is assessed pre-consultation by a study-specific single-item 5-point Likert scale (item T0.11: "How much do you trust the AI's answer?"; anchors 1 = not at all, 5 = completely), administered only to patients who reported pre-consultation AI use (item T0.9 = Yes). Post-consultation, trust change is reassessed by a study-specific 3-level categorical item (item T1.4: increased trust / unchanged / decreased trust). For paired analysis, the post-consultation score is derived by mapping T1.4 categories to integer shifts (+1 / 0 / -1, with floor 1 and ceiling 5) relative to T0.11. Unit of measure: Likert score points on a 1-5 scale (continuous derived score) and proportion of patients per 3-level category. Primary analysis: paired Wilcoxon signed-rank test on the derived continuous score; sensitivity analysis: McNemar test on the 3-level categorical change.
Baseline (within 15 minutes pre-consultation in the orthopaedic outpatient waiting room) and immediately after the consultation (within 15 minutes of consultation exit, same-day index visit).
Patient-physician concordance on artificial intelligence-versus-physician medical advice agreement, measured by Cohen's kappa coefficient between a study-specific 4-category patient item (T1.2) and a study-specific 5-point physician-rated AI medical accu
大体时间:Immediately after the consultation (within 15 minutes of consultation exit), for both patient (T1.2) and physician (H2) forms; same-day index visit.
Concordance is assessed by Cohen's kappa coefficient comparing patient-reported AI-physician concordance (item T1.2: fully concordant / partially concordant / discordant / physician did not address; dichotomized to concordant vs. non-concordant) and physician-reported AI medical accuracy (item H2: 5-point Likert anchored 1 = entirely incorrect to 5 = entirely correct; dichotomized at ≥ 3 as concordant). Unit of measure: kappa coefficient (range -1 to +1) with 95% confidence interval, and percentage of dyads classified as concordant on each instrument.
Immediately after the consultation (within 15 minutes of consultation exit), for both patient (T1.2) and physician (H2) forms; same-day index visit.

次要结果测量

结果测量
措施说明
大体时间
Percentage of enrolled patients reporting pre-consultation artificial intelligence use for the current orthopaedic complaint, measured by a study-specific single-item yes/no question (T0.9).
大体时间:Baseline (within 15 minutes pre-consultation, same-day index visit).
Proportion of enrolled patients responding "Yes" to item T0.9 ("Before today's appointment, did you ask an AI chatbot a question about this health concern?"). Unit of measure: percentage of participants, reported with exact (Clopper-Pearson) 95% confidence interval.
Baseline (within 15 minutes pre-consultation, same-day index visit).
Percentage of pre-consultation artificial-intelligence users whose physician independently confirmed that AI was raised during the consultation, measured by a study-specific yes/no physician item (H1).
大体时间:Baseline (T0.9, pre-consultation) and immediately after the consultation (H1, within 15 minutes of consultation exit), same-day index visit.
Among patients responding "Yes" to T0.9, the proportion in whom the treating physician independently reported "Yes" to item H1 ("Did the patient raise AI during this consultation?"). Unit of measure: percentage of patients with exact 95% confidence interval.
Baseline (T0.9, pre-consultation) and immediately after the consultation (H1, within 15 minutes of consultation exit), same-day index visit.
Percentage of consultations in which the physician reported that the artificial-intelligence discussion shortened, did not change, or prolonged the encounter, measured by a study-specific 3-category physician item (H3).
大体时间:Immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.
Among consultations in which the patient raised AI (H1 = Yes), the physician's categorical rating of effect on consultation duration (H3: "shortened" / "no change" / "prolonged"). Unit of measure: percentage of consultations per category (descriptive).
Immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.
Mean patient rating of how prior artificial-intelligence use facilitated the consultation, measured by a study-specific 5-point Likert item (T1.4b: 1 = much more difficult, 5 = much easier).
大体时间:Immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.
Among patients with T0.9 = Yes, patient-reported facilitation by prior AI use (item T1.4b). Unit of measure: Likert score points (mean with standard deviation), and percentage of participants endorsing scores ≥ 4.
Immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.
Mean patient-reported future intention to use and to recommend artificial intelligence for health information, measured by two study-specific 5-point Likert items (T1.7 future use; T1.8 recommendation to a friend).
大体时间:Immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.
Future-use intention (item T1.7: 1 = definitely will not, 5 = definitely will) and recommendation intention (item T1.8: 1 = definitely will not, 5 = definitely will). Unit of measure: Likert score points (mean with standard deviation), and percentage of participants endorsing scores ≥ 4 on each item.
Immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.
Mean within-patient change in consultation-related anxiety, measured by a study-specific six-item instrument (four-item anxiety subscale, range 4-20).
大体时间:Baseline (within 15 minutes pre-consultation) and immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.
Anxiety is measured immediately before and immediately after the consultation using a study-specific six-item instrument modelled on the Amsterdam Preoperative Anxiety and Information Scale, with the referent moved from an impending procedure to the outpatient consultation. The anxiety subscale is the sum of four items (range 4-20). Protocol change: the 0-10 visual analogue scale originally registered for this outcome (items T0.14 and T1.5) was replaced by this instrument before data collection began and was never administered. Unit of measure: scale points. Analysis: paired t-test, with analysis of covariance for between-group comparison.
Baseline (within 15 minutes pre-consultation) and immediately after the consultation (within 15 minutes of consultation exit), same-day index visit.

其他结果措施

结果测量
措施说明
大体时间
Exploratory association between categorical post-consultation trust change and demographic predictors, estimated by multinomial logistic regression with the study-specific 3-level trust change item (T1.4) as the outcome and age band, sex, education level
大体时间:Through study completion, an average of 12 months from first enrolment.
Multinomial logistic regression model: outcome = T1.4 (decreased / unchanged / increased trust, reference category = unchanged); predictors = age band (5-level), sex (3-level), education (5-level), employment status, weekly internet-use frequency. Unit of measure: adjusted odds ratios with 95% confidence intervals.
Through study completion, an average of 12 months from first enrolment.
Internal consistency of a four-item artificial-intelligence trust subscale, measured by Cronbach's alpha across items T0.11 (baseline trust), T1.4 (post-consultation trust change, linearly recoded), T1.7 (future-use intention), and T1.8 (recommendation
大体时间:Through study completion, an average of 12 months from first enrolment.
Cronbach's alpha is estimated on the final analytic sample using the four trust-related Likert items listed. Unit of measure: alpha coefficient (range 0 to 1) with bootstrap 95% confidence interval.
Through study completion, an average of 12 months from first enrolment.

合作者和调查者

在这里您可以找到参与这项研究的人员和组织。

赞助

调查人员

  • 首席研究员:Utku Gurhan, MD、University of Kyrenia

出版物和有用的链接

负责输入研究信息的人员自愿提供这些出版物。这些可能与研究有关。

一般刊物

研究记录日期

这些日期跟踪向 ClinicalTrials.gov 提交研究记录和摘要结果的进度。研究记录和报告的结果由国家医学图书馆 (NLM) 审查,以确保它们在发布到公共网站之前符合特定的质量控制标准。

研究主要日期

学习开始 (实际的)

2026年6月1日

初级完成 (实际的)

2026年7月31日

研究完成 (实际的)

2026年7月31日

研究注册日期

首次提交

2026年5月18日

首先提交符合 QC 标准的

2026年6月1日

首次发布 (实际的)

2026年6月8日

研究记录更新

最后更新发布 (实际的)

2026年8月31日

上次提交的符合 QC 标准的更新

2026年8月26日

最后验证

2026年8月1日

更多信息

与本研究相关的术语

计划个人参与者数据 (IPD)

计划共享个人参与者数据 (IPD)?

是的

药物和器械信息、研究文件

研究美国 FDA 监管的药品

不

研究美国 FDA 监管的设备产品

不

此信息直接从 clinicaltrials.gov 网站检索,没有任何更改。如果您有任何更改、删除或更新研究详细信息的请求,请联系 register@clinicaltrials.gov. clinicaltrials.gov 上实施更改,我们的网站上也会自动更新.

订阅