このページは自動翻訳されたものであり、翻訳の正確性は保証されていません。を参照してください。 英語版 ソーステキスト用。

Enhancing Readability of Lay Abstracts and Summaries for Medical Knowledge Using Generative Artificial Intelligence (BRIDGE AI 3)

2026年5月4日 更新者:Giovanni Cacciamani、University of Southern California

Enhancing Readability of Lay Abstracts and Summaries for Medical Knowledge Using Generative Artificial Intelligence: A Randomized Controlled Trial (BRIDGE AI 3)

This trial tests if AI can help make medical info clear and readable. Many patients struggle to find medical informations that easy to read and understand from verified medical sources. The study tests if an AI tool can assist health providers to craft clear text for patients more fast than what they do now. Health providers are split at random into two groups-one uses the AI tool and one does not. The trial tests how clear the text is, how correct it is, and how much time is saved. The aim is to see if AI can close the gap between complex research and what patients can grasp.

調査の概要

詳細な説明

This study evaluates whether a generative artificial intelligence (AI) tool can improve the readability and accessibility of lay summaries derived from scientific medical abstracts. Many patients encounter difficulty understanding medical literature due to technical language and complexity, which can limit informed decision-making and engagement with healthcare information.

The BRIDGE-AI (Provider Perspective) initiative aims to address this gap by enabling healthcare professionals and researchers to generate patient-friendly summaries of scientific content using AI-assisted tools. The intervention leverages a generative AI framework (pub2people) designed to translate complex medical terminology into language that is understandable to a general audience.

In this randomized controlled study, participants with experience in scientific publishing will be assigned to either an AI-assisted group or a control group using conventional methods. Participants will be asked to transform scientific abstracts into layperson-friendly summaries. The study compares AI-assisted and manually generated outputs in terms of readability, accuracy, and efficiency.

The primary objective is to determine whether AI-assisted generation improves the readability of lay summaries compared to standard approaches. Secondary objectives include evaluating the accuracy of AI-generated summaries relative to source material and assessing potential time savings associated with AI use.

This study contributes to ongoing efforts to improve health communication by evaluating scalable tools that may enhance the translation of complex medical information into patient-accessible formats.

研究の種類

介入

入学 (実際)

120

段階

  • フェーズ2

連絡先と場所

このセクションには、調査を実施する担当者の連絡先の詳細と、この調査が実施されている場所に関する情報が記載されています。

研究場所

    • California
      • Los Angeles、California、アメリカ、91100
        • University of Southern California

参加基準

研究者は、適格基準と呼ばれる特定の説明に適合する人を探します。これらの基準のいくつかの例は、人の一般的な健康状態または以前の治療です。

適格基準

就学可能な年齢

  • 大人
  • 高齢者

健康ボランティアの受け入れ

はい

説明

Provider Participants

Inclusion:

  • Corresponding authors who have been published in the top 10 journals of urology and medicine
  • All genders
  • Any profession
  • 18+ years of age

Exclusion:

  • Anyone under the age of 18
  • Participants that have not published in the top 10 journals of urology and medicine

研究計画

このセクションでは、研究がどのように設計され、研究が何を測定しているかなど、研究計画の詳細を提供します。

研究はどのように設計されていますか?

デザインの詳細

  • 主な目的:ヘルスサービス研究
  • 割り当て:ランダム化
  • 介入モデル:並列代入
  • マスキング:独身

武器と介入

参加者グループ / アーム
介入・治療
介入なし:Gold Standard
実験的:Pub2Post
Pub2Post, a generative artificial intelligence agent which helps in drafting the layperson abstracts and summaries

この研究は何を測定していますか?

主要な結果の測定

結果測定
メジャーの説明
時間枠
Readability Change
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Flesch Reading Ease Score

Description: Measures text readability based on sentence length and word syllables.

Scale: 0 to 100 Interpretation: Higher scores indicate easier readability (better outcome).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Readability Change
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Flesch-Kincaid Grade Level Description: Estimates U.S. school grade level required to understand the text. Scale: Typically ranges from ~0 to 18+ Interpretation: Lower scores indicate easier readability (better outcome).
The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Readability Change
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Gunning Fog Index Description: Estimates years of formal education needed to understand the text on first reading.

Scale: Typically 0 to 20+ Interpretation: Lower scores indicate easier readability (better outcome).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Readability Change
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
SMOG Index (Simple Measure of Gobbledygook) Description: Estimates years of education required to comprehend the text. Scale: Typically 0 to 20+ Interpretation: Lower scores indicate easier readability (better outcome).
The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Readability Change
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Coleman-Liau Index Description: Readability formula based on characters per word and sentence length.

Scale: Typically 0 to 18+ (grade level equivalent) Interpretation: Lower scores indicate easier readability (better outcome).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Readability Change
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Automated Readability Index (ARI) Description: Estimates grade level required for comprehension using characters and word counts.

Scale: Typically 0 to 14+ Interpretation: Lower scores indicate easier readability (better outcome).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

二次結果の測定

結果測定
メジャーの説明
時間枠
Time Saving
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

To evaluate the time savings achieved by using generative AI compared to traditional methods for generating layperson abstracts and summaries. Time will be recorded in hours, minutes, and seconds. We will collect and compare the total time spent drafting the complete layperson abstract and summaries, as well as the time spent on each individual section - background, methods, results, conclusion, and short summaries. The comparison will be made between summaries created by humans alone versus those created with GAI assistance.

Time will be reported in minutes

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Correctness and meaning retention
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Accuracy Score of Layperson Abstract Sections Description: Degree to which each section (Background, Methods, Results, Conclusion, Short Summary) reflects key information from the original scientific abstract.

Scale: 5-point Likert scale (1 = very inaccurate, 5 = highly accurate) Assessment Method: Two independent reviewers score each section Interpretation: Higher scores indicate better accuracy (better outcome).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Correctness and meaning retention
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Completeness Score of Layperson Abstract Sections Description: Extent to which essential information from the original abstract is included in each section.

Scale: 5-point Likert scale (1 = very incomplete, 5 = fully complete) Assessment Method: Two independent reviewers evaluate each section. Interpretation: Higher scores indicate greater completeness (better outcome).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Correctness and meaning retention
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Clarity Score for Layperson Readability Description: Evaluates simplicity, avoidance of jargon, and coherence for lay audiences.

Scale: 5-point Likert scale (1 = very unclear, 5 = very clear and understandable) Assessment Method: Two independent reviewers evaluate each section. Interpretation: Higher scores indicate better clarity (better outcome).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Correctness and meaning retention
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Section-Level Correctness Rate Description: Proportion of sections rated as "correct," defined as receiving scores ≥4 from both reviewers on accuracy, completeness, and clarity simultaneously.

Scale: 0 to 1 (proportion) or 0% to 100% Interpretation: Higher values indicate better overall section quality.

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Correctness and meaning retention
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.

Hallucination Rate Description: Frequency of false or misleading content in generated layperson abstracts, defined as information not supported by the original abstract.

Scale: Proportion of sections or documents containing hallucinations (0 to 1 or %) Assessment Method: Evaluated separately by reviewers using predefined criteria. Interpretation: Lower values indicate better performance (fewer hallucinations).

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment.
Perceived Task Difficulty
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment

Description: Participant-reported difficulty of completing the lay abstract summarization task.

Scale: 5-point Likert scale (1 = very easy, 5 = very difficult) (adjust anchors if different in your instrument) Assessment Timing: Immediately after task completion Interpretation: Lower scores indicate less perceived difficulty (better outcome)

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment
Perceived Task Duration
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment
Description: Participant perception of time required to complete the task. Scale: 5-point Likert scale (1 = very short, 5 = very long) Assessment Timing: Post-task Interpretation: Lower scores indicate shorter perceived duration (better outcome)
The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment
Perceived Helpfulness of the intervention
時間枠:The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment

Description: Participant-reported usefulness of the generative AI tool in assisting lay abstract creation.

Scale: 5-point Likert scale (1 = not helpful at all, 5 = extremely helpful) Assessment Timing: Post-task Interpretation: Higher scores indicate greater perceived helpfulness (better outcome)

The assessment will be conducted immediately after the study closes, which will occur 4 weeks after enrollment
System Usability Scale (SUS) Score
時間枠:Immediately after completing the system/task (post-use assessment)

Description: Standardized assessment of system usability using the System Usability Scale.

Scale: 0 to 100 Interpretation: Higher scores indicate better usability

  • 70: acceptable usability
  • 90: superior usability
Immediately after completing the system/task (post-use assessment)
Perceived Usefulness (Technology Acceptance Model)
時間枠:Immediately after completing the system/task (post-use assessment)

Description: Degree to which participants believe the GAI tool enhances task performance.

Scale: Likert scale (typically 1-5 or 1-7; specify exact instrument version) Interpretation: Higher scores indicate greater perceived usefulness (better outcome)

Immediately after completing the system/task (post-use assessment)
Perceived Ease of Use (Technology Acceptance Model)
時間枠:Immediately after completing the system/task (post-use assessment)
Description: Degree to which participants find the GAI tool easy to use. Scale: Likert scale (typically 1-5 or 1-7; must match instrument used) Interpretation: Higher scores indicate greater ease of use (better outcome)
Immediately after completing the system/task (post-use assessment)

協力者と研究者

ここでは、この調査に関係する人々や組織を見つけることができます。

研究記録日

これらの日付は、ClinicalTrials.gov への研究記録と要約結果の提出の進捗状況を追跡します。研究記録と報告された結果は、国立医学図書館 (NLM) によって審査され、公開 Web サイトに掲載される前に、特定の品質管理基準を満たしていることが確認されます。

主要日程の研究

研究開始 (実際)

2025年1月16日

一次修了 (実際)

2025年1月22日

研究の完了 (実際)

2025年2月27日

試験登録日

最初に提出

2025年3月12日

QC基準を満たした最初の提出物

2026年4月30日

最初の投稿 (実際)

2026年5月5日

学習記録の更新

投稿された最後の更新 (実際)

2026年5月6日

QC基準を満たした最後の更新が送信されました

2026年5月4日

最終確認日

2026年5月1日

詳しくは

本研究に関する用語

個々の参加者データ (IPD) の計画

個々の参加者データ (IPD) を共有する予定はありますか?

はい

IPD 共有サポート情報タイプ

  • STUDY_PROTOCOL
  • SAP

医薬品およびデバイス情報、研究文書

米国FDA規制医薬品の研究

いいえ

米国FDA規制機器製品の研究

いいえ

この情報は、Web サイト clinicaltrials.gov から変更なしで直接取得したものです。研究の詳細を変更、削除、または更新するリクエストがある場合は、register@clinicaltrials.gov。 までご連絡ください。 clinicaltrials.gov に変更が加えられるとすぐに、ウェブサイトでも自動的に更新されます。

購読する