此页面是自动翻译的,不保证翻译的准确性。请参阅 英文版 对于源文本。

Real-World Data Linkage Research Platform

2026年6月3日 更新者:Kong Yuanyuan、Beijing Friendship Hospital

This study aims to address the lack of intelligent governance tools in clinical data management to promote efficient governance and secure sharing of real-world health data. To achieve this, a self-adaptive, automated governance intelligent agent will be developed based on a High-Order Programming (HOP) architecture, integrating Large Language Models (LLMs) and deep learning techniques. The agent will continuously monitor and correct data quality issues in real time, improving data accuracy and usability.

In parallel, the project will establish a trusted data-sharing framework by integrating AI Confidential Computing (AICC) with Trusted Data Matrix (TDM) technologies. This framework will enable secure, real-time cross-institutional data exchange and collaborative computation while protecting sensitive information.

Overall, the study aims to transform fragmented clinical data into high-quality, standardized, and securely accessible resources, thereby facilitating the circulation of data value and advancing collaborative medical research.

研究概览

详细说明

This multicenter, observational cohort study aims to integrate longitudinal health data from China, including routine health examinations, electronic medical records, and disease registries. The platform is designed to address key data challenges in the medical domain, particularly in chronic diseases and suboptimal health status. It is driven by two primary objectives:

  1. Intelligent and automated data governance To ensure high data quality, the platform will engineer a self-adaptive, automated governance intelligent agent. Integrating Large Language Models (LLMs) and High-Order Programming (HOP), this agent actively monitors and corrects real-world data issues, such as missing values, redundancies, and formatting inconsistencies. Through deep learning, the agent continuously optimizes its governance rules to adapt to complex medical data environments.
  2. Trusted and secure data sharing To facilitate multicenter collaborative research, the study will establish a secure and trusted data-sharing framework. By integrating AI confidential computation (AICC) with Trusted Data Matrix (TDM) technologies, the platform provides hardware-level security guarantees. This ensures that real-time, cross-institutional data exchange and collaborative computation without exposing sensitive patient information.

Overall Objective The platform aims to transform heterogeneous clinical data into standardized, high-quality, and securely accessible resources, thereby enabling efficient data utilization and promoting the value circulation of medical data for real-world evidence research.

研究类型

观察性的

注册 (估计的)

300000

联系人和位置

本节提供了进行研究的人员的详细联系信息,以及有关进行该研究的地点的信息。

学习联系方式

  • 姓名:Yuanyuan Kong, PhD
  • 电话号码:+86 1063139362 +86 15810026760
  • 邮箱:kongyy@ccmu.edu.cn

研究联系人备份

学习地点

    • Beijing Municipality
      • Beijing、Beijing Municipality、中国、100050
        • Beijing Friendship Hospital, Capital Medical University.No. 95, Yongan Road, Xicheng District, Beijing, 100050, China

参与标准

研究人员寻找符合特定描述的人,称为资格标准。这些标准的一些例子是一个人的一般健康状况或先前的治疗。

资格标准

适合学习的年龄

  • 孩子
  • 成人
  • 年长者

接受健康志愿者

是的

取样方法

非概率样本

研究人群

This study establishes a multicenter, observational real-world data platform integrating longitudinal health data from multiple sources across China, including routine health examinations, electronic medical records, and disease registries. The platform is designed to support population-level research without restriction to specific diseases or conditions, enabling inclusive and continuous assessment of health status, disease risk, progression, and outcomes in real-world settings.

All available individuals with usable health-related data are eligible for inclusion, with minimal restrictions to maximize data coverage and representativeness. Both retrospective and prospective data will be incorporated and linked at the individual level using standardized protocols within a secure data governance and privacy protection framework.

描述

Inclusion Criteria:

  • Participants will be eligible for inclusion if they meet all of the following criteria:

    1. Availability of any health-related data generated from routine clinical care, health examinations, or disease surveillance systems, regardless of disease type or health status.
    2. Presence of at least one type of usable data, including but not limited to diagnostic information (structured or unstructured), laboratory results, imaging data, or basic demographic information.
    3. Records contain sufficient information (appropriately anonymized) to allow data organization and, where feasible, linkage at the individual level across time points or data sources.

Exclusion Criteria:

  • Participants or records meeting any of the following criteria will be excluded:

    1. Records lacking minimal essential information required to distinguish individual records or support basic analysis (e.g., completely missing identifiers or time information).
    2. Records confirmed to be invalid, including system-generated test data, corrupted entries, or records that do not represent real clinical or health-related events.
    3. Exact duplicate records that cannot be resolved through standard data processing (only one record will be retained when duplicates are identifiable).

学习计划

本节提供研究计划的详细信息,包括研究的设计方式和研究的衡量标准。

研究是如何设计的?

设计细节

队列和干预

团体/队列
干预/治疗
Data-Link Cohort
The study cohort is derived from a multicenter, population-based real-world data platform that integrates longitudinal data from electronic medical records, disease registries, and routine health examinations across multiple institutions. The platform is designed to support broad, disease-agnostic research and enable dynamic evaluation of health status, disease risk, and outcomes in real-world settings.
This is an observational study. No intervention will be applied.

研究衡量的是什么?

主要结果指标

结果测量
措施说明
大体时间
Accuracy Rate of Automated Data Governance
大体时间:2026.5.30 to 2028.12.31
Using a manually curated gold-standard dataset, the effectiveness of the intelligent agent in improving data accuracy will be evaluated by measuring the proportion of data values that correctly match the gold-standard reference after automated data governance. The accuracy rate will be calculated as the percentage of correctly recorded or corrected data elements among all evaluated data elements. Values range from 0% to 100%, with higher values indicating better data accuracy.
2026.5.30 to 2028.12.31
Completeness Rate of Automated Data Governance
大体时间:2026.5.30 to 2028.12.31
Using a manually curated gold-standard dataset, the effectiveness of the intelligent agent in improving data completeness will be evaluated by measuring the proportion of required data fields that are complete after automated data governance. The completeness rate will be calculated as the percentage of non-missing required data elements among all required data elements. Values range from 0% to 100%, with higher values indicating better data completeness.
2026.5.30 to 2028.12.31

次要结果测量

结果测量
措施说明
大体时间
Correction Accuracy of Automated Data Governance
大体时间:2026.5.30 to 2028.12.31
Using a manually curated gold-standard dataset, the effectiveness of the intelligent agent in resolving identified data quality issues will be evaluated by measuring correction accuracy. Correction accuracy will be calculated as the percentage of identified data quality issues (e.g., missing values, format inconsistencies, and logical conflicts) that are correctly resolved after automated data governance, compared with the gold-standard reference dataset. Values range from 0% to 100%, with higher values indicating better correction performance.
2026.5.30 to 2028.12.31
Data Standardization Rate of Automated Data Governance
大体时间:2026.5.30 to 2028.12.31
Using a manually curated gold-standard dataset, the effectiveness of the intelligent agent in standardizing data will be evaluated by measuring the proportion of data elements that conform to predefined data standards, terminologies, and formatting rules after automated data governance. The data standardization rate will be calculated as the percentage of evaluated data elements that meet standardized data specifications among all assessed data elements. Values range from 0% to 100%, with higher values indicating better data standardization.
2026.5.30 to 2028.12.31
Cross-institutional Data Usability of Automated Data Governance
大体时间:2026.5.30 to 2028.12.31
Using datasets derived from participating institutions, the effectiveness of the intelligent agent in improving cross-institutional data usability will be evaluated by measuring the proportion of governed datasets that can be successfully integrated, interpreted, and used across different institutions according to predefined interoperability and usability criteria after automated data governance. Cross-institutional data usability will be calculated as the percentage of datasets meeting prespecified usability criteria among all evaluated datasets. Values range from 0% to 100%, with higher values indicating better cross-institutional usability.
2026.5.30 to 2028.12.31

合作者和调查者

在这里您可以找到参与这项研究的人员和组织。

调查人员

  • 首席研究员:Yuanyuan Kong、Beijing Friendship Hospital

出版物和有用的链接

负责输入研究信息的人员自愿提供这些出版物。这些可能与研究有关。

一般刊物

研究记录日期

这些日期跟踪向 ClinicalTrials.gov 提交研究记录和摘要结果的进度。研究记录和报告的结果由国家医学图书馆 (NLM) 审查,以确保它们在发布到公共网站之前符合特定的质量控制标准。

研究主要日期

学习开始 (估计的)

2026年5月30日

初级完成 (估计的)

2028年12月31日

研究完成 (估计的)

2030年12月31日

研究注册日期

首次提交

2026年5月20日

首先提交符合 QC 标准的

2026年6月3日

首次发布 (实际的)

2026年6月9日

研究记录更新

最后更新发布 (实际的)

2026年6月9日

上次提交的符合 QC 标准的更新

2026年6月3日

最后验证

2026年5月1日

更多信息

与本研究相关的术语

药物和器械信息、研究文件

研究美国 FDA 监管的药品

不

研究美国 FDA 监管的设备产品

不

此信息直接从 clinicaltrials.gov 网站检索,没有任何更改。如果您有任何更改、删除或更新研究详细信息的请求,请联系 register@clinicaltrials.gov. clinicaltrials.gov 上实施更改,我们的网站上也会自动更新.

订阅