이 페이지는 자동 번역되었으며 번역의 정확성을 보장하지 않습니다. 참조하십시오 영문판 원본 텍스트의 경우.

Clinical Validation of an Artificial Intelligence System for the Interpretation of Laboratory Test Results Combined With a Structured Medical History (LTC-VALID)

2026년 9월 3일 업데이트: Labplus Sp. z o.o.

Telemedicine Technology Supporting the Diagnostic Process Based on Automated Analysis of Laboratory Test Results With a Structured Medical History: A Prospective, Two-Center, Two-Phase Clinical Validation Study

This study evaluates a certified artificial intelligence (AI) based software system that automatically interprets laboratory test results in combination with a structured, dynamically generated medical history questionnaire.

The purpose of the study is to determine how accurately and how safely the system assigns a patient to a category of urgency of medical contact, and how closely the interpretations produced by the system correspond to the assessment of an expert physician.

Adults treated at two university hospitals in Katowice, Poland, who are referred for laboratory testing, complete an electronic medical history questionnaire after their laboratory results become available. The system then generates an interpretation for each laboratory result, including a category of urgency of medical contact, a suggested medical specialty and suggested further laboratory tests. The attending physician independently records a clinical assessment of the same laboratory results and medical history without access to the output of the system. An independent expert physician subsequently establishes the reference assessment, blinded to the authorship of the assessments being compared. The interpretation report is released to the participant only after the assessment of the attending physician has been recorded and locked.

The study is conducted in two phases. The first phase (609 participants) is exploratory and uses the initial version of the system. The second phase (290 participants) is confirmatory and uses the final, frozen version of the system. All outcome measures are assessed in both phases; the pre-specified confirmatory hypothesis test applies to the second phase.

The endpoints of this study are properties of the software system, namely the concordance of its output with an expert reference assessment. Health outcomes of participants are not measured.

연구 개요

상세 설명

Design. LTC-VALID is a prospective, two-center, single-arm, two-phase clinical validation study of a CE-marked class IIa medical device software intended for the automated interpretation of laboratory test results.

Phase structure. Phase I (609 participants, months 1 to 9) is exploratory and serves algorithm development and gap identification using version 1.0 of the software. An interposed optimization stage (months 9 to 10) produces a frozen version 2.0. Phase II (290 participants, months 11 to 13) is confirmatory. All outcome measures listed below are collected identically in both phases and are reported separately by phase. Results from the two phases are not pooled, because the two phases evaluate different versions of the software. The pre-specified confirmatory hypothesis test for the co-primary measures is applied to Phase II data; Phase I results for the same measures are exploratory and are reported descriptively.

Procedures. Eligible participants provide written informed consent, are referred for a mandatory basic laboratory panel and one or two of 29 specialist laboratory panels, and provide a single blood draw at a certified laboratory collection point. After the results become available, the participant completes a dynamically generated electronic medical history questionnaire. The software produces one interpretation per laboratory result; interpretations are not aggregated by the software. The attending physician records an independent clinical assessment of the same data while blinded to the software output; this assessment is locked before the interpretation report is released to the participant. Participants complete a questionnaire evaluating the report, and physicians complete a form evaluating the completeness and relevance of the automated medical history.

Reference standard and comparators. An independent expert physician receives the complete documentation and establishes an own reference assessment before reviewing the assessments to be compared. The assessments of the software, of the attending physician and of large language models are presented in random order and blinded as to authorship. The comparators are comparators of assessment, not study arms; the study is single-arm and no randomization or control group is used.

Reporting standard. The primary analysis follows the Standards for Reporting of Diagnostic Accuracy Studies (STARD). The study is a diagnostic accuracy study and not a study of clinical effectiveness.

Statistical approach. The two primary outcome measures are co-primary and are evaluated using an intersection-union test; both must meet their pre-specified criteria. The pre-specified criteria apply to Phase II: for the safety measure, the lower bound of the one-sided 95 percent confidence interval is at or above 95 percent; for the accuracy measure, at or above 90 percent. Confidence intervals for proportions are calculated using exact methods (Clopper-Pearson). Phase I is exploratory and its data are reported descriptively. Sensitivity for the rare urgency categories is reported with confidence intervals and event counts as a secondary, non-confirmatory measure, because its denominator is not sufficient for formal hypothesis testing at the planned sample size.

연구 유형

중재적

등록 (추정된)

899

단계

  • 해당 없음

연락처 및 위치

이 섹션에서는 연구를 수행하는 사람들의 연락처 정보와 이 연구가 수행되는 장소에 대한 정보를 제공합니다.

연구 연락처

연구 장소

      • Katowice, 폴란드
        • Independent Public Clinical Hospital named after Andrzej Mielecki, Medical University of Silesia
        • 연락하다:
      • Katowice, 폴란드
        • Prof. K. Gibinski University Clinical Center, Medical University of Silesia
        • 연락하다:

참여기준

연구원은 적격성 기준이라는 특정 설명에 맞는 사람을 찾습니다. 이러한 기준의 몇 가지 예는 개인의 일반적인 건강 상태 또는 이전 치료입니다.

자격 기준

공부할 수 있는 나이

  • 성인
  • 고령자

건강한 자원 봉사자를 받아들입니다

아니

설명

Inclusion Criteria:

  • Treated at one of the two participating clinical centers of the Medical University of Silesia
  • Age 18 years or older and under 80 years
  • Presence of symptoms justifying the initiation of a diagnostic work-up
  • Meets the criteria for ordering at least one of the 29 specialist laboratory panels included in the study
  • Able to complete an electronic questionnaire in Polish independently, using a smartphone or a personal computer
  • Holds a Polish national identification number (PESEL)
  • Written informed consent covering all three components of the study

Exclusion Criteria:

  • Pregnancy
  • Age 80 years or older
  • Inability to provide informed consent, including cognitive impairment or a language barrier
  • Participation in another clinical study that could affect the results
  • Refusal of consent to any of the three components of the study

공부 계획

이 섹션에서는 연구 설계 방법과 연구가 측정하는 내용을 포함하여 연구 계획에 대한 세부 정보를 제공합니다.

연구는 어떻게 설계됩니까?

디자인 세부사항

  • 주 목적: 특수 증상
  • 할당: 해당 없음
  • 중재 모델: 단일 그룹 할당
  • 마스킹: 없음(오픈 라벨)

무기와 개입

참가자 그룹 / 팔
개입 / 치료
실험적: AI-Based Interpretation of Laboratory Test Results
Single arm. All participants complete an electronic, dynamically generated medical history questionnaire administered by the software after their laboratory results become available, and subsequently receive a software-generated interpretation report. The report is released to the participant only after the independent clinical assessment of the attending physician has been recorded and locked. No control group is used. The assessments of the attending physician, of an independent expert physician and of large language models are comparators of assessment and do not constitute study arms.
A CE-marked class IIa medical device software for the automated interpretation of laboratory test results, operated within its intended purpose. The software administers a structured medical history questionnaire in which subsequent questions are selected dynamically on the basis of the laboratory results and of previous answers. For each laboratory result the software generates one interpretation, comprising a category of urgency of medical contact (immediate, urgent, routine, or no need for medical contact), a suggested medical specialty, a list of suggested further laboratory tests, and an explanatory text addressed to the patient. Interpretations are not aggregated by the software. The interpretation report is generated once per participant, after completion of the questionnaire, and is released to the participant after the assessment of the attending physician has been recorded and locked. The generated report is not used to direct the clinical management of the participant.
다른 이름들:
  • LabTest Checker

연구는 무엇을 측정합니까?

주요 결과 측정

결과 측정
측정값 설명
기간
Percentage of Participants for Whom the Urgency Category Assigned by the Software Was Not Lower Than the Category Assigned by the Expert Physician (Patient Triage Safety Indicator)
기간: Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
The urgency category assigned by the software is compared with the urgency category assigned by an expert physician serving as the reference standard. The urgency categories, ordered from highest to lowest, are: immediate medical contact, urgent medical contact, routine medical contact, and no need for medical contact. A participant is counted in the numerator when the category assigned by the software is the same as or higher than the category assigned by the expert physician. The measure is the percentage of participants counted in the numerator. Range: 0 to 100 percent; higher values indicate better performance. The measure is assessed separately in each study phase and data from the two phases are not combined. Each comparison is based on data locked at the time of enrollment; the expert assessment is performed retrospectively, in batches, on a locked dataset, and the calendar timing of the batch review does not affect the measured quantity.
Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Percentage of Participants for Whom the Urgency Category Assigned by the Software Was Identical to the Category Assigned by the Expert Physician (Patient Triage Accuracy)
기간: Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
The urgency category assigned by the software is compared with the urgency category assigned by an expert physician serving as the reference standard. The urgency categories, ordered from highest to lowest, are: immediate medical contact, urgent medical contact, routine medical contact, and no need for medical contact. A participant is counted in the numerator when the category assigned by the software is identical to the category assigned by the expert physician. The measure is the percentage of participants counted in the numerator. Range: 0 to 100 percent; higher values indicate better performance. The measure is assessed separately in each study phase and data from the two phases are not combined. Each comparison is based on data locked at the time of enrollment; the expert assessment is performed retrospectively, in batches, on a locked dataset, and the calendar timing of the batch review does not affect the measured quantity.
Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start

2차 결과 측정

결과 측정
측정값 설명
기간
Sensitivity for the Immediate Medical Contact Category
기간: Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Proportion of participants assigned by the expert physician to the immediate medical contact category who were assigned to the same category by the software. Range: 0 to 100 percent; higher values indicate better performance. Reported with a 95 percent confidence interval and the number of events. No formal hypothesis test is performed, because the expected denominator at the planned sample size is not sufficient for confirmatory testing. Assessed separately in both study phases.
Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Sensitivity for the Urgent Medical Contact Category
기간: Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Proportion of participants assigned by the expert physician to the urgent medical contact category who were assigned to the same category by the software. Range: 0 to 100 percent; higher values indicate better performance. Reported with a 95 percent confidence interval and the number of events. No formal hypothesis test is performed, because the expected denominator at the planned sample size is not sufficient for confirmatory testing. Assessed separately in both study phases.
Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Agreement on Urgency Category Measured by Quadratic Weighted Cohen's Kappa
기간: Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Agreement between the urgency category assigned by the software and by the expert physician, expressed as Cohen's kappa with quadratic weights applied to the four ordered urgency categories. Range: -1 to 1; higher values indicate better agreement. Reported with a 95 percent confidence interval. Supplementary to the co-primary measure of accuracy. Assessed separately in both study phases.
Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Per-Result Concordance on Recommended Medical Specialty
기간: Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Concordance between the medical specialty recommended by the software and by the expert physician, evaluated at the level of a single laboratory result. Reported as precision and macro-averaged recall. Range: 0 to 1; higher values indicate better performance. Assessed separately in both study phases.
Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Per-Result Concordance on Recommended Additional Laboratory Tests
기간: Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Concordance between the additional laboratory tests recommended by the software and by the expert physician, evaluated at the level of a single laboratory result. Reported as precision and macro-averaged recall. Range: 0 to 1; higher values indicate better performance. Assessed separately in both study phases.
Through completion of the expert reference assessment for all enrolled participants, up to 16 months after study start
Comparative Concordance of the Software, the Attending Physician and Large Language Models Against the Reference Standard
기간: Months 15 to 18 after study start
Concordance of three sources of assessment - the software, the attending physician and large language models - with the same expert reference assessment, using the measures defined above. Assessments are compared retrospectively on a locked dataset. Reported descriptively for each source; no ranking test is performed.
Months 15 to 18 after study start
Proportion of Participants Rating the Interpretation Report as Comprehensible
기간: Immediately after release of the report
Proportion of participants who, in a study-specific questionnaire completed after release of the interpretation report, indicate that the content of the report was comprehensible. Range: 0 to 100 percent; higher values indicate better performance. Assessed separately in both study phases.
Immediately after release of the report
Proportion of Automated Medical History Records Assessed by the Attending Physician as Complete and Relevant
기간: Immediately after the assessment of the attending physician
Proportion of automated medical history records for which the attending physician reports no objections regarding completeness or relevance, recorded in a study-specific form. Range: 0 to 100 percent; higher values indicate better performance. Assessed separately in both study phases.
Immediately after the assessment of the attending physician

공동 작업자 및 조사자

여기에서 이 연구와 관련된 사람과 조직을 찾을 수 있습니다.

스폰서

수사관

  • 수석 연구원: Jerzy Chudek, MD, PhD, Medical University of Silesia in Katowice

간행물 및 유용한 링크

연구에 대한 정보 입력을 담당하는 사람이 자발적으로 이러한 간행물을 제공합니다. 이것은 연구와 관련된 모든 것에 관한 것일 수 있습니다.

연구 기록 날짜

이 날짜는 ClinicalTrials.gov에 대한 연구 기록 및 요약 결과 제출의 진행 상황을 추적합니다. 연구 기록 및 보고된 결과는 공개 웹사이트에 게시되기 전에 특정 품질 관리 기준을 충족하는지 확인하기 위해 국립 의학 도서관(NLM)에서 검토합니다.

연구 주요 날짜

연구 시작 (추정된)

2026년 8월 19일

기본 완료 (추정된)

2027년 10월 31일

연구 완료 (추정된)

2027년 12월 31일

연구 등록 날짜

최초 제출

2026년 8월 17일

QC 기준을 충족하는 최초 제출

2026년 9월 3일

처음 게시됨 (실제)

2026년 9월 9일

연구 기록 업데이트

마지막 업데이트 게시됨 (실제)

2026년 9월 9일

QC 기준을 충족하는 마지막 업데이트 제출

2026년 9월 3일

마지막으로 확인됨

2026년 9월 1일

추가 정보

이 연구와 관련된 용어

기타 연구 ID 번호

  • LP004
  • FEDS.01.02-IP.01-0078/24 (기타 보조금/기금 번호: European Funds for Lower Silesia 2021-2027 (FEDS))
  • BNW/NWN/0052/KB1/34/26 (기타 식별자: Bioethics Committee, Medical University of Silesia in Katowice)

개별 참가자 데이터(IPD) 계획

개별 참가자 데이터(IPD)를 공유할 계획입니까?

아니요

IPD 계획 설명

Individual participant data will not be shared. The study protocol and the statistical analysis plan may be made available with the primary publication. Individual-level data are subject to trade secret protection and are not covered by the scope of participant consent for redistribution.

약물 및 장치 정보, 연구 문서

미국 FDA 규제 의약품 연구

아니

미국 FDA 규제 기기 제품 연구

아니

이 정보는 변경 없이 clinicaltrials.gov 웹사이트에서 직접 가져온 것입니다. 귀하의 연구 세부 정보를 변경, 제거 또는 업데이트하도록 요청하는 경우 register@clinicaltrials.gov. 문의하십시오. 변경 사항이 clinicaltrials.gov에 구현되는 즉시 저희 웹사이트에도 자동으로 업데이트됩니다. .

구독하다