Large Language Models for Epidural Stimulation Electrode Mapping in Spinal Cord Injury
AI-Assisted Electrode Contact Configuration Mapping for Epidural Electrical Stimulation in Spinal Cord Injury: A Comparative Evaluation of Large Language Models
This observational and methodological study aims to compare the performance of large language models in generating electrode contact configuration recommendations for epidural electrical stimulation in spinal cord injury.
Five standardized synthetic spinal cord injury scenarios will be presented to four large language models: ChatGPT-4o, Claude, Grok 3, and Gemini 2.5 Pro. Each model will receive the same standardized prompt. The generated responses will be anonymized and evaluated independently by experts with experience in spinal cord injury rehabilitation and epidural electrical stimulation.
The responses will be assessed in five main areas: clinical accuracy, technical feasibility, safety awareness, consistency with current clinical guidance, and completeness of the response. Agreement between expert evaluators will also be examined.
No real patients, human participants, clinical interventions, or personal health data are included in this study. The study is designed to explore the potential and current limitations of large language models as artificial intelligence-based clinical decision-support tools in neurorehabilitation.
Přehled studie
Postavení
Postavení
Podmínky
Podmínky
Intervence / Léčba
Intervence / Léčba
Typ studie
Typ studie
Zápis (Aktuální)
Zápis
Kontakty a umístění
Studijní místa
-
-
Istanbul
-
Istanbul, Istanbul, Turecko (Türkiye), 34290
- Istanbul Gelisim University
-
-
Kritéria účasti
Kritéria způsobilosti
Kritéria způsobilosti
Věk způsobilý ke studiu
- Dítě
- Dospělý
- Starší dospělý
Přijímá zdravé dobrovolníky
Metoda odběru vzorků
Studijní populace
Popis
Inclusion Criteria:
- Responses generated for one of the five predefined standardized synthetic spinal cord injury scenarios.
- Responses generated using the identical standardized prompt specified in the study protocol.
- Responses generated by one of the four prespecified large language models.
- Complete responses available for expert evaluation.
Exclusion Criteria:
- Responses generated using prompts that differ from the standardized study prompt.
- Incomplete, interrupted, or technically corrupted model outputs.
- Duplicate responses or outputs not corresponding to a predefined synthetic scenario.
- Any response generated using real patient-identifiable or personal health information.
Studijní plán
Jak je studie koncipována?
Detaily designu
Počet skupin / kohort
Kohorty a intervence
Skupina / kohortaSkupina / kohorta |
Intervence / LéčbaIntervence / Léčba |
|---|---|
|
ChatGPT-4o
Responses generated by ChatGPT-4o for five standardized synthetic spinal cord injury scenarios using the same standardized prompt.
The responses will be evaluated for clinical accuracy, technical feasibility, safety awareness, guideline consistency, and completeness.
|
The large language model receives five standardized synthetic spinal cord injury scenarios using an identical standardized prompt and generates recommendations for epidural electrical stimulation electrode contact configuration mapping.
No intervention is administered to human participants.
|
|
Claude
Responses generated by Claude for five standardized synthetic spinal cord injury scenarios using the same standardized prompt.
The responses will be evaluated for clinical accuracy, technical feasibility, safety awareness, guideline consistency, and completeness.
|
The large language model receives five standardized synthetic spinal cord injury scenarios using an identical standardized prompt and generates recommendations for epidural electrical stimulation electrode contact configuration mapping.
No intervention is administered to human participants.
|
|
Grok 3
Responses generated by Grok 3 for five standardized synthetic spinal cord injury scenarios using the same standardized prompt.
The responses will be evaluated for clinical accuracy, technical feasibility, safety awareness, guideline consistency, and completeness.
|
The large language model receives five standardized synthetic spinal cord injury scenarios using an identical standardized prompt and generates recommendations for epidural electrical stimulation electrode contact configuration mapping.
No intervention is administered to human participants.
|
|
Gemini 2.5 Pro
Responses generated by Grok 3 for five standardized synthetic spinal cord injury scenarios using the same standardized prompt.
The responses will be evaluated for clinical accuracy, technical feasibility, safety awareness, guideline consistency, and completeness.
|
The large language model receives five standardized synthetic spinal cord injury scenarios using an identical standardized prompt and generates recommendations for epidural electrical stimulation electrode contact configuration mapping.
No intervention is administered to human participants.
|
Co je měření studie?
Primární výstupní opatření
Primární výstupní opatření
Měření výsledku |
Popis opatření |
Časové okno |
|---|---|---|
|
Clinical Accuracy Score of Large Language Model Responses
Časové okno: At the time of expert evaluation, within 1 week after study initiation
|
Clinical accuracy of the epidural electrical stimulation electrode contact configuration recommendations generated by each large language model will be independently evaluated by expert reviewers using a 5-point Likert-type rating scale.
Higher scores indicate greater clinical accuracy of the generated recommendations.
|
At the time of expert evaluation, within 1 week after study initiation
|
Sekundární výstupní opatření
Sekundární výstupní opatření
Měření výsledku |
Popis opatření |
Časové okno |
|---|---|---|
|
Technical Feasibility Score of Large Language Model Responses
Časové okno: At expert evaluation, within 1 week after study initiation
|
The technical feasibility of epidural electrical stimulation electrode contact configuration recommendations generated by each large language model will be independently evaluated by expert reviewers using a 5-point Likert-type rating scale.
Higher scores indicate greater technical feasibility and applicability of the generated recommendations.
|
At expert evaluation, within 1 week after study initiation
|
|
Safety Awareness Score of Large Language Model Responses
Časové okno: At expert evaluation, within 1 week after study initiation
|
The safety awareness demonstrated in the epidural electrical stimulation electrode contact configuration recommendations generated by each large language model will be independently evaluated by expert reviewers using a 5-point Likert-type rating scale.
Higher scores indicate greater recognition and consideration of relevant safety issues.
|
At expert evaluation, within 1 week after study initiation
|
|
Clinical Guideline Consistency Score of Large Language Model Responses
Časové okno: At expert evaluation, within 1 week after study initiation
|
The consistency of the generated epidural electrical stimulation electrode contact configuration recommendations with current clinical guidance will be independently evaluated by expert reviewers using a 5-point Likert-type rating scale.
Higher scores indicate greater consistency with current clinical guidance and relevant evidence-based recommendations.
|
At expert evaluation, within 1 week after study initiation
|
|
Response Completeness Score of Large Language Model Responses
Časové okno: At expert evaluation, within 1 week after study initiation
|
The completeness of the epidural electrical stimulation electrode contact configuration recommendations generated by each large language model will be independently evaluated by expert reviewers using a 5-point Likert-type rating scale.
Higher scores indicate more complete and comprehensive responses.
|
At expert evaluation, within 1 week after study initiation
|
Spolupracovníci a vyšetřovatelé
Sponzor
Sponzor
Termíny studijních záznamů
Hlavní termíny studia
Začátek studia (Aktuální)
Začátek studia
Primární dokončení (Odhadovaný)
Primární dokončení
Dokončení studie (Odhadovaný)
Dokončení studie
Termíny zápisu do studia
První předloženo
První předloženo
První předloženo, které splnilo kritéria kontroly kvality
První předloženo, které splnilo kritéria kontroly kvality
První zveřejněno (Aktuální)
První zveřejněno
Aktualizace studijních záznamů
Poslední zveřejněná aktualizace (Aktuální)
Poslední zveřejněná aktualizace
Odeslaná poslední aktualizace, která splnila kritéria kontroly kvality
Odeslaná poslední aktualizace, která splnila kritéria kontroly kvality
Naposledy ověřeno
Naposledy ověřeno
Více informací
Termíny související s touto studií
Další relevantní podmínky MeSH
Další identifikační čísla studie
Další identifikační čísla studie
- EES-LLM-2026-01
Informace o lécích a zařízeních, studijní dokumenty
Studuje lékový produkt regulovaný americkým FDA
Studuje produkt zařízení regulovaný americkým úřadem FDA
Tyto informace byly beze změn načteny přímo z webu clinicaltrials.gov. Máte-li jakékoli požadavky na změnu, odstranění nebo aktualizaci podrobností studie, kontaktujte prosím register@clinicaltrials.gov. Jakmile bude změna implementována na clinicaltrials.gov, bude automaticky aktualizována i na našem webu .