Skip to main content
OpenTrials
Not Yet Recruiting

NCT Number: NCT07740044

Medical Large Language Model-Assisted Diagnosis and Treatment in Primary Care Chronic Disease Management

This study aims to explore the feasibility of using medical large language models to assist in chronic disease management. By setting up experimental and control group interventions and having experts blindly evaluate anonymized cases, it compares different management approaches in chronic disease care, verifying the scientific basis, effectiveness, and potential for broader use of medical large language models in supporting chronic disease management.

Not Yet Recruiting

Trial opening soon.

Get Notified

Key information

Age range

18 year–65 year

Sex eligibility

All sexes

Study type

Interventional

Phase

Not applicable

Primary location

The Affiliated Taizhou People's Hospital of Nanjing Medical University

Taizhou, Jiangsu, 225300, China

Location contact

Guoyu Wang, MD

CONTACT

[email protected]

+8615951154720

Who can participate

Healthy volunteers accepted: No

Only the study team can determine whether someone qualifies for participation.

Inclusion criteria

  • Work experience of 3 years or more;
  • Able to complete the case assessment tasks required by the study;
  • Have a practicing doctor qualification;
  • Work at a township health center or community health service center;
  • Voluntarily participate in the study and sign the informed consent form.

Exclusion criteria

N/A

Treatment and study plan

Medical large language model

Device

The primary care physicians in the intervention group completed the management decisions for chronic disease cases, including disease assessment, examination suggestions, treatment plans, selection of management methods, and health education, with the assistance of the medical LLM.

Primary outcomes

  1. expert overall scores

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale on several aspects: accuracy of condition assessment, reasonableness of test recommendations, reasonableness of treatment plans, appropriateness of management choices, scientific nature of patient education content, feasibility at the primary care level, overall clinical quality. Experts were instructed to rank the seven dimensions from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

Secondary outcomes

  1. accuracy of condition assessment

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale on accuracy of condition assessment from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

  2. reasonableness of test recommendations

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale regarding the reasonableness of test recommendations. Experts were instructed to rank the dimension from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

  3. reasonableness of treatment plans

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale regarding the reasonableness of treatment plans. Experts were instructed to rank the dimension from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

  4. appropriateness of management choices

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale regarding the ppropriateness of management choices. Experts were instructed to rank the dimension from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

  5. scientific nature of patient education content

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale regarding the scientific nature of patient education content. Experts were instructed to rank the dimension from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

  6. feasibility at the primary care level

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale regarding the feasibility at the primary care level. Experts were instructed to rank the dimension from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

  7. the overall clinical quality

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. They rated them using a Likert scale regarding the overall clinical quality. Experts were instructed to rank the dimension from lowest to highest quality. These ordinal rankings were then converted into standardized scores of 1 to 5. Higher scores mean better.

  8. risk advice

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. If there is no risk advice, the expert selects 'No'; otherwise, the expert selects 'Yes' and specifically lists the content of the risk advice.

  9. Any diagnostic or treatment information missing

    Time frame: through study completion, an average of 3 months.

    A total of 200 cases (50 each of coronary heart disease, atrial fibrillation, heart failure, and stroke) were grouped by disease type and independently evaluated in a blinded manner by three subspecialty experts. For each case, the two management plans were anonymized in the same way and then randomly labeled by a computer as Plan A and Plan B. Experts could only view the anonymized case information and the two corresponding anonymized management plans. If there is no omission of diagnostic and treatment information, the expert selects "No"; otherwise, the expert selects "Yes" and specifically lists the omitted diagnostic and treatment content.

  10. feedback regarding the use of medical LLM by primary care physicians

    Time frame: through study completion, an average of 3 months.

    After primary care physicians complete a case assessment assisted by medical LLM , they will fill out an evaluation of the medical LLM usage, including whether the medical LLM helped in disease assessment, assisted in formulating examination plans, aided in developing treatment plans, improved chronic disease management capabilities, supported chronic disease education, assisted in management approach selection, enhanced confidence, saved time, whether there were hallucinations or omissions, whether it provided risky suggestions, and any additional comments on the use of the medical LLM.

Study contacts

Contact information is provided by the study sponsor or research team.

Guoyu Wang, MD

CONTACT

[email protected]

+8615951154720

Sponsors and collaborators

Lead sponsor

Jiangsu Taizhou People's Hospital

Other

Registry information

Official study title

Verification and Application Study of Medical Large Language Model-Assisted Diagnosis and Treatment in Primary Care Chronic Disease Management: A Randomized Controlled Study

Important dates

Study start
2026
Primary completion
2026
Study completion
2026
First posted
Jul 31, 2026
Registry last updated
Jul 31, 2026

OpenTrials presents study information sourced from ClinicalTrials.gov. The official registry record should be consulted for the latest information.

View the official ClinicalTrials.gov record (opens in a new tab)

This listing is for discovery and informational purposes only. It is not medical advice, does not guarantee that a study is recruiting, and does not determine eligibility. Contact the study team and a qualified healthcare professional when considering participation.

Published trials that share one or more normalized conditions with this study.