Image

Large Language Model-Assisted cTNM Annotation From Chinese PSMA PET/CT Reports

Large Language Model-Assisted cTNM Annotation From Chinese PSMA PET/CT Reports

Recruiting
18 years and older
Male
Phase N/A

Powered by AI

Overview

This observational study will develop and validate a large language model-assisted workflow for imaging cTNM staging annotation and uncertainty recognition in prostate cancer using Chinese PSMA PET/CT report texts generated during routine clinical care. The study will use de-identified report texts and necessary baseline clinical information only. No additional imaging examination, blood test, treatment, or follow-up visit will be assigned for this study.

The main objective is to evaluate whether a locally or institutionally controlled large language model can identify report-derived imaging cT, cN, and cM categories, extract supporting evidence from the original report, and recognize uncertainty expressions. Model performance will be assessed using an internal independent validation set, external validation reports from two collaborating hospitals, and a prospective validation set of 100 consecutive routine PSMA PET/CT reports. A human-AI comparison will also be performed using physicians from urology and imaging-related specialties with different seniority levels.

Description

This is a multicenter observational diagnostic accuracy validation study based on Chinese PSMA PET/CT report texts from patients with prostate cancer or suspected prostate cancer. The study is not designed to evaluate a drug, device, surgical procedure, or imaging intervention. PSMA PET/CT examinations will be performed as part of routine clinical care, and the study will only analyze de-identified report texts and necessary baseline information after the reports have been finalized.

The study consists of retrospective and prospective components. Retrospectively, approximately 4,000 PSMA PET/CT reports from the First Affiliated Hospital of Wenzhou Medical University will be systematically annotated to construct a research database. An internal independent validation set of 300 reports, not used for model development or prompt optimization, will be used to evaluate the performance of the large language model. The reference standard for this 300-report validation set will be established by two experienced urologists through joint annotation, with adjudication by a nuclear medicine expert when needed. External validation will be performed using 110 de-identified reports from the First Affiliated Hospital of Ningbo University and 102 de-identified reports from Liuzhou People's Hospital. In addition, after ethics approval, 100 consecutive routine PSMA PET/CT reports from the First Affiliated Hospital of Wenzhou Medical University will be prospectively included to evaluate the accuracy and operational stability of the frozen model and prompt versions.

The large language model workflow will be deployed locally or in an institutionally controlled environment. The model will be instructed to generate structured JSON outputs, including cT\_report, cN\_report, cM\_report, cT\_uncertain, cN\_uncertain, cM\_uncertain, evidence\_T, evidence\_N, and evidence\_M. The target task is report-derived imaging cTNM staging annotation, not pathological TNM staging or overall AJCC stage grouping. The model output will be used only for research evaluation and methodological analysis and will not be used for clinical diagnosis, treatment decision-making, or patient notification.

A human-AI comparison will be conducted on the 300-report internal validation set. Eight human evaluators from urology and imaging-related specialties, including trainees, residents, attending physicians, and associate chief physicians, will independently annotate the reports before and after learning the annotation manual. Annotation time will be recorded for each round. The performance of human evaluators and the large language model will be compared against the expert consensus reference standard.

The main outcome will be the accuracy of the large language model in identifying cT, cN, and cM categories from Chinese PSMA PET/CT reports. Secondary outcomes will include precision, recall, F1-score, macro-F1, micro-F1, complete cTNM triplet matching rate, uncertainty recognition performance, evidence extraction quality, human-AI comparison results, annotation time, external validation performance, prospective validation performance, and error type distribution. Error analysis will focus on local tumor extent, regional versus non-regional lymph node boundaries, bone and visceral metastasis recognition, equivocal wording, treatment-related context, benign or inflammatory alternatives, and lesions not attributable to prostate cancer.

Eligibility

Inclusion Criteria:

  1. Male patients aged 18 years or older.
  2. Patients with clinically diagnosed, pathologically diagnosed, or clinically suspected prostate cancer.
  3. Patients who underwent PSMA PET/CT for initial staging, recurrence assessment, treatment response evaluation, metastatic assessment, or other clinical purposes during routine care.
  4. Complete or basically complete Chinese PSMA PET/CT report text is available, including imaging findings and/or diagnostic impression.
  5. The report text contains information that can be used to evaluate at least one target field, such as local prostate lesion, regional lymph nodes, non-regional lymph nodes, bone metastasis, visceral metastasis, or uncertainty expressions.
  6. The research data can be de-identified and replaced by a study identification number before analysis.

Exclusion Criteria:

  1. PSMA PET/CT reports unrelated to prostate cancer, or reports clearly irrelevant to the research task.
  2. Reports with severely missing, unreadable, or unavailable main text, imaging findings, or diagnostic impression.
  3. Reports that cannot be adequately de-identified or contain residual direct personal identifiers that cannot be safely removed.
  4. Duplicate records, repeated exports of the same examination, or records for which the unique report version cannot be confirmed.
  5. Reports judged by the research team to be of insufficient quality for manual annotation, model evaluation, or statistical analysis.

Study details
    Prostate Cancer

NCT07707232

First Affiliated Hospital of Wenzhou Medical University

25 July 2026

Step 1 Get in touch with the nearest study center
We have submitted the contact information you provided to the research team at {{SITE_NAME}}. A copy of the message has been sent to your email for your records.
Would you like to be notified about other trials? Sign up for Patient Notification Services.
Sign up

Send a message

Enter your contact details to connect with study team

Investigator Avatar

Primary Contact

  Other languages supported:

First name*
Last name*
Email*
Phone number*
Other language

FAQs

Learn more about clinical trials

What is a clinical trial?

A clinical trial is a study designed to test specific interventions or treatments' effectiveness and safety, paving the way for new, innovative healthcare solutions.

Why should I take part in a clinical trial?

Participating in a clinical trial provides early access to potentially effective treatments and directly contributes to the healthcare advancements that benefit us all.

How long does a clinical trial take place?

The duration of clinical trials varies. Some trials last weeks, some years, depending on the phase and intention of the trial.

Do I get compensated for taking part in clinical trials?

Compensation varies per trial. Some offer payment or reimbursement for time and travel, while others may not.

How safe are clinical trials?

Clinical trials follow strict ethical guidelines and protocols to safeguard participants' health. They are closely monitored and safety reviewed regularly.
Add a private note
  • abc Select a piece of text.
  • Add notes visible only to you.
  • Send it to people through a passcode protected link.