GeneralNeuro / Head & NeckChest / ThoracicAI / InformaticsResearch

LLM Extraction from Reports Enables Scalable AI Monitoring in ACR's Assess-AI Registry

Journal of the American College of Radiology : JACR2d ago

In ACR's Assess-AI registry, LLM prompts for extracting findings from radiology reports achieved 0.985 and 0.997 agreement on development cohorts for ICH and PE, but independent validation is lacking.

  • LLM prompts were developed for nine use cases, with high agreement on internal tuning cohorts (0.985 for intracranial hemorrhage, 0.997 for pulmonary embolism).
  • No independent validation was performed; the reported agreement comes from the same cohorts used to refine prompts and reference labels.
  • The workflow demonstrates feasibility of scalable, report-anchored AI performance monitoring, but clinical utility remains unestablished.

Automated summary

RadPigeon summaries are original and for information only. They are not clinical advice.