ViTCAI: A Vision-Language Model for Automated Triage and Captioning of Postoperative Incision Images Submitted by Remote Patients.
basic_science · Level V
Where this comes from
- Record sourced from PubMed, PMID 41359691.
- Also identified by DOI 10.1109/JBHI.2025.3624318.
- No licence information is recorded for this record.
- Because redistribution is not established, this page shows the abstract only. Follow the links below for the full text.
Abstract
Surgical site infections (SSIs) are common postoperative complications that increase patient morbidity, hospital stays, and healthcare costs. Early detection and precise documentation are critical for timely intervention and improved outcomes. While patients often submit images of their wounds through a patient portal for remote monitoring, manual review is challenging due to image variability, high volume, and subjectivity, underscoring the need for automated assessment tools. In this paper, we present Vision Transformer for CAptioning of Incision Images (ViTCAI), a vision-language model designed for automated triage and captioning of postoperative incision images. Fine-tuned on a clinically annotated dataset, ViTCAI improves descriptive accuracy in identifying surgical incisions and SSIs. Our results show that ViTCAI provides consistent, detailed captions that can support clinical decision-making, reducing workload and enhancing diagnostic efficiency in postoperative care.
Medical subject headings
- Triage
- Surgical Wound Infection
- Image Interpretation, Computer-Assisted
- Surgical Wound