メインナビゲーションにスキップ 検索にスキップ メインコンテンツにスキップ

Preliminary Study on Image-Finding Generation and Classification of Lung Nodules in Chest CT Images Using Vision–Language Models

研究成果: ジャーナルへの寄稿学術論文査読

抄録

In the diagnosis of lung cancer, imaging findings of lung nodules are essential for benign and malignant classifications. Although numerous studies have investigated the classification of lung nodules, no method has been proposed for obtaining detailed imaging findings. This study aimed to develop a novel method for generating image findings and classifying benign and malignant nodules in chest computed tomography (CT) images using vision–language models. In this study, we collected chest CT images of 77 patients diagnosed with either benign or malignant tumors at Fujita Health University Hospital. For these images, we cropped the regions of interest around the nodules, and a pulmonologist provided the corresponding image findings. We used vision–language models for image captioning to generate image findings. The findings generated by these two models were grammatically correct, with no deviations in notation, as expected from the image findings. Moreover, the descriptions of benign and malignant characteristics were accurately obtained. The bootstrapping language–image pretraining (BLIP) base model achieved an accuracy of 79.2% in classifying nodules, and the bilingual evaluation understudy-4 score for agreement with physician findings was 0.561. These results suggest that the proposed method may be effective for classifying and generating lung nodule findings.

本文言語英語
論文番号489
ジャーナルComputers
14
11
DOI
出版ステータス出版済み - 11-2025
外部発表はい

UN SDG

この成果は、次の持続可能な開発目標に貢献しています

  1. SDG 3 - すべての人に健康と福祉を
    SDG 3 すべての人に健康と福祉を

All Science Journal Classification (ASJC) codes

  • コンピュータ サイエンス(その他)
  • 人間とコンピュータの相互作用
  • コンピュータ ネットワークおよび通信

フィンガープリント

「Preliminary Study on Image-Finding Generation and Classification of Lung Nodules in Chest CT Images Using Vision–Language Models」の研究トピックを掘り下げます。これらがまとまってユニークなフィンガープリントを構成します。

引用スタイル