精密工学会誌
Online ISSN : 1882-675X
Print ISSN : 0912-0289
ISSN-L : 0912-0289
論文
In-Context Learningを使用した大規模視覚言語モデルによる少数の例示画像付き外観検査
尾下 拓未上野 詩翔山田 悠正中塚 俊介加藤 邦人相澤 宏旭林 良和
著者情報
ジャーナル フリー

2025 年 91 巻 3 号 p. 418-424

詳細
抄録

In this study, we propose a methodology for determining the quality of products based on a few images of non-defective and defective products, along with descriptions that serve as criteria for judgment. Existing Large Vision-Language Models (LVLMs) have demonstrated high performance across a variety of tasks, yet they lack specialized knowledge required for visual inspection. To address this issue, we enhance the LVLM's domain-specific expertise through additional training with a diverse collection of images of non-defective and defective products gathered from the web. Moreover, by utilizing In-Context Learning (ICL), our approach enables inference on inspection images based on a few exemplar images of non-defective and defective products, along with their judgment criteria descriptions, thereby eliminating the need for collecting extensive training samples and training models for each product type as traditionally required. By integrating LVLM with ICL, our method introduces a novel approach to general visual inspection, demonstrating its utility.

著者関連情報
© 2025 公益社団法人 精密工学会
前の記事 次の記事
feedback
Top