Back to Research papers
Research paper index

A Visual Question Answering Model to Automate Nondestructive Evaluation Image Analysis

Mehrdad Shafiei Dizaji, Hoda Azari

arXiv:2608.29408Published August 29, 20260 citations
  • cs.CV
  • cs.LG
  • action

Abstract

This study introduces a Visual Question Answering model designed specifically for nondestructive evaluation applications. VQA models allow inspectors to interactively query NDE images, asking targeted questions like, Is there a crack or Where is the defect located and receive precise answers from the model. Leveraging deep learning and natural language processing, the developed system integrates image feature extraction (via a ResNet-50 model) and language generation capabilities (via GPT-2) to provide accurate, informative feedback. By enabling direct question-and-answer interactions, this VQA model significantly improves inspection efficiency, reduces potential errors, and enhances usability in practical field scenarios.

Read the original paper

This page indexes public paper metadata. The manuscript remains with its original publisher and authors.