Abstract
In the highly competitive steel sector, product quality, particularly in terms of surface integrity, is critical. Surface defect detection (SDD) is essential in maintaining high production standards, as it directly impacts product quality and manufacturing efficiency. Traditional SDD approaches, which rely primarily on manual inspection or traditional computer vision techniques, are plagued with difficulties, including reduced accuracy and potential health concerns to inspectors. This research describes an innovative solution that uses a sequence generation model with transformers to improve the defect detection process while manufacturing hot-rolled steel sheets and generating captions about the defect and its spatial location. This method, which views object detection as a sequence generation problem, allows for a more sophisticated understanding of image content and a complete and contextually rich investigation of surface defects whilst providing captions. While this method can potentially improve detection accuracy, its actual power rests in its scalability and flexibility to various industrial applications. Furthermore, this technique has the potential to be further enhanced for visual question-answering applications, opening up opportunities for interactive and intelligent image analysis.
| Original language | English |
|---|---|
| Title of host publication | Robotics, Computer Vision and Intelligent Systems |
| Subtitle of host publication | 4th International Conference, ROBOVIS 2024, Rome, Italy, February 25–27, 2024, Proceedings |
| Editors | Joaquim Filipe, Juha Röning |
| Place of Publication | Cham, Switzerland |
| Publisher | Springer |
| Chapter | 20 |
| Pages | 316-333 |
| Number of pages | 18 |
| Edition | 1 |
| ISBN (Electronic) | 9783031590573 |
| ISBN (Print) | 9783031590566 |
| DOIs | |
| Publication status | Published - 8 May 2024 |
| Event | Robotics, Computer Vision and Intelligent Systems 2024 - Rome, Italy Duration: 25 Feb 2024 → 27 Feb 2024 |
Publication series
| Name | Communications in Computer and Information Science |
|---|---|
| Publisher | Springer |
| ISSN (Print) | 1865-0929 |
| ISSN (Electronic) | 1865-0937 |
Conference
| Conference | Robotics, Computer Vision and Intelligent Systems 2024 |
|---|---|
| Abbreviated title | ROBOVIS |
| Country/Territory | Italy |
| City | Rome |
| Period | 25/02/24 → 27/02/24 |
UN SDGs
This output contributes to the following UN Sustainable Development Goals (SDGs)
-
SDG 9 Industry, Innovation, and Infrastructure
Keywords
- Image captioning
- Pixel to sequence
- Steel defect detection
Fingerprint
Dive into the research topics of 'MDC-Net: Multimodal Detection And Captioning Network For Steel Surface Defects'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver