Conference
Open access
Sep 2026
Object Surface Defect Segmentation using Vision Transformers and Hybrid Models
An empirical comparison between three deep learning frameworks for pixel-based surface defect detection: a baseline CNN architecture known as U-Net, Vision Transformer (ViT) based on patch-wise attention mechanism, and TransUNet that combines CNN and transformer architecture.
Santhya C, VS Sneha Chowdary, Beaulah Jeyavathana
· Engineering & Technology · 0 citations