A new method for automated concrete bridge damage detection using an efficient Vision Transformer-enhanced anchor-free YOLO (You Only Look Once) has been proposed by researchers from the University of ...
In the last decade, convolutional neural networks (CNNs) have been the go-to architecture in computer vision, owing to their powerful capability in learning representations from images/videos.
Transformers, first proposed in a Google research paper in 2017, were initially designed for natural language processing (NLP) tasks. Recently, researchers applied transformers to vision applications ...