Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Ekdev Rajkitkul

Spatial Transformer Network YOLO Model for Agricultural Object Detection

Jul 31, 2024

Yash Zambre, Ekdev Rajkitkul, Akshatha Mohan, Joshua Peeples

Figure 1 for Spatial Transformer Network YOLO Model for Agricultural Object Detection

Figure 2 for Spatial Transformer Network YOLO Model for Agricultural Object Detection

Figure 3 for Spatial Transformer Network YOLO Model for Agricultural Object Detection

Figure 4 for Spatial Transformer Network YOLO Model for Agricultural Object Detection

Abstract:Object detection plays a crucial role in the field of computer vision by autonomously identifying and locating objects of interest. The You Only Look Once (YOLO) model is an effective single-shot detector. However, YOLO faces challenges in cluttered or partially occluded scenes and can struggle with small, low-contrast objects. We propose a new method that integrates spatial transformer networks (STNs) into YOLO to improve performance. The proposed STN-YOLO aims to enhance the model's effectiveness by focusing on important areas of the image and improving the spatial invariance of the model before the detection process. Our proposed method improved object detection performance both qualitatively and quantitatively. We explore the impact of different localization networks within the STN module as well as the robustness of the model across different spatial transformations. We apply the STN-YOLO on benchmark datasets for Agricultural object detection as well as a new dataset from a state-of-the-art plant phenotyping greenhouse facility. Our code and dataset are publicly available.

* 7 pages, 5 figures, submitted for review

Via

Access Paper or Ask Questions