SODA: Site Object Detection dAtaset for Deep Learning in Construction

AI-generated keywords: Construction Industry Object Detection Dataset Deep Learning Annotation

AI-generated Key Points

Lack of large-scale, open-source datasets for object detection in the construction industry
Development of a new dataset called Site Object Detection dAtaset (SODA)
Dataset contains 19,846 images with 286,201 objects categorized into 15 classes
Dataset is diverse and voluminous, covering different site conditions, weather conditions, construction phases, angles, and perspectives
Feasibility evaluation using two widely-adopted object detection algorithms (YOLO v3/ YOLO v4)
Maximum mean Average Precision (mAP) achieved is 81.47%
Total of 1,246 hours spent on data cleaning and annotation
Data cleaning accounted for 16.4% of the time, data labeling accounted for 68%, and data checking accounted for 15.6%
Most images in SODA have a resolution of 1920 * 1080
Worker labels are the most common while machine and layout labels are less frequent
Use of k-means clustering algorithm to analyze length-width ratio and range of objects in the dataset
Statistical results showing distribution of shooting angles used to capture the images
Contribution of a large-scale image dataset for deep learning-based object detection methods in the construction industry
Performance benchmark set up for evaluating algorithms in this area
Valuable resources provided for further research and development in construction site object detection

Also access our AI generated: Comprehensive summary, Lay summary, Blog-like article; or ask questions about this paper to our AI assistant.

Authors: Rui Duan, Hui Deng, Mao Tian, Yichuan Deng, Jiarui Lin

Automation in Construction, 2022

arXiv: 2202.09554v1 - DOI (cs.CV)

License: CC BY 4.0

Abstract: Computer vision-based deep learning object detection algorithms have been developed sufficiently powerful to support the ability to recognize various objects. Although there are currently general datasets for object detection, there is still a lack of large-scale, open-source dataset for the construction industry, which limits the developments of object detection algorithms as they tend to be data-hungry. Therefore, this paper develops a new large-scale image dataset specifically collected and annotated for the construction site, called Site Object Detection dAtaset (SODA), which contains 15 kinds of object classes categorized by workers, materials, machines, and layout. Firstly, more than 20,000 images were collected from multiple construction sites in different site conditions, weather conditions, and construction phases, which covered different angles and perspectives. After careful screening and processing, 19,846 images including 286,201 objects were then obtained and annotated with labels in accordance with predefined categories. Statistical analysis shows that the developed dataset is advantageous in terms of diversity and volume. Further evaluation with two widely-adopted object detection algorithms based on deep learning (YOLO v3/ YOLO v4) also illustrates the feasibility of the dataset for typical construction scenarios, achieving a maximum mAP of 81.47%. In this manner, this research contributes a large-scale image dataset for the development of deep learning-based object detection methods in the construction industry and sets up a performance benchmark for further evaluation of corresponding algorithms in this area.

Submitted to arXiv on 19 Feb. 2022

Ask questions about this paper to our AI assistant

You can also chat with multiple papers at once here.

AI assistant instructions?

Results of the summarizing process for the arXiv paper: 2202.09554v1

Comprehensive Summary
Key points
Layman's Summary
Blog article

This research paper addresses the lack of large-scale, open-source datasets for object detection in the construction industry. The authors develop a new dataset called Site Object Detection dAtaset (SODA), specifically collected and annotated for construction sites. The dataset contains 19,846 images with 286,201 objects categorized into 15 classes including workers, materials, machines, and layout. The dataset is diverse and voluminous, covering different site conditions, weather conditions, construction phases, angles, and perspectives. The authors also evaluate the feasibility of the dataset by using two widely-adopted object detection algorithms based on deep learning (YOLO v3/ YOLO v4). The evaluation shows that the dataset is suitable for typical construction scenarios and achieves a maximum mean Average Precision (mAP) of 81.47%. The research team spent a total of 1,246 hours on data cleaning and annotation. Data cleaning accounted for 16.4% of the time while data labeling accounted for 68%, and data checking accounted for 15.6%. The dataset analysis reveals that most images in SODA have a resolution of 1920 * 1080. Worker labels are the most common while machine and layout labels are less frequent. The authors also use the k-means clustering algorithm to analyze the length-width ratio and range of objects in the dataset. They provide statistical results showing the distribution of shooting angles used to capture the images. Overall, this research contributes a large-scale image dataset for deep learning-based object detection methods in the construction industry. It sets up a performance benchmark for evaluating algorithms in this area and provides valuable resources for further research and development in construction site object detection.

- Lack of large-scale, open-source datasets for object detection in the construction industry
- Development of a new dataset called Site Object Detection dAtaset (SODA)
- Dataset contains 19,846 images with 286,201 objects categorized into 15 classes
- Dataset is diverse and voluminous, covering different site conditions, weather conditions, construction phases, angles, and perspectives
- Feasibility evaluation using two widely-adopted object detection algorithms (YOLO v3/ YOLO v4)
- Maximum mean Average Precision (mAP) achieved is 81.47%
- Total of 1,246 hours spent on data cleaning and annotation
- Data cleaning accounted for 16.4% of the time, data labeling accounted for 68%, and data checking accounted for 15.6%
- Most images in SODA have a resolution of 1920 * 1080
- Worker labels are the most common while machine and layout labels are less frequent
- Use of k-means clustering algorithm to analyze length-width ratio and range of objects in the dataset
- Statistical results showing distribution of shooting angles used to capture the images
- Contribution of a large-scale image dataset for deep learning-based object detection methods in the construction industry
- Performance benchmark set up for evaluating algorithms in this area
- Valuable resources provided for further research and development in construction site object detection

Summary1. There weren't many big sets of pictures for finding objects in construction. 2. They made a new set called SODA with lots of pictures and things to find. 3. The set has almost 20,000 pictures with over 280,000 objects sorted into groups. 4. The pictures show different conditions and angles on construction sites. 5. They tested two ways to find objects and got good results. Definitions- Dataset: A collection of information or data. - Object detection: Finding and identifying specific things in a picture or video. - Categorized: Sorted into groups based on similarities or characteristics. - Feasibility evaluation: Testing to see if something is possible or practical. - Average Precision (mAP): A way to measure how well something works by calculating the average accuracy.

Large-Scale Open-Source Dataset for Object Detection in Construction Industry

Object detection is a critical task in the construction industry. It can be used to monitor workers, materials, machines, and layout on construction sites. However, there has been a lack of large-scale open-source datasets for object detection in this domain. To address this issue, researchers from the University of Science and Technology Beijing have developed a new dataset called Site Object Detection dAtaset (SODA). This article will discuss the details of SODA and its evaluation results.

Overview of SODA

The SODA dataset contains 19,846 images with 286,201 objects categorized into 15 classes including workers, materials, machines, and layout. The data was collected from different construction sites with diverse conditions such as weather conditions and construction phases. The images were taken at various angles and perspectives to ensure diversity in the dataset. In total 1,246 hours were spent on data cleaning (16.4%), annotation (68%), and checking (15.6%). Most images have a resolution of 1920 * 1080 pixels while worker labels are most common followed by material labels then machine labels then layout labels respectively.

Evaluation Results

To evaluate the feasibility of SODA for typical construction scenarios two widely adopted deep learning based object detection algorithms YOLO v3/YOLO v4 were used to test it out. The evaluation showed that SODA achieved maximum mean Average Precision (mAP) score of 81.47%. Furthermore k-means clustering algorithm was used to analyze length width ratio range as well as shooting angle distribution which revealed valuable insights about the dataset’s contents such as most objects having an aspect ratio between 0:1 - 3:1 with majority being shot at 45 degree angles or higher .

Conclusion

This research paper contributes a large scale image dataset specifically designed for deep learning based object detection methods in the construction industry setting up performance benchmark for evaluating algorithms in this area along with providing valuable resources for further research development in site object detection tasks .

Created on 26 Dec. 2023

Assess the quality of the AI-generated content by voting

Score: 0

The previous summary was created more than a year ago and can be re-run (if necessary) by clicking on the Run button below.

Similar papers summarized with our AI tools

59.8%

Towards large-scale, automated, accurate detection of CCTV camera objects usi…

cs.CV

57.5%

Fast and Accurate Object Detection on Asymmetrical Receptive Field

cs.CV

56.8%

Transfer Learning of Semantic Segmentation Methods for Identifying Buried Arc…

cs.CV

56.0%

Deep-Learning-based Counting Methods, Datasets, and Applications in Agricultu…

cs.CV

55.3%

A Comprehensive Review of Computer Vision in Sports: Open Issues, Future Tren…

cs.CV

53.8%

Continual Object Detection: A review of definitions, strategies, and challeng…

cs.CV

53.7%

Deep learning in agriculture: A survey

cs.LG

Navigate through even more similar papers through a

tree representation

Look for similar papers (in beta version)

By clicking on the button above, our algorithm will scan all papers in our database to find the closest based on the contents of the full papers and not just on metadata. Please note that it only works for papers that we have generated summaries for and you can rerun it from time to time to get a more accurate result while our database grows.

Disclaimer: The AI-based summarization tool and virtual assistant provided on this website may not always provide accurate and complete summaries or responses. We encourage you to carefully review and evaluate the generated content to ensure its quality and relevance to your needs.