Demystifying the Potential of ChatGPT-4 Vision for Construction Progress Monitoring
–arXiv.org Artificial Intelligence
The integration of Large Vision-Language Models (LVLMs) such as OpenAI's GPT-4 Vision into various sectors has marked a significant evolution in the field of artificial intelligence, particularly in the analysis and interpretation of visual data. This paper explores the practical application of GPT-4 Vision in the construction industry, focusing on its capabilities in monitoring and tracking the progress of construction projects. Utilizing high-resolution aerial imagery of construction sites, the study examines how GPT-4 Vision performs detailed scene analysis and tracks developmental changes over time. The findings demonstrate that while GPT-4 Vision is proficient in identifying construction stages, materials, and machinery, it faces challenges with precise object localization and segmentation. Despite these limitations, the potential for future advancements in this technology is considerable. This research not only highlights the current state and opportunities of using LVLMs in construction but also discusses future directions for enhancing the model's utility through domain-specific training and integration with other computer vision techniques and digital twins.
arXiv.org Artificial Intelligence
Dec-20-2024
- Country:
- Africa > Middle East (0.04)
- Asia > Middle East
- Republic of Türkiye
- Ankara Province > Ankara (0.04)
- Istanbul Province > Istanbul (0.04)
- Republic of Türkiye
- Europe
- Middle East > Republic of Türkiye
- Istanbul Province > Istanbul (0.04)
- United Kingdom > England
- Greater London > London (0.04)
- Middle East > Republic of Türkiye
- Genre:
- Research Report > New Finding (0.48)
- Industry:
- Construction & Engineering (1.00)
- Health & Medicine
- Diagnostic Medicine > Imaging (0.47)
- Therapeutic Area > Oncology (0.48)
- Materials > Construction Materials (0.98)
- Technology: