Research graph
References from Label-free evaluation of state-of-the-art vision–language models on scenario understanding tasks: A case study in unstructured indoor construction sites. Local targets link to admitted publications; unresolved targets remain external evidence.
10.1109/iccvw69036.2025.00154
10.1109/iccvw69036.2025.00154 · External reference
Unresolved reference
2025 · External reference
Unresolved reference
2025 · External reference
Convolutional neural networks for construction safety: A technical review of computer vision applications
10.1016/j.asoc.2025.113374 · 2025 · External reference
10.1109/icipw68931.2025.11386061
10.1109/icipw68931.2025.11386061 · External reference
Eagle 2.5: Boosting long-context post-training for frontier vision-language models
2025 · External reference
10.1109/cvpr52733.2024.02283
10.1109/cvpr52733.2024.02283 · External reference
10.22260/isarc2025/0072
10.22260/isarc2025/0072 · External reference
10.22260/isarc2026/0084
10.22260/isarc2026/0084 · External reference
Can large vision-language models understand construction safety? A novel benchmark using construction safety posters
2024 · External reference
Are large pre-trained vision language models effective construction safety inspectors
10.1017/dce.2026.10044 · 2026 · External reference
Unresolved reference
2026 · External reference
Computer vision-based interior construction progress monitoring: A literature review and future research directions
10.1016/j.autcon.2021.103705 · 2021 · External reference
Applications of multimodal large language models in construction industry
10.1016/j.aei.2025.103909 · 2026 · External reference
Integration and evaluation of a 3D LiDAR SLAM system for construction robots in large-scale public building sites
2026 · External reference
Multimodal fusion and vision–language models: A survey for robot vision
10.1016/j.inffus.2025.103652 · 2026 · External reference
Intelligent quality assessment of concrete vibration using computer vision and large language models
10.1016/j.autcon.2025.106507 · 2025 · External reference
Recognizing temporary construction site objects using CLIP-based few-shot learning and multi-modal prototypes
10.1016/j.autcon.2024.105542 · 2024 · External reference
Unresolved reference
2024 · External reference
10.22260/isarc2026/0117
10.22260/isarc2026/0117 · External reference
Real-time safety detection on construction sites using a vision-language and NLP-based model
10.1016/j.aei.2025.103889 · 2026 · External reference
Tandem CCV-VLM: Visual construction safety inspection based on collaborative cross-verification vision-language model
10.1016/j.aei.2026.104841 · 2026 · External reference
Efficient GPT-4V level multimodal large language model for deployment on edge devices
10.1038/s41467-025-61040-5 · 2025 · External reference
Autonomous mobile construction robots in built environment: A comprehensive review
2024 · External reference
Vision-language models for vision tasks: A survey
10.1109/tpami.2024.3369699 · 2024 · External reference
Unresolved reference
2025 · External reference
Training-free few-shot construction tool and material detection using pre-trained vision-language model
10.1111/mice.70129 · 2025 · External reference
Unresolved reference
2025 · External reference
Real-time safety detection on construction sites using a vision-language and NLP-based model
10.1016/j.aei.2025.103889 · ExternalCitation · doi-reference
Applications of multimodal large language models in construction industry
10.1016/j.aei.2025.103909 · ExternalCitation · doi-reference
Tandem CCV-VLM: Visual construction safety inspection based on collaborative cross-verification vision-language model
10.1016/j.aei.2026.104841 · ExternalCitation · doi-reference
Convolutional neural networks for construction safety: A technical review of computer vision applications
10.1016/j.asoc.2025.113374 · ExternalCitation · doi-reference
Computer vision-based interior construction progress monitoring: A literature review and future research directions
10.1016/j.autcon.2021.103705 · ExternalCitation · doi-reference
Recognizing temporary construction site objects using CLIP-based few-shot learning and multi-modal prototypes
10.1016/j.autcon.2024.105542 · ExternalCitation · doi-reference
Intelligent quality assessment of concrete vibration using computer vision and large language models
10.1016/j.autcon.2025.106507 · ExternalCitation · doi-reference
Multimodal fusion and vision–language models: A survey for robot vision
10.1016/j.inffus.2025.103652 · ExternalCitation · doi-reference
Are large pre-trained vision language models effective construction safety inspectors
10.1017/dce.2026.10044 · ExternalCitation · doi-reference
Efficient GPT-4V level multimodal large language model for deployment on edge devices
10.1038/s41467-025-61040-5 · ExternalCitation · doi-reference
10.1109/cvpr52733.2024.02283
10.1109/cvpr52733.2024.02283 · ExternalCitation · doi-reference
10.1109/iccvw69036.2025.00154
10.1109/iccvw69036.2025.00154 · ExternalCitation · doi-reference
10.1109/icipw68931.2025.11386061
10.1109/icipw68931.2025.11386061 · ExternalCitation · doi-reference
Vision-language models for vision tasks: A survey
10.1109/tpami.2024.3369699 · ExternalCitation · doi-reference
Training-free few-shot construction tool and material detection using pre-trained vision-language model
10.1111/mice.70129 · ExternalCitation · doi-reference
10.22260/isarc2025/0072
10.22260/isarc2025/0072 · ExternalCitation · doi-reference
10.22260/isarc2026/0084
10.22260/isarc2026/0084 · ExternalCitation · doi-reference
10.22260/isarc2026/0117
10.22260/isarc2026/0117 · ExternalCitation · doi-reference