paper-with-me

홈 › Papers

AutoRepo: A general framework for multi-modal LLM-based automated construction reporting

2023-10-11 · Hongxu Pu, Xincong Yang, Jing Li, Runhao Guo, Heng Li

Ensuring the safety, quality, and timely completion of construction projects is paramount, with construction inspections serving as a vital instrument towards these goals. Nevertheless, the predominantly manual approach of present-day inspections frequently results in inefficiencies and inadequate information management. Such methods often fall short of providing holistic, exhaustive assessments, consequently engendering regulatory oversights and potential safety hazards. To address this issue, this paper presents a novel framework named AutoRepo for automated generation of construction inspection reports. The unmanned vehicles efficiently perform construction inspections and collect scene information, while the multimodal large language models (LLMs) are leveraged to automatically generate the inspection reports. The framework was applied and tested on a real-world construction site, demonstrating its potential to expedite the inspection process, significantly reduce resource allocation, and produce high-quality, regulatory standard-compliant inspection reports. This research thus underscores the immense potential of multimodal large language models in revolutionizing construction inspection practices, signaling a significant leap forward towards a more efficient and safer construction management paradigm.

📄 PDF Abstract BibTeX arXiv:2310.07944

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

Automated Ensemble Multimodal Machine Learning for Healthcare

2024-07-25 · Fergus Imrie, Stefan Denner, Lucas S. Brunschwig, Klaus Maier-Hein 외

The application of machine learning in medicine and healthcare has led to the creation of numerous diagnostic and prognostic models. However, despite their success, current approaches generally issue predictions using da…

Decision MakingDiagnosticEnsemble Learning

PersonaVlog: Personalized Multimodal Vlog Generation with Multi-Agent Collaboration and Iterative Self-Correction

2025-08-19 · Xiaolu Hou, Bing Ma, Jiaxiang Cheng, Xuhua Ren 외 arxiv

With the growing demand for short videos and personalized content, automated Video Log (Vlog) generation has become a key direction in multimodal content creation. Existing methods mostly rely on predefined scripts, lack…

End-To-End multi-modal sensors fusion system for urban automated driving

2018-10-10 · Ibrahim Sobh, Loay Amin, Sherif Abdelkarim, Khaled Elmadawy 외

In this paper, we present a novel framework for urban automated driving based on multi-modal sensors; LiDAR and Camera. Environment perception through sensors fusion is key to successful deployment of automated driving s…

SegmentationSemantic Segmentation

MODALS: Modality-agnostic Automated Data Augmentation in the Latent Space

2021-01-01 · ICLR 2021 1 · Tsz-Him Cheung, Dit-yan Yeung

Data augmentation is an efficient way to expand a training dataset by creating additional artificial data. While data augmentation is found to be effective in improving the generalization capability of models for various…

Data AugmentationTime SeriesTime Series Analysis

AutoM3L: An Automated Multimodal Machine Learning Framework with Large Language Models

2024-08-01 · Daqin Luo, Chengjian Feng, Yuxuan Nong, Yiqing Shen

Automated Machine Learning (AutoML) offers a promising approach to streamline the training of machine learning models. However, existing AutoML frameworks are often limited to unimodal scenarios and require extensive man…

AutoMLCode GenerationFeature EngineeringHyperparameter Optimization