Scalability in Building Component Data Annotation: Enhancing Facade Material Classification with Synthetic Data
Computer vision models trained on Google Street View images can create material cadastres. However, current approaches need manually annotated datasets that are difficult to obtain and often have class imbalance. To address these challenges, this paper fine-tuned a Swin Transformer model on a synthetic dataset generated with DALL-E and compared the performance to a similar manually annotated dataset. Although manual annotation remains the gold standard, the synthetic dataset performance demonstrates a reasonable alternative. The findings will ease annotation needed to develop material cadastres, offering architects insights into opportunities for material reuse, thus contributing to the reduction of demolition waste.
Code (0)
등록된 구현이 없습니다.
Tasks
Material ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Zero-shot Building Attribute Extraction from Large-Scale Vision and Language Models
Existing building recognition methods, exemplified by BRAILS, utilize supervised learning to extract information from satellite and street-view images for classification and segmentation. However, each task module requir…
AttributeAttribute ExtractionDescriptiveARCH2S: Dataset, Benchmark and Challenges for Learning Exterior Architectural Structures from Point Clouds
Precise segmentation of architectural structures provides detailed information about various building components, enhancing our understanding and interaction with our built environment. Nevertheless, existing outdoor 3D …
3D Scene Reconstruction3D Semantic SegmentationPoint Cloud GenerationPoint Cloud Segmentation+2Alexa Conversations: An Extensible Data-driven Approach for Building Task-oriented Dialogue Systems
Traditional goal-oriented dialogue systems rely on various components such as natural language understanding, dialogue state tracking, policy learning and response generation. Training each component requires annotations…
Dialogue State TrackingGoal-Oriented Dialogue SystemsNatural Language UnderstandingResponse Generation+1GRIT: Graph-Regularized Logit Refinement for Zero-shot Cell Type Annotation
Cell type annotation is a fundamental step in the analysis of single-cell RNA sequencing (scRNA-seq) data. In practice, human experts often rely on the structure revealed by principal component analysis (PCA) followed by…
Space-LLaVA: a Vision-Language Model Adapted to Extraterrestrial Applications
Foundation Models (FMs), e.g., large language models, possess attributes of intelligence which offer promise to endow a robot with the contextual understanding necessary to navigate complex, unstructured tasks in the wil…
Instruction FollowingLanguage ModelingLanguage ModellingNavigate+2