Beyond Segmentation: Structurally Informed Facade Parsing from Imperfect Images
Standard object detectors typically treat architectural elements independently, often resulting in facade parsings that lack the structural coherence required for downstream procedural reconstruction. We address this limitation by augmenting the YOLOv8 training objective with a custom lightweight alignment loss. This regularization encourages grid-consistent arrangements of bounding boxes during training, effectively injecting geometric priors without altering the standard inference pipeline. Experiments on the CMP dataset demonstrate that our method successfully improves structural regularity, correcting alignment errors caused by perspective and occlusion while maintaining a controllable trade-off with standard detection accuracy.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
A MRF Shape Prior for Facade Parsing With Occlusions
We present a new shape prior formalism for segmentation of rectified facade images. It combines the simplicity of split grammars with unprecedented expressive power: the capability of encoding simultaneous alignment in t…
Image SegmentationSegmentationSemantic SegmentationTranslational Symmetry-Aware Facade Parsing for 3D Building Reconstruction
Effectively parsing the facade is essential to 3D building reconstruction, which is an important computer vision problem with a large amount of applications in high precision map for navigation, computer aided design, an…
Semantic ParsingSemantic SegmentationBuilding Facade Parsing R-CNN
Building facade parsing, which predicts pixel-level labels for building facades, has applications in computer vision perception for autonomous vehicle (AV) driving. However, instead of a frontal view, an on-board camera …
Window Detection In Facade Imagery: A Deep Learning Approach Using Mask R-CNN
The parsing of windows in building facades is a long-desired but challenging task in computer vision. It is crucial to urban analysis, semantic reconstruction, lifecycle analysis, digital twins, and scene parsing amongst…
Scene ParsingTransfer LearningUnderOneFacade: Worldwide Facade Semantic Segmentation Benchmark Dataset
Globally consistent semantic digital twins require centimeter-accurate and geographically transferable 3D facade segmentation. However, progress in facade parsing is limited by the lack of large-scale, standardized bench…
Domain GeneralizationSemantic SegmentationPoint Clouds