paper-with-me

홈 › Papers

BiOcularGAN: Bimodal Synthesis and Annotation of Ocular Images

2022-05-03 · Darian Tomašević, Peter Peer, Vitomir Štruc

Current state-of-the-art segmentation techniques for ocular images are critically dependent on large-scale annotated datasets, which are labor-intensive to gather and often raise privacy concerns. In this paper, we present a novel framework, called BiOcularGAN, capable of generating synthetic large-scale datasets of photorealistic (visible light and near-infrared) ocular images, together with corresponding segmentation labels to address these issues. At its core, the framework relies on a novel Dual-Branch StyleGAN2 (DB-StyleGAN2) model that facilitates bimodal image generation, and a Semantic Mask Generator (SMG) component that produces semantic annotations by exploiting latent features of the DB-StyleGAN2 model. We evaluate BiOcularGAN through extensive experiments across five diverse ocular datasets and analyze the effects of bimodal data generation on image quality and the produced annotations. Our experimental results show that BiOcularGAN is able to produce high-quality matching bimodal images and annotations (with minimal manual intervention) that can be used to train highly competitive (deep) segmentation models (in a privacy aware-manner) that perform well across multiple real-world datasets. The source code for the BiOcularGAN framework is publicly available at https://github.com/dariant/BiOcularGAN.

📄 PDF Abstract BibTeX arXiv:2205.01536

Code (1)

dariant/BiOcularGAN 공식 구현 pytorch

Tasks

Image GenerationSegmentation

Methods 이 논문이 사용한 방법론

Path Length Regularization 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
R1 Regularization R_INLINE_MATH_1 Regularization is a regularization technique and gradient penalty for training [generative adversarial…
HuMan(Expedia)||How do I get a human at Expedia? How do I get a human at Expedia? How Do I Get a Human at Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Real-Time Help & Exclusive…
Weight Demodulation 설명 없음

Similar Papers 제목 키워드 기반

CoBiLiRo: A Research Platform for Bimodal Corpora

2020-05-01 · LREC 2020 5 · Dan Cristea, Ionu{\textcommabelow{t}} Pistol, {\textcommabelow{S}}erban Boghiu, Anca-Diana Bibiri 외

This paper describes the on-going work carried out within the CoBiLiRo (Bimodal Corpus for Romanian Language) research project, part of ReTeRom (Resources and Technologies for Developing Human-Machine Interfaces in Roman…

speech-recognitionSpeech Recognition

LabelAny3D: Label Any Object 3D in the Wild

2026-01-04 · Jin Yao, Radowan Mahmud Redoy, Sebastian Elbaum, Matthew B. Dwyer 외 arxiv

Detecting objects in 3D space from monocular input is crucial for applications ranging from robotics to scene understanding. Despite advanced performance in the indoor and autonomous driving domains, existing monocular 3…

Scene UnderstandingAutonomous Driving

EfficientDepth: A Fast and Detail-Preserving Monocular Depth Estimation Model

2025-09-26 · Andrii Litvynchuk, Ivan Livinsky, Anand Ravi, Nima Kalantari 외 arxiv

Monocular depth estimation (MDE) plays a pivotal role in various computer vision applications, such as robotics, augmented reality, and autonomous driving. Despite recent advancements, existing methods often fail to meet…

Monocular Depth EstimationAutonomous Driving3D Reconstruction

DeepFaceFlow: In-the-wild Dense 3D Facial Motion Estimation

2020-05-14 · CVPR 2020 6 · Mohammad Rami Koujan, Anastasios Roussos, Stefanos Zafeiriou

Dense 3D facial motion capture from only monocular in-the-wild pairs of RGB images is a highly challenging problem with numerous applications, ranging from facial expression recognition to facial reenactment. In this wor…

3D ReconstructionFacial Expression RecognitionFacial Expression Recognition (FER)Motion Estimation

Weakly Paired Associative Learning for Sound and Image Representations via Bimodal Associative Memory

2022-01-01 · CVPR 2022 1 · Sangmin Lee, Hyung-Il Kim, Yong Man Ro

Data representation learning without labels has attracted increasing attention due to its nature that does not require human annotation. Recently, representation learning has been extended to bimodal data, especially…

Representation Learning