paper-with-me

Papers

Deep learning for Background Replacement in Video Conferencing

2023-06-12 · International Journal of Network Dynamics and Intelligence 2023 6 · Kiran Shahi, Yongmin Li

Background replacement is one of the most used features in video conferencing applications by many people, perhaps mainly for privacy protection, but also for other purposes such as branding, marketing and promoting professionalism. However, the existing applications in video conference tools have serious limitations. Most applications tend to generate strong artefacts (while there is a slight change in the perspective of the background), or require green screens to avoid such artefacts, which results in an unnatural background or even exposes the original background to other users in the video conference. In this work, we aim to study the relationship between the foreground and background in real-time videos. Three different methods are presented and evaluated, including the baseline U-Net, the lightweight U-Net MobileNet, and the U-Net MobileNet&ConvLSTM models. The above models are trained on public datasets for image segmentation. Experimental results show that both the lightweight U-Net MobileNet and the U-Net MobileNet& ConvLSTM models achieve superior performance as compared to the baseline U-Net model.

📄 PDF Abstract BibTeX

Code (2)

kiranshahi/Real-time-Background-replacement-in-Video-Conferencing 공식 구현 tf
FaceOnLive/Realtime-Background-Changer-SDK-Android

Tasks

Deep LearningImage SegmentationMarketingSemantic SegmentationVideo Background Subtraction

Methods 이 논문이 사용한 방법론

Max Pooling Max Pooling is a pooling operation that calculates the maximum value for patches of a feature map, and uses it to create a downsampled (pooled) feature map. It is usually…
ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…
Concatenated Skip Connection A Concatenated Skip Connection is a type of skip connection that seeks to reuse features by concatenating them to new layers, allowing more information to be retained from…
U-Net 설명 없음
Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
ConvLSTM ConvLSTM is a type of recurrent neural network for spatio-temporal prediction that has convolutional structures in both the input-to-state and state-to-state transitions. The…

Similar Papers 제목 키워드 기반

VCD: A Video Conferencing Dataset for Video Compression

2023-09-14 · Babak Naderi, Ross Cutler, Nabakumar Singh Khongbantabam, Yasaman Hosseinkashi 외

Commonly used datasets for evaluating video codecs are all very high quality and not representative of video typically used in video conferencing scenarios. We present the Video Conferencing Dataset (VCD) for evaluating …

Video Compression

Videoconferencing Software Options for Telemedicine: A Review for Movement Disorder Neurologists

2021-10-11 · Frontiers in Neurology 2021 10 · Esther Cubo, Adrián Arnaiz-Rodríguez, Álvar Arnaiz-González, José Francisco Díez-Pastor 외

Background: The use of telemedicine has increased to address the ongoing healthcare needs of patients with movement disorders. Objective: We aimed to describe the technical and basic security features of the most popu…

Articles

Real or Virtual: A Video Conferencing Background Manipulation-Detection System

2022-04-25 · Ehsan Nowroozi, Yassine Mekdad, Mauro Conti, Simone Milani 외

Recently, the popularity and wide use of the last-generation video conferencing technologies created an exponential growth in its market size. Such technology allows participants in different geographic regions to have a…

Do Not Deceive Your Employer with a Virtual Background: A Video Conferencing Manipulation-Detection System

2021-06-29 · Mauro Conti, Simone Milani, Ehsan Nowroozi, Gabriele Orazi

The last-generation video conferencing software allows users to utilize a virtual background to conceal their personal environment due to privacy concerns, especially in official meetings with other employers. On the oth…

PP-HumanSeg: Connectivity-Aware Portrait Segmentation with a Large-Scale Teleconferencing Video Dataset

2021-12-14 · Lutao Chu, Yi Liu, Zewu Wu, Shiyu Tang 외

As the COVID-19 pandemic rampages across the world, the demands of video conferencing surge. To this end, real-time portrait segmentation becomes a popular feature to replace backgrounds of conferencing participants. Whi…

Portrait SegmentationSegmentationSemantic Segmentation