paper-with-me

홈 › Papers

Flow-based Video Segmentation for Human Head and Shoulders

2021-04-20 · Zijian Kuang, Xinran Tie

Video segmentation for the human head and shoulders is essential in creating elegant media for videoconferencing and virtual reality applications. The main challenge is to process high-quality background subtraction in a real-time manner and address the segmentation issues under motion blurs, e.g., shaking the head or waving hands during conference video. To overcome the motion blur problem in video segmentation, we propose a novel flow-based encoder-decoder network (FUNet) that combines both traditional Horn-Schunck optical-flow estimation technique and convolutional neural networks to perform robust real-time video segmentation. We also introduce a video and image segmentation dataset: ConferenceVideoSegmentationDataset. Code and pre-trained models are available on our GitHub repository: \url{https://github.com/kuangzijian/Flow-Based-Video-Matting}.

📄 PDF Abstract BibTeX arXiv:2104.09752

Code (1)

kuangzijian/Flow-Based-Video-Matting pytorch

Tasks

DecoderImage MattingImage SegmentationOptical Flow EstimationSegmentationSemantic SegmentationVideo MattingVideo SegmentationVideo Semantic Segmentation

Similar Papers 제목 키워드 기반

Gaussian Head & Shoulders: High Fidelity Neural Upper Body Avatars with Anchor Gaussian Guided Texture Warping

2024-05-20 · Tianhao Wu, Jing Yang, Zhilin Guo, Jingyi Wan 외

By equipping the most recent 3D Gaussian Splatting representation with head 3D morphable models (3DMM), existing methods manage to create head avatars with high fidelity. However, most existing methods only reconstruct a…

ShoulderShot: Generating Over-the-Shoulder Dialogue Videos

2025-08-11 · Yuang Zhang, Junqi Cheng, Haoyu Zhao, Jiaxi Gu 외 arxiv

Over-the-shoulder dialogue videos are essential in films, short dramas, and advertisements, providing visual variety and enhancing viewers' emotional connection. Despite their importance, such dialogue scenes remain larg…

Video Generation

1st Place Winner of the 2024 Pixel-level Video Understanding in the Wild (CVPR'24 PVUW) Challenge in Video Panoptic Segmentation and Best Long Video Consistency of Video Semantic Segmentation

2024-06-08 · Qingfeng Liu, Mostafa El-Khamy, Kee-Bong Song

The third Pixel-level Video Understanding in the Wild (PVUW CVPR 2024) challenge aims to advance the state of art in video understanding through benchmarking Video Panoptic Segmentation (VPS) and Video Semantic Segmentat…

BenchmarkingInstance SegmentationPanoptic SegmentationScene Parsing+6

OMG-Avatar: One-shot Multi-LOD Gaussian Head Avatar

2026-03-02 · Jianqiang Ren, Lin Liu, Steven Hoi arxiv

We propose OMG-Avatar, a novel One-shot method that leverages a Multi-LOD (Level-of-Detail) Gaussian representation for animatable 3D head reconstruction from a single image in 0.2s. Our method enables LOD head avatar mo…

Computational Efficiency

Portrait Segmentation Using Deep Learning

2022-02-06 · Sumedh Vilas Datar and, Jesus Gonzales Bernal

A portrait is a painting, drawing, photograph, or engraving of a person, especially one depicting only the face or head and shoulders. In the digital world the portrait of a person is captured by having the person as a s…

Deep LearningPortrait SegmentationSegmentation