paper-with-me

Papers

Efficient Motion Modelling with Variable-sized blocks from Hierarchical Cuboidal Partitioning

2022-08-28 · Priyabrata Karmakar, Manzur Murshed, Manoranjan Paul, David Taubman

Motion modelling with block-based architecture has been widely used in video coding where a frame is divided into fixed-sized blocks that are motion compensated independently. This often leads to coding inefficiency as fixed-sized blocks hardly align with the object boundaries. Although hierarchical block-partitioning has been introduced to address this, the increased number of motion vectors limits the benefit. Recently, approximate segmentation of images with cuboidal partitioning has gained popularity. Not only are the variable-sized rectangular segments (cuboids) readily amenable to block-based image/video coding techniques, but they are also capable of aligning well with the object boundaries. This is because cuboidal partitioning is based on a homogeneity constraint, minimising the sum of squared errors (SSE). In this paper, we have investigated the potential of cuboids in motion modelling against the fixed-sized blocks used in scalable video coding. Specifically, we have constructed motion-compensated current frame using the cuboidal partitioning information of the anchor frame in a group-of-picture (GOP). The predicted current frame has then been used as the base layer while encoding the current frame as an enhancement layer using the scalable HEVC encoder. Experimental results confirm 6.71%-10.90% bitrate savings on 4K video sequences.

📄 PDF Abstract BibTeX arXiv:2208.13137

Code (0)

등록된 구현이 없습니다.

Tasks

4k

Methods 이 논문이 사용한 방법론

BASE 설명 없음
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

Learning Design and Construction with Varying-Sized Materials via Prioritized Memory Resets

2022-04-12 · Yunfei Li, Tao Kong, Lei LI, Yi Wu

Can a robot autonomously learn to design and construct a bridge from varying-sized blocks without a blueprint? It is a challenging task with long horizon and sparse reward -- the robot has to figure out physically stable…

Motion Planning

Towards Hierarchical Discrete Variational Autoencoders

2019-10-16 · pproximateinference AABI Symposium 2019 12 · Valentin Liévin, Andrea Dittadi, Lars Maaløe, Ole Winther

Variational Autoencoders (VAEs) have proven to be powerful latent variable models. How- ever, the form of the approximate posterior can limit the expressiveness of the model. Categorical distributions are flexible and us…

Characterising Behavioural Families and Dynamics of Promotional Twitter Bots via Sequence-Based Modelling

2025-12-19 · Ohoud Alzahrani, Russell Beale, Robert J. Hendley arxiv

This paper asks whether promotional Twitter/X bots form behavioural families and whether members evolve similarly. We analyse 2,798,672 tweets from 2,615 ground-truth promotional bot accounts (2006-2021), focusing on com…

Multiple Sequence Alignment

Hierarchical Codec Diffusion for Video-to-Speech Generation

2026-04-17 · Jiaxin Ye, Gaoxiang Cong, Chenhui Wang, Xin-Cheng Wen 외 arxiv

Video-to-Speech (VTS) generation aims to synthesize speech from a silent video without auditory signals. However, existing VTS methods disregard the hierarchical nature of speech, which spans coarse speaker-aware semanti…

Learngene Search Across Multiple Datasets for Building Variable-Sized Models

2026-05-06 · Boyu Shi, Junbo Zhou, Chang Liu, Xu Yang 외 arxiv

Deep learning methods are widely used under diverse resource constraints, resulting in models of varying sizes, such as the Vision Transformer (ViT) series. Deploying these models typically requires costly pretraining an…