paper-with-me

홈 › Papers

Unsupervised Image to Sequence Translation with Canvas-Drawer Networks

2018-09-21 · Kevin Frans, Chin-Yi Cheng

Encoding images as a series of high-level constructs, such as brush strokes or discrete shapes, can often be key to both human and machine understanding. In many cases, however, data is only available in pixel form. We present a method for generating images directly in a high-level domain (e.g. brush strokes), without the need for real pairwise data. Specifically, we train a "canvas" network to imitate the mapping of high-level constructs to pixels, followed by a high-level "drawing" network which is optimized through this mapping towards solving a desired image recreation or translation task. We successfully discover sequential vector representations of symbols, large sketches, and 3D objects, utilizing only pixel data. We display applications of our method in image segmentation, and present several ablation studies comparing various configurations.

📄 PDF Abstract BibTeX arXiv:1809.08340

Code (1)

wgoldie/canvasdrawer pytorch

Tasks

Image SegmentationSemantic SegmentationTranslation

Similar Papers 제목 키워드 기반

CoDraw: Collaborative Drawing as a Testbed for Grounded Goal-driven Communication

2017-12-15 · ACL 2019 7 · Jin-Hwa Kim, Nikita Kitaev, Xinlei Chen, Marcus Rohrbach 외

In this work, we propose a goal-driven collaborative task that combines language, perception, and action. Specifically, we develop a Collaborative image-Drawing game between two agents, called CoDraw. Our game is grounde…

Imitation Learning

Translation Canvas: An Explainable Interface to Pinpoint and Analyze Translation Systems

2024-10-07 · Chinmay Dandekar, Wenda Xu, Xi Xu, Siqi Ouyang 외

With the rapid advancement of machine translation research, evaluation toolkits have become essential for benchmarking system progress. Tools like COMET and SacreBLEU offer single quality score assessments that are effec…

BenchmarkingMachine TranslationTranslation

CanvasVAE: Learning to Generate Vector Graphic Documents

2021-08-03 · ICCV 2021 10 · Kota Yamaguchi

Vector graphic documents present visual elements in a resolution free, compact format and are often seen in creative applications. In this work, we attempt to learn a generative model of vector graphic documents. We defi…

Length-Adaptive Decoding for Masked Diffusion Machine Translation

2026-08-23 · Yan Zhan, Mengkai Hou, Wanting Zhang, Zhijun Gao hf

Machine translation tests masked diffusion language models (dLLMs) because every source token must be rendered faithfully, while fixed canvas decoding must choose target length before denoising. Existing masked diffusion…

Machine Translation

Multi-task Sequence to Sequence Learning

2015-11-19 · Minh-Thang Luong, Quoc V. Le, Ilya Sutskever, Oriol Vinyals 외

Sequence to sequence learning has recently emerged as a new paradigm in supervised learning. To date, most of its applications focused on only one task and not much work explored this framework for multiple tasks. This p…

Caption GenerationDecoderMachine TranslationMulti-Task Learning+1