paper-with-me

홈 › Papers

Visual-Relation Conscious Image Generation from Structured-Text

2019-08-05 · ECCV 2020 8 · Duc Minh Vo, Akihiro Sugimoto

We propose an end-to-end network for image generation from given structured-text that consists of the visual-relation layout module and the pyramid of GANs, namely stacking-GANs. Our visual-relation layout module uses relations among entities in the structured-text in two ways: comprehensive usage and individual usage. We comprehensively use all available relations together to localize initial bounding-boxes of all the entities. We also use individual relation separately to predict from the initial bounding-boxes relation-units for all the relations in the input text. We then unify all the relation-units to produce the visual-relation layout, i.e., bounding-boxes for all the entities so that each of them uniquely corresponds to each entity while keeping its involved relations. Our visual-relation layout reflects the scene structure given in the input text. The stacking-GANs is the stack of three GANs conditioned on the visual-relation layout and the output of previous GAN, consistently capturing the scene structure. Our network realistically renders entities' details in high resolution while keeping the scene structure. Experimental results on two public datasets show outperformances of our method against state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:1908.01741

Code (0)

등록된 구현이 없습니다.

Tasks

AllImage GenerationRelationText-to-Image Generation

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dogecoin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Bridging the Gap Between Consciousness and Matter: Recurrent Out-of-Body Projection of Visual Awareness Revealed by the Law of Non-Identity

2020-01-29 · Jinsong Meng

Consciousness is an explicit outcome of brain activity. However, the link between consciousness and the material world remains to be explored. We applied a new logic tool, the non-identity law, to the analysis of the vis…

Unbiased Scene Graph Generation via Rich and Fair Semantic Extraction

2020-02-01 · Bin Wen, Jie Luo, Xianglong Liu, Lei Huang

Extracting graph representation of visual scenes in image is a challenging task in computer vision. Although there has been encouraging progress of scene graph generation in the past decade, we surprisingly find that the…

Graph GenerationRelationScene Graph GenerationUnbiased Scene Graph Generation

Image Semantic Relation Generation

2022-10-19 · Mingzhe Du

Scene graphs provide structured semantic understanding beyond images. For downstream tasks, such as image retrieval, visual question answering, visual relationship detection, and even autonomous vehicle technology, scene…

Image RetrievalImage SegmentationImage to textQuestion Answering+8

One-shot Scene Graph Generation

2022-02-22 · Yuyu Guo, Jingkuan Song, Lianli Gao, Heng Tao Shen

As a structured representation of the image content, the visual scene graph (visual relationship) acts as a bridge between computer vision and natural language processing. Existing models on the scene graph generation ta…

Graph GenerationScene Graph GenerationTriplet

Assessment of Unconsciousness for Memory Consolidation Using EEG Signals

2020-05-15 · Gi-Hwan Shin, Minji Lee, Seong-Whan Lee

The assessment of consciousness and unconsciousness is a challenging issue in modern neuroscience. Consciousness is closely related to memory consolidation in that memory is a critical component of conscious experience. …

EEGElectroencephalogram (EEG)