Visual-Relation Conscious Image Generation from Structured-Text
We propose an end-to-end network for image generation from given structured-text that consists of the visual-relation layout module and the pyramid of GANs, namely stacking-GANs. Our visual-relation layout module uses relations among entities in the structured-text in two ways: comprehensive usage and individual usage. We comprehensively use all available relations together to localize initial bounding-boxes of all the entities. We also use individual relation separately to predict from the initial bounding-boxes relation-units for all the relations in the input text. We then unify all the relation-units to produce the visual-relation layout, i.e., bounding-boxes for all the entities so that each of them uniquely corresponds to each entity while keeping its involved relations. Our visual-relation layout reflects the scene structure given in the input text. The stacking-GANs is the stack of three GANs conditioned on the visual-relation layout and the output of previous GAN, consistently capturing the scene structure. Our network realistically renders entities' details in high resolution while keeping the scene structure. Experimental results on two public datasets show outperformances of our method against state-of-the-art methods.
Code (0)
등록된 구현이 없습니다.
Tasks
AllImage GenerationRelationText-to-Image GenerationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Bridging the Gap Between Consciousness and Matter: Recurrent Out-of-Body Projection of Visual Awareness Revealed by the Law of Non-Identity
Consciousness is an explicit outcome of brain activity. However, the link between consciousness and the material world remains to be explored. We applied a new logic tool, the non-identity law, to the analysis of the vis…
Unbiased Scene Graph Generation via Rich and Fair Semantic Extraction
Extracting graph representation of visual scenes in image is a challenging task in computer vision. Although there has been encouraging progress of scene graph generation in the past decade, we surprisingly find that the…
Graph GenerationRelationScene Graph GenerationUnbiased Scene Graph GenerationImage Semantic Relation Generation
Scene graphs provide structured semantic understanding beyond images. For downstream tasks, such as image retrieval, visual question answering, visual relationship detection, and even autonomous vehicle technology, scene…
Image RetrievalImage SegmentationImage to textQuestion Answering+8One-shot Scene Graph Generation
As a structured representation of the image content, the visual scene graph (visual relationship) acts as a bridge between computer vision and natural language processing. Existing models on the scene graph generation ta…
Graph GenerationScene Graph GenerationTripletAssessment of Unconsciousness for Memory Consolidation Using EEG Signals
The assessment of consciousness and unconsciousness is a challenging issue in modern neuroscience. Consciousness is closely related to memory consolidation in that memory is a critical component of conscious experience. …
EEGElectroencephalogram (EEG)