RewriteNet: Reliable Scene Text Editing with Implicit Decomposition of Text Contents and Styles
Scene text editing (STE), which converts a text in a scene image into the desired text while preserving an original style, is a challenging task due to a complex intervention between text and style. In this paper, we propose a novel STE model, referred to as RewriteNet, that decomposes text images into content and style features and re-writes a text in the original image. Specifically, RewriteNet implicitly distinguishes the content from the style by introducing scene text recognition. Additionally, independent of the exact supervisions with synthetic examples, we propose a self-supervised training scheme for unlabeled real-world images, which bridges the domain gap between synthetic and real data. Our experiments present that RewriteNet achieves better generation performances than other comparisons. Further analysis proves the feature decomposition of RewriteNet and demonstrates the reliability and robustness through diverse experiments. Our implementation is publicly available at \url{https://github.com/clovaai/rewritenet}
Code (0)
등록된 구현이 없습니다.
Tasks
Image GenerationScene Text EditingScene Text RecognitionSimilar Papers 제목 키워드 기반
RewriteNets: End-to-End Trainable String-Rewriting for Generative Sequence Modeling
Dominant sequence models like the Transformer represent structure implicitly through dense attention weights, incurring quadratic complexity. We propose RewriteNets, a novel neural architecture built on an alternative pa…
Panoptic Compositional Feature Field for Editable Scene Rendering With Network-Inferred Labels via Metric Learning
Despite neural implicit representations demonstrating impressive high-quality view synthesis capacity, decomposing such representations into objects for instance-level editing is still challenging. Recent works learn…
2D Panoptic SegmentationMetric LearningNovel View SynthesisPanoptic SegmentationExploring Stroke-Level Modifications for Scene Text Editing
Scene text editing (STE) aims to replace text with the desired one while preserving background and styles of the original text. However, due to the complicated background textures and various text styles, existing method…
AttributeScene Text EditingNeural Parameterization for Dynamic Human Head Editing
Implicit radiance functions emerged as a powerful scene representation for reconstructing and rendering photo-realistic views of a 3D scene. These representations, however, suffer from poor editability. On the other hand…
3D geometryNeuMesh: Learning Disentangled Neural Mesh-based Implicit Field for Geometry and Texture Editing
Very recently neural implicit rendering techniques have been rapidly evolved and shown great advantages in novel view synthesis and 3D scene reconstruction. However, existing neural rendering methods for editing purposes…
3D Scene ReconstructionNeural RenderingNovel View Synthesis