paper-with-me

홈 › Papers

SketchParse : Towards Rich Descriptions for Poorly Drawn Sketches using Multi-Task Hierarchical Deep Networks

2017-09-05 · Ravi Kiran Sarvadevabhatla, Isht Dwivedi, Abhijat Biswas, Sahil Manocha, R. Venkatesh Babu

The ability to semantically interpret hand-drawn line sketches, although very challenging, can pave way for novel applications in multimedia. We propose SketchParse, the first deep-network architecture for fully automatic parsing of freehand object sketches. SketchParse is configured as a two-level fully convolutional network. The first level contains shared layers common to all object categories. The second level contains a number of expert sub-networks. Each expert specializes in parsing sketches from object categories which contain structurally similar parts. Effectively, the two-level configuration enables our architecture to scale up efficiently as additional categories are added. We introduce a router layer which (i) relays sketch features from shared layers to the correct expert (ii) eliminates the need to manually specify object category during inference. To bypass laborious part-level annotation, we sketchify photos from semantic object-part image datasets and use them for training. Our architecture also incorporates object pose prediction as a novel auxiliary task which boosts overall performance while providing supplementary information regarding the sketch. We demonstrate SketchParse's abilities (i) on two challenging large-scale sketch datasets (ii) in parsing unseen, semantically related object categories (iii) in improving fine-grained sketch-based image retrieval. As a novel application, we also outline how SketchParse's output can be used to generate caption-style descriptions for hand-drawn sketches.

📄 PDF Abstract BibTeX arXiv:1709.01295

Code (1)

val-iisc/sketch-parse 공식 구현 pytorch

Tasks

Image RetrievalObjectPose PredictionRetrievalSketch-Based Image Retrieval

Similar Papers 제목 키워드 기반

Sketch Me if You Can: Towards Generating Detailed Descriptions of Object Shape by Grounding in Images and Drawings

2019-10-01 · WS 2019 10 · Ting Han, Sina Zarrie{\ss}

A lot of recent work in Language {\&} Vision has looked at generating descriptions or referring expressions for objects in scenes of real-world images, though focusing mostly on relatively simple language like object nam…

AttributeImage CaptioningObject

Deep Plastic Surgery: Robust and Controllable Image Editing with Human-Drawn Sketches

2020-01-09 · ECCV 2020 8 · Shuai Yang, Zhangyang Wang, Jiaying Liu, Zongming Guo

Sketch-based image editing aims to synthesize and modify photos based on the structural information provided by the human-drawn sketches. Since sketches are difficult to collect, previous methods mainly use edge maps ins…

Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval

2026-04-17 · Siyuan Wang, Hanchen Gao, Guangming Zhu, Jiang Lu 외 arxiv

Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-drawn sketches capture complex structural contours, they lack color and …

Image Retrieval

Sketch2Model: View-Aware 3D Modeling from Single Free-Hand Sketches

2021-05-14 · CVPR 2021 1 · Song-Hai Zhang, Yuan-Chen Guo, Qing-Wen Gu

We investigate the problem of generating 3D meshes from single free-hand sketches, aiming at fast 3D modeling for novice users. It can be regarded as a single-view reconstruction problem, but with unique challenges, brou…

A Sketch Is Worth a Thousand Words: Image Retrieval with Text and Sketch

2022-08-05 · Patsorn Sangkloy, Wittawat Jitkrittum, Diyi Yang, James Hays

We address the problem of retrieving images with both a sketch and a text query. We present TASK-former (Text And SKetch transformer), an end-to-end trainable model for image retrieval using a text description and a sket…

Image RetrievalRetrieval