Tensor-to-Image: Image-to-Image Translation with Vision Transformers
Transformers gain huge attention since they are first introduced and have a wide range of applications. Transformers start to take over all areas of deep learning and the Vision transformers paper also proved that they can be used for computer vision tasks. In this paper, we utilized a vision transformer-based custom-designed model, tensor-to-image, for the image to image translation. With the help of self-attention, our model was able to generalize and apply to different problems without a single modification.
Code (1)
Tasks
Image-to-Image TranslationTranslationSimilar Papers 제목 키워드 기반
Tensor Decomposition for Model Reduction in Neural Networks: A Review
Modern neural networks have revolutionized the fields of computer vision (CV) and Natural Language Processing (NLP). They are widely used for solving complex CV tasks and NLP tasks such as image classification, image gen…
image-classificationImage ClassificationImage GenerationMachine Translation+1TESA: Tensor Element Self-Attention via Matricization
Representation learning is a fundamental part of modern computer vision, where abstract representations of data are encoded as tensors optimized to solve problems like image segmentation and inpainting. Recently, self-at…
Image InpaintingImage SegmentationInstance SegmentationRepresentation Learning+1GeometricImageNet: Extending convolutional neural networks to vector and tensor images
Convolutional neural networks and their ilk have been very successful for many learning tasks involving images. These methods assume that the input is a scalar image representing the intensity in each pixel, possibly in …
Synthesizer Based Efficient Self-Attention for Vision Tasks
Self-attention module shows outstanding competence in capturing long-range relationships while enhancing performance on vision tasks, such as image classification and image captioning. However, the self-attention module …
Image Captioningimage-classificationImage ClassificationShow and Tell: Lessons learned from the 2015 MSCOCO Image Captioning Challenge
Automatically describing the content of an image is a fundamental problem in artificial intelligence that connects computer vision and natural language processing. In this paper, we present a generative model based on a …
Image CaptioningSentenceTranslation