paper-with-me

홈 › Papers

Generating Visual Scenes from Touch

2023-09-26 · ICCV 2023 1 · Fengyu Yang, Jiacheng Zhang, Andrew Owens

An emerging line of work has sought to generate plausible imagery from touch. Existing approaches, however, tackle only narrow aspects of the visuo-tactile synthesis problem, and lag significantly behind the quality of cross-modal synthesis methods in other domains. We draw on recent advances in latent diffusion to create a model for synthesizing images from tactile signals (and vice versa) and apply it to a number of visuo-tactile synthesis tasks. Using this model, we significantly outperform prior work on the tactile-driven stylization problem, i.e., manipulating an image to match a touch signal, and we are the first to successfully generate images from touch without additional sources of information about the scene. We also successfully use our model to address two novel synthesis problems: generating images that do not contain the touch sensor or the hand holding it, and estimating an image's shading from its reflectance and touch.

📄 PDF Abstract BibTeX arXiv:2309.15117

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Visual to Sound: Generating Natural Sound for Videos in the Wild

2017-12-04 · CVPR 2018 6 · Yipin Zhou, Zhaowen Wang, Chen Fang, Trung Bui 외

As two of the five traditional human senses (sight, hearing, taste, smell, and touch), vision and sound are basic sources through which humans understand the world. Often correlated during natural events, these two modal…

Touch and Go: Learning from Human-Collected Vision and Touch

2022-11-22 · Fengyu Yang, Chenyang Ma, Jiacheng Zhang, Jing Zhu 외

The ability to associate touch with sight is essential for tasks that require physically interacting with objects in the world. We propose a dataset with paired visual and tactile data called Touch and Go, in which human…

Image Stylization

Draw Like an Artist: Complex Scene Generation with Diffusion Model via Composition, Painting, and Retouching

2024-08-25 · Minghao Liu, Le Zhang, Yingjie Tian, Xiaochao Qu 외

Recent advances in text-to-image diffusion models have demonstrated impressive capabilities in image quality. However, complex scene generation remains relatively unexplored, and even the definition of `complex scene' it…

Scene Generation

Touch-GS: Visual-Tactile Supervised 3D Gaussian Splatting

2024-03-14 · Aiden Swann, Matthew Strong, Won Kyung Do, Gadiel Sznaier Camps 외

In this work, we propose a novel method to supervise 3D Gaussian Splatting (3DGS) scenes using optical tactile sensors. Optical tactile sensors have become widespread in their use in robotics for manipulation and object …

3DGSDepth EstimationMonocular Depth EstimationTransparent objects

Mozart's Touch: A Lightweight Multi-modal Music Generation Framework Based on Pre-Trained Large Models

2024-05-05 · Jiajun Li, Tianze Xu, Xuesong Chen, Xinrui Yao 외

In recent years, AI-Generated Content (AIGC) has witnessed rapid advancements, facilitating the creation of music, images, and other artistic forms across a wide range of industries. However, current models for image- an…

DescriptiveLanguage ModelingLanguage ModellingLarge Language Model+2