paper-with-me

홈 › Papers

Latent Compass: Creation by Navigation

2020-12-20 · Sarah Schwettmann, Hendrik Strobelt, Mauro Martino

In Marius von Senden's Space and Sight, a newly sighted blind patient describes the experience of a corner as lemon-like, because corners "prick" sight like lemons prick the tongue. Prickliness, here, is a dimension in the feature space of sensory experience, an effect of the perceived on the perceiver that arises where the two interact. In the account of the newly sighted, an effect familiar from one interaction translates to a novel context. Perception serves as the vehicle for generalization, in that an effect shared across different experiences produces a concrete abstraction grounded in those experiences. Cezanne and the post-impressionists, fluent in the language of experience translation, realized that the way to paint a concrete form that best reflected reality was to paint not what they saw, but what it was like to see. We envision a future of creation using AI where what it is like to see is replicable, transferrable, manipulable - part of the artist's palette that is both grounded in a particular context, and generalizable beyond it. An active line of research maps human-interpretable features onto directions in GAN latent space. Supervised and self-supervised approaches that search for anticipated directions or use off-the-shelf classifiers to drive image manipulation in embedding space are limited in the variety of features they can uncover. Unsupervised approaches that discover useful new directions show that the space of perceptually meaningful directions is nowhere close to being fully mapped. As this space is broad and full of creative potential, we want tools for direction discovery that capture the richness and generalizability of human perception. Our approach puts creators in the discovery loop during real-time tool use, in order to identify directions that are perceptually meaningful to them, and generate interpretable image translations along those directions.

📄 PDF Abstract BibTeX arXiv:2012.14283

Code (0)

등록된 구현이 없습니다.

Tasks

Image Manipulation

Similar Papers 제목 키워드 기반

Visualizing Temporal Topic Embeddings with a Compass

2024-09-16 · Daniel Palamarchuk, Lemara Williams, Brian Mayer, Thomas Danielson 외

Dynamic topic modeling is useful at discovering the development and change in latent topics over time. However, present methodology relies on algorithms that separate document and word representations. This prevents the …

DiversityDynamic Topic ModelingWord Embeddings

Diffusion Denoiser-Aided Gyrocompassing

2025-07-28 · Gershy Ben-Arie, Daniel Engelsman, Rotem Dror, Itzik Klein arxiv

An accurate initial heading angle is essential for efficient and safe navigation across diverse domains. Unlike magnetometers, gyroscopes can provide accurate heading reference independent of the magnetic disturbances in…

Autonomous Vehicles

Latent Image Animator: Learning to Animate Images via Latent Space Navigation

2022-03-17 · Yaohui Wang, Di Yang, Francois Bremond, Antitza Dantcheva

Due to the remarkable progress of deep generative models, animating images has become increasingly efficient, whereas associated results have become increasingly realistic. Current animation-approaches commonly exploit s…

Latent Image Animator: Learning to animate image via latent space navigation

2021-09-29 · ICLR 2022 4 · Yaohui Wang, Di Yang, Francois Bremond, Antitza Dantcheva

Animating images has become increasingly realistic, as well as efficient due to the remarkable progress of Generative Adversarial Networks (GANs) and auto-encoder. Current animation-approaches commonly exploit structure …

Image AnimationVideo Generation

NeoNav: Improving the Generalization of Visual Navigation via Generating Next Expected Observations

2019-06-17 · Qiaoyun Wu, Dinesh Manocha, Jun Wang, Kai Xu

We propose improving the cross-target and cross-scene generalization of visual navigation through learning an agent that is guided by conceiving the next observations it expects to see. This is achieved by learning a var…

Visual Navigation