paper-with-me

Papers

Interactive Optimization of Generative Image Modeling using Sequential Subspace Search and Content-based Guidance

2019-06-24 · Toby Chong Long Hin, I-Chao Shen, Issei Sato, Takeo Igarashi

Generative image modeling techniques such as GAN demonstrate highly convincing image generation result. However, user interaction is often necessary to obtain the desired results. Existing attempts add interactivity but require either tailored architectures or extra data. We present a human-in-the-optimization method that allows users to directly explore and search the latent vector space of generative image modeling. Our system provides multiple candidates by sampling the latent vector space, and the user selects the best blending weights within the subspace using multiple sliders. In addition, the user can express their intention through image editing tools. The system samples latent vectors based on inputs and presents new candidates to the user iteratively. An advantage of our formulation is that one can apply our method to arbitrary pre-trained model without developing specialized architecture or data. We demonstrate our method with various generative image modeling applications, and show superior performance in a comparative user study with prior art iGAN.

📄 PDF Abstract BibTeX arXiv:1906.09840

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…
Dogecoin Customer Service Number +1-833-534-1729 설명 없음

Similar Papers 제목 키워드 기반

Sequential Attention GAN for Interactive Image Editing

2018-12-20 · Yu Cheng, Zhe Gan, Yitong Li, Jingjing Liu 외

Most existing text-to-image synthesis tasks are static single-turn generation, based on pre-defined textual descriptions of images. To explore more practical and interactive real-life applications, we introduce a new tas…

Image DescriptionImage GenerationText-to-Image Generation

An item is worth one token in Multimodal Large Language Models-based Sequential Recommendation

2025-11-08 · Qiyong Zhong, Jiajie Su, Ming Yang, Yunshan Ma 외 arxiv

Sequential recommendations (SR) predict users' future interactions based on their historical behavior. The rise of Large Language Models (LLMs) has brought powerful generative and reasoning capabilities, significantly en…

Sequential Recommendation

MotionLM: Multi-Agent Motion Forecasting as Language Modeling

2023-09-28 · ICCV 2023 1 · Ari Seff, Brian Cera, Dian Chen, Mason Ng 외

Reliable forecasting of the future behavior of road agents is a critical component to safe planning in autonomous vehicles. Here, we represent continuous trajectories as sequences of discrete motion tokens and cast multi…

Autonomous VehiclesLanguage ModelingLanguage ModellingMotion Forecasting+1

Physically-guided Image Generation for Multi-Projection Mapping

2026-06-21 · Xingyun Liu, Yuqi Li, Jinhui Xiang, Pinyan Tang 외 arxiv

Projection Mapping (PM) enables seamless superimposition of digital content onto real-world 3D objects, serving as a fundamental technique for immersive visualization, digital twins, and interactive art. Although text-to…

Image Generation

Learning the Base Distribution in Implicit Generative Models

2018-03-12 · Cem Subakan, Oluwasanmi Koyejo, Paris Smaragdis

Popular generative model learning methods such as Generative Adversarial Networks (GANs), and Variational Autoencoders (VAE) enforce the latent representation to follow simple distributions such as isotropic Gaussian. In…