paper-with-me

Papers

Matching visual induction effects on screens of different size

2020-05-06 · Trevor D. Canham, Javier Vazquez-Corral, Elise Mathieu, Marcelo Bertalmío

In the film industry, the same movie is expected to be watched on displays of vastly different sizes, from cinema screens to mobile phones. But visual induction, the perceptual phenomenon by which the appearance of a scene region is affected by its surroundings, will be different for the same image shown on two displays of different dimensions. This presents a practical challenge for the preservation of the artistic intentions of filmmakers, as it can lead to shifts in image appearance between viewing destinations. In this work we show that a neural field model based on the efficient representation principle is able to predict induction effects, and how by regularizing its associated energy functional the model is still able to represent induction but is now invertible. From this we propose a method to pre-process an image in a screen-size dependent way so that its perception, in terms of visual induction, may remain constant across displays of different size. The potential of the method is demonstrated through psychophysical experiments on synthetic images and qualitative examples on natural images.

📄 PDF Abstract BibTeX arXiv:2005.02694

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

UI2App: Benchmarking Visual Interaction Inference in Executable Web Application Generation

2026-07-07 · Grace Man Chen, Litao Guo, Yifan Wu, Yiyu Chen 외 arxiv

Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex prompts that impose substantial demands on users and offer limited expres…

VISTA: An End-to-End Benchmark for Visual Spec-to-Web-App Coding Agents

2026-05-22 · JunJia Guo, Yuhang Yao, Jiawei, Zhou 외 arxiv

We present VISTA (VIsual Spec-To-App Benchmark), a benchmark for evaluating the end-to-end web-app generation capabilities of LLM-based agents. Unlike prior code generation benchmarks that focus on algorithmic tasks, VIS…

Code Generation

MaskClaw: Edge-Side Personalized Privacy Arbitration for GUI Agents with Behavior-Driven Skill Evolution

2026-05-27 · Yanqiu Zhao, Dongying Zheng, Kaibo Huang, Yukun Wei 외 arxiv

GUI agents rely on screenshots to infer intent and operate across applications, but these screenshots often contain private messages, medical records, payment credentials, and workplace-specific workflows. Privacy decisi…

Unsupervised Vision-Language Grammar Induction with Shared Structure Modeling

2021-09-29 · ICLR 2022 4 · Bo Wan, Wenjuan Han, Zilong Zheng, Tinne Tuytelaars

We introduce a new task, unsupervised vision-language (VL) grammar induction. Given an image-caption pair, the goal is to extract a shared hierarchical structure for both image and language simultaneously. We argue that…

Contrastive LearningPhrase Grounding

WebGen-Agent: Enhancing Interactive Website Generation with Multi-Level Feedback and Step-Level Reinforcement Learning

2025-09-26 · Zimu Lu, Houxing Ren, Yunqiao Yang, Ke Wang 외 arxiv

Agent systems powered by large language models (LLMs) have demonstrated impressive performance on repository-level code-generation tasks. However, for tasks such as website codebase generation, which depend heavily on vi…

Reinforcement Learning