paper-with-me

Papers

Predicting Perceived Gloss: Do Weak Labels Suffice?

2024-03-26 · Julia Guerrero-Viu, J. Daniel Subias, Ana Serrano, Katherine R. Storrs, Roland W. Fleming, Belen Masia, Diego Gutierrez

Estimating perceptual attributes of materials directly from images is a challenging task due to their complex, not fully-understood interactions with external factors, such as geometry and lighting. Supervised deep learning models have recently been shown to outperform traditional approaches, but rely on large datasets of human-annotated images for accurate perception predictions. Obtaining reliable annotations is a costly endeavor, aggravated by the limited ability of these models to generalise to different aspects of appearance. In this work, we show how a much smaller set of human annotations ("strong labels") can be effectively augmented with automatically derived "weak labels" in the context of learning a low-dimensional image-computable gloss metric. We evaluate three alternative weak labels for predicting human gloss perception from limited annotated data. Incorporating weak labels enhances our gloss prediction beyond the current state of the art. Moreover, it enables a substantial reduction in human annotation costs without sacrificing accuracy, whether working with rendered images or real photographs.

📄 PDF Abstract BibTeX arXiv:2403.17672

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Image Intrinsic Scale Assessment: Bridging the Gap Between Quality and Resolution

2025-02-10 · Vlad Hosu, Lorenzo Agnolucci, Daisuke Iso, Dietmar Saupe

Image Quality Assessment (IQA) measures and predicts perceived image quality by human observers. Although recent studies have highlighted the critical influence that variations in the scale of an image have on its percei…

Image Quality Assessment

Bridging Sign and Spoken Languages: Pseudo Gloss Generation for Sign Language Translation

2025-05-21 · Jianyuan Guo, Peike Li, Trevor Cohn

Sign Language Translation (SLT) aims to map sign language videos to spoken language text. A common approach relies on gloss annotations as an intermediate representation, decomposing SLT into two sub-tasks: video-to-glos…

In-Context LearningLarge Language ModelSign Language TranslationTranslation+1

Massively Multilingual Joint Segmentation and Glossing

2026-01-16 · Michael Ginn, Lindia Tjuatja, Enora Rice, Ali Marashian 외 arxiv

Automated interlinear gloss prediction with neural networks is a promising approach to accelerate language documentation efforts. However, while state-of-the-art models like GlossLM achieve high scores on glossing benchm…

Recurrent Convolutional Neural Networks for Continuous Sign Language Recognition by Staged Optimization

2017-07-01 · CVPR 2017 7 · Runpeng Cui, Hu Liu, Chang-Shui Zhang

This work presents a weakly supervised framework with deep neural networks for vision-based continuous sign language recognition, where the ordered gloss labels but no exact temporal locations are available with the vide…

SentenceSign Language Recognition

Stylistic approaches to predicting Reddit popularity in diglossia

2021-08-01 · ACL 2021 5 · Huikai Chua

Past work investigating what makes a Reddit post popular has indicated that style is a far better predictor than content, where posts conforming to a subreddit{'}s community style are better received. However, what about…