paper-with-me

홈 › Papers

Saving the Sonorine: Photovisual Audio Recovery Using Image Processing and Computer Vision Techniques

2020-05-16 · Kevin Feng

This paper presents a novel technique to recover audio from sonorines, an early 20th century form of analogue sound storage. Our method uses high resolution photographs of sonorines under different lighting conditions to observe the change in reflection behavior of the physical surface features and create a three-dimensional height map of the surface. Sound can then be extracted using height information within the surface's grooves, mimicking a physical stylus on a phonograph. Unlike traditional playback methods, our method has the advantage of being contactless: the medium will not incur damage and wear from being played repeatedly. We compare the results of our technique to a previously successful contactless method using flatbed scans of the sonorines, and conclude with future research that can be applied to this photovisual approach to audio recovery.

📄 PDF Abstract BibTeX arXiv:2005.08944

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Representation-Based Data Quality Audits for Audio

2025-09-30 · Alvaro Gonzalez-Jimenez, Fabian Gröger, Linda Wermelinger, Andrin Bürli 외 arxiv

Data quality issues such as off-topic samples, near duplicates, and label errors often limit the performance of audio-based systems. This paper addresses these issues by adapting SelfClean, a representation-to-rank data …

Learning to Have an Ear for Face Super-Resolution

2019-09-27 · CVPR 2020 6 · Givi Meishvili, Simon Jenni, Paolo Favaro

We propose a novel method to use both audio and a low-resolution image to perform extreme face super-resolution (a 16x increase of the input size). When the resolution of the input image is very low (e.g., 8x8 pixels), t…

Audio Super-ResolutionFace ReconstructionImage ReconstructionSuper-Resolution

Distinguishing Homophenes Using Multi-Head Visual-Audio Memory for Lip Reading

2022-04-04 · The AAAI Conference on Artificial Intelligence (AAAI) 2022 3 · Minsu Kim, Jeong Hun Yeo, Yong Man Ro

Recognizing speech from silent lip movement, which is called lip reading, is a challenging task due to 1) the inherent information insufficiency of lip movement to fully represent the speech, and 2) the existence of homo…

LipreadingLip Reading

Exploring Longitudinal Cough, Breath, and Voice Data for COVID-19 Progression Prediction via Sequential Deep Learning: Model Development and Validation

2022-01-04 · Ting Dang, Jing Han, Tong Xia, Dimitris Spathis 외

Recent work has shown the potential of using audio data (eg, cough, breathing, and voice) in the screening for COVID-19. However, these approaches only focus on one-off detection and detect the infection given the curren…

Specificity

Assessing Fiscal Policy Effectiveness on Household Savings in Hungary, Slovenia, and the Czech Republic during the COVID-19 Crisis: A Markov Switching VAR Approach

2025-03-19 · Tuhin G M Al Mamun

The COVID-19 pandemic significantly disrupted household consumption, savings, and income across Europe, particularly affecting countries like Hungary, Slovenia, and the Czech Republic. This study investigates the effecti…