Saving the Sonorine: Photovisual Audio Recovery Using Image Processing and Computer Vision Techniques
This paper presents a novel technique to recover audio from sonorines, an early 20th century form of analogue sound storage. Our method uses high resolution photographs of sonorines under different lighting conditions to observe the change in reflection behavior of the physical surface features and create a three-dimensional height map of the surface. Sound can then be extracted using height information within the surface's grooves, mimicking a physical stylus on a phonograph. Unlike traditional playback methods, our method has the advantage of being contactless: the medium will not incur damage and wear from being played repeatedly. We compare the results of our technique to a previously successful contactless method using flatbed scans of the sonorines, and conclude with future research that can be applied to this photovisual approach to audio recovery.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
Representation-Based Data Quality Audits for Audio
Data quality issues such as off-topic samples, near duplicates, and label errors often limit the performance of audio-based systems. This paper addresses these issues by adapting SelfClean, a representation-to-rank data …
Learning to Have an Ear for Face Super-Resolution
We propose a novel method to use both audio and a low-resolution image to perform extreme face super-resolution (a 16x increase of the input size). When the resolution of the input image is very low (e.g., 8x8 pixels), t…
Audio Super-ResolutionFace ReconstructionImage ReconstructionSuper-ResolutionDistinguishing Homophenes Using Multi-Head Visual-Audio Memory for Lip Reading
Recognizing speech from silent lip movement, which is called lip reading, is a challenging task due to 1) the inherent information insufficiency of lip movement to fully represent the speech, and 2) the existence of homo…
LipreadingLip ReadingExploring Longitudinal Cough, Breath, and Voice Data for COVID-19 Progression Prediction via Sequential Deep Learning: Model Development and Validation
Recent work has shown the potential of using audio data (eg, cough, breathing, and voice) in the screening for COVID-19. However, these approaches only focus on one-off detection and detect the infection given the curren…
SpecificityAssessing Fiscal Policy Effectiveness on Household Savings in Hungary, Slovenia, and the Czech Republic during the COVID-19 Crisis: A Markov Switching VAR Approach
The COVID-19 pandemic significantly disrupted household consumption, savings, and income across Europe, particularly affecting countries like Hungary, Slovenia, and the Czech Republic. This study investigates the effecti…