Towards Context-Aware Neural Performance-Score Synchronisation
Music can be represented in multiple forms, such as in the audio form as a recording of a performance, in the symbolic form as a computer readable score, or in the image form as a scan of the sheet music. Music synchronisation provides a way to navigate among multiple representations of music in a unified manner by generating an accurate mapping between them, lending itself applicable to a myriad of domains like music education, performance analysis, automatic accompaniment and music editing. Traditional synchronisation methods compute alignment using knowledge-driven and stochastic approaches, typically employing handcrafted features. These methods are often unable to generalise well to different instruments, acoustic environments and recording conditions, and normally assume complete structural agreement between the performances and the scores. This PhD furthers the development of performance-score synchronisation research by proposing data-driven, context-aware alignment approaches, on three fronts: Firstly, I replace the handcrafted features by employing a metric learning based approach that is adaptable to different acoustic settings and performs well in data-scarce conditions. Secondly, I address the handling of structural differences between the performances and scores, which is a common limitation of standard alignment methods. Finally, I eschew the reliance on both feature engineering and dynamic programming, and propose a completely data-driven synchronisation method that computes alignments using a neural framework, whilst also being robust to structural differences between the performances and scores.
Code (0)
등록된 구현이 없습니다.
Tasks
Feature EngineeringFormMetric LearningNavigateSimilar Papers 제목 키워드 기반
VocaLiST: An Audio-Visual Synchronisation Model for Lips and Voices
In this paper, we address the problem of lip-voice synchronisation in videos containing human face and voice. Our approach is based on determining if the lips motion and the voice in a video are synchronised or not, depe…
Audio-Visual SynchronizationMusic Source SeparationSynchronisation-Oriented Design Approach for Adaptive Control
This study presents a synchronisation-oriented perspective towards adaptive control which views model-referenced adaptation as synchronisation between actual and virtual dynamic systems. In the context of adaptation, mod…
Sparse in Space and Time: Audio-visual Synchronisation with Trainable Selectors
The objective of this paper is audio-visual synchronisation of general videos 'in the wild'. For such videos, the events that may be harnessed for synchronisation cues may be spatially small and may occur only infrequent…
Audio-Visual SynchronizationCharacter-aware audio-visual subtitling in context
This paper presents an improved framework for character-aware audio-visual subtitling in TV shows. Our approach integrates speech recognition, speaker diarisation, and character recognition, utilising both audio and visu…
Language ModellingLarge Language Modelspeech-recognitionSpeech RecognitionNovel Approach To Synchronisation Of Wearable IMUs Based On Magnetometers
Synchronisation of wireless inertial measurement units in human movement analysis is often achieved using event-based synchronisation techniques. However, these techniques lack precise event generation and accuracy. An i…
Motion Estimation