Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 9 of 9 for “"Speaker diarization"”.
-
Unsupervised methods for speaker diarization
Given a stream of unlabeled audio data, speaker diarization is the process of determining "who spoke when." We propose a novel approach to solving this problem by taking advantage of the effectiveness of factor analysis as a front-end for extracting speaker-specific features and exploiting the …
-
Robust Speaker Diarization for Single Channel Recorded Meetings
This thesis describes research into speaker diarization for recorded meetings. It explores the algorithms and the implementation of an off-line speaker segmentation and clustering system for meetings that have been recorded using one microphone. Speaker diarization is defined as a process of …
-
Exploiting spatial and spectral information for audio source separation and speaker diarization
… minimizing a non-convex $l_0$-norm function. For speaker diarization where the task is to determine ``who spoke when" in real meetings, a Watson mixture model is optimized using an Expectation-Maximization algorithm in order to detect the probabilistic descriptions, best representing the …
-
Automatic speaker verification and diarization on VoxCeleb data collection
Automatic speaker verification (ASV) is increasingly getting more attention in speech research field in recent years. Because of the importance of cyber-security and personal property security, ASV can be used in many fields in the future in addition to fingerprint and face information. In ASV …
-
Bayesian nonparametric learning with semi-Markovian dynamics
… on synthetic data as well as experiments on a speaker diarization problem and an example of learning the patterns in Morse code.
-
Reducing Costs in Human Assisted Speech Transcription
… improved transcription tool UI and systems for speaker diarization and text correction.</p> <p>This thesis evaluates the effectiveness of these improvements on the human assisted transcription process employed by the Digital Democracy initiative. To facilitate this evaluation, a pipeline for …
-
Audio Segmenting and Natural Language Processing in Oral History Archiving
… using techniques including silence detection and speaker diarization, with the goal of creating a more flexible way to explore interviews in a digital oral history archive. Second, this thesis uses named entity recognition to experiment with metadata extraction for an archive. Next, this thesis …
-
Modeling Temporal and Spatial Data Dependence with Bayesian Nonparametrics
… model is employed on image segmentation and speaker diarization, yielding generally homogeneous segments with sharp boundaries.</p> <p>In addition, we also consider a multi-task learning with each task associated with spatial dependence. For the specific application of co-segmentation with …
-
Breaking down barriers: advancing interdisciplinary speech applications in early children’s development
Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2024-09-16 without embargo terms