Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 9 of 9 for “"Speaker diarization"”.

  1. Unsupervised methods for speaker diarization

    Given a stream of unlabeled audio data, speaker diarization is the process of determining "who spoke when." We propose a novel approach to solving this problem by taking advantage of the effectiveness of factor analysis as a front-end for extracting speaker-specific features and exploiting the …

    mit Repository record for Unsupervised methods for speaker diarization (opens in a new tab)

  2. Robust Speaker Diarization for Single Channel Recorded Meetings

    This thesis describes research into speaker diarization for recorded meetings. It explores the algorithms and the implementation of an off-line speaker segmentation and clustering system for meetings that have been recorded using one microphone. Speaker diarization is defined as a process of …

    whiterose Repository record for Robust Speaker Diarization for Single Channel Recorded Meetings (opens in a new tab)

  3. Exploiting spatial and spectral information for audio source separation and speaker diarization

    … minimizing a non-convex $l_0$-norm function. For speaker diarization where the task is to determine ``who spoke when" in real meetings, a Watson mixture model is optimized using an Expectation-Maximization algorithm in order to detect the probabilistic descriptions, best representing the …

    trento Repository record for Exploiting spatial and spectral information for audio source separation and speaker diarization (opens in a new tab)

  4. Automatic speaker verification and diarization on VoxCeleb data collection

    Automatic speaker verification (ASV) is increasingly getting more attention in speech research field in recent years. Because of the importance of cyber-security and personal property security, ASV can be used in many fields in the future in addition to fingerprint and face information. In ASV …

    gatech Repository record for Automatic speaker verification and diarization on VoxCeleb data collection (opens in a new tab)

  5. Bayesian nonparametric learning with semi-Markovian dynamics

    … on synthetic data as well as experiments on a speaker diarization problem and an example of learning the patterns in Morse code.

    mit Repository record for Bayesian nonparametric learning with semi-Markovian dynamics (opens in a new tab)

  6. Reducing Costs in Human Assisted Speech Transcription

    … improved transcription tool UI and systems for speaker diarization and text correction.</p> <p>This thesis evaluates the effectiveness of these improvements on the human assisted transcription process employed by the Digital Democracy initiative. To facilitate this evaluation, a pipeline for …

    calpoly Repository record for Reducing Costs in Human Assisted Speech Transcription (opens in a new tab)

  7. Audio Segmenting and Natural Language Processing in Oral History Archiving

    … using techniques including silence detection and speaker diarization, with the goal of creating a more flexible way to explore interviews in a digital oral history archive. Second, this thesis uses named entity recognition to experiment with metadata extraction for an archive. Next, this thesis …

    mit Repository record for Audio Segmenting and Natural Language Processing in Oral History Archiving (opens in a new tab)

  8. Modeling Temporal and Spatial Data Dependence with Bayesian Nonparametrics

    … model is employed on image segmentation and speaker diarization, yielding generally homogeneous segments with sharp boundaries.</p> <p>In addition, we also consider a multi-task learning with each task associated with spatial dependence. For the specific application of co-segmentation with …

    duke Repository record for Modeling Temporal and Spatial Data Dependence with Bayesian Nonparametrics (opens in a new tab)

  9. Breaking down barriers: advancing interdisciplinary speech applications in early children’s development

    Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2024-09-16 without embargo terms

    uiuc Repository record for Breaking down barriers: advancing interdisciplinary speech applications in early children’s development (opens in a new tab)