Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 3 of 3 for “"Speaker Diarisation"”.

  1. Attention-Based Encoder-Decoder Models for Speech Processing

    … - speech recognition, confidence estimation and speaker diarisation. Speech recognition technology is widely used in voice assistants and dictation systems. It converts speech signals into text. Traditionally, Hidden Markov Models (HMMs), as a generative sequence-to-sequence model, are widely …

    cambridge Repository record for Attention-Based Encoder-Decoder Models for Speech Processing (opens in a new tab)

  2. Building Speech-Driven Assessment Tools for Afrikaans and isiXhosa Children

    … We integrate Mamela into a pipeline with speaker diarisation, ASR, machine translation and a large language model (LLM) to predict the MAIN scores automatically. Fourth, we integrate the developed MAIN pipeline into a user-friendly mobile application for speech therapists and educators. …

    stellenbosch Repository record for Building Speech-Driven Assessment Tools for Afrikaans and isiXhosa Children (opens in a new tab)

  3. Speech-Based Emotion Modelling and Mental Disorder Detection

    … system is developed which integrates AER with speaker diarisation and speech recognition in a jointly-trained system. Compared to separately optimised cascaded systems, the proposed system achieves not only improved efficiency but also reduced recognition errors for emotional speech. In …

    cambridge Repository record for Speech-Based Emotion Modelling and Mental Disorder Detection (opens in a new tab)