Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 3 of 3 for “"speech foundation models"”.

  1. Speech Foundation Models for Audio Processing

    Foundation models are trained on massive amounts of data and can be adapted to support a wide range of downstream tasks. Their emergence has significantly shaped the development of natural language and speech technologies. Foundation speech models such as Whisper have shown strong performance …

    cambridge Repository record for Speech Foundation Models for Audio Processing (opens in a new tab)

  2. Efficient Knowledge Transfer and Adaptation for Speech and Beyond

    … transfer and adaptation in the realm of speech processing. It is structured to address the limitations of transfer learning in dynamically evolving audio and speech processing contexts, particularly through novel approaches for class-incremental learning, parameter-efficient adaptation, …

    trento Repository record for Efficient Knowledge Transfer and Adaptation for Speech and Beyond (opens in a new tab)

  3. Neural Time Alignment for End-to-End Automatic Speech Recognition Systems

    Automatic Speech Recognition (ASR) is an important component for machines to interact with humans. With the development of deep learning, ASR systems can obtain very low error rates when there is a lot of training data available. Apart from the recognised text, it is also important to determine …

    cambridge Repository record for Neural Time Alignment for End-to-End Automatic Speech Recognition Systems (opens in a new tab)