Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 23 for “"audio processing"”.

  1. Speech Foundation Models for Audio Processing

    … shown strong performance across a variety of audio processing tasks, including automatic speech recognition (ASR) and speech translation. Unlike traditional systems that require task-specific architectures and extensive supervision, these models offer a unified and flexible framework that …

    cambridge Repository record for Speech Foundation Models for Audio Processing (opens in a new tab)

  2. Interconnectable blocks for music and audio processing

    This thesis describes the technical implementation of a set of interconnectable blocks designed to be used by children to explore the possibilities of digital sound manipulation. In contrast to similar modular systems, this project places an emphasis on achieving the minimum possible cost. Every …

    mit Repository record for Interconnectable blocks for music and audio processing (opens in a new tab)

  3. Neuromorphic audio processing through real-time embedded spiking neural networks.

    In this work novel speech recognition and audio processing systems based on a spiking artificial cochlea and neural networks are proposed and implemented. First, the biological behavior of the animal’s auditory system is analyzed and studied, along with the classical mechanisms of audio signal …

    sevilla Repository record for Neuromorphic audio processing through real-time embedded spiking neural networks. (opens in a new tab)

  4. A high accuracy nonlinear model of the human cochlea

    … auditory system is of fundamental importance to audio processing systems and hearing research. Generally, models intended for real-time audio processing are time-efficient but tend to lack grounding in physical reality, while models designed for hearing research may closely fit experimental data …

    uiuc Repository record for A high accuracy nonlinear model of the human cochlea (opens in a new tab)

  5. Neuromorphic auditory computing: towards a digital, event-based implementation of the hearing sense for robotics

    … advance on the development of the neuromorphic audio processing systems in robots through the implementation of an open-source neuromorphic cochlea, event-based models of primary auditory nuclei, and their potential use for real-time robotics applications. First, the main gaps when working with …

    sevilla Repository record for Neuromorphic auditory computing: towards a digital, event-based implementation of the hearing sense for robotics (opens in a new tab)

  6. Multi-modal mixing : gestural control for musical mixing systems

    … give them almost limitless control over audio processing. Devices such as Yamaha's new digital mixers and Digidesign's Pro-Tools computerized editing workstations allow users in a small studio to accomplish tasks which would have required racks full of gear only seven years ago in a …

    mit Repository record for Multi-modal mixing : gestural control for musical mixing systems (opens in a new tab)

  7. Show design and control system for live theater

    … of nine "Operabots"; a novel surround sound and audio processing architecture; a massive string instrument called "The Chandelier"; three 14' tall robotic stage fixtures with displays called "The Walls"; and more. Each component has its own unique challenges, but all share the need for a …

    mit Repository record for Show design and control system for live theater (opens in a new tab)

  8. Practica: A Music Education Application for Learning Jazz Improvisation

    … in the form of color-coded notes. We develop an audio processing and transcription pipeline to generate sheet music for solo recordings. We examine how to present the subjective teaching and evaluation of improvisation in a programmatic manner. Two user studies suggest that Practica successfully …

    mit Repository record for Practica: A Music Education Application for Learning Jazz Improvisation (opens in a new tab)

  9. JamNSync: A User-Friendly, Latency-Agnostic Virtual Rehearsal Platform for Music Ensembles

    … understanding of playback systems and common audio devices. To account for non-deterministic audio latency in the browser, JamNSync provides a user-friendly audio alignment tool that efficiently automates audio processing and produces a group mix ready for immediate feedback. Three rounds of …

    mit Repository record for JamNSync: A User-Friendly, Latency-Agnostic Virtual Rehearsal Platform for Music Ensembles (opens in a new tab)

  10. Spoke : a framework for building speech-enabled websites

    … provides a client-side framework enabling some audio processing in the browser and streaming of the user's audio to a server for recording and backend processing with the aforementioned integrated speech technologies. Spoke's client-side and server-side modules can be used in conjunction to …

    mit Repository record for Spoke : a framework for building speech-enabled websites (opens in a new tab)

  11. Harnessing open sound control for networked music in VST systems

    Professional audio equipment is migrating towards general purpose computers running professional audio software systems such as Virtual Studio Technology [VST) and VST plugins that have been adopted as an informal standard for audio and MIDI plugins. The proliferation of computer networks has …

    cape-town Repository record for Harnessing open sound control for networked music in VST systems (opens in a new tab)

  12. The benefits of acoustic perceptual information for speech processing systems

    … framework has dominated many speech processing systems, such as ASR and AED targeting human speech activities. These systems have little consideration for the science behind speech and treat the task as a simple statistical classification. The framework also assumes each feature …

    uiuc Repository record for The benefits of acoustic perceptual information for speech processing systems (opens in a new tab)

  13. Automatic code generation: from process algebraic architectural descriptions to multithreaded java programs

    … by means of a running example based on an audio processing system. First, we develop an architecture-driven technique for thread coordination management, which is completely automated through a suitable package. Second, we address the translation of the algebraically-specified behavior of …

    bologna Repository record for Automatic code generation: from process algebraic architectural descriptions to multithreaded java programs (opens in a new tab)

  14. Vector-thread architecture and implementation

    … benchmarks, including example kernels for image processing, audio processing, text and data processing, cryptography, network processing, and wireless communication.

    mit Repository record for Vector-thread architecture and implementation (opens in a new tab)

  15. Unsupervised learning of cross-modal mappings between speech and text

    … in computer vision, natural language processing, and speech and audio processing. Current deep learning models, however, rely on signicant amounts of supervision for training to achieve exceptional performance. For example, commercial speech recognition systems are usually trained on …

    mit Repository record for Unsupervised learning of cross-modal mappings between speech and text (opens in a new tab)

  16. Analog Signal Processing Elements for Energy-Constrained Platforms

    Energy constrained processing poses a number of challenges that have resulted in tremendous innovations over the past decade. Shrinking supply voltages and limited clock speeds have placed an emphasis on processing efficiency over the raw throughput of a processor. One of the approaches to increase …

    wvu Repository record for Analog Signal Processing Elements for Energy-Constrained Platforms (opens in a new tab)

  17. Musical interfaces : design and construction of physical manipulatives for musical composition

    … are explored using non-functional form models. Audio processing is performed on a peripheral computer running an audio program written specifically for each system. A "Wizard of Oz" approach was used to study user interactions with each design. Music Blocks are designed to be physical …

    mit Repository record for Musical interfaces : design and construction of physical manipulatives for musical composition (opens in a new tab)

  18. The Importance of Data in RF Machine Learning

    … the last decade, spurred by results in image and audio processing. Machine Learning (ML), and Deep Learning (DL) specifically, are driven by access to relevant data during the training phase of the application due to the learned feature sets that are derived from vast amounts of similar data. …

    vt Repository record for The Importance of Data in RF Machine Learning (opens in a new tab)

  19. End-to-end non-negative auto-encoders: a deep neural alternative to non-negative audio modeling

    … one of the most popular approaches to modeling audio signals. NMF allows us to factorize the magnitude spectrogram to learn representative spectral bases that can be used for a wide range of applications. With the recent advances in deep learning, neural networks (NNs) have surpassed NMF in …

    uiuc Repository record for End-to-end non-negative auto-encoders: a deep neural alternative to non-negative audio modeling (opens in a new tab)

  20. A Study on Unintentional and Intentional Sources of Variability in Nanometer Scale Digital Circuits

    … and synthesis, especially those from image and audio processing domains, are error tolerant, since they pertain to the inherently limited human perception. A new design paradigm called approximate computing, leverages this error-tolerance to implement arithmetic operations through approximate …

    umn Repository record for A Study on Unintentional and Intentional Sources of Variability in Nanometer Scale Digital Circuits (opens in a new tab)

Page 1 of 2