Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 18 of 18 for “"MFCC"”.

  1. DESIGN OF A KEYWORD SPOTTING SYSTEM USING MODIFIED CROSS-CORRELATION IN THE TIME AND THE MFCC DOMAIN

    … spotting was investigated in both the time and MFCC domain. In the time domain the global keyword was cross-correlated with a pitch-normalized utterance. A zero lag ratio (the ratio of the power around the zero lag obtained from a cross correlation to the power in the rest of the signal is …

    temple Repository record for DESIGN OF A KEYWORD SPOTTING SYSTEM USING MODIFIED CROSS-CORRELATION IN THE TIME AND THE MFCC DOMAIN (opens in a new tab)

  2. Machine learning methods for individual acoustic recognition in a species of field cricket

    … Mel-frequency cepstral coefficients (MFCC), raw acoustic samples as well as two temporal features were extracted from each chirp in the cricket recordings and used as inputs to train the machine learning models. The raw acoustic samples were only used in the deep neural network (DNN) …

    cape-town Repository record for Machine learning methods for individual acoustic recognition in a species of field cricket (opens in a new tab)

  3. Metinden bağımsız konuşmacı tanıma sistemlerinin incelenmesi ve gerçekleştirilmesi

    … her bir konuşmacıya ait konuşma sinyalleri için MFCC, ∆MFCC, LPCC kepstral katsayıları çıkarılarak öznitelik vektörler kümesi oluşturulmuştur. Bu vektörler LBG ve Beklentinin Maksimumlaştırılması (BM) algoritmalarıyla modellenmiştir. Eğitim ve test aşamalarında öznitelik katsayılarının sayısı, …

    ankara Repository record for Metinden bağımsız konuşmacı tanıma sistemlerinin incelenmesi ve gerçekleştirilmesi (opens in a new tab)

  4. Robot s autonomním audio-vizuálním řízením

    … pro zpracování obrazu a algoritmy pro výpočet MFCC a DTW pro rozpoznávání hlasových pokynů.

    brno-tech Repository record for Robot s autonomním audio-vizuálním řízením (opens in a new tab)

  5. Nouvelle technique d'analyse acoustique pour la reconnaissance vocale dans les transmissions GSM bruitées

    … pour la reconnaissance de la parole appelée MFCC-XAFE. (Mel-Frequecy Cepstral Coefficients extended Audio Front-End). L'approche présentée dans cette thèse est une alternative de la méthode actuelle ETSI DSR-XAFE. Dans notre approche nous avons ajouté, aux coefficients MFCC, les formants dans …

    moncton Repository record for Nouvelle technique d'analyse acoustique pour la reconnaissance vocale dans les transmissions GSM bruitées (opens in a new tab)

  6. On-device mobile speech recognition

    … and the Mel Frequency Cepstral Coefficients (MFCC) as input features. A speaker dependent approach is presented using the Centre for spoken Language and Understanding (CSLU) database. The results show a very distinct behaviour from conventional speech recognition approaches because the LPC …

    nott-trent Repository record for On-device mobile speech recognition (opens in a new tab)

  7. Productivity Measurement of Call Centre Agents using a Multimodal Classification Approach

    … to use the Mel Frequency Cepstral Coefficient (MFCC) upgraded with Low-Level Descriptors (LLD) to improve classification accuracy. The data modelling architectures for speech and text are based on CNNs, BiLSTMs, and the attention layer. The multimodal approach follows the generated models to …

    sevilla Repository record for Productivity Measurement of Call Centre Agents using a Multimodal Classification Approach (opens in a new tab)

  8. AudioCNN: Audio Event Classification With Deep Learning Based Multi-Channel Fusion Networks

    … (CG), and Mel Frequency Cepstral Coefficient (MFCC), for useful environmental sound classification. We propose the AudioCNN model based on a fusion network consisting of multiple Convolutional Neural Networks (CNN) with aggregation methods for various spectral image spectrogram features and …

    umkc Repository record for AudioCNN: Audio Event Classification With Deep Learning Based Multi-Channel Fusion Networks (opens in a new tab)

  9. Closed-loop auditory-based representation for robust speech recognition

    … noisy speech. Compared with the standard MFCC extraction algorithm, the proposed closed-loop form of feature extraction algorithm provides 9.7%, 9.1% and 11.4% absolution word error rate reduction on average for three kinds of filter banks respectively.

    mit Repository record for Closed-loop auditory-based representation for robust speech recognition (opens in a new tab)

  10. Automatic Phoneme Recognition with Segmental Hidden Markov Models

    … of the Mel-Frequency Cepstral Coefficients (MFCC) and the corresponding Delta and Delta Log Power coefficients is described. Second, we describe the operation of the Baum-Welch re-estimation procedure for the training of the phoneme HMM models, including the K-Means and the …

    vt Repository record for Automatic Phoneme Recognition with Segmental Hidden Markov Models (opens in a new tab)

  11. A Novel Approach to Indoor Environment Assessment: Artificial Intelligence of Things (AIoT) Framework for Improving Occupant Comfort and Health in Educational Facilities

    … IEQ data, non-intrusive occupant feedback (MFCC features from audio recordings, video/thermal features extracted by Vision Transformer (ViT)), and self-reported comfort and health levels, placing a focus on occupant-centric and data-driven decision-making for intelligent educational …

    vt Repository record for A Novel Approach to Indoor Environment Assessment: Artificial Intelligence of Things (AIoT) Framework for Improving Occupant Comfort and Health in Educational Facilities (opens in a new tab)

  12. Percussion Based Detection Method for Localization of Pipe Inspection Gauge using Advanced Machine Learning Classification and Clustering Techniques.

    … localization techniques. Results showed that MFCC feature extraction and Convolutional Neural Network + Long-Short Term Memory Network and Gaussian Mixed Model clustering techniques were able to best classify at high accuracy missing pipe inspection gauges in an experimental pipeline system.

    houston Repository record for Percussion Based Detection Method for Localization of Pipe Inspection Gauge using Advanced Machine Learning Classification and Clustering Techniques. (opens in a new tab)

  13. Art Therapy Program for Children of Battered Women

    … The program will utilize the services of an MFCC Intern who is also trained in Art Therapy. The immediate objectives of the program are to relieve children’s confusion and situational stress, to help them establish healthy coping strategies and ways to remain safe in a volatile family …

    dominican Repository record for Art Therapy Program for Children of Battered Women (opens in a new tab)

  14. Αναγνώριση ομιλητή και ομιλίας με χρήση κυματιδίων

    … οι παράμετροι cepstral με βάση την κλίμακα mel (MFCC). Επιπλέον, στη διατριβή αναλύονται οι ιδιότητες των σημαντικότερων συναρτήσεων κυματιδίων, επιλέγεται η βέλτιστη για την αναπαράσταση του σήματος ομιλίας και πιστοποιείται στην πράξη αυτή η επιλογή. Τέλος, οι δύο πρώτες από τις προαναφερόμενες …

    patras-thes Repository record for Αναγνώριση ομιλητή και ομιλίας με χρήση κυματιδίων (opens in a new tab)

  15. Removing redundancy in speech by modeling forward masking

    … analysis methods, such as LPC, STFT, and MFCC, have been used in speech recognition, the performance of ASR has reached a plateau and the speech decoding problem remains unresolved. Recently, the Human Speech Recognition group (HSR) of the University of Illinois conducted research aimed at …

    uiuc Repository record for Removing redundancy in speech by modeling forward masking (opens in a new tab)

  16. Alignement du chant par rapport à une référence audio en temps réel

    … sont les Mel-frquency Cepstrum Coefficients (MFCC), les Warped Discrete Cosine Transform Coefficients (WDCTC) et les coefficients de l'analyse Perceptual Linear Prediction (PLP). Les résultats obtenus indiquent une meilleure performance pour l'analyse PLP. L'utilisation d'une fonction de …

    sherbrooke Repository record for Alignement du chant par rapport à une référence audio en temps réel (opens in a new tab)

  17. Geometric Multimedia Time Series

    … performs the best, and it also enables us to use MFCC features for this task, which was previously thought not to be possible due to significant timbral differences that can exist between versions. When combined with traditional pitch-based features using similarity metric fusion, we obtain state …

    duke Repository record for Geometric Multimedia Time Series (opens in a new tab)