Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 4 of 4 for “"multi-modal models"”.

  1. Large scale video action understanding

    … dataset called Moments, and train existing/novel models for action recognition. To aid automation of video collection and annotation selection, I trained Convolutional Neural Network models to estimate the likelihood of a desired action appearing in video clips. Selecting clips, which are highly …

    mit Repository record for Large scale video action understanding (opens in a new tab)

  2. Transparent Analysis of Multi-Modal Embeddings

    Vector Space Models of Distributional Semantics – or Embeddings – serve as useful statistical models of word meanings, which can be applied as proxies to learn about human concepts. One of their main benefits is that not only textual, but a wide range of data types can be mapped to a space, where …

    cambridge Repository record for Transparent Analysis of Multi-Modal Embeddings (opens in a new tab)

  3. Data augmentation and data efficiency for low-resource language processing

    Submission published under a 24 month embargo labeled 'U of I Access', the embargo will last until 2027-05-01

    uiuc Repository record for Data augmentation and data efficiency for low-resource language processing (opens in a new tab)

  4. Learning video representations with limited supervision

    Submission original under an indefinite embargo labeled 'Open Access'. The submission was exported from vireo on 2024-03-01 without embargo terms

    uiuc Repository record for Learning video representations with limited supervision (opens in a new tab)