Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 4 of 4 for “"Video Captioning"”.

  1. Connecting vision and language via image retrieval and captioning

    … understanding vision and language, e.g. image/video captioning, multi-modal image retrieval, visual question answering. In this thesis, we consider several problems in cross-modal learning from vision and language. First, the problem of composed query image retrieval is studied. In this …

    manitoba Repository record for Connecting vision and language via image retrieval and captioning (opens in a new tab)

  2. Understanding language through visual imagination

    … language-vision model that can generate videos for sentences and a comprehensive approach for using this capability to solve natural language inference, video captioning and video completion without task-specific training. The only training required is for acquiring a lexicon from …

    mit Repository record for Understanding language through visual imagination (opens in a new tab)

  3. Unifying cross-modal concepts in vision and language

    … We introduce the task of knowledge-aware video captioning for news videos, where models must generate descriptions of videos that leverage interconnected background knowledge pertaining to concepts involved in the videos. We build models that learn to associate patterns of concepts found …

    uiuc Repository record for Unifying cross-modal concepts in vision and language (opens in a new tab)

  4. Adopting the two-branch network to video-text tasks

    This Thesis was approved for publication on 2018-04-23 at 16:33.

    uiuc Repository record for Adopting the two-branch network to video-text tasks (opens in a new tab)