Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 4 of 4 for “"Video semantic segmentation"”.

  1. Towards Comprehensive Visual Understanding via Deep Neural Networks

    … with i) diverse scenes, as well as ii) diverse semantic structures within those scenes. Existing work typically requires extensive annotation for different scenes (domains) and separates the understanding of semantic targets into distinct tasks, designing meticulous networks and corresponding …

    uts Repository record for Towards Comprehensive Visual Understanding via Deep Neural Networks (opens in a new tab)

  2. A Unified Multiscale Encoder-Decoder Transformer for Video Segmentation

    … multiscale encoder-decoder transformer for dense video estimation, with a focus on segmentation. We investigate this direction by exploring unified multiscale processing throughout the processing pipeline of feature encoding, context encoding and object decoding in an encoder-decoder model. …

    york Repository record for A Unified Multiscale Encoder-Decoder Transformer for Video Segmentation (opens in a new tab)

  3. Geometry and Uncertainty in Deep Learning for Computer Vision

    … information than recognition, from images and video. In general, applying these deep learning models from recognition to other problems in computer vision is significantly more challenging. This thesis presents end-to-end deep learning architectures for a number of core computer vision …

    cambridge Repository record for Geometry and Uncertainty in Deep Learning for Computer Vision (opens in a new tab)

  4. Pixel-level video understanding with efficient deep models

    The ability to understand videos at the level of pixels plays a key role in a wide range of computer vision applications. For example, a robot or autonomous vehicle relies on classifying each pixel in the video stream into semantic categories to holistically understand the surrounding environment, …

    bu Repository record for Pixel-level video understanding with efficient deep models (opens in a new tab)