Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 20 of 30 for “"Visual Understanding"”.

  1. Toward ontological visual understanding

    … remains when it comes to automatic multimedia understanding. One difference between the human cognition system and state-of-the-art machine vision algorithms is that the former possesses and uses high-level semantic knowledge, or ontology. In this thesis, we present our work on image-level …

    uiuc Repository record for Toward ontological visual understanding (opens in a new tab)

  2. Toward visual understanding of everyday object

    … using large scale data. However, for any visual system to interact with objects, it needs to understand much more than simply recognizing where the objects are. The goal of my research is to explore and solve object understanding tasks for interaction - finding an object's pose in 3D, …

    mit Repository record for Toward visual understanding of everyday object (opens in a new tab)

  3. Label-Efficient Visual Understanding with Consistency Constraints

    … networks are proficient at solving various visual recognition and understanding tasks, as long as a sufficiently large labeled dataset is available during the training time. However, the progress of these visual tasks is limited by the number of manual annotations. On the other hand, it is …

    vt Repository record for Label-Efficient Visual Understanding with Consistency Constraints (opens in a new tab)

  4. Towards Comprehensive Visual Understanding via Deep Neural Networks

    … (DNNs) have made significant advancements in visual scene understanding, demonstrating great potential for applications in downstream tasks such as autonomous driving, robotic navigation, and human-computer interaction. Despite these successes, generalization ability remains a major obstacle …

    uts Repository record for Towards Comprehensive Visual Understanding via Deep Neural Networks (opens in a new tab)

  5. LEARNING STRUCTURED ALIGNMENT: FROM VISUAL UNDERSTANDING TO ROBOT CONTROL

    The rapid growth of visual data and the increasing demand for intelligent robotic systemshave created a pressing need for methods that can establish meaningful correspondences and relationships across diverse visual modalities and robotic tasks. This dissertation addresses the fundamental challenge …

    maryland Repository record for LEARNING STRUCTURED ALIGNMENT: FROM VISUAL UNDERSTANDING TO ROBOT CONTROL (opens in a new tab)

  6. Visual Understanding of Hybrid Circuits via Procedural Models: An Approach to Automatic Inspection of Hybrid Circuits

    Made available in DSpace on 2014-12-12T20:54:12Z (GMT). No. of bitstreams: 1 7524408.pdf: 5803201 bytes, checksum: 19eeae58595e41f38a37da6d9b2483b2 (MD5) Previous issue date: 1975

    uiuc Repository record for Visual Understanding of Hybrid Circuits via Procedural Models: An Approach to Automatic Inspection of Hybrid Circuits (opens in a new tab)

  7. Assessing the development of high school chemistry students' conceptual and visual understanding of dimensional analysis via supplemental use of a proprietary interactive software program

    … “Conversionoes,” on students' conceptual and visual understanding of dimensional analysis. The participants in the study were high school general chemistry students enrolled in two public high schools with different demographics (School A and School B) in the Chattanooga, Tennessee, …

    lsu-thes Repository record for Assessing the development of high school chemistry students' conceptual and visual understanding of dimensional analysis via supplemental use of a proprietary interactive software program (opens in a new tab)

  8. ARGILE: A Conceptual Framework Combining Augmented Reality with Agile Philosophy for the UK Construction Industry

    … agile project management and augmented reality visualisation technology. It enables an understanding on how the use of the proposed, designed, and developed ARGILE framework (integration of Agile and Augmented Reality) enhances the collaboration, communication, decision-making, and visual

    liverpool-jm Repository record for ARGILE: A Conceptual Framework Combining Augmented Reality with Agile Philosophy for the UK Construction Industry (opens in a new tab)

  9. Learning from minimally labeled data with accelerated convolutional neural networks

    … to be satisfied by the mapping function for visual understanding. First, the function should produce good representations of the visual world, which will be able to recognize images independently of pose, scale and illumination. Furthermore, the designed artificial vision system has to learn …

    purdue-thes Repository record for Learning from minimally labeled data with accelerated convolutional neural networks (opens in a new tab)

  10. Predictive vision

    … an extensive set of rules that enable richer visual understanding. While large annotated datasets fuel progress in object recognition, the knowledge required for event understanding is vast and potentially ambiguous. To tackle this challenge, we develop predictive vision algorithms that …

    mit Repository record for Predictive vision (opens in a new tab)

  11. Deep Learning Multimodal Extraction of Reaction Data

    … allow powerful large language models to leverage visual understanding. Our findings indicate that VLMs still require additional work in order to meet the performance of our bespoke models.

    mit Repository record for Deep Learning Multimodal Extraction of Reaction Data (opens in a new tab)

  12. Unsupervised Compositional Image Decompositionwith Diffusion Models

    Our visual understanding of the world is factorized and compositional. With just a single observation, we can ascertain both global and local attributes in a scene, such as lighting, weather, and underlying objects. These attributes are highly compositional and can be combined in various ways to …

    mit Repository record for Unsupervised Compositional Image Decompositionwith Diffusion Models (opens in a new tab)

  13. Modeling Dynamic Objects in Scenes with Generative Particle Systems

    … serves as a critical signal for recognition and understanding. Inspired by this ability, we propose a generative model that represents 3D matter as small Gaussians (“particles”) drawn from clusters capturing groups of coherently moving matter. We develop an e!cient inference algorithm based on …

    mit Repository record for Modeling Dynamic Objects in Scenes with Generative Particle Systems (opens in a new tab)

  14. Generalization of deep neural networks to unseen attribute combinations

    Visual understanding results from a combined understanding of primitive visual attributes such as color, texture, and shape. This allows humans and other primates to generalize their understanding of objects to new combinations of attributes. For instance, one can understand that a pink elephant is …

    mit Repository record for Generalization of deep neural networks to unseen attribute combinations (opens in a new tab)

  15. Dependency models as a basis for analyzing software product platform modularity : a case study in strategic software design rationalization

    … large software systems, which can be used to aid visual understanding of their structure, and which also form the basis for quantitative metrics for evaluating the degree of modularity of a software system (MacCormack et al., 2005; Sangal et al., 2005; Sullivan et al., 2006).

    mit Repository record for Dependency models as a basis for analyzing software product platform modularity : a case study in strategic software design rationalization (opens in a new tab)

  16. AI Commentator: Narrating Sports Games through Multimodal Perception and Large Language Models

    Automated visual understanding is an essential part of the sports industry, particularly in the context of major sports tournaments. The scale of generated video footage necessitates the use of automated systems to generate insights and enhance fan experiences. One area where this is particularly …

    mit Repository record for AI Commentator: Narrating Sports Games through Multimodal Perception and Large Language Models (opens in a new tab)

  17. Bridging the Realism Gap for CAD-Based Visual Recognition

    … are applied to guide manufacturing robots, to visually monitor plants, or to assist human operators in recognizing specific components. Recent progress in computer vision has been dominated by deep artificial neural network, i.e., machine learning methods simulating the way that information …

    passau-thes Repository record for Bridging the Realism Gap for CAD-Based Visual Recognition (opens in a new tab)

  18. Seeing the God of New Mexico: Mary Austin's Starry Adventure and the Optic of Enchantment

    … Mexican modernism/primitivism movement in the visual arts. I explore what I call Austin's "optic of enchantment," a visual experience of divinity that is uniquely tied to the New Mexican landscape. I break down this optic of enchantment into three distinct and definitive facets: First - a …

    denver Repository record for Seeing the God of New Mexico: Mary Austin's Starry Adventure and the Optic of Enchantment (opens in a new tab)

  19. Dense visual learning for robot manipulation

    … the scalable ability to understand the dense visual state of their varied environments. The limitations have both been in the state representations used, and how to acquire them without significant human labeling effort. In this thesis we present work that leverages self-supervision, …

    mit Repository record for Dense visual learning for robot manipulation (opens in a new tab)

  20. Computational strategies for understanding underwater optical image datasets

    … the time at which operators are able to gain a visual understanding of the survey environment. Typical missions can generate imagery at rates hundreds of times greater than highly compressed images can be transmitted acoustically, delaying that understanding until after the vehicle has been …

    woods-hole Repository record for Computational strategies for understanding underwater optical image datasets (opens in a new tab)

Page 1 of 2