Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 20 of 30 for “"Visual understanding"”.
-
Toward ontological visual understanding
… remains when it comes to automatic multimedia understanding. One difference between the human cognition system and state-of-the-art machine vision algorithms is that the former possesses and uses high-level semantic knowledge, or ontology. In this thesis, we present our work on image-level …
-
Toward visual understanding of everyday object
… using large scale data. However, for any visual system to interact with objects, it needs to understand much more than simply recognizing where the objects are. The goal of my research is to explore and solve object understanding tasks for interaction - finding an object's pose in 3D, …
-
Label-Efficient Visual Understanding with Consistency Constraints
… networks are proficient at solving various visual recognition and understanding tasks, as long as a sufficiently large labeled dataset is available during the training time. However, the progress of these visual tasks is limited by the number of manual annotations. On the other hand, it is …
-
Towards Comprehensive Visual Understanding via Deep Neural Networks
… (DNNs) have made significant advancements in visual scene understanding, demonstrating great potential for applications in downstream tasks such as autonomous driving, robotic navigation, and human-computer interaction. Despite these successes, generalization ability remains a major obstacle …
-
LEARNING STRUCTURED ALIGNMENT: FROM VISUAL UNDERSTANDING TO ROBOT CONTROL
The rapid growth of visual data and the increasing demand for intelligent robotic systemshave created a pressing need for methods that can establish meaningful correspondences and relationships across diverse visual modalities and robotic tasks. This dissertation addresses the fundamental challenge …
-
Visual Understanding of Hybrid Circuits via Procedural Models: An Approach to Automatic Inspection of Hybrid Circuits
Made available in DSpace on 2014-12-12T20:54:12Z (GMT). No. of bitstreams: 1 7524408.pdf: 5803201 bytes, checksum: 19eeae58595e41f38a37da6d9b2483b2 (MD5) Previous issue date: 1975
-
Assessing the development of high school chemistry students' conceptual and visual understanding of dimensional analysis via supplemental use of a proprietary interactive software program
… “Conversionoes,” on students' conceptual and visual understanding of dimensional analysis. The participants in the study were high school general chemistry students enrolled in two public high schools with different demographics (School A and School B) in the Chattanooga, Tennessee, …
-
ARGILE: A Conceptual Framework Combining Augmented Reality with Agile Philosophy for the UK Construction Industry
… agile project management and augmented reality visualisation technology. It enables an understanding on how the use of the proposed, designed, and developed ARGILE framework (integration of Agile and Augmented Reality) enhances the collaboration, communication, decision-making, and visual …
-
Learning from minimally labeled data with accelerated convolutional neural networks
… to be satisfied by the mapping function for visual understanding. First, the function should produce good representations of the visual world, which will be able to recognize images independently of pose, scale and illumination. Furthermore, the designed artificial vision system has to learn …
-
Predictive vision
… an extensive set of rules that enable richer visual understanding. While large annotated datasets fuel progress in object recognition, the knowledge required for event understanding is vast and potentially ambiguous. To tackle this challenge, we develop predictive vision algorithms that …
-
Deep Learning Multimodal Extraction of Reaction Data
… allow powerful large language models to leverage visual understanding. Our findings indicate that VLMs still require additional work in order to meet the performance of our bespoke models.
-
Unsupervised Compositional Image Decompositionwith Diffusion Models
Our visual understanding of the world is factorized and compositional. With just a single observation, we can ascertain both global and local attributes in a scene, such as lighting, weather, and underlying objects. These attributes are highly compositional and can be combined in various ways to …
-
Modeling Dynamic Objects in Scenes with Generative Particle Systems
… serves as a critical signal for recognition and understanding. Inspired by this ability, we propose a generative model that represents 3D matter as small Gaussians (“particles”) drawn from clusters capturing groups of coherently moving matter. We develop an e!cient inference algorithm based on …
-
Generalization of deep neural networks to unseen attribute combinations
Visual understanding results from a combined understanding of primitive visual attributes such as color, texture, and shape. This allows humans and other primates to generalize their understanding of objects to new combinations of attributes. For instance, one can understand that a pink elephant is …
-
Dependency models as a basis for analyzing software product platform modularity : a case study in strategic software design rationalization
… large software systems, which can be used to aid visual understanding of their structure, and which also form the basis for quantitative metrics for evaluating the degree of modularity of a software system (MacCormack et al., 2005; Sangal et al., 2005; Sullivan et al., 2006).
-
AI Commentator: Narrating Sports Games through Multimodal Perception and Large Language Models
Automated visual understanding is an essential part of the sports industry, particularly in the context of major sports tournaments. The scale of generated video footage necessitates the use of automated systems to generate insights and enhance fan experiences. One area where this is particularly …
-
Bridging the Realism Gap for CAD-Based Visual Recognition
… are applied to guide manufacturing robots, to visually monitor plants, or to assist human operators in recognizing specific components. Recent progress in computer vision has been dominated by deep artificial neural network, i.e., machine learning methods simulating the way that information …
-
Seeing the God of New Mexico: Mary Austin's Starry Adventure and the Optic of Enchantment
… Mexican modernism/primitivism movement in the visual arts. I explore what I call Austin's "optic of enchantment," a visual experience of divinity that is uniquely tied to the New Mexican landscape. I break down this optic of enchantment into three distinct and definitive facets: First - a …
-
Dense visual learning for robot manipulation
… the scalable ability to understand the dense visual state of their varied environments. The limitations have both been in the state representations used, and how to acquire them without significant human labeling effort. In this thesis we present work that leverages self-supervision, …
-
Computational strategies for understanding underwater optical image datasets
… the time at which operators are able to gain a visual understanding of the survey environment. Typical missions can generate imagery at rates hundreds of times greater than highly compressed images can be transmitted acoustically, delaying that understanding until after the vehicle has been …
Page 1 of 2