Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 20 of 22 for “"Fault recovery"”.
-
Eventual fault recovery strategies for Byzantine failures
Byzantine faults in distributed systems can have very destructive consequences for services built on top of these systems but are not commonly tolerated in production systems due to the overhead and scalability limitations with existing approaches such as Byzantine fault tolerance. This work …
-
FastRecover: simple and effective fault recovery in a distributed operator-based stream processing engine
Fault tolerance is a key requirement in large-scale distributed stream processing engines (SPEs), especially those that run atop commodity hardware. Currently, fault tolerance in popular distributed SPEs is either inadequate (e.g., those without automatic recovery of operator states) or complex and …
-
Dynamic service recovery in a grid environment
… of services. These two cause services to be fault-prone. Therefore, there is a need to develop an autonomic fault recovery mechanism that will effectively monitor, diagnose and recover a running service from failure. In addressing the above mentioned challenge, a dynamic service recovery …
-
Dynamic service recovery in a grid environment
… of services. These two cause services to be fault prone. Therefore, there is a need to develop an autonomic fault recovery mechanism that will effectively monitor, diagnose and recover a running service from failure. In addressing the above mentioned challenge, a dynamic service recovery …
-
Fault detection, isolation, and recovery for autonomous parafoils
… transportation methods. The occurrence of a fault during a flight can severely degrade vehicle performance, effectively nullifying the value of the guided system, or worse. Quickly detecting and identifying faults enables the choice of an appropriate recovery strategy, potentially mitigating …
-
Recovery from transient faults in wavefront processor arrays
A transient fault in an array of processing elements results in an inconsistent or incorrect state in the processing element. If the erroneous information has already propagated before detection occurs, then the neighboring processing elements can also be in an incorrect state. Restarting the …
-
In-System Testing of Configurable Logic Blocks in Xilinx 7-Series FPGAs
FPGA fault recovery techniques, such as bitstream scrubbing, are only limited to detecting and correcting soft errors that corrupt the configuration memory. Scrubbing and related techniques cannot detect permanent faults within the FPGA fabric, such as short circuits and open circuits in FPGA …
-
Manual and compiler assisted methods for generating fault-tolerant parallel programs
Algorithm-based fault-tolerance (ABFT) is an inexpensive method of incorporating fault-tolerance into existing applications. Applications are modified to operate on encoded data and produce encoded results which may then be checked for correctness. An attractive feature of the scheme is that it …
-
Intelligent Application of Flexible AC Transmission System Components in an Evolving Power Grid
… power grid, including renewable energy sources, fault protection, and SMART grid technology. The addition of new energy sources has led to the decommissioning of inefficient energy sources. The implementation of new technologies and power load on a large scale, coupled with the removal of grid …
-
Detecting and Recovering from In-Core Hardware Faults Through Software Anomaly Treatment
… solution that effectively handles hardware faults while incurring low cost during the common mode of fault-free operations. SWAT is based on two key observations about the design of resilient systems. First, only those hardware faults that affect software need to be handled and second, since …
-
Microrid operation strategy for improved recovery and inertial response after large disturbances
… strategy, two large disturbances are studied: a fault in the distribution network that creates a reactive power imbalance due to induction motor stalling, and a sudden change in generation or consumption that leads to a real power imbalance. In the first part, a framework is created to study …
-
Efficient parallel processing and fault tolerance in a streaming join system
… distributed system, specifically sharding and fault tolerance. In particular, we are able to leverage RiverJoin's automatic parallelization mechanism to provide a data stream sharding interface that allows for close-to linear speedup in stream joins, allowing arbitrarily fast joins. Stream …
-
Software Reliability Issues: An Experimental Approach
… experiments.</p> <p>We show that controlling fault recovery order as represented by the data input to some well-known reliability models can enable them to produce more accurate predictions and can mitigate anomalous effects we attribute to manifestations of the fault interaction phenomenon. …
-
Generalized Task Structure Learning for Collaborative Multi-Robot/Human-Robot Task Allocation
… to ensure the task can be completely without faults. In order to alleviate these concerns, we have developed a generalized task structure which is able to transfer skills of a learned task to teams of heterogeneous robots. This system uses a small number of human demonstrations to learn a …
-
DC fault protection in HVDC system and the impact on AC frequency response
… duration of power outage in the event of a DC fault. The current definition of maximum loss-of-infeed for an AC network does not consider the duration of the power outage, and the impacts of DC fault protection arrangements which result in different speed of power restoration, on the system …
-
Orbital Maneuvers and Interplanetary Trajectory Design via Reinforcement Learning
… Potential applications include autonomous fault recovery, initial transfer design, and support for onboard decision-making during deep space operations.</p> <p>This dissertation contributes to the growing body of research on machine learning for space applications by systematically …
-
Automatic recovery for request oriented systems
Gracefully recovering from software and hardware faults is important to ensuring highly reliable and available systems. Operating systems have privileged access to all aspects of system operation, thus a fault related to them is able to affect the entire system. Existing approaches to operating …
-
Infrastructure sharing of 5G mobile core networks on an SDN/NFV platform
… of the VNFs allow for automation, scalability, fault recovery, and security to be evaluated. The testbed developed is readily re-creatable and based on open-source software.
-
Testing the effects of violating component axioms in validation of complex aircraft systems
This thesis focuses on estimating faults in complex large-scale integrated aircraft systems, especially where they interact with, and control, the aircraft dynamics. A general assumption considered in the reliability of such systems is that any component level fault will be monitored, detected and …
-
Detection of network anomalies and novel attacks in the internet via statistical network traffic separation and normality prediction
… and performance degradations is a key to rapid fault recovery and robust networking, and has been receiving increasing attention lately. In this dissertation we present a network anomaly detection methodology, which relies on the analysis of network traffic and the characterization of the …
Page 1 of 2