{"id":{"repo_id":"vt","oai_identifier":"oai:vtechworks.lib.vt.edu:10919/137496"},"canonical_url":"https://search.dev.ndltd.org/etd/vt/oai:vtechworks.lib.vt.edu:10919/137496","repository":{"repo_id":"vt","name":"Virginia Tech","base_url":"https://vtechworks.lib.vt.edu/oai/request"},"display":{"title":"The Advancements on the Interface of Statistical Computing, Survival Analysis, and Degradation Analysis","abstract":"In survival analysis and degradation analysis, analyzing complex data often requires advanced computing methods to carry out model estimation and inference. Motivated by computational challenges in survival analysis and a specific dataset in degradation analysis, this dissertation focuses on the applications of advanced computational methods. Furthermore, this dissertation develops novel computational methods to meet the demands of these applications. In competing risk analysis, computing the exact partial likelihood becomes challenging due to the presence of tied events. The primary challenge comes from computing the denominator of the partial likelihood, which is the risk score for the risk set at a given failure time. Historically, no method has efficiently addressed this issue. Approximation methods have been the predominate approach for decades even though their estimates for coefficients are biased. This dissertation presents a novel computational method for the partial likelihood. The method re-represents the denominator as a part of the probability mass function (pmf) of a Poisson multinomial distribution (PMD). However, efficient methods for computing the pmf of the PMD have yet to be developed. To bridge this gap, this dissertation introduces three distinct methods to calculate the pmf of the PMD under different circumstances. Additionally, this dissertation explores the potential applications of the PMD in voting theory, ecological inference, and machine learning. In degradation analysis, a dataset from a field study presents challenges in degradation modeling with the existence of multiple degradation characteristics (DCs) and dynamic covariates. A nonlinear general path model with random effects is employed to capture the data's complexity. A Bayesian framework helps solve the estimation difficulties in the model. Subsequently, Markov Chain Monte Carlo (MCMC) is employed to draw posterior inference conclusions. This dissertation also introduces a spline-based method to describe the effects of covariates. Unlike traditional spline regression methods, this approach does not suffer from overfitting and can model covariate effects under shape constraints.","abstract_html":"In survival analysis and degradation analysis, analyzing complex data often requires advanced computing methods to carry out model estimation and inference. Motivated by computational challenges in survival analysis and a specific dataset in degradation analysis, this dissertation focuses on the applications of advanced computational methods. Furthermore, this dissertation develops novel computational methods to meet the demands of these applications. In competing risk analysis, computing the exact partial likelihood becomes challenging due to the presence of tied events. The primary challenge comes from computing the denominator of the partial likelihood, which is the risk score for the risk set at a given failure time. Historically, no method has efficiently addressed this issue. Approximation methods have been the predominate approach for decades even though their estimates for coefficients are biased. This dissertation presents a novel computational method for the partial likelihood. The method re-represents the denominator as a part of the probability mass function (pmf) of a Poisson multinomial distribution (PMD). However, efficient methods for computing the pmf of the PMD have yet to be developed. To bridge this gap, this dissertation introduces three distinct methods to calculate the pmf of the PMD under different circumstances. Additionally, this dissertation explores the potential applications of the PMD in voting theory, ecological inference, and machine learning. In degradation analysis, a dataset from a field study presents challenges in degradation modeling with the existence of multiple degradation characteristics (DCs) and dynamic covariates. A nonlinear general path model with random effects is employed to capture the data&#x27;s complexity. A Bayesian framework helps solve the estimation difficulties in the model. Subsequently, Markov Chain Monte Carlo (MCMC) is employed to draw posterior inference conclusions. This dissertation also introduces a spline-based method to describe the effects of covariates. Unlike traditional spline regression methods, this approach does not suffer from overfitting and can model covariate effects under shape constraints.","abstract_has_math":false,"creators":["Lin, Zhengzhi"],"institution":"Virginia Tech","degree_name":"Doctor of Philosophy","degree_level":"doctoral","degree_discipline":"Statistics","degree_department":"Statistics","school":null,"contributors":[],"advisors":[],"committee_chairs":["Hong, Yili"],"committee_members":["Franck, Christopher Thomas","Zhu, Hongxiao","Deng, Xinwei","Kim, Inyoung"],"year":2025,"date_issued":"2025-08-13","date_published":"2025-08-13","updated_at":"2026-07-22T22:19:13Z","subjects":["Competing Risk","Poisson Multinomial Distribution","Bayesian Framework","Multiple Degradation Characteristics","General Path Models"],"languages":["en"],"rights":["Creative Commons Attribution-NonCommercial 4.0 International"],"rights_urls":["http://creativecommons.org/licenses/by-nc/4.0/"],"identifier_entries":[{"key":"dc:identifier.other","label":"Dc Identifier Other","values":["vt_gsexam:44459"],"render_values":[{"text":"vt_gsexam:44459","href":null,"code":true}]}]},"links":{"outbound_url":"https://hdl.handle.net/10919/137496","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.committeechair","label":"Committee Chair","values":["Hong, Yili"]},{"key":"dc:contributor.committeemember","label":"Committee Member","values":["Franck, Christopher Thomas","Zhu, Hongxiao","Deng, Xinwei","Kim, Inyoung"]},{"key":"dc:contributor.department","label":"Department","values":["Statistics"]},{"key":"dc:creator","label":"Author","values":["Lin, Zhengzhi"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2025-08-14T08:00:30Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2025-08-14T08:00:30Z"]},{"key":"dc:date.issued","label":"Date","values":["2025-08-13"]},{"key":"dc:publisher","label":"Institution","values":["Virginia Tech"]},{"key":"dc:type","label":"Dc Type","values":["Dissertation"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Statistics"]},{"key":"thesis:degree_level","label":"Degree Level","values":["doctoral"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Doctor of Philosophy"]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["Virginia Polytechnic Institute and State University"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Competing Risk","Poisson Multinomial Distribution","Bayesian Framework","Multiple Degradation Characteristics","General Path Models"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["en"]},{"key":"dc:rights","label":"Dc Rights","values":["Creative Commons Attribution-NonCommercial 4.0 International"]},{"key":"dc:rights.uri","label":"Rights URI","values":["http://creativecommons.org/licenses/by-nc/4.0/"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.other","label":"Dc Identifier Other","values":["vt_gsexam:44459"]},{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://hdl.handle.net/10919/137496"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["In survival analysis and degradation analysis, analyzing complex data often requires advanced computing methods to carry out model estimation and inference. Motivated by computational challenges in survival analysis and a specific dataset in degradation analysis, this dissertation focuses on the applications of advanced computational methods. Furthermore, this dissertation develops novel computational methods to meet the demands of these applications. In competing risk analysis, computing the exact partial likelihood becomes challenging due to the presence of tied events. The primary challenge comes from computing the denominator of the partial likelihood, which is the risk score for the risk set at a given failure time. Historically, no method has efficiently addressed this issue. Approximation methods have been the predominate approach for decades even though their estimates for coefficients are biased. This dissertation presents a novel computational method for the partial likelihood. The method re-represents the denominator as a part of the probability mass function (pmf) of a Poisson multinomial distribution (PMD). However, efficient methods for computing the pmf of the PMD have yet to be developed. To bridge this gap, this dissertation introduces three distinct methods to calculate the pmf of the PMD under different circumstances. Additionally, this dissertation explores the potential applications of the PMD in voting theory, ecological inference, and machine learning. In degradation analysis, a dataset from a field study presents challenges in degradation modeling with the existence of multiple degradation characteristics (DCs) and dynamic covariates. A nonlinear general path model with random effects is employed to capture the data's complexity. A Bayesian framework helps solve the estimation difficulties in the model. Subsequently, Markov Chain Monte Carlo (MCMC) is employed to draw posterior inference conclusions. This dissertation also introduces a spline-based method to describe the effects of covariates. Unlike traditional spline regression methods, this approach does not suffer from overfitting and can model covariate effects under shape constraints."]},{"key":"dc:description.abstractgeneral","label":"General Abstract","values":["Modern data in biostatistics, engineering, and social science often involve tracking how long it takes for events to happen—like the failure of a machine or the recovery of a patient. Analyzing this kind of data is challenging, especially when events can happen for different reasons or when the data is complex. This dissertation aims at such challenges by developing and applying novel computational tools. One major part of the work focuses on understanding risks when different types of failure can occur. A long-standing problem in this area has been how to handle situations when many events happen at the same time. Traditional methods use approximations that can lead to errors. While approximation methods have been standard for decades, they introduce bias in parameter estimates. This dissertation proposes a new approach that reformulates the partial likelihood function using the probability mass function (pmf) of the Poisson multinomial distribution (PMD). Since no general-purpose method currently exists for efficiently computing the PMD's pmf, this dissertation develops three computational strategies to handle different practical scenarios. These innovations also extend the applicability of PMD-based inference to domains such as voting theory, ecological inference, and machine learning. In degradation analysis, this dissertation investigates a field dataset characterized by multiple degradation paths and time-varying environmental factors such as temperature and humidity. This dissertation proposes a model to describe this complexity. Estimation is performed within a Bayesian framework, and a novel spline-based method is introduced to flexibly model the factors' effects under given domain assumptions, reducing the risk of overfitting commonly seen in traditional spline approaches."]},{"key":"dc:description.degree","label":"Dc Description Degree","values":["Doctor of Philosophy"]},{"key":"dc:format.medium","label":"Dc Format Medium","values":["ETD"]},{"key":"dc:title","label":"Title","values":["The Advancements on the Interface of Statistical Computing, Survival Analysis, and Degradation Analysis"]}]}],"canonical_facts":{"dc:contributor.committeechair":["Hong, Yili"],"dc:contributor.committeemember":["Franck, Christopher Thomas","Zhu, Hongxiao","Deng, Xinwei","Kim, Inyoung"],"dc:contributor.department":["Statistics"],"dc:creator":["Lin, Zhengzhi"],"dc:date.accessioned":["2025-08-14T08:00:30Z"],"dc:date.available":["2025-08-14T08:00:30Z"],"dc:date.issued":["2025-08-13"],"dc:description.abstract":["In survival analysis and degradation analysis, analyzing complex data often requires advanced computing methods to carry out model estimation and inference. Motivated by computational challenges in survival analysis and a specific dataset in degradation analysis, this dissertation focuses on the applications of advanced computational methods. Furthermore, this dissertation develops novel computational methods to meet the demands of these applications. In competing risk analysis, computing the exact partial likelihood becomes challenging due to the presence of tied events. The primary challenge comes from computing the denominator of the partial likelihood, which is the risk score for the risk set at a given failure time. Historically, no method has efficiently addressed this issue. Approximation methods have been the predominate approach for decades even though their estimates for coefficients are biased. This dissertation presents a novel computational method for the partial likelihood. The method re-represents the denominator as a part of the probability mass function (pmf) of a Poisson multinomial distribution (PMD). However, efficient methods for computing the pmf of the PMD have yet to be developed. To bridge this gap, this dissertation introduces three distinct methods to calculate the pmf of the PMD under different circumstances. Additionally, this dissertation explores the potential applications of the PMD in voting theory, ecological inference, and machine learning. In degradation analysis, a dataset from a field study presents challenges in degradation modeling with the existence of multiple degradation characteristics (DCs) and dynamic covariates. A nonlinear general path model with random effects is employed to capture the data's complexity. A Bayesian framework helps solve the estimation difficulties in the model. Subsequently, Markov Chain Monte Carlo (MCMC) is employed to draw posterior inference conclusions. This dissertation also introduces a spline-based method to describe the effects of covariates. Unlike traditional spline regression methods, this approach does not suffer from overfitting and can model covariate effects under shape constraints."],"dc:description.abstractgeneral":["Modern data in biostatistics, engineering, and social science often involve tracking how long it takes for events to happen—like the failure of a machine or the recovery of a patient. Analyzing this kind of data is challenging, especially when events can happen for different reasons or when the data is complex. This dissertation aims at such challenges by developing and applying novel computational tools. One major part of the work focuses on understanding risks when different types of failure can occur. A long-standing problem in this area has been how to handle situations when many events happen at the same time. Traditional methods use approximations that can lead to errors. While approximation methods have been standard for decades, they introduce bias in parameter estimates. This dissertation proposes a new approach that reformulates the partial likelihood function using the probability mass function (pmf) of the Poisson multinomial distribution (PMD). Since no general-purpose method currently exists for efficiently computing the PMD's pmf, this dissertation develops three computational strategies to handle different practical scenarios. These innovations also extend the applicability of PMD-based inference to domains such as voting theory, ecological inference, and machine learning. In degradation analysis, this dissertation investigates a field dataset characterized by multiple degradation paths and time-varying environmental factors such as temperature and humidity. This dissertation proposes a model to describe this complexity. Estimation is performed within a Bayesian framework, and a novel spline-based method is introduced to flexibly model the factors' effects under given domain assumptions, reducing the risk of overfitting commonly seen in traditional spline approaches."],"dc:description.degree":["Doctor of Philosophy"],"dc:format.medium":["ETD"],"dc:identifier.other":["vt_gsexam:44459"],"dc:identifier.uri":["https://hdl.handle.net/10919/137496"],"dc:language.iso":["en"],"dc:publisher":["Virginia Tech"],"dc:rights":["Creative Commons Attribution-NonCommercial 4.0 International"],"dc:rights.uri":["http://creativecommons.org/licenses/by-nc/4.0/"],"dc:subject":["Competing Risk","Poisson Multinomial Distribution","Bayesian Framework","Multiple Degradation Characteristics","General Path Models"],"dc:title":["The Advancements on the Interface of Statistical Computing, Survival Analysis, and Degradation Analysis"],"dc:type":["Dissertation"],"thesis:degree_discipline":["Statistics"],"thesis:degree_level":["doctoral"],"thesis:degree_name":["Doctor of Philosophy"],"thesis:institution_name":["Virginia Polytechnic Institute and State University"]},"updated_at":"2026-07-22T22:19:13Z"}