{"id":{"repo_id":"brazil-ufv","oai_identifier":"oai:locus.ufv.br:123456789/32302"},"canonical_url":"https://search.dev.ndltd.org/etd/brazil-ufv/oai:locus.ufv.br:123456789/32302","repository":{"repo_id":"brazil-ufv","name":"Brazil UFV","base_url":"https://locus.ufv.br/server/oai/request"},"display":{"title":"Statistical genetics tools for empowered data-driven decisions","abstract":"The pressure to accelerate results in plant breeding programs is intensifying. Conversely, there is a concerning decline in the genetic diversity of staple crops, making it increasingly diﬃcult to achieve genetic gains. Consequently, eﬃcient resource allocation within breeding programs requires the strategic implementation of statistical genetics tools. This shift necessitates data-driven decision-making, placing professionals proﬁcient in this toolkit at a signiﬁcant advantage for addressing both traditional and emerging challenges. This thesis serves as a practical demonstration of utilizing statistical genetics in various plant breeding endeavours. Divided into six chapters, each with distinct objectives, the work showcases a range of applications. In Chapter 1, we determined the optimal number of harvests for selection in cacao breeding, considering both recommendation and recombination. Chapter 2 explores the application of covariance structure modelling in two common scenarios of perennial plant breeding: multi-harvest and multi-site data analysis. Chapter 3 demonstrates the use of factor analytic mixed models in maize breeding, including the incorporation of selection tools for streamlined decision-making. Notably, this chapter highlights the advantage of seasonal selection for achieving greater genetic gains compared to a combined approach. In Chapter 4, we evaluated the eﬃcacy of the reciprocal recurrent selection (RRS) scheme within a eucalyptus breeding program. This chapter acknowledges the extended timeframe associated with RRS but also demonstrates its success in enhancing the hybrid population. Additionally, the chapter emphasizes the importance of considering dominance eﬀects during the selection process. Chapter 5 oﬀers a comprehensive tutorial on conducting linear mixed model analyses in perennial plant breeding. The chapter covers various analyses, including individual trials, multi- environment trials, spatial analysis, and competition analysis. Finally, Chapter 6 introduces the R package ProbBreed, which utilizes Bayesian principles and probabilistic concepts to support selection in multi-environment trials. ProbBreed estimates the risk associated with selecting candidates, empowering more informed decision-making. This chapter also introduces a novel multi-location-year model and compares the outcomes of ProbBreed and ASReml-R using simulated data. By showcasing the applications of statistical genetics tools and facilitating knowledge sharing through open-source code and reproducible examples, this thesis emphasizes the versatility and importance of this ﬁeld in tackling diverse challenges within the dynamic ﬁeld of plant breeding. Keywords: Data Analysis. Linear Mixed Models. Bayesian models. Genotype-by-Environment interaction. Spatial Analysis. Reciprocal Recurrent Selection","abstract_html":"The pressure to accelerate results in plant breeding programs is intensifying. Conversely, there is a concerning decline in the genetic diversity of staple crops, making it increasingly diﬃcult to achieve genetic gains. Consequently, eﬃcient resource allocation within breeding programs requires the strategic implementation of statistical genetics tools. This shift necessitates data-driven decision-making, placing professionals proﬁcient in this toolkit at a signiﬁcant advantage for addressing both traditional and emerging challenges. This thesis serves as a practical demonstration of utilizing statistical genetics in various plant breeding endeavours. Divided into six chapters, each with distinct objectives, the work showcases a range of applications. In Chapter 1, we determined the optimal number of harvests for selection in cacao breeding, considering both recommendation and recombination. Chapter 2 explores the application of covariance structure modelling in two common scenarios of perennial plant breeding: multi-harvest and multi-site data analysis. Chapter 3 demonstrates the use of factor analytic mixed models in maize breeding, including the incorporation of selection tools for streamlined decision-making. Notably, this chapter highlights the advantage of seasonal selection for achieving greater genetic gains compared to a combined approach. In Chapter 4, we evaluated the eﬃcacy of the reciprocal recurrent selection (RRS) scheme within a eucalyptus breeding program. This chapter acknowledges the extended timeframe associated with RRS but also demonstrates its success in enhancing the hybrid population. Additionally, the chapter emphasizes the importance of considering dominance eﬀects during the selection process. Chapter 5 oﬀers a comprehensive tutorial on conducting linear mixed model analyses in perennial plant breeding. The chapter covers various analyses, including individual trials, multi- environment trials, spatial analysis, and competition analysis. Finally, Chapter 6 introduces the R package ProbBreed, which utilizes Bayesian principles and probabilistic concepts to support selection in multi-environment trials. ProbBreed estimates the risk associated with selecting candidates, empowering more informed decision-making. This chapter also introduces a novel multi-location-year model and compares the outcomes of ProbBreed and ASReml-R using simulated data. By showcasing the applications of statistical genetics tools and facilitating knowledge sharing through open-source code and reproducible examples, this thesis emphasizes the versatility and importance of this ﬁeld in tackling diverse challenges within the dynamic ﬁeld of plant breeding. Keywords: Data Analysis. Linear Mixed Models. Bayesian models. Genotype-by-Environment interaction. Spatial Analysis. Reciprocal Recurrent Selection","abstract_has_math":false,"creators":["Chaves, Saulo Fabrício da Silva"],"institution":"Universidade Federal de Viçosa","degree_name":null,"degree_level":null,"degree_discipline":null,"degree_department":null,"school":null,"contributors":["Dias, Kaio Olimpio das Graças","Alves, Rodrigo Silva"],"advisors":["Dias, Luiz Antônio dos Santos"],"committee_chairs":[],"committee_members":[],"year":2024,"date_issued":"2024-03-18","date_published":"2024-03-18","updated_at":"2026-07-24T01:21:22Z","subjects":["Plantas - Melhoramento genético","Genética quantitativa"],"languages":["eng"],"rights":["Acesso Aberto"],"rights_urls":[],"identifier_entries":[{"key":"dc:identifier.doi","label":"DOI","values":["https://doi.org/10.47328/ufvbbt.2024.122"],"render_values":[{"text":"https://doi.org/10.47328/ufvbbt.2024.122","href":"https://doi.org/10.47328/ufvbbt.2024.122","code":true}]}]},"links":{"outbound_url":"https://locus.ufv.br//handle/123456789/32302","outbound_label":"Repository record","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Dias, Kaio Olimpio das Graças","Alves, Rodrigo Silva"]},{"key":"dc:contributor.advisor","label":"Advisor","values":["Dias, Luiz Antônio dos Santos"]},{"key":"dc:creator","label":"Author","values":["Chaves, Saulo Fabrício da Silva"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2024-06-05T15:33:56Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2024-06-05T15:33:56Z"]},{"key":"dc:date.issued","label":"Date","values":["2024-03-18"]},{"key":"dc:publisher","label":"Institution","values":["Universidade Federal de Viçosa"]},{"key":"dc:type","label":"Dc Type","values":["Tese"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Plantas - Melhoramento genético","Genética quantitativa"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["eng"]},{"key":"dc:rights","label":"Dc Rights","values":["Acesso Aberto"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.doi","label":"DOI","values":["https://doi.org/10.47328/ufvbbt.2024.122"]},{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://locus.ufv.br//handle/123456789/32302"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["The pressure to accelerate results in plant breeding programs is intensifying. Conversely, there is a concerning decline in the genetic diversity of staple crops, making it increasingly diﬃcult to achieve genetic gains. Consequently, eﬃcient resource allocation within breeding programs requires the strategic implementation of statistical genetics tools. This shift necessitates data-driven decision-making, placing professionals proﬁcient in this toolkit at a signiﬁcant advantage for addressing both traditional and emerging challenges. This thesis serves as a practical demonstration of utilizing statistical genetics in various plant breeding endeavours. Divided into six chapters, each with distinct objectives, the work showcases a range of applications. In Chapter 1, we determined the optimal number of harvests for selection in cacao breeding, considering both recommendation and recombination. Chapter 2 explores the application of covariance structure modelling in two common scenarios of perennial plant breeding: multi-harvest and multi-site data analysis. Chapter 3 demonstrates the use of factor analytic mixed models in maize breeding, including the incorporation of selection tools for streamlined decision-making. Notably, this chapter highlights the advantage of seasonal selection for achieving greater genetic gains compared to a combined approach. In Chapter 4, we evaluated the eﬃcacy of the reciprocal recurrent selection (RRS) scheme within a eucalyptus breeding program. This chapter acknowledges the extended timeframe associated with RRS but also demonstrates its success in enhancing the hybrid population. Additionally, the chapter emphasizes the importance of considering dominance eﬀects during the selection process. Chapter 5 oﬀers a comprehensive tutorial on conducting linear mixed model analyses in perennial plant breeding. The chapter covers various analyses, including individual trials, multi- environment trials, spatial analysis, and competition analysis. Finally, Chapter 6 introduces the R package ProbBreed, which utilizes Bayesian principles and probabilistic concepts to support selection in multi-environment trials. ProbBreed estimates the risk associated with selecting candidates, empowering more informed decision-making. This chapter also introduces a novel multi-location-year model and compares the outcomes of ProbBreed and ASReml-R using simulated data. By showcasing the applications of statistical genetics tools and facilitating knowledge sharing through open-source code and reproducible examples, this thesis emphasizes the versatility and importance of this ﬁeld in tackling diverse challenges within the dynamic ﬁeld of plant breeding. Keywords: Data Analysis. Linear Mixed Models. Bayesian models. Genotype-by-Environment interaction. Spatial Analysis. Reciprocal Recurrent Selection","A pressão para acelerar os resultados dos programas de melhoramento de plantas está a intensiﬁcar-se. Em contraste, há um declínio preocupante na diversidade genética das principais culturas, tornando cada vez mais difícil obter ganhos genéticos consistentes. Consequentemente, a alocação eﬁciente de recursos nos programas de melhoramento requer a implementação estratégica de ferramentas de genética-estatística. Esta mudança exige uma tomada de decisão baseada em dados, colocando os proﬁssionais proﬁcientes neste conjunto de ferramentas numa vantagem signiﬁcativa para enfrentar desaﬁos tradicionais e emergentes. Esta tese serve como uma demonstração prática da utilização da genética-estatística em várias vertentes dod melhoramento de plantas. Dividido em seis capítulos, cada um com objetivos distintos, o trabalho apresenta uma gama de aplicações. No Capítulo 1 determinamos o número ideal de colheitas para seleção no melhoramento de cacau, considerando recomendação e recombinação. O Capítulo 2 explora a aplicação da modelagem de estrutura de covariância em dois cenários comuns de melhoramento de plantas perenes: análise de dados de múltiplas colheitas e de múltiplos locais. O Capítulo 3 demonstra o uso de modelos fator analítico no melhoramento de milho, incluindo a incorporação de ferramentas de seleção para agilizar a tomada de decisões. Notavelmente, este capítulo destaca a vantagem da seleção sazonal para alcançar maiores ganhos genéticos em comparação com uma abordagem combinada. No Capítulo 4 avaliamos a eﬁcácia do esquema de seleção recorrente recíproca (SRR) dentro de um programa de melhoramento genético de eucalipto. Este capítulo reconhece o prazo alargado associado ao SRR, mas também demonstra o seu sucesso no aumento da população híbrida. Além disso, o capítulo enfatiza a importância de considerar os efeitos de dominância durante o processo de seleção. O Capítulo 5 oferece um tutorial abrangente sobre a condução de análises baseadas em modelos lineares mistos no melhoramento de plantas perenes. O capítulo cobre várias análises, incluindo ensaios individuais, ensaios multiambientais, análise espacial e análise de concorrência. Finalmente, o Capítulo 6 apresenta o pacote do R ProbBreed, que utiliza princípios bayesianos e conceitos probabilísticos para apoiar a seleção em ensaios multiambientes. ProbBreed estima o risco associado à seleção de candidatos, capacitando uma tomada de decisão mais informada. Este capítulo também apresenta um novo modelo de múltiplos anos e locais e compara os resultados dos pacotes ProbBreed e ASReml-R usando dados simulados. Ao mostrar as aplicações de ferramentas de genética-estatística e facilitar o compartilhamento de conhecimento através de códigos e exemplos reproduzíveis, esta tese enfatiza a versatilidade e a importância deste campo no enfrentamento de diversos desaﬁos no dinâmico campo do melhoramento de plantas. Palavras-chave: Análise de dados. Modelos Lineares Mistos. Modelos Bayesianos. Interação Genótipos-Ambientes. Análise Espacial. Seleção Recorrente Recíproca"]},{"key":"dc:title","label":"Title","values":["Statistical genetics tools for empowered data-driven decisions","Ferramentas de genética estatística para decisões qualiﬁcadas baseadas em dados"]}]}],"canonical_facts":{"dc:contributor":["Dias, Kaio Olimpio das Graças","Alves, Rodrigo Silva"],"dc:contributor.advisor":["Dias, Luiz Antônio dos Santos"],"dc:creator":["Chaves, Saulo Fabrício da Silva"],"dc:date.accessioned":["2024-06-05T15:33:56Z"],"dc:date.available":["2024-06-05T15:33:56Z"],"dc:date.issued":["2024-03-18"],"dc:description.abstract":["The pressure to accelerate results in plant breeding programs is intensifying. Conversely, there is a concerning decline in the genetic diversity of staple crops, making it increasingly diﬃcult to achieve genetic gains. Consequently, eﬃcient resource allocation within breeding programs requires the strategic implementation of statistical genetics tools. This shift necessitates data-driven decision-making, placing professionals proﬁcient in this toolkit at a signiﬁcant advantage for addressing both traditional and emerging challenges. This thesis serves as a practical demonstration of utilizing statistical genetics in various plant breeding endeavours. Divided into six chapters, each with distinct objectives, the work showcases a range of applications. In Chapter 1, we determined the optimal number of harvests for selection in cacao breeding, considering both recommendation and recombination. Chapter 2 explores the application of covariance structure modelling in two common scenarios of perennial plant breeding: multi-harvest and multi-site data analysis. Chapter 3 demonstrates the use of factor analytic mixed models in maize breeding, including the incorporation of selection tools for streamlined decision-making. Notably, this chapter highlights the advantage of seasonal selection for achieving greater genetic gains compared to a combined approach. In Chapter 4, we evaluated the eﬃcacy of the reciprocal recurrent selection (RRS) scheme within a eucalyptus breeding program. This chapter acknowledges the extended timeframe associated with RRS but also demonstrates its success in enhancing the hybrid population. Additionally, the chapter emphasizes the importance of considering dominance eﬀects during the selection process. Chapter 5 oﬀers a comprehensive tutorial on conducting linear mixed model analyses in perennial plant breeding. The chapter covers various analyses, including individual trials, multi- environment trials, spatial analysis, and competition analysis. Finally, Chapter 6 introduces the R package ProbBreed, which utilizes Bayesian principles and probabilistic concepts to support selection in multi-environment trials. ProbBreed estimates the risk associated with selecting candidates, empowering more informed decision-making. This chapter also introduces a novel multi-location-year model and compares the outcomes of ProbBreed and ASReml-R using simulated data. By showcasing the applications of statistical genetics tools and facilitating knowledge sharing through open-source code and reproducible examples, this thesis emphasizes the versatility and importance of this ﬁeld in tackling diverse challenges within the dynamic ﬁeld of plant breeding. Keywords: Data Analysis. Linear Mixed Models. Bayesian models. Genotype-by-Environment interaction. Spatial Analysis. Reciprocal Recurrent Selection","A pressão para acelerar os resultados dos programas de melhoramento de plantas está a intensiﬁcar-se. Em contraste, há um declínio preocupante na diversidade genética das principais culturas, tornando cada vez mais difícil obter ganhos genéticos consistentes. Consequentemente, a alocação eﬁciente de recursos nos programas de melhoramento requer a implementação estratégica de ferramentas de genética-estatística. Esta mudança exige uma tomada de decisão baseada em dados, colocando os proﬁssionais proﬁcientes neste conjunto de ferramentas numa vantagem signiﬁcativa para enfrentar desaﬁos tradicionais e emergentes. Esta tese serve como uma demonstração prática da utilização da genética-estatística em várias vertentes dod melhoramento de plantas. Dividido em seis capítulos, cada um com objetivos distintos, o trabalho apresenta uma gama de aplicações. No Capítulo 1 determinamos o número ideal de colheitas para seleção no melhoramento de cacau, considerando recomendação e recombinação. O Capítulo 2 explora a aplicação da modelagem de estrutura de covariância em dois cenários comuns de melhoramento de plantas perenes: análise de dados de múltiplas colheitas e de múltiplos locais. O Capítulo 3 demonstra o uso de modelos fator analítico no melhoramento de milho, incluindo a incorporação de ferramentas de seleção para agilizar a tomada de decisões. Notavelmente, este capítulo destaca a vantagem da seleção sazonal para alcançar maiores ganhos genéticos em comparação com uma abordagem combinada. No Capítulo 4 avaliamos a eﬁcácia do esquema de seleção recorrente recíproca (SRR) dentro de um programa de melhoramento genético de eucalipto. Este capítulo reconhece o prazo alargado associado ao SRR, mas também demonstra o seu sucesso no aumento da população híbrida. Além disso, o capítulo enfatiza a importância de considerar os efeitos de dominância durante o processo de seleção. O Capítulo 5 oferece um tutorial abrangente sobre a condução de análises baseadas em modelos lineares mistos no melhoramento de plantas perenes. O capítulo cobre várias análises, incluindo ensaios individuais, ensaios multiambientais, análise espacial e análise de concorrência. Finalmente, o Capítulo 6 apresenta o pacote do R ProbBreed, que utiliza princípios bayesianos e conceitos probabilísticos para apoiar a seleção em ensaios multiambientes. ProbBreed estima o risco associado à seleção de candidatos, capacitando uma tomada de decisão mais informada. Este capítulo também apresenta um novo modelo de múltiplos anos e locais e compara os resultados dos pacotes ProbBreed e ASReml-R usando dados simulados. Ao mostrar as aplicações de ferramentas de genética-estatística e facilitar o compartilhamento de conhecimento através de códigos e exemplos reproduzíveis, esta tese enfatiza a versatilidade e a importância deste campo no enfrentamento de diversos desaﬁos no dinâmico campo do melhoramento de plantas. Palavras-chave: Análise de dados. Modelos Lineares Mistos. Modelos Bayesianos. Interação Genótipos-Ambientes. Análise Espacial. Seleção Recorrente Recíproca"],"dc:identifier.doi":["https://doi.org/10.47328/ufvbbt.2024.122"],"dc:identifier.uri":["https://locus.ufv.br//handle/123456789/32302"],"dc:language.iso":["eng"],"dc:publisher":["Universidade Federal de Viçosa"],"dc:rights":["Acesso Aberto"],"dc:subject":["Plantas - Melhoramento genético","Genética quantitativa"],"dc:title":["Statistical genetics tools for empowered data-driven decisions","Ferramentas de genética estatística para decisões qualiﬁcadas baseadas em dados"],"dc:type":["Tese"]},"updated_at":"2026-07-24T01:21:22Z"}