Structural equation modeling and genome-wide selection for multiple traits to enhance arabica coffee breeding programs.

No hay miniatura disponible

Fecha

Título de la revista

ISSN de la revista

Título del volumen

Editor

Resumen

Descripción

Recognizing the interrelationship among variables becomes critical in genetic breeding programs, where the goal is often to optimize selection for multiple traits. Conventional multi-trait models face challenges such as convergence issues, and they fail to account for cause-and-effect relationships. To address these challenges, we conducted a comprehensive analysis involving confirmatory factor analysis (CFA), Bayesian networks (BN), structural equation modeling (SEM), and genome-wide selection (GWS) using data from 195 arabica coffee plants. These plants were genotyped with 21,211 single nucleotide polymorphism markers as part of the Coffea arabica breeding program at UFV/EPAMIG/EMBRAPA. Traits included vegetative vigor (VV), canopy diameter (CD), number of vegetative nodes (NVN), number of reproductive nodes (NRN), leaf length (LL), and yield (Y). CFA established the following latent variables: vigor latent (VL) explaining VV and CD; nodes latent (NL) explaining NVN and NRN; leaf length latent (LLL) explaining LL; and yield latent (YL) explaining Y. These were integrated into the BN model, revealing the following key interrelationships: LLL → VL, LLL → NL, LLL → YL, VL → NL, and NL → YL. SEM estimated structural coefficients, highlighting the biological importance of VL → NL and NL → YL connections. Genomic predictions based on observed and latent variables showed that using VL to predict NVN and NRN traits resulted in similar gains to using NL. Predicting gains in Y using NL increased selection gains by 66.35% compared to YL. The SEM-GWS approach provided insights into selection strategies for traits linked with vegetative vigor, nodes, leaf length, and coffee yield, offering valuable guidance for advancing Arabica coffee breeding programs.

Palabras clave

Coffea Arábica, Genome, Plant breeding, Structural equation modeling, Bayesian theory, Nucleotides

Citación