Search PubMedSearch

PubMed · 40914749

Optimising parent selection in plant breeding: comparing metaheuristic algorithms for genotype building.

Abstract

Stacking desirable haplotypes across the genome to develop superior genotypes has been implemented in several crop species. A major challenge in Optimal Haplotype Selection is identifying a set of parents that collectively contain all desirable haplotypes, a complex combinatorial problem with countless possibilities. In this study, we evaluated the performance of metaheuristic search algorithms (MSAs)-genetic algorithm (GA), differential evolution (DE), particle swarm optimisation (PSO), and simulated annealing (SA) for optimising parent selection under two genotype building (GB) objectives: Optimal Haplotype Selection (OHS) and Optimal Population Value (OPV). Using a diverse wheat population of 583 lines genotyped for 29,972 SNPs, forming 7645 haplotype blocks and phenotyped for stripe rust scores, we assessed each algorithm's performance across fitness optimisation, convergence speed, and computational efficiency. GA consistently achieved high fitness and rapid convergence, while DE showed robustness but required longer runtime and careful tuning. PSO performed well under the OHS criterion but was less effective for OPV. SA, although computationally lighter, was less consistent in finding optimal solutions. Simulation over 100 breeding cycles showed that OHS outperformed both OPV and GEBV-based selection in long-term genetic gain and diversity retention. OHS maintained heterozygosity and additive variance, which are key for sustainable improvement, while GEBV selection led to early allele fixation. Our findings underscore the potential of GB strategies that prioritise the collective performance of parent sets rather than individual ranking to enhance selection outcomes in genomic-assisted breeding programmes.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

S Yadav, S Dillon, M McNeil, E Dinglasan, R Mago, P Dodds, L Hickey, B J Hayes. 2025-09-06. Optimising parent selection in plant breeding: comparing metaheuristic algorithms for genotype building.. https://doi.org/10.1007/s00122-025-05028-1

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

A comparative study highlights superiority of LSTM in crop genomic prediction.

We systematically evaluated three key determinants affecting prediction accuracy and the algorithm performance differences based on fifteen state-of-the-art GP methods, and found LSTM suitable for capturing additive and epistatic effects. Genomic prediction (GP) has been developed as an important method supporting crop breeding. By utilizing the phenotype values result from GP, breeders could make decisions in the seedling stage that consequently benefit for cost saving. In recent years, machine learning emerged as an efficient technology to solve modeling problems in many fields, including crop breeding. However, numerous modeling approaches have hindered the application of GP since breeders struggle to choose. Therefore, a comprehensively methodological research with guiding significance is extremely necessary. In the present study, we systematically evaluated three key determinants affecting prediction accuracy and the algorithm performance differences based on fifteen state-of-the-art GP methods. As for genomic feature processing, we found feature selection (SNP filtering approach) performed better than feature extraction (PCA method). Specifically, the feature relationship dependent methods (GBLUP, RNN, and LSTM) as well as DNN architecture showed superior performance with feature selection. Marker density analysis showed positive correlation with prediction accuracy in a limited threshold. Comparison on effect of population size demonstrated a positive correlation between trait genetic complexity and the optimal population size required. By testing fifteen modeling methods, we found LSTM network displayed superior performance, achieving the highest average STScore (0.967) across six datasets. Further research using all cell states or the latest cell states of LSTM inputs demonstrated its architecture particularly adept with capturing additive and epistatic QTL effects among SNPs. In conclusion, our findings provide basic principles for implementing GP in breeding project to maximize prediction accuracy while maintaining cost-effectiveness.

Plant Breeding