Genomic and phenomic prediction for soybean seed yield, protein, and oil

Liza Van der Laan
Kyle Parmley
Mojdeh Saadati
Hernan Torres Pacin
Srikanth Panthulugiri
Soumik Sarkar
Baskar Ganapathysubramanian
Aaron Lorenz
Asheesh K. Singh

Read the full article

Discuss this preprint

Start a discussion What are Sciety discussions?

Listed in

This article is not in any list yet, why not save it to one of your lists.

Abstract

Developments in genomics and phenomics have provided valuable tools for use in cultivar development. Genomic prediction (GP) has been used in commercial soybean [ Glycine max L. (Merr.)] breeding programs to predict grain yield and seed composition traits. Phenomic prediction (PP) is a rapidly developing field that holds the potential to be used for the selection of genotypes early in the growing season. The objectives of this study were to compare the performance of GP and PP for predicting soybean seed yield, protein, and oil. We additionally conducted genome‐wide association studies (GWAS) to identify significant single‐nucleotide polymorphisms (SNPs) associated with the traits of interest. The GWAS panel of 292 diverse accessions was grown in six environments in replicated trials. Spectral data were collected at two time points during the growing season. A genomic best linear unbiased prediction (GBLUP) model was trained on 269 accessions, while three separate machine learning (ML) models were trained on vegetation indices (VIs) and canopy traits. We observed that PP had a higher correlation coefficient than GP for seed yield, while GP had higher correlation coefficients for seed protein and oil contents. VIs with high feature importance were used as covariates in a new GBLUP model, and a new random forest model was trained with the inclusion of selected SNPs. These models did not outperform the original GP and PP models. These results show the capability of using ML for in‐season predictions for specific traits in soybean breeding and provide insights on PP and GP inclusions in breeding programs.

Version published to 10.1002/tpg2.70002
Feb 19, 2025
Version published to 10.1101/2024.11.01.621550 on bioRxiv
Nov 2, 2024

Nutritional Genomics of Tepary Bean (Phaseolus acutifolius): Genome‑wide association analysis and genomic prediction of seed nutritional traits and size

This article has 3 authors:
1. Sri Kiran Reddy Alla
2. Benedict Analin
3. Vijay Joshi
This article has no evaluationsLatest version Mar 10, 2026
Multi-Trait Selection Index for the Improvement of Agronomic and Yield Traits in Rice (Oryza sativa L.)

This article has 13 authors:
1. Elizabeth Norkor Nartey
2. Hygienus Godswill
3. Bernard Sakyiamah
4. Vincent Agyemang Opoku
5. Braima Amadu
6. Priscilla Francisco Ribeiro
7. Kirpal Agyemang Ofosu
8. Felix Frimpong
9. Daniel Dzorkpe Gamenyah
10. Stephen John Ayeh
11. Phyllis Aculey
12. Sober Ernest Boadu
13. Maxwell Darko Asante
This article has no evaluationsLatest version Jan 24, 2026
Integrating Meta-QTL Analysis and Genome-Wide Association Mapping in Ethiopian Sesame (Sesamum indicum L.) Reveals Novel Loci for Plant Height and Seed Coat Color

This article has 2 authors:
1. Adane Gebeyehu
2. Rodomiro Ortiz
This article has no evaluationsLatest version Feb 2, 2026

Discuss this preprint

Listed in

Abstract

Article activity feed

Related articles

Nutritional Genomics of Tepary Bean (Phaseolus acutifolius): Genome‑wide association analysis and genomic prediction of seed nutritional traits and size

Multi-Trait Selection Index for the Improvement of Agronomic and Yield Traits in Rice (Oryza sativa L.)

Integrating Meta-QTL Analysis and Genome-Wide Association Mapping in Ethiopian Sesame (Sesamum indicum L.) Reveals Novel Loci for Plant Height and Seed Coat Color