Searching the Resource Information Network

Our searching services are busy right now. Please try again later

  • Register
X
Forgot Password

If you have forgotten your password you can enter your email here and get a temporary password sent to your email.

X

Leaving Community

Are you sure you want to leave this community? Leaving the community will revoke any permissions you have been granted in this community.

No
Yes

Genetic variation among 481 diverse soybean accessions, inferred from genomic re-sequencing.

Babu Valliyodan | Anne V Brown | Juexin Wang | Gunvant Patil | Yang Liu | Paul I Otyama | Rex T Nelson | Tri Vuong | Qijian Song | Theresa A Musket | Ruth Wagner | Pradeep Marri | Sam Reddy | Allen Sessions | Xiaolei Wu | David Grant | Philipp E Bayer | Manish Roorkiwal | Rajeev K Varshney | Xin Liu | David Edwards | Dong Xu | Trupti Joshi | Steven B Cannon | Henry T Nguyen
Scientific data | 2021

We report characteristics of soybean genetic diversity and structure from the resequencing of 481 diverse soybean accessions, comprising 52 wild (Glycine soja) selections and 429 cultivated (Glycine max) varieties (landraces and elites). This data was used to identify 7.8 million SNPs, to predict SNP effects relative to genic regions, and to identify the genetic structure, relationships, and linkage disequilibrium. We found evidence of distinct, mostly independent selection of lineages by particular geographic location. Among cultivated varieties, we identified numerous highly conserved regions, suggesting selection during domestication. Comparisons of these accessions against the whole U.S. germplasm genotyped with the SoySNP50K iSelect BeadChip revealed that over 95% of the re-sequenced accessions have a high similarity to their SoySNP50K counterparts. Probable errors in seed source or genotype tracking were also identified in approximately 5% of the accessions.

Pubmed ID: 33558550

Research resources used in this publication

None found

Antibodies used in this publication

None found

Associated grants

None

Publication data is provided by the National Library of Medicine ® and PubMed ®. Data is retrieved from PubMed ® on a weekly schedule. For terms and conditions see the National Library of Medicine Terms and Conditions.

This is a list of tools and resources that we have found mentioned in this publication.


PLINK (tool)

RRID:SCR_001757

Open source whole genome association analysis toolset, designed to perform range of basic, large scale analyses in computationally efficient manner. Used for analysis of genotype/phenotype data. Through integration with gPLINK and Haploview, there is some support for subsequent visualization, annotation and storage of results. PLINK 1.9 is improved and second generation of the software.

View all literature mentions

SnpEff (tool)

RRID:SCR_005191

Genetic variant annotation and effect prediction software toolbox that annotates and predicts effects of variants on genes (such as amino acid changes). By using standards, such as VCF, SnpEff makes it easy to integrate with other programs.

View all literature mentions

Phytozome (tool)

RRID:SCR_006507

A comparative platform for green plant genomics. Families of orthologous and paralogous genes that represent the modern descendents of ancestral gene sets are constructed at key phylogenetic nodes. These families allow easy access to clade specific orthology / paralogy relationships as well as clade specific genes and gene expansions. As of release v9.1, Phytozome provides access to forty-one sequenced and annotated green plant genomes which have been clustered into gene families at 20 evolutionarily significant nodes. Where possible, each gene has been annotated with PFAM, KOG, KEGG, and PANTHER assignments, and publicly available annotations from RefSeq, UniProt, TAIR, JGI are hyper-linked and searchable.

View all literature mentions

Picard (tool)

RRID:SCR_006525

Java toolset for working with next generation sequencing data in the BAM format.

View all literature mentions

FISHER (tool)

RRID:SCR_009181

THIS RESOURCE IS NO LONGER IN SERVICE, documented on February 1st, 2022. Software application for genetic analysis of classical biometric traits like blood pressure or height that are caused by a combination of polygenic inheritance and complex environmental forces. (entry from Genetic Analysis Software)

View all literature mentions

CyVerse (tool)

RRID:SCR_014531

A google drive interface for scientific big data. CyVerse cyberinfrastructure is applicable to all life sciences disciplines and works equally well on data from plants, animals, or microbes. It provides life scientists with computational infrastructure to handle large datasets and complex analyses. Its extensible platforms provide data storage, bioinformatics tools, image analyses, cloud services, and APIs.

View all literature mentions