{"@type": "dcat:Dataset", "accessLevel": "public", "bureauCode": ["005:18"], "contactPoint": {"fn": "Coyne, Clarice", "hasEmail": "mailto:Clarice.Coyne@usda.gov"}, "description": "<p>Included in this dataset are SNP and fasta data for the Pea Single Plant Plus Collection (PSPPC) and the PSPPC augmented with 25 <em>P. fulvum</em> accessions. </p>\n<p>These 6 datasets can be roughly divided into two groups. Group 1 consists of three datasets labeled PSPPC which refer to SNP data pertaining to the USDA Pea Single Plant Plus Collection. Group 2 consists of three datasets labeled PSPPC + <em>P. fulvum</em> which refer to SNP data pertaining to the USDA PSPPC with 25 accessions of <em>Pisum fulvum</em> added. SNPs for each of these groups were called independently; therefore SNP names that are shared between the PSPPC and PSPPC + <em>P. fulvum</em> groups should NOT be assumed to refer to the same locus.</p>\n<p>For analysis, SNP data is available in two widely used formats: hapmap and vcf. These formats can be successfully loaded into TASSEL v. 5.2.25 (<a href=\"http://www.maizegenetics.net/tassel\">http://www.maizegenetics.net/tassel</a>). Explanations of fields (columns) in the VCF files are contained within commented (##) rows at the top of the file. </p>\n<p>Descriptions of the first 11 columns in the hapmap file are as follows:</p>\n<ul>\n<li>rs#- Name of locus (i.e. SNP name)</li>\n<li>alleles- Indicates the SNPs for each allele at the locus</li>\n<li>chrom- Irrelevant for these datasets, since markers are unordered.</li>\n<li>pos- Irrelevant for these datasets, since markers are unordered.</li>\n<li>strand- Irrelevant for these datasets, since markers are unordered</li>\n<li>assembly#- required field for hapmap format. NA for these datasets</li>\n<li>center- required field for hapmap format. NA for these datasets</li>\n<li>protLSID- required field for hapmap format. NA for these datasets</li>\n<li>assayLSID- required field for hapmap format. NA for these datasets</li>\n<li>panel- required field for hapmap format. NA for these datasets</li>\n<li>QCcode- required field for hapmap format. NA for these datasets</li>\n</ul>\n<p>The fasta sequences containing the SNPs are also available for such downstream applications as development of primers for platform-specific markers.</p>\n<p>For more information about this dataset, contact Clarice Coyne at Clarice.Coyne@usda.gov or coynec@wsu.edu. </p><div><br>Resources in this dataset:</div><br><ul><li><p>Resource Title: PSPPC SNPs in hapmap format.</p> <p>File Name: PSPPC.hmp<em>.txt</em></p><p><em>Resource Description: 66591 unanchored SNPs for the PSPPC collection in hapmap format</em></p><p><em>Resource Software Recommended: TASSEL,url: <a href=\"http://www.maizegenetics.net/tassel\">http://www.maizegenetics.net/tassel</a> </em></p></li><em><br></em><li><em><p>Resource Title: PSPPC SNP FASTA Sequences.</p> </em><p><em>File Name: PSPPC.fa</em>.txt</p><p>Resource Description: FASTA sequences for each allele of the PSPPC SNP dataset</p></li><br><li><p>Resource Title: PPSPPC + P. fulvum SNPs in hapmap format.</p> <p>File Name: PSPPC+fulvums.hmp<em>.txt</em></p><p><em>Resource Description: 67400 SNPs from the PSPPC augmented with 25 P. fulvum accessions in hapmap format. SNP names are independent and unrelated to plain PSPPC SNP files.</em></p><p><em>Resource Software Recommended: TASSEL,url: <a href=\"http://www.maizegenetics.net/tassel\">http://www.maizegenetics.net/tassel</a> </em></p></li><em><br></em><li><em><p>Resource Title: PSPPC + P. fulvum SNP FASTA Sequences.</p> </em><p><em>File Name: PSPPC+fulvums.fa</em>.txt</p><p>Resource Description: FASTA sequences for each allele of the PSPPC + P. fulvum SNP dataset. SNP names are independent and unrelated to plain PSPPC SNP files.</p></li><br><li><p>Resource Title: PSPPC + P. fulvum SNPs in vcf format.</p> <p>File Name: PSPPC+fulvums.vcf<em>.txt</em></p><p><em>Resource Description: 67400 SNPs from the PSPPC augmented with 25 P. fulvum accessions in vcf format. SNP names are independent and unrelated to plain PSPPC SNP files.</em></p><p><em>Resource Software Recommended: TASSEL,url: <a href=\"http://www.maizegenetics.net/tassel\">http://www.maizegenetics.net/tassel</a> </em></p></li><em><br></em><li><em><p>Resource Title: PSPPC SNPs in vcf format.</p> </em><p><em>File Name: PSPPC.vcf</em>.txt</p><p>Resource Description: 66591 SNPs from the PSPPC in vcf format</p><p>Resource Software Recommended: TASSEL,url: <a href=\"http://www.maizegenetics.net/tassel\">http://www.maizegenetics.net/tassel</a> </p></li><br><li><p>Resource Title: README.</p> <p>File Name: Data Dictionary.docx</p><p>Resource Description: These data are for the Pea Single Plant Plus Collection (PSPPC) and the PSPPC augmented with 25 <em>P. fulvum </em>accessions.</p>\n<p>The 6 datasets can be divided into two groups. Group 1 consists of 3 datasets labeled \u201cPSPPC\u201d which refer to SNP data pertaining to the USDA Pea Single Plant Plus Collection. Group 2 consists of 3 datasets labeled \u201cPSPPC + <em>P. fulvum</em>\u201d which refer to SNP data pertaining to the PSPPC with 25 accessions of <em>Pisum fulvum </em>added. SNPs for each of these groups were called independently; therefore any SNP name that is shared between the PSPPC and PSPPC + <em>P. fulvum </em>groups should NOT be assumed to refer to the same locus.</p>\n<p>For analysis, SNP data is available in two widely used formats: hapmap and vcf. These files were successfully loaded into the standalone version of TASSEL v. 5.2.25 (<a href=\"http://www.maizegenetics.net/tassel\">http://www.maizegenetics.net/tassel</a>). </p>\n<p>Explanations of fields (columns) in the VCF files are contained within commented (##) rows at the top of the file. </p>\n<p>The first 11 columns required for the hapmap format are as follows:\nrs#- Name of locus (i.e. SNP name)\nalleles- Indicates the SNPs for each allele at the locus\nchrom- N/A, since markers are unordered.\npos- N/A, since markers are unordered.\nstrand- N/A, since markers are unordered\nassembly#- N/A\ncenter- N/A\nprotLSID- N/A\nassayLSID- N/A\npanel- N/A\nQCcode- N/A</p>\n<p>The fasta sequences containing the SNPs are also available here for such downstream applications as development of primers for platform-specific markers.\n</p></li></ul><p></p>", "distribution": [{"@type": "dcat:Distribution", "downloadURL": "https://ndownloader.figshare.com/files/44335712", "format": "txt", "mediaType": "text/plain", "title": "PSPPC.hmp_.txt"}, {"@type": "dcat:Distribution", "downloadURL": "https://ndownloader.figshare.com/files/44335715", "format": "txt", "mediaType": "text/plain", "title": "PSPPC.vcf__0.txt"}, {"@type": "dcat:Distribution", "downloadURL": "https://ndownloader.figshare.com/files/44335718", "format": "txt", "mediaType": "text/plain", "title": "PSPPC.fa_.txt"}, {"@type": "dcat:Distribution", "downloadURL": "https://ndownloader.figshare.com/files/44335721", "format": "txt", "mediaType": "text/plain", "title": "PSPPC+fulvums.vcf__0.txt"}, {"@type": "dcat:Distribution", "downloadURL": "https://ndownloader.figshare.com/files/44335736", "format": "txt", "mediaType": "text/plain", "title": "PSPPC+fulvums.hmp_.txt"}, {"@type": "dcat:Distribution", "downloadURL": "https://ndownloader.figshare.com/files/44335739", "format": "txt", "mediaType": "text/plain", "title": "PSPPC+fulvums.fa_.txt"}, {"@type": "dcat:Distribution", "downloadURL": "https://ndownloader.figshare.com/files/44335742", "format": "docx", "mediaType": "application/vnd.openxmlformats-officedocument.wordprocessingml.document", "title": "Data Dictionary.docx"}], "identifier": "10.15482/USDA.ADC/1347137", "keyword": ["ARS", "NP301", "data.gov"], "license": "https://www.usa.gov/publicdomain/label/1.0/", "modified": "2025-11-21", "programCode": ["005:040"], "publisher": {"@type": "org:Organization", "name": "Agricultural Research Service"}, "spatial": "{\"type\": \"Polygon\", \"coordinates\": [[[-166.640625, -59.987997631212], [-166.640625, 83.254516804633], [194.765625, 83.254516804633], [194.765625, -59.987997631212], [-166.640625, -59.987997631212]]]}", "temporal": "2013-01-01/2014-12-31", "title": "Data from: A Community Resource for Exploring and Utilizing Genetic Diversity in the USDA Pea Single Plant Plus Collection"}