U.S. flag

An official website of the United States government

i5k Workspace

About the i5k Workspace@NAL

The i5k Workspace (https://i5k.nal.usda.gov) is an inclusive genome portal for any arthropod genome project that would like to make use of our resources. We provide download services, BLAST, the JBrowse genome browser, and the Apollo manual curation service. Over 50 arthropod genomes are now part of the i5k Workspace, and users are encouraged to browse the genomes that we host, and contribute to the curation of each genome. For more information about the i5k Workspace, you can read our paper on the i5k Workspace, view our posters and talks, and find our software projects on github. The Ag Data Commons is now hosting a growing number of i5k Workspace datasets.

About the i5k initiative

The i5k initiative is a transformative project that aims to sequence and analyze the genomes of 5,000 arthropod species. The National Agricultural Library has partnered with the i5k initiative to create the i5k Workspace@NAL, which serves any ‘orphaned’ arthropod genome project's hosting needs. For more information about the i5k initiative, read the paper and visit the website.

i5k Datasets

25 datasets

Halyomorpha halys Official Gene Set v1.0

    This dataset presents the *Halyomorpha halys* Official Gene Set (OGS) v1.0. The OGS is an integration of automatic gene predictions from NCBI's eukaryotic annotation pipeline, [NCBI Halyomorpha halys Annotation Release 100](https://www.ncbi.nlm.nih.gov/genome/annotation_euk/Halyomorpha_halys/100/), with manual annotations by the research community (performed via the Apollo manual curation software, http://genomearchitect.org/).

    Frankliniella occidentalis Official Gene Set OGSv1.0

      The *Frankliniella occidentalis* genome was recently sequenced and annotated as part of the i5k pilot project by the Baylor College of Medicine. The *Frankliniella occidentalis* research community has manually reviewed and curated the computational gene predictions and generated an official gene set, OGSv1.0. OGSv1.0 was generated by merging gene set FOCC-V0.5.3-Models generated by the Baylor College of Medicine, and community-curated models in the Apollo software, after QC of the Apollo output. After the merge, scaffolds that were likely bacterial contamination were identified by John H. Werren, and gene models overlapping with these contaminated regions were removed from the OGS.

      Blattella germanica Official Gene Set OGSv1.0

        The *Blattella germanica* genome was recently sequenced and annotated as part of the i5k pilot project by the Baylor College of Medicine. The *Blattella germanica* research community has manually reviewed and curated the computational gene predictions and generated an official gene set, OGSv1.0.

        Agrilus planipennis genome annotations v0.5.3

          This dataset presents the Agrilus planipennis gene set BCM_v_0.5.3. RNA-Seq data was used with additional protein homology data for a MAKER automated annotation of the Agrilus planipennis genome assembly 1.0. This dataset is free for all use.

          Pachypsylla venusta genome annotations v0.5.3

            This dataset presents the Pachypsylla venusta gene set BCM_v_0.5.3. RNA-Seq data was used with additional protein homology data for a MAKER automated annotation of the Pachypsylla venusta genome assembly 1.0.

            Anoplophora glabripennis Official Gene Set OGSv1.2

              The *Anoplophora glabripennis* genome was recently sequenced, assembled and annotated as part of the i5k pilot project by the Baylor College of Medicine, in collaboration with the McKenna Laboratory at the University of Memphis. The *Anoplophora glabripennis* research community has manually reviewed and curated the computational gene predictions and generated an official gene set, OGSv1.2. OGSv1.2 was generated by merging gene set AGLA-c0.5.3-Models generated by the Baylor College of Medicine, and community-curated models in the Apollo software, after QC of the Apollo output.

              Athalia rosae genome annotations v0.5.3

                This dataset presents the Athalia rosae gene set BCM_v_0.5.3. RNA-Seq data was used with additional protein homology data for a MAKER automated annotation of the Athalia rosae genome assembly 1.0.

                Blattella germanica genome annotations v0.5.3

                  This dataset presents the Blattella germanica gene set BCM_v_0.5.3. RNA-Seq data was used with additional protein homology data for a MAKER automated annotation of the Blattella germanica genome assembly 1.0. This dataset is free for all use.

                  Catajapyx aquilonaris genome annotations v0.5.3

                    This dataset presents the Catajapyx aquilonaris gene set BCM_v_0.5.3. RNA-Seq data was used with additional protein homology data for a MAKER automated annotation of the Catajapyx aquilonaris genome assembly 1.0.This dataset is free for all use.

                    Centruroides sculpturatus genome annotations v0.5.3

                      This dataset presents the Centruroides sculpturatus gene set BCM_v_0.5.3. RNA-Seq data was used with additional protein homology data for a MAKER automated annotation of the Centruroides sculpturatus genome assembly 1.0. This dataset is free for all use.