Juan F Medrano, Dario Cantu, Andrea Minio, Christian Dreischer, Theodore Gibbons, Jason Chin, Shiyu Chen, Allen Van Deynze, Amanda M Hulse-Kemp
{"title":"De novo whole-genome assembly and annotation of Coffea arabica var. Geisha, a high-quality coffee variety from the primary origin of coffee.","authors":"Juan F Medrano, Dario Cantu, Andrea Minio, Christian Dreischer, Theodore Gibbons, Jason Chin, Shiyu Chen, Allen Van Deynze, Amanda M Hulse-Kemp","doi":"10.1093/g3journal/jkae262","DOIUrl":null,"url":null,"abstract":"<p><p>Geisha coffee is recognized for its unique aromas and flavors and accordingly, has achieved the highest prices in the specialty coffee markets. We report the development of a chromosome-level, well-annotated, genome assembly of Coffea arabica var. Geisha. Geisha is considered an Ethiopian landrace that represents germplasm from the Ethiopian center of origin of coffee. We used a hybrid de novo assembly approach combining two long-reads single molecule sequencing technologies, Oxford Nanopore and Pacific Biosciences, together with scaffolding with Hi-C libraries. The final assembly is 1.03GB in size with BUSCO assessment of the assembly completeness of 97.7% of single-copy orthologs clusters. RNAseq and IsoSeq data were used as transcriptional experimental evidence for annotation and gene prediction revealing the presence of 47,062 gene loci encompassing 53,273 protein-coding transcripts. Comparison of the assembly to the progenitor subgenomes separated the set of chromosome sequences inherited from C. canephora from those of C. eugenioides. Corresponding orthologs between the two Arabica varieties, Geisha and Red Bourbon, had a 99.67% median identity, higher than what we observe with the progenitor assemblies (median 97.28%). Both Geisha and Red Bourbon contain a recombination event on Chromosome 10 relative to the two progenitors that must have happened before the geographical separation of the two varieties, consistent with a single allopolyploidization event giving rise to C. arabica. Broadening the availability of high-quality genome assemblies of Coffea arabica varieties, paves the way for understanding the evolution and domestication of coffee, as well as the genetic basis and environmental interactions of why a variety like Geisha is capable of producing beans with such exceptional and unique high-quality.</p>","PeriodicalId":12468,"journal":{"name":"G3: Genes|Genomes|Genetics","volume":" ","pages":""},"PeriodicalIF":2.1000,"publicationDate":"2024-11-15","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"G3: Genes|Genomes|Genetics","FirstCategoryId":"99","ListUrlMain":"https://doi.org/10.1093/g3journal/jkae262","RegionNum":3,"RegionCategory":"生物学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q3","JCRName":"GENETICS & HEREDITY","Score":null,"Total":0}
引用次数: 0
Abstract
Geisha coffee is recognized for its unique aromas and flavors and accordingly, has achieved the highest prices in the specialty coffee markets. We report the development of a chromosome-level, well-annotated, genome assembly of Coffea arabica var. Geisha. Geisha is considered an Ethiopian landrace that represents germplasm from the Ethiopian center of origin of coffee. We used a hybrid de novo assembly approach combining two long-reads single molecule sequencing technologies, Oxford Nanopore and Pacific Biosciences, together with scaffolding with Hi-C libraries. The final assembly is 1.03GB in size with BUSCO assessment of the assembly completeness of 97.7% of single-copy orthologs clusters. RNAseq and IsoSeq data were used as transcriptional experimental evidence for annotation and gene prediction revealing the presence of 47,062 gene loci encompassing 53,273 protein-coding transcripts. Comparison of the assembly to the progenitor subgenomes separated the set of chromosome sequences inherited from C. canephora from those of C. eugenioides. Corresponding orthologs between the two Arabica varieties, Geisha and Red Bourbon, had a 99.67% median identity, higher than what we observe with the progenitor assemblies (median 97.28%). Both Geisha and Red Bourbon contain a recombination event on Chromosome 10 relative to the two progenitors that must have happened before the geographical separation of the two varieties, consistent with a single allopolyploidization event giving rise to C. arabica. Broadening the availability of high-quality genome assemblies of Coffea arabica varieties, paves the way for understanding the evolution and domestication of coffee, as well as the genetic basis and environmental interactions of why a variety like Geisha is capable of producing beans with such exceptional and unique high-quality.
期刊介绍:
G3: Genes, Genomes, Genetics provides a forum for the publication of high‐quality foundational research, particularly research that generates useful genetic and genomic information such as genome maps, single gene studies, genome‐wide association and QTL studies, as well as genome reports, mutant screens, and advances in methods and technology. The Editorial Board of G3 believes that rapid dissemination of these data is the necessary foundation for analysis that leads to mechanistic insights.
G3, published by the Genetics Society of America, meets the critical and growing need of the genetics community for rapid review and publication of important results in all areas of genetics. G3 offers the opportunity to publish the puzzling finding or to present unpublished results that may not have been submitted for review and publication due to a perceived lack of a potential high-impact finding. G3 has earned the DOAJ Seal, which is a mark of certification for open access journals, awarded by DOAJ to journals that achieve a high level of openness, adhere to Best Practice and high publishing standards.