Submitted:
23 July 2026
Posted:
23 July 2026
You are already at the latest version
Abstract
Keywords:
1. Introduction
2. Materials and Methods
2.1. Fungal Strains and Culture Conditions
2.2. Nucleic Acid Extraction and Quality Assessment
2.3. Library Construction and Sequencing Platform Integration
2.4. Genome Survey and Preliminary Characterization
2.5. Data Quality Control and T2T Assembly Pipeline
2.6. Assembly Quality Validation
2.7. Repetitive Element Annotation
2.8. Gene Prediction, Functional Annotation, and CAZyme Detection
2.9. Genome Visualization and Circos Plot Generation
2.10. Comparative Genomics Framework
2.11. Gene Family Evolution and Functional Enrichment
2.12. Synteny Analysis and Ks-Based Duplication Assessment
2.13. Data Visualization and Statistical Software
3. Results
3.1. Sequencing Data Generation and Coverage Statistics
3.2. T2T Assembly Quality, Structural Validation, and Telomere Integrity
3.3. Chromosome-Level Genomic Landscapes
3.4. Repetitive Element Landscapes
3.5. Gene Prediction, Functional Annotation, and CAZyme Repertoire
3.6. Comparative Genomics and Orthologous Group Identification
3.7. Phylogenetic Reconstruction, Divergence Time Estimation, and Gene-Family Evolution
3.8. GO and KEGG Enrichment of Expanded Gene Families
3.9. Genomic Synteny and Ks-Based Evolutionary Trajectory Detection
4. Discussion
5. Conclusions
Supplementary Materials
Author Contributions
Funding
Conflicts of Interest
References
- Li, H.; Durbin, R. Genome assembly in the telomere-to-telomere era. Nat. Rev. Genet. 2024, 25, 658–670. [Google Scholar] [CrossRef]
- Nurk, S.; et al. The complete sequence of a human genome. Science 2022, 376, 44–53. [Google Scholar] [CrossRef]
- Bowyer, P.; Currin, A.; Delneri, D.; Fraczek, M.G. Telomere-to-telomere genome sequence of the model mould pathogen Aspergillus fumigatus. Nat. Commun. 2022, 13, 5394. [Google Scholar] [CrossRef]
- Sonnenberg, A.S.M.; et al. Telomere-to-telomere assembled and centromere annotated genomes of the two main subspecies of the button mushroom Agaricus bisporus reveal especially polymorphic chromosome ends. Sci. Rep. 2020, 10, 14653. [Google Scholar] [CrossRef] [PubMed]
- Wang, M.; et al. Telomere-to-telomere genome assembly of Tibetan medicinal mushroom Ganoderma leucocontextum and the first Copia centromeric retrotransposon in macro-fungi genome. J. Fungi 2024, 10, 15. [Google Scholar] [CrossRef] [PubMed]
- Han, J.-N.; et al. Haplotype-resolved telomere-to-telomere genome assembly of the dikaryotic fungus pathogen Rhizoctonia cerealis. Sci. Data 2025, 12, 951. [Google Scholar] [CrossRef] [PubMed]
- Redhead, S.A.; Vilgalys, R.; Moncalvo, J.-M.; Johnson, J.; Hopple, J.S., Jr. Coprinus Pers. and the disposition of Coprinus species sensu lato. Taxon 2001, 50, 203–241. [Google Scholar] [CrossRef] [PubMed]
- Nagy, L.G.; Hazi, J.; Vagvolgyi, C.; Papp, T. Phylogeny and species delimitation in the genus Coprinellus with special emphasis on the haired species. Mycologia 2012, 104, 254–275. [Google Scholar] [CrossRef] [PubMed]
- Stajich, J.E.; et al. Insights into evolution of multicellular fungi from the assembled chromosomes of the mushroom Coprinopsis cinerea (Coprinus cinereus). Proc. Natl. Acad. Sci. USA 2010, 107, 11889–11894. [Google Scholar] [CrossRef] [PubMed]
- Fabros, R.J.P.; Dulay, R.M.R. Status review of the distribution, biological compounds, and bioactivities of Coprinellus mushrooms, and their medicinal and biotechnological prospects. Stud. Fungi 2025, 10, e0250024. [Google Scholar] [CrossRef]
- Marcais, G.; Kingsford, C. A fast, lock-free approach for efficient parallel counting of occurrences of k-mers. Bioinformatics 2011, 27, 764–770. [Google Scholar] [CrossRef] [PubMed]
- Ranallo-Benavidez, T.R.; Jaron, K.S.; Schatz, M.C. GenomeScope 2.0 and Smudgeplot for reference-free profiling of polyploid genomes. Nat. Commun. 2020, 11, 1432. [Google Scholar] [CrossRef] [PubMed]
- Chen, S.; et al. fastp: an ultra-fast all-in-one FASTQ preprocessor. Bioinformatics 2018, 34, i884–i890. [Google Scholar] [CrossRef] [PubMed]
- Krueger, F. Trim Galore: A wrapper tool around Cutadapt and FastQC to consistently apply quality and adapter trimming to FastQ files. Babraham Bioinformatics. 2015. Available online: https://github.com/FelixKrueger/TrimGalore.
- Wick, R.R. Filtlong: A tool for filtering long reads by quality. GitHub. 2017. Available online: https://github.com/rrwick/Filtlong.
- Chen, Y.; et al. Efficient assembly of nanopore reads via highly accurate and intact error correction. Nat. Commun. 2021, 12, 60. [Google Scholar] [CrossRef] [PubMed]
- Vaser, R.; et al. Fast and accurate de novo genome assembly from long uncorrected reads. Genome Res. 2017, 27, 737–746. [Google Scholar] [CrossRef] [PubMed]
- Li, H. Minimap2: pairwise alignment for nucleotide sequences. Bioinformatics 2018, 34, 3094–3100. [Google Scholar] [CrossRef] [PubMed]
- Vasimuddin, M.; et al. Efficient architecture-aware acceleration of BWA-MEM for multicore systems. In 2019 IEEE International Parallel and Distributed Processing Symposium (IPDPS); IEEE: Rio de Janeiro, Brazil, 2019; pp. 314–324. [Google Scholar] [CrossRef]
- Li, H.; et al. The Sequence Alignment/Map format and SAMtools. Bioinformatics 2009, 25, 2078–2079. [Google Scholar] [CrossRef] [PubMed]
- Walker, B.J.; et al. Pilon: an integrated tool for comprehensive microbial variant detection and assembly improvement. PLoS ONE 2014, 9, e112963. [Google Scholar] [CrossRef] [PubMed]
- Durand, N.C.; et al. Juicer provides a one-click system for analyzing loop-resolution Hi-C experiments. Cell Syst. 2016, 3, 95–98. [Google Scholar] [CrossRef] [PubMed]
- Dudchenko, O.; et al. De novo assembly of the Aedes aegypti genome using Hi-C yields chromosome-length scaffolds. Science 2017, 356, 92–95. [Google Scholar] [CrossRef] [PubMed]
- Dudchenko, O.; et al. The Juicebox Assembly Tools module facilitates de novo assembly of mammalian genomes with chromosome-length scaffolds for under $1000. bioRxiv 2018. [Google Scholar] [CrossRef]
- Simao, F.A.; et al. BUSCO: assessing genome assembly and annotation completeness with single-copy orthologs. Bioinformatics 2015, 31, 3210–3212. [Google Scholar] [CrossRef] [PubMed]
- Gurevich, A.; et al. QUAST: quality assessment tool for genome assemblies. Bioinformatics 2013, 29, 1072–1075. [Google Scholar] [CrossRef] [PubMed]
- Rhie, A.; et al. Merqury: reference-free quality, completeness, and phasing assessment for genome assemblies. Genome Biol. 2020, 21, 245. [Google Scholar] [CrossRef] [PubMed]
- Wolff, J.; et al. Galaxy HiCExplorer 3: a web server for reproducible Hi-C, capture Hi-C and single-cell Hi-C data analysis, quality control and visualization. Nucleic Acids Res. 2020, 48, W177–W184. [Google Scholar] [CrossRef] [PubMed]
- Flynn, J.M.; et al. RepeatModeler2 for automated genomic discovery of transposable element families. Proc. Natl. Acad. Sci. USA 2020, 117, 9451–9457. [Google Scholar] [CrossRef] [PubMed]
- Tarailo-Graovac, M.; Chen, N. Using RepeatMasker to identify repetitive elements in genomic sequences. Curr. Protoc. Bioinform. 2009, Chapter 4, 4.10.1–4.10.14. [Google Scholar] [CrossRef] [PubMed]
- Bao, W.; Kojima, K.K.; Kohany, O. Repbase Update, a database of repetitive elements in eukaryotic genomes. Mob. DNA 2015, 6, 11. [Google Scholar] [CrossRef] [PubMed]
- Kim, D.; et al. Graph-based genome alignment and genotyping with HISAT2 and HISAT-genotype. Nat. Biotechnol. 2019, 37, 907–915. [Google Scholar] [CrossRef] [PubMed]
- Hoff, K.J.; Lange, S.; Lomsadze, A.; Borodovsky, M.; Stanke, M. BRAKER1: Unsupervised RNA-Seq-Based Genome Annotation with GeneMark-ET and AUGUSTUS. Bioinformatics 2016, 32, 767–769. [Google Scholar] [CrossRef] [PubMed]
- Cantarel, B.L.; et al. MAKER: an easy-to-use annotation pipeline designed for emerging model organism genomes. Genome Res. 2008, 18, 188–196. [Google Scholar] [CrossRef] [PubMed]
- Bushmanova, E.; et al. rnaSPAdes: a de novo transcriptome assembler and its application to RNA-Seq data. GigaScience 2019, 8, giz100. [Google Scholar] [CrossRef] [PubMed]
- Stanke, M.; et al. AUGUSTUS: a web server for gene finding in eukaryotes. Nucleic Acids Res. 2004, 32, W309–W312. [Google Scholar] [CrossRef] [PubMed]
- Cantalapiedra, C.P.; et al. eggNOG-mapper v2: functional annotation, orthology assignments, and domain prediction at the metagenomic scale. Mol. Biol. Evol. 2021, 38, 5825–5829. [Google Scholar] [CrossRef] [PubMed]
- Zhang, H.; et al. dbCAN2: a meta server for automated carbohydrate-active enzyme annotation. Nucleic Acids Res. 2018, 46, W95–W101. [Google Scholar] [CrossRef] [PubMed]
- Eddy, S.R. Accelerated profile HMM searches. PLoS Comput. Biol. 2011, 7, e1002195. [Google Scholar] [CrossRef] [PubMed]
- Buchfink, B.; Reuter, K.; Drost, H.G. Sensitive protein alignments at tree-of-life scale using DIAMOND. Nat. Methods 2021, 18, 366–368. [Google Scholar] [CrossRef] [PubMed]
- Busk, P.K.; et al. Homology to peptide pattern for annotation of carbohydrate-active enzymes and prediction of function. BMC Bioinform. 2017, 18, 214. [Google Scholar] [CrossRef] [PubMed]
- Krzywinski, M.; et al. Circos: an information aesthetic for comparative genomics. Genome Res. 2009, 19, 1639–1645. [Google Scholar] [CrossRef] [PubMed]
- Quinlan, A.R.; Hall, I.M. BEDTools: a flexible suite of utilities for comparing genomic features. Bioinformatics 2010, 26, 841–842. [Google Scholar] [CrossRef] [PubMed]
- Emms, D.M.; Kelly, S. OrthoFinder: phylogenetic orthology inference for comparative genomics. Genome Biol. 2019, 20, 238. [Google Scholar] [CrossRef] [PubMed]
- Katoh, K.; et al. MAFFT: a novel method for rapid multiple sequence alignment based on fast Fourier transform. Nucleic Acids Res. 2002, 30, 3059–3066. [Google Scholar] [CrossRef] [PubMed]
- Stamatakis, A. RAxML version 8: a tool for phylogenetic analysis and post-analysis of large phylogenies. Bioinformatics 2014, 30, 1312–1313. [Google Scholar] [CrossRef] [PubMed]
- Zhang, C.; et al. ASTRAL-III: polynomial time species tree reconstruction from partially resolved gene trees. BMC Bioinform. 2018, 19, 153. [Google Scholar] [CrossRef] [PubMed]
- Yang, Z. PAML 4: phylogenetic analysis by maximum likelihood. Mol. Biol. Evol. 2007, 24, 1586–1591. [Google Scholar] [CrossRef] [PubMed]
- Kumar, S.; et al. TimeTree 5: an expanded resource for species divergence times. Mol. Biol. Evol. 2022, 39, msac174. [Google Scholar] [CrossRef] [PubMed]
- De Bie, T.; Cristianini, N.; Demuth, J.P.; Hahn, M.W. CAFE: a computational tool for the study of gene family evolution. Bioinformatics 2006, 22, 1269–1271. [Google Scholar] [CrossRef] [PubMed]
- Wu, T.; et al. clusterProfiler 4.0: a universal enrichment tool for interpreting omics data. Innovation 2021, 2, 100141. [Google Scholar] [CrossRef] [PubMed]
- Tang, H.; et al. JCVI: a versatile toolkit for comparative genomics analysis. iMeta 2024, 3, e211. [Google Scholar] [CrossRef] [PubMed]
- Zwaenepoel, A.; Van de Peer, Y. wgd: simple command line tools for the analysis of ancient whole-genome duplications. Bioinformatics 2019, 35, 2153–2155. [Google Scholar] [CrossRef] [PubMed]
- Zachos, J.C.; Dickens, G.R.; Zeebe, R.E. An early Cenozoic perspective on greenhouse warming and carbon-cycle dynamics. Nature 2008, 451, 279–283. [Google Scholar] [CrossRef] [PubMed]
- Westerhold, T.; et al. An astronomically dated record of Earth’s climate and its predictability over the last 66 million years. Science 2020, 369, 1383–1387. [Google Scholar] [CrossRef] [PubMed]
- McInerney, F.A.; Wing, S.L. The Paleocene-Eocene Thermal Maximum: A perturbation of carbon cycle, climate, and biosphere with implications for the future. Annu. Rev. Earth Planet. Sci. 2011, 39, 489–516. [Google Scholar] [CrossRef]
- Jaramillo, C.; et al. Effects of rapid global warming at the Paleocene-Eocene boundary on Neotropical vegetation. Science 2010, 330, 957–961. [Google Scholar] [CrossRef] [PubMed]







| Species | Data Type | TotalBase of Raw Data (G) | TotalBase of Clean Data (G) | TotalReads of Raw Data (M) | TotalReads of Clean Data (M) | Coverage (X) |
|---|---|---|---|---|---|---|
| C. xanthothrix | Short-read WGS | 15.22 | 15.12 | 101.48 | 101.47 | 324X |
| C. xanthothrix | Long-read | 10.76 | 10.38 | 1.77 | 1.62 | 228X |
| C. xanthothrix | Hi-C | 10.85 | 10.84 | 36.18 | 36.18 | 231X |
| C. xanthothrix | RNA-seq | 5.49 | / | 0.11 | / | / |
| C. saccharinus | Short-read WGS | 10.28 | 10.03 | 68.52 | 67.92 | 187X |
| C. saccharinus | Long-read | 12.57 | 11.96 | 0.54 | 0.50 | 229X |
| C. saccharinus | Hi-C | 19.11 | 19.09 | 63.69 | 63.69 | 349X |
| C. saccharinus | RNA-seq | 9.12 | / | 0.19 | / | / |
| Feature | C. xanthothrix | C. saccharinus |
|---|---|---|
| Assembly level | T2T | T2T |
| Total genome size (Mb) | 46.92 | 54.81 |
| Chromosome number | 13 | 13 |
| Largest chromosome (Mb) | 5.08 (Chr01) | 7.31 (Chr01) |
| Smallest chromosome (Mb) | 2.22 (Chr13) | 0.84 (Chr13) |
| Gap number | 0 | 0 |
| GC content (%) | 53.72% | 53.79% |
| N50 | 4017987 | 5059299 |
| N90 | 2528967 | 3343676 |
| L50 | 6 | 5 |
| L90 | 11 | 11 |
| QV | 38.36 | 39.26 |
| Telomere motif | AACCCT | AACCCT |
| Telomere-capped ends | 26/26 (100%) | 26/26 (100%) |
| BUSCO (fungi_odb10) | 99.20% | 99.10% |
| NCBI/GSA Accession | PRJNA1450334 / PRJCA038423 | PRJNA1450334 / PRJCA038423 |
| Species | Element Type | Number of elements | Length (bp) | Percentage of sequence (%) |
|---|---|---|---|---|
| C. xanthothrix | Retroelements | 1286 | 1649099 | 3.51 |
| DNA transposons | 142 | 160597 | 0.34 | |
| Rolling-circles | 101 | 106931 | 0.23 | |
| Unclassified | 3806 | 1925552 | 4.10 | |
| Total interspersed repeats | / | 3735248 | 7.96 | |
| Small RNA | 41 | 11155 | 0.02 | |
| Simple repeats | 4609 | 220747 | 0.47 | |
| Low complexity | 1037 | 58988 | 0.13 | |
| C. saccharinus | Retroelements | 2757 | 4424077 | 8.07 |
| DNA transposons | 181 | 160697 | 0.29 | |
| Rolling-circles | 13 | 28761 | 0.05 | |
| Unclassified | 4803 | 2033022 | 3.71 | |
| Total interspersed repeats | / | 6617796 | 12.07 | |
| Small RNA | 127 | 36128 | 0.07 | |
| Simple repeats | 5741 | 255121 | 0.47 | |
| Low complexity | 1322 | 73680 | 0.13 |
| Species | Category | Count | Percentage (%) |
|---|---|---|---|
| C. xanthothrix | Total Genes | 11205 | 100.00 |
| GO Annotated | 3645 | 32.53 | |
| KEGG KO Annotated | 4563 | 40.72 | |
| KEGG Pathway Annotated | 2862 | 25.54 | |
| COG Annotated | 8325 | 74.30 | |
| Unique GO Terms | 10173 | / | |
| Unique KEGG KOs | 3298 | / | |
| Unique KEGG Pathways | 750 | / | |
| Unique COG Categories | 24 | / | |
| C. saccharinus | Total Genes | 11881 | 100.00 |
| GO Annotated | 3690 | 31.06 | |
| KEGG KO Annotated | 4717 | 39.70 | |
| KEGG Pathway Annotated | 2929 | 24.65 | |
| COG Annotated | 8733 | 73.50 | |
| Unique GO Terms | 10220 | / | |
| Unique KEGG KOs | 3318 | / | |
| Unique KEGG Pathways | 752 | / | |
| Unique COG Categories | 24 | / |
| Species | Genes | Species-specific | Single-copy |
|---|---|---|---|
| A. bisporus | 10140 | 1546 | 4742 |
| C. saccharinus | 15802 | 109 | 6662 |
| C. xanthothrix | 13046 | 45 | 5702 |
| C. cinerea | 11869 | 687 | 5422 |
| C.disseminatus | 13929 | 180 | 5757 |
| C. domesticus | 11784 | 32 | 5718 |
| C. marcescibilis | 13294 | 775 | 5404 |
| C. micaceus | 20758 | 1068 | 5977 |
| C. radians | 12332 | 32 | 5713 |
| P. aberdarensis | 13952 | 913 | 4614 |
| S. cerevisiae | 4399 | 659 | 2636 |
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
© 2026 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license (http://creativecommons.org/licenses/by/4.0/).