Abstract
Shiraia bambusicola has long been used as a traditional Chinese medicine and its major medicinal active metabolite is hypocrellin, which exhibits outstanding antiviral and antitumor properties. Here we report the 32 Mb draft genome sequence of S. bambusicola S4201, encoding 11,332 predicted genes. The genome of S. bambusicola is enriched in carbohydrate-active enzymes (CAZy) and pathogenesis-related genes. The phylogenetic tree of S. bambusicola S4201 and nine other sequenced species was constructed and its taxonomic status was supported (Pleosporales, Dothideomycetes). The genome contains a rich set of secondary metabolite biosynthetic gene clusters, suggesting that strain S4201 has a remarkable capacity to produce secondary metabolites. Overexpression of the zinc finger transcription factor zftf, which is involved in hypocrellin A (HA) biosynthesis, increases HA production when compared with wild type. In addition, a new putative HA biosynthetic pathway is proposed. These results provide a framework to study the mechanisms of infection in bamboo and to understand the phylogenetic relationships of S. bambusicola S4201. At the same time, knowledge of the genome sequence may potentially solve the puzzle of HA biosynthesis and lead to the discovery of novel genes and secondary metabolites of importance in medicine and agriculture.
Introduction
Shiraia bambusicola, belonging to the genus Shiraia and the phylum Ascomycota, is an important pathogenic and parasitic fungus that grows on the twigs of bamboos. It adversely affects the growth of bamboo and causes significant economic losses every year. S. bambusicola is widely distributed in many provinces of southern China and Japan (; ). The fruiting body of S. bambusicola has been used as a traditional Chinese medicine for the treatment of sciatica, pertussis, tracheitis, and rheumatic arthritis (). On the other hand, S. bambusicola is known as the main hyprocrellin-producing species.
The hyprocrellin family includes hypocrellin A (HA), hypocrellin B (HB), hypocrellin C (HC), and hypocrellin D (HD), which are isolated from the parasitic fungi S. bambusicola () and Hypocrella bambusae (). Resembling many other natural perylenequinones (Figure 1), such as cercosporin and elsinochrome, the hypocrellin family is characterized by a helical chiral pentacyclic conjugated core structure combined with C7, C7′-substitutions, and possessing centrochiral stereochemistry (). Because of its high potential pharmacological value, HA has attracted much attention. It has been shown to possess antibacterial, antiviral, antitumor, and anti-inflammatory bioactivities (; ). As a photosensitizer, it can be activated under light irradiation and can produce reactive oxygen species (ROS), which can destroy DNA, proteins, and lipids (). Many researchers are studying the biological characteristics and fermentation production processes of S. bambusicola, as well as the applicability of photodynamic therapy of HA. Although many pharmacological effects of HA are being investigated, HA biosynthetic genes and pathways are still not well understood.
FIGURE 1
Fungi have specialized in using plant biomass as a carbon source by producing enzymes that degrade cell wall polysaccharides into metabolizable sugars for nutrition (; ). Carbohydrate-active enzymes (CAZy) are enzymes involved in the synthesis, metabolism, modification, and transport of carbohydrates. It is believed that the number of CAZymes in a fungal species correlates with its life strategies and the nutritional availability (). The Pathogen–Host Interaction database (PHI-base) contains expertly curated molecular and biological information on genes proven to affect the outcome of pathogen-host interactions. Cytochrome P450s (CYPs) are involved in the synthesis of toxins and pathogenesis (). The S. bambusicola S4201 genome data were entered in the databases mentioned above and we found some genes in S. bambusicola that may be involved in pathogenesis, providing a foundation for future studies of its lifestyle.
The taxonomic status of S. bambusicola has been reclassified several times over the past one hundred years. The genus Shiraia was initially classified under Nectriaceae, Hypocreales, and Pyrenomycetes, but was later transferred to the Hypocreaceae family based on the larger fleshy stroma. This taxonomic treatment was popular for several decades, but then Shiraia was transferred to the order Pleosporales based on its bitunicate, as opposed to unitunicate asci. According to the ninth edition of the fungal dictionary, Shiraia was classified under Dothideales incertae sedis (). Previous taxonomic classifications of Shiraia were mainly based on morphological characteristics of the ascostromata, ascus, ascospores, etc. However, several traditional morphological features are not class unique and DNA sequence comparisons are important to define the class. Recent studies sequencing the 18S rDNA and ITS-5.8S rDNA regions indicated that Shiraia should be classified under the order Pleosporales (). More recently, a new family, Shiraiaceae, has been proposed with Shiraia as the type genus. Its epitype was redescribed based on morphological characters and partial LSU-rDNA, EF, and RPB gene sequencing data, suggesting that it is a new family of Pleosporales (). With the availability of more microbial genome sequencing data, researchers have begun to study phylogeny at the genomic level. However, the phylogeny of S. bambusicola has not been analyzed based on data at the genomic level. A phylogenetic analysis based on a small number of concatenated genes in any genome has a high probability of supporting conflicting topologies, while analysis of the entire data set of concatenated genes may provide a single, fully resolved species tree with maximum support ().
Fungal genomics and comparative analysis of different genomes are helpful in understanding phylogenetic and evolutionary relationships, gene function, lifestyle and strategies, biocontrol mechanisms, pathogenicity, and the development and utilization of secondary metabolites. Given the importance of these studies, we sequenced the S. bambusicola S4201 genome with Illumina HiSeq sequencing technology combined with PacBio single molecular long-read sequencing. A rich repertoire of CAZymes and pathogenesis-related genes were discovered. The taxonomic position of S. bambusicola was reconfirmed based on genomic data. A large set of secondary metabolic biosynthetic gene clusters and core genes were identified. Furthermore, we identified the HA gene cluster and proposed a new putative biosynthetic pathway. The information contained in this study could be useful for understanding the pathogenicity, taxonomic status, and diversity of the secondary metabolites of S. bambusicola S4201.
Materials and Methods
Strain, DNA, and RNA Isolation
Shiraia bambusicola S4201 was isolated in China from the fruiting bodies of S. bambusicola P. Henn, a strain which has shown excellent HA production and which has been used for sequencing (). The strain S4201 was incubated in potato dextrose broth (PDB, 200 g/L potato extract, 20 g/L glucose, pH 7.0) medium for 120 h at 28°C under agitation (150 r/min). The fermentation liquid was centrifuged at 3000 r/min for 5 min at room temperature and the supernatant was removed. The mycelium was washed with sterile water and frozen in liquid nitrogen. Genomic DNA was isolated using the cetyl trimethyl ammonium bromide (CTAB) method (). The quality of the total DNA was confirmed by means of a Qubit Fluorometer, NanoDrop spectrophotometer, and by agarose gel electrophoresis before further processing.
Genome Sequencing and Assembly
The S. bambusicola S4201 genome sequences included short reads from the Illumina library and long reads from a PacBio single-molecule library at the Beijing Genomics Institute (BGI) in Wuhan, China. A library with 410 bp inserts was constructed and sequenced using an Illumina Hiseq 4000 platform. To obtain clean data, 20 low-quality bases, 10% Ns, 15 bp overlap between adapter and duplications in the raw Illumina sequencing data were filtered. Total DNA (10 μg) was used to construct a 20 kb DNA library for sequencing with the PacBio platform. The PacBio data were also cleaned by removing adapter sequences, low-quality polymerase reads, and by discarding trimmed reads with lengths less than 1000 bp. The abundance of 15-mers was measured to obtain a preliminary estimate of the genome size, heterozygosity, and repetitive sequences information. Sequences were assembled using various assembly software and according to the following steps: (i) correcting the subreads to obtain the corrected reads (Proovread 2.12, -t 4 –coverage 60 –mode sr); (ii) assembling the corrected reads (Celera Assembler 8.3, doTrim_initialQualityBased = 1, doTrim_finalEvidenceBased = 1, doRemoveSpurReads = 1, doRemoveChimericReads = 1, -d properties -U); (iii) correcting the data using Illumina short reads (GATK v1.6-13, -cluster 2 -window 5 -stand_call_conf 50 -stand_emit_conf 10.0 -dcov 200 MQ0 ≥ 4) (; ; ; ; ).
Gene Prediction and Annotation
The gene components were predicted by using Genewise (), SNAP (), Genemark-ES (), and Augustus () software. rRNA was aligned to an rRNA database or predicted by means of RNAmmer () software. tRNA and tRNA secondary structures were predicted with tRNA scan () and sRNA was matched with the Rfam () database by using Infernal software. RepeatMasker and RepeatProteinMasker software were employed to examine the repeats present in S. bambusicola S4201. BuildXDFDatabase, RepeatModeler, and Repeatmasker were used for de novo detection. The tandem repeat sequences were identified using Tandem Repeat Finder () software. The protein-encoding genes were annotated through BLASTp searches in the Cluster of Orthologous Groups of proteins (COG) (; ), Gene Ontology (GO) (), Kyoto Encyclopedia of Genes and Genomes (KEGG) (), Non-Redundant (NR) Protein Database, SwissProt Databases and the best hits were filtered (E-value < 0.00001). CAZymes database () was used to identify proteins involved in carbohydrate metabolism. Pathogenicity and virulence-associated genes were predicted by means of the PHI database (). CYPs were annotated using the Fungal CYP Database ().
Comparative Genomics and Phylogenetic Tree
Since S. bambusicola S4201 is a single genus and species, we searched for several closely related species (Leptosphaeria maculans, Parastagonospora nodorum, Paraphaeosphaeria sporulosa, Cercospora zeina, Stagonospora sp. SRC1lsM3a, Ascochyta rabiei, Diplodia corticola, Trichoderma citrinoviride, Neurospora crassa) for evolutionary analysis. All of the related species genome sequences were downloaded from the National Center for Biotechnology Information (NCBI) database. To avoid the effects of alternative splicing, we chose the longest transcripts to represent the coding sequences. Gene families were defined by TreeFam methodology and then clustered with Hcluster_sg software (; ). The orthologous groups and single-copy orthologs of these fungi were detected using the Perl script modified according to the programing code of . Multiple sequence alignments of protein sequences were generated for each gene family using MUSCLE (), which converted the CDS alignments into protein alignments. A maximum likelihood phylogenetic tree was created using TreeBest1 with WAG amino acid substitution model. Synteny analysis and core-pan genes analysis were performed on four fungal species (S. bambusicola, P. sporulosa, P. nodorum, and L. maculans).
Analysis of Core Genes and Gene Clusters Involved in Secondary Metabolism
The web-based prediction tool antibiotics and Secondary Metabolite Analysis Shell (antiSMASH) were used to predict secondary metabolite gene clusters and core genes (). Based on previous RNA-Seq data from S. bambusicola S4201 (NCBI accession numbers: SRR2352154 and SRR2153022), we found the FPKM measurements of core genes which were commonly used to present gene expression levels. The PKS domain structures of different species were assigned according to the Conserved Domain Database (CDD) from the NCBI Search database and antiSMASH. The KS domain is the most conserved and is commonly used to infer the phylogenetic relationship between PKS genes. The predicted KS domains of S. bambusicola S4201 and other fungi were aligned by using ClustalW () and the neighbor-joining tree was created with MEGA7, with 1000 bootstrap replicates ().
Expression Vector Construction, Overexpression, and HPLC Profiling
The promoter GPD1, selection marker gene ben, promoter GPD2, and zinc finger transcription factor gene (zftf, as specialized to be involved in HA biosynthesis) were separately amplified from the S4201 genome and from pDHt-Ben using the primers OE-GPD-F1/OE-GPD-R1, OE-Ben-F1/OE-Ben-R1, OE-GPD-F2/OE-GPD-R2, and OE-zftf-F1/OE-zftf-R1. Then, the GPD1-Ben and GPD2-Zftf fragments were generated by overlapping PCR using PrimeSTAR® Max DNA Polymerase (TaKaRa, Dalian, China) and primers OE-GPD-F1/OE-Ben-R1 and OE-GPD-F2/OE-zftf-R1. In addition, the GPD1-Ben fragment was cloned into the pEASY-Blunt Zero vector (TransGen Biotech) to generate the plasmid pGPD1-Ben. The GPD2-Zftf fragment was subcloned into the PmeI site of pGPD1-Ben, forming the plasmid pOE-zftf. S. bambusicola S4201 was grown in PDA at 28°C for 120 h. The spores were collected in sterile water, and the protoplasts were prepared by adding an enzyme mixture. The plasmids were introduced into S. bambusicola S4201 by the polyethylene glycol-calcium chloride (PEG-CaCl2) transformation method, according to published procedures (). After culturing for 120 h at 28°C, the transformants were identified by diagnostic PCR and the sequences were analyzed using the DNA of the mycelium. Primer sequences are shown in the supporting information section (Supplementary Table S1). Cultures of 1 × 105 conidia per mL of S4201 and the overexpressing transformants were grown in PDB at 28°C under constant agitation at 150 r/min for 120 h. Next, the mycelia were used for detection of expression by qRT-PCR. The culture supernatant was extracted with ethyl acetate until it turned colorless. Then, the extracts were concentrated to dryness and the residue was dissolved in acetonitrile. The HA and elsinochrome A (EA) concentrations in the extracts were measured by high-performance liquid chromatography (HPLC) using standard reagents. HPLC was performed using an Agilent Technologies 1220 Infinity LC instrument. The HA and EA of S. bambusicola S4201 and the overexpression transformant were analyzed by HPLC at 30°C using a C18 column (5 μm, 4.6 × 250 mm) (SunFireTM, Waters, United States), with a flow rate of 1 mL/min and injection volume of 10 μL.
Quantitative Real-Time PCR Analysis
The glyceraldehyde-3-phosphate dehydrogenase gene (GAPDH) was used as the internal reference. Total RNA was isolated using TRIzol reagent according to the instructions of the manufacturer (Vazyme, Nanjing, China). The RNA quality was assessed with 1.0% agarose gels and verified using NanoDrop 2000 (Thermo Fisher Scientific, United States). cDNA was synthesized using the HiScript II First Strand cDNA Synthesis Kit (+ gDNA wiper) (Vazyme, Nanjing, China). qRT-PCR was performed using a StepOnePlusTM Real-Time PCR System (Applied Biosystems, United States) and AceQ qPCR SYBR Green Master Mix (Vazyme, Nanjing, China). The results were calculated using the 2–ΔΔCt method. All the gene-specific primers are listed in Supplementary Table S2. All the experimental data were obtained from three biological replicates.
Results
De novo Genome Sequencing, Assembly, and Annotation
We isolated endophytic S. bambusicola S4201 fungi from the fruiting bodies of S. bambusicola P. Henn and demonstrated that it produced HA, based on HPLC analysis (). To gain insight into the pathogenicity, classification, and secondary metabolites of S. bambusicola S4201, we de novo sequenced and assembled its genome. By integrating Illumina HiSeq and PacBio single molecule long-read sequencing, we generated 2649 Mb of raw data and 4359 Mb of polymerase read bases. After removing the adapter sequence and low-quality reads, we obtained 2053 Mb of clean data and 4335 Mb of subread bases. The final S. bambusicola S4201 genome assembly included 32 Mb in 57 contigs, with a contig N50 of 1,565,644 bp (Figure 2). The genomic information was compared with that of Shiraia sp. slf14 (GenBank: AXZN00000000.1, no annotation, unpublished results) in the NCBI database (Table 1). Moreover, the gene length distributions and open reading frames of the genome were predicted based on the assembly sequences (Supplementary Figure S1 and Supplementary Table S3). Non-coding RNA (ncRNA) statistics are shown in Supplementary Table S4. The repeats, long terminal repeats (LTRs), long interspersed elements (LINE), and short interspersed elements (SINE) are shown in Supplementary Tables S5, S6. The genome sequence of S. bambusicola S4201 has been deposited at GenBank under accession number SRR8379567.
FIGURE 2
TABLE 1
| Sample | Shiraia bambusicola S4201 | Shiraia sp. slf14 |
| Genome size (bp) | 32,655,727 | 32,067,383 |
| Contig | 57 | 288 |
| N50 (bp) | 1,565,644 | 525,954 |
| GC content | 47.91 | 48 |
| Chromosome number | 57 | 0 |
| Genome coverage | 100× | 55× |
| Gene number | 11,332 | – |
| ncRNA number | 438 | – |
| Repeat size (bp) | 1,294,534 | – |
| Annotation number | 10,307 (90.95%) | – |
| Sequencing technology | Illumina HiSeq + PacBio | Illumina HiSeq |
Characteristics of the genome of Shiraia bambusicola.
To conduct functional annotation of the S4201 gene model, we used the blast search function to enter the putative protein-coding sequences into the GO, KEGG, COG, Swiss-Prot, Trembl, GenBank NR, EggNOG (), and TransportDB databases. There were 6358 (56.10%) annotated genes obtained for the three main GO categories of biological process, cellular component, and molecular function, including 48 sub-categories (Supplementary Figure S2). In total, 4368 (38.54%) genes were annotated and assigned to 45 different KEGG pathways (Supplementary Figure S3). “Carbohydrate metabolism” was the most enriched pathway, followed by “Amino acid metabolism” and “Translation.” In summary, 1318 (11.63%) genes were annotated into the COG database (Supplementary Figure S4). Among the 25 COG functional categories, the cluster for “General function prediction only” was the largest, followed by “Amino acid transport and metabolism” and “Carbohydrate transport and metabolism.” Only a few genes were assigned to the “Extracellular structures,” “Chromatin structure and dynamics,” and “RNA processing and modification” groups.
Genes Involved in Pathogenicity
To produce a successful infection, phytopathogenic fungi often have to break down plant cell walls and penetrate into host tissues by means of CAZymes (). The S. bambusicola S4201 genome encodes 414 putative CAZymes, including 166 glycoside hydrolases (GHs), 52 glycosyltransferases (GTs), nine polysaccharide lyases (PLs), 13 carbohydrate esterases (CEs), 72 auxiliary activities (AAs), and 102 carbohydrate-binding modules (CBMs). A heatmap based on the classification of CAZymes of 20 fungal species (N. crassa OR74A, Aspergillus oryzae RIB40, Debaryomyces hansenii CBS767, Fusarium fujikuroi IMI 58289, Hypocrella siamensis MTCC 10142, L. maculans JN3, Melanopsichium pennsylvanicum 4, Thielavia terrestris NRRL 8126, Thermothelomyces thermophila ATCC 42464, Ustilago bromivora UB2112, Xanthophyllomyces dendrorhous, Yarrowia lipolytica CLIB122, Zygosaccharomyces bailii CLIB 213, Zymoseptoria tritici ST99CH_1A5, Z. tritici ST99CH_1E4, Arthrobotrys oligospora ATCC 24927, Aspergillus nidulans FGSC A4, Aspergillus niger CBS 513.88, Blumeria graminis f. sp. tritici, S. bambusicola S4201) is shown in Figure 3. The AA coding ability of S. bambusicola S4201 was higher than that of other strains, and the CBM coding ability was second only to A. oligospora ATCC 24927. However, its GT coding ability was the weakest, lower than other selected strains. The proteins containing CAZy domains in the GH, PL, and CE groups may act as plant polysaccharide degradation (PDD) enzymes (). Regarding the PDD enzyme families, 99 sequences were predicted, distributed in 34 families (Supplementary Table S7). As a component of the plant cell wall and the intercellular spaces, pectin can provide nutrients to pathogenic fungi. S. bambusicola S4201 has many candidate pectinases and includes almost all pectinase families known in fungi, including PL1, PL3, PL4, GH78, GH88, GH95, GH105, and GH115. Moreover, some putative pathogenic factors were predicted, such as cutinase genes, which may be needed to catalyze plant cuticle degradation. In addition, proteins belonging to the GH6, GH7, GH39, and GH51 families, involved in the degradation of cellulose, hemi-cellulose, and pectin in plant cell walls, were also predicted. Eight hundred ninety genes in S. bambusicola S4201 were putatively involved in pathogenicity and virulence when analyzed with the PHI database, and this accounts for 7.85% of the total predicted genes (Supplementary Table S8). One thousand twenty-six putative P450s (9.05%) were predicted in S. bambusicola S4201 (Supplementary Table S9). In total, 74 major facilitator superfamily (MFS) transporters (Supplementary Table S10) and 46 ATP-Binding Cassette (ABC) transporters (Supplementary Table S11) were predicted.
FIGURE 3
Comparative Genomics and Classification
In S. bambusicola S4201 and nine other selected fungal species (L. maculans, P. nodorum, P. sporulosa, C. zeina, Stagonospora sp. SRC1lsM3a, A. rabiei, D. corticola, T. citrinoviride, N. crassa), a total of 120,799 genes clustered into 10,851 gene families, including 3423 (31.55%) gene families shared by all species and 2105 (19.40%) that were single copy orthologs (Supplementary Table S12). The differences and quantities of ortholog groups among the fungi are shown in Supplementary Figure S5. The taxonomic status of S. bambusicola among other nine fungal species was evaluated. Based on highly conserved single copy orthologs, the phylogenetic tree (Figure 4) was constructed with the maximum likelihood method. The S. bambusicola S4201 genome was further characterized by comparative gene functional analyses with the following closely related species: P. sporulosa, P. nodorum, and L. maculans. Comparative analyses of the four available genomes revealed that P. nodorum had the largest number of genes. Synteny analysis showed that S. bambusicola S4201 and P. nodorum genome shared many large areas of synteny (Supplementary Figure S6) and they also share more homologous proteins. Comparative analysis of core and pan genes among the four species showed that there are 3480 genes classified as core genes, 35,395 genes classified as pan genes, and 5275 genes classified as dispensable genes (Supplementary Figure S7). The heatmap of dispensable genes in these species suggests that S. bambusicola S4201 has the fewest species specific genes (Supplementary Figure S8). It shows that S. bambusicola S4201 is more closely related to P. nodorum (Phaeosphaeriaceae, Pleosporales, Dothideomycetes) than to other species. The analysis revealed that S. bambusicola S4201 belongs among the species of Pleosporales, Dothideomycetes.
FIGURE 4
Secondary Metabolites
Filamentous fungi can produce many secondary metabolites, such as bioactive compounds or mycotoxins, which have been used for the synthesis of pharmaceuticals. In fungi, the core genes, accessory genes, regulators, transporters, and other genes involved in the biosynthesis and modification of secondary metabolites are organized in clusters (). Using the scaffolds as the query sequences for the antiSMASH 4.1.0 platform, 73 putative secondary biosynthetic gene clusters were predicted. In total, 15 polyketide synthases (PKS, 14 T1PKS, and one T3PKS), six non-ribosomal peptide synthases (NRPS), one linaridin, two terpenes, one indole, one cf_fatty_acid, 28 cf_putative, and two other gene clusters were identified in the S. bambusicola S4201 genome, distributed among 20 scaffolds. Cluster 22 shared 26% gene similarity with the patulin biosynthetic gene cluster, cluster 60 is likely to be an asperfuranone biosynthetic gene cluster (27% of genes show similarity) and cluster 72 acts as a cyclochlorotine biosynthetic gene cluster (25% of genes show similarity). These results indicate that these predicted gene clusters may produce the mentioned secondary metabolites in S. bambusicola S4201.
To further understand the expression patterns of core biosynthetic genes involved in secondary metabolite biosynthesis, we estimated the expression of these genes in fragments per kb per million reads (FPKM) using previous RNA-seq data from S. bambusicola S4201 cultured for 82 h in PDB medium with Triton X-100 (). The secondary metabolite biosynthetic gene clusters are shown in Table 2. Phylogenetic analysis based on the KS domain amino acid sequences of PKSs in S. bambusicola S4201 and the products of known PKSs, was divided according to three main classes of PKSs, including highly reduced (HR) PKSs, partially reduced (PR) PKSs, and non-reduced (NR) PKSs (Figure 5). The reduced PKSs contain the reductive domains dehydratase (DH), enoyl reductase (ER), and keto-reductase (KR), while the NR-PKSs do not. ctg2_orf583, ctg16_orf306, ctg9_orf6, ctg18_orf232, ctg13_orf57, ctg17_orf70, and ctg12_orf33 were grouped with lovastatin and compactin synthase from Aspergillus terreus and Penicillium citrinum in the HR-PKS class. On the other hand, ctg21_orf150, ctg7_orf10, and ctg2_orf331 were nested in the PR-PKS clade. In addition, ctg1_orf376, ctg20_orf61, and ctg2_orf265 were distributed in the NR-PKS class, which include the polyketides cercosporin, aflatoxin, elsinochrome, and a putative polyketide involved in HA biosynthesis (ctg14_orf277).
TABLE 2
| Cluster ID | Core biosynthetic gene ID | Transcript ID | Type | FPKM | Domain structure |
| cluster2 | ctg1_orf376 | No hits found | T1PKS | – | KS-AT |
| cluster7 | ctg2_orf265 | comp5421_c0_ seq3 | T1PKS | 1.98 | KS-AT-TE |
| cluster8 | ctg2_orf331 | No hits found | T1PKS | – | KS-AT-DH-cMT |
| cluster9 | ctg2_orf583 | comp12424_c0_seq1 | T1PKS | 5.09 | KS-AT-DH-cMT-KR |
| cluster22 | ctg7_orf10 | comp5995_c0_ seq2 | T1PKS | 1.39 | KS-AT-DH-KR |
| cluster29 | ctg9_orf6 | comp21554_c0_seq1 | T1PKS | 1.48 | KS-AT-DH-cMT-KR-TD |
| cluster46 | ctg12_orf33 | comp22754_c0_seq1 | T1PKS | 0.90 | KS-AT-DH-cMT-ER-KR |
| cluster50 | ctg13_orf57 | No hits found | T1PKS | – | AT-DH-cMT-KR |
| cluster54 | ctg14_orf277 | comp12566_c1_seq1 | T1PKS | 1,058.05 | KS-AT-TE |
| cluster60 | ctg16_orf306 | comp5618_c0_ seq1 | T1PKS | 39.77 | KS-AT-DH-cMT-ER-KR |
| cluster63 | ctg17_orf70 | comp19386_c0_seq1 | T1PKS | 1.27 | KS-AT-DH-ER-KR |
| cluster66 | ctg18_orf232 | comp9563_c0_ seq1 | T1PKS | 1.62 | AT-DH-ER |
| cluster70 | ctg20_orf61 | No hits found | T1PKS | – | KS-AT-TE |
| cluster73 | ctg21_orf150 | comp10038_c0_seq3 | T1PKS | 2.24 | KS-AT-DH-KR |
| cluster64 | ctg17_orf101 | comp611_c0_ seq2 | T3PKS | 1.08 | – |
| cluster4 | ctg1_orf613 | comp20525_c0_seq1 | NRPS | 0.00 | A-TD |
| cluster11 | ctg3_orf196 | comp11842_c0_seq1 | NRPS | 20.45 | C-A-C-A-E-C-A-C-A-E-C-C |
| cluster31 | ctg9_orf134 | comp70_c0_ seq1 | NRPS | 1.19 | C |
| cluster36 | ctg10_orf5 | No hits found | NRPS | – | C-A-E-C-A-C-A-C-C-A-C-A-C-A-C |
| cluster39 | ctg10_orf231 | comp6571_c0_ seq2 | NRPS | 1.83 | A-C-C |
| cluster72 | ctg20_orf238 | comp1287_c0_ seq1 | NRPS | 1.47 | A |
| cluster3 | ctg1_orf437 | comp7504_c0_ seq1 | Linaridin | 110.33 | – |
| cluster16 | ctg5_orf244 | comp12952_c0_seq1 | Terpene | 129.04 | – |
| cluster65 | ctg17_orf160 | comp13731_c0_seq1 | Terpene | 50.50 | – |
| cluster40 | ctg11_orf2 | No hits found | Indole | – | – |
| cluster42 | ctg11_orf76 | comp12067_c0_seq2 | Other | 4.12 | A-NAD |
| cluster48 | ctg12_orf304 | comp8656_c0_ seq1 | Other | 2.76 | A-TD-KR |
| cluster34 | ctg9_orf346 | comp12432_c0_seq5 | Cf_fatty_acid | 3.88 | – |
Secondary metabolite biosynthetic gene clusters in Shiraia bambusicola S4201.
T1PKS, Type I PKS (Polyketide synthase); T3PKS, Type III PKS; NRPS, non-ribosomal peptide synthetase cluster; Linaridin, linaridin cluster; Terpene, terpene; Indole, indole cluster; Other, cluster containing a secondary metabolite-related protein that does not fit into any other category; Cf_fatty_acid, possible fatty acid cluster.
FIGURE 5
The putative HA biosynthesis genes were classified into cluster 54, which was located on scaffold 14 (location: 1101201–1146665 nt) and included 13 genes (Figure 6). The genes in cluster 54 probably encode a FAD/FMN-dependent oxidoreductase (ctg14_orf271), a hydroxylase (ctg14_orf272), a zinc finger transcription factor (ctg14_orf273), an o-methyltransferase (ctg14_orf274), an MFS transporter (ctg14_orf275), an o-methyltransferase/FAD-dependent monooxygenase (ctg14_orf276), a polyketide synthase (ctg14_orf277), a dynamin GTPase domain (ctg14_orf278), and hypothetical proteins (ctg14_orf279–ctg14_orf283) (Supplementary Table S13).
FIGURE 6
Overexpression of the Zinc Finger Transcription Factor zftf
A powerful approach to enhance the production of secondary metabolites is to overexpress Zn(II)Cys6-type pathway-specific transcription factors embedded in the corresponding gene clusters (; ). To determine the function of zftf, we cloned it into a plasmid, then transformed S. bambusicola S4201 with pOE-zftf, and the OE mutants were selected by diagnostic PCR amplification of the expression cassette. HPLC analysis with standard reagents (Supplementary Figure S9) showed that HA production by the OE mutant (206.75 mg/L) was higher than in wild type (82.60 mg/L). The expression levels of zftf in the mycelia were analyzed by quantitative real-time PCR (qRT-PCR), and the OE mutant showed more than threefold upregulation on average. These results show that the zinc finger transcription factor zftf plays an important regulatory role in HA production.
Quantitative Real-Time PCR Analysis of Several Genes
Seven genes (encoding polyketide synthase, o-methyltransferase/FAD-dependent monooxygenase, o-methyltransferase, FAD/FMN-dependent oxidoreductase, hydroxylase, MFS transporter, and zinc finger transcription factor) involved in HA biosynthesis were selected for qRT-PCR analysis to validate changes in gene expression. qRT-PCR analysis of total RNA extracted from wild type and the OE mutant grown in PDA showed that expression of these genes was upregulated in the OE mutant (Figure 7) indicating that the ZFTF transcriptional activator controls HA production by controlling gene transcript levels. Based on the previous RNA-Seq data (), the heatmap analysis of the expression of these genes in S4201-D1 and S4201-W was consistent with that in S4201-W and OE mutant. S. bambusicola S4201-W shows abundant HA production, while the S4201-D1 mutant does not. Thus, the HA biosynthesis core gene cluster is apparently coregulated through zftf, which encodes a Zn(II)Cys6 transcriptional activator, as expression of relevant genes was enhanced in the OE mutant.
FIGURE 7
Discussion
Shiraia bambusicola is an important bamboo ascomycete pathogen. In this study, we sequenced, assembled, and analyzed the genome of this pathogen. The genome data are the initial step to understand the biology of S. bambusicola.
The chemical composition of bamboo fibers is mainly lignin, cellulose, and hemicellulose. Lignin accounts for the hardness and yellowness of bamboo fibers and is present at different concentrations in different layers of the cell wall (
Dothideomycetes is regarded as the largest and most diverse class of Ascomycete fungi, and comprises 11 orders, 90 families, 1300 genera, and more than 19,000 known species. The species are taxonomically classified into three subclasses, 11 orders: Capnodiales, Dothideales and Myriangiales (Dothideomycetidae); Hysteriales, Jahnulales, Mytilinidiales and Pleosporales (Pleosporomycetidae); Botryosphaeriales, Microthyriales, Patellariales, and Trypetheliales (Incertae sedis). Among them, the most diverse fungal order in Dothideomycetes is Pleosporales, which represent roughly a quarter of all Dothideomycetous species (
In the secretion and transport stages, secondary metabolites act as virulence factors, and transport is also important for pathogenesis. These secondary metabolites, particularly mycotoxins in S. bambusicola, may be responsible for its adaptation to an exclusively parasitic life cycle. There are mainly two transport systems, ABC-transporter (
To study the phylogenetic relationship of PKSs in the genome of S. bambusicola, the structures of the synthetic products and the PKS reduction types, a phylogenetic tree was constructed based on the ketoacyl synthase domain (KS domain) amino acid sequence of the polyketide synthases of S. bambusicola and 16 other polyketide synthases. Because some of the predicted polyketide synthases of S. bambusicola were similar to the polyketide synthases producing known compounds, and clustered in the phylogenetic tree of the same branch, we speculated that S. bambusicola can produce the same compounds or similar structural analogs (Figure 5). We selected one of these compounds, EA, for HPLC detection. EA was detected when S. bambusicola S4201 was cultured in PDB for 120 h (Supplementary Figure S9). Since S. bambusicola can produce EA and HA, and their structures are similar (Figure 1), we believe that EA is also a type of photosensitizer and virulence factor that can cause bamboo disease.
Some HA biosynthetic genes in the cluster we predicted by antiSMASH differed from the ones in Shiraia sp. slf14 (GenBank: KM434884.1). Compared with Shiraia sp. slf14, the gene clusters we predicted lacked the NADPH-dependent oxidoreductase gene, MFS gene, fasciclin gene, putative multicopper oxidase gene, and three hypothetical proteins genes. In contrast, they included a dynamin GTPase domain, an unknown gene, and four hypothetical proteins genes. Comparison of putative HA biosynthetic gene cluster in S. bambusicola S4201 and Shiraia sp. slf14 is shown in Figure 6. In order to finally determine the function of these genes and whether they are involved in HA biosynthesis, further experiments such as gene knockout are still needed. The putative S. bambusicola S4201 PKS (cluster 54_ctg14_orf277) involved in HA biosynthesis are clustered in the same branch as the Cercospora nicotianae PKS, and their molecular structure and biosynthetic gene clusters are similar. Comparing the CTB and HA cluster genes, NADPH-dependent oxidoreductase (CTB6) and FAD-dependent monooxygenase (CTB7) appeared to be unique to the CTB cluster. Considering the function of these enzymes, it makes sense that they are absent in the HA cluster, given that these structural features are non-existent (
The Zn(II)Cys6 family of transcription factors can regulate various cellular processes in fungi (
FIGURE 8

Schematic diagram of the central carbons of hypocrellin A showing its synthesis in Shiraia bambusicola. This flow chart simulates the process of HA synthesis. The highlighted and dotted lines on the left indicate the sources of HA substrates, and the genes marked with red on the right were upregulated in the transcriptome analysis we performed (
Current studies are mainly focused on the fermentation of S. bambusicola and the extraction of secondary metabolites. Future efforts should focus on identifying the genes involved in HA biosynthesis, the intermediate metabolites, and the synthesis pathways by gene knockout experiments and expression of heterologous proteins.
Conclusion
Shiraia bambusicola is an important parasitic pathogen that grows on the twigs of bamboos and which has traditionally been considered to have anti-inflammatory and antiviral properties. We have de novo sequenced and assembled the genome of S. bambusicola S4201, which may serve as a basis for understanding its pathogenicity and exclusive parasitic lifestyle. S. bambusicola is capable of encoding a large and diverse set of CAZy, primary and secondary metabolism enzymes and transporters. The genome comparative study facilitated our understanding of the phylogeny of S. bambusicola and the evolutionary relationships between S. bambusicola and other sequenced species. Numerous secondary metabolite biosynthetic gene clusters were found. In addition, biosynthetic gene clusters for unknown products show the enormous potential that exists for drug discovery in S. bambusicola. Overexpression of the transcription factor zftf increased the production of HA by regulating the expression levels of other genes in the HA biosynthetic gene cluster. A new putative HA biosynthetic pathway was proposed. Our analyses provide some new insights into the pathogenicity, phylogeny, and secondary metabolites of S. bambusicola. In the future, this study will be an important resource for gene discovery and to increase HA production through further genetic manipulation.
Statements
Data availability statement
The genome data datasets generated for this study can be found in the NCBI Repository with accession number SRR8379567. The RNA-Seq datasets from S. bambusicola S4201 generated for this study can be found in the NCBI Repository with accession numbers SRR2352154 and SRR2153022.
Author contributions
S-LC, S-ZY, and NZ conceived and designed the experiments. NZ, DL, and B-JG performed the experiments. NZ and XT analyzed the data. XL contributed reagents, materials, and analysis tools. NZ wrote the manuscript.
Funding
This study was supported by the funds of the National Key Technology Research and Development Program of the Ministry of Science and Technology of China (No. 2012BAD36B05) and the Priority Academic Program Development of Jiangsu Higher Education Institutions.
Acknowledgments
We are sincerely grateful to the referees whose comments made our results more accessible to the readers. We are thankful to Dr. Cheng-Shu Wang (Institute of Plant Physiology and Ecology, SIBS, CAS) for donating plasmid pDHt-Ben.
Conflict of interest
The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.
Supplementary material
The Supplementary Material for this article can be found online at: https://www.frontiersin.org/articles/10.3389/fmicb.2020.00643/full#supplementary-material
References
1
AliS. M.OlivoM. (2002). Efficacy of hypocrellin pharmacokinetics in phototherapy.Int. J. Oncol.211229–1237.
2
BadouinH.HoodM. E.GouzyJ.AguiletaG.SiguenzaS.PerlinM. H.et al (2015). Chaos of rearrangements in the mating-type chromosomes of the anther-smut fungus Microbotryum lychnidis-dioicae.Genetics2001275–1284. 10.1534/genetics.115.177709
3
BennettJ. W. (2005). Fungal secondary metabolism-from biochemistry to genomics.Nat. Rev. Microbiol.3937–947. 10.1038/nrmicro1286
4
BensonG. (1999). Tandem repeats finder: a program to analyze DNA sequences.Nucleic Acids Res.27573–580. 10.1093/nar/27.2.573
5
BergmannS.SchümannJ.ScherlachK.LangeC.BrakhageA. A.HertweckC. (2007). Genomics-driven discovery of PKS-NRPS hybrid metabolites from Aspergillus nidulans.Nat. Chem. Biol.3213–217. 10.1038/nchembio869
6
BirneyE.ClampM.DurbinR. (2004). Genewise and genomewise.Genome Res.14988–995. 10.1101/gr.1865504
7
BlinK.WolfT.ChevretteM. G.LuX. W.SchwalenC. J.KautsarS. A.et al (2017). antiSMASH 4.0-improvements in chemistry prediction and gene cluster boundary identification.Nucleic Acids Res.45W36–W41.
8
BorastonA. B.BolamD. N.GilbertH. J.DaviesG. J. (2004). Carbohydrate-binding modules: fine-tuning polysaccharide recognition.Biochem. J.382769–781. 10.1042/bj20040892
9
CantarelB. L.CoutinhoP. M.CorinneR.ThomasB.VincentL.BernardH. (2009). The Carbohydrate-Active EnZymes database (CAZy): an expert resource for glycogenomics.Nucleic Acids Res.37D233–D238.
10
ChengT. F.JiaX. M.MaX. H.LinH. P.ZhaoY. H. (2004). Phylogenetic study on Shiraia bambusicola by rDNA sequence analyses.J. Basic Microb.44339–350. 10.1002/jobm.200410434
11
ChooiY. H.ZhangG.HuJ.Muria-GonzalezM. J.TranP. N.PettittA.et al (2017). Functional genomics-guided discovery of a light-activated phytotoxin in the wheat pathogen Parastagonospora nodorum via pathway activation.Environ. Microbiol.191975–1986. 10.1111/1462-2920.13711
12
DaubM. E.ChungK. R. (2009). Photoactivated perylenequinone toxins in plant pathogenesis.Plant Relationsh.252197–206. 10.1016/j.femsle.2005.08.033
13
EdgarR. C. (2004). MUSCLE: multiple sequence alignment with high accuracy and high throughput.Nucleic Acids Res.32113–132.
14
FainoL.SeidlM. F.DatemaE.BergG. C. M. V. D.JanssenA.WittenbergA. H. J.et al (2015). Single-molecule Real-Time sequencing combined with optical mapping yields completely finished fungal genome.mBio61–11.
15
FischerM.KnollM.SirimD.WagnerF.FunkeS.PleissJ. (2007). The cytochrome P450 engineering database: a navigation and prediction tool for the cytochrome P450 protein family.Bioinformatics232015–2017. 10.1093/bioinformatics/btm268
16
GalperinM. Y.MakarovaK. S.WolfY. I.KooninE. V. (2015). Expanded microbial genome coverage and improved protein family annotation in the COG database.Nucleic Acids Res.43D261–D269.
17
GaoL.MaY. Y.LiX. T.ZhangL. P.ZhangC.ChenQ. Y.et al (2019). Research on the roles of genes coding ATP-binding cassette transporters in Porphyromonas gingivalis pathogenicity.J Cell Biochem.12193–102. 10.1002/jcb.28887
18
GardnerP. P.DaubJ.TateJ. G.NawrockiE. P.KolbeD. L.LindgreenS.et al (2009). Rfam: updates to the RNA families database.Nucleic Acids Res.37136–140.
19
GuX. T.ZhouJ. H.FengY. Y.WeiS. H.WangX. S.ZhangB. W. (2005). Determination of Hypocrellin A in Hypocrella Bambusae by HPLC/DAD.J. Instru. Anal.24130–132.
20
GuoL. Y.YanS. Z.LiQ.XuQ.LinX.QiS. S.et al (2017). Poly(lactic-co-glycolic) acid nanoparticles improve oral bioavailability of hypocrellin A in rat.RSC Adv.742073–42082. 10.1039/c7ra04748g
21
HacquardS.KracherB.HirumaK.MünchP. C.Garrido-OterR.ThonM. R.et al (2016). Survival trade-offs in plant roots during colonization by closely related beneficial and pathogenic fungi.Nat. Commun.7:11362.
22
HanX.ChakraborttiA.ZhuJ.LiangZ. X.LiJ. (2016). Sequencing and functional annotation of the whole genome of the filamentous fungus Aspergillus westerdijkiae.BMC Genomics17:633. 10.1186/s12864-016-2974-x
23
Huerta-CepasJ.SzklarczykD.ForslundK.CookH.HellerD.WalterM. C.et al (2016). eggNOG 4.5: a hierarchical orthology framework with improved functional annotations for eukaryotic, prokaryotic and viral sequences.Nucleic Acids Res.44286–293.
24
JiangH.ShenY.LiuW.LingL. (2014). Deletion of the putative stretch-activated ion channel Mid1 is hypervirulent in Aspergillus fumigatus.Fungal Genet Biol.6262–70. 10.1016/j.fgb.2013.11.003
25
JohnsonA. D.HandsakerR. E.PulitS. L.NizzariM. M.O’DonnellC. J.BakkerP. I. W. D. (2008). SNAP: a web-based tool for identification and annotation of proxy SNPs using HapMap.Bioinformatics242938–2939. 10.1093/bioinformatics/btn564
26
KanehisaM.SatoY.KawashimaM.FurumichiM.TanabeM. (2016). KEGG as a reference resource for gene and protein annotation.Nucleic Acids Res.44D457–D462.
27
KhalilH. P. S. A.BhatI. U. H.JawaidM.ZaidonA.HermawanD.HadiY. S. (2012). Bamboo fibre reinforced biocomposites: a review.Mater. Design.42353–368. 10.1016/j.matdes.2012.06.015
28
KimK. E.PelusoP.BabayanP.YeadonP. J.YuC.FisherW. W.et al (2014). Long-read, whole-genome shotgun sequence data for five model organisms.Sci. Data1140045–140055.
29
KirkP. M.CannonP. F.DavidJ. C.StalpersJ. A. (2008). Ainsworth and Bisby’s Dictionary of the Fungi, 10th Edn, Wallingford: CAB International.
30
KumarS.StecherG.TamuraK. (2016). MEGA7: molecular evolutionary genetics analysis version 7.0 for bigger datasets.Mol. Biol. Evol.331870–1874. 10.1093/molbev/msw054
31
LagesenK.HallinP.RodlandE.StaerfeldtH.RognesT.UsseryD. (2007). RNAmmer: consistent and rapid annotation of ribosomal RNA genes.Nucleic Acids Res.353100–3108. 10.1093/nar/gkm160
32
LairsonL. L.HenrissatB.DaviesG. J.WithersS. G. (2008). Glycosyltransferases: structures, functions, and mechanisms.Annu. Rev. Biochem.77521–555. 10.1146/annurev.biochem.76.061005.092322
33
LarkinM. A.BlackshieldsG.BrownN. P.ChennaR.McgettiganP. A.McwilliamH.et al (2007). Clustal W and clustal X version 2.0.Bioinformatics232947–2948. 10.1093/bioinformatics/btm404
34
LevasseurA.DrulaE.LombardV.CoutinhoP. M.HenrissatB. (2013). Expansion of the enzymatic repertoire of the CAZy database to integrate auxiliary redox enzymes.Biotechnol. Biofuels641–55.
35
LiH.CoghlanA.RuanJ.CoinL. J.HérichéJ. K.OsmotherlyL.et al (2006). TreeFam: a curated database of phylogenetic trees of animal gene families.Nucleic Acids Res.34D572–D580.
36
LinX.YanS. Z.QiS. S.XuQ.HanS. S.GuoL. Y.et al (2017). Transferrin-modified nanoparticles for photodynamic therapy enhance the antitumor efficacy of hypocrellin A.Front. Pharmacol.8:815. 10.3389/fphar.2017.00815
37
LiuB.BaoJ. Y.ZhangZ. B.YanR. M.WangY.YangH. L.et al (2018). Enhanced production of perylenequinones in the endophytic fungus Shiraia sp. Slf14 by calcium/calmodulin signal transduction.Appl. Microbiol. Biot.102153–163. 10.1007/s00253-017-8602-0
38
LiuY. X.HydeK. D.AriyawansaH.LiW. J.ZhouD. Q.YangY. L.et al (2013). Shiraiaceae, new family of Pleosporales (Dothideomycetes, Ascomycota).Phytotaxa10351–60.
39
LoweT. M.EddyS. R. (1997). tRNAscan-SE: a program for improved detection of transfer RNA genes in genomic sequence.Nucleic Acids Res.25955–964. 10.1093/nar/25.5.955
40
MakarovaK. S.WolfY. I.KooninE. V. (2015). Archaeal clusters of orthologous genes (arCOGs): an update and application for analysis of shared features between Thermococcales, Methanococcales, and Methanobacteriales.Life5818–840. 10.3390/life5010818
41
MorakotkarnD.KawasakiH.SekiT. (2007). Molecular diversity of bamboo-associated fungi isolated from Japan.FEMS Microbiol. Lett.26610–19. 10.1111/j.1574-6968.2006.00489.x
42
NelsonD. R. (2011). Progress in tracing the evolutionary paths of cytochrome P450.BBA Proteins Proteom.181414–18. 10.1016/j.bbapap.2010.08.008
43
NewmanA. G.TownsendC. A. (2016). Molecular characterization of the cercosporin biosynthetic pathway in the fungal plant pathogen Cercospora nicotianae.J. Am. Chem. Soc.1384219–4228. 10.1021/jacs.6b00633
44
O’BrienE. M.MorganB. J.MulrooneyC. A.CarrollP. J.KozlowskiM. C. (2010). Perylenequinone natural products: total synthesis of hypocrellin A.J. Org. Chem.7557–68. 10.1021/jo901386d
45
OhmR. A.NicolasF.BernardH.SchochC. L.HorwitzB. A.BarryK. W.et al (2012). Diverse lifestyles and strategies of plant pathogenesis encoded in the genomes of eighteen Dothideomycetes fungi.PLoS Pathog.8:e1003037. 10.1371/journal.ppat.1003037
46
Ospina-GiraldoM. D.GriffithJ. G.LairdE. W.MingoraC. (2010). The CAZyome of Phytophthora spp.: a comprehensive analysis of the gene complement coding for carbohydrate-active enzymes in species of the genus phytophthora.BMC Genomics11:525. 10.1186/1471-2164-11-525
47
QiS. S.LinX.ZhangM. M.YanS. Z.YuS. Q.ChenS. L. (2014). Preparation and evaluation of hypocrellin A loaded poly (lactic-co-glycolic acid) nanoparticles for photodynamic therapy.RSC Adv.440085–40094. 10.1039/c4ra05796a
48
RobertsonC. A.EvansD. H.AbrahamseH. (2009). Photodynamic therapy (PDT): a short review on cellular mechanisms and cancer research applications for PDT.J. Photoch. Photobio. B961–8. 10.1016/j.jphotobiol.2009.04.001
49
RogersS. O.BendichA. J. (1994). “Extraction of total cellular DNA from plants, algae and fungi,” in Plant Molecular Biology Manual, edsGelvinS. B.SchilperoortR. A. (Dordrecht: Springer), 183–190. 10.1007/978-94-011-0511-8_12
50
RokasA.WilliamsB. L.NicoleK.CarrollS. B. (2003). Genome-scale approaches to resolving incongruence in molecular phylogenies.Nature425798–804. 10.1038/nature02053
51
RuanJ.LiH.ChenZ. Z.CoghlanA.CoinL. J.GuoY.et al (2008). Treefam: 2008 update.Nucleic Acids Res.36D735–D740.
52
RytiojaJ.HildénK.YuzonJ.HatakkaA.VriesR. P. D.MäkeläM. R. (2014). Plant-polysaccharide-degrading enzymes from Basidiomycetes.Microbiol. Mol. Biol. R.78614–649. 10.1128/mmbr.00035-14
53
ShenX. Y.LiT.ChenS.FanL.GaoJ.HouC. L. (2015). Characterization and phylogenetic analysis of the mitochondrial genome of Shiraia bambusicola reveals special features in the order of pleosporales.PLoS One10:0116466. 10.1371/journal.pone.0116466
54
SherlockG. (2009). Gene ontology: tool for the unification of biology.Nat. Genet.2525–29.
55
SitC. S.RuzziniA. C.Van ArnamE. B.RamadharT. R.CurrieC. R.ClardyJ. (2015). Variable genetic architectures produce virtually identical molecules in bacterial symbionts of fungus-growing ants.Proc. Natl. Acad. Sci. U.S.A.11213150–13154. 10.1073/pnas.1515348112
56
StankeM.DiekhansM.BaertschR.HausslerD. (2008). Using native and syntenically mapped cDNA alignments to improve de novo gene finding.Bioinformatics24637–644. 10.1093/bioinformatics/btn013
57
Ter-HovhannisyanV.LomsadzeA.ChernoffY. O.BorodovskyM. (2008). Gene prediction in novel fungal genomes using an ab initio algorithm with unsupervised training.Genome Res.181979–1990. 10.1101/gr.081612.108
58
ThiemeK. G.JenniferG.SasseC.ValeriusO.ThiemeS.KarimiR.et al (2018). Velvet domain protein VosA represses the zinc cluster transcription factor SclB regulatory network for Aspergillus nidulans asexual development, oxidative stress response and secondary metabolism.PLoS Genet.14:e1007511. 10.1371/journal.pgen.1007511
59
Torto-AlaliboT.CollmerC. W.Gwinn-GiglioM. (2009). The plant-associated microbe gene ontology (PAMGO) consortium: community development of new gene ontology terms describing biological processes involved in microbe-host interactions.BMC Microbiol.9:S1. 10.1186/1471-2180-9-S1-S1
60
TsujiM.KudohS.HoshinoT. (2015). Draft genome sequence of cryophilic basidiomycetous yeast Mrakia blollopis SK-4, isolated from an algal mat of Naga-ike lake in the Skarvsnes ice-free area, east Antarctica.Genome Announc.3:e1454.
61
WangQ. H.ChenD. P.WuM. C.ZhuJ. D.JiangC.XuJ. R.et al (2018). MFS transporters and GABA metabolism are involved in the Self-defense against DON in Fusarium graminearum.Front. Plant Sci.9:438. 10.3389/fpls.2018.00438
62
YongZ.KangZ.FangA.HanY.YangJ.XueM.et al (2014). Specific adaptation of Ustilaginoidea virens in occupying host florets revealed by comparative and functional genomics.Nat. Commun.51–12.
63
ZhaoN.LinX.QiS. S.LuoZ. M.ChenS. L. (2016). De novo transcriptome assembly in Shiraia bambusicola to investigate putative genes involved in the biosynthesis of hypocrellin A.Int. J. Mol. Sci.17311–324.
64
ZhaoZ.LiuH.WangC.XuJ. R. (2013). Comparative analysis of fungal genomes reveals different plant cell wall degrading capacity in fungi.BMC Genomics14:274. 10.1186/1471-2164-14-274
Summary
Keywords
Shiraia bambusicola, pathogenicity, phylogeny, secondary metabolites, biosynthesis
Citation
Zhao N, Li D, Guo B-J, Tao X, Lin X, Yan S-Z and Chen S-L (2020) Genome Sequencing and Analysis of the Hypocrellin-Producing Fungus Shiraia bambusicola S4201. Front. Microbiol. 11:643. doi: 10.3389/fmicb.2020.00643
Received
28 April 2019
Accepted
20 March 2020
Published
09 April 2020
Volume
11 - 2020
Edited by
David W. Ussery, University of Arkansas for Medical Sciences, United States
Reviewed by
Kathryn Bushley, University of Minnesota Twin Cities, United States; Michael H. Perlin, University of Louisville, United States
Updates

Check for updates
Copyright
© 2020 Zhao, Li, Guo, Tao, Lin, Yan and Chen.
This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.
*Correspondence: Shu-Zhen Yan, yanshuzhen@njnu.edu.cnShuang-Lin Chen, chenshuanglin@njnu.edu.cn; nnucylab@126.com
This article was submitted to Evolutionary and Genomic Microbiology, a section of the journal Frontiers in Microbiology
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.