Literature DB >> 25017872

Identification of tissue-enriched novel transcripts and novel exons in mice.

Seong-Eui Hong, Hong Ki Song, Do Han Kim1.   

Abstract

BACKGROUND: RNA sequencing (RNA-seq) has revolutionized the detection of transcriptomic signatures due to its high-throughput sequencing ability. Therefore, genomic annotations on different animal species have been rapidly updated using information from tissue-enriched novel transcripts and novel exons.
RESULTS: 34 putative novel transcripts and 236 putative tissue-enriched exons were identified using RNA-Seq datasets representing six tissues available in mouse databases. RT-PCR results indicated that expression of 21 and 2 novel transcripts were enriched in testes and liver, respectively, while 31 of the 39 selected novel exons were detected in the testes or heart. The novel isoforms containing the identified novel exons exhibited more dominant expression than the known isoforms in heart and testes. We also identified an example of pathology-associated exclusion of heart-enriched novel exons such as Sorbs1 and Cluh during pressure-overload cardiac hypertrophy.
CONCLUSION: The present study depicted tissue-enriched novel transcripts, a tissue-specific isoform switch, and pathology-associated alternative splicing in a mouse model, suggesting tissue-specific genomic diversity and plasticity.

Entities:  

Mesh:

Substances:

Year:  2014        PMID: 25017872      PMCID: PMC4111849          DOI: 10.1186/1471-2164-15-592

Source DB:  PubMed          Journal:  BMC Genomics        ISSN: 1471-2164            Impact factor:   3.969


Background

Information on the spatial and temporal signatures of transcriptomes is essential for diagnosis and treatment of severe diseases such as cardiomyopathies and malignant cancers. For the past several decades, high-throughput (HTP) data generated using the microarray method have contributed significantly to the discovery of quantitative signatures of various diseases. However, the microarray method has critical limitations, such as spatial bias, uneven probe problems, low sensitivity, and dependency on the probes spotted. Therefore, large-scale transcriptomic analyses using the microarray method have been superseded by the RNA-Seq generated through application of the recently developed next-generation sequencing (NGS) method. RNA-Seq is a revolutionary method useful for transcriptomic signatures, since it can elucidate both quantitative and qualitative signatures (e.g., alternative splicing, AS) by de novo analysis, and it has therefore made possible the large-scale discovery of novel transcripts, such as noncoding RNAs. AS is an important event for proteome complexity and proteome diversity. However, current approaches using microarray or serial analysis of gene expression (SAGE) tags have faced limitations, such as probe dependency and low coverage. The robust sequencing capacity of RNA-Seq has dramatically increased our knowledge of dynamic alternation via AS. For instance, RNA-seq has revealed the subtype-specific novel isoforms for the most common breast cancers (e.g. triple negative breast cancer (TNBC), non-TNBC, and human epidermal growth factor receptor 2 (HER2)-positive breast cancer [1]). Information related to novel exons, recognized in the intronic regions, has rapidly increased owing to RNA-Seq [2-4]. De novo analyses of RNA-Seq datasets have rapidly updated the genome annotations of different species through examination of novel transcripts [5-7]. Furthermore, the detection of novel non-coding RNAs by RNA-Seq has identified them as important functional molecules regulating various biological processes [8-10]. The present study employed RNA-seq data to identify novel exons and novel transcripts enriched in different tissues in mice (here “novel” means “new” exons or “new” transcripts not identified in mice so far), leading to the discovery of novel transcripts expressed in testes or liver, and recognition that the novel isoforms containing the novel exons were dominantly expressed in testes or heart. These results should contribute to a more sophisticated annotation of the mouse genome, as well as improved understanding of tissue-specific gene regulation.

Results and Discussion

In silicoanalysis of tissue-enriched novel transcripts and exons

In order to identify tissue-enriched novel transcripts and exons in mice, the RNA-seq datasets for six tissues (i.e., GSE30352 for brain, cerebrum, heart, kidney, liver and testes) [11] were analyzed using the pipeline ‘Tophat-Cufflinks-Cuffcompare’ [12, 13]. As a result, 76,250 and 77,784 transcribed loci were constructed using UCSC and ENSEMBL, respectively. Among the transcribed loci, 184 transcripts located in the intergenic region were collected as putative novel transcripts (Additional file 1: Table S1). From this list of putative novel transcripts, we further examined the tissue-enriched transcripts using DESeq [14]. Novel transcripts exhibiting significant enrichment (P < 0.05) in the specific tissue were eventually defined as tissue-enriched novel transcripts. As a result, 32 and 2 novel transcripts were found to be significantly enriched in testes and liver, respectively (Table 1).
Table 1

Summary of novel exons

IDPosition (mm9)# of exonsTissue p-value min 1 p-value max 2 Homologous proteinNon-mouse geneEST
tNT-1chr1:99242925-99375659+13Testis2.82E-102.38E-06XP_001475034.3Slc06d1-
tNT-2chr1:163105753-163167403+18Testis2.17E-070.000138XP_344166.4Slc9c2-
tNT-3chr1:121786157-121796422-5Testis0.0006050.0083NP_001102853--
tNT-4chr10:86118523-86133752+9Testis6.79E-076.28E-05XP_487135.3--
tNT-5chr10:85157135-85173678-10Testis4.18E-050.002125XP_896558.3--
tNT-6chr10:85989049-86004244-9Testis2.73E-079.99E-06XP_896769.1--
tNT-7chr10:111578812-111597369-6Testis1.13E-060.000297XP_001480681.1--
tNT-8chr13:56527964-56535585+5Testis1.49E-081.44E-05XP_001475551RGD1562024O
tNT-9chr13:97569388-97679749+17Testis3.30E-082.86E-06XP_005065555Ankrd31O
tNT-10chr15:25984096-25992242-5Testis0.0008570.014133--O
tNT-11chr15:76363652-76365439-5Testis1.13E-085.45E-05XP_988010.2Tmem249O
tNT-12chr18:13666918-13682257+5Testis0.0001830.005416YP_480919--
tNT-13chr18:32317886-32322406-5Testis4.44E-070.000236WP_005016571-O
tNT-14chr18:32617748-32622397-8Testis0.0002020.014407ELW62217-O
tNT-15chr19:40823242-40903550+22Testis4.82E-080.001078XP_004749675-O
tNT-16chr2:170290671-170296220-5Testis4.46E-080.000182XP_004246409-O
tNT-17chr2:173112853-173116672-5Testis1.93E-060.00039YP_003981506-O
tNT-18chr3:31543868-31589424-14Testis1.07E-076.66E-05NP_001028651.1--
tNT-19chr5:28278582-28303869+6Testis6.41E-111.92E-06XP_003085546--
tNT-20chr5:129869449-129872910+5Testis9.84E-092.13E-05YP_001641156-O
tNT-21chr5:117435484-117468048-5Testis0.0008160.017383XP_001524870.1-O
tNT-22chr6:16406558-16419928-6Testis7.12E-050.00307--O
tNT-23chr6:44030493-44033356-5Testis0.0012820.018015--O
tNT-24chr7:120126477-120132935+5Testis0.0044190.041732EGV91268-O
tNT-25chr7:127696533-127711472+5Testis7.17E-050.002835EDL17209.1-O
tNT-26chr7:36029889-36060558-10Testis1.14E-084.01E-06XP_001480194WDR88-
tNT-27chr8:74348600-74377254-11Testis9.06E-060.001105EDL28738-O
tNT-28chrX:98891146-98901100+5Testis0.000370.034496XP_005095122--
tNT-29chrX:43597928-43606686-5Testis1.97E-101.94E-06---
tNT-30chr12:44067864-44135252-5Testis0.0022540.017821EDL38698.1-
tNT-31chr17:14128560-14192168-6Testis6.01E-114.13E-06EDL20486.1--
tNT-32chr17:21191727-21199635-6Testis7.30E-113.14E-06EDL20488.1-O
lNT-1chr10:111026064-111048582+5Liver1.19E-050.021035EDL21734.1-O
lNT-2chr12:73709901-73729881-6Liver4.69E-060.009759XP_003512062.1Dhrc7O

1,2Indicate the minimum and maximum p-values, respectively, when the expression of novel transcripts in testis compared to other tissues in a pairwise manner.

Summary of novel exons 1,2Indicate the minimum and maximum p-values, respectively, when the expression of novel transcripts in testis compared to other tissues in a pairwise manner. In addition to the novel transcripts, we examined the tissue-enriched novel exons for known genes and the novel junctions for the obtained de novo transcripts. In total, 5,582 novel exons were identified from 6 tissues (Additional file 2: Table S2). To examine tissue-enrichment of the novel exons, the read numbers for the novel exons were counted and compared across the 6 tissues in a pairwise manner using DESeq. Of the 236 novel exons evaluated, 197 were expressed in testes (Additional file 2: Table S2), which was consistent with a study by Howald et al. reporting that these novel transcripts are mainly identified in the testes of humans [15].

Experimental confirmation of testes- and liver-enriched novel transcripts

Enrichment of the putative novel transcripts in testes and liver was further examined experimentally using mouse heart, testes, liver, kidney, brain and lung tissues by qRT-PCR and RT-PCR, to determine the expression levels and patterns. Among the 32 testes- and 2 liver-enriched novel transcripts (32 tNT and 2 lNT), enrichment of 21 tNTs and 2 lNTs were experimentally confirmed (Figure 1). We were unable to detect 11 tNTs, including tNT-5, −31, and −32, by RT-PCR. Although highly specific expression of tNT-13 was found in testes using qRT-PCR, we could not detect the expression using RT-PCR, which may have been due to low expression levels.
Figure 1

Testes- and liver-enriched expression of the novel transcripts. The expressions of testes- and liver-enriched novel transcripts were experimentally confirmed by (A) RT-PCR and (B) qRT-PCR for 6 tissues (H: Heart, T: Testes, Lv: Liver, K: Kidney, B: Brain, Lu: Lung). tNT and lNT indicate testes- and liver-enriched novel transcripts, respectively. tNTs and lNTs with blue and orange circles, respectively, were experimentally confirmed by both RT-PCR and qRT-PCR experiments. Values on the X axis indicate the relative expression of tNTs and lNTs in testes and liver, respectively, compared to the expressions in brain (log10(2-△△Ct)). The values on the Y axis indicate the relative expression of the novel transcripts when compared to the expressions of 18S in testes (log10(2-△Ct)).

Testes- and liver-enriched expression of the novel transcripts. The expressions of testes- and liver-enriched novel transcripts were experimentally confirmed by (A) RT-PCR and (B) qRT-PCR for 6 tissues (H: Heart, T: Testes, Lv: Liver, K: Kidney, B: Brain, Lu: Lung). tNT and lNT indicate testes- and liver-enriched novel transcripts, respectively. tNTs and lNTs with blue and orange circles, respectively, were experimentally confirmed by both RT-PCR and qRT-PCR experiments. Values on the X axis indicate the relative expression of tNTs and lNTs in testes and liver, respectively, compared to the expressions in brain (log10(2-△△Ct)). The values on the Y axis indicate the relative expression of the novel transcripts when compared to the expressions of 18S in testes (log10(2-△Ct)). Expression levels of tNTs and lNTs were generally enriched in testes and liver (e.g. 8.3–1,328-fold higher than in the brain). However, the most specific expression in testes was observed in tNT-2 (1,328–3,502 fold higher than other tissues), a homolog to Slc9c2, which is a human Na+/H+ exchanger. tNT-7 was the most abundantly expressed of the tNTs (Figure 1B). No expressed sequence tag (EST) for tNT-7 has been reported to date, however, it is predicted to be homologous to cysteine-rich secretory protein (CRISP) involved in sperm-egg fusion [16]. Most of the tNTs encoding proteins with MW values ranging from 6–389 kDa exhibited a broad range of similarity (19–100%) between the species (Table 1). Despite the absence of a matched mouse gene or EST, tNT-1 was identical to the predicted protein model, XP_001475034.3, and shared high sequence identity with rat Slco6d1 (~80%), suggesting that it may function as an ion transporter in testes. tNT-18 seems to encode a protein identical to NP001028651.1 encoded by Gm1516 in chromosome 3. tNT-18 is located 3Mbps away from Gm1516 in chromosome 3, indicating that Gm1516 and tNT-18 are paralogs encoding the same protein sequence. Many of these novel transcripts are predicted to encode functional domains or highly homologous proteins in other species, as well (Table 1). Conversely, two testes-enriched novel transcripts (tNT-10 and −22) likely represented noncoding transcripts. Noncoding transcripts are also important regulatory molecules involved in diverse processes such as gene-specific transcription [17], regulation of basal transcriptional machinery [18], splicing [19], and translation [20]. The in-depth functional characterization of the confirmed testes- and liver-enriched novel transcripts is expected to lead to important information regarding tissue-specific gene regulation.

Experimental confirmation of testes-enriched novel exons

Among 197 testis-enriched novel exons, 26 novel exons were selected for experimental validation, on the basis of their read number (expression level), easiness of primer design, and straightforward exon structures. Among the 26 testes-enriched novel exons (hereafter, tNE), the strong enrichment of 24 tNEs in testes was confirmed by qRT-PCR and RT-PCR (Figure 2A and B). tNE-17 of Ms4a5 was the most abundantly and specifically expressed in testes, whereas tNE-6 was barely expressed in testis. tNE-1, −13, −15 and −22 were strongly expressed in testes, whereas little or no expression was observed in other tissues. Multiple novel exons were identified for Eya4 (tNE-2, −12 and −26), Fam71d (tNE-4, −5 and −7) and Pkm2 (tNE-10 and −21). We further examined the expression of the genes containing tNEs to determine whether the expression was due to testes-specific genes. Results indicated that most of the genes containing tNEs were ubiquitously expressed in different tissues (Figure 2C and D). However, the expressions of genes such as Skp2, Eya4, Scamp2, and Zfp385a were significantly lower in testes than in the brain (Figure 2D), despite strong expression of the tNEs (i.e., tNE-2, −3, −9, −12, −18 and −26) in testes, while the strong expressions of tNEs of Fam71d, Ms4a5 and 1700025F22Rik were assumed to be due to the testis-specific expression of the genes.
Figure 2

Testes-enriched novel exons. The expressions of testes-enriched novel exons (tNEs) were experimentally confirmed by (A) RT-PCR and (B) qRT-PCR. Blue circles indicate the tNEs confirmed by RT-PCR. The expression levels of the genes containing tNEs were measured by (C) RT-PCR and (D) qRT-PCR.

Testes-enriched novel exons. The expressions of testes-enriched novel exons (tNEs) were experimentally confirmed by (A) RT-PCR and (B) qRT-PCR. Blue circles indicate the tNEs confirmed by RT-PCR. The expression levels of the genes containing tNEs were measured by (C) RT-PCR and (D) qRT-PCR. We hypothesized that the insertion of novel exons could produce new UTRs or protein variants, as listed in Table 2. More than half of the testes-enriched novel exons (n = 112, 56.8%) were identified as alternative 5′-UTRs that would likely result in the differential regulation of transcription or translation in testes. Several studies have demonstrated that testes-specific 5′-UTRs include regulatory elements, such as the upstream open reading frames (uORFs), for translational regulation [21-23]. We also found that testes-enriched novel 5′-UTRs have abundant uORFs (n = 56, 50%) with some of 197 novel exons in the testes, suggesting a testes-specific regulatory role in translation. For example, more than 5 uORFs were found to be the testis-enriched 5′-UTRs of Nt5c2, Lrrc8b, Mllt11, Mphosph9, Kdm5b, Proca1, and 5730559C18Rik in 197 testis-enriched novel exons. Additionally, the inclusion of tNE-2, 3, 20 and 21 of Eya4, Skp2, Higd1a and Pkm2 could contribute to the 5′UTRs forming G-quadruplex, which is involved in translational control [24].
Table 2

Summary of testis or heart-enriched novel exons

IDGenePosition (mm9)LengthProteinPrediction 1
hNE-1Cluhchr11:74467029-74467303275+Q5SW1938 AA shorter
hNE-2Mylk4chr13:32820204-32820624421-Q5SUV5
hNE-3Clasp1chr1:120451862-120451963102+
hNE-4Schip1chr3:68388089-68388346258+Q3TI5327 AA different
hNE-5Mylk4chr13:32868030-32868501472+Q5SUV585 AA different
hNE-6Mylk4chr13:32818712-32819001290+Q5SUV5
hNE-7Clasp1chr1:120378050-120378669620+
hNE-8Trdnchr10:33086092-330872141123+Same as 51 kDa skeletal Trdn
hNE-9Sorbs1chr19:40452144-40452869726-Q62417241 AA longer
hNE-10Csde1chr3:102840498-102840644147+Q91W5046 AA longer
hNE-11Larp5chr13:9127241-91303703130+Q80UQ3105 AA longer
hNE-12Nedd5lchr18:65243095-652444481354+
hNE-13Nexnchr3:151927873-151928180308-Q7TPW1
tNE-1Lrrc8bchr5:105881814-1058831281315+Q5DU41Different 5′UTR
tNE-2Eya4chr10:22905057-22905389333-Q9Z191Different 5′UTR
tNE-3Skp2chr15:9082539-9082780242-Q9Z0Z3Different 5′UTR
tNE-4Fam71dchr12:79824797-79824939143+D3YV9226 shorter AA, different C term
tNE-5Fam71dchr12:79796826-79797030205+D3YV92Different 5′UTR
tNE-6Rfx1chr8:86608465-86608794330+
tNE-7Fam71dchr12:79823117-79823280164+D3YV92
tNE-8Pfkmchr15:97925522-97925641120+Q1LZL770 A.A longer
tNE-9Scamp2chr9:57426081-57426209129+Q9ERN044 AA longer
tNE-10Pkm2chr9:59510847-59510963117+P52480Different 5′UTR
tNE-11Vapachr17:65936384-65936506123-Q9WV5541 AA longer
tNE-12Eya4chr10:22903219-22903421203-Q9Z191Different 5′UTR
tNE-13Efr3achr15:65696232-65696453222+Q8BG67131 AA shorter (C-term)
tNE-141700001C02Rikchr5:30779031-30779154124+Q9DAS2N-term 15 AA
tNE-15Mtmr6chr14:60909543-60909656114+Q8VE1138 AA longer
tNE-16Mbtd1chr11:93800835-93801026192+
tNE-17Ms4a5chr19:11352451-11352587137-Q810P8C-term 97 AA shorter
tNE-18Zfp385achr15:103151501-103151619119-Q8VD12N-term 50 AA shorter
tNE-191700025F22Rikchr19:11233536-11233685150-Q6P8I056 AA longer
tNE-20Higd1achr9:121765839-121765990152-Q9JLR9Different 5′UTR
tNE-21Pkm2chr9:59506806-59506960155+P52480Different 5′UTR
tNE-22Dnahc2chr11:69331069-69331230162-Q9P22554 AA longer
tNE-231700006A11Rikchr3:124105142-124105398257-B9EHI371 AA longer (C-term)
tNE-24Zfand6chr7:91790796-91790947152-Q9DCH6Different 5′UTR
tNE-25Pcbp2chr15:102303428-102303532105+Q61990Different 5′UTR
tNE-26Eya4chr10:22902544-2290263087-Q9Z19131 AA shorter

1Changes of amino acid sequences and UTRs due to insertions of novel exons were predicted using the free software ‘Translate’ provided by ExPASy [39].

Summary of testis or heart-enriched novel exons 1Changes of amino acid sequences and UTRs due to insertions of novel exons were predicted using the free software ‘Translate’ provided by ExPASy [39]. Insertions of tNEs may lead to dramatic changes in protein expression. (Example 1) The C-terminal truncation (~50%) of MS4A5 is related to the insertion of tNE-17. MS4A5 is known to have four membrane-spanning domains [25], but insertion of tNE-17 results in the loss of two domains. (Example 2) Prediction by cNLS mapper [26] suggests that the novel isoform of EFR3A lacks the C-terminal 131 residue sequence containing one of the nuclear localization signals (NLSs). It is also possible that tNE-13 plays an important role in the regulation of EFR3A localization in testes [27]. (Example 3) For tNE-11 belonging to Vapa, two variants showing a 9-bp difference were identified by Cufflinks (Additional file 3: Figure S1A) and were predicted to encode 38–41 additional amino acids, GKTPPGIASTVASLSSVSSAVATPASYHLKNDPRELKE (VKQ). Interestingly, it is likely that this sequence contributes to the membrane-spanning region in a testes-specific manner by the prediction using TopPred [27] (Additional file 3: Figure S1B). The function of VAPA in neurons is known to be associated with ER and microtubules [28], and tNE-11 might confer testes-specific functions via the membrane-spanning region. Collectively, these data suggest that the testes-enriched novel exons could be involved in dramatic structural changes.

Experimental confirmation of the heart-specific novel exons

Among 26 heart-enriched novel exons, 13 novel exons (hereafter, hNE) were selected for experimental validation, on the basis of their read number (expression level), easiness of primer design, and straightforward exon structures and the enrichment of 10 hNEs in heart was experimentally confirmed by qRT-PCR and RT-PCR (Figure 3C and D). Most hNEs were strongly expressed in the heart, except for hNE-1 and −9. Multiple novel exons (i.e., hNE-2, −5 and −6) were identified in Mylk4 and predicted to produce a different 5′UTR with slightly different N-terminal regions. Similar to the tNEs, the alternative 5′-UTRs containing 1–2 uORFs were observed in the hNEs for Cluh, Mylk4, Schip1, Larp5, and Nexn, suggesting heart-specific post-transcriptional regulation.
Figure 3

Heart-enriched novel exons. The expressions of heart-enriched novel exons (hNEs) were experimentally confirmed by (A) RT-PCR and (B) qRT-PCR. The expression levels of the genes containing the hNEs were measured by (C) RT-PCR and (D) qRT-PCR.

Heart-enriched novel exons. The expressions of heart-enriched novel exons (hNEs) were experimentally confirmed by (A) RT-PCR and (B) qRT-PCR. The expression levels of the genes containing the hNEs were measured by (C) RT-PCR and (D) qRT-PCR. Among the variants identified, hNE-8 of Trdn is likely to result in truncation of the C-terminal region. A total of six isoforms were identified for Trdn, and their estimated sizes were approximately 1.3, 4.3, and 5 kb in the heart, and 5, 5.5, and 7 kb in skeletal muscle [29]. In addition, hNE-8 was specifically expressed in the heart and inserted in the transcripts expressed in skeletal muscle, which could result in the C-terminus-truncated TRDN. Based on analysis of data using Cufflinks, the relative expression of the isoform containing hNE-8 was predicted to be considerably lower than the known cardiac-specific isoforms (Additional file 4), suggesting a restricted role for hNE-8 of Trdn in the heart. Several dramatic changes were predicted in the case of Sorbs1 variants containing hNE-9. This predicted additional exon was highly enriched in proline residues such as PPPAPPPDPP, PPCLPFP, PKPYIPPSTP, and PSLPTPTSVP. Proline-rich residues was known to be important for binding the SH3 domains in signaling cascades [30, 31], therefore it suggested the insertion of hNE-9 might be involved in the regulation of signaling cascade in a heart-specific manner. At present, seven known isoforms of Sorbs1 have been identified [32-34] and hNE-9 is novel and appears highly enriched in the heart. Additionally, data suggested that hNE-10 from Csde1 likely encoded a serine-rich region consisting of 46 additional residues (MENMLTVSSDPQPTPAAPPSLSLPLSSSSTSSWTKKQKRTPTYQRS). Interestingly, Ser-32 and Thr-34 of hNE-10 were predicted to be phosphorylated by PKC according to NetPhosK [35], suggesting heart-specific signal regulation.

Alternative splicing patterns of the novel isoforms containing the novel exons

We then compared the expression levels of the novel isoforms containing the novel exons to those of the known isoforms. As seen in Figure 4, at least 10 novel isoforms exhibited dominant expression when compared with the previously known isoforms in the heart or testes. More than 90% of the expressions of Scamp2, Vapa, Zfp385a, 1700001C02Rik, Fam71d, 1700025F22Rik, and Mtmr6 were identified in the novel isoforms in testes, suggesting testes-specific roles of the isoforms. For Mtmr6, a recent study reported that the testes-specific MTMR6 protein had a slightly higher molecular weight than the known protein, but the similar portions of the novel and known isoforms were observed at a protein level [36].
Figure 4

Alternative splicing patterns of the novel isoforms containing the novel exons. (A) Multiple isoforms either with or without the novel exons were produced by RT-PCR. Exon structures amplified by the primers are seen in the right panel, while novel exons are highlighted with blue boxes. (B) The relative expression levels of the isoforms were quantified by ImageJ.

Alternative splicing patterns of the novel isoforms containing the novel exons. (A) Multiple isoforms either with or without the novel exons were produced by RT-PCR. Exon structures amplified by the primers are seen in the right panel, while novel exons are highlighted with blue boxes. (B) The relative expression levels of the isoforms were quantified by ImageJ. Conversely, the novel isoforms for Pkm2, Ms4a5, Pfkm, and Skp2 were expressed at relatively low levels in the testes. Unexpected isoforms were observed in Zfp385a, 1700001C02Rik, Fam71d, Mtmr6, and Zfand6, implying incomplete coverage in spite of the high-resolution of NGS. However, the rapidly accumulating datasets will help complete a mouse gene annotation.

Expressional changes of heart-specific novel exons during cardiac hypertrophy

For the identified hNEs, we investigated the alternative splicing patterns occurring during cardiac hypertrophy induced by transverse aortic constriction (TAC). The number of reads mapped to all exons, including hNEs, were calculated using our RNA-Seq dataset (E-MTAB-727) on cardiac hypertrophy [37], and the differential expression levels of hNEs were identified using DEXSeq [38]. Two differentially expressed hNEs (hNE-1 and −9 for Cluh and Sorbs1, respectively) were obtained (p < 0.05) (Additional file 5: Table S3) from the analysis. As seen in Figure 5A, the expression of Cluh was significantly decreased by ~36% (p = 0.015), while the expression of hNE-1 decreased by ~65% during cardiac hypertrophy (p = 0.007), indicating that hNE-1 in Cluh was alternatively spliced during cardiac hypertrophy (gene vs. hNE-1 in TAC, p = 0.046). Cufflinks analysis (Additional file 6: Figure S3) indicated that the portion of the novel isoform containing hNE-1 represented approximately 33% of the expression of Cluh in the heart, and that the predicted protein derived from the isoform was 38 residues shorter than the known isoform. The expression of the heart-specific minor isoform containing hNE-1 was thought to be down-regulated during cardiac hypertrophy.
Figure 5

Alternative splicing patterns of the novel exons of and during cardiac hypertrophy. (A) The expression levels of Cluh and hNE-1, (B) Sorbs1 and hNE-9, measured by qRT-PCR. Black and gray bars indicate the expression levels in Sham and TAC (n = 3 for each model), respectively, and (C) expression patterns measured by RT-PCR.

Alternative splicing patterns of the novel exons of and during cardiac hypertrophy. (A) The expression levels of Cluh and hNE-1, (B) Sorbs1 and hNE-9, measured by qRT-PCR. Black and gray bars indicate the expression levels in Sham and TAC (n = 3 for each model), respectively, and (C) expression patterns measured by RT-PCR. The expression of hNE-9 in Sorbs1 was also significantly decreased during cardiac hypertrophy. While the expression of Sorbs1 gene was not changed (p = 0.34), the expression of hNE-9 was significantly decreased by ~36% (p = 0.037) (Figure 5B). Thus, hNE-9 was thought to be excluded during cardiac hypertrophy suggesting a disease-related function associated with hNE-9 in the heart. Therefore, we examined the relationship between cardiac hypertrophy and hNEs, and further experimentally validated the significant exclusion of hNE-1 and −9 of Cluh and Sorbs1 during TAC-induced cardiac hypertrophy. As no changes were indicated in exercise-induced cardiac hypertrophy (Table S3), we concluded that the exclusion of these exons could be related to pathology of the heart.

Conclusions

The results of this study will contribute to updating mouse gene annotation through the identification of specific tissue-enriched novel transcripts and novel exons. Tissue-specific isoform switches mediated by novel exons could provide important insights into the tissue-specific roles of the novel exons. The exclusion of the hNEs during cardiac hypertrophy also suggested sensitivity of the novel exons to pathological status. Our findings emphasize the necessity of this approach to identify tissue-specific novel transcripts and exons.

Methods

Ethics Statement

All animal experiments and animal ethics were approved by the GIST Institutional Animal Care and Use Committee (IACUC) (Permit number: GIST-2013-22).

Identification of novel transcripts and exons

‘Tophat-Cufflinks-Cuffompare’ pipeline was used to identify novel transcripts and exons. Fastq files for six mouse tissues (brain, cerebrum, kidney, heart, liver and testis) in GSE30352 were downloaded and the reads were further aligned to mouse genome (UCSC mm9 version) using ‘Tophat’. Using resultant Bam files, de novo assembly was performed to construct the transcripts using ‘Cufflinks’. All transcripts were then compared to the predefined gene annotations such as UCSC and ENSEMBL using ‘Cuffcompare’. To identify the novel transcripts, we collected the transcripts located in intergenic region and classified as “unknown” in both UCSC and ENSEMBL as putative novel transcripts. In case of novel exons, we searched the consecutive novel junctions in-between the known neighbouring exons, thereby deduced the novel exons spanning the novel junctions.

Tissue-specificity of novel transcripts and exons

Numbers of the reads for the novel transcripts and exons were counted for heart, testis, liver, kidney, cerebrum and brain using HTSeq [39]. Multiple tests for the novel transcripts and exons were performed for a specific tissue vs. remaining tissues using DESeq and DEXSeq [14, 38], respectively, in a pairwise manner. The novel transcripts and exons significantly enriched in a specific tissue compared to all other tissues (p < 0.05) were collected.

Alternative splicing of heart-enriched novel exons during cardiac hypertrophy

Differential expression of the heart-enriched novel exons during cardiac hypertrophy were analysed using ‘Tophat-HTSeq-DEXSeq’ pipeline. Fastq files of previously reported RNA-Seq datasets on TAC-induced cardiac hypertrophy were aligned to mouse genome (UCSC mm9 version) using Tophat [12]. Number of the reads mapped onto the genes containing the heart-enriched novel exons were counted using HTSeq. Then, the differential expression of the heart-enriched exons during cardiac hypertrophy were analysed using DEXSeq (p < 0.05).

Transverse aortic constriction operation

Cardiac hypertrophy was induced by TAC operation under anesthesia with intraperitoneal injection of avertin, 2-2-2 tribromoethanol (Sigma, St. Louis, MO) dissolved in tert-amyl alcohol (Sigma, St. Louis, MO). The procedure of operation was followed as previously described [37]. As a control group, sham operation (same procedure except for tying) was done. 1 week after operation, mice were sacrificed, and hearts were removed, and then stored in deep freezer at -80°C before RNA extraction.

Tissue preparation and RNA isolation

Adult (8 weeks old) C57BL6 mouse heart, testes, liver, kidney, brain and lung were snap frozen in liquid nitrogen, stores at -80°C, and homogenized in liquid nitrogen using a mortar and pestle. Approximately 450–700 mg of grinded whole mouse heart was used for extraction of total RNA with 1 ml Trizol Reagent® (Invitrogen, Carlsbad, CA) following the manufacturer’s instructions.

RT-PCR and qRT-PCR

First-strand cDNA was synthesized from 2 μg of total RNA with Random hexamer using Omniscript® reverse transcription (Qiagen, Valencia, CA) according to the manufacturer’s instruction. Briefly, qRT-PCR assays were performed using TOPreal™ qPCR premix (Enzynomics, Korea) under the following two-step conditions: denaturation at 95°C for 15 seconds followed by annealing and extension at 60°C for 40 seconds, for a total of 40 cycles. The 18S transcript was used as an endogenous reference to assess the relative level of mRNA transcript. RT-PCR assays were performed on a ABI thermal cycler TP600 (TaKaRa, Japan) using nTaq-HOT DNA polymerase (Enzynomics, Daejeon, South Korea) under the following 3 step conditions: denaturation at 94°C for 30s, annealing at 55-60°C for 30s and extension at 72°C for 40s with total 35–37 cycles. All primer pairs are listed in Additional file 7: Table S4.

Availability

GSE30352 and E-MTAB-727 are publicly available in the Gene Expression Omnibus (GEO) and European Nucleotide Archive (ENA) databases, respectively. Additional file 1: Table S1: Tissue-specific novel transcripts. Detailed information on 184 novel transcripts are listed. The novel transcripts were identified by the pipeline of ‘Tophat-Cufflinks-Cuffcompare’. Multiple isoforms transcribed from the same transcribed loci are in the list. (XLSX 30 KB) Additional file 2: Table S2: List for the number of novel exons for the 6 tissues. (XLSX 10 KB) Additional file 3: Figure S1: Structure of Vapa (A) Structures of Vapa and magnified image for the isoforms were illustrated using UCSC Genome browser (B) Predicted hydrophobicity of the novel exon of Vapa suggest the membrane spanning ability. TopPred was applied to predict hydrophobicity. (PNG 209 KB) Additional file 4: Figure S2: Expression level of hNE-9 estimated by Cufflinks. (PNG 32 KB) Additional file 5: Table S3: Expression levels of heart-specific novel exons (hNEs) during cardiac hypertrophy. Alternative splicing of hNEs during either transverse aortic constriction (TAC) or exercise-induced cardiac hypertrophy was determined using DEXSeq. (XLSX 10 KB) Additional file 6: Figure S3: Relative expression levels of the isoforms for Cluh and Sorbs1 in heart. Relative expression levels of the isoforms were measured by FPKM of Cufflinks. Green bars indicate the expression levels of the isoforms containing novel exons. (PNG 32 KB) Additional file 7: Table S4: Primers used. (XLSX 21 KB)
  39 in total

1.  The evolution of gene expression levels in mammalian organs.

Authors:  David Brawand; Magali Soumillon; Anamaria Necsulea; Philippe Julien; Gábor Csárdi; Patrick Harrigan; Manuela Weier; Angélica Liechti; Ayinuer Aximu-Petri; Martin Kircher; Frank W Albert; Ulrich Zeller; Philipp Khaitovich; Frank Grützner; Sven Bergmann; Rasmus Nielsen; Svante Pääbo; Henrik Kaessmann
Journal:  Nature       Date:  2011-10-19       Impact factor: 49.962

2.  Structural basis for the binding of proline-rich peptides to SH3 domains.

Authors:  H Yu; J K Chen; S Feng; D C Dalgarno; A W Brauer; S L Schreiber
Journal:  Cell       Date:  1994-03-11       Impact factor: 41.582

3.  Identification of novel transcripts in annotated genomes using RNA-Seq.

Authors:  Adam Roberts; Harold Pimentel; Cole Trapnell; Lior Pachter
Journal:  Bioinformatics       Date:  2011-06-21       Impact factor: 6.937

4.  Analysis of transcriptome complexity through RNA sequencing in normal and failing murine hearts.

Authors:  Jae-Hyung Lee; Chen Gao; Guangdun Peng; Christopher Greer; Shuxun Ren; Yibin Wang; Xinshu Xiao
Journal:  Circ Res       Date:  2011-10-27       Impact factor: 17.367

5.  Detecting differential usage of exons from RNA-seq data.

Authors:  Simon Anders; Alejandro Reyes; Wolfgang Huber
Journal:  Genome Res       Date:  2012-06-21       Impact factor: 9.043

6.  A novel BAT3 sequence generated by alternative RNA splicing of exon 11B displays cell type-specific expression and impacts on subcellular localization.

Authors:  Nadine Kämper; Jörg Kessler; Sebastian Temme; Claudia Wegscheid; Johannes Winkler; Norbert Koch
Journal:  PLoS One       Date:  2012-04-25       Impact factor: 3.240

Review 7.  5'-UTR RNA G-quadruplexes: translation regulation and targeting.

Authors:  Anthony Bugaut; Shankar Balasubramanian
Journal:  Nucleic Acids Res       Date:  2012-02-20       Impact factor: 16.971

8.  Revealing the missing expressed genes beyond the human reference genome by RNA-Seq.

Authors:  Geng Chen; Ruiyuan Li; Leming Shi; Junyi Qi; Pengzhan Hu; Jian Luo; Mingyao Liu; Tieliu Shi
Journal:  BMC Genomics       Date:  2011-12-02       Impact factor: 3.969

9.  Deep RNA sequencing reveals novel cardiac transcriptomic signatures for physiological and pathological hypertrophy.

Authors:  Hong Ki Song; Seong-Eui Hong; Taeyong Kim; Do Han Kim
Journal:  PLoS One       Date:  2012-04-16       Impact factor: 3.240

10.  Combining RT-PCR-seq and RNA-seq to catalog all genic elements encoded in the human genome.

Authors:  Cédric Howald; Andrea Tanzer; Jacqueline Chrast; Felix Kokocinski; Thomas Derrien; Nathalie Walters; Jose M Gonzalez; Adam Frankish; Bronwen L Aken; Thibaut Hourlier; Jan-Hinnerk Vogel; Simon White; Stephen Searle; Jennifer Harrow; Tim J Hubbard; Roderic Guigó; Alexandre Reymond
Journal:  Genome Res       Date:  2012-09       Impact factor: 9.043

View more
  6 in total

Review 1.  What to compare and how: Comparative transcriptomics for Evo-Devo.

Authors:  Julien Roux; Marta Rosikiewicz; Marc Robinson-Rechavi
Journal:  J Exp Zool B Mol Dev Evol       Date:  2015-04-10       Impact factor: 2.656

2.  Systematic identification and comparison of expressed profiles of lncRNAs and circRNAs with associated co-expression and ceRNA networks in mouse germline stem cells.

Authors:  Xiaoyong Li; Junping Ao; Ji Wu
Journal:  Oncotarget       Date:  2017-04-18

3.  Isoform Evolution in Primates through Independent Combination of Alternative RNA Processing Events.

Authors:  Shi-Jian Zhang; Chenqu Wang; Shouyu Yan; Aisi Fu; Xuke Luan; Yumei Li; Qing Sunny Shen; Xiaoming Zhong; Jia-Yu Chen; Xiangfeng Wang; Bertrand Chin-Ming Tan; Aibin He; Chuan-Yun Li
Journal:  Mol Biol Evol       Date:  2017-10-01       Impact factor: 16.240

4.  Relative Abundance of Transcripts ( RATs): Identifying differential isoform abundance from RNA-seq.

Authors:  Kimon Froussios; Kira Mourão; Gordon Simpson; Geoff Barton; Nicholas Schurch
Journal:  F1000Res       Date:  2019-02-24

5.  Integrated Analysis of Long Non-Coding RNA and mRNA Expression Profiles in Testes of Calves and Sexually Mature Wandong Bulls (Bos taurus).

Authors:  Hongyu Liu; Ibrar Muhammad Khan; Huiqun Yin; Xinqi Zhou; Muhammad Rizwan; Jingyi Zhuang; Yunhai Zhang
Journal:  Animals (Basel)       Date:  2021-07-05       Impact factor: 3.231

6.  Systematic Analysis of Long Non-Coding RNAs and mRNAs in the Ovaries of Duroc Pigs During Different Follicular Stages Using RNA Sequencing.

Authors:  Yi Liu; Mengxun Li; Xinwen Bo; Tao Li; Lipeng Ma; Tenjiao Zhai; Tao Huang
Journal:  Int J Mol Sci       Date:  2018-06-11       Impact factor: 5.923

  6 in total

北京卡尤迪生物科技股份有限公司 © 2022-2023.