Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 SCOPE++: sequence classification of homoPolymer emissions.

Literature DB >> 25087770

SCOPE++: sequence classification of homoPolymer emissions.

James T Morton¹, Patricia Abrudan², Nathanial Figueroa³, Chun Liang⁴, John E Karro⁵.

Abstract

BACKGROUND: mRNA polyadenylation, the addition of a poly(A) tail to the 3'-end of pre-mRNA, is a process critical to gene expression and regulation in eukaryotes. To understand the molecular mechanisms governing polyadenylation and other relevant biological processes, it is important to identify these poly(A) tails accurately in transcriptome sequencing data and differentiate them from artificial adapter sequences added in the sequencing process. But the annotation of these tails is complicated by the presence of sequencing errors and post-transcriptional modifications. While determining that a tail is present in a given transcript fragment is straight-forward, these obfuscations make the problem of boundary identification a challenge; conventional seed-and-extend algorithms struggle to accurately identify these poly(A) tail end-points. Further, all existing tools that we are aware of focus exclusively on the trimming of poly(A) tails, failing to provide the detailed information needed for studying the polyadenylation process.
RESULTS: We have created SCOPE++, an open-source tool for finding the precise border of poly(A) tails and other homopolymers in raw mRNA sequence reads. Based on a Hidden Markov Model (HMM) approach, SCOPE++ accurately identifies specific homopolymer sequences in error-prone EST/cDNA data or RNA-Seq data at a speed appropriate for large sequence sets.
CONCLUSIONS: We demonstrate that our tool can precisely identify poly(A) tails with near perfect accuracy at the speed required for high-throughput applications, providing a valuable resource for polyadenylation research.

Entities: Chemical Disease Gene Species

Keywords: Hidden Markov Model; Polyadenylation; Transcriptome

Mesh：

Substances：
RNA, Messenger

Year: 2014 PMID： 25087770 PMCID： PMC4165746 DOI： 10.1016/j.ygeno.2014.07.005

Source DB: PubMed Journal: Genomics ISSN： 0888-7543 Impact factor: 5.736

Keyword Cloud
References

16 in total

SCOPE++: sequence classification of homoPolymer emissions.

1. EMBOSS: the European Molecular Biology Open Software Suite.

Review 2. Translational control by CPEB: a means to the end.

3. Nontemplated nucleotide addition prior to polyadenylation: a comparison of Arabidopsis cDNA and genomic sequences.

4. Novel extraction strategy of ribosomal RNA and genomic DNA from cheese for PCR-based investigations.

5. Genome-wide landscape of polyadenylation in Arabidopsis provides evidence for extensive alternative polyadenylation.

Review 6. Ending the message: poly(A) signals then and now.

7. Comprehensive polyadenylation site maps in yeast and human reveal pervasive alternative polyadenylation.

8. SeqTrim: a high-throughput pipeline for pre-processing any type of sequence read.

9. Widespread shortening of 3'UTRs by alternative cleavage and polyadenylation activates oncogenes in cancer cells.

10. Poly(A)-tail profiling reveals an embryonic switch in translational control.