Back to search

Article

Generation of ENSEMBL-based proteogenomics databases boosts the identification of non-canonical peptides

2021-06-09

Abstract excerpt

<h4>Summary</h4> We have implemented the pypgatk package and the pgdb workflow to create proteogenomics databases based on ENSEMBL resources. The tools allow the generation of protein sequences from novel protein-coding transcripts by performing a three-frame translation of pseudogenes, lncRNAs, and other non-canonical transcripts, such as those produced by alternative splicing events. It also includes exonic o...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
2008cbc0-c6e3-5eae-b1e0-6e16367edd75
DOI
10.1101/2021.06.08.447496
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Generation of ENSEMBL-based proteogenomics databases boosts the identification of non-canonical peptidesDOI 10.1101/2021.06.08.447496
Select a neighboring publication to make it the new centre.