ProteoDisco: A flexible R approach to generate customized protein databases for extended search space of novel and variant proteins in proteogenomic studies

preprint OA: closed CC-BY-4.0
📄 Open PDF View at publisher
AI-generated summary by claude@2026-07, 2026-07-16

ProteoDisco is an R package that enables the flexible creation of custom protein databases by incorporating genomic variants, fusion genes, and transcriptomic variants for proteogenomic studies.

One-sentence paraphrase of the abstract; not a substitute for reading it. No clinical advice. How this works

Abstract

Summary We present an R-based open-source software termed ProteoDisco that allows for flexible incorporation of genomic variants, fusion-genes and (aberrant) transcriptomic variants from standardized formats into protein variant sequences. ProteoDisco allows for a flexible step-by-step workflow allowing for in-depth customization to suit a myriad of research approaches in the field of proteogenomics, on all organisms for which a reference genome and transcript annotations are available. Availability and Implementation ProteoDisco (R package version ≥ 0.99) is available from https://github.com/ErasmusMC-CCBC/ProteoDisco/ . Contact [email protected] Supplementary information Supplementary table, figures and data files available.

My notes (saved in your browser only)

Citation neighborhood (no data yet)

We don't have any in-corpus citations linked to this paper yet. The paper's references may be in our DB but unresolved to ``paper_id`` (resolution happens at ingest when the cited DOI matches a row we already have). Run the cross-source citation reconcile pass to retry.

Source provenance

europepmc
last seen: 2026-05-19T01:45:01.086888+00:00
unpaywall
last seen: 2026-05-24T02:00:01.246996+00:00
License: CC-BY-4.0