• Home
  • Search
  • MGcount: a total RNA-seq quantification tool to address multi-mapping and multi-overlapping alignments ambiguity in non-coding transcripts
  • Open Access IconOpen Access
  • Cite Icon17
  • https://doi.org/10.1186/s12859-021-04544-3Copy DOI Icon

MGcount: a total RNA-seq quantification tool to address multi-mapping and multi-overlapping alignments ambiguity in non-coding transcripts

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

BackgroundTotal-RNA sequencing (total-RNA-seq) allows the simultaneous study of both the coding and the non-coding transcriptome. Yet, computational pipelines have traditionally focused on particular biotypes, making assumptions that are not fullfilled by total-RNA-seq datasets. Transcripts from distinct RNA biotypes vary in length, biogenesis, and function, can overlap in a genomic region, and may be present in the genome with a high copy number. Consequently, reads from total-RNA-seq libraries may cause ambiguous genomic alignments, demanding for flexible quantification approaches.ResultsHere we present Multi-Graph count (MGcount), a total-RNA-seq quantification tool combining two strategies for handling ambiguous alignments. First, MGcount assigns reads hierarchically to small-RNA and long-RNA features to account for length disparity when transcripts overlap in the same genomic position. Next, MGcount aggregates RNA products with similar sequences where reads systematically multi-map using a graph-based approach. MGcount outputs a transcriptomic count matrix compatible with RNA-sequencing downstream analysis pipelines, with both bulk and single-cell resolution, and the graphs that model repeated transcript structures for different biotypes. The software can be used as a python module or as a single-file executable program.ConclusionsMGcount is a flexible total-RNA-seq quantification tool that successfully integrates reads that align to multiple genomic locations or that overlap with multiple gene features. Its approach is suitable for the simultaneous estimation of protein-coding, long non-coding and small non-coding transcript concentration, in both precursor and processed forms. Both source code and compiled software are available at https://github.com/hitaandrea/MGcount.

Loading PDF

Similar Papers
  • Research Article
  • Citations1

Object Stitching by Clustering of Adjacent Regions for accurate quantification of three-dimensional tissues.

  • Sep 15, 2025
  • Journal of cell science
  • Mario Ledesma-Terrón +5
  • PDF
  • Research Article
  • Citations3

Organization and expression analysis of 5S and 45S ribosomal DNA clusters in autotetraploid Carassius auratus

  • Nov 05, 2021
  • BMC Ecology and Evolution
  • Chun Zhao +10
  • Research Article
  • Citations38

Technical Note: In silico imaging tools from the VICTRE clinical trial.

  • Jul 17, 2019
  • Medical Physics
  • Diksha Sharma +7
  • PDF
  • Research Article
  • Citations116

Identification of large intergenic non-coding RNAs in bovine muscle using next-generation transcriptomic sequencing

  • Jun 19, 2014
  • BMC Genomics
  • Coline Billerey +8
  • Research Article
  • Citations84

Whole-Cell Impedance Analysis for Highly and Poorly Metastatic Cancer Cells

  • Aug 01, 2009
  • Journal of Microelectromechanical Systems
  • Younghak Cho +5
  • Research Article
  • Citations1

HISSTA: a human in situ single-cell transcriptome atlas.

  • Mar 29, 2025
  • Bioinformatics (Oxford, England)
  • Jiwon Yu +10
  • Research Article

The Genetic and Phylogenetic Analysis of the D-Loop Region in Mitochondrial Genome of Najdi Goat

  • Oct 22, 2020
  • SHILAP Revista de lepidopterología
  • Ameneh Bashiri +2
  • Peer Review Report

Decision letter: Epigenetic conservation at gene regulatory elements revealed by non-methylated DNA profiling in seven vertebrates

  • Dec 10, 2012
  • Anne Ferguson-Smith
  • Research Article

Disc-Hub: a python package for benchmarking machine learning strategies in DIA-MS identification

  • Sep 30, 2025
  • Bioinformatics Advances
  • Yiwen Yu +2
  • Research Article
  • Citations3

Detection and classification of long terminal repeat sequences in plant LTR-retrotransposons and their analysis using explainable machine learning

  • Dec 18, 2024
  • BioData Mining
  • Jakub Horvath +5
  • Research Article
  • Citations9

Target site analysis of RTE1_LA and its AfroSINE partner in the elephant genome

  • Aug 28, 2008
  • Gene
  • Clément Gilbert +2
  • Research Article
  • Citations78

MetaQuant: a tool for the automatic quantification of GC/MS-based metabolome data

  • Oct 17, 2006
  • Bioinformatics
  • Boyke Bunk +6
  • Research Article
  • Citations21

Oarfish: enhanced probabilistic modeling leads to improved accuracy in long read transcriptome quantification

  • Jul 01, 2025
  • Bioinformatics
  • Zahra Zare Jousheghani +2
  • Research Article
  • Citations11

Spatial Analysis and Modeling Tool Version 2 (SAMT2), a spatial modeling tool kit written in Python

  • Aug 08, 2015
  • Ecological Informatics
  • Ralf Wieland +3
  • Book Chapter
  • Citations5

Deep Cap Analysis of Gene Expression (CAGE): Genome-Wide Identification of Promoters, Quantification of Their Activity, and Transcriptional Network Inference

  • Jan 01, 2017
  • Alexandre Fort +1
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.