bioRxiv Science⌕ Search

Biology subjects

Paniagua, A.

Publications and source records attributed to Paniagua, A..

3 recordsLinked to original sources

SQANTI-browser: visualization and curation of SQANTI3-classified long-read transcriptomes within the UCSC Genome Browser

Long-read sequencing enables transcriptome-wide isoform discovery. However, it generates substantial technical and structural ambiguity that complicates transcript interpretation. Here, we present SQANTI-browser, a classification-aware visualization framework that converts SQANTI3 outputs into interactive UCSC Genome Browser Track Hubs, preserving full transcript structural metadata. By integrating SQANTI classifications directly within the UCSC ecosystem, SQANTI-browser enables dynamic filtering and evidence-guided curation alongside public resource tracks. Furthermore, its adaptive architecture natively supports non-reference genomes, orthogonal data, and custom metadata fields. Applied to clinical, noisy, and synthetic datasets, SQANTI-browser resolves alignment artifacts and rescues actionable novel isoforms, providing a robust framework for long-read transcriptome curation.

bioinformatics↗

To join or not to join: handling biological replicates in long-read RNA sequencing data

Long-read RNA sequencing (lrRNA-seq) has revolutionized transcriptomics facilitating the study of alternative splicing and resulting in identification of thousands of novel transcripts. While isoform identification has received significant attention, the handling of biologically replicated lrRNA-seq datasets remains less explored. However, how multiple samples are combined in a lrRNA-seq study may strongly impact transcript identification. This study defines and evaluates two strategies for obtaining consensus transcriptomes from multi-sample lrRNA-seq data: "Join & Call", where reads from all samples are combined before transcript identification, and "Call & Join", where transcript identification is performed on individual samples before combining the resulting annotations. We applied these strategies to a highly replicated dataset of mouse brain and kidney tissues, using both PacBio and ONT technologies, across six widely used transcript reconstruction tools. Our results indicate that the optimal strategy depends on the chosen computational tool and research objective. We found that Join & Call is generally more suitable for discovering rarely occurring, novel isoforms, as pooling evidence increases confidence in calling lowly-expressed transcripts. Conversely, Call & Join is computationally more efficient and often preferable for highly replicated datasets when the investigation of rare novel transcripts is not the primary objective. Our findings provide a conceptual and practical framework for multi-sample transcriptome reconstruction, guiding best practices in the context of increasingly large-scale lrRNA-seq studies.

bioinformatics↗

Transcriptome Universal Single-isoform COntrol: A Framework for Evaluating Transcriptome reconstruction Quality

Long-read sequencing (LRS) platforms, such as Oxford Nanopore and Pacific Biosciences, enable comprehensive transcriptome analysis but face challenges such as sequencing errors, sample quality variability, and library preparation biases. Current benchmarking approaches address these issues insufficiently: BUSCO assesses transcriptome completeness using conserved single-copy orthologs but can misinterpret alternative splicing as gene duplications, while spike-ins (SIRVs, ERCCs) oversimplify real- sample complexity, neglecting RNA degradation and RNA extraction artifacts, thus inflating performance metrics. Simulation algorithms are limited to recapitulate this complexity. To overcome these limitations, we introduce the Transcriptome Universal Single-isoform Control (TUSCO), a curated internal reference set of genes lacking alternative isoforms. TUSCO evaluates precision by identifying transcripts deviating from reference annotations and assesses sensitivity by verifying detection completeness in human and mouse samples. Masking TUSCO transcripts--and optionally inserting decoy splice variants--creates a novel- isoform challenge that assesses recovery of the true, now-unannotated isoforms. Our validation demonstrates that TUSCO provides accurate and reliable benchmarking without external controls, significantly improving quality control standards for transcriptome reconstruction using LRS.

bioinformatics↗