k-mer-based Upstream Preprocessing of long reads for Isoform Discovery.
KuPID preprocessing enhances isoform discovery accuracy by up to 11.6 points and halves runtime in long-read RNA-seq analysis.
- Why it matters: Accurate annotation of splice junctions is essential for identifying novel isoforms, but current methods rely on time-consuming dynamic programming alignments, limiting efficiency and scalability.
- What they did: The authors developed KuPID, a k-mer sketching-based prefilter that rapidly pseudo-aligns long reads to known isoforms, reducing the need for full alignments to only the most relevant reads.
- The result: This approach accelerates isoform discovery pipelines and improves downstream accuracy, enabling more efficient and precise transcriptome analysis, with optional modes for combined isoform discovery and quantification.