科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Nature Communications2025-12-03· Transcriptome

Long-read transcriptomics of a diverse human cohort reveals ancestry bias in gene annotation

Pau Clavell-Revelles, Fairlie Reese, Sílvia Carbonell Sala, Fabien Degalez, Carme Arnan, Winona Oliveros, Emilio Palumbo, Tamara Perteghella, Roderic Guigó, Marta Melé

原始摘要(英文原文)· Original abstract
Accurate gene annotations are fundamental for interpreting genetic variation, cellular function, and disease mechanisms. However, current human gene annotations are largely derived from transcriptomic data of individuals with European ancestry, leaving gaps of annotation that remain uncharacterized. Here, we generate over 800 million full-length reads with long-read RNA-seq in 43 lymphoblastoid cell line samples from eight genetically-diverse human populations and build a cross-ancestry gene annotation. We demonstrate that transcripts from non-European samples are underrepresented in reference gene annotations, leading to incomplete characterization in allele-specific transcript usage. Furthermore, we show that personal genome assemblies enhance transcript discovery compared to the generic GRCh38 reference assembly, even though genomic regions unique to each individual are heavily depleted of genes. These findings underscore the urgent need for a more inclusive gene annotation framework that accurately represents global transcriptome diversity. Current human gene annotations are predominantly based on individuals of European ancestry. Here, the authors use long-read RNA sequencing across diverse human populations to reveal ancestry-specific transcripts, highlighting the need for more inclusive gene annotation frameworks.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Long-read transcriptomics of a diverse human cohort reveals ancestry bias in gene annotation — 科研速览 Science Skim