paper-with-me

홈 › Papers

The status of the human gene catalogue

2023-03-24 · Paulo Amaral, Silvia Carbonell-Sala, Francisco M. De La Vega, Tiago Faial, Adam Frankish, Thomas Gingeras, Roderic Guigo, Jennifer L Harrow, Artemis G. Hatzigeorgiou, Rory Johnson, Terence D. Murphy, Mihaela Pertea, Kim D. Pruitt, Shashikant Pujar, Hazuki Takahashi, Igor Ulitsky, Ales Varabyou, Christine A. Wells, Mark Yandell, Piero Carninci, Steven L. Salzberg

Scientists have been trying to identify all of the genes in the human genome since the initial draft of the genome was published in 2001. Over the intervening years, much progress has been made in identifying protein-coding genes, and the estimated number has shrunk to fewer than 20,000, although the number of distinct protein-coding isoforms has expanded dramatically. The invention of high-throughput RNA sequencing and other technological breakthroughs have led to an explosion in the number of reported non-coding RNA genes, although most of them do not yet have any known function. A combination of recent advances offers a path forward to identifying these functions and towards eventually completing the human gene catalogue. However, much work remains to be done before we have a universal annotation standard that includes all medically significant genes, maintains their relationships with different reference genomes, and describes clinically relevant genetic variants.

📄 PDF Abstract BibTeX arXiv:2303.13996

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Masader: Metadata Sourcing for Arabic Text and Speech Data Resources

2021-10-13 · LREC 2022 6 · Zaid Alyafeai, Maraim Masoud, Mustafa Ghaleb, Maged S. Al-shaibani

The NLP pipeline has evolved dramatically in the last few years. The first step in the pipeline is to find suitable annotated datasets to evaluate the tasks we are trying to solve. Unfortunately, most of the published da…

Virtual Process Dossier: A Process-Aware Data Catalogue

2026-07-30 · Lukas Kubelka, Alexander Bott, Frank Döhner, Saksham Kiroriwal 외 arxiv

We propose the Virtual Process Dossier (VPD), a Knowledge Graph-based data catalogue that also captures workflow provenance. We developed VPD for multi-stage manufacturing use-cases where downstream AI-based optimization…

Galaxy classification: A machine learning analysis of GAMA catalogue data

2019-03-18 · Aleke Nolte, Lingyu Wang, Maciej Bilicki, Benne Holwerda 외

We present a machine learning analysis of five labelled galaxy catalogues from the Galaxy And Mass Assembly (GAMA): The SersicCatVIKING and SersicCatUKIDSS catalogues containing morphological features, the GaussFitSimple…

BIG-bench Machine LearningClassificationGeneral ClassificationQuantization

Guiding Catalogue Enrichment with User Queries

2024-06-11 · Yupei Du, Jacek Golebiowski, Philipp Schmidt, Ziawasch Abedjan

Techniques for knowledge graph (KGs) enrichment have been increasingly crucial for commercial applications that rely on evolving product catalogues. However, because of the huge search space of potential enrichment, pred…

Language Diversity: Visible to Humans, Exploitable by Machines

2022-03-09 · ACL 2022 5 · Gábor Bella, Erdenebileg Byambadorj, Yamini Chandrashekar, Khuyagbaatar Batsuren 외

The Universal Knowledge Core (UKC) is a large multilingual lexical database with a focus on language diversity and covering over a thousand languages. The aim of the database, as well as its tools and data catalogue, is …

Diversity