paper-with-me

홈 › Papers

LAPIS is a fast web API for massive open virus sequencing databases

2022-06-02 · Chaoran Chen, Alexander Taepper, Fabian Engelniederhammer, Jonas Kellerer, Cornelius Roemer, Tanja Stadler

Background: Recent epidemic outbreaks such as the SARS-CoV-2 pandemic and the mpox outbreak in 2022 have demonstrated the value of genomic sequencing data for tracking the origin and spread of pathogens. Laboratories around the globe generated new sequences at unprecedented speed and volume and bioinformaticians developed new tools and dashboards to analyze this wealth of data. However, a major challenge that remains is the lack of simple and efficient approaches for accessing and processing sequencing data. Results: The Lightweight API for Sequences (LAPIS) facilitates rapid retrieval and analysis of genomic sequencing data through a REST API. It supports complex mutation- and metadata-based queries and can perform aggregation operations on massive datasets. LAPIS is optimized for typical questions relevant to genomic epidemiology. Using a newly-developed in-memory database engine, it has a high speed and throughput: between 25 January and 4 February 2023, the SARS-CoV-2 instance of LAPIS, which contains 14.5 million sequences, processed over 20 million requests with a mean response time of 411 ms and a median response time of 1 ms. LAPIS is the core engine behind our dashboards on genspectrum.org and we currently maintain public LAPIS instances for SARS-CoV-2 and mpox. Conclusions: Powered by an optimized database engine and available through a web API, LAPIS enhances the accessibility of genomic sequencing data. It is designed to serve as a common backend for dashboards and analyses with the potential to be integrated into common database platforms such as GenBank.

📄 PDF Abstract BibTeX arXiv:2206.01210

Code (1)

cevo-public/lapis 공식 구현

Tasks

EpidemiologyRetrieval

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

XVir: A Transformer-Based Architecture for Identifying Viral Reads from Cancer Samples

2023-08-28 · Shorya Consul, John Robertson, Haris Vikalo

It is estimated that approximately 15% of cancers worldwide can be linked to viral infections. The viruses that can cause or increase the risk of cancer include human papillomavirus, hepatitis B and C viruses, Epstein-Ba…

Diversity

MEEPTOOLS: A maximum expected error based FASTQ read filtering and trimming toolkit

2015-12-10

Next generation sequencing technology rapidly produces massive volume of data and quality control of this sequencing data is essential to any genomic analysis. Here we present MEEPTOOLS, which is a collection of open-sou…

Identification of Repurposable Drugs and Adverse Drug Reactions for Various Courses of COVID-19 Based on Single-Cell RNA Sequencing Data

2020-05-16

Coronavirus disease 2019 (COVID-19) has impacted almost every part of human life worldwide, posing a massive threat to human health. There is no specific drug for COVID-19, highlighting the urgent need for the developmen…

Unexpected novel Merbecovirus discoveries in agricultural sequencing datasets from Wuhan, China

2021-04-04 · Daoyu Zhang, Adrian Jones, Yuri Deigin, Karl Sirotkin 외

In this study we document the unexpected discovery of multiple coronaviruses and a BSL-3 pathogen in agricultural cotton and rice sequencing datasets. In particular, we have identified a novel HKU5-related Merbecovirus i…

Virology

LAPIS: A novel dataset for personalized image aesthetic assessment

2025-04-10 · Anne-Sofie Maerten, Li-Wei Chen, Stefanie De Winter, Christophe Bossens 외

We present the Leuven Art Personalized Image Set (LAPIS), a novel dataset for personalized image aesthetic assessment (PIAA). It is the first dataset with images of artworks that is suitable for PIAA. LAPIS consists of 1…