Have a personal or library account? Click to login
The correctness of large scale analysis of genomic data Cover

The correctness of large scale analysis of genomic data

Open Access
|Dec 2021

Abstract

Implementing a large genomic project is a demanding task, also from the computer science point of view. Besides collecting many genome samples and sequencing them, there is processing of a huge amount of data at every stage of their production and analysis. Efficient transfer and storage of the data is also an important issue. During the execution of such a project, there is a need to maintain work standards and control quality of the results, which can be difficult if a part of the work is carried out externally. Here, we describe our experience with such data quality analysis on a number of levels - from an obvious check of the quality of the results obtained, to examining consistency of the data at various stages of their processing, to verifying, as far as possible, their compatibility with the data describing the sample.

DOI: https://doi.org/10.2478/fcds-2021-0024 | Journal eISSN: 2300-3405 | Journal ISSN: 0867-6356
Language: English
Page range: 423 - 436
Submitted on: Oct 2, 2021
|
Accepted on: Nov 27, 2021
|
Published on: Dec 17, 2021
In partnership with: Paradigm Publishing Services
Publication frequency: 4 issues per year

© 2021 Pawel Wojciechowski, Karol Krause, Piotr Lukasiak, Jacek Blazewicz, published by Poznan University of Technology
This work is licensed under the Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 License.