Abstract

The exponential increase in sequencing data calls for conceptual and computational advances to extract useful biological insights. One such advance, minimizers, allows for reducing the quantity of data handled while maintaining some of its key properties. We provide a basic introduction to minimizers, cover recent methodological developments, and review the diverse applications of minimizers to analyze genomic data, including de novo genome assembly, metagenomics, read alignment, read correction, and pangenomes. We also touch on alternative data sketching techniques including universal hitting sets, syncmers, or strobemers. Minimizers and their alternatives have rapidly become indispensable tools for handling vast amounts of data.

Details

Title
When less is more: sketching with minimizers in genomics
Author
Ndiaye, Malick; Prieto-Baños, Silvia; Fitzgerald, Lucy M; Kharrazi, Ali Yazdizadeh; Oreshkov, Sergey; Dessimoz, Christophe; Sedlazeck, Fritz J; Glover, Natasha; Majidian, Sina
Pages
1-35
Section
Review
Publication year
2024
Publication date
2024
Publisher
BioMed Central
ISSN
14747596
e-ISSN
1474760X
Source type
Scholarly Journal
Language of publication
English
ProQuest document ID
3201607032
Copyright
© 2024. This work is licensed under http://creativecommons.org/licenses/by/4.0/ (the “License”). Notwithstanding the ProQuest Terms and Conditions, you may use this content in accordance with the terms of the License.