Lien vers Pubmed [PMID] – 35904548
Lien DOI – 10.1093/bioinformatics/btac528
Bioinformatics 2022 Jul; ():
Bioinformatics applications increasingly rely on ad-hoc disk storage of k-mer sets, e.g. for de Bruijn graphs or alignment indexes. Here we introduce the K-mer File Format (KFF) as a general lossless framework for storing and manipulating k-mer sets, realizing space savings of 3-5x compared to other formats, and bringing interoperability across tools.Format specification, C ++/Rust API, tools: https://github.com/Kmer-File-Format/.