Link to Pubmed [PMID] – 35904548
Link to HAL – inria-03891410
Link to DOI – 10.1093/bioinformatics/btac528
Bioinformatics, 2022, 38 (18), pp.4423-4425. ⟨10.1093/bioinformatics/btac528⟩
Bioinformatics applications increasingly rely on ad hoc disk storage of k-mer sets, e.g. for de Bruijn graphs or alignment indexes. Here, we introduce the K-mer File Format as a general lossless framework for storing and manipulating k-mer sets, realizing space savings of 3–5× compared to other formats, and bringing interoperability across tools. Availability: Format specification, C++/Rust API, tools: https://github.com/Kmer-File-Format/.