Plos iconPlosSep 17, 2026 ~1 min source read

PanDelos-plus: A parallel algorithm for computing sequence homology in pangenomic analysis

However, the increasing availability of genomic data requires tools that can scale efficiently to larger datasets. Benchmarks on synthetic datasets show that PanDelos-plus achieves up to 14x faster execution and reduces memory usage by up to 96%, while maintaining consistency with the original algorithm.

PanDelos-plus: A parallel algorithm for computing sequence homology in pangenomic analysis

Share this story

Send the public story page.

Useful takeaways from this story.

PanDelos addresses this challenge with an alignment-free and parameter-free approach based on k-mer profiles, combining high speed, ease of use, and competitive accuracy with state-of-the-art methods.

However, the increasing availability of genomic data requires tools that can scale efficiently to larger datasets.

Benchmarks on synthetic datasets show that PanDelos-plus achieves up to 14x faster execution and reduces memory usage by up to 96%, while maintaining consistency with the original algorithm.

Building the complete brief

The page is ready to read now. The fuller skim-friendly version will appear here automatically.

The useful part

PanDelos addresses this challenge with an alignment-free and parameter-free approach based on k-mer profiles, combining high speed, ease of use, and competitive accuracy with state-of-the-art methods. However, the increasing availability of genomic data requires tools that can scale efficiently to larger datasets. To address this need, we present PanDelos-plus, a fully parallel, gene-centric redesign of PanDelos.

How it works

  • Benchmarks on synthetic datasets show that PanDelos-plus achieves up to 14x faster execution and reduces memory usage by up to 96%, while maintaining consistency with the original algorithm.
  • The algorithm parallelizes the most computationally intensive phases (Best Hit detection and Bidirectional Best Hit extraction) through data decomposition and a thread pool strategy, while employing...
  • These improvements allow the PanDelos methodology to be applied to population-scale comparative genomics, thus enabling more precise characterisation of pangenome structure and dynamics.

Details worth keeping

by Simone Colli, Emiliano Maresi, Vincenzo Bonnici The identification of homologous gene families across multiple genomes is a central task in bacterial pangenomics traditionally requiring computationally demanding all-against-all comparisons. PanDelos-plus is available at github.com/synbionics/PanDelos-plus.

Keep reading in the app

Open the app view to save this story, compare related coverage, and continue from the same source.

Open in app