Sampling solution traces for the problem of sorting permutations by signed reversals

Abstract : Background
Traditional algorithms to solve the problem of sorting by signed reversals output just one optimal solution while the space of all optimal solutions can be huge. A so-called trace represents a group of solutions which share the same set of reversals that must be applied to sort the original permutation following a partial ordering. By using traces, we therefore can represent the set of optimal solutions in a more compact way. Algorithms for enumerating the complete set of traces of solutions were developed. However, due to their exponential complexity, their practical use is limited to small permutations. A partial enumeration of traces is a sampling of the complete set of traces and can be an alternative for the study of distinct evolutionary scenarios of big permutations. Ideally, the sampling should be done uniformly from the space of all optimal solutions. This is however conjectured to be ♯P-complete.
Results
We propose and evaluate three algorithms for producing a sampling of the complete set of traces that instead can be shown in practice to preserve some of the characteristics of the space of all solutions. The first algorithm (RA) performs the construction of traces through a random selection of reversals on the list of optimal 1-sequences. The second algorithm (DFALT) consists in a slight modification of an algorithm that performs the complete enumeration of traces. Finally, the third algorithm (SWA) is based on a sliding window strategy to improve the enumeration of traces. All proposed algorithms were able to enumerate traces for permutations with up to 200 elements.
Conclusions
We analysed the distribution of the enumerated traces with respect to their height and average reversal length. Various works indicate that the reversal length can be an important aspect in genome rearrangements. The algorithms RA and SWA show a tendency to lose traces with high average reversal length. Such traces are however rare, and qualitatively our results show that, for testable-sized permutations, the algorithms DFALT and SWA produce distributions which approximate the reversal length distributions observed with a complete enumeration of the set of traces.
Type de document :
Article dans une revue
Algorithms for Molecular Biology, BioMed Central, 2012, 7 (1), pp.18. 〈10.1186/1748-7188-7-18〉
Liste complète des métadonnées

Littérature citée [28 références]  Voir  Masquer  Télécharger

https://hal.inria.fr/hal-00784400
Contributeur : Ed. Bmc <>
Soumis le : lundi 4 février 2013 - 12:59:15
Dernière modification le : jeudi 28 juin 2018 - 14:38:44
Document(s) archivé(s) le : lundi 17 juin 2013 - 18:37:09

Fichiers

1748-7188-7-18.pdf
Fichiers éditeurs autorisés sur une archive ouverte

Identifiants

Collections

Citation

Christian Baudet, Zanoni Dias, Marie-France Sagot. Sampling solution traces for the problem of sorting permutations by signed reversals. Algorithms for Molecular Biology, BioMed Central, 2012, 7 (1), pp.18. 〈10.1186/1748-7188-7-18〉. 〈hal-00784400〉

Partager

Métriques

Consultations de la notice

477

Téléchargements de fichiers

354