Accéder directement au contenu Accéder directement à la navigation
Article dans une revue

Bios2mds: an R package for comparing orthologous protein families by metric multidimensional scaling

Abstract :

BACKGROUND: The distance matrix computed from multiple alignments of homologous sequences is widely used by distance-based phylogenetic methods to provide information on the evolution of protein families. This matrix can also be visualized in a low dimensional space by metric multidimensional scaling (MDS). Applied to protein families, MDS provides information complementary to the information derived from tree-based methods. Moreover, MDS gives a unique opportunity to compare orthologous sequence sets because it can add supplementary elements to a reference space. RESULTS: The R package bios2mds (from BIOlogical Sequences to MultiDimensional Scaling) has been designed to analyze multiple sequence alignments by MDS. Bios2mds starts with a sequence alignment, builds a matrix of distances between the aligned sequences, and represents this matrix by MDS to visualize a sequence space. This package also offers the possibility of performing K-means clustering in the MDS derived sequence space. Most importantly, bios2mds includes a function that projects supplementary elements (a.k.a. "out of sample" elements) onto the space defined by reference or "active" elements. Orthologous sequence sets can thus be compared in a straightforward way. The data analysis and visualization tools have been specifically designed for an easy monitoring of the evolutionary drift of protein sub-families. CONCLUSIONS: The bios2mds package provides the tools for a complete integrated pipeline aimed at the MDS analysis of multiple sets of orthologous sequences in the R statistical environment. In addition, as the analysis can be carried out from user provided matrices, the projection function can be widely used on any kind of data.

Type de document :
Article dans une revue
Liste complète des métadonnées

https://hal.univ-angers.fr/hal-03403932
Contributeur : Okina Univ Angers Connectez-vous pour contacter le contributeur
Soumis le : mardi 26 octobre 2021 - 13:50:17
Dernière modification le : vendredi 29 octobre 2021 - 17:46:26

Lien texte intégral

Identifiants

Collections

Citation

J. Pele, J. Becu, Hervé Abdi, Marie Chabbert. Bios2mds: an R package for comparing orthologous protein families by metric multidimensional scaling. BMC Bioinformatics, BioMed Central, 2012, 13, Non spécifié. ⟨10.1186/1471-2105-13-133⟩. ⟨hal-03403932⟩

Partager

Métriques

Consultations de la notice

5