Catalogue Search | MBRL
Search Results Heading
Explore the vast range of titles available.
MBRLSearchResults
-
DisciplineDiscipline
-
Is Peer ReviewedIs Peer Reviewed
-
Item TypeItem Type
-
SubjectSubject
-
YearFrom:-To:
-
More FiltersMore FiltersSourceLanguage
Done
Filters
Reset
11,436
result(s) for
"phylogenetic trees"
Sort by:
Treemmer: a tool to reduce large phylogenetic datasets with minimal loss of diversity
by
Gygli, Sebastian M.
,
Loiseau, Chloé
,
Menardo, Fabrizio
in
Algorithms
,
Biogeography
,
Bioinformatics
2018
Background
Large sequence datasets are difficult to visualize and handle. Additionally, they often do not represent a random subset of the natural diversity, but the result of uncoordinated and convenience sampling. Consequently, they can suffer from redundancy and sampling biases.
Results
Here we present Treemmer, a simple tool to evaluate the redundancy of phylogenetic trees and reduce their complexity by eliminating leaves that contribute the least to the tree diversity.
Conclusions
Treemmer can reduce the size of datasets with different phylogenetic structures and levels of redundancy while maintaining a sub-sample that is representative of the original diversity. Additionally, it is possible to fine-tune the behavior of Treemmer including any kind of meta-information, making Treemmer particularly useful for empirical studies.
Journal Article
Common Methods for Phylogenetic Tree Construction and Their Implementation in R
2024
A phylogenetic tree can reflect the evolutionary relationships between species or gene families, and they play a critical role in modern biological research. In this review, we summarize common methods for constructing phylogenetic trees, including distance methods, maximum parsimony, maximum likelihood, Bayesian inference, and tree-integration methods (supermatrix and supertree). Here we discuss the advantages, shortcomings, and applications of each method and offer relevant codes to construct phylogenetic trees from molecular data using packages and algorithms in R. This review aims to provide comprehensive guidance and reference for researchers seeking to construct phylogenetic trees while also promoting further development and innovation in this field. By offering a clear and concise overview of the different methods available, we hope to enable researchers to select the most appropriate approach for their specific research questions and datasets.
Journal Article
‘gitana’ (phyloGenetic Imaging Tool for Adjusting Nodes and other Arrangements), a tool for plotting phylogenetic trees into ready-to-publish figures
by
de la Haba, Rafael R.
,
Galisteo, Cristina
in
Algorithms
,
Bioinformatics
,
Biomedical and Life Sciences
2025
Background
Phylogenetic trees are essential diagrams used in different sciences, such as evolutionary biology or taxonomy, and they depict the relationships between a given set of taxa sharing a common ancestor. So far, a multitude of tools have already been developed to infer phylogeny, and even more to visualize the resulting trees. However, editing generated graphical plots to obtain ready-to-publish figures is still a major issue. Most available tools do not take into consideration important aspects in nomenclature, such as the use of italics for taxon names or the superscript
T
that must be displayed after the strain/specimen designation to denote the type strain/specimen, at least not automatically. A gap also exists to easily highlight tree branches conserved across different phylogenies containing the same taxa. The lack of available tools to achieve these tasks is challenging for scientists, since manual formatting of phylogenetic trees is very time-consuming.
Results
Here, we present a tool named ‘gitana’, running in Linux/Windows/Mac operating systems with R software installed. It creates ready-to-publish trees with formatting taxon nomenclature and editing options such as rerooting, clade highlighting or collapsing, among other features. Moreover, ‘gitana’ performs node comparisons among phylogenies comprising the same taxa to identify conserved branches.
Conclusions
‘gitana’ is a user-friendly tool to output high-quality and ready-to-publish phylogenetic trees for users without R-coding skills. It combines dedicated functions of popular R packages for phylogeny and graphical visualization into an easy one-line-command. The users’ manual and source code are freely available at
https://github.com/cristinagalisteo/gitana
.
Journal Article
PhyloScape: interactive and scalable visualization platform for phylogenetic trees
2025
Background
With the accumulation of phylogenomic data and the growing demand for bioinformatics analyses, it has become increasingly important and complex to construct evolutionary relationships for different research purposes. Therefore, the ability to support multiple scenarios has become an essential need for phylogenetic visualization.
Results
In this study, we present PhyloScape, a web-based application for interactive visualization of phylogenetic trees that can be used stand-alone or as a toolkit deployed on the users’ website. The platform supports customizable multiple visualization features and is equipped with a flexible metadata annotation system, providing researchers with publishable, interactive views of trees. PhyloScape extensions include views of amino acid identity, geometry, and protein structure, which are applicable to various areas such as microbial taxonomy, pathogen phylogeny, and plant conservation. Trees published on the website can be efficiently shared and integrated into the users’ own system via a unique address.
Conclusions
As a scalable platform, PhyloScape provides a variety of online plug-ins that users can easily combine for specific scenarios. PhyloScape is freely available at
http://darwintree.cn/PhyloScape
.
Journal Article
PhyloMissForest: a random forest framework to construct phylogenetic trees with missing data
by
Santander-Jimenéz, Sergio
,
Pinheiro, Diogo
,
Ilic, Aleksandar
in
Algorithms
,
Amino acid sequence
,
Amino acids
2022
Background
In the pursuit of a better understanding of biodiversity, evolutionary biologists rely on the study of phylogenetic relationships to illustrate the course of evolution. The relationships among natural organisms, depicted in the shape of phylogenetic trees, not only help to understand evolutionary history but also have a wide range of additional applications in science. One of the most challenging problems that arise when building phylogenetic trees is the presence of missing biological data. More specifically, the possibility of inferring wrong phylogenetic trees increases proportionally to the amount of missing values in the input data. Although there are methods proposed to deal with this issue, their applicability and accuracy is often restricted by different constraints.
Results
We propose a framework, called PhyloMissForest, to impute missing entries in phylogenetic distance matrices and infer accurate evolutionary relationships. PhyloMissForest is built upon a random forest structure that infers the missing entries of the input data, based on the known parts of it. PhyloMissForest contributes with a robust and configurable framework that incorporates multiple search strategies and machine learning, complemented by phylogenetic techniques, to provide a more accurate inference of lost phylogenetic distances. We evaluate our framework by examining three real-world datasets, two DNA-based sequence alignments and one containing amino acid data, and two additional instances with simulated DNA data. Moreover, we follow a design of experiments methodology to define the hyperparameter values of our algorithm, which is a concise method, preferable in comparison to the well-known exhaustive parameters search. By varying the percentages of missing data from 5% to 60%, we generally outperform the state-of-the-art alternative imputation techniques in the tests conducted on real DNA data. In addition, significant improvements in execution time are observed for the amino acid instance. The results observed on simulated data also denote the attainment of improved imputations when dealing with large percentages of missing data.
Conclusions
By merging multiple search strategies, machine learning, and phylogenetic techniques, PhyloMissForest provides a highly customizable and robust framework for phylogenetic missing data imputation, with significant topological accuracy and effective speedups over the state of the art.
Journal Article
Genetic Analysis of Dengue Virus in Severe and Non-Severe Cases in Dhaka, Bangladesh, in 2018–2022
by
Hasan, Abu
,
Rahman, Mizanur
,
Biswas, Suma Mita
in
Antigens
,
Bangladesh
,
Bangladesh - epidemiology
2023
Dengue virus (DENV) infections have unpredictable clinical outcomes, ranging from asymptomatic or minor febrile illness to severe and fatal disease. The severity of dengue infection is at least partly related to the replacement of circulating DENV serotypes and/or genotypes. To describe clinical profiles of patients and the viral sequence diversity corresponding to non-severe and severe cases, we collected patient samples from 2018 to 2022 at Evercare Hospital Dhaka, Bangladesh. Serotyping of 495 cases and sequencing of 179 cases showed that the dominant serotype of DENV shifted from DENV2 in 2017 and 2018 to DENV3 in 2019. DENV3 persisted as the only representative serotype until 2022. Co-circulation of clades B and C of the DENV2 cosmopolitan genotype in 2017 was replaced by circulation of clade C alone in 2018 with all clones disappearing thereafter. DENV3 genotype I was first detected in 2017 and was the only genotype in circulation until 2022. We observed a high incidence of severe cases in 2019 when the DENV3 genotype I became the only virus in circulation. Phylogenetic analysis revealed clusters of severe cases in several different subclades of DENV3 genotype I. Thus, these serotype and genotype changes in DENV may explain the large dengue outbreaks and increased severity of the disease in 2019.
Journal Article
A new fast method for inferring multiple consensus trees using k-medoids
by
Tahiri, Nadia
,
Willems, Matthieu
,
Makarenkov, Vladimir
in
Adaptation (Biology)
,
Algorithms
,
Animal Systematics/Taxonomy/Biogeography
2018
Background
Gene trees carry important information about specific evolutionary patterns which characterize the evolution of the corresponding gene families. However, a reliable species consensus tree cannot be inferred from a multiple sequence alignment of a single gene family or from the concatenation of alignments corresponding to gene families having different evolutionary histories. These evolutionary histories can be quite different due to horizontal transfer events or to ancient gene duplications which cause the emergence of paralogs within a genome. Many methods have been proposed to infer a single consensus tree from a collection of gene trees. Still, the application of these tree merging methods can lead to the loss of specific evolutionary patterns which characterize some gene families or some groups of gene families. Thus, the problem of inferring multiple consensus trees from a given set of gene trees becomes relevant.
Results
We describe a new fast method for inferring multiple consensus trees from a given set of phylogenetic trees (i.e. additive trees or
X
-trees) defined on the same set of species (i.e. objects or taxa). The traditional consensus approach yields a single consensus tree. We use the popular
k
-medoids partitioning algorithm to divide a given set of trees into several clusters of trees. We propose novel versions of the well-known Silhouette and Caliński-Harabasz cluster validity indices that are adapted for tree clustering with
k
-medoids. The efficiency of the new method was assessed using both synthetic and real data, such as a well-known phylogenetic dataset consisting of 47 gene trees inferred for 14 archaeal organisms.
Conclusions
The method described here allows inference of multiple consensus trees from a given set of gene trees. It can be used to identify groups of gene trees having similar intragroup and different intergroup evolutionary histories. The main advantage of our method is that it is much faster than the existing tree clustering approaches, while providing similar or better clustering results in most cases. This makes it particularly well suited for the analysis of large genomic and phylogenetic datasets.
Journal Article
Genome-Wide Characterization, Identification and Expression Profile of MYB Transcription Factor Gene Family during Abiotic and Biotic Stresses in Mango (Mangifera indica)
2022
Mango (Mangifera indica) is an economically important fruit tree, and is cultivated in tropical, subtropical, and dry-hot valley areas around the world. Mango fruits have high nutritional value, and are mainly consumed fresh and used for commercial purposes. Mango is affected by various environmental factors during its growth and development. The MYB transcription factors participates in various physiological activities of plants, such as phytohormone signal transduction and disease resistance. In this study, 54 MiMYB transcription factors were identified in the mango genome (371.6 Mb). A phylogenetic tree was drawn based on the amino acid sequences of 222 MYB proteins of mango and Arabidopsis. The phylogenetic tree showed that the members of the mango MYB gene family were divided into 7 group, including Groups 1, -3, -4, -5, -6, -8, and -9. Ka/Ks ratios generally indicated that the MiMYBs of mango were affected by negative or positive selection. Quantitative real-time PCR showed that the transcription levels of MiMYBs were different under abiotic and biotic stresses, including salicylic acid, methyl jasmonate, and H2O2 treatments, and Colletotrichum gloeosporioides and Xanthomonas campestris pv. mangiferaeindicae infection, respectively. The transcript levels of MiMYB5, -35, -36, and -54 simultaneously responded positively to early treatments with salicylic acid, methyl jasmonate, and H2O2. The transcript level of MiMYB54 was activated by pathogenic fungal and bacterial infection. These results are beneficial for future interested researchers aiming to understand the biological functions and molecular mechanisms of MiMYB genes.
Journal Article
Testing the effectiveness of rbcLa DNA-barcoding for species discrimination in tropical montane cloud forest vascular plants (Oaxaca, Mexico) using BLAST, genetic distance, and tree-based methods
by
Trujillo-Argueta, Sonia
,
del Castillo, Rafael F.
,
Velasco-Murguía, Abril
in
Analysis
,
Animals
,
Bar codes
2022
DNA-barcoding is a species identification tool that uses a short section of the genome that provides a genetic signature of the species. The main advantage of this novel technique is that it requires a small sample of tissue from the tested organism. In most animal groups, this technique is very effective. However, in plants, the recommended standard markers, such as rbcL a, may not always work, and their efficacy remains to be tested in many plant groups, particularly from the Neotropical region. We examined the discriminating power of rbcL a in 55 tropical cloud forest vascular plant species from 38 families (Oaxaca, Mexico). We followed the CBOL criteria using BLASTn, genetic distance, and monophyly tree-based analyses (neighbor-joining, NJ, maximum likelihood, ML, and Bayesian inference, BI). rbcL a universal primers amplified 69.0% of the samples and yielded 91.3% bi-directional sequences. Sixty-three new rbcL a sequences were established. BLAST discriminates 80.8% of the genus but only 15.4% of the species. There was nil minimum interspecific genetic distances in Quercus, Oreopanax , and Daphnopsis . Contrastingly, Ericaceae (5.6%), Euphorbiaceae (4.6%), and Asteraceae (3.3%) species displayed the highest within-family genetic distances. According to the most recent angiosperm classification, NJ and ML trees successfully resolved (100%) monophyletic species. ML trees showed the highest mean branch support value (87.3%). Only NJ and ML trees could successfully discriminate Quercus species belonging to different subsections: Quercus martinezii (white oaks) from Q. callophylla and Q. laurina (red oaks). The ML topology could distinguish species in the Solanaceae clade with similar BLAST matches. Also, the BI topology showed a polytomy in this clade, and the NJ tree displayed low-support values. We do not recommend genetic-distance approaches for species discrimination. Severe shortages of rbcL a sequences in public databases of neotropical species hindered effective BLAST comparisons. Instead, ML tree-based analysis displays the highest species discrimination among the tree-based analyses. With the ML topology in selected genera, rbcL a helped distinguish infra-generic taxonomic categories, such as subsections, grouping affine species within the same genus, and discriminating species. Since the ML phylogenetic tree could discriminate 48 species out of our 55 studied species, we recommend this approach to resolve tropical montane cloud forest species using rbcL a, as an initial step and improve DNA amplification methods.
Journal Article