The topic of my diploma thesis is the statistical evaluation of biological sequences with the help of phylogenic trees. In the theoretical part we will create a literary recherche of estimation methodology concerning the course of phylogeny on the basis of the similarity of biological sequences (DNA and proteins) and we will focus on the inaccuracies of the estimation, their causes and the possibilities of their elimination. Afterwards, we will compare the methods for the statistical evaluation of the correctness of the course of phylogeny. In the practical part of the thesis we will suggest algorithms that will be used for testing the correctness of the phylogenic trees on the basis of bootstrapping, jackknifing, OTU jackknifing and PTP test which are able to the capture phylogenic tree with the method neighbor joining from the biological sequences in FASTA code. It is also possible to change the distance model and the substitution matrix. To be able to use these algorithms for the statistical support of phylogenic trees we have to verify their right function. This verification will be evaluated on the theoretical sequences of the amino acids. For the verification of the correct function of the algorithms, we will carry out single statistical tests on real 10 sequences of mammalian ubiquitin. These results will be analysed and appropriately discussed.
Identifer | oai:union.ndltd.org:nusl.cz/oai:invenio.nusl.cz:220015 |
Date | January 2013 |
Creators | Zembol, Filip |
Contributors | Provazník, Ivo, Škutková, Helena |
Publisher | Vysoké učení technické v Brně. Fakulta elektrotechniky a komunikačních technologií |
Source Sets | Czech ETDs |
Language | Czech |
Detected Language | English |
Type | info:eu-repo/semantics/masterThesis |
Rights | info:eu-repo/semantics/restrictedAccess |
Page generated in 0.0021 seconds