Genome Clustering: From Linguistic Models to Classification of Genetic Texts (Studies in Computational Intelligence)

admin
Image

This book deals with the methods of text comparison which are based on different techniques of converting the text into a distribution on a certain finite support, be it a genetic text or a text of some other type. Such distribution is usually referred to as “spectrum”. The measure of dissimilarity of two texts is formally expressed as a certain “distance” between the spectra of these texts. Such definition implies that the similarity of the texts results from the similarity of the random processes generating the texts.