Skip Navigation

This Article
Right arrow FREE Full Text (Print PDF) Freely available
Right arrow Comments: Submit a response
Right arrow Alert me when this article is cited
Right arrow Alert me when Comments are posted
Right arrow Alert me if a correction is posted
Services
Right arrow Email this article to a friend
Right arrow Similar articles in this journal
Right arrow Similar articles in ISI Web of Science
Right arrow Similar articles in PubMed
Right arrow Alert me to new issues of the journal
Right arrow Add to My Personal Archive
Right arrow Download to citation manager
Right arrow Search for citing articles in:
ISI Web of Science (9)
Right arrowRequest Permissions
Google Scholar
Right arrow Articles by Sneath, P. H.
Right arrow Search for Related Content
PubMed
Right arrow PubMed Citation
Right arrow Articles by Sneath, P. H.
Social Bookmarking
 Add to CiteULike   Add to Connotea   Add to Del.icio.us  
What's this?

Bioinformatics, Vol 14, 608-616, Copyright © 1998 by Oxford University Press


ARTICLES

The effect of evenly spaced constant sites on the distribution of the random division of a molecular sequence

PH Sneath
Department of Microbiology and Immunology, Leicester University, Leicester LE1 9HN, UK. phas1@le.ac.uk

MOTIVATION: A modified Sherman statistic can be used to test whether the differences between two aligned sequences are distributed at random along the sequences, or whether they are clustered, which suggests anomalies of evolution such as partial gene recombination or functional constraints. The presence of evenly spaced constant sites (such as constancy at the second codon position in genes coding for proteins) lowers the statistic and makes the significance less than it should be. RESULTS: The magnitude of the constant-site effect is shown by simulation to depend mainly on the proportion of differences between two sequences and on the number of constant sites that are added after each variable site. This latter number can be estimated from the variance of sites in a sequence matrix at the first, second and third codon positions, to obtain a ratio that corrects the statistic. When expressed as standard errors, the uncorrected results are too low (typically half to one unit when almost all the variation is at the third codon position). Correction raises the standard errors to levels close to expectation. If the data show no marked ternary periodicity, the correction is very small. The method is illustrated with biological data that show close to random behaviour, and with data that exhibit strong clustering. AVAILABILITY: The software is available from the author and has also been placed on the EMBL file server (Software@embl- ebi.ac.uk). CONTACT: phas1@le.ac.uk
Add to CiteULike CiteULike   Add to Connotea Connotea   Add to Del.icio.us Del.icio.us    What's this?


This article has been cited by other articles:


Home page
GeneticsHome page
M. F. Boni, D. Posada, and M. W. Feldman
An Exact Nonparametric Method for Inferring Mosaic Structure in Sequence Triplets
Genetics, June 1, 2007; 176(2): 1035 - 1047.
[Abstract] [Full Text] [PDF]


Home page
Mol Biol EvolHome page
D. Posada
Evaluation of Methods for Detecting Recombination from DNA Sequences: Empirical Data
Mol. Biol. Evol., May 1, 2002; 19(5): 708 - 717.
[Abstract] [Full Text] [PDF]


Home page
Proc. Natl. Acad. Sci. USAHome page
D. Posada and K. A. Crandall
Evaluation of methods for detecting recombination from DNA sequences: Computer simulations
PNAS, November 20, 2001; 98(24): 13757 - 13762.
[Abstract] [Full Text] [PDF]



Disclaimer: Please note that abstracts for content published before 1996 were created through digital scanning and may therefore not exactly replicate the text of the original print issues. All efforts have been made to ensure accuracy, but the Publisher will not be held responsible for any remaining inaccuracies. If you require any further clarification, please contact our Customer Services Department.