Back to Search Start Over

Joint modeling of DNA sequence and physical properties to improve eukaryotic promoter recognition.

Authors :
Ohler U
Niemann H
Liao Gc
Rubin GM
Source :
Bioinformatics (Oxford, England) [Bioinformatics] 2001; Vol. 17 Suppl 1, pp. S199-206.
Publication Year :
2001

Abstract

We present an approach to integrate physical properties of DNA, such as DNA bendability or GC content, into our probabilistic promoter recognition system McPROMOTER. In the new model, a promoter is represented as a sequence of consecutive segments represented by joint likelihoods for DNA sequence and profiles of physical properties. Sequence likelihoods are modeled with interpolated Markov chains, physical properties with Gaussian distributions. The background uses two joint sequence/profile models for coding and non-coding sequences, each consisting of a mixture of a sense and an anti-sense submodel. On a large Drosophila test set, we achieved a reduction of about 30% of false positives when compared with a model solely based on sequence likelihoods.

Details

Language :
English
ISSN :
1367-4803
Volume :
17 Suppl 1
Database :
MEDLINE
Journal :
Bioinformatics (Oxford, England)
Publication Type :
Academic Journal
Accession number :
11473010
Full Text :
https://doi.org/10.1093/bioinformatics/17.suppl_1.s199