Developing a gene model for simulations that incorporates multi-species conservation
Name
894242021-MIT.pdf
Description
Full printable version
Size
7.68 MB
Format
Adobe PDF
Checksum (MD5)
84addd5e3d000d3d70bd86f8649879d2
Author(s)
Liu, Brendan F
Advisor(s)
David Altshuler.
Date Issued
2014
Publisher
Massachusetts Institute of Technology
Abstract
The genetic architecture, the number, frequency, and effect size of disease causing alleles for many common diseases including Type 2 Diabetes is not fully understood. Genetic simulations can be used to make predictions under specified genetic architecture models. Models whose predictions are inconsistent with empirical data can be rejected. We extended a gene simulation model previously published by our lab. The distribution of number and length of coding and intron regions of each simulated gene was consistent with the distribution in the human genome. Selection pressure against mutations was modeled by utilizing the cross-species conservation of each region. The combined distribution of variants by their frequency over 500 genes was compared between the simulated genes and the corresponding empirical data. This distribution of variants between the simulated and empirical data was found to be consistent.
Description
Thesis: M. Eng., Massachusetts Institute of Technology, Department of Electrical Engineering and Computer Science, 2014.
Cataloged from PDF version of thesis.
Includes bibliographical references (pages 69-71).
Subjects
Electrical Engineering and Computer Science.
MIT Department
Massachusetts Institute of Technology. Department of Electrical Engineering and Computer Science
Terms of Use
M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission.
Persistent DSpace Link