Hierarchical Dirichlet process model for gene expression clustering
Tóm tắt
Từ khóa
Tài liệu tham khảo
Schena M, Shalon D, Davis R, Brown P: Quantitative monitoring of gene expression patterns with a complementary DNA microarray. Science 1995,270(5235):467-470. 10.1126/science.270.5235.467
Cho R, Campbell M, Winzeler E, Steinmetz L, Conway A, Wodicka L, Wolfsberg T, Gabrielian A, Landsman D, Lockhart D: A genome-wide transcriptional analysis of the mitotic cell cycle. Mol. Cell 1998, 2: 65-73. 10.1016/S1097-2765(00)80114-8
Hughes J, Estep P, Tavazoie S, Church G: Computational identification of cis-regulatory elements associated with groups of functionally related genes in Saccharomyces cerevisiae. J. Mol. Biol 2000,296(5):1205-1214. 10.1006/jmbi.2000.3519
Eisen M, Spellman P, Brown P, Botstein D: Cluster analysis and display of genome-wide expression patterns. Proc. Natl. Acad. Sci 1998,95(25):14863-14868. 10.1073/pnas.95.25.14863
MacQueen J: Some methods for classification and analysis of multivariate observations. In Proceedings of the Fifth Berkeley Symposium on Mathematical Statistics and Probability, vol. 1. California: University of California Press; 1967:281-297.
Jiang D, Tang C, Zhang A: Cluster analysis for gene expression data: a survey. IEEE Trans. Knowledge Data Eng 2004,16(11):1370-1386. 10.1109/TKDE.2004.68
Dempster A, Laird N, Rubin D: Maximum likelihood from incomplete data via the EM algorithm. J. R. Stat. Soc. Ser. B (Methodological) 1977, 39: 1-38.
Fraley C, Raftery A, clustering Model-based, analysis discriminant, Am densityestimation. J.: Stat. Assoc. 2002,97(458):611-631. 10.1198/016214502760047131
Yeung K, Fraley C, Murua A, Raftery A, Ruzzo W: Model-based clustering and data transformations for gene expression data. Bioinformatics 2001,17(10):977-987. 10.1093/bioinformatics/17.10.977
Akaike H: A new look at the statistical model identification. IEEE Trans Autom. Control 1974,19(6):716-723. 10.1109/TAC.1974.1100705
Medvedovic M, Sivaganesan S: Bayesian infinite mixture model based clustering of gene expression profiles. Bioinformatics 2002,18(9):1194-1206. 10.1093/bioinformatics/18.9.1194
Ferguson T: A Bayesian analysis of some nonparametric problems. Ann. Stat 1973,1(2):209-230. 10.1214/aos/1176342360
Neal R: Markov chain sampling methods for Dirichlet process mixture models. J. Comput. Graph. Stat 2000,9(2):249-265.
Pitman J: Some developments of the Blackwell-MacQueen urn scheme. Lecture Notes-Monograph Series 1996, 245-267.
Kaufman L, Rousseeuw P: Finding Groups in Data: An Introduction to Cluster Analysis. Wiley Online Library; 1990.
Jiang D, Pei J, Zhang A: DHC: a density-based hierarchical clustering method for time series gene expression data. In Proceedings of Third IEEE Symposium on Bioinformatics and Bioengineering. Bethesda: IEEE; 2003:393-400.
Piatigorsky J: Gene Sharing and Evolution: The Diversity of Protein Functions. Cambridge: Harvard University Press; 2007.
Teh Y, Jordan M, Beal M, Blei D: Hierarchical Dirichlet processes. J. Am. Stat. Assoc 2006,101(476):1566-1581. 10.1198/016214506000000302
Sethuraman J: A constructive definition of Dirichlet priors. Stat. Sinica 1991, 4: 639-650.
Aldous D: Exchangeability and related topics. École d’Été de Probabilités de Saint-Flour XIII 1985, 1-198.
Casella G, George E: Explaining the Gibbs sampler. Am. Stat 1992,46(3):167-174.
Blackwell D, MacQueen J: Ferguson distributions via Pólya urn schemes. Ann. Stat 1973,1(2):353-355. 10.1214/aos/1176342372
Brooks S: Markov chain Monte Carlo method and its application. J. R. Stat. Soc. Ser. D (The Statistician) 1998, 47: 69-100. 10.1111/1467-9884.00117
Rousseeuw PJ: Silhouettes: a graphical aid to the interpretation and validation of cluster analysis. J. Comput. Appl. Math 1987, 20: 53-65.
Yeung KY, Ruzzo WL: Principal component analysis for clustering gene expression data. Bioinformatics 2001,17(9):763-774. 10.1093/bioinformatics/17.9.763
Yeung K, Medvedovic M, Bumgarner R: Clustering gene-expression data with repeated measurements. Genome Biol 2003,4(5):R34. 10.1186/gb-2003-4-5-r34
Chu S, DeRisi J, Eisen M, Mulholland J, Botstein D, Brown PO, Herskowitz I: The transcriptional program of sporulation in budding yeast. Science 1998,282(5389):699-705.
Iyer VR, Eisen MB, Ross DT, Schuler G, Moore T, Lee JC, Trent JM, Staudt LM, Hudson J, Boguski MS: The transcriptional program in the response of human fibroblasts to serum. Science 1999,283(5398):83-87. 10.1126/science.283.5398.83
Spellman P, Sherlock G, Zhang M, Iyer V, Anders K, Eisen M, Brown P, Botstein D, Futcher B: Comprehensive identification of cell cycle-regulated genes of the yeast Saccharomyces cerevisiae by microarray hybridization. Mol. Biol. Cell 1998,9(12):3273.
Blei D, Ng A, Jordan M: Latent Dirichlet allocation. J. Mach. Learn. Res 2003, 3: 993-1022.
Fraley C, Raftery A: MCLUST: software for model-based cluster analysis. J. Classif 1999,16(2):297-306. 10.1007/s003579900058
Furey T, Cristianini N, Duffy N, Bednarski D, Schummer M, Haussler D: Support vector machine classification and validation of cancer tissue samples using microarray expression data. Bioinformatics 2000,16(10):906-914. 10.1093/bioinformatics/16.10.906
Tavazoie S, Hughes JD, Campbell MJ, Cho RJ, Church GM: Systematic determination of genetic network architecture. Nat. Genetics 1999, 22: 281-285. 10.1038/10343
Chung F, Lu L CBMS Lecture Series no. 107. In Complex Graphs and Networks. Providence: American Mathematical Society; 2006.
Ashburner M, Ball C, Blake J, Botstein D, Butler H, Cherry J, Davis A, Dolinski K, Dwight S, Eppig J: Gene ontology: tool for the unification of biology. Nat. Genet 2000, 25: 25-29. 10.1038/75556
Stanford University: Yeast cell cycle datasets http://genome-www.stanford.edu/cellcycle/data/rawdata
Lukashin A, Fuchs R: Analysis of temporal gene expression profiles: clustering by simulated annealing and determining the optimal number of clusters. Bioinformatics 2001,17(5):405-414. 10.1093/bioinformatics/17.5.405
