Person: Kong, Sek Won
Email Address
AA Acceptance Date
Birth Date
Research Projects
Organizational Units
Job Title
Last Name
First Name
Name
Search Results
Publication Network-Based Analysis of Affected Biological Processes in Type 2 Diabetes Models
(Public Library of Science, 2007) Liu, Manway; Liberzon, Arthur; Kong, Sek Won; Lai, Weil R; Park, Peter; Kohane, Isaac; Kasif, SimonType 2 diabetes mellitus is a complex disorder associated with multiple genetic, epigenetic, developmental, and environmental factors. Animal models of type 2 diabetes differ based on diet, drug treatment, and gene knockouts, and yet all display the clinical hallmarks of hyperglycemia and insulin resistance in peripheral tissue. The recent advances in gene-expression microarray technologies present an unprecedented opportunity to study type 2 diabetes mellitus at a genome-wide scale and across different models. To date, a key challenge has been to identify the biological processes or signaling pathways that play significant roles in the disorder. Here, using a network-based analysis methodology, we identified two sets of genes, associated with insulin signaling and a network of nuclear receptors, which are recurrent in a statistically significant number of diabetes and insulin resistance models and transcriptionally altered across diverse tissue types. We additionally identified a network of protein–protein interactions between members from the two gene sets that may facilitate signaling between them. Taken together, the results illustrate the benefits of integrating high-throughput microarray studies, together with protein–protein interaction networks, in elucidating the underlying biological processes associated with a complex disorder.
Publication Divergent dysregulation of gene expression in murine models of fragile X syndrome and tuberous sclerosis
(BioMed Central, 2014) Kong, Sek Won; Sahin, Mustafa; Collins, Christin D.; Wertz, Mary H; Campbell, Malcolm G; Leech, Jarrett D; Krueger, Dilja; Bear, Mark F; Kunkel, Louis; Kohane, IsaacBackground: Fragile X syndrome and tuberous sclerosis are genetic syndromes that both have a high rate of comorbidity with autism spectrum disorder (ASD). Several lines of evidence suggest that these two monogenic disorders may converge at a molecular level through the dysfunction of activity-dependent synaptic plasticity. Methods: To explore the characteristics of transcriptomic changes in these monogenic disorders, we profiled genome-wide gene expression levels in cerebellum and blood from murine models of fragile X syndrome and tuberous sclerosis. Results: Differentially expressed genes and enriched pathways were distinct for the two murine models examined, with the exception of immune response-related pathways. In the cerebellum of the Fmr1 knockout (Fmr1-KO) model, the neuroactive ligand receptor interaction pathway and gene sets associated with synaptic plasticity such as long-term potentiation, gap junction, and axon guidance were the most significantly perturbed pathways. The phosphatidylinositol signaling pathway was significantly dysregulated in both cerebellum and blood of Fmr1-KO mice. In Tsc2 heterozygous (+/−) mice, immune system-related pathways, genes encoding ribosomal proteins, and glycolipid metabolism pathways were significantly changed in both tissues. Conclusions: Our data suggest that distinct molecular pathways may be involved in ASD with known but different genetic causes and that blood gene expression profiles of Fmr1-KO and Tsc2+/− mice mirror some, but not all, of the perturbed molecular pathways in the brain.
Publication Characteristics and Predictive Value of Blood Transcriptome Signature in Males with Autism Spectrum Disorders
(Public Library of Science, 2012) Kong, Sek Won; Collins, Christin D.; Shimizu-Motohashi, Yuko; Holm, Ingrid; Campbell, Malcolm G.; Lee, In-Hee; Brewster, Stephanie J.; Hanson, Ellen; Harris, Heather; Lowe, Kathryn R.; Saada, Adrianna; Mora, Andrea; Madison, Kimberly; Hundley, Rachel; Egan, Jessica; McCarthy, Jillian; Eran, Ally; Galdzicki, Michal; Rappaport, Leonard; Kunkel, Louis; Kohane, IsaacAutism Spectrum Disorders (ASD) is a spectrum of highly heritable neurodevelopmental disorders in which known mutations contribute to disease risk in 20% of cases. Here, we report the results of the largest blood transcriptome study to date that aims to identify differences in 170 ASD cases and 115 age/sex-matched controls and to evaluate the utility of gene expression profiling as a tool to aid in the diagnosis of ASD. The differentially expressed genes were enriched for the neurotrophin signaling, long-term potentiation/depression, and notch signaling pathways. We developed a 55-gene prediction model, using a cross-validation strategy, on a sample cohort of 66 male ASD cases and 33 age-matched male controls (P1). Subsequently, 104 ASD cases and 82 controls were recruited and used as a validation set (P2). This 55-gene expression signature achieved 68% classification accuracy with the validation cohort (area under the receiver operating characteristic curve (AUC): 0.70 [95% confidence interval [CI]: 0.62–0.77]). Not surprisingly, our prediction model that was built and trained with male samples performed well for males (AUC 0.73, 95% CI 0.65–0.82), but not for female samples (AUC 0.51, 95% CI 0.36–0.67). The 55-gene signature also performed robustly when the prediction model was trained with P2 male samples to classify P1 samples (AUC 0.69, 95% CI 0.58–0.80). Our result suggests that the use of blood expression profiling for ASD detection may be feasible. Further study is required to determine the age at which such a test should be deployed, and what genetic characteristics of ASD can be identified.
Publication Pathway-based outlier method reveals heterogeneous genomic structure of autism in blood transcriptome
(BioMed Central, 2013) Campbell, Malcolm G; Kohane, Isaac; Kong, Sek WonBackground: Decades of research strongly suggest that the genetic etiology of autism spectrum disorders (ASDs) is heterogeneous. However, most published studies focus on group differences between cases and controls. In contrast, we hypothesized that the heterogeneity of the disorder could be characterized by identifying pathways for which individuals are outliers rather than pathways representative of shared group differences of the ASD diagnosis. Methods: Two previously published blood gene expression data sets – the Translational Genetics Research Institute (TGen) dataset (70 cases and 60 unrelated controls) and the Simons Simplex Consortium (Simons) dataset (221 probands and 191 unaffected family members) – were analyzed. All individuals of each dataset were projected to biological pathways, and each sample’s Mahalanobis distance from a pooled centroid was calculated to compare the number of case and control outliers for each pathway. Results: Analysis of a set of blood gene expression profiles from 70 ASD and 60 unrelated controls revealed three pathways whose outliers were significantly overrepresented in the ASD cases: neuron development including axonogenesis and neurite development (29% of ASD, 3% of control), nitric oxide signaling (29%, 3%), and skeletal development (27%, 3%). Overall, 50% of cases and 8% of controls were outliers in one of these three pathways, which could not be identified using group comparison or gene-level outlier methods. In an independently collected data set consisting of 221 ASD and 191 unaffected family members, outliers in the neurogenesis pathway were heavily biased towards cases (20.8% of ASD, 12.0% of control). Interestingly, neurogenesis outliers were more common among unaffected family members (Simons) than unrelated controls (TGen), but the statistical significance of this effect was marginal (Chi squared P < 0.09). Conclusions: Unlike group difference approaches, our analysis identified the samples within the case and control groups that manifested each expression signal, and showed that outlier groups were distinct for each implicated pathway. Moreover, our results suggest that by seeking heterogeneity, pathway-based outlier analysis can reveal expression signals that are not apparent when considering only shared group differences.
Publication The MedSeq Project: a randomized trial of integrating whole genome sequencing into clinical medicine
(BioMed Central, 2014) Vassy, Jason; Lautenbach, Denise M; McLaughlin, Heather M; Kong, Sek Won; Christensen, Kurt; Krier, Joel; Kohane, Isaac; Feuerman, Lindsay Z; Blumenthal-Barby, Jennifer; Roberts, J Scott; Lehmann, Lisa Soleymani; Ho, Carolyn; Ubel, Peter A; MacRae, Calum; Seidman, Christine; Murray, Michael F; McGuire, Amy L; Rehm, Heidi; Green, RobertBackground: Whole genome sequencing (WGS) is already being used in certain clinical and research settings, but its impact on patient well-being, health-care utilization, and clinical decision-making remains largely unstudied. It is also unknown how best to communicate sequencing results to physicians and patients to improve health. We describe the design of the MedSeq Project: the first randomized trials of WGS in clinical care. Methods/Design This pair of randomized controlled trials compares WGS to standard of care in two clinical contexts: (a) disease-specific genomic medicine in a cardiomyopathy clinic and (b) general genomic medicine in primary care. We are recruiting 8 to 12 cardiologists, 8 to 12 primary care physicians, and approximately 200 of their patients. Patient participants in both the cardiology and primary care trials are randomly assigned to receive a family history assessment with or without WGS. Our laboratory delivers a genome report to physician participants that balances the needs to enhance understandability of genomic information and to convey its complexity. We provide an educational curriculum for physician participants and offer them a hotline to genetics professionals for guidance in interpreting and managing their patients’ genome reports. Using varied data sources, including surveys, semi-structured interviews, and review of clinical data, we measure the attitudes, behaviors and outcomes of physician and patient participants at multiple time points before and after the disclosure of these results. Discussion The impact of emerging sequencing technologies on patient care is unclear. We have designed a process of interpreting WGS results and delivering them to physicians in a way that anticipates how we envision genomic medicine will evolve in the near future. That is, our WGS report provides clinically relevant information while communicating the complexity and uncertainty of WGS results to physicians and, through physicians, to their patients. This project will not only illuminate the impact of integrating genomic medicine into the clinical care of patients but also inform the design of future studies. Trial registration ClinicalTrials.gov identifier NCT01736566
Publication Combining Gene Expression Data from Different Generations of Oligonucleotide Arrays
(BioMed Central, 2004) Hwang, Kyu-Baek; Kong, Sek Won; Greenberg, Steven; Park, PeterBackground: One of the important challenges in microarray analysis is to take full advantage of previously accumulated data, both from one's own laboratory and from public repositories. Through a comparative analysis on a variety of datasets, a more comprehensive view of the underlying mechanism or structure can be obtained. However, as we discover in this work, continual changes in genomic sequence annotations and probe design criteria make it difficult to compare gene expression data even from different generations of the same microarray platform. Results: We first describe the extent of discordance between the results derived from two generations of Affymetrix oligonucleotide arrays, as revealed in cluster analysis and in identification of differentially expressed genes. We then propose a method for increasing comparability. The dataset we use consists of a set of 14 human muscle biopsy samples from patients with inflammatory myopathies that were hybridized on both HG-U95Av2 and HG-U133A human arrays. We find that the use of the probe set matching table for comparative analysis provided by Affymetrix produces better results than matching by UniGene or LocusLink identifiers but still remains inadequate. Rescaling of expression values for each gene across samples and data filtering by expression values enhance comparability but only for few specific analyses. As a generic method for improving comparability, we select a subset of probes with overlapping sequence segments in the two array types and recalculate expression values based only on the selected probes. We show that this filtering of probes significantly improves the comparability while retaining a sufficient number of probe sets for further analysis. Conclusions: Compatibility between high-density oligonucleotide arrays is significantly affected by probe-level sequence information. With a careful filtering of the probes based on their sequence overlaps, data from different generations of microarrays can be combined more effectively.
Publication Integration of Heterogeneous Expression Data Sets Extends the Role of the Retinol Pathway in Diabetes and Insulin Resistance
(Oxford University Press, 2009) Park, Peter; Kong, Sek Won; Tebaldi, Toma; Lai, Weil R.; Kasif, Simon; Kohane, IsaacMotivation: Type 2 diabetes is a chronic metabolic disease that involves both environmental and genetic factors. To understand the genetics of type 2 diabetes and insulin resistance, the DIabetes Genome Anatomy Project (DGAP) was launched to profile gene expression in a variety of related animal models and human subjects. We asked whether these heterogeneous models can be integrated to provide consistent and robust biological insights into the biology of insulin resistance. Results: We perform integrative analysis of the 16 DGAP data sets that span multiple tissues, conditions, array types, laboratories, species, genetic backgrounds and study designs. For each data set, we identify differentially expressed genes compared with control. Then, for the combined data, we rank genes according to the frequency with which they were found to be statistically significant across data sets. This analysis reveals RetSat as a widely shared component of mechanisms involved in insulin resistance and sensitivity and adds to the growing importance of the retinol pathway in diabetes, adipogenesis and insulin resistance. Top candidates obtained from our analysis have been confirmed in recent laboratory studies. Contact: Isaac_kohane@harvard.edu