Rockefeller University
UniversityNew York, United States
Research output, citation impact, and the most-cited recent papers from Rockefeller University (United States). Aggregated across the NobleBlocks index of 300M+ scholarly works.
Top-cited papers from Rockefeller University
Functional elucidation of causal genetic variants and elements requires precise genome editing technologies. The type II prokaryotic CRISPR (clustered regularly interspaced short palindromic repeats)/Cas adaptive immune system has been shown to facilitate RNA-guided site-specific DNA cleavage. We engineered two different type II CRISPR/Cas systems and demonstrate that Cas9 nucleases can be directed by short RNAs to induce precise cleavage at endogenous genomic loci in human and mouse cells. Cas9 can also be converted into a nicking enzyme to facilitate homology-directed repair with minimal mutagenic activity. Lastly, multiple guide sequences can be encoded into a single CRISPR array to enable simultaneous editing of several sites within the mammalian genome, demonstrating easy programmability and wide applicability of the RNA-guided nuclease technology.
A 2.91-billion base pair (bp) consensus sequence of the euchromatic portion of the human genome was generated by the whole-genome shotgun sequencing method. The 14.8-billion bp DNA sequence was generated over 9 months from 27,271,853 high-quality sequence reads (5.11-fold coverage of the genome) from both ends of plasmid clones made from the DNA of five individuals. Two assembly strategies-a whole-genome assembly and a regional chromosome assembly-were used, each combining sequence data from Celera and the publicly funded genome effort. The public data were shredded into 550-bp segments to create a 2.9-fold coverage of those genome regions that had been sequenced, without including biases inherent in the cloning and assembly procedure used by the publicly funded group. This brought the effective coverage in the assemblies to eightfold, reducing the number and size of gaps in the final assembly over what would be obtained with 5.11-fold coverage. The two assembly strategies yielded very similar results that largely agree with independent mapping data. The assemblies effectively cover the euchromatic regions of the human chromosomes. More than 90% of the genome is in scaffold assemblies of 100,000 bp or more, and 25% of the genome is in scaffolds of 10 million bp or larger. Analysis of the genome sequence revealed 26,588 protein-encoding transcripts for which there was strong corroborating evidence and an additional approximately 12,000 computationally derived genes with mouse matches or other weak supporting evidence. Although gene-dense clusters are obvious, almost half the genes are dispersed in low G+C sequence separated by large tracts of apparently noncoding sequence. Only 1.1% of the genome is spanned by exons, whereas 24% is in introns, with 75% of the genome being intergenic DNA. Duplications of segmental blocks, ranging in size up to chromosomal lengths, are abundant throughout the genome and reveal a complex evolutionary history. Comparative genomic analysis indicates vertebrate expansions of genes associated with neuronal function, with tissue-specific developmental regulation, and with the hemostasis and immune systems. DNA sequence comparisons between the consensus sequence and publicly funded genome data provided locations of 2.1 million single-nucleotide polymorphisms (SNPs). A random pair of human haploid genomes differed at a rate of 1 bp per 1250 on average, but there was marked heterogeneity in the level of polymorphism across the genome. Less than 1% of all SNPs resulted in variation in proteins, but the task of determining which SNPs have functional consequences remains an open challenge.
Plasmodium falciparum can now be maintained in continuous culture in human erythrocytes incubated at 38 degrees C in RPMI 1640 medium with human serum under an atmosphere with 7 percent carbon dioxide and low oxygen (1 or 5 percent). The original parasite material, derived from an infected Aotus trivirgatus monkey, was diluted more than 100 million times by the addition of human erythrocytes at 3- or 4-day intervals. The parasites continued to reproduce in their normal asexual cycle of approximately 48 hours but were no longer highly synchronous. The have remained infective to Aotus.
The proliferation of large-scale DNA-sequencing projects in recent years has driven a search for alternative methods to reduce time and cost. Here we describe a scalable, highly parallel sequencing system with raw throughput significantly greater than that of state-of-the-art capillary electrophoresis instruments. The apparatus uses a novel fibre-optic slide of individual wells and is able to sequence 25 million bases, at 99% or better accuracy, in one four-hour run. To achieve an approximately 100-fold increase in throughput over current Sanger sequencing technology, we have developed an emulsion method for DNA amplification and an instrument for sequencing by synthesis using a pyrosequencing protocol optimized for solid support and picolitre-scale volumes. Here we show the utility, throughput, accuracy and robustness of this system by shotgun sequencing and de novo assembly of the Mycoplasma genitalium genome with 96% coverage at 99.96% accuracy in one run of the machine. The race is on for a big prize: the job of providing the world's DNA sequencing laboratories with the successor to the ‘Sanger-based’ technology that gave us the first wave of genome sequences. One technology in the frame is that produced by 454 Life Sciences Corporation of Branford, Connecticut. Today's technology reads 67,000 base pairs per hour; this new approach is 100 times faster, reading 6 million base pairs per hour. The improved performance results from using picolitre-sized chemical reactors, enhanced light-emitting sequencing chemistries and complex informatics. Further miniaturization of the system is planned. Such leaps in technology may one day make it possible to analyse an individual's genome before designing therapy: the ultimate in personalized medicine.
The potassium channel from Streptomyces lividans is an integral membrane protein with sequence similarity to all known K+ channels, particularly in the pore region. X-ray analysis with data to 3.2 angstroms reveals that four identical subunits create an inverted teepee, or cone, cradling the selectivity filter of the pore in its outer end. The narrow selectivity filter is only 12 angstroms long, whereas the remainder of the pore is wider and lined with hydrophobic amino acids. A large water-filled cavity and helix dipoles are positioned so as to overcome electrostatic destabilization of an ion in the pore at the center of the bilayer. Main chain carbonyl oxygen atoms from the K+ channel signature sequence line the selectivity filter, which is held open by structural constraints to coordinate K+ ions but not smaller Na+ ions. The selectivity filter contains two K+ ions about 7.5 angstroms apart. This configuration promotes ion conduction by exploiting electrostatic repulsive forces to overcome attractive forces between K+ ions and the selectivity filter. The architecture of the pore establishes the physical principles underlying selective K+ conduction.
Over 60 years ago, Selye1 recognized the paradox that the physiologic systems activated by stress can not only protect and restore but also damage the body. What links these seemingly contradictory roles? How does stress influence the pathogenesis of disease, and what accounts for the variation in vulnerability to stress-related diseases among people with similar life experiences? How can stress-induced damage be quantified? These and many other questions still challenge investigators.This article reviews the long-term effect of the physiologic response to stress, which I refer to as allostatic load.2 Allostasis — the ability to achieve stability through change3 — . . .
Through the study of transcriptional activation in response to interferon alpha (IFN-alpha) and interferon gamma (IFN-gamma), a previously unrecognized direct signal transduction pathway to the nucleus has been uncovered: IFN-receptor interaction at the cell surface leads to the activation of kinases of the Jak family that then phosphorylate substrate proteins called STATs (signal transducers and activators of transcription). The phosphorylated STAT proteins move to the nucleus, bind specific DNA elements, and direct transcription. Recognition of the molecules involved in the IFN-alpha and IFN-gamma pathway has led to discoveries that a number of STAT family members exist and that other polypeptide ligands also use the Jak-STAT molecules in signal transduction.
In 2008 we published the first set of guidelines for standardizing research in autophagy. Since then, research on this topic has continued to accelerate, and many new scientists have entered the field. Our knowledge base and relevant new technologies have also been expanding. Accordingly, it is important to update these guidelines for monitoring autophagy in different organisms. Various reviews have described the range of assays that have been used for this purpose. Nevertheless, there continues to be confusion regarding acceptable methods to measure autophagy, especially in multicellular eukaryotes. For example, a key point that needs to be emphasized is thatthere is a difference between measurements that monitor the numbers or volume of autophagic elements (e.g., autophagosomes or autolysosomes) at any stage of the autophagic process versus those that measure flux through the autophagy pathway (i.e., the completeprocess including the amount and rate of cargo sequestered and degraded). In particular, a block in macroautophagy that results in autophagosome accumulation must be differentiated from stimuli that increase autophagic activity, defined as increasedautophagy induction coupled with increased delivery to, and degradation within, lysosomes (inmost higher eukaryotes and some protists such as Dictyostelium) or the vacuole (in plants and fungi). In other words, it is especially important that investigators new to the field understand that the appearance of more autophagosomes does not necessarily equate with more autophagy. In fact, in manycases, autophagosomes accumulate because of a block in trafficking to lysosomes without a concomitant change in autophagosome biogenesis, whereas an increase in autolysosomes may reflect a reduction in degradative activity. It is worth emphasizing here that lysosomal digestion is a stage of autophagy and evaluating its competence is a crucial part of the evaluation of autophagic flux, or complete autophagy. Here, we present a set of guidelines for the selection and interpretation of methods for use by investigators who aim to examine macroautophagy and related processes, as well as forreviewers who need to provide realistic and reasonable critiques of papers that are focused on these processes. These guidelines are not meant to be a formulaic set of rules, because the appropriate assays depend in part on the question being asked and the system being used. In addition, we emphasize that no individual assay is guaranteed to be the most appropriate one in every situation, and we strongly recommend the use of multipleassays to monitor autophagy. Along these lines, because of the potential for pleiotropic effects due to blocking autophagy through genetic manipulation, it is imperative to target by gene knockout or RNA interference more than one autophagyrelated protein. In addition, some individual Atg proteins, or groups of proteins, are involved in other cellular pathways implying that not all Atg proteins can be used as a specific marker for an autophagic process. In these guidelines, we consider these various methods of assessing autophagy and what information can, or cannot, be obtained from them. Finally, by discussing the merits and limits of particular assays, we hope to encourage technical innovation in the field.
Clonal populations of cells exhibit substantial phenotypic variation. Such heterogeneity can be essential for many biological processes and is conjectured to arise from stochasticity, or noise, in gene expression. We constructed strains of Escherichia coli that enable detection of noise and discrimination between the two mechanisms by which it is generated. Both stochasticity inherent in the biochemical process of gene expression (intrinsic noise) and fluctuations in other cellular components (extrinsic noise) contribute substantially to overall variation. Transcription rate, regulatory dynamics, and genetic factors control the amplitude of noise. These results establish a quantitative foundation for modeling noise in genetic networks and reveal how low intracellular copy numbers of molecules can fundamentally limit the precision of gene regulation.
Adaptation in the face of potentially stressful challenges involves activation of neural, neuroendocrine and neuroendocrine-immune mechanisms. This has been called "allostasis" or "stability through change" by Sterling and Eyer (Fisher S., Reason J. (eds): Handbook of Life Stress, Cognition and Health. J. Wiley Ltd. 1988, p. 631), and allostasis is an essential component of maintaining homeostasis. When these adaptive systems are turned on and turned off again efficiently and not too frequently, the body is able to cope effectively with challenges that it might not otherwise survive. However, there are a number of circumstances in which allostatic systems may either be overstimulated or not perform normally, and this condition has been termed "allostatic load" or the price of adaptation (McEwen and Stellar, Arch. Int. Med. 1993; 153: 2093.). Allostatic load can lead to disease over long periods. Types of allostatic load include (1) frequent activation of allostatic systems; (2) failure to shut off allostatic activity after stress; (3) inadequate response of allostatic systems leading to elevated activity of other, normally counter-regulated allostatic systems after stress. Examples will be given for each type of allostatic load from research pertaining to autonomic, CNS, neuroendocrine, and immune system activity. The relationship of allostatic load to genetic and developmental predispositions to disease is also considered.
The brain is the key organ of the response to stress because it determines what is threatening and, therefore, potentially stressful, as well as the physiological and behavioral responses which can be either adaptive or damaging. Stress involves two-way communication between the brain and the cardiovascular, immune, and other systems via neural and endocrine mechanisms. Beyond the "flight-or-fight" response to acute stress, there are events in daily life that produce a type of chronic stress and lead over time to wear and tear on the body ("allostatic load"). Yet, hormones associated with stress protect the body in the short-run and promote adaptation ("allostasis"). The brain is a target of stress, and the hippocampus was the first brain region, besides the hypothalamus, to be recognized as a target of glucocorticoids. Stress and stress hormones produce both adaptive and maladaptive effects on this brain region throughout the life course. Early life events influence life-long patterns of emotionality and stress responsiveness and alter the rate of brain and body aging. The hippocampus, amygdala, and prefrontal cortex undergo stress-induced structural remodeling, which alters behavioral and physiological responses. As an adjunct to pharmaceutical therapy, social and behavioral interventions such as regular physical activity and social support reduce the chronic stress burden and benefit brain and body health and resilience.
Dendritic cells are a system of antigen presenting cells that function to initiate several immune responses such as the sensitization of MHC-restricted T cells, the rejection of organ transplants, and the formation of T-dependent antibodies. Dendritic cells are found in many nonlymphoid tissues but can migrate via the afferent lymph or the blood stream to the T-dependent areas of lymphoid organs. In skin, the immunostimulatory function of dendritic cells is enhanced by cytokines, especially GM-CSF. After foreign proteins are administered in situ, dendritic cells are a principal reservoir of immunogen. In vitro studies indicate that dendritic cells only process proteins for a short period of time, when the rate of synthesis of MHC products and content of acidic endocytic vesicles are high. Antigen processing is selectively dampened after a day in culture, but the capacity to stimulate responses to surface bound peptides and mitogens remains strong. Dendritic cells are motile, and efficiently cluster and activate T cells that are specific for stimuli on the cell surface. High levels of MHC class-I and -II products and several adhesins, such as ICAM-1 and LFA-3, likely contribute to these functions. Therefore dendritic cells are specialized to mediate several physiologic components of immunogenicity such as the acquisition of antigens in tissues, the migration to lymphoid organs, and the identification and activation of antigen-specific T cells. The function of these presenting cells in immunologic tolerance is just beginning to be studied.
Age-related macular degeneration (AMD) is a major cause of blindness in the elderly. We report a genome-wide screen of 96 cases and 50 controls for polymorphisms associated with AMD. Among 116,204 single-nucleotide polymorphisms genotyped, an intronic and common variant in the complement factor H gene (CFH) is strongly associated with AMD (nominal P value <10(-7)). In individuals homozygous for the risk allele, the likelihood of AMD is increased by a factor of 7.4 (95% confidence interval 2.9 to 19). Resequencing revealed a polymorphism in linkage disequilibrium with the risk allele representing a tyrosine-histidine change at amino acid 402. This polymorphism is in a region of CFH that binds heparin and C-reactive protein. The CFH gene is located on chromosome 1 in a region repeatedly linked to AMD in family-based studies.
The gene product of the ob locus is important in the regulation of body weight. The ob product was shown to be present as a 16-kilodalton protein in mouse and human plasma but was undetectable in plasma from C57BL/6J ob/ob mice. Plasma levels of this protein were increased in diabetic (db) mice, a mutant thought to be resistant to the effects of ob. Daily intraperitoneal injections of either mouse or human recombinant OB protein reduced the body weight of ob/ob mice by 30 percent after 2 weeks of treatment with no apparent toxicity but had no effect on db/db mice. The protein reduced food intake and increased energy expenditure in ob/ob mice. Injections of wild-type mice twice daily with the mouse protein resulted in a sustained 12 percent weight loss, decreased food intake, and a reduction of body fat from 12.2 to 0.7 percent. These data suggest that the OB protein serves an endocrine function to regulate body fat stores.
Oligonucleotide arrays can provide a broad picture of the state of the cell, by monitoring the expression level of thousands of genes at the same time. It is of interest to develop techniques for extracting useful information from the resulting data sets. Here we report the application of a two-way clustering method for analyzing a data set consisting of the expression patterns of different cell types. Gene expression in 40 tumor and 22 normal colon tissue samples was analyzed with an Affymetrix oligonucleotide array complementary to more than 6,500 human genes. An efficient two-way clustering algorithm was applied to both the genes and the tissues, revealing broad coherent patterns that suggest a high degree of organization underlying gene expression in these tissues. Coregulated families of genes clustered together, as demonstrated for the ribosomal proteins. Clustering also separated cancerous from noncancerous tissue and cell lines from in vivo tissues on the basis of subtle distributed patterns of genes even when expression of individual genes varied only slightly between the tissues. Two-way clustering thus may be of use both in classifying genes into functional groups and in classifying tissues based on gene expression.
Leukocytes respond to lipopolysaccharide (LPS) at nanogram per milliliter concentrations with secretion of cytokines such as tumor necrosis factor-alpha (TNF-alpha). Excess secretion of TNF-alpha causes endotoxic shock, an often fatal complication of infection. LPS in the bloodstream rapidly binds to the serum protein, lipopolysaccharide binding protein (LBP), and cellular responses to physiological levels of LPS are dependent on LBP. CD14, a differentiation antigen of monocytes, was found to bind complexes of LPS and LBP, and blockade of CD14 with monoclonal antibodies prevented synthesis of TNF-alpha by whole blood incubated with LPS. Thus, LPS may induce responses by interacting with a soluble binding protein in serum that then binds the cell surface protein CD14.
STATs (signal transducers and activators of transcription) are a family of latent cytoplasmic proteins that are activated to participate in gene control when cells encounter various extracellular polypeptides. Biochemical and molecular genetic explorations have defined a single tyrosine phosphorylation site and, in a dimeric partner molecule, an Src homology 2 (SH2) phosphotyrosine-binding domain, a DNA interaction domain, and a number of protein-protein interaction domains (with receptors, other transcription factors, the transcription machinery, and perhaps a tyrosine phosphatase). Mouse genetics experiments have defined crucial roles for each known mammalian STAT. The discovery of a STAT in Drosophila, and most recently in Dictyostelium discoideum, implies an ancient evolutionary origin for this dual-function set of proteins.
MicroRNAs (miRNAs) interact with target mRNAs at specific sites to induce cleavage of the message or inhibit translation. The specific function of most mammalian miRNAs is unknown. We have predicted target sites on the 3' untranslated regions of human gene transcripts for all currently known 218 mammalian miRNAs to facilitate focused experiments. We report about 2,000 human genes with miRNA target sites conserved in mammals and about 250 human genes conserved as targets between mammals and fish. The prediction algorithm optimizes sequence complementarity using position-specific rules and relies on strict requirements of interspecies conservation. Experimental support for the validity of the method comes from known targets and from strong enrichment of predicted targets in mRNAs associated with the fragile X mental retardation protein in mammals. This is consistent with the hypothesis that miRNAs act as sequence-specific adaptors in the interaction of ribonuclear particles with translationally regulated messages. Overrepresented groups of targets include mRNAs coding for transcription factors, components of the miRNA machinery, and other proteins involved in translational regulation, as well as components of the ubiquitin machinery, representing novel feedback loops in gene regulation. Detailed information about target genes, target processes, and open-source software for target prediction (miRanda) is available at http://www.microrna.org. Our analysis suggests that miRNA genes, which are about 1% of all human genes, regulate protein production for 10% or more of all human genes.
BACKGROUND: The recent discoveries of microRNA (miRNA) genes and characterization of the first few target genes regulated by miRNAs in Caenorhabditis elegans and Drosophila melanogaster have set the stage for elucidation of a novel network of regulatory control. We present a computational method for whole-genome prediction of miRNA target genes. The method is validated using known examples. For each miRNA, target genes are selected on the basis of three properties: sequence complementarity using a position-weighted local alignment algorithm, free energies of RNA-RNA duplexes, and conservation of target sites in related genomes. Application to the D. melanogaster, Drosophila pseudoobscura and Anopheles gambiae genomes identifies several hundred target genes potentially regulated by one or more known miRNAs. RESULTS: These potential targets are rich in genes that are expressed at specific developmental stages and that are involved in cell fate specification, morphogenesis and the coordination of developmental processes, as well as genes that are active in the mature nervous system. High-ranking target genes are enriched in transcription factors two-fold and include genes already known to be under translational regulation. Our results reaffirm the thesis that miRNAs have an important role in establishing the complex spatial and temporal patterns of gene activity necessary for the orderly progression of development and suggest additional roles in the function of the mature organism. In addition the results point the way to directed experiments to determine miRNA functions. CONCLUSIONS: The emerging combinatorics of miRNA target sites in the 3' untranslated regions of messenger RNAs are reminiscent of transcriptional regulation in promoter regions of DNA, with both one-to-many and many-to-one relationships between regulator and target. Typically, more than one miRNA regulates one message, indicative of cooperative translational control. Conversely, one miRNA may have several target genes, reflecting target multiplicity. As a guide to focused experiments, we provide detailed online information about likely target genes and binding sites in their untranslated regions, organized by miRNA or by gene and ranked by likelihood of match. The target prediction algorithm is freely available and can be applied to whole genome sequences using identified miRNA sequences.
Since its initial release in 2000, the human reference genome has covered only the euchromatic fraction of the genome, leaving important heterochromatic regions unfinished. Addressing the remaining 8% of the genome, the Telomere-to-Telomere (T2T) Consortium presents a complete 3.055 billion-base pair sequence of a human genome, T2T-CHM13, that includes gapless assemblies for all chromosomes except Y, corrects errors in the prior references, and introduces nearly 200 million base pairs of sequence containing 1956 gene predictions, 99 of which are predicted to be protein coding. The completed regions include all centromeric satellite arrays, recent segmental duplications, and the short arms of all five acrocentric chromosomes, unlocking these complex regions of the genome to variational and functional studies.