Max Planck Institute of Molecular Cell Biology and Genetics
facilityDresden, Saxony, Germany
Research output, citation impact, and the most-cited recent papers from Max Planck Institute of Molecular Cell Biology and Genetics (Germany). Aggregated across the NobleBlocks index of 300M+ scholarly works.
Top-cited papers from Max Planck Institute of Molecular Cell Biology and Genetics
The many functional partnerships and interactions that occur between proteins are at the core of cellular processing and their systematic characterization helps to provide context in molecular systems biology. However, known and predicted interactions are scattered over multiple resources, and the available data exhibit notable differences in terms of quality and completeness. The STRING database (http://string-db.org) aims to provide a critical assessment and integration of protein-protein interactions, including direct (physical) as well as indirect (functional) associations. The new version 10.0 of STRING covers more than 2000 organisms, which has necessitated novel, scalable algorithms for transferring interaction information between organisms. For this purpose, we have introduced hierarchical and self-consistent orthology annotations for all interacting proteins, grouping the proteins into families at various levels of phylogenetic resolution. Further improvements in version 10.0 include a completely redesigned prediction pipeline for inferring protein-protein associations from co-expression data, an API interface for the R computing environment and improved statistical analysis for enrichment tests in user-provided networks.
Cell membranes display a tremendous complexity of lipids and proteins designed to perform the functions cells require. To coordinate these functions, the membrane is able to laterally segregate its constituents. This capability is based on dynamic liquid-liquid immiscibility and underlies the raft concept of membrane subcompartmentalization. Lipid rafts are fluctuating nanoscale assemblies of sphingolipid, cholesterol, and proteins that can be stabilized to coalesce, forming platforms that function in membrane signaling and trafficking. Here we review the evidence for how this principle combines the potential for sphingolipid-cholesterol self-assembly with protein specificity to selectively focus membrane bioactivity.
In sexually reproducing organisms, embryos specify germ cells, which ultimately generate sperm and eggs. In Caenorhabditis elegans, the first germ cell is established when RNA and protein-rich P granules localize to the posterior of the one-cell embryo. Localization of P granules and their physical nature remain poorly understood. Here we show that P granules exhibit liquid-like behaviors, including fusion, dripping, and wetting, which we used to estimate their viscosity and surface tension. As with other liquids, P granules rapidly dissolved and condensed. Localization occurred by a biased increase in P granule condensation at the posterior. This process reflects a classic phase transition, in which polarity proteins vary the condensation point across the cell. Such phase transitions may represent a fundamental physicochemical mechanism for structuring the cytoplasm.
Since its initial release in 2000, the human reference genome has covered only the euchromatic fraction of the genome, leaving important heterochromatic regions unfinished. Addressing the remaining 8% of the genome, the Telomere-to-Telomere (T2T) Consortium presents a complete 3.055 billion-base pair sequence of a human genome, T2T-CHM13, that includes gapless assemblies for all chromosomes except Y, corrects errors in the prior references, and introduces nearly 200 million base pairs of sequence containing 1956 gene predictions, 99 of which are predicted to be protein coding. The completed regions include all centromeric satellite arrays, recent segmental duplications, and the short arms of all five acrocentric chromosomes, unlocking these complex regions of the genome to variational and functional studies.
Cells organize many of their biochemical reactions in non-membrane compartments. Recent evidence has shown that many of these compartments are liquids that form by phase separation from the cytoplasm. Here we discuss the basic physical concepts necessary to understand the consequences of liquid-like states for biological functions.
Abstract High-quality and complete reference genome assemblies are fundamental for the application of genomics to biology, disease, and biodiversity conservation. However, such assemblies are available for only a few non-microbial species 1–4 . To address this issue, the international Genome 10K (G10K) consortium 5,6 has worked over a five-year period to evaluate and develop cost-effective methods for assembling highly accurate and nearly complete reference genomes. Here we present lessons learned from generating assemblies for 16 species that represent six major vertebrate lineages. We confirm that long-read sequencing technologies are essential for maximizing genome quality, and that unresolved complex repeats and haplotype heterozygosity are major sources of assembly error when not handled correctly. Our assemblies correct substantial errors, add missing sequence in some of the best historical reference genomes, and reveal biological discoveries. These include the identification of many false gene duplications, increases in gene sizes, chromosome rearrangements that are specific to lineages, a repeated independent chromosome breakpoint in bat genomes, and a canonical GC-rich pattern in protein-coding genes and their regulatory regions. Adopting these lessons, we have embarked on the Vertebrate Genomes Project (VGP), an international effort to generate high-quality, complete reference genomes for all of the roughly 70,000 extant vertebrate species and to help to enable a new era of discovery across the life sciences.
Accurate profiling of lipidomes relies upon the quantitative and unbiased recovery of lipid species from analyzed cells, fluids, or tissues and is usually achieved by two-phase extraction with chloroform. We demonstrated that methyl-tert-butyl ether (MTBE) extraction allows faster and cleaner lipid recovery and is well suited for automated shotgun profiling. Because of MTBE's low density, lipid-containing organic phase forms the upper layer during phase separation, which simplifies its collection and minimizes dripping losses. Nonextractable matrix forms a dense pellet at the bottom of the extraction tube and is easily removed by centrifugation. Rigorous testing demonstrated that the MTBE protocol delivers similar or better recoveries of species of most all major lipid classes compared with the "gold-standard" Folch or Bligh and Dyer recipes.
MOTIVATION: Modern anatomical and developmental studies often require high-resolution imaging of large specimens in three dimensions (3D). Confocal microscopy produces high-resolution 3D images, but is limited by a relatively small field of view compared with the size of large biological specimens. Therefore, motorized stages that move the sample are used to create a tiled scan of the whole specimen. The physical coordinates provided by the microscope stage are not precise enough to allow direct reconstruction (Stitching) of the whole image from individual image stacks. RESULTS: To optimally stitch a large collection of 3D confocal images, we developed a method that, based on the Fourier Shift Theorem, computes all possible translations between pairs of 3D images, yielding the best overlap in terms of the cross-correlation measure and subsequently finds the globally optimal configuration of the whole group of 3D images. This method avoids the propagation of errors by consecutive registration steps. Additionally, to compensate the brightness differences between tiles, we apply a smooth, non-linear intensity transition between the overlapping images. Our stitching approach is fast, works on 2D and 3D images, and for small image sets does not require prior knowledge about the tile configuration. AVAILABILITY: The implementation of this method is available as an ImageJ plugin distributed as a part of the Fiji project (Fiji is just ImageJ: http://pacific.mpi-cbg.de/).
eggNOG is a public resource that provides Orthologous Groups (OGs) of proteins at different taxonomic levels, each with integrated and summarized functional annotations. Developments since the latest public release include changes to the algorithm for creating OGs across taxonomic levels, making nested groups hierarchically consistent. This allows for a better propagation of functional terms across nested OGs and led to the novel annotation of 95 890 previously uncharacterized OGs, increasing overall annotation coverage from 67% to 72%. The functional annotations of OGs have been expanded to also provide Gene Ontology terms, KEGG pathways and SMART/Pfam domains for each group. Moreover, eggNOG now provides pairwise orthology relationships within OGs based on analysis of phylogenetic trees. We have also incorporated a framework for quickly mapping novel sequences to OGs based on precomputed HMM profiles. Finally, eggNOG version 4.5 incorporates a novel data set spanning 2605 viral OGs, covering 5228 proteins from 352 viral proteomes. All data are accessible for bulk downloading, as a web-service, and through a completely redesigned web interface. The new access points provide faster searches and a number of new browsing and visualization capabilities, facilitating the needs of both experts and less experienced users. eggNOG v4.5 is available at http://eggnog.embl.de.
Although fluorescence microscopy provides a crucial window into the physiology of living specimens, many biological processes are too fragile, are too small, or occur too rapidly to see clearly with existing tools. We crafted ultrathin light sheets from two-dimensional optical lattices that allowed us to image three-dimensional (3D) dynamics for hundreds of volumes, often at subsecond intervals, at the diffraction limit and beyond. We applied this to systems spanning four orders of magnitude in space and time, including the diffusion of single transcription factor molecules in stem cell spheroids, the dynamic instability of mitotic microtubules, the immunological synapse, neutrophil motility in a 3D matrix, and embryogenesis in Caenorhabditis elegans and Drosophila melanogaster. The results provide a visceral reminder of the beauty and the complexity of living systems.
BACKGROUND: PacBio high fidelity (HiFi) sequencing reads are both long (15-20 kb) and highly accurate (> Q20). Because of these properties, they have revolutionised genome assembly leading to more accurate and contiguous genomes. In eukaryotes the mitochondrial genome is sequenced alongside the nuclear genome often at very high coverage. A dedicated tool for mitochondrial genome assembly using HiFi reads is still missing. RESULTS: MitoHiFi was developed within the Darwin Tree of Life Project to assemble mitochondrial genomes from the HiFi reads generated for target species. The input for MitoHiFi is either the raw reads or the assembled contigs, and the tool outputs a mitochondrial genome sequence fasta file along with annotation of protein and RNA genes. Variants arising from heteroplasmy are assembled independently, and nuclear insertions of mitochondrial sequences are identified and not used in organellar genome assembly. MitoHiFi has been used to assemble 374 mitochondrial genomes (368 Metazoa and 6 Fungi species) for the Darwin Tree of Life Project, the Vertebrate Genomes Project and the Aquatic Symbiosis Genome Project. Inspection of 60 mitochondrial genomes assembled with MitoHiFi for species that already have reference sequences in public databases showed the widespread presence of previously unreported repeats. CONCLUSIONS: MitoHiFi is able to assemble mitochondrial genomes from a wide phylogenetic range of taxa from Pacbio HiFi data. MitoHiFi is written in python and is freely available on GitHub ( https://github.com/marcelauliano/MitoHiFi ). MitoHiFi is available with its dependencies as a Docker container on GitHub (ghcr.io/marcelauliano/mitohifi:master).
The genomes of prokaryotes and eukaryotic organelles are usually circular as are most plasmids and viral genomes. In contrast, the nuclear genomes of eukaryotes are organized on linear chromosomes, which require mechanisms to protect and replicate DNA ends. Eukaryotes navigate these problems with the advent of telomeres, protective nucleoprotein complexes at the ends of linear chromosomes, and telomerase, the enzyme that maintains the DNA in these structures. Mammalian telomeres contain a specific protein complex, shelterin, that functions to protect chromosome ends from all aspects of the DNA damage response and regulates telomere maintenance by telomerase. Recent experiments, discussed here, have revealed how shelterin represses the ATM and ATR kinase signaling pathways and hides chromosome ends from nonhomologous end joining and homology-directed repair.
Interactions between proteins and small molecules are an integral part of biological processes in living organisms. Information on these interactions is dispersed over many databases, texts and prediction methods, which makes it difficult to get a comprehensive overview of the available evidence. To address this, we have developed STITCH ('Search Tool for Interacting Chemicals') that integrates these disparate data sources for 430 000 chemicals into a single, easy-to-use resource. In addition to the increased scope of the database, we have implemented a new network view that gives the user the ability to view binding affinities of chemicals in the interaction network. This enables the user to get a quick overview of the potential effects of the chemical on its interaction partners. For each organism, STITCH provides a global network; however, not all proteins have the same pattern of spatial expression. Therefore, only a certain subset of interactions can occur simultaneously. In the new, fifth release of STITCH, we have implemented functionality to filter out the proteins and chemicals not associated with a given tissue. The STITCH database can be downloaded in full, accessed programmatically via an extensive API, or searched via a redesigned web interface at http://stitch.embl.de.
Caveolae are plasma membrane invaginations that may play an important role in numerous cellular processes including transport, signaling, and tumor suppression. By targeted disruption of caveolin-1, the main protein component of caveolae, we generated mice that lacked caveolae. The absence of this organelle impaired nitric oxide and calcium signaling in the cardiovascular system, causing aberrations in endothelium-dependent relaxation, contractility, and maintenance of myogenic tone. In addition, the lungs of knockout animals displayed thickening of alveolar septa caused by uncontrolled endothelial cell proliferation and fibrosis, resulting in severe physical limitations in caveolin-1-disrupted mice. Thus, caveolin-1 and caveolae play a fundamental role in organizing multiple signaling pathways in the cell.
Unwanted side effects of drugs are a burden on patients and a severe impediment in the development of new drugs. At the same time, adverse drug reactions (ADRs) recorded during clinical trials are an important source of human phenotypic data. It is therefore essential to combine data on drugs, targets and side effects into a more complete picture of the therapeutic mechanism of actions of drugs and the ways in which they cause adverse reactions. To this end, we have created the SIDER ('Side Effect Resource', http://sideeffects.embl.de) database of drugs and ADRs. The current release, SIDER 4, contains data on 1430 drugs, 5880 ADRs and 140 064 drug-ADR pairs, which is an increase of 40% compared to the previous version. For more fine-grained analyses, we extracted the frequency with which side effects occur from the package inserts. This information is available for 39% of drug-ADR pairs, 19% of which can be compared to the frequency under placebo treatment. SIDER furthermore contains a data set of drug indications, extracted from the package inserts using Natural Language Processing. These drug indications are used to reduce the rate of false positives by identifying medical terms that do not correspond to ADRs.
P granules and other RNA/protein bodies are membrane-less organelles that may assemble by intracellular phase separation, similar to the condensation of water vapor into droplets. However, the molecular driving forces and the nature of the condensed phases remain poorly understood. Here, we show that the Caenorhabditis elegans protein LAF-1, a DDX3 RNA helicase found in P granules, phase separates into P granule-like droplets in vitro. We adapt a microrheology technique to precisely measure the viscoelasticity of micrometer-sized LAF-1 droplets, revealing purely viscous properties highly tunable by salt and RNA concentration. RNA decreases viscosity and increases molecular dynamics within the droplet. Single molecule FRET assays suggest that this RNA fluidization results from highly dynamic RNA-protein interactions that emerge close to the droplet phase boundary. We demonstrate than an N-terminal, arginine/glycine rich, intrinsically disordered protein (IDP) domain of LAF-1 is necessary and sufficient for both phase separation and RNA-protein interactions. In vivo, RNAi knockdown of LAF-1 results in the dissolution of P granules in the early embryo, with an apparent submicromolar phase boundary comparable to that measured in vitro. Together, these findings demonstrate that LAF-1 is important for promoting P granule assembly and provide insight into the mechanism by which IDP-driven molecular interactions give rise to liquid phase organelles with tunable properties.
The field of image denoising is currently dominated by discriminative deep learning methods that are trained on pairs of noisy input and clean target images. Recently it has been shown that such methods can also be trained without clean targets. Instead, independent pairs of noisy images can be used, in an approach known as Noise2Noise (N2N). Here, we introduce Noise2Void (N2V), a training scheme that takes this idea one step further. It does not require noisy image pairs, nor clean target images. Consequently, N2V allows us to train directly on the body of data to be denoised and can therefore be applied when other methods cannot. Especially interesting is the application to biomedical image data, where the acquisition of training targets, clean or noisy, is frequently not possible. We compare the performance of N2V to approaches that have either clean target images and/or noisy image pairs available. Intuitively, N2V cannot be expected to outperform methods that have more information available during training. Still, we observe that the denoising performance of Noise2Void drops in moderation and compares favorably to training-free denoising methods.
For most intracellular structures with larger than molecular dimensions, little is known about the connection between underlying molecular activities and higher order organization such as size and shape. Here, we show that both the size and shape of the amphibian oocyte nucleolus ultimately arise because nucleoli behave as liquid-like droplets of RNA and protein, exhibiting characteristic viscous fluid dynamics even on timescales of < 1 min. We use these dynamics to determine an apparent nucleolar viscosity, and we show that this viscosity is ATP-dependent, suggesting a role for active processes in fluidizing internal contents. Nucleolar surface tension and fluidity cause their restructuring into spherical droplets upon imposed mechanical deformations. Nucleoli exhibit a broad distribution of sizes with a characteristic power law, which we show is a consequence of spontaneous coalescence events. These results have implications for the function of nucleoli in ribosome subunit processing and provide a physical link between activity within a macromolecular assembly and its physical properties on larger length scales.
Although the exact etiology of Alzheimer's disease (AD) is a topic of debate, the consensus is that the accumulation of beta-amyloid (Abeta) peptides in the senile plaques is one of the hallmarks of the progression of the disease. The Abeta peptide is formed by the amyloidogenic cleavage of the amyloid precursor protein (APP) by beta- and gamma-secretases. The endocytic system has been implicated in the cleavages leading to the formation of Abeta. However, the identity of the intracellular compartment where the amyloidogenic secretases cleave and the mechanism by which the intracellularly generated Abeta is released into the extracellular milieu are not clear. Here, we show that beta-cleavage occurs in early endosomes followed by routing of Abeta to multivesicular bodies (MVBs) in HeLa and N2a cells. Subsequently, a minute fraction of Abeta peptides can be secreted from the cells in association with exosomes, intraluminal vesicles of MVBs that are released into the extracellular space as a result of fusion of MVBs with the plasma membrane. Exosomal proteins were found to accumulate in the plaques of AD patient brains, suggesting a role in the pathogenesis of AD.
Cholesterol plays an indispensable role in regulating the properties of cell membranes in mammalian cells. Recent advances suggest that cholesterol exerts many of its actions mainly by maintaining sphingolipid rafts in a functional state. How rafts contribute to cholesterol metabolism and transport in the cell is still an open issue. It has long been known that cellular cholesterol levels are precisely controlled by biosynthesis, efflux from cells, and influx of lipoprotein cholesterol into cells. The regulation of cholesterol homeostasis is now receiving a new focus, and this changed perspective may throw light on diseases caused by cholesterol excess, the prime example being atherosclerosis.