Abstract
The cytoplasm is a densely packed environment filled with macromolecules with hindered diffusion. Molecular simulation of the diffusion of biomolecules under such macromolecular crowding conditions requires the definition of a simulation cell with a cytoplasmic-like composition. This has been previously done for prokaryote cells (E. coli) but not for eukaryote cells such as yeast as a model organism. Yeast proteomics datasets vary widely in terms of cell growth conditions, the technique used to determine protein composition, the reported relative abundance of proteins, and the units in which abundances are reported. We determined that the gene ontology profiles of the most abundant proteins across these datasets are similar, but their abundances vary greatly. To overcome this problem, we chose five mass spectrometry proteomics datasets that fulfilled the following criteria: high internal consistency, consistency with published experimental data, and freedom from GFP-tagging artifacts. Using these datasets, the contents of a simulation cell containing a single 80S ribosome were defined, such that the macromolecular density and the mass ratio of ribosomal-to-cytoplasmic proteins were consistent with experiment and chosen datasets. Finally, multiple tRNAs were added, consistent with their experimentally-determined number in the yeast cell. The resulting composition can be readily used in molecular simulations representative of yeast cytoplasmic macromolecular crowding conditions to characterize a variety of phenomena, such as protein diffusion, protein-protein interactions and biological processes such as protein translation.
Introduction
The environment inside cells is densely packed, termed macromolecular crowding, the extent of which varies throughout the different growth and differentiation stages of the cell, as well as according to its type and volume (Nakano et al., ). A typical cell has a macromolecular concentration in the range 100–450 g/L, with 5–40% of its volume being occupied by macromolecules (Feig et al., ). Therefore, the space available for the free diffusion of metabolites and other macromolecules is greatly reduced, leading to what is known as an excluded volume effect. This reduces diffusion and favors more compact protein conformations and protein association. Transient aggregation of proteins is favored in crowded systems and is correlated with slower diffusion (Nawrocki et al., ). Macromolecules reduce the amount of bulk-like water in the cell by reducing the amount of water molecules present beyond the second solvation layer (Harada et al., ). As a consequence, a 40% reduction in the dielectric constant of yeast cells compared to that of a dilute solution has been determined (Asami et al., ; Tanizaki et al., ), leading to an increase in electrostatic interactions between molecules. Hindered diffusion due to macromolecular crowding, on the other hand, increases the probability of ligands being in the vicinity of their receptors in what is termed caging effect, which enhances reaction rates (Feig et al., ). Cells are believed to maintain their macromolecular concentration within a very small range in a process now termed “homeocrowding” (Van Den Berg et al., 2017). Moreover, it has been shown that the diffusion coefficient of molecules depends not only on the macromolecular concentration but also on the composition of the solution (Wang et al., 2010). Molecular crowding inside cells affects various biochemical processes such as protein translation. The diffusion of tRNA complexes in the cytoplasmic environment is hindered by crowding, in turn affecting the rate of translation (Klumpp et al., ).
Molecular dynamics (MD) simulations can be used to characterize the complex nature of the effects of macromolecular crowding, including effects on the diffusion of tRNAs and their binding to cytoplasmic ribosomes during translation. Two prior studies of the cytoplasm have focused on prokaryotic systems (E. coli). In one study, 118 protein molecules were chosen on the basis of their mole percentage in the cytosol, with the number of ribosomes being scaled down based on abundances reported at cell level and a total macromolecular density of 340 g/L (Ridgway et al., ). Each protein molecule was represented as a sphere, whilst tRNAs were not included at all (Ridgway et al., ). In a second study, 51 different types of macromolecules were considered, out of which 45 were proteins and which accounted for 86% of the total cytoplasmic protein mass reported by the proteomics dataset used with a macromolecular concentration of 275 g/L. The simulation cell also included three types of tRNAs (tRNA-Gln, tRNA-Phe, and tRNA-Cys) and 10 ribosomes in their corresponding subunits. The volume corresponding to lipids, lipopolysaccharides, mRNA, DNA, murein, and glycogen was accounted for by increasing the concentration of protein in the simulation cell (McGuffee and Elcock, ). In a more recent cytoplasmic model, developed for Mycoplasma genitalium, the simulation cell comprised more than 1,000 protein molecules, 275 tRNAs, nucleotides, metabolites, ions, and a total of 26 million water molecules represented atomistically with a macromolecular density of 291.5 g/L (Feig et al., ). To our knowledge, an equivalent representative definition of the eukaryotic cytoplasm has not been reported in the literature. The key challenges in defining such a simulation cell include identification of the required proteomics datasets and defining appropriate criteria to minimize the size of the cell whilst retaining the properties of the cytoplasmic environment.
In this study, we sought to address the lack of a standard molecular simulation environment for eukaryotes by defining the contents of a simulation cell based on the abundances of proteins, tRNAs and ribosomes in the yeast cytoplasm. A recent yeast proteomics dataset (Ho et al., ) unified abundance data from 21 different datasets, comprising a range of mass spectrometry (MS)-derived datasets, datasets based on green fluorescent protein (GFP)-tagging of yeast proteins and GFP flow cytometry and also a tandem affinity purification (TAP-tagging)-immunoblot dataset. We employed an in-depth proteomics survey of these datasets in order to define a molecular simulation environment for a model eukaryote cell. However, these datasets vary in terms of the growth conditions used to culture the cells, the cellular growth phase, the units in which abundances are reported, and the technique used to measure them. It was therefore necessary to investigate how these factors affect protein abundances reported across the range of datasets. We characterized the internal consistency amongst the datasets and their agreement with other published experimental data, leading to the selection of a proteome composition for the yeast cytoplasmic environment. Consideration of additional experimental data on the macromolecular density and the mass ratio of ribosomal-to-cytoplasmic proteins in the cytoplasm was also used, allowing the definition of the contents of a molecular simulation cell representative of the yeast cytoplasm.
Methods
Definition of a Eukaryote Cell Simulation Environment
Previous reports of the number of ribosomes in yeast cytoplasm were taken from cell population scale experiments (Waldron and Lacroute, 1975) and from cell tomography experiments at single cell level (Yamaguchi et al., 2011), and were compared with the numbers calculated from proteomics datasets. The volume percentage of individual components of the yeast cell were also obtained from cell tomography studies (Yamaguchi et al., 2011), which are in agreement with other cell tomography experiments (Wei et al., 2012). Furthermore, we used the recently published unified yeast proteomics dataset that covers a total of 5,391 proteins (Ho et al., ).
Proteins associated with the nucleus, cell wall, ribosomes, mitochondria, endoplasmic reticulum, and vacuoles were removed from the dataset with the help of GO-slim annotations (http://www.yeastgenome.org/) to assign cellular location to a given protein. Gene ontology analysis of the function of encoded proteins was performed using the webserver Funcassociate 3.0 (http://llama.mshri.on.ca/funcassociate/) (Berriz et al., ).
Statistical Analysis
The abundances reported for individual ribosomal proteins by any dataset were treated as multiple observations of the number of ribosomes (described in detail in the Results section). Based on this, pairwise statistical two-tailed t-tests for unequal variances between proteomics datasets were performed using an in-house code in MATLAB (https://github.com/BMMG-Curtin/FMOLB) to quantitatively understand the differences and similarities between datasets (Figure S1). Where multiple pairwise t-tests were conducted, the Bonferroni correction was applied to address type-I errors, whereby the critical alpha value is divided by the number of pairwise tests. In addition, p-values were adjusted using the Benjamini-Hochberg approach to address type-I errors and the results obtained were found to be qualitatively the same (Figure S1). The data was assumed to be normally distributed whilst conducting the above t-tests; therefore, a non-parametric Mann–Whitney U-test with the Bonferroni correction was also employed (Figure S2). The results of the U-test were also found to be qualitatively similar to the results obtained with the t-tests. Pairwise correlations between the functional ontological classes of proteins across different datasets were quantified using the Pearson's correlation coefficient. The Jaccard index was used to quantify the similarities between the ontological profiles obtained for each of the datasets.
Results
Analysis of Internal Consistency of Yeast Proteomics Datasets
In order to define the protein composition of a eukaryote molecular simulation cell, the recently published unified yeast proteomics dataset was used (Ho et al., ). This covers 5,391 genes with a total protein mass per yeast cell of 2.7 × 1012 Da, which is in good agreement with the total protein mass of a yeast cell previously reported to be 3 × 1012 Da (Sasidharan et al., ). This proteomics dataset comprises data integrated from 21 different datasets, which vary in the type of growth medium used to culture cells, their growth phase and the technique used to measure protein abundances.
The top 200 most abundant proteins were taken from each of the 21 datasets based on their mass (i.e., molecular mass multiplied by their abundance) and were found to account for ~70% of the total cytoplasmic protein mass (Figure 1). In order to assess the possible influence of cell culture conditions, growth phase and the method used to measure protein abundance on the composition of the yeast cytoplasm, the ontological classes of these proteins were assessed. The systematic names of these proteins were submitted to the Funcassociate 3.0 webserver, which detects over-representation of gene ontologies in a gene list. The number of proteins associated with each gene ontology class was identified for every dataset. Each pair of datasets was then compared by calculating the Pearson's correlation coefficient between the number of proteins associated with each gene ontology class. The Jaccard index was used to quantify the similarities between the sets of gene ontology classes obtained for every dataset. Despite the above differences between the datasets, a similar ontological landscape for the top 200 proteins in each of the datasets was observed, except for one dataset that used N-terminal GFP tagging, YOF (Yofe et al., 2016; Figure 2).
Figure 1
Figure 2
Although the gene ontology profiles of the top 200 cytoplasmic proteins are similar across datasets, significant differences in protein abundances were observed. For example, the average coefficient of variation (CV) (measured across the 21 datasets) for the cytoplasmic proteins is 78%. The differences are more marked in the case of ribosomal proteins (CV = 106%).
In order to investigate the internal consistency of the proteomic datasets and their agreement with other published data, ribosomal proteins were examined separately. The protein composition of ribosomes can be assumed to be fixed (Perry,
Depending on the consistency between datasets, the numbers reported for a given ribosomal protein across different datasets are expected to vary showing patterns in terms of experimental conditions. In order to test this, the abundances of different ribosomal proteins were compared across different datasets. Given the 1:1 stoichiometry for each ribosomal protein with respect to the ribosome (Warner, 1999), the abundance of each ribosomal protein in each dataset provided an estimate of the number of ribosomes per cell. The average number of ribosomal proteins was therefore calculated to derive an average ribosome per cell value for each dataset. The resulting values were then compared between datasets by performing multiple pairwise t-tests to determine any patterns arising from the growth media, growth phase or the technique used to measure protein abundance (Figure 3). High p-values were observed in the pairwise tests between the datasets derived from GFP-tagging of proteins, indicating consistency between them. On the other hand, no clear consistency was apparent within the MS datasets, and no patterns were observed that might be accounted for by the growth media or growth phase used during cell culture.
Figure 3

Testing of statistical difference between the abundance of ribosomal proteins in each of the datasets. Mass spectrometry-based datasets are shown in red on the axes, GFP datasets are shown in green on the axes and the TAP-immunoblot dataset is shown in white. Ribosomal protein numbers were not reported in the YOF dataset and, therefore, it is not included. The results of t-tests with p > (0.05/190) are colored dark blue and all others are colored light blue. GFP datasets exhibit a high level of consistency. There is also consistency among the first five MS datasets. However, there are no discernible patterns in terms of the growth media, growth phase or protein abundance units.
It has previously been reported that there are ribosomal proteins with extra-ribosomal functions in yeast (Lu et al.,
Selection of Datasets
Whilst the gene ontology profiles of the proteomics datasets are similar, they vary widely in the protein abundances reported. The ratio of the median of abundances reported by GFP datasets to the median of MS datasets was calculated for cytoplasmic and ribosomal proteins. We determined that for 74% of cytoplasmic proteins and 84% of ribosomal proteins the medians differ by more than 25%. The differences in the individual protein abundances between the GFP and MS datasets were reported to be possibly due to changes in protein or mRNA stability following GFP tagging (Ho et al.,
The number of ribosomes, calculated by taking the median of all ribosomal proteins reported in the GFP datasets, revealed an estimated 51,800 ribosomes per cell, whereas previously reported figures are 150,000–300,000 (Waldron and Lacroute, 1975) and 169,000–265,000 (Yamaguchi et al., 2011) ribosomes per cell. As discussed earlier, the abundances of ribosomal proteins reported in the GFP datasets are also widely spread, with an average CV of 103%, in contrast to the average CV of 69% in the MS datasets. It was thus decided to omit the GFP datasets from further consideration.
The first five (LU, PENG, KUL, LAW, and LAHT) MS datasets report abundances in absolute numbers, whereas the other MS datasets report normalized abundances (with respect to the average of the five MS datasets) (Ho et al.,
Constraints for the Definition of the Contents of a Simulation Cell
A molecular simulation cell should be designed to mimic the environment of the yeast cytoplasm. This requires the inclusion of three important constraints: macromolecular density, the mass ratio of ribosomal-to-cytoplasmic proteins, and the number of ribosomes in the simulation cell.
Macromolecular density is an indirect measure of the excluded volume and, therefore, crowding. The volume of yeast cell has been reported to be 42 μm3 (Jorgensen et al.,
It has been reported that the fractions of ribosomal protein (R-protein), translation protein (T-protein), fixed protein (Q), the proportion of which is independent of growth rate, and metabolic protein (P-protein), given by, ΦR, ΦT, ΦQ, and ΦP, respectively, are unique for a specific growth rate (Klumpp et al.,
where A is the total protein mass and C is the growth rate specific constant. The total Q- and P-protein content can be divided into cytoplasmic and non-cytoplasmic fractions. Therefore, the previous equation can be rewritten as
The last Equation (3) states the assumption that the mass ratio of cytoplasmic to non-cytoplasmic proteins is constant at a given growth rate, from which it follows that cytoplasmic fraction in Q- and P-proteins remains constant. Since the T-protein fraction is a growth rate-dependent constant, the mass ratio of ribosomal-to-total cytoplasmic proteins is constant at a given growth rate. This is the second constraint for the definition of the contents of a simulation cell. The mass ratio of ribosomal-to-cytoplasmic proteins (rib/cyt) was determined to be 0.2229.
The crystal structure of the ribosome is composed of 75 ribosomal proteins (Ben-Shem et al.,
Definition of the Contents of the Simulation Cell
The choice of five MS datasets reduced the number of cytoplasmic proteins with abundance data from 1,594 to 1,374; however, when calculating the macromolecular density of the cytoplasm, data from all 1,594 proteins was considered. The total mass of cytoplasmic proteins calculated using abundances in the unified dataset is 7.56 × 1011 Da. The median of the number of molecules reported for a given protein by the five chosen MS datasets was taken as the measure of its abundance in a typical yeast cell. The total mass of a given type of protein was calculated by multiplying its abundance (number of proteins per cell) by its molecular mass, and the protein list was then sorted in descending order of total mass. The top 200 proteins contribute, as mentioned earlier, about 70% of the total cytoplasmic protein mass. The top proteins from the list were chosen due to their significant contribution to the protein mass in the cytoplasm and their abundances were subsequently scaled down to their corresponding value in proportion to only one ribosome (calculated as the abundance “n” of a protein divided by the 126,213 ribosomes predicted in the MS datasets).
Each of the less abundant cytoplasmic proteins does not contribute significantly to the overall protein mass. However, their collective removal results in a significant loss in protein mass which needs to be accounted for in order to maintain the desired macromolecular density of the simulation cell. Additionally, a number of proteins will contribute to the cytoplasm in fractional units that are lost due to rounding. The number of protein molecules of each of the cytoplasmic proteins was thus multiplied by a scaling factor aimed at maintaining the overall macromolecular density of the simulation cell. The number of protein types was chosen such that their total mass contribution reflects the expected value of the rib/cyt ratio. This was achieved by testing multiple scaling factors under the above-described constraints. Use of a large scaling factor (e.g., 3.0) meant that the rib/cyt ratio could be reached with just 20 different types of proteins, amounting to 119 protein molecules. By contrast, the rib/cyt ratio could not be reached with very low scaling factors (e.g., <1.8). Although the total number of protein molecules remained in the range 120–130 with all of the scaling factors tested, the observed protein composition was affected significantly with the use of large scaling factors. A range of scaling factors meet the constraints of macromolecular density, rib/cyt ratio and the presence of one ribosome in the simulation cell. However, in order to maintain the most representative composition of cytoplasmic proteins, the lowest possible scaling factor of 1.803 was chosen. This resulted in a final list containing 128 protein molecules belonging to 70 types of proteins (Table S2).
Based on the constraint that there should be only one ribosome, the size of the simulation cell was calculated. A total of 126,213 ribosomes are assumed to be present in the cytoplasm, which has a volume of 27.3 μm3. This volume was scaled down to one ribosome unit, which for a cubic simulation cell results in a length of 560 Å. The number of tRNAs was scaled down from 3 million units per cell to the volume of the simulation box, resulting in 22 tRNA units. With one 80S ribosome, 128 protein molecules and 22 tRNAs, the resulting simulation cell has the required total macromolecular density of 90 g/L.
Discussion
This study shows that the ontological profiles of the most abundant proteins in yeast remains constant despite differences in growth medium and growth phase, indicating that the most abundant proteins constitute the fundamental biochemical framework of the cell. The abundances reported in GFP datasets are affected by tagging, particularly in the case of ribosomal proteins. This has been explained previously on the basis that ribosomal proteins form a compact structure in a single ribosome molecule and the tag attached to them affects their packing. Although this explains the low numbers of ribosomal proteins reported, the cause of the high CV of ribosomal proteins in GFP datasets (CV = 103%), indicating a selective effect of tagging, compared with that of MS datasets (CV = 69%) remains unclear. Moreover, the average number of ribosomes calculated using MS datasets that report abundances in relative units is very low (30,500 units). The causes behind this remain undetermined, although normalization of the data is a possible factor.
Unlike prokaryotic cells, eukaryotic cells have a sophisticated organization of cellular machinery into different organelles with varying macromolecular environments. In order to study the influence of this macromolecular environment, an accurate description of its composition is needed. This was achieved by assigning the cellular location of a protein from its gene annotation data (GO-slim data) and determining the volume percentage of cytoplasm in yeast from cell tomography experiments. The macromolecular density of yeast cytoplasm was found to be 90 g/L, which is three times lower than that of the cytoplasm of E. coli. Measurements of the diffusion coefficient of GFP in eukaryotic and prokaryotic cells indicate that the eukaryotic cytoplasm is less crowded (Ellis,
In conclusion, a simulation cell was defined such that the yeast cellular composition of proteins, the ribosome-to-cytoplasmic protein mass ratio and the macromolecular density are retained. This was achieved by increasing the relative proportion of the most abundant proteins under specific constraints. The resulting simulation cell contains 128 protein molecules belonging to 70 protein types, 22 tRNAs and one 80s ribosome within a cubic cell of 560 Å in length. The simulation cell contents act as a generic representation of the cytoplasm that can be used to study the diffusion and interactions of molecules in the yeast cytoplasmic environment.
Statements
Data availability statement
The datasets generated for this study can be found in the https://github.com/BMMG-Curtin/FMOLB.
Author contributions
VK conducted all of the analyses. VK and RM wrote the manuscript, which was further proofread by MR and IS. All authors conceived and designed this study and performed the interpretation of data.
Funding
VK gratefully acknowledges the receipt of a scholarship under the Aberdeen-Curtin Alliance collaborative Ph.D. program.
Acknowledgments
We thank Prof. Grant Brown (University of Toronto) for making the yeast proteomics datasets available to us.
Conflict of interest
The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.
Supplementary material
The Supplementary Material for this article can be found online at: https://www.frontiersin.org/articles/10.3389/fmolb.2019.00097/full#supplementary-material
References
1
AsamiK.HanaiT.KoizumiN. (1976). Dielectric properties of yeast cells. J. Membr. Biol.28, 169–180. 10.1007/BF01869695
2
Ben-ShemA.De LoubresseN. G.MelnikovS.JennerL.YusupovaG.YusupovM. (2011). The structure of the eukaryotic ribosome at 3.0 Å resolution. Science334, 1524–1529. 10.1126/science.1212642
3
BerrizG. F.BeaverJ. E.CenikC.TasanM.RothF. P. (2009). Next generation software for functional trend analysis. Bioinformatics25, 3043–3044. 10.1093/bioinformatics/btp498
4
BrekerM.GymrekM.SchuldinerM. (2013). A novel single-cell screening platform reveals proteome plasticity during yeast stress responses. J. Cell Biol.200, 839–850. 10.1083/jcb.201301120
5
ChongY. T.KohJ. L. Y.FriesenH.DuffyK.CoxM. J.MosesA.et al. (2015). Yeast proteome dynamics from single cell imaging and automated analysis. Cell161, 1413–1424. 10.1016/j.cell.2015.04.051
6
DavidsonG. S.JoeR. M.RoyS.MeirellesO.AllenC. P.WilsonM. R.et al. (2011). The proteomics of quiescent and nonquiescent cell differentiation in yeast stationary-phase cultures. Mol. Biol. Cell22, 988–998. 10.1091/mbc.e10-06-0499
7
De GodoyL. M. F.OlsenJ. V.CoxJ.NielsenM. L.HubnerN. C.FröhlichF.et al. (2008). Comprehensive mass-spectrometry-based proteome quantification of haploid versus diploid yeast. Nature455, 1251–1254. 10.1038/nature07341
8
DenervaudN.BeckerJ.Delgado-GonzaloR.DamayP.RajkumarA. S.UnserM.et al. (2013). A chemostat array enables the spatio-temporal analysis of the yeast proteome. Proc. Natl. Acad. Sci. U.S.A.110, 15842–15847. 10.1073/pnas.1308265110
9
EllisR. J. (2001). Macromolecular crowding: an important but neglected aspect of the intracellular environment. Curr. Opin. Struct. Biol.11, 114–119. 10.1016/S0959-440X(00)00172-X
10
FeigM.HaradaR.MoriT.YuI.TakahashiK.SugitaY. (2015). Complete atomistic model of a bacterial cytoplasm for integrating physics, biochemistry, and systems biology. J. Mol. Graph. Model.58, 1–9. 10.1016/j.jmgm.2015.02.004
11
FeigM.YuI.WangP. H.NawrockiG.SugitaY. (2017). Crowding in cellular environments at an atomistic level from computer simulations. J. Phys. Chem. B121, 8009–8025. 10.1021/acs.jpcb.7b03570
12
GhaemmaghamiS.HuhW. K.BowerK.HowsonR. W.BelleA.DephoureN.et al. (2003). Global analysis of protein expression in yeast. Nature425, 737–741. 10.1038/nature02046
13
HaradaR.SugitaY.FeigM. (2012). Protein crowding affects hydration structure and dynamics. J. Am. Chem. Soc.134, 4842–4849. 10.1021/ja211115q
14
HoB.BaryshnikovaA.BrownG. W. (2018). Unification of protein abundance datasets yields a quantitative Saccharomyces cerevisiae proteome. Cell Syst.6, 192–205.e3. 10.1016/j.cels.2017.12.004
15
JorgensenP.NishikawaJ. L.BreitkreutzB. J.TyersM. (2002). Systematic identification of pathways that couple cell growth and division in yeast. Science297, 395–400. 10.1126/science.1070850
16
KlumppS.ScottM.PedersenS.HwaT. (2013). Molecular crowding limits translation and cell growth. Proc. Natl. Acad. Sci. U.S.A.110, 16754–16759. 10.1073/pnas.1310377110
17
KulakN. A.PichlerG.ParonI.NagarajN.MannM. (2014). Minimal, encapsulated proteomic-sample processing applied to copy-number estimation in eukaryotic cells. Nat. Methods11, 319–324. 10.1038/nmeth.2834
18
LahtveeP. J.SánchezB. J.SmialowskaA.KasvandikS.ElsemmanI. E.GattoF.et al. (2017). Absolute quantification of protein and mRNA abundances demonstrate variability in gene-specific translation efficiency in yeast. Cell Syst.4, 495–504.e5. 10.1016/j.cels.2017.03.003
19
LawlessC.HolmanS. W.BrownridgeP.LanthalerK.HarmanV. M.WatkinsR.et al. (2016). Direct and absolute quantification of over 1800 yeast proteins via selected reaction monitoring. Mol. Cell. Proteomics15, 1309–1322. 10.1074/mcp.M115.054288
20
LeeM. V.TopperS. E.HublerS. L.HoseJ.WengerC. D.CoonJ. J.et al. (2011). A dynamic model of proteome changes reveals new roles for transcript alteration in yeast. Mol. Syst. Biol.7:514. 10.1038/msb.2011.48
21
LeeM. W.KimB. J.ChoiH. K.RyuM. J.KimS. B.KangK. M.et al. (2007). Global protein expression profiling of budding yeast in response to DNA damage. Yeast24, 145–154. 10.1002/yea.1446
22
LuH.ZhuY. F.XiongJ.WangR.JiaZ. (2015). Potential extra-ribosomal functions of ribosomal proteins in Saccharomyces cerevisiae. Microbiol. Res.177, 28–33. 10.1016/j.micres.2015.05.004
23
LuP.VogelC.WangR.YaoX.MarcotteE. M. (2007). Absolute protein expression profiling estimates the relative contributions of transcriptional and translational regulation. Nat. Biotechnol.25, 117–124. 10.1038/nbt1270
24
MazumderA.PesudoL. Q.McReeS.BatheM.SamsonL. D. (2013). Genome-wide single-cell-level screen for protein abundance and localization changes in response to DNA damage in S. cerevisiae. Nucleic Acids Res.41, 9310–9324. 10.1093/nar/gkt715
25
McGuffeeS. R.ElcockA. H. (2010). Diffusion, crowding & protein stability in a dynamic molecular model of the bacterial cytoplasm. PLoS Comput. Biol.6:e1000694. 10.1371/journal.pcbi.1000694
26
NagarajN.Alexander KulakN.CoxJ.NeuhauserN.MayrK.HoerningO.et al. (2012). System-wide perturbation analysis with nearly complete coverage of the yeast proteome by single-shot ultra HPLC runs on a bench top orbitrap. Mol. Cell. Proteomics11:M111.013722. 10.1074/mcp.M111.013722
27
NakanoS. I.MiyoshiD.SugimotoN. (2014). Effects of molecular crowding on the structures, interactions, and functions of nucleic acids. Chem. Rev.114, 2733–2758. 10.1021/cr400113m
28
NawrockiG.WangP. H.YuI.SugitaY.FeigM. (2017). Slow-down in diffusion in crowded protein solutions correlates with transient cluster formation. J. Phys. Chem. B121, 11072–11084. 10.1021/acs.jpcb.7b08785
29
NewmanJ. R. S.GhaemmaghamiS.IhmelsJ.BreslowD. K.NobleM.DeRisiJ. L.et al. (2006). Single-cell proteomic analysis of S. cerevisiae reveals the architecture of biological noise. Nature441, 840–846. 10.1038/nature04785
30
PengM.TaouatasN.CappadonaS.Van BreukelenB.MohammedS.ScholtenA.et al. (2012). Protease bias in absolute protein quantitation. Nat. Methods9, 524–525. 10.1038/nmeth.2031
31
PerryR. P. (2007). Balanced production of ribosomal proteins. Gene401, 1–3. 10.1016/j.gene.2007.07.007
32
PicottiP.Clément-ZizaM.LamH.CampbellD. S.SchmidtA.DeutschE. W.et al. (2013). A complete mass-spectrometric map of the yeast proteome applied to quantitative trait analysis. Nature494, 266–270. 10.1038/nature11835
33
RidgwayD.BroderickG.Lopez-CampistrousA.Ru'AiniM.WinterP.HamiltonM.et al. (2008). Coarse-grained molecular simulation of diffusion and reaction kinetics in a crowded virtual cytoplasm. Biophys. J.94, 3748–3759. 10.1529/biophysj.107.116053
34
SasidharanK.AmarieiC.TomitaM.MurrayD. B. (2012). Rapid DNA, RNA and protein extraction protocols optimized for slow continuously growing yeast cultures. Yeast29, 311–322. 10.1002/yea.2911
35
TanizakiS.CliffordJ.ConnellyB. D.FeigM. (2008). Conformational sampling of peptides in cellular environments. Biophys. J.94, 747–759. 10.1529/biophysj.107.116236
36
ThakurS. S.GeigerT.ChatterjeeB.BandillaP.FröhlichF.CoxJ.et al. (2011). Deep and highly sensitive proteome coverage by LC-MS/MS without prefractionation. Mol. Cell. Proteomics10:M110.003699. 10.1074/mcp.M110.003699
37
TkachJ. M.YimitA.LeeA. Y.RiffleM.CostanzoM.JaschobD.et al. (2012). Dissecting DNA damage response pathways by analysing protein localization and abundance changes during DNA replication stress. Nat. Cell Biol.14, 966–976. 10.1038/ncb2549
38
Van Den BergJ.BoersmaA. J.PoolmanB. (2017). Microorganisms maintain crowding homeostasis. Nat. Rev. Microbiol.15, 309–318. 10.1038/nrmicro.2017.17
39
von der HaarT. (2008). A quantitative estimation of the global translational activity in logarithmically growing yeast cells. BMC Syst. Biol.2:87. 10.1186/1752-0509-2-87
40
WaldronC.LacrouteF. (1975). Effect of growth rate on the amounts of ribosomal and transfer ribonucleic acids in yeast. J. Bacteriol.122, 855–865.
41
WangY.LiC.PielakG. J. (2010). Effects of proteins on protein diffusion. J. Am. Chem. Soc.132, 9392–9397. 10.1021/ja102296k
42
WarnerJ. R. (1999). The economics of ribosome biosynthesis in yeast. Trends Biochem. Sci.24, 437–440. 10.1016/S0968-0004(99)01460-7
43
WebbK. J.XuT.ParkS. K.YatesJ. R. (2013). Modified MuDPIT separation identified 4488 proteins in a system-wide analysis of quiescence in yeast. J. Proteome Res.12, 2177–2184. 10.1021/pr400027m
44
WeiD.JacobsS.ModlaS.ZhangS.YoungC. L.CirinoR.et al. (2012). High-resolution three-dimensional reconstruction of a whole yeast cell using focused-ion beam scanning electron microscopy. Biotechniques53, 41–48. 10.2144/000113850
45
YamaguchiM.NamikiY.OkadaH.MoriY.FurukawaH.WangJ.et al. (2011). Structome of Saccharomyces cerevisiae determined by freeze-substitution and serial ultrathin-sectioning electron microscopy. J. Electron Microsc.60, 321–335. 10.1093/jmicro/dfr052
46
YofeI.WeillU.MeurerM.ChuartzmanS.ZalckvarE.GoldmanO.et al. (2016). One library to make them all: Streamlining the creation of yeast libraries via a SWAp-Tag strategy. Nat. Methods13, 371–378. 10.1038/nmeth.3795
Summary
Keywords
macromolecular crowding, proteomics, protein translation, yeast, molecular dynamics
Citation
Kompella VPS, Stansfield I, Romano MC and Mancera RL (2019) Definition of the Minimal Contents for the Molecular Simulation of the Yeast Cytoplasm. Front. Mol. Biosci. 6:97. doi: 10.3389/fmolb.2019.00097
Received
18 July 2019
Accepted
11 September 2019
Published
02 October 2019
Volume
6 - 2019
Edited by
Valentina Tozzini, Nanosciences Institute, National Research Council, Italy
Reviewed by
Chia-en Chang, University of California, Riverside, United States; Abhigyan Satyam, Harvard Medical School, United States
Updates

Check for updates
Copyright
© 2019 Kompella, Stansfield, Romano and Mancera.
This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.
*Correspondence: Ricardo L. Mancera r.mancera@curtin.edu.au
This article was submitted to Biological Modeling and Simulation, a section of the journal Frontiers in Molecular Biosciences
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.