Skip to main content
PLOS One logoLink to PLOS One
. 2014 Jul 23;9(7):e102580. doi: 10.1371/journal.pone.0102580

Protein Distributions from a Stochastic Model of the lac Operon of E. coli with DNA Looping: Analytical solution and comparison with experiments

Krishna Choudhary 1, Stefan Oehler 1, Atul Narang 1,*
Editor: Mark Isalan2
PMCID: PMC4108355  PMID: 25055040

Abstract

Although noisy gene expression is widely accepted, its mechanisms are subjects of debate, stimulated largely by single-molecule experiments. This work is concerned with one such study, in which Choi et al., 2008, obtained real-time data and distributions of Lac permease in E. coli. They observed small and large protein bursts in strains with and without auxiliary operators. They also estimated the size and frequency of these bursts, but these were based on a stochastic model of a constitutive promoter. Here, we formulate and solve a stochastic model accounting for the existence of auxiliary operators and DNA loops. We find that DNA loop formation is so fast that small bursts are averaged out, making it impossible to extract their size and frequency from the data. In contrast, we can extract not only the size and frequency of the large bursts, but also the fraction of proteins derived from them. Finally, the proteins follow not the negative binomial distribution, but a mixture of two distributions, which reflect the existence of proteins derived from small and large bursts.

Introduction

Data from many independent experiments show that the abundance of any given protein varies among individual cells of isogenic populations growing under identical conditions [1][3]. Early experiments with fluorescent reporters showed that such non-uniformity in protein abundance was due to the inherent stochasticity of gene expression (intrinsic noise) and various forms of cell-to-cell variation (extrinsic noise) [4], [5]. The subsequent development of single-molecule techniques has led to deeper insights into the molecular mechanisms generating the noise [6], [7]. By measuring the number of mRNAs in single cells, Golding et al. showed that transcription was too bursty to be modeled as a Poisson process [8]. Cai et al. [9] and Yu et al. [10] developed two different methods for measuring the number of proteins in single cells. The real-time data of both studies showed that protein synthesis was bursty, and the burst size was exponentially distributed. Under this condition, the steady state protein distribution follows the Gamma distribution, Inline graphic, where Inline graphic and Inline graphic denote the mean burst frequency and burst size [11]. Cai et al. and Yu et al. showed that the Gamma distribution could fit their steady state data, and the values of the mean burst frequency and size derived from the steady state data agreed well with those obtained from real-time measurements.

Armed with these results, Choi et al. [12] attacked a long-standing problem. When non-induced cells of E. coli are exposed to small concentrations of the gratuitous inducer TMG, the lac operon is induced by stochastic switching of individual cells from the non-induced to the induced state [13]. Choi et al. sought the molecular mechanism of this stochastic switching. To this end, they first quantified the minimum number of LacY molecules required to switch a cell to the induced state, and found this threshold to be 375 molecules. They then suggested a molecular mechanism capable of yielding this threshold by appealing to the known mechanisms of repression and transcription of the lac operon. Repression is mediated by the stable DNA loops formed when the Lac repressor is simultaneously bound to the main and auxiliary operators (Fig. 1). Transcription can take place either due to partial dissociations, which occur when a repressor trapped in a DNA loop dissociates from the main operator, but not the auxiliary operator; or complete dissociations, which occur when the repressor dissociates completely from the DNA. Choi et al. hypothesized that since a partially dissociated repressor remains attached to the DNA, it rapidly rebinds to the main operator, thus limiting the number of transcription events. Although the evidence suggests that no more than one mRNA is made during a partial dissociation, it is conceivable that multiple transcripts are made during a partial dissociation despite its short lifetime, thus leading to a small transcriptional burst. In contrast, a completely dissociated repressor takes a relatively long time to find an operator, which results in a large transcriptional burst. These large transcriptional bursts can provide enough proteins to cross the threshold for stochastic switching.

Figure 1. Structure and states of the lac operon.

Figure 1

The repressor Inline graphic can bind to any of the three operators, namely the main operator Inline graphic, and the two auxiliary operators Inline graphic, Inline graphic. The repressor-free state is enclosed by the lower dashed box. The repressor-bound states, enclosed by the upper dashed box, consist of the following 5 states (clockwise from the left): the Inline graphic -bound state Inline graphic, the looped state Inline graphic, the Inline graphic -bound state Inline graphic, the looped state, Inline graphic, and the Inline graphic-bound state Inline graphic. Transcription occurs only if the operon is in the repressor-free state or the repressor-bound state Inline graphic. Small bursts occur whenever the repressor dissociates from the looped state Inline graphic to form the Inline graphic -bound state Inline graphic. Large bursts occur whenever the repressor dissociates from the DNA to form the repressor-free state. Transitions between repressor-free and repressor-bound states occur with propensities Inline graphic and Inline graphic.

Choi et al. tested the foregoing hypotheses as follows. The statistics of small transcriptional bursts were obtained with strain SX701, a Inline graphic strain that exhibits mostly small bursts. To capture the statistics of large bursts, they deleted the auxiliary operators of their Inline graphic cells, thus creating strain SX703 which yields only large bursts. The statistics of the small and large bursts were quantified by measuring the steady-state protein distributions for both strains at various inducer concentrations. They then concluded, based on the model of Friedman et al. [11], that if Inline graphic denote the mean and variance of a protein distribution obtained with strain SX701, then the Fano factor, Inline graphic, and the reciprocal of the noise, Inline graphic, represent the size and frequency of the small bursts. Likewise, if Inline graphic denote the mean and variance for SX703, then Inline graphic represent the size and frequency of the large bursts. Analysis of the data for SX703 with this method showed that Inline graphic did not change with inducer levels, but Inline graphic increased dramatically (Fig. 2a), thus confirming their hypothesis that large bursts can generate enough proteins to trigger stochastic switching. Surprisingly, analysis of the data for SX701 also yielded similar trends (Fig. 2b), but this was attributed to the distortions created by the few cells exhibiting large bursts. Indeed, if the data were filtered by removing the contribution of large bursts, Inline graphic and Inline graphic did not change much with the inducer concentration (Fig. 2c), leading the authors to conclude that the small burst frequency and size were independent of the inducer level.

Figure 2. The variation of the Fano factor and the reciprocal of the noise with the inducer level [12].

Figure 2

(a) Derived from data for strain SX703, which exhibits only large transcriptional bursts, since it lacks both auxiliary operators. Choi et al. proposed that Inline graphic and Inline graphic represent the size and frequency of large transcriptional bursts. (b) Derived from raw data for strain SX701, which exhibits mostly small transcriptional bursts, since it has both auxiliary operators. Choi et al. did not consider this data on the grounds that the occurrence of large bursts in a few cells distorted the statistics of the small transcriptional bursts. (c) Derived from data for strain SX701 that was filtered by rejecting the data corresponding to the few cells exhibiting large bursts. Choi et al. proposed that this Inline graphic and Inline graphic represent the size and frequency of small transcriptional bursts. (d) Mean size of large transcriptional bursts in strain SX701, Inline graphic, (full red curve) and fraction of proteins derived from such bursts, Inline graphic, (full blue curve) estimated from the data in (b). The ordinate of the dashed red line is one-third of the ordinate of the Inline graphic vs. [TMG] line shown in (a), and therefore represents one-third of the (large) transcriptional burst size in strain SX703. The proximity of the full and dashed red lines implies that the mean size of large transcriptional bursts in strain SX701 is approximately one-third of the transcriptional burst size in strain SX703, which is consistent with our model predictions.

Choi et al. also explained these results by appealing to the known states of the lac operon (Fig. 1). However, the mathematical model of Friedman et al., which forms the basis of their data analysis, does not account for these complexities — it only considers a constitutive (unregulated) promoter. Consequently, there is no strong support for the assumption that the proteins follow the Gamma distribution; Inline graphic represent the size of small and large bursts; and Inline graphic represent the frequency of small and large bursts. The goal of this study is to verify the validity of these assumptions by formulating a stochastic model accounting for the known states of the operon, and deriving analytical expressions for the steady state protein distribution, Fano factor, and noise.

There are stochastic models accounting for the details shown in Fig. 1 [14][16], but these studies do not give analytical expressions for the steady state protein distribution. The literature also contains several stochastic models of gene regulation for which analytical solutions were obtained [11], [17][24], but they do not account for the presence of multiple auxiliary operators and DNA looping. Our model fills this gap in the theoretical literature, and its analysis yields deeper insights into the experimental data. Specifically, we show that the size and frequency of small bursts cannot be extracted from the data for strain SX701 because they are averaged out. However, we can extract not only the size and frequency of the large bursts, but also their contribution to total protein synthesis, provided the data is not filtered (Fig. 2d). This result also yields tests for the consistency of the model by providing relationships between the size and frequency of large bursts in strains SX701 and SX703. Finally, we show that neither one of the two strains follow the negative binomial (or Gamma) distribution.

The paper is organized as follows. In the Analysis section, we describe the model, derive the master equation, and explain the key approximations used to obtain the steady state protein distribution. In the Results section, we perform simulations to check the validity of the analytical expression for the protein distribution, and we derive the expressions for mean and the variance of the distribution. We also show that the mean, variance, and hence, the Fano factor and the reciprocal of the noise, can be expressed in terms of the size and frequency of the transcriptional and translational bursts. In the Discussion section, the latter are compared with the assumptions of Choi et al. We also show that negative binomial distributions are obtained only if the size of the large transcriptional bursts is relatively small.

Analysis

The model scheme, shown in Figure 1, is based on the following facts enunciated by Oehler et al. [25], [26]. The lac operon of E. coli contains three operators, namely the main operator Inline graphic, and the two auxiliary operators Inline graphic, lying downstream and upstream of Inline graphic. The lac operon rarely entertains more than one Lac repressor, and this single repressor Inline graphic can bind to any one of the operators, thus forming the operon states, Inline graphic, Inline graphic, and Inline graphic. Since the tetrameric repressor is a "dimer of dimers,'' it has a free dimer even after it is bound to one of the operators. This free dimer can bind to one of the remaining two free operators, thus forming a DNA loop. In principle, three looped states are feasible, namely, Inline graphic, Inline graphic, and Inline graphic, but the last one is very unlikely to form. We are therefore led to consider only six feasible states of the operon — the repressor-free state, and the five repressor-bound states, Inline graphic, Inline graphic, Inline graphic, Inline graphic, and Inline graphic. Only three of these six states permit transcriptional activity, namely, the repressor-free state and the repressor-bound states, Inline graphic and Inline graphic. The first two states permit full transcriptional activity. The last state can be neglected since it permits only 3–5% of the full transcriptional activity.

The model kinetics are based on the following assumptions. All cells have the same number of repressors, Inline graphic, which is tantamount to neglecting extrinsic noise [4]. Since association of a cytosolic repressor to an operator is diffusion-limited, we assume that a cytosolic repressor has the same propensity, Inline graphic, for association with each of the operators. In contrast, the propensity for dissociation of operator-bound repressor does depend on the identity of the operator, and we denote the propensity for dissociation of Inline graphic -bound repressor by Inline graphic. Next, we consider the kinetics of looping. The looped state Inline graphic can be formed from either Inline graphic or Inline graphic, but both pathways have the same propensity because they are driven by the same local concentration effect [26]. Thus, we denote the propensity for formation of Inline graphic from Inline graphic or Inline graphic by the same symbol, Inline graphic. Similarly, we denote the propensity for formation of Inline graphic from Inline graphic or Inline graphic by the same symbol, Inline graphic. Finally, we let Inline graphic denote the propensities for mRNA synthesis and degradation, and Inline graphic denote the propensities for protein synthesis and dilution.

Equations

We take a master equation approach to describe the system, our state variables being the number of mRNAs, Inline graphic, the number of proteins, Inline graphic, and the six states of the operon shown in Figure 1. We let Inline graphic denote the probability of Inline graphic mRNAs and Inline graphic proteins when the operon is in state Inline graphic. Here, Inline graphic when the operon is free, and Inline graphic or Inline graphic when the operon is repressor-bound, where Inline graphic are integers identifying the operator(s) to which the repressor is bound (e.g., Inline graphic denotes the state Inline graphic and Inline graphic denotes the state Inline graphic). Then the master equations for the kinetic scheme in Figure 1 are

graphic file with name pone.0102580.e090.jpg (1)
graphic file with name pone.0102580.e091.jpg (2)
graphic file with name pone.0102580.e092.jpg (3)
graphic file with name pone.0102580.e093.jpg (4)
graphic file with name pone.0102580.e094.jpg (5)
graphic file with name pone.0102580.e095.jpg (6)

Our goal is to derive the steady state protein distribution corresponding to these equations.

Parameter values

Table 1 shows the parameter values in the absence of the inducer. The parameters Inline graphic and Inline graphic reflect the experimental values measured by Yu et al. [10]. The parameter Inline graphic was chosen such that the the mean burst size, Inline graphic, agreed with the measured value Inline graphic, reported by Yu et al. The parameter Inline graphic was estimated by assuming that the mean burst frequency of fully induced cells, Inline graphic, is 600. The rationale for this assumption is as follows. An uninduced cell contains, on average, 0.5 molecules of the tetrameric LacZ [9], and hence, is expected to contain 2 molecules of the monomeric LacY. Since the number of LacY and LacZ molecules increases Inline graphic1200-fold in fully induced cells [25], there are 2400 LacY molecules in such cells, i.e., Inline graphic, which implies that Inline graphic. All other parameter values were estimated using the method of Vilar & Leibler [15]. They estimated all the equilibrium constants using the repression data of Oehler et al. [26]. Then, given an experimental estimate of any one parameter, they could find all other parameter values. They took that one parameter to be the dissociation rate constant, Inline graphic, and assigned to it the value obtained from in vitro data [27]. Based on this procedure, the association rate, Inline graphic, was found to be 0.73 Inline graphic. However, recent in vivo measurement show that the association rate for a dimeric repressor is 0.014 Inline graphic [28]. If the dimeric and tetrameric repressor associate at the same rate, and each cell contains 10 repressors [29], the estimated value of Inline graphic from these measurements is 0.14 Inline graphic. We assumed Inline graphic Inline graphic, and chose Inline graphic, Inline graphic, Inline graphic, Inline graphic, Inline graphic to ensure consistency with the repression data. As we show later, these parameter values yield good fits of the experimental data.

Table 1. Parameter values in the absence of inducer.

Parameter Value (in Inline graphic) Parameter Value (in Inline graphic)
Inline graphic 0.011 Inline graphic 0.0016
Inline graphic 0.0002 Inline graphic 0.019
Inline graphic 0.12 Inline graphic 0.73
Inline graphic 0.044 Inline graphic 4
Inline graphic 0.07 Inline graphic 24

Since we are also concerned with protein distributions in the presence of the inducer, it is necessary to identify the parameters that change under these conditions. We assume that Inline graphic, Inline graphic, Inline graphic, and Inline graphic are independent of the inducer level. The propensities for looping, Inline graphic Inline graphic, are also unlikely to change in the presence of small inducer concentrations because a partially dissociated repressor has too little time to interact with the inducer: In the presence of 10 µM IPTG (considered equivalent to 100 µM TMG), the pseudo-first-order rate constant for repressor-inducer binding is 0.1 Inline graphic [30], which is negligible compared to the looping rate constant of 4 Inline graphic. Thus, the only parameters that can change with the inducer concentration are the association rate, Inline graphic, and the dissociation rates, Inline graphic. Based on the analysis of their experimental protein distributions, Choi et al. concluded that the dissociation rates are independent of the inducer concentration, while the association rate decreases with the inducer concentration. We shall also assume that this is the case. This assumption holds only if the concentration of TMG is significantly below 1 mM [31], [32], a condition satisfied by all the concentrations used by Choi et al., except possibly the highest concentration of 200 µM.

Model reduction

The determination of the steady state protein distribution corresponding to eqs. (1)(6) is facilitated by the fact that loop formation and mRNA degradation are relatively fast.

Rapid loop formation

Table 1 shows that in the absence of the inducer, Inline graphic are much greater than all other propensities, and as explained above, this persists even in the presence of low inducer concentrations. It follows that the repressor-bound states rapidly equilibrate on the fast time scale Inline graphic, after which there are relatively infrequent transitions between the repressor-free and repressor-bound states. To capture this physical fact, we replace eq. (2) with the equation for the slow variable

graphic file with name pone.0102580.e143.jpg (7)

which represents the probability of Inline graphic mRNAs and Inline graphic proteins when the operon is repressor-bound. We then apply the quasi-steady state approximation to the fast variables, Inline graphic, Inline graphic, Inline graphic, Inline graphic, and find that the probabilities of the equilibrated bound states are given by the relations

graphic file with name pone.0102580.e150.jpg (8)
graphic file with name pone.0102580.e151.jpg (9)
graphic file with name pone.0102580.e152.jpg (10)
graphic file with name pone.0102580.e153.jpg (11)
graphic file with name pone.0102580.e154.jpg (12)

which express the physical fact that after the bound states reach quasi-equilibrium, they obey the principle of detailed balance and are almost always in one of the looped states (Table 2). Moreover, the slow variables follow the equations

Table 2. Magnitudes of important derived parameters in the absence of the inducer.
Parameter Value Parameter Value
Inline graphic 0.86 Inline graphic Inline graphic Inline graphic
Inline graphic 0.14 Inline graphic 0.22 Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic 600
Inline graphic Inline graphic Inline graphic 4
graphic file with name pone.0102580.e172.jpg (13)
graphic file with name pone.0102580.e173.jpg (14)

where

graphic file with name pone.0102580.e174.jpg (15)
graphic file with name pone.0102580.e175.jpg (16)
graphic file with name pone.0102580.e176.jpg (17)

Equations (13)–(14) describe the evolution of the reduced model containing only two operon states — the free and the equilibrated bound states — between which are transitions with propensities, Inline graphic, which are slow compared to the propensities for looping (Table 2). This is highlighted in Figure 1 by enclosing the free and bound states in dashed boxes, and drawing dashed arrows with labels, Inline graphic and Inline graphic, to denote the transitions between them. The reduced model is similar to Shahrezaei & Swain's three-stage model for a regulated promoter [22], but there is an important difference. Both operon states are transcriptionally active: The transcription rates in the free and bound states are Inline graphic and Inline graphic, respectively, where Inline graphic is the probability of the Inline graphic state. Even though Inline graphic (Table 2), we cannot neglect the transcription from the bound state, since it captures the effect of the small transcriptional bursts, which can account, as we show later, for almost 80% of the mRNAs synthesized per cell cycle.

Table 2 shows that in the absence of the inducer, Inline graphic, so that the free state occurs infrequently and lasts for very short periods of time, i.e., Inline graphic. We shall show later that this persists in the presence of the low inducer concentrations (Inline graphic 200 µM TMG) used by Choi et al. Hence, under the experimental conditions of interest, the conditional probabilities in (8)–(12) are essentially equal to the absolute probabilities.

Rapid mRNA degradation

The second approximation appeals to the fact that mRNA degradation is rapid compared to protein dilution, i.e., Inline graphic. To apply this approximation, we follow Shahrezaei & Swain [22]. Thus, we begin by rescaling time with respect to the time scale for protein degradation. Letting Inline graphic transforms the reduced equations to the form

graphic file with name pone.0102580.e190.jpg (18)
graphic file with name pone.0102580.e191.jpg (19)

where Inline graphic and Inline graphic are the frequencies of transitions between the free and bound operator states, Inline graphic is the frequency of unregulated transcription (in the absence of the repressor), Inline graphic is the translational burst size, i.e., the average number of proteins produced per mRNA, and Inline graphic is the ratio of protein and mRNA lifetimes. Next, we define the generating functions, Inline graphic and Inline graphic, to obtain the partial differential equations

graphic file with name pone.0102580.e199.jpg (20)
graphic file with name pone.0102580.e200.jpg (21)

where Inline graphic and Inline graphic. Since Inline graphic, we have the quasi-steady state approximation, Inline graphic. The steady state protein distribution is therefore given by the equations

graphic file with name pone.0102580.e205.jpg (22)
graphic file with name pone.0102580.e206.jpg (23)

Since we are interested in the generating function, Inline graphic, it is convenient to rewrite these equations as

graphic file with name pone.0102580.e208.jpg (24)
graphic file with name pone.0102580.e209.jpg (25)

which reduce to the second-order differential equation

graphic file with name pone.0102580.e210.jpg (26)

We solve this equation with the initial condition, Inline graphic, and revert to Inline graphic as the independent variable, to obtain the following generating function for the steady state protein distribution

graphic file with name pone.0102580.e213.jpg (27)

where Inline graphic denotes the Gaussian hypergeometric function and

graphic file with name pone.0102580.e215.jpg (28)

As expected, if Inline graphic, (27) reduces to the generating function of the negative hypergeometric distribution [22]. In general, however, (27) is the generating function for a mixture of the negative binomial and negative hypergeometric distributions, which reflects, as we show below, the existence of two sub-populations of proteins, namely those derived from small and large transcriptional bursts.

Results

Analytical expressions for the statistics of the protein distributions

Strain with auxiliary operators

The generating function (27) yields the following expressions for the mean, Inline graphic, and variance, Inline graphic, of the protein distribution

graphic file with name pone.0102580.e219.jpg (29)
graphic file with name pone.0102580.e220.jpg (30)

Since Inline graphic represents the mean number of proteins synthesized per mRNA, (29) implies that Inline graphic is the mean frequency of regulated transcription. The two terms of Inline graphic also have simple physical interpretations: Since Inline graphic and Inline graphic are the probabilities of the Inline graphic and free states, Inline graphic and Inline graphic represent the mean number of mRNAs produced per cell cycle due to small and large transcriptional bursts.

Expanding Inline graphic about Inline graphic yields the steady state protein distribution

graphic file with name pone.0102580.e231.jpg (31)

Figure 3 shows that the protein distributions obtained from this expression agree well with those obtained by simulating the full model with the Optimized Direct Method implementation of Gillespie's Stochastic Simulation Algorithm [33] provided in the simulation package StochKit2 [34]. The protein distribution in the absence of the inducer, shown in Fig. 3a, was obtained with the parameter values in Table 1. The distributions in the presence of the inducer were obtained by decreasing the association rate, Inline graphic, 10-fold (Fig. 3b) and 20-fold (Fig. 3c). Evidently, (31) is a good approximation to the exact solutions in all three cases. We conclude that our approximate solution is accurate down to a 20-fold reduction of the association rate.

Figure 3. Despite a 20-fold change in the repressor association rate, Inline graphic, the protein distributions derived from the analytical expression (31) (grey squares) are in good agreement with those obtained from stochastic simulations of the model (black disks).

Figure 3

(a) Parameter values in Table 1. (b) Inline graphic is 1/10th of the value in Table 1; other parameter values as in Table 1. (c) Inline graphic is 1/20th of the value in Table 1; other parameter values as in Table 1.

Table 2 shows that in the absence of the inducer, Inline graphic. These relations remain valid at the relatively low inducer levels studied by Choi et al. (Inline graphic200 µM TMG). Indeed, under these conditions, the operon is expressed to no more than 1% of the fully induced level [12], i.e.,

graphic file with name pone.0102580.e238.jpg (32)

and (29)–(30) can be rewritten as

graphic file with name pone.0102580.e239.jpg (33)
graphic file with name pone.0102580.e240.jpg (34)

It is worth noting that due to rapid loop formation, small transcriptional bursts are very bursty (pulsatile). Moreover, under the weakly inducing conditions used in the experiments (Inline graphic 200 µM TMG), Inline graphic is relatively large, and hence, the large transcriptional bursts are also quite bursty. It follows that under these conditions, (33)–(34) should be expressible in terms of the size and frequency of the small and large transcriptional bursts. We shall show below that this is indeed the case.

Strain without auxiliary operators

In the absence of auxiliary operators, the operon fluctuates between the free and the Inline graphic-bound state, and only the former allows transcription. This is identical to Shahrezaei & Swain's 3-stage model of a regulated promoter [22], and corresponds to the special case, Inline graphic, Inline graphic, Inline graphic of our model. It follows that the generating function for the steady state protein distribution is the Gaussian hypergeometric function

graphic file with name pone.0102580.e247.jpg (35)

where

graphic file with name pone.0102580.e248.jpg (36)

and

graphic file with name pone.0102580.e249.jpg (37)

Moreover, the protein distribution is given by the expression

graphic file with name pone.0102580.e250.jpg (38)

and the mean and variance are

graphic file with name pone.0102580.e251.jpg (39)
graphic file with name pone.0102580.e252.jpg (40)

At TMG concentrations of Inline graphic 100 µM, which are equivalent to an IPTG concentration of Inline graphic 10 µM, the operon is expressed to no more than 5% of the fully induced level [35]. It follows that under the experimental conditions of interest

graphic file with name pone.0102580.e255.jpg (41)

and Inline graphic can be approximated by the expressions

graphic file with name pone.0102580.e257.jpg (42)
graphic file with name pone.0102580.e258.jpg (43)

Expressing the statistics in terms of the burst size and frequency

Choi et al. assumed that the quantities Inline graphic and Inline graphic represent the size and frequency of small transcriptional bursts, and Inline graphic and Inline graphic represent the size and frequency of large transcriptional bursts. To check the validity of these assumptions, we shall express (33)–(34) and (42)–(43) in terms of the size and frequency of the transcriptional bursts. Given these expressions, we can immediately infer the dependence of Inline graphic on the size and frequency of the transcriptional bursts, and then compare them to the assumptions made by Choi et al.

Strain with auxiliary operators

To express Inline graphic in terms of the size and frequency of the transcriptional bursts, we begin by recalling that Inline graphic consists of two terms, Inline graphic and Inline graphic, which represent the mean frequency of transcription due to partial and complete dissociations of the repressor, respectively. Since partial dissociations occur when a repressor trapped in the Inline graphic -loop dissociates from Inline graphic, we define the number of the partial dissociations per cell cycle as

graphic file with name pone.0102580.e270.jpg (44)

where we have appealed to the detailed balance between the operon states Inline graphic and Inline graphic. We also define the number of mRNAs synthesized per partial dissociation as

graphic file with name pone.0102580.e273.jpg (45)

since the time for rebinding of a partially dissociated repressor to Inline graphic is on the order of Inline graphic Inline graphic. It follows from these definitions that

graphic file with name pone.0102580.e277.jpg (46)

i.e., we have successfully expressed the first term of Inline graphic in terms of frequency and mRNA burst size due to partial dissociations. We now proceed to express the second term of Inline graphic in terms of the frequency and mRNA burst size due to complete dissociations. Since complete dissociations occur whenever the operon becomes repressor-free, it is natural to define the number of complete dissociations per cell cycle as

graphic file with name pone.0102580.e280.jpg (47)

We also define the number of mRNAs synthesized per complete dissociation as

graphic file with name pone.0102580.e281.jpg (48)

because the time for rebinding of a completely dissociated repressor to an operator is on the order of Inline graphic. Evidently

graphic file with name pone.0102580.e283.jpg (49)

and we conclude that

graphic file with name pone.0102580.e284.jpg (50)

Hence, (33)–(34) can be rewritten as

graphic file with name pone.0102580.e285.jpg (51)
graphic file with name pone.0102580.e286.jpg (52)

which imply that

graphic file with name pone.0102580.e287.jpg (53)
graphic file with name pone.0102580.e288.jpg (54)

where

graphic file with name pone.0102580.e289.jpg (55)

is the fraction of proteins derived from complete dissociations. It follows from (53) that the total burstiness, Inline graphic, is entirely due to translational and large transcriptional bursts. Moreover, the burstiness of large transcriptional bursts depends on their intrinsic burstiness, Inline graphic, suitably weighted by Inline graphic, the fraction of proteins derived from such bursts. Importantly, Inline graphic is completely determined by Inline graphic, the equilibrium constant for dissociation of the repressor from Inline graphic. In the absence of the inducer, this equilibrium constant is 0.25 [25], [26], and hence, Inline graphic, i.e., 20% of the proteins are derived from large transcriptional bursts. As the inducer concentration increases, Inline graphic increases because Inline graphic decreases.

Strain without auxiliary operators

In this case, if we define the number of complete dissociations per cell cycle as

graphic file with name pone.0102580.e299.jpg (56)

and the number of mRNAs synthesized per complete dissociation as

graphic file with name pone.0102580.e300.jpg (57)

the mean frequency of regulated transcription can be rewritten as

graphic file with name pone.0102580.e301.jpg (58)

It follows that (42)–(43) can be rewritten as

graphic file with name pone.0102580.e302.jpg (59)
graphic file with name pone.0102580.e303.jpg (60)

which imply that

graphic file with name pone.0102580.e304.jpg (61)
graphic file with name pone.0102580.e305.jpg (62)

We are now ready to address questions concerning the physical meaning of the parameters of the distribution and their variation with inducer concentration [12].

Discussion

Interpretation of the protein distribution data

Strain with auxiliary operators. Interpretation of Inline graphic and Inline graphic derived from filtered data

Choi et al. assumed that Inline graphic and Inline graphic derived from the filtered data (Fig. 2c) represent the size and frequency of small transcriptional bursts. In terms of our model, these assumptions have the form

graphic file with name pone.0102580.e310.jpg (63)
graphic file with name pone.0102580.e311.jpg (64)

However, (53)–(54) imply that this Inline graphic and Inline graphic, obtained by eliminating the contribution of the large transcriptional bursts, have a different physical meaning. Indeed, (53) implies that the Fano factor obtained from the filtered data has the form, Inline graphic, which represents the size of the translational, rather than small transcriptional, bursts. Similarly, (54) implies that the reciprocal of the noise derived from the filtered data has the form, Inline graphic, which is proportional to Inline graphic, the average number of mRNAs derived from small bursts, rather than the frequency of the small bursts. Since Inline graphic (Fig. 2c) and Inline graphic, our interpretation of the filtered data implies that Inline graphic, which is close to the estimate obtained from the model (Table 3).

Table 3. Burst frequency and size in uninduced cells with and without auxiliary operators.
Strain With auxiliary operators Without auxiliary operators
Bursts due to Partial dissociations Complete dissociations Complete dissociations
Burst properties Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Model-based value Inline graphic 6.9 0.03 0.20 0.1 0.55 0.05 8 1.65 13
Data-based value Inline graphic 1.25 0.2 3

Inline graphic Calculated from eqs. (44)(45), (47)(48), and (56)(57) with the parameter values in Table 1.

Inline graphic Determined from the data in Figs. 2 a–c by appealing to eqs. (53)(54) and (61)(62).

Evidently, there is a discrepancy between the assumptions of Choi et al. and the implications of our model. To understand its origin, observe that their assumptions are equivalent to the relations

graphic file with name pone.0102580.e333.jpg (65)
graphic file with name pone.0102580.e334.jpg (66)

i.e., they assumed, in effect, that both the mean and the variance are dominated by contributions from small transcriptional bursts. In contrast, (51)–(52) show that small bursts contribute to the mean, but not to the variance. This difference arises because we assumed that looping is so fast that the rapid fluctuations due to partial dissociations are averaged out on the slow time scale of the other processes. This averaging process preserves the contribution of small transcriptional bursts to the mean, but eliminates their contribution to the variance.

The assumption Inline graphic appears to be implausible. Indeed, (53) implies that translational bursts contribute the term Inline graphic to the Fano factor. For the small bursts to make a significant, let alone dominant, contribution to the Fano factor, it is clear that Inline graphic, i.e., on average, approximately one mRNA must be synthesized per partial dissociation. However, looping is so fast compared to transcription that Inline graphic in the absence of the inducer (Table 3). Moreover, Inline graphic is unlikely to change even in the presence of the inducer since Inline graphic and Inline graphic are constant over the range of inducer concentrations used in the experiments. We conclude that the bursts due to partial dissociations are so small that they cannot be the dominant source of burstiness.

Interpretation of Inline graphic and Inline graphic derived from raw data

Choi et al. rejected the raw data shown in Fig. 2b since the occurrence of large bursts in a few cells distorted the statistics of the small bursts. We show below that these data are a valuable source of information about the statistics of large bursts. Specifically, (53)–(54) predict the observed variation of Inline graphic and Inline graphic derived from the raw data, and thus provide a method for estimating not only the size and frequency of the large transcriptional bursts, but also the fraction of proteins derived from them. This method is particularly useful because, as we show below, there are simple relationships between the size and frequency of the large bursts in strains SX701 and SX703, but they are not identical.

The analysis of the raw data shows that the total burstiness, Inline graphic, increases with inducer concentration (Fig. 2b). Eq. (53) implies that this is due to the growing burstiness of the large transcriptional bursts: Since both Inline graphic and Inline graphic increase with inducer level, so does Inline graphic. This increase occurs so rapidly that at 100 µM TMG, large trancriptional bursts become the dominant source of burstiness, i.e, Inline graphic. Indeed, assuming Inline graphic, (53) implies Inline graphic whenever Inline graphic. Inspection of Fig. 2b shows that at 100 µM TMG, Inline graphic, and hence, Inline graphic. We shall show below that at such inducer levels, Inline graphic and Inline graphic.

In contrast to the total burstiness, Inline graphic, the reciprocal of the total noise, Inline graphic, decreases with inducer concentration until it reaches a constant value (Fig. 2b). The model suggests that this is because both Inline graphic and Inline graphic increase with inducer level, but Inline graphic increases faster than Inline graphic: Indeed, both Inline graphic and Inline graphic increase with inducer level, and Eq. (54) shows that Inline graphic is proportional to the ratio Inline graphic, whereas Inline graphic increases with the product Inline graphic. The decreasing trend of Inline graphic continues until the inducer levels become so high that large bursts account for all the proteins (Inline graphic) and burstiness (Inline graphic). Under these conditions Inline graphic approaches Inline graphic, the frequency of large bursts, which is independent of inducer concentration. Comparison with the data in Fig. 2b then implies that Inline graphic.

Given Inline graphic and Inline graphic, (53)–(54) provide a method for estimating the variation of Inline graphic and Inline graphic with inducer levels from the raw data for SX701. To see this, it is convenient to rewrite (53)–(54) in the form

graphic file with name pone.0102580.e380.jpg (67)
graphic file with name pone.0102580.e381.jpg (68)

Since the variation of Inline graphic and Inline graphic with the inducer concentration is known (Fig. 2b), we can solve the above equations to obtain Inline graphic and Inline graphic as a function of the inducer concentration. These calculated profiles, shown in Fig. 2d, agree with the claims above: Both Inline graphic and Inline graphic increase with the inducer level, and the latter approaches 1 at 100 µM TMG.

Strain without auxiliary operators. Interpretation of Inline graphic and Inline graphic

Choi et al. assumed that the Inline graphic and Inline graphic shown in Fig. 2a represent the size and frequency of large transcriptional bursts, i.e.,

graphic file with name pone.0102580.e392.jpg (69)
graphic file with name pone.0102580.e393.jpg (70)

Our model implies that these relations are valid at all non-zero inducer concentrations used in the experiments. Indeed, since Inline graphic, (61)–(62) imply that the above relations are valid whenever Inline graphic, which is satisfied (Inline graphic) at all the non-zero inducer concentrations used in the experiments (Fig. 2a). In particular, comparison with the data in Fig. 2a implies that Inline graphic.

Relationships between the statistics of large bursts in the strains with and without auxiliary operators

The model predicts simple relationships between the size and frequency of the large transcriptional bursts in strains SX701 and SX703, which provide tests for checking the consistency of the model. Indeed, it follows from (48) and (57) that Inline graphic, a relationship that is also mirrored by the data (compare full and dashed lines in Fig. 2d). Similarly, (47) and (56) imply that

graphic file with name pone.0102580.e399.jpg (71)

a ratio estimated to be 1/80 based on the values in Table 1, which is of the same order of magnitude as the value 1/15, obtained from the experimentally determined values of Inline graphic and Inline graphic.

Condition for the negative binomial distribution

Choi et al. assumed that the protein distributions of both strains follow the Gamma distribution, the continuous analog of the negative binomial distribution. We have shown above that neither one of the strains follows the negative binomial distribution. Here, we demonstrate that the distributions can reduce to the negative binomial distribution, but only only if the large burst size is negligibly small, i.e., the association rate Inline graphic, is much larger than the transcription rate Inline graphic. Under this condition, even the large bursts are averaged out, and they contribute to the mean, but not the variance or the burstiness.

We begin by considering the strain without auxiliary operators. Under the weakly induced conditions used in the experiments, Inline graphic, and the generating function for the protein distribution is the negative hypergeometric function

graphic file with name pone.0102580.e405.jpg (72)

which reduces to the generating function for the negative binomial distribution precisely when Inline graphic or Inline graphic. Now (37) implies that

graphic file with name pone.0102580.e408.jpg (73)
graphic file with name pone.0102580.e409.jpg (74)

The condition Inline graphic can never be satisfied since Inline graphic. However, Inline graphic precisely when Inline graphic, in which case Inline graphic and

graphic file with name pone.0102580.e415.jpg (75)

which is the generating function for the negative binomial distribution

graphic file with name pone.0102580.e416.jpg (76)

It is worth noting that under this condition

graphic file with name pone.0102580.e417.jpg (77)

i.e., large transcriptional bursts make no contribution to the burstiness.

A similar argument shows that the generating function for the strain with auxiliary operators reduces to

graphic file with name pone.0102580.e418.jpg (78)

precisely when Inline graphic. Under this condition, the proteins follow the negative binomial distribution

graphic file with name pone.0102580.e420.jpg (79)

and

graphic file with name pone.0102580.e421.jpg (80)

i.e., even the large transcriptional bursts do not contribute to the burstiness.

We have shown above that the proteins follow the negative binomial distribution only if the large bursts are, in fact, rather small, and hence, do not contribute to the burstiness. But it follows from the data in Figs. 2a,b that these bursts do contribute significantly to the burstiness of strains SX701 and SX703 — if this was not true, (77) and (80) imply that the burstiness would be independent of inducer concentration, which contradicts the data. The negative binomial distribution is therefore unlikely to provide good fits to the raw data for both strains, but will fit the filtered data well, since the contribution of large bursts has been eliminated from it. The fits in Choi et al. are consistent with this conclusion. The Gamma distribution fits the filtered data for strain SX701 rather well. However, this is less so for the protein distributions obtained with strain SX703, which exhibits only large bursts. Figure 4 shows that better fits are obtained with the negative hypergeometric distribution (38).

Figure 4. Protein distribution data for strain SX703 (full circles) at various TMG concentrations fitted with the Gamma distribution by Choi et al. (dashed curve) and the negative hypergeometric distribution (full curve).

Figure 4

The negative hypergeometric distribution was fitted with the parameter values in Table 1, except Inline graphic, which was decreased with increasing inducer concentration. (a) Data obtained at 50 µM TMG fitted with Inline graphic Inline graphic. (b) Data obtained at 100 µM TMG fitted with Inline graphic Inline graphic. (c) Data obtained at 200 µM TMG fitted with Inline graphic Inline graphic.

Conclusions

We formulated and solved a stochastic model of lac expression accounting for auxiliary operators and DNA looping. Based on a comparison of our expressions for the Fano factor, noise, and protein distribution of strains SX701 (with auxiliary operators) and SX703 (without auxiliary operators) with those proposed by Choi et al., we arrive at the following conclusions:

  1. The physical interpretations of the Fano factor Inline graphic and reciprocal noise Inline graphic for strain SX703 are identical to those proposed by Choi et al., namely Inline graphic and Inline graphic represent the size and frequency of (large) transcriptional bursts.

  2. The physical interpretations of the Fano factor Inline graphic and reciprocal noise Inline graphic derived from the filtered data for SX701 differ from those given by Choi et al., namely Inline graphic and Inline graphic represent the size and frequency of small transcriptional bursts. Instead, we find that Inline graphic represents the size of translational bursts, and Inline graphic is proportional to the mean number of mRNAs derived from small transcriptional bursts. Our interpretation is different because we assume that looping is so fast that fluctuations due to small transcriptional bursts are averaged out — small bursts therefore contribute to the mean, but not the burstiness, of the protein distribution. This has two consequences:

    1. The information lost due to the averaging implies that the small burst size and frequency cannot be separately extracted from the data. At best, we can only determine the product of the small burst size and frequency, which represents the mean number of mRNAs derived from small bursts.

    2. The burstiness is entirely due to translational and large transcriptional bursts. In particular, the burst size derived from the filtered data for strain SX701, from which the contribution of the large bursts has been deliberately eliminated, yields the size of translational, rather than small transcriptional, bursts.

  3. Choi et al. did not consider the raw data for SX701 because large bursts, although rare, contributed significantly to protein synthesis. This is consistent with our model: Even in uninduced cells, 20% of the proteins are derived from large bursts. We find that the raw data contains valuable information about the statistics of large bursts. By analyzing this data with our model, we isolate not only the size and frequency of large bursts, but also the fraction of proteins derived from them. The large burst size obtained in this manner is consistent with another prediction of the model, namely, it is one-third of the (large) burst size in strain SX703. The model also predicts that the fraction of proteins derived from large bursts is completely determined by a measurable quantity, namely the dissociation constant for binding of the repressor to the auxiliary operator Inline graphic.

  4. The protein distributions for both strains are not negative binomial: SX703 follows a negative hypergeometric distribution, and SX701 follows a mixture of the negative binomial and negative hypergeometric distributions that reflects the existence of two sub-populations of proteins, namely, those derived from small and large bursts. Negative binomial distributions are attained only if large bursts are insignificant, a condition that holds only if the data are filtered by eliminating the contribution of such bursts.

These results imply that interpretation of the steady state protein distributions depends crucially on the details of the regulatory mechanisms.

Acknowledments

We are grateful to Sayantari Ghosh for critical comments and help with the fits of the experimental data.

Data Availability

The authors confirm that all data underlying the findings are fully available without restriction. Data are from the Choi et al, Science, 322, 442, 2008 study whose authors may be contacted at xie@chemistry.harvard.edu

Funding Statement

AN received grant number SR/SO/BB-79/2010 funded by Department of Science & Technology (www.dst.gov.in). The funders had no role in study design, data collection and analysis, decision to publish, or preparation of the manuscript.

References

  • 1. Balzsi G, van Oudenaarden A, Collins JJ (2011) Cellular decision making and biological noise: from microbes to mammals. Cell 144: 910–925. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 2. Li GW, Xie XS (2011) Central dogma at the single-molecule level in living cells. Nature 475: 308–315. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 3. Snijder B, Pelkmans L (2011) Origins of regulated cell-to-cell variability. Nat Rev Mol Cell Biol 12: 119–125. [DOI] [PubMed] [Google Scholar]
  • 4. Elowitz MB, Levine AJ, Siggia ED, Swain PS (2002) Stochastic gene expression in a single cell. Science 297: 1183–1186. [DOI] [PubMed] [Google Scholar]
  • 5. Ozbudak EM, Thattai M, Kurtser I, Grossman AD, van Oudenaarden A (2002) Regulation of noise in the expression of a single gene. Nat Genet 31: 69–73. [DOI] [PubMed] [Google Scholar]
  • 6. Raj A, van Oudenaarden A (2009) Single-molecule approaches to stochastic gene expression. Annu Rev Biophys 38: 255–270. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 7. Xie XS, Choi PJ, Li GW, Lee NK, Lia G (2008) Single-molecule approach to molecular biology in living bacterial cells. Annu Rev Biophys 37: 417–44. [DOI] [PubMed] [Google Scholar]
  • 8. Golding I, Paulsson J, Zawilski SM, Cox EC (2005) Real-time kinetics of gene activity in individual bacteria. Cell 123: 1025–36. [DOI] [PubMed] [Google Scholar]
  • 9. Cai L, Friedman N, Xie XS (2006) Stochastic protein expression in individual cells at the single molecule level. Nature 440: 358–62. [DOI] [PubMed] [Google Scholar]
  • 10. Yu J, Xiao J, Ren X, Lao K, Xie XS (2006) Probing gene expression in live cells, one protein molecule at a time. Science (New York, NY) 311: 1600–3. [DOI] [PubMed] [Google Scholar]
  • 11. Friedman N, Cai L, Xie X (2006) Linking stochastic dynamics to population distribution: An analytical framework of gene expression. Phys Rev Lett 97: 168302. [DOI] [PubMed] [Google Scholar]
  • 12. Choi PJ, Cai L, Frieda K, Xie XS (2008) A stochastic single-molecule event triggers phenotype switching of a bacterial cell. Science (New York, NY) 322: 442–6. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 13. Novick A, Weiner M (1957) Enzyme induction as an all-or-none phenomenon. Proc Nat Acad Sci USA 43: 553–566. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 14. Earnest TM, Roberts E, Assaf M, Dahmen K, Luthey-Schulten Z (2013) DNA looping increases the range of bistability in a stochastic model of the lac genetic switch. Phys Biol 10: 026002. [DOI] [PubMed] [Google Scholar]
  • 15. Vilar JMG, Leibler S (2003) DNA looping and physical constraints on transcription regulation. J Mol Biol 331: 981–989. [DOI] [PubMed] [Google Scholar]
  • 16. Stamatakis M, Mantzaris NV (2009) Comparison of deterministic and stochastic models of the lac operon genetic network. Biophys J 96: 887–906. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 17. Berg OG (1978) A model for the statistical fluctuations of protein numbers in a microbial population. J Theor Biol 71: 587–603. [DOI] [PubMed] [Google Scholar]
  • 18. Kepler TB, Elston TC (2001) Stochasticity in transcriptional regulation: origins, consequences, and mathematical representations. Biophys J 81: 3116–36. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 19. Peccoud J, Ycart B (1995) Markovian modeling of gene product synthesis. Theor Popul Biol 48: 222–234. [Google Scholar]
  • 20. Rigney DR (1979) Stochastic model of constitutive protein levels in growing and dividing bacterial cells. J Theor Biol 76: 453–80. [DOI] [PubMed] [Google Scholar]
  • 21. Raj A, Peskin CS, Tranchina D, Vargas DY, Tyagi S (2006) Stochastic mRNA synthesis in mammalian cells. PLoS Biol 4: e309. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 22. Shahrezaei V, Swain PS (2008) Analytical distributions for stochastic gene expression. Proc Natl Acad Sci U S A 105: 17256–61. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 23. Swain PS, Elowitz MB, Siggia ED (2002) Intrinsic and extrinsic contributions to stochasticity in gene expression. Proc Natl Acad Sci U S A 99: 12795–800. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 24. Thattai M, van Oudenaarden a (2001) Intrinsic noise in gene regulatory networks. Proc Natl Acad Sci U S A 98: 8614–9. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 25. Oehler S, Eismann ER, Krmer H, Mller-Hill B (1990) The three operators of the lac operon cooperate in repression. EMBO J 9: 973–979. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 26. Oehler S, Amouyal M, Kolkhof P, von Wilcken-Bergmann B, Mller-Hill B (1994) Quality and position of the three lac operators of E. coli define efficiency of repression. EMBO J 13: 3348–3355. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 27. Goeddel DV, Yansura DG, Caruthers MH (1978) How lac repressor recognizes lac operator. Proc Natl Acad Sci U S A 75: 3578–82. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 28. Hammar P, Leroy P, Mahmutovic A, Marklund EG, Berg OG, et al. (2012) The Lac repressor displays facilitated diffusion in living cells. Science (New York, NY) 336: 1595–8. [DOI] [PubMed] [Google Scholar]
  • 29. Gilbert W, Mller-Hill B (1966) Isolation of the lac repressor. Proc Natl Acad Sci U S A 56: 1891–1898. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 30. Dunaway M, Manly SP, Matthews KS (1980) Model for lactose repressor protein and its interaction with ligands. Proc Natl Acad Sci U S A 77: 7181–7185. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 31. Barkley MD, Riggs AD, Jobe A, Bourgeois S (1975) Interaction of effecting ligands with lac repressor and repressor-operator complex. Biochemistry 14: 1700–1712. [DOI] [PubMed] [Google Scholar]
  • 32. Dunaway M, Olson JS, Rosenberg JM, Kallai OB, Dickerson RE, et al. (1980) Kinetic studies of inducer binding to lac repressor operator complex. J Biol Chem 255: 10115–10119. [PubMed] [Google Scholar]
  • 33. Cao Y, Li H, Petzold L (2004) Efficient formulation of the stochastic simulation algorithm for chemically reacting systems. J Chem Phys 121: 4059–4067. [DOI] [PubMed] [Google Scholar]
  • 34. Sanft KR, Wu S, Roh M, Fu J, Lim RK, et al. (2011) StochKit2: software for discrete stochastic simulation of biochemical systems with events. Bioinformatics (Oxford, England) 27: 2457–8. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 35. Oehler S, Alberti S, Mller-Hill B (2006) Induction of the lac promoter in the absence of DNA loops and the stoichiometry of induction. Nucleic Acids Res 34: 606–612. [DOI] [PMC free article] [PubMed] [Google Scholar]

Associated Data

This section collects any data citations, data availability statements, or supplementary materials included in this article.

Data Availability Statement

The authors confirm that all data underlying the findings are fully available without restriction. Data are from the Choi et al, Science, 322, 442, 2008 study whose authors may be contacted at xie@chemistry.harvard.edu


Articles from PLoS ONE are provided here courtesy of PLOS

RESOURCES