Abstract
Viruses have an extraordinary ability to diversify and evolve. For segmented viruses, reassortment can introduce drastic genomic and phenotypic changes by allowing a direct exchange of genetic material between coinfecting strains. For instance, multiple influenza pandemics were caused by reassortments of viruses typically found in separate hosts. What is unclear, however, are the underlying mechanisms driving these events and the level of intrinsic bias in the diversity of strains that emerge from coinfection. To address this problem, previous experiments looked for correlations between segments of strains that coinfect cells in vitro. Here, we present an information theory approach as the natural mathematical framework for this question. We study, for influenza and other segmented viruses, the extent to which a virus’s segments can communicate strain information across an infection and among one another. Our approach goes beyond previous association studies and quantifies how much the diversity of emerging strains is altered by patterns in reassortment, whether biases are consistent across multiple strains and cell types, and if significant information is shared among more than two segments. We apply our approach to a new experiment that examines reassortment patterns between the 2009 H1N1 pandemic and seasonal H1N1 strains, contextualizing its segmental information sharing by comparison with previously reported strain reassortments. We find evolutionary patterns across classes of experiments and previously unobserved higher-level structures. Finally, we show how this approach can be combined with virulence potentials to assess pandemic threats.
Keywords: viral evolution, systems biology, emerging infectious disease
Reassortment of segmented viruses is a key mechanism for rapid novel virus creation. At least two human influenza pandemics in the last century were linked to lineages where some number of genomic segments reassorted with a genome of nonhuman origin (1, 2). This fact was reinforced by the emergence of the 2009 H1N1 pandemic (2009 pdm) virus (3–5). Novel reassortant strains can evade adaptive immunity by introducing antigens to a naïve host population or overly stimulate innate immunity by presenting a new host with abundant nonself molecular signals (6–10). Moreover, both sequence database studies and in vitro experiments have shown that genome reassortment between strains happens nonrandomly: If two strains coinfect the same cell, their progeny may not sample all possible strain/segment combinations uniformly (11–14). These analyses focused on whether it is more likely that pairs of segments from the same strain appear together in reassortments, typically using chi-square tests to establish significance.
Because influenza has eight segments, there are 256 possible reassortant viruses when a cell is coinfected by two strains. Each strain type and host cellular environment can influence reassortment differently, so it would seem impossible to predict whether a new pandemic strain can form. However, not every possible progeny combination may occur or survive. As we show, this problem can be reformulated using information theory, determining the information content of a segment’s strain of origin distribution and the information shared among segments. Our approach uncovers new aspects of the process of reassortment, demonstrating significant combinations that occur commonly in several different crosses of viral strains. We show that virus strain and host cell-type influence outcomes. We include a novel experiment with the latest pandemic strain, 2009 pdm, and the seasonal H1N1 strain circulating prior to its introduction, expanding the number of reassortment examples analyzed to date and comparing this case to other analogous experiments within our framework.
Information theory is a general mathematical framework for quantifying the transmission and exchange of information (15, 16). Within this framework, we can separate multiple levels of information transfer and exchange within viral segment replication and reassortment. At the same time, this formalism allows us to show how these different levels constrain one another and to relate information theoretic quantities, such as entropy and mutual information, to the likely diversity of viral populations produced by a host coinfection. We further provide a nonparametric permutation test to assess the statistical significance of these quantities. We show which segments share meaningful amounts of information across all experiments, implying general segregation rules in influenza, and which segments only share significant information for particular strain pairings. Significantly, we quantify how much information they actually share—a key component in determining the diversity of progeny. Finally, we extend our method to reoviruses, a member of the reoviridae family, which includes rotavirus, the leading cause of acute childhood diarrhea worldwide (17).
In a typical experiment, a relevant cell type is coinfected with two different strains, and the repertoire of progeny viruses is explored. These experiments separate intrinsic biases from those observed in circulating strains, in refs. 11 and 12, that may have additional causes. Suppose two strains are introduced to cells in culture at equal multiplicities of infection (MOI), a typical experimental scenario. MOI is defined as the ratio of infectious agents to host targets, so each strain, ideally, is equally likely to infect a target cell. After an experiment, the output probability that a segment comes from a given parental strain may no longer be the input value of one-half. We quantify this effect as the entropy change per segment between the input probability distribution that a segment came from a given strain versus the output distribution.
If bias exists toward how pairs of segments appear together in the output virus, such as may arise from packaging effects, this will be captured by the mutual information shared between those two segments. The entropy per segment constrains this quantity: The mutual information between segments is always less than the minimum entropy per segment. For these quantities we have designed a nonparametric “channel scrambling” test for generating p values. Furthermore, we use the total correlation to capture structures of a higher order than pairwise, a feature not found in previous analyses. A hypothetical case is represented in Fig, 1, where segments 2, 3, and 7 have a significant total correlation. In this case the segments, taken individually, are equally likely to come from the same strain of origin. Yet, if one segment has a given strain of origin, the other two segments will also come from the same strain.
Fig. 1.
Representation of a reassortment structure. Two parental strains (indicated as red and blue) produce a set of eight reassortant viruses. One can interpret the relation between the segments as an exchange of information. Segments 2, 3, and 7 constitute the most statistical significant structure (Total Correlation = 2, p value = 8.2 × 10-4), more significant than any pair correlation. Although its strain is uniformly distributed when examined on its own, whenever segment 2 is a given color, segments 3 and 7 are the same color.
By formalizing the mathematical analysis for this process and testing that analysis on both new and existing reassortment data, we may better understand reassortment outcomes and predict limitations upon which virus will emerge, resulting in better preparedness. That is the goal of this program. While a full exploration of the true set of reassortment biases requires a large-scale exploration of all combinations of infecting strains, infected cell type, and cell-type species of origin, in a manner that faithfully reflects the likely backgrounds and cellular response states in which coinfecting strains could replicate, our approach makes the problem more quantitative and informative. In doing so, we demonstrate both how this method can be used in future experiments to assess pandemic risk and to uncover fundamental limits on the ability to communicate strain information between viral segments.
Results
Entropy Change per Segment in Coinfection Experiments.
In the classic coinfection experiment, first outlined in Lubeck et al., two strains of equal MOI of 1–5 PFU (plaque forming units) per cell are introduced to MDCK cell culture (11). However, the total concentration of a given segment in the progeny viruses may be far from uniform. If two strains are introduced to these cells, and each segment has an initial probability of ½ for having come from a given strain, then each segment will have an initial entropy per segment of 1 bit. We assume that there are initially M strains introduced to a cell with an equal probability, although this does not need to be the case, and pn(s) is the probability that a given segment, n, from the output viruses came from strain, s. If the entropy of a segment, n, for output viruses is defined as
![]() |
the entropy change for that segment will be
Typically, there are two equally probable coinfecting strains and the first term will then be equal to 1. The above quantity measures how much the output distribution deviates from uniformity. For a given segment, a value near zero would indicate that, in the output viruses, one is equally likely to see a segment come from either strain. If the value is close to 1, it indicates that this segment is dominantly from one of the two input strains in the output viruses. Hence, a change in entropy implies that one type of progeny virus is now more likely to appear than another, whereas previously that was not the case.
We analyze this quantity for an original experiment in which MDCK cells were coinfected at MOI of 1 PFU/cell of seasonal H1N1 (A/Hong Kong/226654/07) and 2009 pdm (A/California/4/09). MDCK cells were incubated with virus inoculums for 1 h at room temperature and then were briefly washed by acidic buffer to inactivate nonincorporated wild-type parental viruses. The virus was allowed to grow, and the supernatant was collected at 72 h post infection and used to perform standard plaque assays. Plaques were purified and their genotypes were identified by RT-PCR using segment specific primers for both strains as previously described (18). A detailed description of these procedures is available in SI Text, along with a full table of results.
The results of our original experiment were compared to two previous experiments in which MDCK cells were coinfected with influenza strains. MDCK cells are commonly used to measure infection of a cell by multiple possible strains of origin, given the ability of a wide variety of influenza subtypes to infect and grow productively in these cells (19). Because of this, they provide a standard context for comparison of intrinsic reassortment potential. The first comparison experiment, performed, by Li et al., cotransfected seasonal human H3N2 and equine H7N7 expression plasmids into 293T cells (20). The data in this study was derived by reverse genetics techniques, and the viral repertoire studied was generated by cotransfection of viral segments producing plasmids for H3N2 and H7N7 strains as well as the viral polymerase complex protein expressing plasmids for H1N1. The second is the aforementioned Lubeck experiment in which seasonal human H1N1 and H3N2 strains coinfect MDCK cells (11). It should be noted that these experiments used different experimental settings and time courses to generate recombinant viruses. Given that there may be different experimental biases in each of these settings, using data generated from different experimental platforms might help us to reduce the overall bias in our analysis and would make trends that are observed across experimental platforms all the more convincing. A comparison of the different approaches used by the published influenza experiments under analysis in this work is also provided in SI Text.
Several noteworthy features are presented in Table 1 and Fig. 2. Entropy changes of greater than ½ for our experiment correspond to segments 4 (encoding HA) and 7 (M), respectively. In comparison, the experiment of Li et al., with two strains of very different species origin compared to the other two experiments, shows four segments with changes greater than 0.4 and Lubeck et al., with two human seasonal strains, shows no changes greater than 0.1515 bits. Comparative examination of these results shows that, in our case and in the case of Li et al., HA dominantly occurs from one strain, which also happens once for PB2 in the Li et al. experiment. Presumably these HA had an advantage over another as the reassortant viruses reinfected. It is also interesting to note that segments 6 (NA) and 8 (NS) show very little change in entropy across all three experiments, implying that one strain’s version of these proteins is typically not favored. Segment 8, consisting of two nonstructural proteins (NS1/NS2) and one of the more conserved sequences in influenza, may not be expected to have divergent function across strains, implying that one strain’s proteins have no real strain specific advantage.
Table 1.
Entropy per segment for MDCK reassortment experiments
Fig. 2.
Entropy change per segment in parallel reassortment experiments. In these three comparable reassortment experiments, the entropy change for each segment is shown for each of the eight viral segments. In all of these cases the initial entropy is 1 bit, the maximum possible value. Any entropy change is therefore an entropy loss.
Mutual Information Among Replication-Biased Segments.
Given that the entropy per segment is altered from a uniform probability, the opportunity for viral segments to communicate their strain to one another will become limited. For instance, in the case of the H3N2-H7N7 reassortment experiment of ref. 20, segment 1 has lost all entropy per segment because every progeny strain has segment 1 from the same parental strain. Hence, even if another segment would gain an advantage from pairing with a segment 1 from H7N7, such communication about strain type is not possible because segment 1 is always from H3N2. The ability of segment 1 to communicate information to another segment is closed.
In information theoretic terms, the mutual dependence between two random variables is quantified by the mutual information, applied to this problem as
![]() |
where pmn(si,sj) is the joint probability that segments m and n come from strains si and sj, respectively, and, therefore, Hmn is the entropy for the two-segment joint probability. The mutual information and entropy per segment constrain one another via the inequality
due to the fact that the joint entropy is always greater than the maximum individual entropy, as summarized in SI Text (15).
Two issues now present themselves. The first is to determine that a quantity of mutual information between two segments should be deemed significant, as opposed to when it could have been achieved randomly. We use a nonparametric test for associating a p value to the mutual information shared between segments. To obtain a p value, we randomly permute the order of the output strains for each segment—“scrambling” the message. We count the number of times, out of the number of permutations, that the observed mutual information is greater than the empirical value (following the convention of counting those on the boundary half of the time). Our procedure is further described in SI Text. Second, because there are eight segments in the virus, there are multiple pairs of segments that can share information. Multiple hypothesis correction must be taken into account. To focus on the clearest associations, we use Bonferroni corrections. There are 28 possible segment pairs, so we use a p value of 0.05/28 = 0.00179 for significance. The results of those segment combinations that pass the Bonferroni corrections, under 105 permutations, are listed in Table 2. In this table significant pairs are listed, along with their mutual information and p value. Also listed is the normalized mutual information: the mutual information divided by the maximum possible value given by the above inequality.
Table 2.
Mutual information for significant segment pairs for MDCK reassortants
As is clear from Table 2 and Fig. 3, the most consistently significant pairing is between segments 2 and 3 (PA). In each case this pair is significant, sharing between 0.1303 and 0.1912 bits of information per segment, with the more closely related strains, seasonal H1N1 and H3N2, typically sharing the most information. Each strain also has strain specific pairings, which may indicate an association that was significant to the biology of that particular strain combination but not to others. For instance, segments 3 and 8 show significant communication for 2009 pdm and seasonal H1N1, marginal association for H1N1 and H3N2, and no association for H3N2 and H7N7. The constancy of the 2–3 pairing across strains is important. In the original experiment on the subject, ref. 11, the entire polymerase complex was significantly associated—that is, segments 1, 2, and 3. However, it is clear from this work that if, in fact, polymerase segments pair preferentially by strain, only segments 2 and 3 pair as a general rule. In H3N2—H7N7, the H7N7 PB2 is completely dominant, while for H1N1—2009 pdm segments 1 and 2 make no significant associations.
Fig. 3.
Mutual information between segments versus significance. The amount of mutual information shared between segments is plotted versus the negative logarithm of the associated p value. The p value is calculated by the permutation test described in the text. Those mutual information values that meet the Bonferroni criterion separate at the far right from the group on the far right. The colors correspond to the same strain crosses as in Fig. 2.
Total Correlation For More Than Two Segments.
Our approach can be generalized for correlations between higher order complexes, such as triples of segments and so on. However, there are two drawbacks. First, there are many ways to extend the concept of mutual information to multivariate distributions (21). No one function captures all aspects of the two-dimensional mutual information for multivariate distributions. Second, looking for associations among greater combinations of segments can create more hypotheses to test, necessitating more experiments. For significant association among three segments, there are 56 possible hypotheses, and the Bonferroni correction would yield a p value of 0.05/56 = 0.000893.
Here we focus on the total correlation (22). This measure captures the notion that the proper measure of dependence is the Kullback–Leibler divergence between the multivariate distribution and the independent distribution. Another commonly used quantity, the multivariate mutual information, is a natural generalization of the idea that the relevant quantity in an N-dimensional information measure is the difference in information between an N - 1 dimensional subset and the probability distribution of that subset conditioned on an additional variable. This quantity, although containing useful insights, can be difficult to interpret in ambiguous cases (23). Unlike the total correlation and two-dimensional mutual information, it can be negative and, as such, is ill-suited for our significance test because there is not a monotonic interpretation of the meaningfulness of this quantity. Both quantities reduce to the mutual information in the two-dimensional case.
The total correlation is the difference between the total entropy of all single variables and the entropy of the N variable independent probability distribution. It is nonnegative, and one can define a nonparametric “channel scrambling” test as in the previous section. For three segments, l, m, and n, the total correlation is defined as
This quantity has attractive qualities in terms of convergence and decomposition (22). The equivalent of the aforementioned inequality for mutual information is that
as derived in SI Text. Significant total correlations are listed in Table 3 and all total correlations are plotted against the logarithm of their p values in Fig. 4. They follow the same pattern as the mutual information: More closely related strains have higher total correlations in general. While there is no completely consistent pattern, in all but one case significant total correlations contain two segments from the polymerase complex. This suggests that other segments may group with these polymerase complex pairs preferentially rather than with their individual segments. This is particularly true for the H3N2-H7N7 pairing, where the segment 2–3 pair, itself significant, appears significantly associated with every other segment. The distance between these two strains, one equine and one human, may therefore enhance the need for other segments to associate with this pair in a strain-specific manner, because the two strains evolved in comparatively different host backgrounds.
Table 3.
Significant total correlation values for three segment groupings
| Segments | Total correlation | p value |
| H1N1-2009 pdm | ||
| 238 | 0.3871 | 0.000015 |
| 237 | 0.2429 | 0.00063 |
| 378 | 0.2308 | 0.000855 |
| H3N2-H7N7 (20) | ||
| 236 | 0.1619 | 0.00001 |
| 235 | 0.1476 | < 10-5 |
| 234 | 0.1384 | 0.000025 |
| 238 | 0.136 | 0.00011 |
| 237 | 0.1305 | 0.00024 |
| 123 | 0.1303 | 0.00001 |
| 578 | 0.1185 | 0.00066 |
| H1N1-H3N2 (11) | ||
| 123 | 0.5065 | 0.00002 |
| 126 | 0.4542 | 0.00008 |
| 128 | 0.3921 | 0.00033 |
Fig. 4.
Total correlation values versus significance. The amount of total correlation between segment triplexes is plotted versus the negative logarithm of the associated p value. As in Fig. 3, the p value is calculated by permutation testing. Those mutual information values that meet the Bonferroni criterion separate at the far right from the group on the far right. The colors correspond to the same strain crosses as in Fig. 2.
Comparison to Other Experimental Settings and Segmented Viruses.
The above experiments deal with the important setting of equiprobable strain reassortment in MDCK cells. We now examine three variants on the previous experiments. In Octaviani et al., reassortants between H5N1 and 2009 pdm were examined in MDCK cells (24). For some strains, even when one infectious unit is provided per target cell for an infection, many cells may not be coinfected with both viruses due to a particular strain’s dominance. For example, in the pilot experiment of ref. 24 when H5N1 and 2009 pdm were recombined at equal MOI of 1 PFU/cell, the H5N1 was highly dominant, so the authors increased the MOI ratio to 5∶1. Clearly this is a significant study, because a combination of H5N1’s high case fatality rate with the 2009 pdm quick transmission ability could cause a significant public health crisis (25, 26). Because the initial probability distribution for a segment coming from a given strain is no longer (1/2, 1/2), but is now (1/6, 5/6), the full entropy formula is used for the initial distribution. The initial entropy is now 0.6836 bits per segment, rather than the previous 1 bit per segment. Because this is no longer the maximum entropy, the possibility exists that the entropy can either increase or decrease over the course of the experiment. In Table S2, we see that in most cases the entropy increased. That is, output viruses were typically more random per segment than the input viruses. For segment 5 (NP) the trend was reversed and for segment 1 it was minimal. In Varich et al., a human H1N1 virus (A/WSN/33) and an avian H4N6 virus (A/Duck/Czechoslovakia/56) coinfect chicken embryos at high, equal MOI of 5–10 PFU/cell (27). Both experiments again show the potential variability of the HA containing segment from the random distribution—in each case one particular HA is preferred. Likewise, we calculated the entropy change per segment for equal MOI of 5–10 PFU/cell reassortment experiments on two mammalian reovirus strains performed by Nibert et al. (28). Reoviruses are 10 segment dsRNA viruses, so their p-value cutoff is altered accordingly, showing the strongest associations between segments 3 and 7.
An interesting aspect of these variants comes when the mutual information is examined between influenza segments, as shown in Table S3. In the case of H5N1—2009 pdm, there are many unique associations that are not seen for other strains. However, some association between segments 2 and 3 continues to persist—it is just below our cutoff. For the H4N6-H1N1 chicken egg reassortment, the highest amount of mutual information in any experiment is recorded, between segments 3 and 5 and 4 and 7. Although this suggests fundamentally different interactions in avian cells, other experiments are needed before drawing such conclusions.
Discussion
Quantifying reassortment bias is critical for understanding potential future strains and gaining insight into the environment in which segmented viruses operate. As we have shown, information theory provides a natural approach to interpreting and contextualizing experimental results across multiple settings. Our method is applicable to any segmented virus, or any system where genetic segments from multiple sources combine in progeny. From it one gains the insights that come from showing that this problem can be interpreted in an information theoretic context, biological information from experiments directly studied in this work, and a template for future studies that will assist practical and theoretical endeavors. We summarize these features below.
At the single segment level, the entropy per segment captures a change from the input probability of that segment coming from a particular strain. If input segments were all equally likely to have come from one of two strains, it would be expected that, for N output viruses, there are 2N equally likely outputs for each segment. If the entropy, H, is reduced and one has a reliable estimate of its value, there will “typically” be 2HN equally likely experimental outputs. For all eight segments the number of readouts will therefore change from
In these cases, one needs solid estimates, rather than just ascertaining significance. Because these quantities are biased estimators, to minimize error one wants to ensure, in such a planned experiment, that all strain probabilities, pi, satisfy Npi≫1 (29).
Further gain comes from separating the preceding step from the mutual information between segments and beyond to total correlations. Assuming the mutual information between segments 2 and 3 is significant across strains, then, in all cases, the contribution of those terms to the diversity of strains will further change from
If one assumes that enough trials have been performed, as indicated by significance under our method, one can now get a handle on which reassortants are most probable. The ratio of the likelihood of one reassortant to another would give a gross sense of relative fitness (this would not be exact because the number of generations is unclear). If the mutual information is significant, then the probability that two segments came from a given strain is to be replaced by the joint probability, and, likewise, a significant total correlation would indicate replacement of the independent probabilities of, say, three segments, with their joint probabilities.
The power of these methods can be increased when combined with approaches such as that of Li et al. (30, 31). In this work, all 256 possible output viruses were examined for their virulence potential. Because our approach gives a better handle on which reassortants are most likely to be produced, one can examine the overlap between the set of most probable viruses and a viral phenotype. We could not find a strain cross that has been assessed in the literature for both reassortment and phenotype of all strain crosses. However, the methods employed here inform future experiments, where virulence could also be combined or substituted with transmissibility or another relevant phenotype.
If a likely reassortant also has a stronger than normal phenotype, then a higher priority can be placed on preventive measures, such as vaccine preparation, surveillance, or eradication of an infected bird or pig host. When such a case exists, that strain can also be examined more closely, such as by querying its immunostimulatory potential in relevant cell types. In this way, one can attempt to characterize potentially virulent, pandemic strains before they can enter a population. This is one of the real values of our approach. Only a rare reassortant may occur with the probability and virulence to cause a dangerous pandemic. With the proper experiments one could estimate this probability, identify any viral genomic alterations that could put a population at risk, and thus respond accordingly. If, through surveillance, two strains are found in the same location, with a suitable host for mixing these strains in the wild, we could estimate the diversity of reassortant strains prior to this mixing occurring. These predictions could well provide measured responses depending upon the threat level.
Biologically, these experiments open a window into the cellular environment in which influenza replicates. The largest mutual information was associated with accessory proteins to the viral polymerase, PB1 significantly associated with PA across all similar experiments, but not consistently with PB2 as some had found. One possibility is related to the work of ref. 32, which indicates that a PB1-PA dimer is formed separately from a monomeric PB2 and assembled in the nucleus. Our results suggest that this pairing may well be the most consistently significant effect on viral progeny diversity across strains. However, it remains to be determined experimentally if this advantage comes from the favorability of dimerization or an optimization of one dimer over another for, say, more efficient polymerization. In addition to the association of segments 2 and 3 by strain, we also note that segments 6 and 8 seem to show indifference to their strain of origin. Moreover, segment 4, containing HA, is highly variable between experiments, ranging from random to almost complete dominance. In this case, because HA binds to sialic acid for viral entry, one would be inclined to attribute the change in progeny diversity to the fitness of one HA over another. Future reassortment experiments at a single cell level would make an interesting point of comparison.
Theoretically, if a sufficient quantity of strain combination experiments were studied, the maximum mutual information would give an indication of the channel capacity between strains, giving a bound on how much information can be possibly communicated between two segments in a noisy environment. If the rate of reassortment is pushed past this limit, it could disrupt viral segment communication, a possible novel defense strategy. As more consistent groupings are found, over many experiments, investigators can narrow in on these consistent sets to see whether the increase in likelihood of certain progeny being found over others in a given environment is due to factors such as the fitness of a reassortant, functional constraints on the proteins, or genome packaging sequence differences. Understanding how much these segments transfer information about their strain of origin, and to what extent this is possible, can ultimately lead to novel antiviral strategies. We provide a quantitative framework, along with specific examples of what can be discovered, giving a greater sense of how to assess preparedness and the limits of viral segment communication.
Supplementary Material
Acknowledgments.
We would like to thank Vladimir Trifonov, Hossein Khianbanian, Edo Kussell, Yoshihiro Kawaoka, and Gabriele Neumann for helpful discussions and comments. We also thank Polly Mak for her technical support. R.R. is supported by the Northeast Biodefence Center (U54-AI057158), the National Institutes of Health (U54 CA121852-05), and the National Library of Medicine (1R01LM010140-01). The laboratory work was supported by the Area of Excellence Scheme of the University Grants Committee Hong Kong (AoE/M-12/06) and the Research Fund for the Control of Infectious Disease Commissioned Project from Food and Health Bureau, Hong Kong. B.D.G. is the Eric and Wendy Schmidt Member in Biology at the Institute for Advanced Study and would like to thank them for their support.
Footnotes
The authors declare no conflict of interest.
This article is a PNAS Direct Submission.
This article contains supporting information online at www.pnas.org/lookup/suppl/doi:10.1073/pnas.1113300109/-/DCSupplemental.
References
- 1.Taubenberger JK, Morens DM. 1918 influenza: The mother of all pandemics. Rev Biomed. 2006;17:69–79. doi: 10.3201/eid1201.050979. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 2.Kilbourne ED. Infleunza pandemics of the twentieth century. Emerg Infect Dis. 2006;12:9–14. doi: 10.3201/eid1201.051254. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 3.Novel Swine-Origin Influenza A (H1N1) Virus Investigation Team. Emergence of a novel swine-origin influenza A (H1N1) virus in humans. New Engl J Med. 2009;360:2605–2615. doi: 10.1056/NEJMoa0903810. [DOI] [PubMed] [Google Scholar]
- 4.Trifonov V, Khiabanian H, Rabadan R. Geographic dependence, surveillance, and origins of the 2009 influenza A (H1N1) virus. New Engl J Med. 2009;361:115–119. doi: 10.1056/NEJMp0904572. [DOI] [PubMed] [Google Scholar]
- 5.Smith GJ, et al. Origins and evolutionary genomics of the 2009 swine-origin H1N1 influenza A epidemic. Nature. 2009;459:1122–1125. doi: 10.1038/nature08182. [DOI] [PubMed] [Google Scholar]
- 6.Maines TR, et al. Pathogenesis of emerging avian influenza viruses in mammals and the host innate immune response. Immunol Rev. 2008;225:68–84. doi: 10.1111/j.1600-065X.2008.00690.x. [DOI] [PubMed] [Google Scholar]
- 7.Kobasa D, et al. Aberrant innate immune response in lethal infection of macaques with the 1918 influenza virus. Nature. 2007;445:319–323. doi: 10.1038/nature05495. [DOI] [PubMed] [Google Scholar]
- 8.Diebold SS, et al. Innate antiviral responses by means of TLR7-mediated recognition of single-stranded RNA. Science. 2004;303:1529–1531. doi: 10.1126/science.1093616. [DOI] [PubMed] [Google Scholar]
- 9.Greenbaum BD, Rabadan R, Levine A. Patterns of oligonucleotide sequences in viral and host cell RNA identify mediators of the host innate immune system. PLoS One. 2009;4:e5969. doi: 10.1371/journal.pone.0005969. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 10.Jimenez-Baranda S, et al. Oligonucleotide motifs that disappear during the evolution of influenza in humans increase IFN-alpha secretion by plasmacytoid dendritic cells. J Virol. 2011;85:3893–3904. doi: 10.1128/JVI.01908-10. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 11.Lubeck MD, Palese P, Schulman JL. Nonrandom association of parental genes in influenza A virus recombinants. Virology. 1979;95:269–274. doi: 10.1016/0042-6822(79)90430-6. [DOI] [PubMed] [Google Scholar]
- 12.Muramoto Y, et al. Hierarchy among viral RNA (vRNA) segments in their role in vRNA incorporation into influenza A virions. J Virol. 2006;80:2318–2325. doi: 10.1128/JVI.80.5.2318-2325.2006. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 13.Rabadan R, Levine AJ, Krasnitz M. Non-random reassortment in human influenza A viruses. Influenza Other Respi Viruses. 2008;2:9–22. doi: 10.1111/j.1750-2659.2007.00030.x. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 14.Khiabanian H, Trifonov V, Rabadan R. Reassortment patterns in Swine influenza viruses. PLoS One. 2009;4:e7366. doi: 10.1371/journal.pone.0007366. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 15.Cover TM, Thomas JA. Elements of Information Theory. New York: Wiley; 1991. [Google Scholar]
- 16.Shannon C. A Mathematical Theory of Communication. Bell Syst Tech J. 1948;27:379–423. [Google Scholar]
- 17.United Nations Children’s Fund/World Health Organization. Diarrhoea: Why children are still dying and what can be done. 2009. (United Nations Children’s Fund, New York; World Health Organization, Geneva)
- 18.Poon LL, et al. Rapid detection of reassortment of pandemic H1N1/2009 influenza virus. Clin Chem. 2010;56:1340–1344. doi: 10.1373/clinchem.2010.149179. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 19.Gaush CR, Smith TF. Replication and plaque assay of influenza virus in an established line of canine kidney cells. Appl Microbiol. 1968;16:588–594. doi: 10.1128/am.16.4.588-594.1968. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 20.Li C, Hatta M, Watanabe S, Neumann G, Kawaoka Y. Compatibility among polymerase subunit proteins is a restricting factor in reassortment between equine H7N7 and human H3N2 viruses. J Virol. 2008;82:11880–11888. doi: 10.1128/JVI.01445-08. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 21.Han TS. Multiple mutual informations and multiple interactions in frequency data. Inform Control. 1980;46:26–45. [Google Scholar]
- 22.Watanabe S. Information theoretical analysis of multivariate correlation. IBM J Res Dev. 1960;4:66–82. [Google Scholar]
- 23.McGill WJ. Multivariate information transmission. IEEE Trans Inf Theory. 1954;4:93–111. [Google Scholar]
- 24.Octaviani CP, Ozawa M, Yamada S, Goto H, Kawaoka Y. High level of genetic compatibility between swine-origin H1N1 and highly pathogenic avian H5N1 influenza viruses. J Virol. 2010;84:10918–10922. doi: 10.1128/JVI.01140-10. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 25.Li FCK, Choi BCK, Sly T, Pak AWP. Finding the real case-fatality rate of H5N1 avian influenza. J Epidemiol Community Health. 2008;62:555–559. doi: 10.1136/jech.2007.064030. [DOI] [PubMed] [Google Scholar]
- 26.Fiebig L, et al. Avian influenza A (H5N1) in humans: new insights from a line list of World Health Organization confirmed cases, September 2006 to August 2010. Euro Surveill. 2011;11:19941. [PubMed] [Google Scholar]
- 27.Varich NL, Gitelman AK, Shilov AA, Smirnov YA, Kaverin NV. Deviation from the random distribution pattern of influenza A virus gene segments in reassortants produced under non-selective conditions. Arch Virol. 2008;153:1149–1154. doi: 10.1007/s00705-008-0070-5. [DOI] [PubMed] [Google Scholar]
- 28.Nibert ML, Margraf RL, Coombs KM. Nonrandom segregation of parental alleles in reovirus reassortants. J Virol. 1996;70:7295–7300. doi: 10.1128/jvi.70.10.7295-7300.1996. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 29.Carlton AG. On the bias of information estimates. Psychol Bull. 1969;71:108–109. [Google Scholar]
- 30.Li C, et al. Reassortment between avian H5N1 and human H3N2 influenza viruses creates hybrid viruses with substantial virulence. Proc Natl Acad Sci USA. 2010;107:4687–4692. doi: 10.1073/pnas.0912807107. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 31.Sun Y, et al. High genetic compatibility and increased pathogenicity of reassortants derived from avian H9N2 and pandemic H1N1/2009 influenza viruses. Proc Natl Acad Sci USA. 2011;108:4164–4169. doi: 10.1073/pnas.1019109108. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 32.Deng T, Sharps J, Fodor E, Brownlee GG. In vitro assembly of PB2 with a PB1-PA dimer supports a new model of assembly of influenza A virus polymerase subunits into a functional trimeric complex. J Virol. 2005;79:8669–8674. doi: 10.1128/JVI.79.13.8669-8674.2005. [DOI] [PMC free article] [PubMed] [Google Scholar]
Associated Data
This section collects any data citations, data availability statements, or supplementary materials included in this article.






