Skip to main content
eLife logoLink to eLife
. 2022 Oct 11;11:e79570. doi: 10.7554/eLife.79570

Variation in ubiquitin system genes creates substrate-specific effects on proteasomal protein degradation

Mahlon A Collins 1,, Gemechu Mekonnen 1,, Frank Wolfgang Albert 1,
Editors: Magnus Nordborg2, David Ron3
PMCID: PMC9634822  PMID: 36218234

Abstract

Precise control of protein degradation is critical for life, yet how natural genetic variation affects this essential process is largely unknown. Here, we developed a statistically powerful mapping approach to characterize how genetic variation affects protein degradation by the ubiquitin-proteasome system (UPS). Using the yeast Saccharomyces cerevisiae, we systematically mapped genetic influences on the N-end rule, a UPS pathway in which protein N-terminal amino acids function as degradation-promoting signals. Across all 20 possible N-terminal amino acids, we identified 149 genomic loci that influence UPS activity, many of which had pathway- or substrate-specific effects. Fine-mapping of four loci identified multiple causal variants in each of four ubiquitin system genes whose products process (NTA1), recognize (UBR1 and DOA10), and ubiquitinate (UBC6) cellular proteins. A cis-acting promoter variant that modulates UPS activity by altering UBR1 expression alters the abundance of 36 proteins without affecting levels of the corresponding mRNA transcripts. Our results reveal a complex genetic basis of variation in UPS activity.

Research organism: S. cerevisiae

Introduction

Protein degradation is an essential biological process that occurs continuously throughout the life of a cell. Degradative protein turnover helps maintain protein homeostasis by regulating protein abundance and eliminating misfolded and damaged proteins from cells (Varshavsky, 2011; Collins and Goldberg, 2017; Hanna and Finley, 2007). In eukaryotes, most protein degradation occurs through the concerted actions of the ubiquitin system and the proteasome, together known as the ubiquitin-proteasome system (UPS) (Coux et al., 1996; Collins and Goldberg, 2017; Hershko and Ciechanover, 1998; Bachmair et al., 1986; Ciechanover et al., 2000). Ubiquitin system enzymes bind degradation-promoting signal sequences, termed degrons (Varshavsky, 1991), in cellular proteins and mark them for degradation by covalently attaching chains of the small protein ubiquitin (Bett, 2016; Hershko and Ciechanover, 1998; Finley et al., 2012). The proteasome binds poly-ubiquitinated proteins, then processively deubiquitinates, unfolds, and degrades them to small peptides (Kisselev et al., 1999). The UPS degrades a wide array of proteins spanning diverse biological functions and subcellular localizations (Schwanhäusser et al., 2011; Kong et al., 2021; Christiano et al., 2020). By controlling the turnover of a large fraction of the cellular proteome, the UPS regulates numerous aspects of cellular physiology and function, including gene expression, protein homeostasis, cell growth and division, stress responses, and energy metabolism (Varshavsky, 2011; Hershko and Ciechanover, 1998; Hanna and Finley, 2007; Pohl and Dikic, 2019).

Because of the central role of UPS protein degradation in regulating protein abundance, variation in UPS activity can influence a variety of cellular and organismal phenotypes (Varshavsky, 2011; Schwartz and Ciechanover, 1999; Hanna and Finley, 2007; Schmidt and Finley, 2014). Physiological variation in UPS activity enables cells to respond to changes in their internal and external environments. For example, UPS activity increases when misfolded or oxidatively damaged proteins accumulate, preventing these molecules from damaging the cell (Sontag et al., 2014; Grimm et al., 2012; Finley and Prado, 2020). Conversely, UPS activity decreases during nutrient deprivation, when the energetic demands of UPS protein degradation would be costly to the cell (Waite et al., 2016; Laporte et al., 2008; Bajorek et al., 2003). Variation in UPS activity may also create discrepancies between protein degradation and the proteolytic needs of the cell, leading to adverse phenotypic outcomes. For example, age-related declines in UPS activity exacerbate the accumulation of damaged and misfolded proteins that occurs during aging, compromising protein homeostasis and, in turn, cellular viability (Stolzing and Grune, 2001; Baraibar and Friguet, 2012; Shringarpure and Davies, 2002). Understanding the sources of variation in UPS activity thus has considerable implications for our understanding of the many traits influenced by protein degradation.

A handful of examples have shown that variation in UPS activity can be caused by individual genetic differences. Rare mutations that ablate or diminish the function of ubiquitin system or proteasome genes impair UPS protein degradation and cause a variety of incurable syndromes. For example, nonsense and frameshift mutations in UBR1, an E3 ubiquitin ligase that targets proteins for proteasomal degradation, cause the developmental disorder Johanson-Blizzard Syndrome (Zenker et al., 2005). Several UBR1 missense mutations that moderately decrease Ubr1 activity cause less severe forms of Johanson-Blizzard Syndrome (Hwang et al., 2011), suggesting a continuum of variant effects on UPS activity, similar to other genetically complex traits. More recently, proteasome gene missense mutations that impair proteasome assembly and reduce proteasome activity were shown to cause the autoimmune disorder proteasome-associated autoinflammatory syndrome (Brehm et al., 2015; Arima et al., 2011; Liu et al., 2012), further establishing individual genetic differences as a potentially important source of variation in UPS activity. Genome-wide association studies have also linked variation in ubiquitin system (Xia et al., 2014; Diskin et al., 2012) and proteasome genes (Cho et al., 2011) to a variety of disorders, but have neither established the individual causal variants nor tested their effects on UPS activity.

Our understanding of how natural genetic variation affects the UPS comes largely from these limited examples, leaving critical knowledge gaps in several key areas. First, a focus on rare, large-effect mutations linked to Mendelian syndromes likely provides a narrow, incomplete view of the genetics of UPS activity. Most traits are genetically complex, shaped by many loci of small effect and few loci of large effect throughout the genome (Mackay et al., 2009; Ehrenreich et al., 2009), suggesting that variants that completely or largely ablate UPS gene functions represent only one extreme of a continuum of genetic effects on UPS activity. Second, we have virtually no knowledge of how natural variation in non-UPS genes affects UPS activity. Third, variation in UPS activity can differentially affect the degradation of distinct UPS substrates (Christiano et al., 2020; Kong et al., 2021). Whether genetic effects on UPS activity affect the turnover of distinct proteins consistently or in a substrate-specific manner remains a fundamentally open question. Finally, we do not know how genetic effects on UPS activity influence other traits. For example, many genetic effects on gene expression influence protein levels without altering mRNA abundance for the same gene (Battle et al., 2015; Ghazalpour et al., 2011; Mirauta et al., 2020; Albert et al., 2014; Brion et al., 2020; Chick et al., 2016; Cenik et al., 2015; Abell et al., 2022; Foss et al., 2011). These protein-specific effects could arise through differences in UPS activity, but there have been no efforts to understand how natural variation that alters UPS activity influences global gene expression at the protein and RNA levels.

Technical challenges have precluded a comprehensive view of the genetics of UPS activity. Mapping genetic influences on a trait with high statistical power requires assaying large, genetically diverse populations of thousands of individuals (Bloom et al., 2013). At this scale, in vitro biochemical assays of UPS activity are impractical. Several synthetic reporter systems can measure UPS activity with high-throughput in vivo (Geffen et al., 2016; Yu et al., 2016; Yen et al., 2008). However, these systems use genetically encoded fluorescent proteins coupled to degrons to measure UPS activity. When deployed in genetically diverse populations, their output is likely confounded from genetic effects on reporter expression levels.

Here, we leveraged advances in synthetic reporter design to obtain high-throughput, reporter expression level-independent measurements of UPS activity in millions of live, single cells. We use these measurements to map genetic influences on the N-end rule, a UPS pathway that recognizes degrons in protein N-termini (N-degrons) (Varshavsky, 1991) of thousands of endogenous cellular proteins (Kats et al., 2018; Bartel et al., 1990; Hwang et al., 2010; Varshavsky, 2011). Different N-degrons are processed by one of two distinct targeting systems (Figure 1A), which allowed us to test for potential pathway-specific effects of natural genetic variation on UPS activity. Systematic, statistically powerful genetic mapping revealed the complex, polygenic genetic architecture of UPS activity. Across the set of 20 N-degrons, we identified 149 loci influencing UPS activity, many of which had pathway- or substrate-specific effects. Resolving causal nucleotides at four loci identified regulatory and missense variants in ubiquitin system genes whose products process, recognize, and ubiquitinate cellular proteins. By measuring the effect of a causal variant in the UBR1 promoter on the transcriptome and proteome, we implicate genetic influences on UPS activity as a potentially prominent source of post-translational variation in gene expression.

Figure 1. UPS N-end rule activity reporters and genetic mapping method.

(A) Schematic of the production and degradation of UPS activity reporters according to the UPS N-end rule. (B) Density plots of the log2 RFP / GFP ratio from 10,000 cells for each of 8 independent biological replicates per strain per reporter for representative Arg/N-end and Ac/N-end pathway reporters. "BY" and "RM" are genetically divergent yeast strains. "BY rpn4Δ", "BY ubr1Δ”, and "BY doa10Δ" carry the indicated gene deletions in the BY background and were used as reporter control strains. (C) The median from each biological replicate in B. was scaled, normalized, and plotted as a stripchart such that y axis values are directly proportional to UPS activity. (D). Heatmap for all strains and N-degrons using data generated as in C. Symbols above the heatmap denote significant UPS activity differences between BY and RM. "*" indicates 0.05 > Tukey HSD p > 1e-6; “#” indicates Tukey HSD p < 1e-6. (E) Schematic of the bulk segregant analysis genetic mapping method used to identify UPS activity QTLs. (F) Density plot of the UPS activity distribution for a genetically diverse mapping population harboring the tryptophan (Trp) N-degron reporter. Dashed vertical lines show the thresholds used to collect cells with extreme UPS activity, which correspond to the high and low UPS activity pools denoted in E. (G) Backplot of the cells collected in F. onto a scatter plot of GFP and RFP.

Figure 1—source data 1. Results of all between-strain comparisons for all N-degron TFTs.

Figure 1.

Figure 1—figure supplement 1. Comparison of UPS activity between strains across N-degron reporters.

Figure 1—figure supplement 1.

The -log2 RFP / GFP ratio value was extracted from 10,000 cells from each of 8 independent biological replicates per strain per reporter and converted to Z-scores. High values correspond to high UPS activity and low values correspond to low UPS activity. Tukey HSD p-values for each between strain comparison for each reporter are listed in Figure 1—source data 1.

Figure 1—figure supplement 2. Overview of the constructs and strain construction steps used to generate yeast strains harboring TFT UPS activity reporters.

Figure 1—figure supplement 2.

Results

Single-cell measurements identify heritable variation in UPS activity

To understand how genetic variation influences UPS activity, we focused on the N-end rule, in which a protein’s N-terminal amino acid functions as an N-degron that results in a protein’s ubiquitination and proteasomal degradation (Figure 1A). The UPS N-end rule can be subdivided into the Arg/N-end and Ac/N-end pathways based on the molecular properties and recognition mechanisms of each pathway’s constituent N-degrons (Figure 1A; Varshavsky, 2011). We reasoned that the breadth of degradation signals and recognition mechanisms encompassed in the N-end rule would allow us to identify diverse genetic influences on UPS activity and that the well-characterized effectors of the N-end rule would aid in defining the molecular mechanisms of variant effects. We used a previously described approach (Varshavsky, 2005) to generate constructs containing each of the 20 possible N-degrons and appended these sequences to tandem fluorescent timers (TFTs; Figure 1A; Khmelinskii et al., 2012). TFTs are fusions of a rapidly maturing green fluorescent protein (GFP) and a slower maturing red fluorescent protein (RFP) (Khmelinskii et al., 2012; Khmelinskii and Knop, 2014). The TFT’s output, expressed as the -log2 RFP / GFP ratio, is directly proportional to its degradation rate and, when fused to N-degrons, measures UPS N-end rule activity (Kats et al., 2018; Kong et al., 2021; Khmelinskii et al., 2012). Because the TFT is expressed as a single protein construct, the output of the TFT is also independent of its expression level (Kats et al., 2018; Khmelinskii et al., 2014; Khmelinskii et al., 2012; Kong et al., 2021), enabling its use in genetically diverse populations.

We characterized the performance of our TFTs by measuring their output in yeast strains with gene deletions that alter UPS activity towards N-end rule substrates. As expected, deleting the E3 ubiquitin ligases of the Arg/N-end (UBR1) and the Ac/N-end (DOA10) pathways specifically stabilized N-degron TFTs from these pathways (corrected p < 0.05 vs. the BY strain, Figure 1B–D, Figure 1—figure supplement 1, Figure 1—source data 1). Deleting RPN4, which encodes a transcription factor for proteasome genes, reduces proteasome activity (Xie and Varshavsky, 2001) and stabilized reporters from both the Arg/N-end and Ac/N-end pathways (corrected p < 0.05 vs. the BY strain, Figure 1B–D, Figure 1—figure supplement 1, Figure 1—source data 1). These results show that our TFTs provide sensitive, quantitative, substrate-specific measures of UPS N-end rule activity.

To understand how natural genetic variation influences UPS activity, we compared two genetically divergent S. cerevisiae strains, the "BY" laboratory strain and the "RM" vineyard strain (Ehrenreich et al., 2009). RM had higher UPS activity than BY for 9 of 12 Arg/N-degrons and 6 of 8 Ac/N-degrons (corrected p < 0.05, Figure 1D, Figure 1—figure supplement 1, Figure 1—source data 1). BY had higher UPS activity than RM for the phenlyalanine, tryptophan, and tyrosine Arg/N-degrons (corrected p < 0.05, Figure 1D, Figure 1—figure supplement 1, Figure 1—source data 1). BY and RM had similar activity towards the methionine and proline Ac/N-degrons (corrected p > 0.05, Figure 1D, Figure 1—figure supplement 1, Figure 1—source data 1). Together, these results show that individual genetic differences create heritable, substrate-specific variation in UPS activity.

Genetic mapping reveals a complex, polygenic genetic architecture for UPS activity

We mapped quantitative trait loci (QTLs) for UPS activity using bulk segregant analysis (Figure 1E; Michelmore et al., 1991; Ehrenreich et al., 2010; Albert et al., 2014). In our implementation, this approach attains high statistical power by comparing whole-genome sequence data from pools of thousands of single cells with extreme UPS activity selected from a large population of haploid, recombinant progeny obtained by crossing BY and RM (Figure 1E–G; Ehrenreich et al., 2010; Albert et al., 2014). Using this method, we reproducibly identified 149 UPS activity QTLs across the set of 20 N-degrons at a false discovery rate of 0.5% (Figure 2A/B, Figure 2—source data 1, Supplementary file 1, Appendix 1). The number of QTLs per reporter ranged from 1 (for the Ile N-degron) to 15 (for the Ala N-degron) with a median of 7 (Figure 2B, Figure 2—source data 1). Using the absolute difference in allele frequency between the high and low UPS activity pools as a measure of effect size, we found that most QTLs had small effects, with only 5 loci (3%) causing an allele frequency difference greater than 0.5 (Figure 2C, Figure 2—source data 1). Thus, UPS activity is a complex, polygenic trait, shaped by variation throughout the genome.

Figure 2. UPS activity QTL mapping results.

Figure 2.

(A) Results from the alanine (Ala) N-degron reporter illustrate the results and reproducibility of the method. Asterisks denote QTLs, colored by biological replicate. (B) QTL mapping results for the 20 N-degrons. Colored blocks of 100 kb denote QTLs detected in each of two independent biological replicates, colored according to the direction and magnitude of the effect size (RM allele frequency difference between high and low UPS activity pools). Experimentally validated (boxed) and candidate (unboxed) causal genes for select QTLs are annotated above the plot. (C) Cumulative distributions of the effect size and direction for Arg/N-end and Ac/N-end QTLs. (D) Cumulative distribution of LOD scores for Arg/N-end and Ac/N-end QTLs.

Figure 2—source data 1. All N-end rule QTLs.
"chr" = chromosome, "LOD" = logarithm of the odds, "QTL_CI_left" = left index of QTL confidence interval, "QTL_peak" = peak position of QTL, "QTL_CI_right" = right index of QTL confidence interval.
Figure 2—source data 2. All distinct N-end rule QTL regions.
"chr" = chromosome, "QTL_CI_left" = left index of QTL confidence interval, "QTL_peak" = peak position of QTL, "QTL_right_CI" = right index of QTL confidence interval, "LOD" = logarithm of the odds, "RM_AFD" = RM allele frequency difference between high and low UPS activity pools.

Analysis of the set of UPS QTLs revealed several patterns. First, the RM allele was associated with higher UPS activity in a significant majority of UPS QTLs (89 out of 149, 60%, binominal test p = 0.021, Figure 2C), a result that is consistent with our observation that RM had higher UPS activity for 15 of 20 N-degrons (Figure 1D, Figure 1—source data 1). Second, the number and patterns of QTLs differed between the Ac/N-end and Arg/N-end pathways (Figure 2B, Figure 2—source data 1). The Ac/N-end pathway was affected by a significantly higher number of QTLs per reporter than the Arg/N-end pathway (9 vs 7, respectively, Wilcoxon test p = 0.021), while the QTLs with the largest effect sizes were found for the Arg/N-end pathway (Figure 2C / D, Figure 2—source data 1).

Third, multiple QTLs for distinct N-degrons occurred in close proximity and had the same direction of effect (Figure 2B), suggesting these QTLs may result from the same causal genes or variants. To better understand potential pleiotropy among the set of UPS activity QTLs, we computed overlap among the set of 149 UPS activity QTLs. We considered QTLs for distinct N-degrons overlapping when their peak position occurred within 100 kb and they had the same direction of effect (the sign of the RM allele frequency between the high and low UPS activity pools). Applying these criteria revealed that the 149 UPS activity QTLs were located at 35 distinct QTL regions (Figure 2—source data 2). Of these 35 regions, 23 (66%) affected only reporters from either the Arg/N-end (12) or Ac/N-end (11) pathways of the N-end rule (Figure 2—source data 2). Five of the 23 pathway-specific QTL regions affected only individual N-degrons (Figure 2—source data 2). Use of more lenient LOD score thresholds for QTL detection did not alter these general conclusions (Figure 2—source data 2, Supplementary file 2). Thus, the majority of QTLs for the N-end rule are pathway-specific, revealing considerable complexity in the genetics of UPS protein degradation.

Multiple causal DNA variants in UBR1 create substrate-specific effects on UPS activity

We leveraged the high degree of pathway specificity in our N-end rule QTLs to aid in the identification of causal genes in broad genomic QTL regions. A QTL on chromosome VII detected with 8 of 12 Arg/N-degron reporters (Figure 2B) was centered on UBR1, the E3 ligase that recognizes Arg/N-degrons and targets them for UPS protein degradation (Figure 3A). To determine whether UBR1 contains causal DNA variants for UPS activity towards Arg/N-degrons, we used the CRISPR-swap allelic engineering method (Lutz et al., 2019) to create BY strains with RM UBR1 alleles (see ‘Materials and methods’). Arg/N-degrons are classified as Type 1 or 2 depending on their Ubr1 binding site (Varshavsky, 2011; Bartel et al., 1990). The RM allele at the chromosome VII QTL was associated with decreased UPS activity towards Type 1 Arg/N-degrons and increased UPS activity towards Type 2 Arg/N-degrons (Figure 2B). We therefore tested the effects of the RM UBR1 alleles on two Type 1 (asparagine [Asn] and aspartate [Asp]) and two Type 2 Arg/N-degrons (tryptophan [Trp] and phenylalanine [Phe]).

Figure 3. Substrate-specific effects of UBR1 variants on the degradation of Arg/N-degrons.

(A) Schematic illustrating Ubr1’s role in Arg/N-degron recognition. (B) Multiple causal DNA variants in UBR1 shape UPS activity towards the Trp N-degron. The BY strain was engineered to contain full or partial RM UBR1 alleles as indicated and UPS activity towards the Trp N-degron TFT was measured by flow cytometry. UPS activity was Z-score normalized and scaled relative to the median of a control BY strain engineered to contain the full BY UBR1 allele. Each point in the plot shows the median of 10,000 cells for each of 16 independent biological replicates per strain. p-values at the top of the plot display the Benjamini-Hochberg corrected p-value for the t-test of the indicated strain versus the strain with the BY UBR1 allele. Box plot center lines, box boundaries, and whiskers display the median, interquartile range, and 1.5 times the interquartile range, respectively. (C). Barchart summarizing the effects of RM UBR1 alleles on UPS activity towards the indicated Type 1 and 2 Arg/N-degrons using data generated as in B. p-values in the plot display the Benjamini-Hochberg corrected p-value for the t-test of the indicated strain versus the control strain engineered to contain the BY UBR1 allele. (D) Diagram of the individual BY / RM UBR1 promoter variants. (E) as in C., but for the RM UBR1 promoter and individual BY / RM UBR1 promoter variants. (F) Sequence logo of the Hap5 binding motif created by the causal –469A>T UBR1 promoter variant. (G) Multi-species alignment of the UBR1 promoter at the causal –469A>T variant. Abbreviations: ‘S. para.’, Saccharomyces paradoxus; ‘S. mik.’, Saccharomyces mikatae;S. bay.’, Saccharomyces bayanus; ‘S. arb’, Saccharomyces arboricola; ‘S. pas.’, Saccharomyces pastorianus; ‘S. jur’, Saccharomyces jurei.

Figure 3.

Figure 3—figure supplement 1. Raw UBR1 full gene fine-mapping data.

Figure 3—figure supplement 1.

Raw UBR1 full gene fine-mapping results. The BY strain was engineered to contain full or partial RM UBR1 alleles as indicated and UPS activity towards the indicated Type 1 and Type 2 Arg/N-degron TFTs was measured by flow cytometry. UPS activity was Z-score normalized and scaled relative to the median of a control BY strain engineered to contain the full BY UBR1 allele. Each point in the plot shows the median of 10,000 cells for each of 16 independent biological replicates per strain per reporter. p-values at the top of the plot display the Benjamini-Hochberg-corrected p-value for the t-test of the indicated strain versus the strain with the BY UBR1 allele. Box plot center lines, box boundaries, and whiskers display the median, interquartile range, and 1.5 times the interquartile range, respectively. (A) UPS activity towards the indicated Type 1 Arg/N-degrons. (B) UPS activity towards the indicated Type 2 Arg/N-degrons.
Figure 3—figure supplement 2. Raw UBR1 promoter fine-mapping data.

Figure 3—figure supplement 2.

Fine-mapping the causal nucleotide in the UBR1 promoter. The BY strain was engineered to carry the RM UBR1 promoter or one of the two single nucleotide BY / RM UBR1 promoter variants as indicated and UPS activity towards the Trp and Phe Type 2 Arg/N-degrons was measured by flow cytometry. UPS activity was Z-score normalized and scaled relative to the median of a control BY strain engineered to contain the full BY UBR1 allele. Each point in the plot shows the median of 10,000 cells for each of 16 independent biological replicates per strain per reporter. p-values at the top of the plot display the Benjamini-Hochberg-corrected p-value for the t-test of the indicated strain versus the strain with the BY UBR1 allele. Box plot center lines, box boundaries, and whiskers display the median, interquartile range, and 1.5 times the interquartile range, respectively.
Figure 3—figure supplement 3. Population frequency and distribution of the causal UBR1 –469A>T variant.

Figure 3—figure supplement 3.

Population frequencies and distribution of causal variants. Tree diagrams show genetic distance among a global panel of S. cerevisiae isolates with branches colored according to which allele a strain carries. Indicated clades with the BY allele for a causal DNA variant are outlined.

Consistent with our QTL mapping results, The RM UBR1 allele significantly decreased the degradation rate of Type 1 Arg/N-degrons and increased the degradation rate of Type 2 Arg/N-degrons (corrected p < 0.05, Figure 3B/C, Figure 3—figure supplement 1). Thus, UBR1 is a causal gene for the chromosome VII QTL, and BY / RM variants in UBR1 differentially affect the degradation of Type 1 and 2 substrates of the Arg/N-end pathway.

QTL causal genes may contain multiple causal variants, making it necessary to test the effects of individual gene regions and variants in isolation (Lutz et al., 2019; Abell et al., 2022; Laurie-Ahlberg and Stam, 1987). We used CRISPR-swap to test the effect of partial RM UBR1 alleles on UPS activity towards Type 1 Arg/N-degrons. The RM open-reading frame (ORF) significantly decreased the degradation of the Asn, but not the Asp TFT (Figure 3C, Figure 3—figure supplement 1). The RM UBR1 promoter and terminator did not affect UPS activity towards either reporter (corrected p > 0.05, Figure 3C, Figure 3—figure supplement 1). Thus, variants in the RM UBR1 ORF are the main determinant of the gene’s effects on the Asn N-degron, while the effects of the RM UBR1 alleles on the Asp N-degron may be driven by epistatic interactions between variants in the promoter, ORF, and terminator.

The partial RM UBR1 alleles had drastically different effects on the degradation of Type 2 Arg/N-degrons (Figure 3C). Both the RM UBR1 promoter and ORF significantly increased UPS activity towards the Type 2 Trp and Phe Arg/N-degrons (corrected p < 0.05, Figure 3B/C, Figure 3—figure supplement 1). The RM UBR1 terminator did not affect the degradation of either Type 2 Arg/N-degron (corrected p > 0.05, Figure 3B / C, Figure 3—figure supplement 1). Thus, the RM UBR1 promoter and ORF each contain at least one causal variant that increases UPS activity towards Type 2 Arg/N-degron-containing substrates. Together with our Type 1 Arg/N-degron fine-mapping, these results establish that UPS activity QTLs can contain multiple causal DNA variants in a single gene that can differentially affect the turnover of distinct UPS substrates.

To identify individual causal variants, we tested the effect of the two BY / RM UBR1 promoter variants (Figure 3D) on UPS activity towards Type 2 Arg/N-degrons. The –469A>T variant significantly increased the degradation rate of the Trp and Phe N-degrons (corrected p < 0.05, Figure 3E, Figure 3—figure supplement 2). By contrast, the –197T>G variant had no effect on either N-degron, establishing –469A>T as the causal nucleotide in the UBR1 promoter (corrected p > 0.05, Figure 3E, Figure 3—figure supplement 2). The magnitude of the effect caused by –469A>T suggests that this variant accounts for the majority of UBR1 effects on the degradation of Type 2 Arg/N-end substrates (Figure 3B/C/E, Figure 3—figure supplements 1 and 2).

To gain further insight into the causal –469A>T variant, we examined its molecular properties, evolutionary history and population frequency using genome sequence data from a panel of 1,011 S. cerevisiae strains (Peter et al., 2018). The BY allele of the causal –469A>T variant in the UBR1 promoter disrupts a predicted binding site for the transcription activator Hap5 (Figure 3F) and decreased the output of a synthetic reporter in a massively parallel study of yeast promoter variants (Renganaath et al., 2020). BY carries the derived ‘A’ allele at –469A>T, which occurs in a poly(T) motif that is highly conserved across yeast species (Figure 3G). The population frequency of –469A>T is 1% and the variant is found in only in the Mosaic Region 1 clade that contains the BY strain (Figure 3—figure supplement 3). These results suggest the BY allele decreases UPS activity by decreasing UBR1 expression, which we subsequently validated at the RNA and protein levels (Figure 5). Moreover, the derived status and low population frequency of the BY allele at position –469 suggests that it may negatively impact organismal fitness, a notion consistent with the generally deleterious consequences of reduced UBR1 activity or expression (Zenker et al., 2006; Zenker et al., 2005; Chen et al., 2006).

Causal variants in functionally diverse ubiquitin system genes influence UPS activity

Some of the QTLs with the largest effects were specific to distinct N-end rule pathways or substrates and centered on known ubiquitin system genes (Figure 2B). We used allelic engineering to test whether these genes contained causal DNA variants for UPS activity.

A QTL on chromosome X was specific to the Type 1 asparagine (Asn) N-degron of the Arg/N-end pathway (Figure 2B). The QTL’s peak occurred within NTA1, which encodes an amidase that converts N-terminal asparagine and glutamine residues to aspartate and glutamate, respectively (Figure 4A). This processing is necessary to convert Asn and Gln N-ends into functional N-degrons (Baker and Varshavsky, 1995). NTA1 contains multiple BY / RM promoter variants and two missense variants that alter amino acids on the protein’s exterior surface (Figure 4B/D). Consistent with the chromosome X QTL effect, the full RM NTA1 allele significantly increased the degradation rate of the Asn TFT (corrected p < 0.05, Figure 4C, Figure 4—figure supplement 1). The RM NTA1 promoter did not alter the degradation rate of the Asn TFT (corrected p > 0.05, Figure 4C). Instead, the two BY / RM NTA1 missense variants, D111E and E129G, both influenced degradation of the Asn TFT, but in opposite directions. D111E decreased the Asn TFT’s degradation, while, E129G increased it (corrected p < 0.05, Figure 4C, Figure 4—figure supplement 1). The effect of E129G was in the same direction as that of the chromosome X QTL and was approximately threefold greater than that of the effect of D111E (Figure 4C, Figure 4—figure supplement 1). Thus, at NTA1, one causal variant's large effect masks the opposing, smaller effect of a second causal variant.

Figure 4. Identification of causal DNA variants for UPS activity in functionally diverse ubiquitin system genes.

(A, E, and I). Schematics showing the role of Nta1 (A), Doa10 (E), and Ubc6 (I) in UPS substrate processing, recognition, and ubiquitination, respectively. (B, F, and J). Location of regulatory and missense BY / RM variants, as well as active sites and functional domains in the proteins encoded by NTA1 (B), DOA10 (F), and UBC6 (J). C., G., and K. Fine-mapping results for NTA1 (C), DOA10 (G), and UBC6 (K). Benjamini-Hochberg corrected p-values are shown for the t-test of the indicated strain versus a control BY strain engineered to contain the BY allele of each gene. AlphaFold predicted protein structures for Nta1 (D), Doa10 (H), and Ubc6 (L) are shown with causal DNA variants, functional domains, active sites, and transmembrane helices highlighted. The inset in L. shows a predicted hydrogen bonding network at residue 229 in the BY Ubc6 protein.

Figure 4.

Figure 4—figure supplement 1. Raw NTA1 fine-mapping data.

Figure 4—figure supplement 1.

The BY strain was engineered to contain full or partial RM NTA1 alleles as indicated and UPS activity towards the Asn Arg/N-degron was measured by flow cytometry. UPS activity was Z-score normalized and scaled relative to the median of a control BY strain engineered to contain the full BY NTA1 allele. Each point in the plot shows the median of 10,000 cells for each of 16 independent biological replicates per strain per reporter. p-values at the top of the plot display the Benjamini-Hochberg-corrected p-value for the t-test of the indicated strain versus the strain with the BY NTA1 allele. Box plot center lines, box boundaries, and whiskers display the median, interquartile range, and 1.5 times the interquartile range, respectively.
Figure 4—figure supplement 2. Raw DOA10 fine-mapping data.

Figure 4—figure supplement 2.

The BY strain was engineered to contain full or partial RM DOA10 alleles as indicated and UPS activity towards the Gly and Thr Ac/N-degrons was measured by flow cytometry. UPS activity was Z-score normalized and scaled relative to the median of a control BY strain engineered to contain the full BY DOA10 allele. Each point in the plot shows the median of 10,000 cells for each of 16 independent biological replicates per strain per reporter. p-values at the top of the plot display the Benjamini-Hochberg-corrected p-value for the t-test of the indicated strain versus the strain with the BY DOA10 allele. Box plot center lines, box boundaries, and whiskers display the median, interquartile range, and 1.5 times the interquartile range, respectively.
Figure 4—figure supplement 3. Raw UBC6 fine-mapping data.

Figure 4—figure supplement 3.

The BY strain was engineered to contain full or partial RM UBC6 alleles as indicated and UPS activity towards the Thr and Ala Ac/N-degrons was measured by flow cytometry. UPS activity was Z-score normalized and scaled relative to the median of a control BY strain engineered to contain the full BY UBC6 allele. Each point in the plot shows the median of 10,000 cells for each of 16 independent biological replicates per strain per reporter. p-values at the top of the plot display the Benjamini-Hochberg-corrected p-value for the t-test of the indicated strain versus the strain with the BY UBC6 allele. Box plot center lines, box boundaries, and whiskers display the median, interquartile range, and 1.5 times the interquartile range, respectively.
Figure 4—figure supplement 4. Population frequencies and distributions of causal variants.

Figure 4—figure supplement 4.

Tree diagrams show genetic distance among a global panel of S. cerevisiae isolates with branches colored according to which allele a strain carries. Indicated clades with the BY allele for a causal DNA variant are outlined.

A QTL on chromosome IX detected for 6 of 8 Ac/N-end degrons contained DOA10, the E3 ligase of the Ac/N-end rule pathway (Figure 4E). The effect size of this QTL varied between Ac/N-degrons. We therefore tested the glycine (Gly) and threonine (Thr) reporters to determine whether BY / RM DOA10 variants exert substrate-specific effects on UPS activity. The RM DOA10 allele contains three missense variants, Q410E, K1012N, and Y1186F, and does not contain promoter or terminator variants (Figure 4F/H). The full RM DOA10 allele significantly increased the degradation of both the Gly and Thr TFTs (corrected p < 0.05, Figure 4G, Figure 4—figure supplement 2). When tested in isolation, all three BY / RM DOA10 missense variants increased the degradation of the Gly TFT (corrected p < 0.05, Figure 4G, Figure 4—figure supplement 2). In contrast, only the Y1186F variant significantly increased the degradation of the Thr TFT (Figure 4G, Figure 4—figure supplement 2). The multiple causal variants and substrate-specific effects of individual DOA10 variants further highlights the complex effects of variation in E3 ubiquitin ligases on UPS activity.

A QTL on chromosome V detected for 7 of 8 Ac/N-degrons contained UBC6, the E2 ubiquitin-conjugating enzyme of the Ac/N-end pathway. Ubc6 pairs with Doa10 to ubiquitinate substrates of the Ac/N-end pathway (Figure 4I; Chen et al., 1993; Sommer and Jentsch, 1993; Swanson et al., 2001). The RM UBC6 allele contains a 3 base pair deletion in the promoter, one missense variant, and one terminator variant (Figure 4F). Using the threonine (Thr) and alanine (Ala) TFTs, we established that UBC6 is a causal gene for this QTL (corrected p < 0.05, Figure 4K, Figure 4—figure supplement 3). The D229G missense variant altered UPS activity towards the alanine (Ala) and threonine (Thr) Ac/N-degrons (corrected p < 0.05, Figure 4K, Figure 4—figure supplement 3). The UBC6 RM terminator also significantly increased the degradation rate of the Ala, but not Thr TFT, further establishing the substrate-specific effects of genetic variation on UPS activity (Figure 4K, Figure 4—figure supplement 3). Our results with UBC6 show that genetic variation influences UPS activity through effects on substrate ubiquitination by E2 ubiquitin conjugating enzymes, as well as substrate recognition by the E3 ligases Ubr1 and Doa10 described above.

Knowledge of the causal nucleotides in NTA1, DOA10, and UBC6 allowed us to examine their molecular properties, evolutionary histories, and population frequencies. A notable feature of causal missense variants was the distal location of their encoded amino acids relative to the active site of the corresponding protein (Figure 4D/H/L). The amino acids encoded by the NTA1 causal variants occur on the protein’s exterior surface (Figure 4D), while those for the DOA10 and UBC6 causal variants occur in or near transmembrane helices that anchor these proteins to the endoplasmic reticulum membrane (Figure 4H/L). Thus, in addition to effects on ubiquitin system gene expression, such as with UBR1 –469A>T, the molecular basis for the continuous distribution of variant effects on UPS activity may also involve subtle alterations to the stability, localization, or physical interactions of ubiquitin system proteins (Oh et al., 2020).

The BY alleles of the causal DOA10 Y1186F and UBC6 D229G variants are each derived, at low population frequencies (5.1% and 2.2%, respectively), and occur in only 3 non-BY clades (Figure 4—figure supplement 4), similar to the –469A>T UBR1 variant. Given the generally deleterious effects of reduced UPS activity (Pohl and Dikic, 2019; Schwartz and Ciechanover, 1999), these variants may be subject to purifying selection. In contrast, the RM allele of the causal NTA1 E129G variant is derived, common (51.5% population frequency), and found in most clades (Figure 4—figure supplement 4). The derived RM allele of the causal NTA1 E129G variant may have been able to rise to comparatively high population frequency because deleting NTA1 does not decrease competitive fitness (Baker and Varshavsky, 1995).

We examined additional QTLs to nominate candidate causal genes. The most frequently observed UPS QTL was detected for 8 of 8 Ac/N-end and 6 of 12 Arg/N-end TFTs and was located on chromosome XII in the immediate vicinity of a Ty1 insertion in the HAP1 transcription factor in the BY strain (Figure 2B, Figure 2—source data 1; Gaisne et al., 1999). The Ty1 insertion in HAP1 exerts highly pleiotropic effects on gene expression, altering the expression of 3,755 genes (Albert et al., 2018). Similarly, a QTL on chromosome XIV affected 10 of our 20 N-degron TFTs and contained the MKT1 gene (Figure 2B, Figure 2—source data 1). MKT1 encodes a multi-functional RNA binding protein involved in the post-transcriptional regulation of gene expression and is the causal gene for other QTLs previously mapped in the BY/RM cross (Jain et al., 2016; Wickner, 1987; Icho et al., 1986). HAP1 and MKT1 are the likely causal genes for the chromosome XII and XIV QTLs, showing that genetic variation may also shape UPS activity through indirect effects on genes with no known connection to the UPS.

Taken together, our analysis of causal genes and nucleotides illustrates the breadth and diversity of genetic influences on UPS activity. Each fine-mapped causal gene harbored multiple causal variants that may differentially affect distinct UPS substrates. Regulatory and missense variants in ubiquitin system genes that shape the full sequence of molecular events in protein ubiquitination, including substrate processing, recognition, and ubiquitination, alter UPS activity.

Protein-specific effects of UBR1 -469A>T on gene expression

Previous efforts to understand how genetic variation influences gene expression have revealed considerable discrepancies between genetic effects on mRNA versus protein abundance. Many gene expression QTLs alter protein abundance without detectable effects on mRNA levels (Battle et al., 2015; Ghazalpour et al., 2011; Mirauta et al., 2020; Albert et al., 2014; Brion et al., 2020; Chick et al., 2016; Cenik et al., 2015; Abell et al., 2022; Foss et al., 2011). We reasoned that protein-specific gene expression QTLs could arise through effects on UPS protein degradation. To test this idea and explore how variant effects on UPS activity influence other aspects of cellular physiology, we measured global gene expression at the protein and RNA levels in the same cultures of the BY strain and a BY strain engineered to contain the causal –469A>T RM allele in the UBR1 promoter ("BY UBR1 –469A>T"). As expected, the derived BY allele decreased UBR1 protein and RNA levels (Figure 5A/B).

Figure 5. Proteomic and RNA-seq analysis of the effect of the UBR1 –469A>T promoter variant on gene expression.

(A) Protein fold-change versus statistical significance for BY versus BY UBR1 –469A>T for all detected proteins. Differentially abundant proteins are shown in blue. (B) RNA fold-change versus statistical significance for BY versus BY UBR1 –469A>T for all detected transcripts. Differentially expressed transcripts are shown in yellow. (C) Scatterplot comparing changes in protein and RNA abundance caused by UBR1 –469A>T. "LFC" = log2 fold change.

Figure 5—source data 1. Full proteomics results.
elife-79570-fig5-data1.xlsx (153.2KB, xlsx)
Figure 5—source data 2. Full RNA-seq results.
elife-79570-fig5-data2.xlsx (317.8KB, xlsx)

Figure 5.

Figure 5—figure supplement 1. Over-represented GO biological processes and Reactome pathways in the set of differentially expressed transcripts.

Figure 5—figure supplement 1.

Barchart of significantly over-represented Gene Ontology Biological Process and Reactome Pathway terms identified using the list of differentially expressed mRNA transcripts between the wild-type BY strain and BY UBR1 –469A>T. The numbers in the bars denote the observed (left) and expected (right) number of genes for a given process or pathway.

Out of 3,046 proteins quantified by mass spectrometry, 39 proteins were differentially abundant at a 10% FDR (Figure 5A, Figure 5—source data 1). Consistent with the reduced UPS activity conferred by the BY UBR1 allele, a significant majority (28 of 39, 71%, binomial test p = 9.5e-3) of differentially abundant proteins were increased by the BY allele (Figure 5A, Figure 5—source data 1). The median log2 fold change across all proteins was –0.012, while for differentially abundant proteins, the median log2 fold change was 0.37 (Figure 5A, Figure 5—source data 1). No Gene Ontology or Reactome pathway terms were enriched in our set of differentially abundant proteins. This result is consistent with recent observations that sequence features, rather than biological function or subcellular localization are the primary determinants of substrate targeting by E3 ligases. (Kong et al., 2021; Christiano et al., 2020).

To determine whether differences in protein abundance were reflected at the mRNA level, we used RNA-seq to quantify the levels of 5,675 transcripts. A total of 78 transcripts were differentially expressed between BY and BY UBR1 –469A>T at a 10% FDR (Figure 5B, Figure 5—source data 2). Only three genes, UBR1, HSP26, and TMA10, showed significant and concordant changes at the RNA and protein levels (Figure 5C) and the overall correlation of log2 fold changes at the protein and mRNA levels was low, albeit significant (Pearson r = 0.064, p = 7.2e-4). In contrast to our proteomics results, the BY allele tended to decrease mRNA abundance, causing lower expression at a significant majority of differentially expressed genes (67 / 78, 86%, binomial test p = 2e-6, Figure 5B, Figure 5—source data 2). The median log2 fold change across all transcripts was 0.0024, while at differentially expressed transcripts it was –0.20 (Figure 5B, Figure 5—source data 2). Multiple GO proteostasis-related pathways were enriched among the differentially abundant transcripts (Figure 5—figure supplement 1), driven by the decreased transcript abundance for genes such as UBR1 and the chaperones HSP26, HSP30, HSP31, and HSP82 in BY. Our results add to the emerging view of complex, protein-specific influences of genetic variation on gene expression (Battle et al., 2015; Ghazalpour et al., 2011; Mirauta et al., 2020; Albert et al., 2014; Brion et al., 2020; Chick et al., 2016; Cenik et al., 2015; Abell et al., 2022; Foss et al., 2011). Specifically, a non-coding variant that decreases expression of a single E3 ubiquitin ligase increases the levels of dozens of proteins without detectable effects on transcript abundance, implicating genetic effects on UPS activity as a potentially prominent source of post-translational variation in gene expression.

Discussion

Protein degradation by the UPS is an essential biological process that influences virtually all aspects of eukaryotic cellular physiology (Hanna and Finley, 2007; Varshavsky, 2011; Schwartz and Ciechanover, 1999; Finley and Prado, 2020). Understanding the sources of variation in UPS activity thus has considerable implications for our understanding of numerous cellular and organismal traits, including human health and disease (Schmidt and Finley, 2014; Petrucelli and Dawson, 2004; Gomes, 2013; Schwartz and Ciechanover, 1999). Our statistically powerful, systematic genetic mapping of the N-end rule has revealed that individual genetic differences create heritable variation in UPS protein degradation. Genetic effects on UPS activity are numerous and comprise a continuous distribution of many loci with small effects and few loci of large effect (Figure 2), similar to other complex traits (Mackay et al., 2009; Ehrenreich et al., 2009). Previous efforts to understand how individual genetic differences cause variation in UPS activity have focused on individual disease-causing mutations in UPS genes (Gomes, 2013; Agarwal et al., 2010; Zenker et al., 2006; Deng et al., 2011; Kröll-Hermi et al., 2020). Our results show that these large-effect mutations in UPS genes sit atop one extreme of a continuous distribution of variant effects that is dominated by many loci of small effect. Aberrant UPS activity is a hallmark of many common diseases with a poorly-understood, complex genetic basis (Schmidt and Finley, 2014; Petrucelli and Dawson, 2004; Zheng et al., 2014). Our results raise the possibility that the effects of many common, small-effect alleles may contribute to the risk of these diseases through their effects on UPS activity.

Using genome engineering, we experimentally identified causal regulatory and missense variants in four functionally distinct ubiquitin system genes. A major function of the ubiquitin system is conferring specificity to UPS protein degradation (Komander and Rape, 2012; Johnson et al., 1992; Bett, 2016). Non-ubiquitinated proteins are blocked by the proteasome’s 19S regulatory particle from degradation by the 20S catalytic core (Inobe and Matouschek, 2014). The selective binding of ubiquitinated substrates by the 19S regulatory particle ensures that only proteins targeted for degradation enter the proteasome. The activity of the ubiquitin system towards distinct substrates is highly variable, even for proteins degraded by the same UPS pathway (Bachmair et al., 1986; Kats et al., 2018; Christiano et al., 2020). Consistent with these observations, the effects of causal ubiquitin system gene variants were highly substrate-specific (Figures 3 and 4). Our results raise the question of whether UPS protein degradation is also shaped by variation in proteasome genes and whether any such effects would be less substrate-specific than those in the ubiquitin system. Given the multiple QTLs arising from ubiquitin system genes, detecting genetic influences on proteasome activity may benefit from assays that can measure proteasome activity independently of the ubiquitin system.

The remarkable complexity in causal variants we uncovered underscores the challenge of predicting variant effects on UPS protein degradation. Similar to recent results (Lutz et al., 2022; Abell et al., 2022), each of the four QTL regions we fine-mapped contained multiple causal variants in a single gene (Figures 3 and 4). In the case of NTA1, we observed that the effect of the D111E variant was likely masked during QTL mapping by the larger effect of the E129G variant (Figure 4C), highlighting the need to test individual variants in isolation. Causal variants may also exert substrate-specific effects on UPS protein degradation. We observed multiple instances where the magnitude of a causal variant’s effect varied between substrates. In the case of UBR1, the RM UBR1 ORF exerted significant, discrepant effects on the degradation of Arg/N-degrons (Figure 3C). Recent efforts have established that a protein’s sequence critically determines how its degradation is altered by changes in UPS activity (Christiano et al., 2020). Thus, a complete understanding of a given variant’s influence on UPS protein degradation will require testing its effect on the turnover of multiple substrates with diverse sequence compositions.

Our results suggest that genetic effects on UPS activity are an important source of post-translational variation in gene expression. A promoter variant that reduces UPS activity by decreasing UBR1 expression alters the abundance of dozens of proteins without detectable effects on levels of the corresponding mRNA transcripts (Figure 5). Ubr1 and Doa10 target distinct sets of cellular proteins (Kats et al., 2018; Kong et al., 2021; Christiano et al., 2020). Their genes each contain multiple causal variants that differentially affected individual N-degrons and thus, potentially, endogenous cellular proteins. Similar effects arising from variation in the approximately 100 E3 ubiquitin ligases encoded in the S. cerevisiae genome (Finley et al., 2012) may help explain the numerous protein-specific gene expression QTLs (Battle et al., 2015; Brion et al., 2020; Albert et al., 2014; Mirauta et al., 2020). Such effects could be even more prevalent in the human genome, which encodes an estimated 600 E3 ubiquitin ligases (Li et al., 2008).

We have developed a generalizable framework for mapping genetic influences on protein degradation. Our results lay important groundwork for future efforts to understand how heritable differences in UPS activity contribute to variation in complex cellular and organismal traits, including the many diseases marked by aberrant UPS activity.

Materials and methods

Key resources table.

Reagent type (species) or resource Designation Source or reference Identifiers Additional information
Gene (Saccharomyces cerevisiae) UBR1 Saccharomyces Genome Database (SGD) YGR184C edited to contain
alternative alleles / variants
Gene (S. cerevisiae) DOA10 SGD YIL030C edited to contain
alternative alleles / variants
Gene (S. cerevisiae) UBC6 SGD YER100W edited to contain
alternative alleles / variants
Gene (S. cerevisiae) NTA1 SGD YJR062C edited to contain
alternative alleles / variants
Gene (S. cerevisiae) HIS3 SGD YOR202W selectable marker for genome engineering
Gene (S. cerevisiae) LYP1 SGD YNL268W selectable marker for genome engineering
Strain, strain
background (S. cerevisiae)
BY4741 Leonid Kruglyak YFA0040 Supplementary file 5
Strain, strain
background (S. cerevisiae)
RM11.1a Leonid Kruglyak YFA0039 Supplementary file 5
Strain, strain
background (S. cerevisiae)
recombinant progeny of BY4741 x RM11.1a this study SFA- Supplementary file 5
Strain, strain
background (S. cerevisiae)
strains with tandem fluorescent timer reporters this study YFA- Supplementary file 5
Strain, strain
background (S. cerevisiae)
strains lacking individual ubiquitin-proteasome system genes this study YFA- Supplementary file 5
Strain, strain
background (S. cerevisiae)
strains with alternative UPS gene alleles / variants this study YFA- Supplementary file 5
Strain, strain
background (Escherichia coli)
DH5α New England Biolabs for plasmid cloning and propagation
Recombinant DNA reagent 23 plasmids this study PFA- Supplementary file 4
Recombinant DNA reagent backbone plasmid Addgene 35121
Recombinant DNA reagent backbone plasmid Addgene 41030
Recombinant DNA reagent KanMX cassette Wach et al., 1994;
10.1002/yea.320101310
selectable marker for genome engineering
Recombinant DNA reagent NatMX cassette Wach et al., 1994;
10.1002/yea.320101310
selectable marker for genome engineering
Sequence-based reagent 102 oligonucleotides Integrated DNA
Techologies
OFA- Supplementary file 3
Commercial assay or kit Nextera DNA
Library Prep Kit
Illumina FC-121–
1030
Commercial assay or kit EB Ultra II
Directional RNA
library kit for Illumina
New England Biolabs E7760
Commercial assay or kit Monarch Gel
Extraction kit
New England Biolabs T1010L
Commercial assay or kit HiFi DNA Assembly
Cloning Kit
New England Biolabs E5520S
Commercial assay or kit TMT10plex Isobaric
Label Reagent Set
ThermoFisher Scientific 90110
Commercial assay or kit ZR Fungal / Bacterial
RNA Miniprep kit
Zymo Research R2014
Commercial assay or kit Quick-96 DNA Plus kit Zymo Research
10.1093/bioinformatics/btp324
D4070
Software, algorithm MULTIPOOL Edwards and Gifford, 2012;
10.1186/1471-2105-13-S6-S8
Software, algorithm trimmomatic Bolger et al., 2014
10.1093/bioinformatics/btu170
Software, algorithm kallisto Bray et al., 2016;
10.1038/nbt.3519
Software, algorithm PANTHER Mi et al., 2021;
10.1093/nar/gkaa1106
Software, algorithm fastp Chen et al., 2018;
10.1093/bioinformatics/bty560
Software, algorithm RSeQC Wang et al., 2012;
10.1093/bioinformatics/bts356
Software, algorithm Scaffold https://www.proteomesoftware.com/
Software, algorithm Proteome Discoverer Thermo Scientific
Software, algorithm AlphaFold Jumper et al., 2021;
10.1038/s41586-021-03819-2
Software, algorithm Inkscape https://inkscape.org
Other LSR II Flow Cytometer BD flow cytometry
Other FACSAria II Cell Sorter BD cell sorting
Other Orbitrap Fusion Tribrid
MS-MS instrument
Thermo Scientific mass spectrometry
Other Next-Seq 550 Illumina DNA / RNAsequencing

Tandem fluorescent timer ubiquitin-proteasome system activity reporters

We used tandem fluorescent timers (TFTs) to measure ubiquitin-proteasome system (UPS) activity. TFTs are fusions of two fluorescent proteins (FPs) with distinct spectral profiles and maturation kinetics (Khmelinskii et al., 2012; Khmelinskii and Knop, 2014). In the most common implementation, a TFT consists of a faster maturing green fluorescent protein (GFP) and a slower maturing red fluorescent protein (RFP). Because the FPs in the TFT mature at different rates, the RFP / GFP ratio changes over time. If the degradation rate of a TFT exceeds the maturation rate of the RFP, the -log2 RFP / GFP ratio is directly proportional to the construct’s degradation rate (Khmelinskii and Knop, 2014; Khmelinskii et al., 2012). When fused to N-degrons, the TFT’s RFP / GFP ratio measures UPS N-end rule activity (Khmelinskii et al., 2012; Khmelinskii et al., 2014). The RFP / GFP ratio is also independent of the TFT’s expression level, (Khmelinskii et al., 2012; Khmelinskii and Knop, 2014; Kong et al., 2021) preventing confounding from genetic effects on reporter expression in genetically diverse cell populations.

We used fluorescent proteins from previously characterized TFTs in our experiments (Khmelinskii et al., 2016; Khmelinskii and Knop, 2014; Khmelinskii et al., 2012; Khmelinskii et al., 2014). superfolder GFP (Pédelacq et al., 2006) (sfGFP) was used as the faster maturing FP in all TFTs. sfGFP matures in approximately 5 min and has excitation and emission maximums of 485 nm and 510 nm, respectively (Pédelacq et al., 2006). The slower maturing FP in each TFT was either mCherry or mRuby. mCherry matures in approximately 40 min and has excitation and emission maximums of 587 nm and 610 nm, respectively (Shaner et al., 2004). mRuby matures in approximately 170 min and has excitation and emission maximums of 558 nm and 605 nm, respectively (Kredel et al., 2009). All TFT fluorescent proteins are monomeric. We separated green and red FPs in each TFT with an unstructured 35 amino acid linker sequence to minimize fluorescence resonance energy transfer (Khmelinskii et al., 2012).

Construction of Arg/N-end and Ac/N-end pathway TFTs

To generate TFT constructs with defined N-terminal amino acids, we used the ubiquitin-fusion technique (Bachmair et al., 1986; Varshavsky, 2005; Varshavsky, 2011), which involves placing a ubiquitin moiety immediately upstream of a sequence encoding the desired N-degron. During translation, ubiquitin-hydrolases cleave the ubiquitin moiety, exposing the N-degron (Figure 1A). We synthesized DNA (Integrated DNA Technologies [IDT], Coralville, Iowa, USA) encoding the Saccharomyces cerevisiae ubiquitin sequence and a peptide linker sequence derived from Escherichia coli ß-galactosidase previously used to identify components of the Arg/N-end and Ac/N-end pathways (Bachmair et al., 1986). The peptide linker sequence is unstructured and contains internal lysine residues required for ubiquitination and degradation by the UPS (Bachmair et al., 1986; Hwang et al., 2010). Peptide linkers encoding the 20 possible N-terminal amino acids were made by PCR amplifying the linker sequence using oligonucleotides encoding each unique N-terminal amino acid (Supplementary file 3).

We then devised a general strategy to assemble TFT-containing plasmids with defined N-terminal amino acids (Figure 1—figure supplement 2). We first obtained sequences encoding each reporter element by PCR or DNA synthesis. We codon-optimized the sfGFP, mCherry, mRuby, and the TFT linker sequences for expression in S. cerevisiae using the Java Codon Adaptation Tool (JCaT) (Grote et al., 2005) and synthesized DNA fragments encoding each sequence (IDT). We used the TDH3 promoter to drive expression of each TFT reporter. The TDH3 promoter was PCR-amplified from Addgene plasmid #67639 (a gift from John Wyrick). We used the ADH1 terminator in all TFT constructs, which we PCR amplified from Addgene plasmid #67639. We used the KanMX cassette (Wach et al., 1994), which confers resistance to G418, as the selection module for all TFT constructs and obtained this sequence by PCR amplification from Addgene plasmid #41030 (a gift from Michael Boddy). Thus, each construct has the general structure of TDH3 promoter, N-degron, linker sequence, TFT, ADH1 terminator, and the KanMX resistance cassette (Figure 1—figure supplement 2). Based on the half-lives of N-degrons (Bachmair et al., 1986; Hwang et al., 2010; Varshavsky, 2011), we used the mCherry-sfGFP TFT for all Arg/N-end constructs and the mRuby-sfGFP TFT for all Ac/N-end constructs.

We used Addgene plasmid #35121 (a gift from John McCusker) to construct all TFT plasmids. Digesting this plasmid with BamHI and EcoRV restriction enzymes produces a 2,451 bp fragment that we used as a vector backbone for TFT plasmid assembly. We obtained a DNA fragment containing 734 bp of sequence upstream of the LYP1 start codon, a SwaI restriction site, and 380 bp of sequence downstream of the LYP1 stop codon by DNA synthesis (IDT). We performed isothermal assembly cloning using the New England Biolabs (NEB; Ipswich, MA, USA) HiFi Assembly Cloning Kit (NEB) to insert the LYP1 homology sequence into the BamHI/EcoRV digest of Addgene plasmid #35121 to create the final backbone plasmid BFA0190 (Supplementary file 4). We then combined SwaI digested BFA0190 and the components of each TFT reporter and used the NEB HiFi Assembly Kit (NEB) to produce each TFT plasmid. The 5’ and 3’ LYP1 sequences in each TFT contain naturally-occurring SacI and BglII restriction sites, respectively. We digested each TFT plasmid with SacI and BglII (NEB) to obtain a linear DNA transformation fragment (Figure 1—figure supplement 2). The flanking LYP1 homology and KanMX module in each TFT construct allows selection for reporter integration at the LYP1 locus using G418 (Goldstein and McCusker, 1999) and the toxic amino acid analogue thialysine (S-(2-aminoethyl)-L-cysteine hydrochloride) (Zwolshen and Bhattacharjee, 1981; Baryshnikova et al., 2010; Kuzmin et al., 2016). The sequence identity of all assembled plasmids was verified by Sanger sequencing. The full list of plasmids used in this study is found in Supplementary file 4.

Yeast strain handling

We used two strains of the yeast Saccharomyces cerevisiae to characterize our TFT reporters and perform genetic mapping of UPS activity. The haploid BY strain (genotype: MATa his3Δ hoΔ) is closely related to the S. cerevisiae S288C laboratory strain. The second mapping strain, RM, was originally isolated from a California vineyard and is haploid with genotype MATa can1Δ::STE2pr-SpHIS5 his3Δ::NatMX AMN1-BY hoΔ::HphMX URA3-FY. BY and RM differ at 1 nucleotide per 200 base pairs on average, such that approximately 45,000 single nucleotide variants (SNVs) between the strains can serve as markers in a genetic mapping experiment (Albert et al., 2014; Brion et al., 2020; Ehrenreich et al., 2009; Ehrenreich et al., 2010).

We built additional strains for characterizing our UPS activity reporters by deleting individual UPS genes from the BY strain. Each deletion strain was constructed by replacing the targeted gene with the NatMX cassette (Goldstein and McCusker, 1999), which confers resistance to the antibiotic nourseothricin. We PCR amplified the NatMX cassette from Addgene plasmid #35121 using primers with homology to the 5’ upstream and 3’ downstream sequences of the targeted gene. The oligonucleotides for each gene deletion cassette amplification are listed in Supplementary file 3. We created a BY strain lacking the UBR1 gene, which encodes the Arg/N-end pathway E3 ligase Ubr1. We refer to this strain hereafter as ‘BY ubr1Δ’. We created a BY strain (‘BY doa10Δ’) lacking the DOA10 gene that encodes the Ac/N-end pathway E3 ligase Doa10. Finally, we created a BY strain (‘BY rpn4Δ’) lacking the RPN4 that encodes the proteasome transcription factor Rpn4. Table 1 lists these strains and their full genotypes. Supplementary file 5 contains the complete list of strains used in this study.

Table 1. Strain genotypes.

Short Name Genotype Antibiotic Resistance Auxotrophies
BY MATa his3Δ hoΔ histidine
RM MATα can1Δ::STE2pr-SpHIS5 clonNAT, hygromycin histidine
his3Δ::NatMX hoΔ::HphMX
BY rpn4Δ MATa his3Δ hoΔ rpn4Δ::NatMX clonNAT histidine
BY ubr1Δ MATa his3Δ hoΔ ubr1Δ::NatMX clonNAT histidine
BY doa10Δ MATa his3Δ hoΔ doa10Δ::NatMX clonNAT histidine

Table 2 describes the media formulations used for all experiments. Synthetic complete amino acid powders (SC -lys and SC -his -lys -ura) were obtained from Sunrise Science (Knoxville, TN, USA). Where indicated, we added the following reagents at the indicated concentrations to yeast media: G418, 200 mg/mL (Fisher Scientific, Pittsburgh, PA, USA); clonNAT (nourseothricin sulfate, Fisher Scientific), 50 mg/L; thialysine (S-(2-aminoethyl)-L-cysteine hydrochloride; MilliporeSigma, St. Louis, MO, USA), 50 mg/L; canavanine (L-canavanine sulfate, MilliporeSigma), 50 mg/L.

Table 2. Media formulations.

Media Name Abbreviation Formulation
Yeast-Peptone-Dextrose YPD 10 g/L yeast extract
20 g/L peptone
20 g/L dextrose
Synthetic Complete SC 6.7 g/L yeast nitrogen base
1.96 g/L amino acid mix -lys
20 g/L dextrose
Haploid Selection SGA 6.7 g/L yeast nitrogen base
1.74 g/L amino acid mix -his -lys -ura
20 g/L dextrose
Sporulation SPO 1 g/L yeast extract
10 g/L potassium acetate
0.5 g/L dextrose

Yeast transformation

We used a standard yeast transformation protocol to construct reporter control strains and build strains with UPS activity reporters (Gietz and Schiestl, 2007). In brief, we inoculated yeast strains growing on solid YPD medium into 5 mL of YPD liquid medium for overnight growth at 30°C. The following morning, we diluted 1 mL of saturated culture into 50 mL of fresh YPD and grew the cells for 4 hr. The cells were then successively washed in sterile ultrapure water and transformation solution 1 (10 mM Tris HCl [pH 8.0], 1 mM EDTA [pH 8.0], and 0.1 M lithium acetate). At each step, we pelleted the cells by centrifugation at 3000 rpm for 2 min in a benchtop centrifuge and discarded the supernatant. The cells were suspended in 100 μL of transformation solution 1 along with 50 μg of salmon sperm carrier DNA and 300 ng of transforming DNA. The cells were incubated at 30 for 30 min and 700 μL of transformation solution 2 (10 mM Tris HCl [pH 8.0], 1 mM EDTA [pH 8.0], and 0.1 M lithium acetate in 40% polyethylene glycol [PEG]) was added to each tube, followed by a 30-min heat shock at 42°C. We then washed the transformed cells in sterile, ultrapure water. We added 1 mL of liquid YPD medium to each tube and incubated the tubes for 90 min with rolling at 30°C to allow for expression of the antibiotic resistance cassettes. After washing with sterile, ultrapure water, we plated 200 μL of cells on solid SC -lys medium with G418 and thialysine, and, for strains with the NatMX cassette, clonNAT. For each strain, we streaked multiple independent colonies (biological replicates) from the transformation plate for further analysis as indicated in the text. We verified reporter integration at the targeted genomic locus by colony PCR (Ward, 1992). The primers used for these experiments are listed in Supplementary file 3.

Yeast mating and segregant populations

We created populations of genetically variable, recombinant cells ("segregants") for genetic mapping using a modified synthetic genetic array (SGA) approach (Baryshnikova et al., 2010; Kuzmin et al., 2016). We first mated BY strains with a given UPS activity reporter to RM by mixing freshly streaked cells of each strain on solid YPD medium. For each UPS activity reporter, we mated two independently-derived clones (biological replicates) to the RM strain. Cells were grown overnight at 30°C and we selected for diploid cells (successful BY-RM matings) by streaking mated cells onto solid YPD medium with G418 (which selects for the KanMX cassette in the TFT in the BY strain) and clonNAT (which selects for the NatMX cassette in the RM strain). We inoculated 5 mL of YPD with freshly streaked diploid cells for overnight growth at 30°C. The next day, we pelleted the cultures, washed them with sterile, ultrapure water, and resuspended the cells in 5 mL of SPO liquid medium (Table 2). We sporulated the cells by incubating them at room temperature with rolling for 9 days. After confirming sporulation by brightfield microscopy, we pelleted 2 mL of culture, washed cells with 1 mL of sterile, ultrapure water, and resuspended cells in 300 μL of 1 M sorbitol containing 3 U of Zymolyase lytic enzyme (United States Biological, Salem, MA, USA) to degrade ascal walls. Digestions were carried out at 30°C with rolling for 2 hr. We then washed the spores with 1 mL of 1 M sorbitol, vortexed for 1 min at the highest intensity setting, resuspended the cells in sterile ultrapure water, and confirmed the release of cells from ascii by brightfield microscopy. We plated 300 μl of cells onto solid SGA medium containing G418 and canavanine. This media formulation selects for haploid cells with (1) a UPS activity reporter via G418, (2) the MATa mating type via the Schizosaccharomyces pombe HIS5 gene under the control of the STE2 promoter (which is only active in MATa cells), and (3) replacement of the CAN1 gene with S. pombe HIS5 via the toxic arginine analog canavanine (Baryshnikova et al., 2010; Kuzmin et al., 2016). Haploid segregant populations were grown for 2 days at 30°C and harvested by adding 10 mL of sterile, ultrapure water and scraping the cells from each plate. We pelleted each cell suspension by centrifugation at 3000 rpm for 10 min and resuspended the cells in 1 mL of SGA medium. We added 450 μL of 40% (v/v) sterile glycerol solution to 750 μL of segregant culture and stored samples in screw cap cryovials at -80°C. We stored two independent sporulations of each reporter (derived from our initial matings) as independent biological replicates.

Flow cytometry

We measured UPS activity by flow cytometry as follows. Yeast strains were manually inoculated into 400 μL of liquid SC -lys medium with G418 and grown overnight in 2 mL 96-well plates at 30°C with 1000 rpm mixing using a MixMate (Eppendorf, Hamburg, Germany). The following morning, we inoculated a fresh 400 μL of G418-containing SC -lys media with 4 μL of each saturated culture. Cells were grown for an additional 3 hr prior to analysis by flow cytometry. All flow cytometry experiments were performed on an LSR II flow cytometer (BD, Franklin Lakes, NJ, USA) equipped with a 20 mW 488 nm laser with 488/10 and 525/50 filters for measuring forward/side scatter and sfGFP, respectively, as well as a 40 mW 561 nm laser and a 610/20 filter for measuring mCherry and mRuby. Table 3 lists the parameters and settings that were used for all flow cytometry and fluorescence-activated cell sorting (FACS) experiments. We recorded 10,000 cells each from 8 independent biological replicates per strain for our analyses of BY, RM, and reporter control strains.

Table 3. Flow cytometry and FACS settings.

Parameter Laser Line (nm) Laser Setting (V) Filter
forward scatter (FSC) 488 500 488/10
side scatter (SSC) 488 275 488/10
sfGFP 488 500 525/50
mCherry 561 615 610/20
mRuby 561 615 610/20

We analyzed flow cytometry data using R (R Foundation for Statistical Computing, Vienna Austria) and the flowCore R package (Hahne et al., 2009). We first filtered each flow cytometry dataset to include only those cells within 10% ± the forward scatter (a proxy for cell size) median. We empirically determined that this gating approach captured the central peak of cells in the FSC histogram. It also removed cellular debris, aggregates of multiple cells, and restricted our analyses to cells of the same approximate size. We observed that the TFT’s output changed with the passage of time during flow cytometry experiments. We used the residuals of a loess regression of the TFT’s output on time to correct for this effect, similar to a previously-described approach (Brion et al., 2020).

To characterize our TFT reporters, we used the following analysis steps. We extracted the median -log2 RFP / GFP ratio from each of 10,000 cells per strain per reporter. These values were Z-score normalized relative to the sample lowest degradation rate (typically the E3 ligase deletion strain). Following this transformation, the strain with lowest degradation rate has a degradation rate of approximately 0 and the now-scaled RFP / GFP ratio is directly proportional to the construct’s degradation rate. To compare degradation rates between strains and individual UPS activity reporters, we then converted scaled RFP/GFP ratios to Z scores, which we report as "Normalized UPS Activity". Statistical significance was assessed using a one-way ANOVA with Tukey’s HSD post-hoc test.

For fine-mapping causal genes and variants for UPS activity QTLs, we used the following approach. We extracted the median -log2 RFP / GFP ratio from each of 10,000 cells per strain per reporter. These values were Z-score normalized relative to the median of the control strain (a BY strain engineered to contain the BY allele of a candidate causal gene). Statistical significance was assessed using a t-test of each experimental strain versus the control strain with Benjamini-Hochberg correction for multiple testing (Benjamini and Hochberg, 1995).

Fluorescence-activated cell sorting

We selected populations of segregants for QTL mapping using a previously described approach for isolating phenotypically extreme cell populations by FACS (Albert et al., 2014; Brion et al., 2020). Segregant populations were thawed approximately 16 hr prior to cell sorting and grown overnight in 5 mL of SGA medium containing G418 and canavanine. The following morning, 1 mL of cells from each segregant population was diluted into a fresh 4 mL of SGA medium containing G418 and canavanine. Segregant cultures were then grown for an additional 4 hours prior to sorting. All FACS experiments were carried out using a FACSAria II cell sorter (BD). We used plots of side scatter (SSC) height by SSC width and forward scatter (FSC) height by FSC width to remove doublets from each sample. We then filtered cells on the basis of FSC area, restricting our sorts to ±7.5% of the central FSC peak, which we empirically determined excluded cellular debris and aggregates while encompassing the primary haploid cell population. Finally, we defined a fluorescence-positive population by comparing each segregant population to negative control BY and RM strains without TFTs. We collected pools of 20,000 cells each from three gates drawn on each segregant population:

  1. The 2% lower tail of the UPS activity distribution

  2. The 2% upper tail of the UPS activity distribution

  3. Fluorescence-positive cells without selection on UPS activity (“null pools”), which were used to determine the false positive rate of the QTL mapping method (see below)

We collected cell pools from two independent biological replicates (spore preparations) for each reporter. Each pool of 20,000 cells was collected into sterile 1.5 mL polypropylene tubes containing 1 mL of SGA medium and grown overnight at 30°C with rolling. The next day, we mixed 750 μL of cells with 450 μL of 40% (v/v) glycerol and stored this mixture in 2 mL 96-well plates at −80°C.

Genomic DNA isolation and library preparation

We extracted genomic DNA from sorted segregant pools for whole-genome sequencing. Deep-well plates containing glycerol stocks of sorted segregant pools were thawed and 800 μL of each sample was pelleted by centrifugation at 3,700 rpm for 10 min. We discarded the supernatant and resuspended cell pellets in 800 μL of a 1 M sorbitol solution containing 0.1 M EDTA, 14.3 mM β-mercaptoethanol, and 500 U of Zymolyase lytic enzyme to digest cell walls prior to DNA extraction. The digestion reaction was carried out by resuspending cell pellets with mixing at 1,000 rpm for 2 min followed by incubation for 2 hr at 37°C. When the digestion reaction finished, we discarded the supernatant, resuspended cells in 50 μL of phosphate buffered saline, and used the Quick-DNA 96 Plus kit (Zymo Research, Irvine, CA, USA) to extract genomic DNA. We followed the manufacturer’s protocol to extract genomic DNA with the following modifications. We incubated cells in a 20 mg/mL proteinase K solution overnight with incubation at 55°C. After completing the DNA extraction protocol, we eluted DNA using 40 μL of DNA elution buffer (10 mM Tris-HCl [pH 8.5], 0.1 mM EDTA). The DNA concentration for each sample was determined using the Qubit dsDNA BR assay kit (Thermo Fisher Scientific, Waltham, MA, USA) in a 96 well format using a Synergy H1 plate reader (BioTek Instruments, Winooski, VT, USA).

We used a previously-described approach to prepare libraries for short-read whole-genome sequencing on the Illumina Next-Seq platform (Albert et al., 2014; Brion et al., 2020). We used the Nextera DNA library kit (Illumina, San Diego, CA, USA) according to the manufacturer’s instructions with the following modifications. For the tagmentation reaction, 5 ng of genomic DNA from each sample was diluted in a master mix containing 4 μL of Tagment DNA buffer, 1 μL of sterile molecular biology grade water, and 5 μL of Tagment DNA enzyme diluted 1:20 in Tagment DNA buffer. The tagmentation reaction was run on a SimpliAmp thermal cycler (Thermo Fisher Scientific) using the following parameters: 55°C temperature, 20 μL reaction volume, 10 min incubation. To prepare libraries for sequencing, we added 10 μL of the tagmentation reaction to a master mix containing 1 μL of an Illumina i5 and i7 index primer pair mixture, 0.375 μL of ExTaq polymerase (Takara Bio, Mountain View, CA, USA), 5 μL of ExTaq buffer, 4 μL of a dNTP mixture, and 29.625 μL of sterile molecular biology grade water. We generated all 96 possible index oligo combinations using 8 i5 and 12 i7 index primers. The library amplification reaction was run on a SimpliAmp thermal cycler with the following parameters: initial denaturation at 95°C for 30 s, then 17 cycles of 95°C for 10 s (denaturation), 62°C for 30 s (annealing), and 72°C for 3 min (extension). We quantified the DNA concentration of each reaction using the Qubit dsDNA BR assay kit (Thermo Fisher Scientific) and pooled 10 μL of each reaction. This pooled mixture was run on a 2% agarose gel and we extracted and purified DNA in the 400 bp to 600 bp region using the Monarch Gel Extraction Kit (NEB) according to the manufacturer’s instructions.

Whole-genome sequencing

We submitted pooled, purified DNA libraries to the University of Minnesota Genomics Center (UMGC) for Illumina sequencing. Prior to sequencing, UMGC staff performed three quality control (QC) assays. Library concentration was determined using the PicoGreen dsDNA quantification reagent (Thermo Fisher Scientific) with libraries at a concentration of 1 ng/μL passing QC. Library size was determined using the Tapestation electrophoresis system (Agilent Technologies, Santa Clara, CA, USA) with libraries in the range of 200–700 bp passing QC. Library functionality was determined using the KAPA DNA Library Quantification kit (Roche, Penzberg, Germany), with libraries with a concentration greater than 2 nM passing. All submitted libraries passed each QC assay. We submitted 7 libraries for sequencing at different times. Libraries were sequenced on a NextSeq 550 instrument (Illumina). Depending on the number of samples, we used the following output settings. For libraries with 70 or more samples (2 libraries), 75 bp paired end sequencing was performed in high-output mode to generate approximately 360 × 106 reads. For libraries with 50 or fewer samples (5 libraries), 75 bp paired end sequencing was performed in mid-output mode to generate approximately 120 × 106 reads. Average read coverage of the genome ranged from 9 to 35 with a median coverage of 28 across all libraries. Sequence data de-multiplexing was performed by UMGC. Whole-genome sequencing data have been deposited into the NIH Sequence Read Archive under Bioproject accession PRJNA881749.

Raw whole-genome sequencing data processing

We calculated allele frequencies from our whole-genome sequencing data using the following pipeline. We initially filtered reads to include only those reads with mapping quality scores greater than 30. We aligned the filtered reads to the S. cerevisiae reference genome (version sacCer3) using BWA (Li and Durbin, 2009a) (command: ‘mem -t 24’). We then used samtools (Li et al., 2009b) to remove unaligned reads, non-uniquely aligned reads, and PCR duplicates (command: ‘samtools rmdup -S’). Finally, we produced vcf files containing coverage and allelic read counts at each of 18,871 high-confidence, reliable SNPs (Bloom et al., 2013; Ehrenreich et al., 2010) (command: ‘samtools mpileup -vu -t INFO/AD -l’). Because the BY strain is closely related to the S288C genome reference S. cerevisiae strain, we considered BY alleles reference and RM alleles alternative alleles.

QTL mapping

We identified QTLs from sequence data following established procedures for bulk segregant analysis (Ehrenreich et al., 2010; Albert et al., 2014; Brion et al., 2020). Allele counts in the vcf files generated above were provided to the MULTIPOOL algorithm (Edwards and Gifford, 2012). MULTIPOOL computes logarithm of the odds (LOD) scores by comparing two models: (1) a model in which the high and low UPS activity pools come from one from common population and thus share the same frequency of the BY and RM allele, and (2) a model in which these pools come from two populations with two different allele frequencies, indicating the presence of a QTL. We identified QTLs as genomic regions exceeding an empirically-derived significance threshold (see below). We used MULTIPOOL with the following settings: bp per centiMorgan = 2,200, bin size = 100 bp, effective pool size = 1,000. As in previous QTL mapping in the BY/RM cross by bulk segregant analysis (Albert et al., 2014; Brion et al., 2020), we excluded variants with allele frequencies higher than 0.9 or lower than 0.1 (Albert et al., 2014; Brion et al., 2020). We also used MULTIPOOL to estimate confidence intervals for each significant QTL, which we defined as a 2-LOD drop from the QTL peak position. To visualize QTLs and gauge their effects, we also computed the RM allele frequency differences (ΔAF) at each site between our high and low UPS activity pools. Because allele frequencies are affected by random counting noise, we used loess regression to smooth the allele frequency for each sample before computing ΔAF. We used the smoothed values to plot the ΔAF distribution along the genome and as a measure of QTL effect size.

Null sorts and empirical false discovery rate estimation

We used "null" segregant pools (fluorescence-positive cells with no selection on UPS activity) to empirically estimate the false discovery rate (FDR) of our QTL mapping method. Because these cells are obtained as two pools from the same null population in the same sample, any ΔAF differences between them are the result of technical noise or random variation. We permuted these null comparisons across segregant pools with the same UPS activity reporter for a total of 112 null comparisons. We define the "null QTL rate" at a given LOD threshold as the number of QTLs that exceeded the threshold in these comparisons divided by the number of null comparisons. To determine the FDR for a given LOD score, we then determined the number of QTLs for our experimental comparisons (high UPS activity versus low UPS activity). We define the "experimental QTL rate" as the number of experimental QTLs divided by the number of experimental comparisons. The FDR is thus computed as follows:

nullQTLrate=n.nullQTLsn.nullcomparisons
experimentalQTLrate=n.experimentalQTLsn.experimentalcomparisons
FDR=nullQTLrateexperimentalQTLrate

We evaluated the FDR over a LOD range of 2.5–10 in 0.5 LOD increments. We found that a LOD value of 4.5 led to a null QTL rate of 0.0625 and an FDR of 0.507%. We used this value as our significance threshold for QTL mapping and further filtered our QTL list by excluding QTLs that were not detected in each of two independent biological replicates. Replicating QTLs were defined as those whose peaks were within 100 kb of each other on the same chromosome with the same direction (positive or negative) of RM allele frequency difference between high and low UPS activity pools.

QTL fine-mapping by allelic engineering

We used ‘CRISPR-Swap’ (Lutz et al., 2019), a two-step method for scarless allelic editing, to fine-map QTLs to the level of their causal genes and nucleotides. In the first step of CRISPR-Swap, a gene of interest (GOI) is deleted and replaced with a selectable marker. In the second step, cells are co-transformed with (1) a plasmid that expresses CRISPR-cas9 and a guide RNA targeting the selectable marker and (2) a repair template encoding the desired allele of the GOI.

We used CRISPR-Swap to generate BY strains harboring either RM alleles or chimeric BY/RM alleles of several genes, as described below. To do so, we first replaced the gene of interest in BY with the NatMX selectable marker by transforming a PCR product encoding the NatMX cassette with 40 bp overhangs at the 5’ and 3’ ends of the targeted gene. To generate GOIΔ::NatMX transformation fragments, we PCR amplified NatMX from Addgene plasmid #35121 with the primers listed in Supplementary file 3 using Phusion Hot Start Flex DNA polymerase (NEB). The NatMX cassette was transformed into the BY strain using the methods described above and transformants were plated onto YPD medium containing clonNAT. We verified the deletion of each gene of interest from single-colony purified transformants by colony PCR (primer sequences listed in Supplementary file 3).

We then modified the original CRISPR-Swap plasmid (PFA0055, Addgene plasmid #131774) to replace its LEU2 selectable marker with the HIS3 selectable marker, creating plasmid PFA0227 (Supplementary file 4). To build PFA0277, we first digested PFA0055 with restriction enzymes BsmBI-v2 and HpaI to remove the LEU2 selectable marker. We synthesized the S. cerevisiae HIS3 selectable marker from plasmid pRS313 (Sikorski and Hieter, 1989) with 20 base pairs of overlap to BsmBI-v2/HpaI-digested PFA0055 on both ends. We used this synthetic HIS3 fragment and BsmBI-v2/HpaI-digested PFA0055 to create plasmid PFA0227 by isothermal assembly cloning using the HiFi Assembly Cloning Kit (NEB) according to the manufacturer’s instructions. In addition to the HIS3 selectable marker, PFA0227 contains the cas9 gene driven by the constitutively active TDH3 promoter and a guide RNA, gCASS5a, that directs cleavage of a site immediately upstream of the TEF promoter used to drive expression of the MX series of selectable markers (Goldstein and McCusker, 1999; Lutz et al., 2019). We verified the sequence of PFA0227 by Sanger sequencing.

We used genomic DNA from BY and RM strains as a template to PCR amplify repair templates for CRISPR-Swap. Genomic DNA was extracted from BY and RM strains using the ‘10-min prep’ protocol (Hoffman and Winston, 1987). We amplified full-length repair templates from RM and BY containing each GOI’s promoter, open-reading frame (ORF), and terminator using Phusion Hot Start Flex DNA polymerase (NEB). We also created chimeric repair templates containing combinations of BY and RM alleles using PCR splicing by overlap extension (Horton et al., 1989). Table 4 lists the repair templates used for CRISPR swap. The sequence of all repair templates was verified by Sanger sequencing.

Table 4. CRISPR-swap repair templates.

Gene Allele Name Promoter ORF Terminator
UBR1 UBR1 BY BY BY BY
UBR1 UBR1 RM RM RM RM
UBR1 UBR1 RM promoter RM BY BY
UBR1 UBR1 RM ORF BY RM BY
UBR1 UBR1 RM terminator BY BY RM
UBR1 UBR1 -469A>T –469, RM; all other, BY BY BY
UBR1 UBR1 -197T>G –197, RM; all other, BY BY BY
DOA10 DOA10 BY BY BY BY
DOA10 DOA10 RM RM RM RM
DOA10 DOA10 Q410E BY 1228, RM; all other, BY BY
DOA10 DOA10 K1012N BY 3036, RM; all other, BY BY
DOA10 DOA10 Y1186F BY 3557, RM; all other, BY BY
NTA1 NTA1 BY BY BY BY
NTA1 NTA1 RM RM RM RM
NTA1 NTA1 RM promoter RM BY BY
NTA1 NTA1 D111E RM 331, RM; all other, BY BY
NTA1 NTA1 E129G RM 386, RM; all other, BY BY
UBC6 UBC6 BY BY BY BY
UBC6 UBC6 RM RM RM RM
UBC6 UBC6 RM promoter RM BY BY
UBC6 UBC6 D229G BY 1686, RM; all other, BY BY
UBC6 UBC6 RM terminator BY BY RM

To create allele swap strains, we co-transformed BY strains with 200 ng of plasmid PFA0227 and 1.5 μg of GOI repair template. Transformants were selected and single colony purified on synthetic complete medium lacking histidine and then patched onto solid YPD medium. We tested each strain for the desired exchange of the NatMX selectable marker with a UBR1 allele by patching strains onto solid YPD medium containing clonNAT. We then verified allelic exchange in strains lacking clonNAT resistance by colony PCR (primers listed in Supplementary file 3). We kept 16 independently-derived biological replicates of each allele swap strain. To test the effects of each allele swap, we transformed UPS activity reporters into our allele swap strains and characterized reporter activity by flow cytometry using the methods described above.

We tested whether a QTL on chromosome V results from variation in UBC6 using CRISPR-Swap. Deleting UBC6 caused a large growth defect relative to the wild-type BY strain. Providing cells with multiple UBC6 alleles, including the BY allele, did not correct the growth rate defect. We did not observe growth defects in any other fine-mapping strains.

RNA isolation

We isolated total RNA from 5 independent biological replicates each of the wild-type BY strain and a BY strain edited to contain the –469A>T RM variant in the UBR1 promoter (hereafter "UBR1 –469A>T BY"). All 10 samples were grown and harvested at the same time. BY and UBR1 –469A>T BY strains were grown overnight at 30°C in 5 mL of SC medium. The following day, the cultures were diluted to an OD of 0.05 in 100 mL of fresh SC medium and grown for approximately 7 hours. When the optical density (OD) of each culture was approximately 0.40, the cells were pelleted by centrifugation at 3,000 rpm for 10 min. Pellets were then washed by resuspending them in 1 mL of sterile ultrapure water, followed by centrifugation at 3,000 rpm for 3 min to again pellet the cells. Following this step, cell pellets were resuspended in 1 mL of ultrapure water and split into 4 aliquots, each containing 250 μL. After re-centrifuging and discarding the supernatant, the pellets were snap frozen by immersion in liquid nitrogen, followed by storage at −80°C. Pellets were subsequently used for RNA isolation and mass spectrometric proteomic analysis, as described below.

Total RNA was extracted from frozen cell pellets using the ZR Fungal/Bacterial miniprep kit (Zymo), according to the manufacturer’s instructions. Briefly, total RNA was isolated from cell pellets in two batches, each containing equal numbers of BY and UBR1 –469A>T BY samples. After thawing, pellets were resuspended in lysis buffer and transferred to screwcap lysis tubes containing glass beads. Tubes were secured in a Mini-BeadBeater (BioSpec Products, Bartlesville, OK, USA) and cells were processed in 5 cycles of 2 min of agitation followed by 2 min at −80°C. The cell lysate/bead mixture was centrifuged for 1 min at 16,000 x g and 400 μL of 95% ethanol was added to the cleared supernatant followed by mixing. Samples were then spun through a binding column and on-column DNA digest was performed with DNase I (Zymo) according to the manufacturer’s instructions. Total RNA was eluted from columns using 50 μL of RNase-free ultrapure water. The concentration of each sample was quantified using RiboGreen; all samples had a concentration greater than 300 ng/μL. The integrity of each sample was assessed at UMGC using the Tapestation (Agilent) and an RNA ScreenTape. RNA integrity numbers ranged from 9.7 to 10.0 (where 10.0 is the maximum possible score), with a median value of 9.9. All RNA samples were stored at −80°C.

RNA-seq

We isolated mRNA from each total RNA sample using the 550 ng of total RNA input and the NEBNext Poly(A) mRNA Magnetic Isolation Module (NEB). All samples were processed in a single batch and the isolated mRNA from each sample was used to prepare RNA sequencing libraries using the NEBNext Ultra II Directional RNA Library Prep kit (NEB) according to the manufacturer’s instructions. Libraries were amplified using NEBNext Ultra II Q5 polymerase and unique combinations of primers from the NEBNext Multiplex Oligos for Illumina (NEB). The following amplification protocol was used: initial denaturation at 98°C for 30 s, followed by 10 cycles of 98°C (10 s; denaturation), 65°C (75 s; annealing and extension), and a 65°C final extension for 5 min. PCR reactions were pooled using equal amounts of DNA and submitted to UMGC for three quality control assays, which measured the library concentration by PicoGreen, library functionality by KAPA qPCR, and library size using the Tapestation electrophoresis system (Agilent). The resulting library contained a small amount of adapter dimer (approximately 9%), which was subsequently removed via a bead-based cleanup. The final, cleaned library passed all three QC assays and was sequenced on a Next-Seq 2000 instrument (Illumina) in paired-end mode with 150 bp reads. The sequencing run generated 1,367,252,076 reads with an average of 136,725,207 (range: 112,285,619–152,571,763) reads per sample.

RNA-seq data processing and analysis

We performed quality control and preprocessing of RNA-seq data using fastp (Chen et al., 2018). Our initial processing removed reads with a length less than 36 bp and any reads where the mean quality dropped below a mean quality score of 15 in a 4 bp window. We also used fastp to trim adapter sequences from the ends of all reads. We then used Kallisto (Bray et al., 2016) to pseudoalign processed reads to the S. cerevisiae transcriptome, which was obtained from Ensembl (version 96) (Howe et al., 2021).

To identify differentially expressed transcripts, we used the estimated counts obtained from Kallisto as a measure of gene expression and filtered the estimated counts using the following procedures. First, we computed a transcript Transcript Integrity Number (TIN) for each gene using the RSeqQC (Wang et al., 2012) and removed any genes with a TIN less than 1 for any sample. We also removed any genes that Kallisto estimated to have an effective length less of less than 1 and those genes whose estimated counts were less than 10 in any sample. The resulting dataset comprised 5,676 expressed genes. Raw RNA-seq reads and processed counts were deposited in the NIH Gene Expression Omnibus database under accession number GSE213689. We used DESeq2 (Love et al., 2014) to perform statistical analysis of the resulting dataset. We used the RNA harvest batch and OD at time of sample harvest as covariates in our analysis. To further control for confounding sample-to-sample variation, we used surrogate variable analysis (Leek et al., 2012; Leek and Storey, 2007), which identified two significant surrogate variables that were subsequently added to our statistical model. We corrected for multiple testing using the Benjamini-Hochberg method (Benjamini and Hochberg, 1995) and considered significant differences as those with a corrected p-value less than 0.1.

To link differences in transcript abundance to biological pathways, we performed gene ontology enrichment analysis using PANTHER (Mi et al., 2021). The ‘statistical overrepresentation test’ was used to search for gene ontology (GO) biological processes and Reactome pathways enriched in our set of 78 transcripts differentially expressed between BY and UBR1 –469A>T BY. We used the 5,676 genes quantified in our RNA-seq statistical analysis as the reference set and used the Benjamini-Hochberg method (Benjamini and Hochberg, 1995) to correct for multiple testing. GO terms and Reactome pathways with a corrected p-value less than 0.05 were considered significant in our analysis.

Protein isolation and proteomic analysis by mass spectrometry

To quantify gene expression at the protein level, we submitted five cell pellets each from the same BY and UBR1 –469A>T BY cultures used for RNA-seq analysis to the University of Minnesota Center for Mass Spectrometry and Proteomics (CMSP) for proteomic analysis by mass spectrometry. Cell pellets were resuspended for protein extraction in a protein extraction buffer containing 7 M urea, 2 M thiourea, 0.4 M triethylammonium bicarbonate pH 8.5, 20% acetonitrile, and 4 mM tris(2-carboxyethyl)phosphine. Cell lysis and protein extraction was then performed using the Barocycler NEP2320 (Pressure Biosciences, Medford, MA, USA).

Samples were prepared and analyzed by mass spectrometry as follows. CMSP first labeled individual samples using the tandem mass tag (TMT) 10plex labeling kit (Thermo). After tagging, samples were pooled for analysis by mass spectrometry on an Orbitrap Tribrid Eclipse instrument (Thermo). Database searching was performed using the Proteome Discoverer software and the statistical analysis of protein abundance was performed in Scaffold (Proteome Software, Portland, OR, USA). We considered proteins to be differentially abundant between strains if they had a and a Benjamini-Hochberg corrected p-value less than 0.1. We performed ontological enrichment analysis of differentially abundant proteins using PANTHER as described above, except that the set of 3,046 detected proteins was used as the reference set.

Evolutionary analysis of variants

We inferred the allelic status of individual variants by comparing them to two outgroups: a likely-ancestral Taiwanese S. cerevisiae isolate and the sister species Saccharomyces paradoxus. We classified variants as ancestral if they were found in at least one outgroup. All alleles analyzed in this study could be unambiguously classified using this approach. We extracted the population frequency of all analyzed variants using genome sequence data from a panel of 1,011 S. cerevisiae isolates (Peter et al., 2018).

Data and statistical analysis

All data were analyzed using R (version 3.6.1; R Project for Statistical Computing). For all boxplots, the center line shows the median, the box excludes the upper and lower quartiles, the whiskers extend to 1.5 times the interquartile range. Protein structure predictions were obtained from the AlphaFold Protein Structure Database (Jumper et al., 2021) and visualized using ChimeraX (Pettersen et al., 2021). DNA binding motifs were determined using the Yeast Transcription Factor Specificity Compendium database (de Boer and Hughes, 2012). Final figures and illustrations were made using Inkscape (version 0.92; Inkscape Project).

Computational scripts used to process data, for statistical analysis, and to generate figures are available at: https://www.github.com/mac230/N-end_Rule_QTL_paper; copy archived at swh:1:rev:24baa12af4e9c45691be2590ab30b2c1faf0c497 (Collins, 2022).

Acknowledgements

We thank Leonid Kruglyak for the BY and RM yeast strains and Michael Knop for technical assistance in implementing the TFT reporter system. We thank the University of Minnesota’s Flow Cytometry Resource, Genomics Center, and Center for Mass Spectrometry and Proteomics for their contributions to the project. We thank Margaret Kliebhan for the BY / RM variant file used for QTL mapping. We thank the members of the Albert laboratory and the BioKansas Scientific Writing Program for critical feedback on the manuscript. This work was supported by NIH grants F32-GM128302 to MAC and R35-GM124676 to FWA, as well as a Pew Scholarship in the Biomedical Sciences from the Pew Charitable Trusts to FWA.

Appendix 1

Genetic Influences on the Proline N-degron

We observed that, consistent with previous results (Hwang et al., 2010; Gilchrist et al., 1997), the proline N-degron TFT was only partially stabilized in BY DOA10Δ (Figure 1D, Figure 1—figure supplement 1, Figure 1—source data 1). Ubiquitin is inefficiently cleaved in the ubiquitin-fusion technique when followed by a proline N-degron (Varshavsky, 2011; Varshavsky, 2005), leading to the production of two species, proline N-degron constructs and constructs with uncleaved N-terminal ubiquitin moieties. N-terminal ubiquitin functions as a degron and is recognized and degraded by the ubiquitin-fusion degradation (UFD) UPS pathway (Johnson et al., 1995). The proline N-end TFT thus measures the activities of the Ac/N-end pathway towards the proline N-degron and the UFD pathway towards the N-terminal ubiquitin fusion degron.

Despite this partial limitation, we were able to map genetic influences on the proline N-degron. We detected 5 Ac/N-end pathway-specific QTLs using our proline reporter, including those resulting from variation in DOA10 and UBC6. There were no additional QTLs affecting the majority of Ac/N-end reporters that were not detected with the proline reporter (Figure 2B, Figure 2—source data 2). We therefore conclude that the QTLs identified with the proline reporter correspond primarily to true genetic influences on the proline N-degron.

Three QTLs, on one chromosome XII and two on chromosome XVI were detected only with the proline reporter (Figure 2B, Figure 2—source data 2). The intervals for these QTLs did not contain genes previously linked to either the proline N-degron, the Ac/N-end pathway, or the UFD pathway.

Funding Statement

The funders had no role in study design, data collection and interpretation, or the decision to submit the work for publication.

Contributor Information

Mahlon A Collins, Email: mahlon@umn.edu.

Frank Wolfgang Albert, Email: falbert@umn.edu.

Magnus Nordborg, Gregor Mendel Institute, Austria.

David Ron, University of Cambridge, United Kingdom.

Funding Information

This paper was supported by the following grants:

  • National Institutes of Health F32-GM128302 to Mahlon A Collins.

  • National Institutes of Health R35-GM124676 to Frank Wolfgang Albert.

  • Pew Charitable Trusts Scholarship in the Biomedical Sciences to Frank Wolfgang Albert.

Additional information

Competing interests

No competing interests declared.

No competing interests declared.

Author contributions

Conceptualization, Software, Formal analysis, Supervision, Funding acquisition, Validation, Investigation, Visualization, Methodology, Writing - original draft, Project administration, Writing – review and editing.

Formal analysis, Validation, Investigation.

Conceptualization, Resources, Supervision, Funding acquisition, Methodology, Writing – review and editing.

Additional files

Supplementary file 1. Allele frequency difference and LOD score traces from QTL mapping experiments.

The plots show the loess-smoothed RM allele frequency difference (high UPS activity pool minus low UPS activity pool) and LOD score traces for the 20 N-degrons. QTLs are marked with asterisks, which are colored by biological replicate.

elife-79570-supp1.pdf (10MB, pdf)
Supplementary file 2. Influence of LOD score significance threshold on QTL pathway specificity.

The LOD score and RM allele frequency difference (QTL effect direction) traces for two independent biological replicates of each N-degron are shown for each of 23 pathway-specific QTL regions. Dashed lines at distinct LOD scores illustrate how changing the significance threshold changes the pathway-specificity of a given QTL region.

elife-79570-supp2.pdf (31.7MB, pdf)
Supplementary file 3. Oligonucleotides.

Table listing oligonucleotides used in this study.

elife-79570-supp3.xlsx (8.1KB, xlsx)
Supplementary file 4. Plasmids.

Table of plasmids used in this study.

elife-79570-supp4.xlsx (6.2KB, xlsx)
Supplementary file 5. Yeast strains.

Table listing all yeast strains used in the study.

elife-79570-supp5.xlsx (12.2KB, xlsx)
MDAR checklist

Data availability

Raw sequencing reads from QTL mapping experiments are available from the NIH Sequence Read Archive under the Bioproject Accession PRJNA881749. Raw and processed RNA-seq data is available from the NIH Gene Expression Omnibus under the accession GSE213689. These datasets are fully available without restriction. Computational scripts used to process data, for statistical analysis, and to generate figures are available at: https://www.github.com/mac230/N-end_Rule_QTL_paper, (copy archived at swh:1:rev:24baa12af4e9c45691be2590ab30b2c1faf0c497).

The following datasets were generated:

Collins MA. 2022. Variation in ubiquitin system genes creates substrate-specific effects on proteasomal protein degradation. NCBI Gene Expression Omnibus. GSE213689

Collins MA. 2022. Variation in ubiquitin system genes creates substrate-specific effects on proteasomal protein degradation. NCBI BioProject. PRJNA881749

References

  1. Abell NS, DeGorter MK, Gloudemans MJ, Greenwald E, Smith KS, He Z, Montgomery SB. Multiple causal variants underlie genetic associations in humans. Science. 2022;375:1247–1254. doi: 10.1126/science.abj5117. [DOI] [PMC free article] [PubMed] [Google Scholar]
  2. Agarwal AK, Xing C, DeMartino GN, Mizrachi D, Hernandez MD, Sousa AB, Martínez de Villarreal L, dos Santos HG, Garg A. PSMB8 encoding the β5i proteasome subunit is mutated in joint contractures, muscle atrophy, microcytic anemia, and panniculitis-induced lipodystrophy syndrome. American Journal of Human Genetics. 2010;87:866–872. doi: 10.1016/j.ajhg.2010.10.031. [DOI] [PMC free article] [PubMed] [Google Scholar]
  3. Albert FW, Treusch S, Shockley AH, Bloom JS, Kruglyak L. Genetics of single-cell protein abundance variation in large yeast populations. Nature. 2014;506:494–497. doi: 10.1038/nature12904. [DOI] [PMC free article] [PubMed] [Google Scholar]
  4. Albert FW, Bloom JS, Siegel J, Day L, Kruglyak L. Genetics of trans-regulatory variation in gene expression. eLife. 2018;7:e35471. doi: 10.7554/eLife.35471. [DOI] [PMC free article] [PubMed] [Google Scholar]
  5. Arima K, Kinoshita A, Mishima H, Kanazawa N, Kaneko T, Mizushima T, Ichinose K, Nakamura H, Tsujino A, Kawakami A, Matsunaka M, Kasagi S, Kawano S, Kumagai S, Ohmura K, Mimori T, Hirano M, Ueno S, Tanaka K, Tanaka M, Toyoshima I, Sugino H, Yamakawa A, Tanaka K, Niikawa N, Furukawa F, Murata S, Eguchi K, Ida H, Yoshiura KI. Proteasome assembly defect due to a proteasome subunit beta type 8 (PSMB8) mutation causes the autoinflammatory disorder, nakajo-nishimura syndrome. PNAS. 2011;108:14914–14919. doi: 10.1073/pnas.1106015108. [DOI] [PMC free article] [PubMed] [Google Scholar]
  6. Bachmair A, Finley D, Varshavsky A. In vivo half-life of a protein is a function of its amino-terminal residue. Science. 1986;234:179–186. doi: 10.1126/science.3018930. [DOI] [PubMed] [Google Scholar]
  7. Bajorek M, Finley D, Glickman MH. Proteasome disassembly and downregulation is correlated with viability during stationary phase. Current Biology. 2003;13:1140–1144. doi: 10.1016/s0960-9822(03)00417-2. [DOI] [PubMed] [Google Scholar]
  8. Baker RT, Varshavsky A. Yeast N-terminal amidase A new enzyme and component of the N-end rule pathway. The Journal of Biological Chemistry. 1995;270:12065–12074. doi: 10.1074/jbc.270.20.12065. [DOI] [PubMed] [Google Scholar]
  9. Baraibar MA, Friguet B. Changes of the proteasomal system during the aging process. Progress in Molecular Biology and Translational Science. 2012;109:249–275. doi: 10.1016/B978-0-12-397863-9.00007-9. [DOI] [PubMed] [Google Scholar]
  10. Bartel B, Wünning I, Varshavsky A. The recognition component of the N-end rule pathway. The EMBO Journal. 1990;9:3179–3189. doi: 10.1002/j.1460-2075.1990.tb07516.x. [DOI] [PMC free article] [PubMed] [Google Scholar]
  11. Baryshnikova A, Costanzo M, Dixon S, Vizeacoumar FJ, Myers CL, Andrews B, Boone C. Synthetic genetic array (SGA) analysis in Saccharomyces cerevisiae and Schizosaccharomyces pombe. Methods in Enzymology. 2010;470:145–179. doi: 10.1016/S0076-6879(10)70007-0. [DOI] [PubMed] [Google Scholar]
  12. Battle A, Khan Z, Wang SH, Mitrano A, Ford MJ, Pritchard JK, Gilad Y. Genomic variation. Impact of regulatory variation from RNA to protein. Science. 2015;347:664–667. doi: 10.1126/science.1260793. [DOI] [PMC free article] [PubMed] [Google Scholar]
  13. Benjamini Y, Hochberg Y. Controlling the false discovery rate: a practical and powerful approach to multiple testing. Journal of the Royal Statistical Society. 1995;57:289–300. doi: 10.1111/j.2517-6161.1995.tb02031.x. [DOI] [Google Scholar]
  14. Bett JS. Proteostasis regulation by the ubiquitin system. Essays in Biochemistry. 2016;60:143–151. doi: 10.1042/EBC20160001. [DOI] [PubMed] [Google Scholar]
  15. Bloom JS, Ehrenreich IM, Loo WT, Lite T-LV, Kruglyak L. Finding the sources of missing heritability in a yeast cross. Nature. 2013;494:234–237. doi: 10.1038/nature11867. [DOI] [PMC free article] [PubMed] [Google Scholar]
  16. Bolger AM, Lohse M, Usadel B. Trimmomatic: a flexible trimmer for illumina sequence data. Bioinformatics. 2014;30:2114–2120. doi: 10.1093/bioinformatics/btu170. [DOI] [PMC free article] [PubMed] [Google Scholar]
  17. Bray NL, Pimentel H, Melsted P, Pachter L. Near-optimal probabilistic RNA-seq quantification. Nature Biotechnology. 2016;34:525–527. doi: 10.1038/nbt.3519. [DOI] [PubMed] [Google Scholar]
  18. Brehm A, Liu Y, Sheikh A, Marrero B, Omoyinmi E, Zhou Q, Montealegre G, Biancotto A, Reinhardt A, Almeida de Jesus A, Pelletier M, Tsai WL, Remmers EF, Kardava L, Hill S, Kim H, Lachmann HJ, Megarbane A, Chae JJ, Brady J, Castillo RD, Brown D, Casano AV, Gao L, Chapelle D, Huang Y, Stone D, Chen Y, Sotzny F, Lee C-CR, Kastner DL, Torrelo A, Zlotogorski A, Moir S, Gadina M, McCoy P, Wesley R, Rother KI, Rother K, Hildebrand PW, Brogan P, Krüger E, Aksentijevich I, Goldbach-Mansky R. Additive loss-of-function proteasome subunit mutations in CANDLE/PRAAS patients promote type I IFN production. The Journal of Clinical Investigation. 2015;125:4196–4211. doi: 10.1172/JCI81260. [DOI] [PMC free article] [PubMed] [Google Scholar]
  19. Brion C, Lutz SM, Albert FW. Simultaneous quantification of mRNA and protein in single cells reveals post-transcriptional effects of genetic variation. eLife. 2020;9:e60645. doi: 10.7554/eLife.60645. [DOI] [PMC free article] [PubMed] [Google Scholar]
  20. Cenik C, Cenik ES, Byeon GW, Grubert F, Candille SI, Spacek D, Alsallakh B, Tilgner H, Araya CL, Tang H, Ricci E, Snyder MP. Integrative analysis of RNA, translation, and protein levels reveals distinct regulatory variation across humans. Genome Research. 2015;25:1610–1621. doi: 10.1101/gr.193342.115. [DOI] [PMC free article] [PubMed] [Google Scholar]
  21. Chen P, Johnson P, Sommer T, Jentsch S, Hochstrasser M. Multiple ubiquitin-conjugating enzymes participate in the in vivo degradation of the yeast MAT alpha 2 repressor. Cell. 1993;74:357–369. doi: 10.1016/0092-8674(93)90426-q. [DOI] [PubMed] [Google Scholar]
  22. Chen E, Kwon YT, Lim MS, Dubé ID, Hough MR. Loss of UBR1 promotes aneuploidy and accelerates B-cell lymphomagenesis in TLX1/HOX11-transgenic mice. Oncogene. 2006;25:5752–5763. doi: 10.1038/sj.onc.1209573. [DOI] [PubMed] [Google Scholar]
  23. Chen S, Zhou Y, Chen Y, Gu J. Fastp: an ultra-fast all-in-one FASTQ preprocessor. Bioinformatics. 2018;34:i884–i890. doi: 10.1093/bioinformatics/bty560. [DOI] [PMC free article] [PubMed] [Google Scholar]
  24. Chick JM, Munger SC, Simecek P, Huttlin EL, Choi K, Gatti DM, Raghupathy N, Svenson KL, Churchill GA, Gygi SP. Defining the consequences of genetic variation on a proteome-wide scale. Nature. 2016;534:500–505. doi: 10.1038/nature18270. [DOI] [PMC free article] [PubMed] [Google Scholar]
  25. Cho YS, Chen C-H, Hu C, Long J, Ong RTH, Sim X, Takeuchi F, Wu Y, Go MJ, Yamauchi T, Chang Y-C, Kwak SH, Ma RCW, Yamamoto K, Adair LS, Aung T, Cai Q, Chang L-C, Chen Y-T, Gao Y, Hu FB, Kim H-L, Kim S, Kim YJ, Lee JJ-M, Lee NR, Li Y, Liu JJ, Lu W, Nakamura J, Nakashima E, Ng DP-K, Tay WT, Tsai F-J, Wong TY, Yokota M, Zheng W, Zhang R, Wang C, So WY, Ohnaka K, Ikegami H, Hara K, Cho YM, Cho NH, Chang T-J, Bao Y, Hedman ÅK, Morris AP, McCarthy MI, Takayanagi R, Park KS, Jia W, Chuang L-M, Chan JCN, Maeda S, Kadowaki T, Lee J-Y, Wu J-Y, Teo YY, Tai ES, Shu XO, Mohlke KL, Kato N, Han B-G, Seielstad M, DIAGRAM Consortium. MuTHER Consortium Meta-analysis of genome-wide association studies identifies eight new loci for type 2 diabetes in east asians. Nature Genetics. 2011;44:67–72. doi: 10.1038/ng.1019. [DOI] [PMC free article] [PubMed] [Google Scholar]
  26. Christiano R, Arlt H, Kabatnik S, Mejhert N, Lai ZW, Farese RV, Walther TC. A systematic protein turnover map for decoding protein degradation. Cell Reports. 2020;33:108378. doi: 10.1016/j.celrep.2020.108378. [DOI] [PMC free article] [PubMed] [Google Scholar]
  27. Ciechanover A, Orian A, Schwartz AL. Ubiquitin-Mediated proteolysis: biological regulation via destruction. BioEssays. 2000;22:442–451. doi: 10.1002/(SICI)1521-1878(200005)22:5<442::AID-BIES6>3.0.CO;2-Q. [DOI] [PubMed] [Google Scholar]
  28. Collins GA, Goldberg AL. The logic of the 26S proteasome. Cell. 2017;169:792–806. doi: 10.1016/j.cell.2017.04.023. [DOI] [PMC free article] [PubMed] [Google Scholar]
  29. Collins M. N-end_rule_qtl_paper. swh:1:rev:24baa12af4e9c45691be2590ab30b2c1faf0c497Software Heritage. 2022 https://archive.softwareheritage.org/swh:1:dir:9e4de720d8485ef11cd52171952e8a87baee3e09;origin=https://www.github.com/mac230/N-end_Rule_QTL_paper;visit=swh:1:snp:8fba304ebe4d095e13449fce533c22fe254f2f71;anchor=swh:1:rev:24baa12af4e9c45691be2590ab30b2c1faf0c497
  30. Coux O, Tanaka K, Goldberg AL. Structure and functions of the 20S and 26S proteasomes. Annual Review of Biochemistry. 1996;65:801–847. doi: 10.1146/annurev.bi.65.070196.004101. [DOI] [PubMed] [Google Scholar]
  31. de Boer CG, Hughes TR. YeTFaSCo: a database of evaluated yeast transcription factor sequence specificities. Nucleic Acids Research. 2012;40:D169–D179. doi: 10.1093/nar/gkr993. [DOI] [PMC free article] [PubMed] [Google Scholar]
  32. Deng HX, Chen W, Hong ST, Boycott KM, Gorrie GH, Siddique N, Yang Y, Fecto F, Shi Y, Zhai H, Jiang H, Hirano M, Rampersaud E, Jansen GH, Donkervoort S, Bigio EH, Brooks BR, Ajroud K, Sufit RL, Haines JL, Mugnaini E, Pericak-Vance MA, Siddique T. Mutations in UBQLN2 cause dominant X-linked juvenile and adult-onset ALS and ALS/dementia. Nature. 2011;477:211–215. doi: 10.1038/nature10353. [DOI] [PMC free article] [PubMed] [Google Scholar]
  33. Diskin SJ, Capasso M, Schnepp RW, Cole KA, Attiyeh EF, Hou C, Diamond M, Carpenter EL, Winter C, Lee H, Jagannathan J, Latorre V, Iolascon A, Hakonarson H, Devoto M, Maris JM. Common variation at 6q16 within Hace1 and Lin28b influences susceptibility to neuroblastoma. Nature Genetics. 2012;44:1126–1130. doi: 10.1038/ng.2387. [DOI] [PMC free article] [PubMed] [Google Scholar]
  34. Edwards MD, Gifford DK. High-resolution genetic mapping with pooled sequencing. BMC Bioinformatics. 2012;13 Suppl 6:S8. doi: 10.1186/1471-2105-13-S6-S8. [DOI] [PMC free article] [PubMed] [Google Scholar]
  35. Ehrenreich IM, Gerke JP, Kruglyak L. Genetic dissection of complex traits in yeast: insights from studies of gene expression and other phenotypes in the byxrm cross. Cold Spring Harbor Symposia on Quantitative Biology; 2009. pp. 145–153. [DOI] [PMC free article] [PubMed] [Google Scholar]
  36. Ehrenreich IM, Torabi N, Jia Y, Kent J, Martis S, Shapiro JA, Gresham D, Caudy AA, Kruglyak L. Dissection of genetically complex traits with extremely large pools of yeast segregants. Nature. 2010;464:1039–1042. doi: 10.1038/nature08923. [DOI] [PMC free article] [PubMed] [Google Scholar]
  37. Finley D, Ulrich HD, Sommer T, Kaiser P. The ubiquitin-proteasome system of Saccharomyces cerevisiae. Genetics. 2012;192:319–360. doi: 10.1534/genetics.112.140467. [DOI] [PMC free article] [PubMed] [Google Scholar]
  38. Finley D, Prado MA. The proteasome and its network: engineering for adaptability. Cold Spring Harbor Perspectives in Biology. 2020;12:a033985. doi: 10.1101/cshperspect.a033985. [DOI] [PMC free article] [PubMed] [Google Scholar]
  39. Foss EJ, Radulovic D, Shaffer SA, Goodlett DR, Kruglyak L, Bedalov A. Genetic variation shapes protein networks mainly through non-transcriptional mechanisms. PLOS Biology. 2011;9:e1001144. doi: 10.1371/journal.pbio.1001144. [DOI] [PMC free article] [PubMed] [Google Scholar]
  40. Gaisne M, Bécam AM, Verdière J, Herbert CJ. A “ natural ” mutation in Saccharomyces cerevisiae strains derived from S288C affects the complex regulatory gene Hap1 (CYP1) Current Genetics. 1999;36:195–200. doi: 10.1007/s002940050490. [DOI] [PubMed] [Google Scholar]
  41. Geffen Y, Appleboim A, Gardner RG, Friedman N, Sadeh R, Ravid T. Mapping the landscape of a eukaryotic degronome. Molecular Cell. 2016;63:1055–1065. doi: 10.1016/j.molcel.2016.08.005. [DOI] [PubMed] [Google Scholar]
  42. Ghazalpour A, Bennett B, Petyuk VA, Orozco L, Hagopian R, Mungrue IN, Farber CR, Sinsheimer J, Kang HM, Furlotte N, Park CC, Wen PZ, Brewer H, Weitz K, Camp DG, Pan C, Yordanova R, Neuhaus I, Tilford C, Siemers N, Gargalovic P, Eskin E, Kirchgessner T, Smith DJ, Smith RD, Lusis AJ. Comparative analysis of proteome and transcriptome variation in mouse. PLOS Genetics. 2011;7:e1001393. doi: 10.1371/journal.pgen.1001393. [DOI] [PMC free article] [PubMed] [Google Scholar]
  43. Gietz RD, Schiestl RH. High-Efficiency yeast transformation using the liac/SS carrier DNA/PEG method. Nature Protocols. 2007;2:31–34. doi: 10.1038/nprot.2007.13. [DOI] [PubMed] [Google Scholar]
  44. Gilchrist CA, Gray DA, Baker RT. A ubiquitin-specific protease that efficiently cleaves the ubiquitin-proline bond. The Journal of Biological Chemistry. 1997;272:32280–32285. doi: 10.1074/jbc.272.51.32280. [DOI] [PubMed] [Google Scholar]
  45. Goldstein AL, McCusker JH. Three new dominant drug resistance cassettes for gene disruption in Saccharomyces cerevisiae. Yeast. 1999;15:1541–1553. doi: 10.1002/(SICI)1097-0061(199910)15:14<1541::AID-YEA476>3.0.CO;2-K. [DOI] [PubMed] [Google Scholar]
  46. Gomes AV. Genetics of proteasome diseases. Scientifica. 2013;2013:637629. doi: 10.1155/2013/637629. [DOI] [PMC free article] [PubMed] [Google Scholar]
  47. Grimm S, Höhn A, Grune T. Oxidative protein damage and the proteasome. Amino Acids. 2012;42:23–38. doi: 10.1007/s00726-010-0646-8. [DOI] [PubMed] [Google Scholar]
  48. Grote A, Hiller K, Scheer M, Münch R, Nörtemann B, Hempel DC, Jahn D. JCat: a novel tool to adapt codon usage of a target gene to its potential expression host. Nucleic Acids Research. 2005;33:W526–W531. doi: 10.1093/nar/gki376. [DOI] [PMC free article] [PubMed] [Google Scholar]
  49. Hahne F, LeMeur N, Brinkman RR, Ellis B, Haaland P, Sarkar D, Spidlen J, Strain E, Gentleman R. FlowCore: a Bioconductor package for high throughput flow cytometry. BMC Bioinformatics. 2009;10:106. doi: 10.1186/1471-2105-10-106. [DOI] [PMC free article] [PubMed] [Google Scholar]
  50. Hanna J, Finley D. A proteasome for all occasions. FEBS Letters. 2007;581:2854–2861. doi: 10.1016/j.febslet.2007.03.053. [DOI] [PMC free article] [PubMed] [Google Scholar]
  51. Hershko A, Ciechanover A. The ubiquitin system. Annual Review of Biochemistry. 1998;67:425–479. doi: 10.1146/annurev.biochem.67.1.425. [DOI] [PubMed] [Google Scholar]
  52. Hoffman CS, Winston F. A ten-minute DNA preparation from yeast efficiently releases autonomous plasmids for transformation of Escherichia coli. Gene. 1987;57:267–272. doi: 10.1016/0378-1119(87)90131-4. [DOI] [PubMed] [Google Scholar]
  53. Horton RM, Hunt HD, Ho SN, Pullen JK, Pease LR. Engineering hybrid genes without the use of restriction enzymes: gene splicing by overlap extension. Gene. 1989;77:61–68. doi: 10.1016/0378-1119(89)90359-4. [DOI] [PubMed] [Google Scholar]
  54. Howe KL, Achuthan P, Allen J, Allen J, Alvarez-Jarreta J, Amode MR, Armean IM, Azov AG, Bennett R, Bhai J, Billis K, Boddu S, Charkhchi M, Cummins C, Da Rin Fioretto L, Davidson C, Dodiya K, El Houdaigui B, Fatima R, Gall A, Garcia Giron C, Grego T, Guijarro-Clarke C, Haggerty L, Hemrom A, Hourlier T, Izuogu OG, Juettemann T, Kaikala V, Kay M, Lavidas I, Le T, Lemos D, Gonzalez Martinez J, Marugán JC, Maurel T, McMahon AC, Mohanan S, Moore B, Muffato M, Oheh DN, Paraschas D, Parker A, Parton A, Prosovetskaia I, Sakthivel MP, Salam AIA, Schmitt BM, Schuilenburg H, Sheppard D, Steed E, Szpak M, Szuba M, Taylor K, Thormann A, Threadgold G, Walts B, Winterbottom A, Chakiachvili M, Chaubal A, De Silva N, Flint B, Frankish A, Hunt SE, IIsley GR, Langridge N, Loveland JE, Martin FJ, Mudge JM, Morales J, Perry E, Ruffier M, Tate J, Thybert D, Trevanion SJ, Cunningham F, Yates AD, Zerbino DR, Flicek P. Ensembl 2021. Nucleic Acids Research. 2021;49:D884–D891. doi: 10.1093/nar/gkaa942. [DOI] [PMC free article] [PubMed] [Google Scholar]
  55. Hwang CS, Shemorry A, Varshavsky A. N-Terminal acetylation of cellular proteins creates specific degradation signals. Science. 2010;327:973–977. doi: 10.1126/science.1183147. [DOI] [PMC free article] [PubMed] [Google Scholar]
  56. Hwang C-S, Sukalo M, Batygin O, Addor M-C, Brunner H, Aytes AP, Mayerle J, Song HK, Varshavsky A, Zenker M, Nordheim A. Ubiquitin ligases of the N-end rule pathway: assessment of mutations in UBR1 that cause the johanson-blizzard syndrome. PLOS ONE. 2011;6:e24925. doi: 10.1371/journal.pone.0024925. [DOI] [PMC free article] [PubMed] [Google Scholar]
  57. Icho T, Lee HS, Sommer SS, Wickner RB. Molecular characterization of chromosomal genes affecting double-stranded RNA replication in Saccharomyces cerevisiae. Basic Life Sciences. 1986;40:165–171. doi: 10.1007/978-1-4684-5251-8_13. [DOI] [PubMed] [Google Scholar]
  58. Inobe T, Matouschek A. Paradigms of protein degradation by the proteasome. Current Opinion in Structural Biology. 2014;24:156–164. doi: 10.1016/j.sbi.2014.02.002. [DOI] [PMC free article] [PubMed] [Google Scholar]
  59. Jain S, Wheeler JR, Walters RW, Agrawal A, Barsic A, Parker R. ATPase-modulated stress granules contain a diverse proteome and substructure. Cell. 2016;164:487–498. doi: 10.1016/j.cell.2015.12.038. [DOI] [PMC free article] [PubMed] [Google Scholar]
  60. Johnson ES, Bartel B, Seufert W, Varshavsky A. Ubiquitin as a degradation signal. The EMBO Journal. 1992;11:497–505. doi: 10.1002/j.1460-2075.1992.tb05080.x. [DOI] [PMC free article] [PubMed] [Google Scholar]
  61. Johnson ES, Ma PCM, Ota IM, Varshavsky A. A proteolytic pathway that recognizes ubiquitin as A degradation signal. The Journal of Biological Chemistry. 1995;270:17442–17456. doi: 10.1074/jbc.270.29.17442. [DOI] [PubMed] [Google Scholar]
  62. Jumper J, Evans R, Pritzel A, Green T, Figurnov M, Ronneberger O, Tunyasuvunakool K, Bates R, Žídek A, Potapenko A, Bridgland A, Meyer C, Kohl SAA, Ballard AJ, Cowie A, Romera-Paredes B, Nikolov S, Jain R, Adler J, Back T, Petersen S, Reiman D, Clancy E, Zielinski M, Steinegger M, Pacholska M, Berghammer T, Bodenstein S, Silver D, Vinyals O, Senior AW, Kavukcuoglu K, Kohli P, Hassabis D. Highly accurate protein structure prediction with alphafold. Nature. 2021;596:583–589. doi: 10.1038/s41586-021-03819-2. [DOI] [PMC free article] [PubMed] [Google Scholar]
  63. Kats I, Khmelinskii A, Kschonsak M, Huber F, Knieß RA, Bartosik A, Knop M. Mapping degradation signals and pathways in a eukaryotic N-terminome. Molecular Cell. 2018;70:488–501. doi: 10.1016/j.molcel.2018.03.033. [DOI] [PubMed] [Google Scholar]
  64. Khmelinskii A., Keller PJ, Bartosik A, Meurer M, Barry JD, Mardin BR, Kaufmann A, Trautmann S, Wachsmuth M, Pereira G, Huber W, Schiebel E, Knop M. Tandem fluorescent protein timers for in vivo analysis of protein dynamics. Nature Biotechnology. 2012;30:708–714. doi: 10.1038/nbt.2281. [DOI] [PubMed] [Google Scholar]
  65. Khmelinskii A., Blaszczak E, Pantazopoulou M, Fischer B, Omnus DJ, Le Dez G, Brossard A, Gunnarsson A, Barry JD, Meurer M, Kirrmaier D, Boone C, Huber W, Rabut G, Ljungdahl PO, Knop M. Protein quality control at the inner nuclear membrane. Nature. 2014;516:410–413. doi: 10.1038/nature14096. [DOI] [PMC free article] [PubMed] [Google Scholar]
  66. Khmelinskii A, Knop M. Analysis of protein dynamics with tandem fluorescent protein timers. Methods in Molecular Biology. 2014;1174:195–210. doi: 10.1007/978-1-4939-0944-5_13. [DOI] [PubMed] [Google Scholar]
  67. Khmelinskii A, Meurer M, Ho CT, Besenbeck B, Füller J, Lemberg MK, Bukau B, Mogk A, Knop M. Incomplete proteasomal degradation of green fluorescent proteins in the context of tandem fluorescent protein timers. Molecular Biology of the Cell. 2016;27:360–370. doi: 10.1091/mbc.E15-07-0525. [DOI] [PMC free article] [PubMed] [Google Scholar]
  68. Kisselev AF, Akopian TN, Woo KM, Goldberg AL. The sizes of peptides generated from protein by mammalian 26 and 20 S proteasomes. Implications for understanding the degradative mechanism and antigen presentation. The Journal of Biological Chemistry. 1999;274:3363–3371. doi: 10.1074/jbc.274.6.3363. [DOI] [PubMed] [Google Scholar]
  69. Komander D, Rape M. The ubiquitin code. Annual Review of Biochemistry. 2012;81:203–229. doi: 10.1146/annurev-biochem-060310-170328. [DOI] [PubMed] [Google Scholar]
  70. Kong K-YE, Fischer B, Meurer M, Kats I, Li Z, Rühle F, Barry JD, Kirrmaier D, Chevyreva V, San Luis B-J, Costanzo M, Huber W, Andrews BJ, Boone C, Knop M, Khmelinskii A. Timer-based proteomic profiling of the ubiquitin-proteasome system reveals a substrate receptor of the GID ubiquitin ligase. Molecular Cell. 2021;81:2460–2476. doi: 10.1016/j.molcel.2021.04.018. [DOI] [PMC free article] [PubMed] [Google Scholar]
  71. Kredel S, Oswald F, Nienhaus K, Deuschle K, Röcker C, Wolff M, Heilker R, Nienhaus GU, Wiedenmann J, Gladfelter AS. MRuby, a bright monomeric red fluorescent protein for labeling of subcellular structures. PLOS ONE. 2009;4:e4391. doi: 10.1371/journal.pone.0004391. [DOI] [PMC free article] [PubMed] [Google Scholar]
  72. Kröll-Hermi A, Ebstein F, Stoetzel C, Geoffroy V, Schaefer E, Scheidecker S, Bär S, Takamiya M, Kawakami K, Zieba BA, Studer F, Pelletier V, Eyermann C, Speeg-Schatz C, Laugel V, Lipsker D, Sandron F, McGinn S, Boland A, Deleuze JF, Kuhn L, Chicher J, Hammann P, Friant S, Etard C, Krüger E, Muller J, Strähle U, Dollfus H. Proteasome subunit PSMC3 variants cause neurosensory syndrome combining deafness and cataract due to proteotoxic stress. EMBO Molecular Medicine. 2020;12:e11861. doi: 10.15252/emmm.201911861. [DOI] [PMC free article] [PubMed] [Google Scholar]
  73. Kuzmin E, Costanzo M, Andrews B, Boone C. Synthetic genetic array analysis. Cold Spring Harbor Protocols. 2016;2016:pdb.prot088807. doi: 10.1101/pdb.prot088807. [DOI] [PubMed] [Google Scholar]
  74. Laporte D, Salin B, Daignan-Fornier B, Sagot I. Reversible cytoplasmic localization of the proteasome in quiescent yeast cells. The Journal of Cell Biology. 2008;181:737–745. doi: 10.1083/jcb.200711154. [DOI] [PMC free article] [PubMed] [Google Scholar]
  75. Laurie-Ahlberg CC, Stam LF. Use of P-element-mediated transformation to identify the molecular basis of naturally occurring variants affecting Adh expression in Drosophila melanogaster. Genetics. 1987;115:129–140. doi: 10.1093/genetics/115.1.129. [DOI] [PMC free article] [PubMed] [Google Scholar]
  76. Leek JT, Storey JD. Capturing heterogeneity in gene expression studies by surrogate variable analysis. PLOS Genetics. 2007;3:e30161. doi: 10.1371/journal.pgen.0030161. [DOI] [PMC free article] [PubMed] [Google Scholar]
  77. Leek JT, Johnson WE, Parker HS, Jaffe AE, Storey JD. The SVA package for removing batch effects and other unwanted variation in high-throughput experiments. Bioinformatics. 2012;28:882–883. doi: 10.1093/bioinformatics/bts034. [DOI] [PMC free article] [PubMed] [Google Scholar]
  78. Li W, Bengtson MH, Ulbrich A, Matsuda A, Reddy VA, Orth A, Chanda SK, Batalov S, Joazeiro CAP. Genome-Wide and functional annotation of human E3 ubiquitin ligases identifies Mulan, a mitochondrial E3 that regulates the organelle’s dynamics and signaling. PLOS ONE. 2008;3:e1487. doi: 10.1371/journal.pone.0001487. [DOI] [PMC free article] [PubMed] [Google Scholar]
  79. Li H, Durbin R. Fast and accurate short read alignment with burrows-wheeler transform. Bioinformatics. 2009a;25:1754–1760. doi: 10.1093/bioinformatics/btp324. [DOI] [PMC free article] [PubMed] [Google Scholar]
  80. Li H, Handsaker B, Wysoker A, Fennell T, Ruan J, Homer N, Marth G, Abecasis G, Durbin R, 1000 Genome Project Data Processing Subgroup The sequence alignment/map format and samtools. Bioinformatics. 2009b;25:2078–2079. doi: 10.1093/bioinformatics/btp352. [DOI] [PMC free article] [PubMed] [Google Scholar]
  81. Liu Y, Ramot Y, Torrelo A, Paller AS, Si N, Babay S, Kim PW, Sheikh A, Lee C-CR, Chen Y, Vera A, Zhang X, Goldbach-Mansky R, Zlotogorski A. Mutations in proteasome subunit β type 8 cause chronic atypical neutrophilic dermatosis with lipodystrophy and elevated temperature with evidence of genetic and phenotypic heterogeneity. Arthritis and Rheumatism. 2012;64:895–907. doi: 10.1002/art.33368. [DOI] [PMC free article] [PubMed] [Google Scholar]
  82. Love MI, Huber W, Anders S. Moderated estimation of fold change and dispersion for RNA-seq data with deseq2. Genome Biology. 2014;15:12. doi: 10.1186/s13059-014-0550-8. [DOI] [PMC free article] [PubMed] [Google Scholar]
  83. Lutz S, Brion C, Kliebhan M, Albert FW. Dna variants affecting the expression of numerous genes in trans have diverse mechanisms of action and evolutionary histories. PLOS Genetics. 2019;15:11. doi: 10.1371/journal.pgen.1008375. [DOI] [PMC free article] [PubMed] [Google Scholar]
  84. Lutz S, Van Dyke K, Feraru MA, Albert FW. Multiple epistatic DNA variants in a single gene affect gene expression in trans. Genetics. 2022;220:iyab208. doi: 10.1093/genetics/iyab208. [DOI] [PMC free article] [PubMed] [Google Scholar]
  85. Mackay TFC, Stone EA, Ayroles JF. The genetics of quantitative traits: challenges and prospects. Nature Reviews. Genetics. 2009;10:565–577. doi: 10.1038/nrg2612. [DOI] [PubMed] [Google Scholar]
  86. Mi H, Ebert D, Muruganujan A, Mills C, Albou LP, Mushayamaha T, Thomas PD. Panther version 16: a revised family classification, tree-based classification tool, enhancer regions and extensive API. Nucleic Acids Research. 2021;49:D394–D403. doi: 10.1093/nar/gkaa1106. [DOI] [PMC free article] [PubMed] [Google Scholar]
  87. Michelmore RW, Paran I, Kesseli RV. Identification of markers linked to disease-resistance genes by bulked segregant analysis: a rapid method to detect markers in specific genomic regions by using segregating populations. PNAS. 1991;88:9828–9832. doi: 10.1073/pnas.88.21.9828. [DOI] [PMC free article] [PubMed] [Google Scholar]
  88. Mirauta BA, Seaton DD, Bensaddek D, Brenes A, Bonder MJ, Kilpinen H, Stegle O, Lamond AI, HipSci Consortium Population-scale proteome variation in human induced pluripotent stem cells. eLife. 2020;9:e57390. doi: 10.7554/eLife.57390. [DOI] [PMC free article] [PubMed] [Google Scholar]
  89. Oh JH, Hyun JY, Chen SJ, Varshavsky A. Five enzymes of the arg/N-degron pathway form a targeting complex: the concept of superchanneling. PNAS. 2020;117:10778–10788. doi: 10.1073/pnas.2003043117. [DOI] [PMC free article] [PubMed] [Google Scholar]
  90. Pédelacq J-D, Cabantous S, Tran T, Terwilliger TC, Waldo GS. Engineering and characterization of a superfolder green fluorescent protein. Nature Biotechnology. 2006;24:79–88. doi: 10.1038/nbt1172. [DOI] [PubMed] [Google Scholar]
  91. Peter J, De Chiara M, Friedrich A, Yue J-X, Pflieger D, Bergström A, Sigwalt A, Barre B, Freel K, Llored A, Cruaud C, Labadie K, Aury J-M, Istace B, Lebrigand K, Barbry P, Engelen S, Lemainque A, Wincker P, Liti G, Schacherer J. Genome evolution across 1,011 Saccharomyces cerevisiae isolates. Nature. 2018;556:339–344. doi: 10.1038/s41586-018-0030-5. [DOI] [PMC free article] [PubMed] [Google Scholar]
  92. Petrucelli L, Dawson TM. Mechanism of neurodegenerative disease: role of the ubiquitin proteasome system. Annals of Medicine. 2004;36:315–320. doi: 10.1080/07853890410031948. [DOI] [PubMed] [Google Scholar]
  93. Pettersen EF, Goddard TD, Huang CC, Meng EC, Couch GS, Croll TI, Morris JH, Ferrin TE. UCSF chimerax: structure visualization for researchers, educators, and developers. Protein Science. 2021;30:70–82. doi: 10.1002/pro.3943. [DOI] [PMC free article] [PubMed] [Google Scholar]
  94. Pohl C, Dikic I. Cellular quality control by the ubiquitin-proteasome system and autophagy. Science. 2019;366:818–822. doi: 10.1126/science.aax3769. [DOI] [PubMed] [Google Scholar]
  95. Renganaath K, Cheung R, Day L, Kosuri S, Kruglyak L, Albert FW. Systematic identification of cis-regulatory variants that cause gene expression differences in a yeast cross. eLife. 2020;9:e62669. doi: 10.7554/eLife.62669. [DOI] [PMC free article] [PubMed] [Google Scholar]
  96. Schmidt M, Finley D. Regulation of proteasome activity in health and disease. Biochimica et Biophysica Acta - Molecular Cell Research. 2014;1843:13–25. doi: 10.1016/j.bbamcr.2013.08.012. [DOI] [PMC free article] [PubMed] [Google Scholar]
  97. Schwanhäusser B, Busse D, Li N, Dittmar G, Schuchhardt J, Wolf J, Chen W, Selbach M. Global quantification of mammalian gene expression control. Nature. 2011;473:337–342. doi: 10.1038/nature10098. [DOI] [PubMed] [Google Scholar]
  98. Schwartz AL, Ciechanover A. The ubiquitin-proteasome pathway and pathogenesis of human diseases. Annual Review of Medicine. 1999;50:57–74. doi: 10.1146/annurev.med.50.1.57. [DOI] [PubMed] [Google Scholar]
  99. Shaner NC, Campbell RE, Steinbach PA, Giepmans BNG, Palmer AE, Tsien RY. Improved monomeric red, orange and yellow fluorescent proteins derived from Discosoma sp. red fluorescent protein. Nature Biotechnology. 2004;22:1567–1572. doi: 10.1038/nbt1037. [DOI] [PubMed] [Google Scholar]
  100. Shringarpure R, Davies KJA. Protein turnover by the proteasome in aging and disease. Free Radical Biology & Medicine. 2002;32:1084–1089. doi: 10.1016/s0891-5849(02)00824-9. [DOI] [PubMed] [Google Scholar]
  101. Sikorski RS, Hieter P. A system of shuttle vectors and yeast host strains designed for efficient manipulation of DNA in Saccharomyces cerevisiae. Genetics. 1989;122:19–27. doi: 10.1093/genetics/122.1.19. [DOI] [PMC free article] [PubMed] [Google Scholar]
  102. Sommer T, Jentsch S. A protein translocation defect linked to ubiquitin conjugation at the endoplasmic reticulum. Nature. 1993;365:176–179. doi: 10.1038/365176a0. [DOI] [PubMed] [Google Scholar]
  103. Sontag EM, Vonk WIM, Frydman J. Sorting out the trash: the spatial nature of eukaryotic protein quality control. Current Opinion in Cell Biology. 2014;26:139–146. doi: 10.1016/j.ceb.2013.12.006. [DOI] [PMC free article] [PubMed] [Google Scholar]
  104. Stolzing A, Grune T. The proteasome and its function in the ageing process. Clinical and Experimental Dermatology. 2001;26:566–572. doi: 10.1046/j.1365-2230.2001.00867.x. [DOI] [PubMed] [Google Scholar]
  105. Swanson R, Locher M, Hochstrasser M. A conserved ubiquitin ligase of the nuclear envelope/endoplasmic reticulum that functions in both ER-associated and MATalpha2 repressor degradation. Genes & Development. 2001;15:2660–2674. doi: 10.1101/gad.933301. [DOI] [PMC free article] [PubMed] [Google Scholar]
  106. Varshavsky A. Naming a targeting signal. Cell. 1991;64:13–15. doi: 10.1016/0092-8674(91)90202-a. [DOI] [PubMed] [Google Scholar]
  107. Varshavsky A. Ubiquitin fusion technique and related methods. Methods in Enzymology. 2005;399:777–799. doi: 10.1016/S0076-6879(05)99051-4. [DOI] [PubMed] [Google Scholar]
  108. Varshavsky A. The N-end rule pathway and regulation by proteolysis. Protein Science. 2011;20:1298–1345. doi: 10.1002/pro.666. [DOI] [PMC free article] [PubMed] [Google Scholar]
  109. Wach A, Brachat A, Pöhlmann R, Philippsen P. New heterologous modules for classical or PCR-based gene disruptions in Saccharomyces cerevisiae. Yeast. 1994;10:1793–1808. doi: 10.1002/yea.320101310. [DOI] [PubMed] [Google Scholar]
  110. Waite KA, De-La Mota-Peynado A, Vontz G, Roelofs J. Starvation induces proteasome autophagy with different pathways for core and regulatory particles. The Journal of Biological Chemistry. 2016;291:3239–3253. doi: 10.1074/jbc.M115.699124. [DOI] [PMC free article] [PubMed] [Google Scholar]
  111. Wang L, Wang S, Li W. RSeQC: quality control of RNA-seq experiments. Bioinformatics. 2012;28:2184–2185. doi: 10.1093/bioinformatics/bts356. [DOI] [PubMed] [Google Scholar]
  112. Ward AC. Rapid analysis of yeast transformants using colony-PCR. BioTechniques. 1992;13:350. [PubMed] [Google Scholar]
  113. Wickner RB. MKT1, a nonessential Saccharomyces cerevisiae gene with a temperature-dependent effect on replication of M2 double-stranded RNA. Journal of Bacteriology. 1987;169:4941–4945. doi: 10.1128/jb.169.11.4941-4945.1987. [DOI] [PMC free article] [PubMed] [Google Scholar]
  114. Xia K, Guo H, Hu Z, Xun G, Zuo L, Peng Y, Wang K, He Y, Xiong Z, Sun L, Pan Q, Long Z, Zou X, Li X, Li W, Xu X, Lu L, Liu Y, Hu Y, Tian D, Long L, Ou J, Liu Y, Li X, Zhang L, Pan Y, Chen J, Peng H, Liu Q, Luo X, Su W, Wu L, Liang D, Dai H, Yan X, Feng Y, Tang B, Li J, Miedzybrodzka Z, Xia J, Zhang Z, Luo X, Zhang X, St Clair D, Zhao J, Zhang F. Common genetic variants on 1p13.2 associate with risk of autism. Molecular Psychiatry. 2014;19:1212–1219. doi: 10.1038/mp.2013.146. [DOI] [PubMed] [Google Scholar]
  115. Xie Y, Varshavsky A. Rpn4 is a ligand, substrate, and transcriptional regulator of the 26S proteasome: a negative feedback circuit. PNAS. 2001;98:3056–3061. doi: 10.1073/pnas.071022298. [DOI] [PMC free article] [PubMed] [Google Scholar]
  116. Yen H-CS, Xu Q, Chou DM, Zhao Z, Elledge SJ. Global protein stability profiling in mammalian cells. Science. 2008;322:918–923. doi: 10.1126/science.1160489. [DOI] [PubMed] [Google Scholar]
  117. Yu H, Singh Gautam AK, Wilmington SR, Wylie D, Martinez-Fonts K, Kago G, Warburton M, Chavali S, Inobe T, Finkelstein IJ, Babu MM, Matouschek A. Conserved sequence preferences contribute to substrate recognition by the proteasome. The Journal of Biological Chemistry. 2016;291:14526–14539. doi: 10.1074/jbc.M116.727578. [DOI] [PMC free article] [PubMed] [Google Scholar]
  118. Zenker M, Mayerle J, Lerch MM, Tagariello A, Zerres K, Durie PR, Beier M, Hülskamp G, Guzman C, Rehder H, Beemer FA, Hamel B, Vanlieferinghen P, Gershoni-Baruch R, Vieira MW, Dumic M, Auslender R, Gil-da-Silva-Lopes VL, Steinlicht S, Rauh M, Shalev SA, Thiel C, Ekici AB, Winterpacht A, Kwon YT, Varshavsky A, Reis A. Deficiency of UBR1, a ubiquitin ligase of the N-end rule pathway, causes pancreatic dysfunction, malformations and mental retardation (johanson-blizzard syndrome) Nature Genetics. 2005;37:1345–1350. doi: 10.1038/ng1681. [DOI] [PubMed] [Google Scholar]
  119. Zenker M, Mayerle J, Reis A, Lerch MM. Genetic basis and pancreatic biology of johanson-blizzard syndrome. Endocrinology and Metabolism Clinics of North America. 2006;35:243–253. doi: 10.1016/j.ecl.2006.02.013. [DOI] [PubMed] [Google Scholar]
  120. Zheng C, Geetha T, Babu JR. Failure of ubiquitin proteasome system: risk for neurodegenerative diseases. Neuro-Degenerative Diseases. 2014;14:161–175. doi: 10.1159/000367694. [DOI] [PubMed] [Google Scholar]
  121. Zwolshen JH, Bhattacharjee JK. Genetic and biochemical properties of thialysine-resistant mutants of Saccharomyces cerevisiae. Journal of General Microbiology. 1981;122:281–287. doi: 10.1099/00221287-122-2-281. [DOI] [PubMed] [Google Scholar]

Editor's evaluation

Magnus Nordborg 1

The authors use an elegant experimental design to study genetic variation in the ubiquitin-proteasome degradation system in yeast. They identify a large number of QTLs for naturally occurring variation, and they elucidate the causal variants and likely functional mechanisms of several of these. The paper illustrates an innovative new approach to high-throughput QTL mapping for specific molecular processes.

Decision letter

Editor: Magnus Nordborg1
Reviewed by: Magnus Nordborg2

Our editorial process produces two outputs: (i) public reviews designed to be posted alongside the preprint for the benefit of readers; (ii) feedback on the manuscript for the authors, including requests for revisions, shown below. We also include an acceptance summary that explains what the editors found interesting or important about the work.

Decision letter after peer review:

Thank you for submitting your article "Variation in Ubiquitin System Genes Creates Substrate-Specific Effects on Proteasomal Protein Degradation" for consideration by eLife. Your article has been reviewed by 4 peer reviewers, , including Magnus Nordborg as the Reviewing Editor and Reviewer #1 and the evaluation has been overseen by David Ron as the Senior Editor.

The reviewers have discussed their reviews with one another and agreed this work is elegant, and absolutely deserves to be published. The Reviewing Editor has drafted this to help you prepare a revised submission – please also see the individual reviews for further suggestions for improvements.

There is an overall need to quantify several statements (see individual reviews). This should be straightforward and requires no additional experiments.

In addition, we would like to encourage you to put a bit more thought into the evolutionary analysis. For example, the issue of pleiotropy could be discussed further (as noted by Reviewer 1), and so could the persistent shifts in effect between the strains (Reviewer 4). It may be worth considering Hunter Fraser's use of sign tests to detect selection, e.g.,

https://www.pnas.org/doi/10.1073/pnas.0912245107 in this context.

Reviewer #1 (Recommendations for the authors):

I trust more mechanistic reviewers to comment on your experiments. From my point of view, I have only a few suggestions for improvements:

1) You argue (in connection with Figure 2)1 that "many QTLs were found only for individual pathways", but this is surely sensitive to significance thresholds. A more sophisticated analysis of pleiotropy would be in order.

2) You never quantify how much of the total variation your QTL explains.

3) You never quantify how much of each QTL your identified causal SNPs explain.

4) When dissecting your UBR1 alleles, would be nice to pay homage to this classic:

Laurie-Ahlberg, C. C., and L. F. Stam. 1987. "Use of P-Element-Mediated Transformation to Identify the Molecular Basis of Naturally Occurring Variants Affecting Adh Expression in Drosophila melanogaster." Genetics 115 (1): 129-40.

Reviewer #2 (Recommendations for the authors):

The only thing I would have appreciated is the validation of some of their findings using orthogonal approaches. There are a lot of readily established biochemical assays they can use to support their claims.

Reviewer #3 (Recommendations for the authors):

1. Figure 1 was super helpful in explaining the question and the methods used in this paper.

2. In this sentence, I am not sure that the second portion follows from the first: "These results are consistent with our observation that RM had higher UPS activity for 15 of 20 N-degrons (Figure 1D, Supplementary Table 1), suggesting that the QTLs we have mapped underlie a substantial portion of the heritable UPS activity difference between BY and RM." Maybe you missed the majority of the genetic variation that influences UPS activity, but you got lucky and the portion you detected just happened to be consistent with the overall trend. I think the authors need a different type of analysis to draw a conclusion about whether they sampled "a substantial portion" of heritable differences or whether there is missing heritability here. Can the authors use their data to calculate the narrow sense heritability and then report how much of that heritability is explained by the detected variants?

3. Perhaps reconsider using the term N-degron in the abstract. It's a bit specific. The intro and figures do a great job at defining an N-degron.

4. I did not know that each of the 20 amino acids is dealt with by one of two UPS systems. This was eventually clear in figure 1B where it seems Trp and Met N-degrons are degraded by different systems. Perhaps make this clearer in the text where N-degrons are described and defined.

5. I was mildly confused by the "italics vs plain" statement in figure 3D. Is it intended for only this panel or for the whole figure, or for the whole paper?

6. Some of the focus of the manuscript might be shifted away from general "straw man" questions, like whether this trait is mendelian and controlled by large effect variants (intro lines 55 – 66). The discovery of smaller effect variants is presented as a major finding, but it seems obvious that these exist. Perhaps instead focus on a more quantitative analysis of the power of the current study. How much heritability is explained by previously known genes or variants, and how much additional heritability is explained by the previously unknown genes/variants detected here? Even if a heritability analysis is not feasible, it think a shift in focus from a qualitative statement – we detect small effect variants – to a more quantitative or nuanced statement, would be appropriate.

Reviewer #4 (Recommendations for the authors):

I'm interested in the overall pattern that BY seems to have systematically lower UPS activity than RM. BY carries a rare allele at 3 out of the 4 examples in Figure S5, which hints that perhaps there has been an adaptation of BY for lower UPS. It would be interesting to explore this hypothesis further, or at least to discuss it. Has there been an overall adaptation for BY to be less transcriptionally active, coupled with lower degradation rates? Perhaps this might be explored using the previous data sets on mRNA in these lines – presumably, changes in transcription under this model could be either cis or trans. (That analysis is distinct from the analysis reported at L312, which focuses on specific trans-effects of the UPS variants.) And perhaps the authors may have other ideas as well to explore the evolutionary questions.

The paper reports 149 loci, but if I am reading this correctly it looks like this may double-count shared loci. It would be nice to discuss a bit more about likely sharing (ie pleiotropy) of the hits. (It's also a bit hard to see from 2B which of these are likely shared.)

L143: it should be possible to make this statement quantitative.

Para 303, 312: It seems likely that you may be looking for an effect that is small at individual genes but widespread. I would suggest looking more carefully at the overall distribution: eg in the protein data is there an upward shift in the mean? You could also use a more sophisticated method like ashR to study the distribution of changes. For Figure 5 you could also consider adding information about global patterns in means and correlations, which are difficult to read from these plots.

eLife. 2022 Oct 11;11:e79570. doi: 10.7554/eLife.79570.sa2

Author response


Reviewer #1 (Recommendations for the authors):

I trust more mechanistic reviewers to comment on your experiments. From my point of view, I have only a few suggestions for improvements:

1) You argue (in connection with Figure 2)1 that "many QTLs were found only for individual pathways", but this is surely sensitive to significance thresholds. A more sophisticated analysis of pleiotropy would be in order.

The revised manuscript includes an expanded analysis of pleiotropy and how the significance threshold used for QTL detection influences QTL pathway specificity. The high degree of QTL pathway specificity we observed was robust to these additional analyses.

As described in the revised manuscript (pg. 6, para. 3, line 180), we considered a QTL's effect common to multiple N-degrons when QTL peaks for distinct N-degrons were within 100 kb of each other and had the same direction of effect. Using these criteria, the 149 QTLs detected for the 20 N-degrons correspond to 35 distinct QTL regions that each affect between 1 and 11 N-degrons. At the LOD score threshold of 4.5 used in the manuscript, 23 of these 35 QTL regions specifically affected N-degrons from an individual pathway in the N-end Rule, with 11 regions affecting only Ac/N-degrons and 12 regions affecting only Arg/N-degrons. Five of the 23 distinct, pathway-specific QTL regions exclusively affected individual N-degrons. The remaining 12 of the 35 distinct QTL regions have the same direction of effect for at least one reporter each from the Arg/N-end and Ac/N-end pathways. The revised manuscript includes additional text in the Results section (pg. 6, para. 3) discussing pleiotropy among QTLs and a new table (Figure 2—figure supplement 2), which provides detailed information on the 35 distinct QTL regions, including their pathway- and N-degron specificity.

To understand the influence of significance threshold on these results, we analyzed how relaxing the LOD score threshold affected the number of pathway- or N-degron-specific QTL regions. The results of this analysis are summarized in Author response table 1. Supplementary file 2 shows the LOD score and allele frequency difference traces for all 20 N-degrons at the 23 N-end Rule pathway- or N-degron-specific QTL regions at multiple significance thresholds. Author response table 2 and Supplementary file 2 provide detailed information on how the N-end Rule pathway- and N-degron-specificity of each of the 35 distinct QTL regions is influenced by LOD score threshold.

Author response table 1. Influence of LOD score significance threshold on the fraction of N-end rule pathway- and N-degron-specific QTLs.

LOD 4.5 LOD 3.5 LOD 2.5
N-end Rule Pathway-specific QTLs 23 / 35 (66%) 19 / 35 (54%) 16 / 35 (46%)
N-degron-specific QTLs 5 / 35 (14%) 3 / 35 (9%) 3 / 35 (9%)

Author response table 2. Shared QTL Regions and N-degrons Affected.

N QTL chr QTL_CI_left QTL_peak QTL_CI_right LOD RM_AFD N. N-degrons Affected N-degrons Affected Pathway-Specific (LOD 4.5) Pathway-Specific (LOD 3.5) Pathway-Specific (LOD 2.5) N-degron-Specific (LOD 4.5) N-degron-Specific (LOD 3.5) N-degron-Specific (LOD 2.5)
1 chr_1_a 1 5021 39400 52122 45 -0.325 7 Ala, Cys, Gly, Pro, Ser, Thr, Val Ac/N-end Ac/N-end Ac/N-end no no no
2 chr_2_a 2 473450 515516 573934 11.1 0.13 3 Arg, Lys, Phe Arg/N-end Arg/N-end no no no no
3 chr_2_b 2 534970 585630 640010 8.9 0.112 5 Ala, Arg, Asp, Lys, Met no no no no no no
4 chr_4_a 4 29500 87662 116525 7.3 -0.108 4 Gln, Glu, Gly, His no no no no no no
5 chr_4_b 4 273175 327300 362600 17.5 0.152 2 Trp, Tyr Arg/N-end Arg/N-end no no no no
6 chr_4_c 4 376817 425550 463567 14.3 0.156 6 Ala, Phe, Pro, Ser, Thr, Trp no no no no no no
7 chr_4_d 4 392275 498200 557000 8.2 -0.107 2 Asn, Asp Arg/N-end Arg/N-end Arg/N-end no no no
8 chr_5_a 5 351400 372136 396250 37.5 0.248 7 Ala, Cys, Gly, Pro, Ser, Thr, Val Ac/N-end Ac/N-end Ac/N-end no no no
13 chr_7_a 7 42075 62775 103225 12.5 -0.155 2 Asn, Thr no no no no no no
9 chr_7_b 7 98178 132779 172279 14.3 -0.156 7 Ala, Cys, Gly, Pro, Ser, Thr, Val Ac/N-end no no no no no
10 chr_7_c 7 389941 418641 451318 18 0.167 11 Ala, Asn, Cys, Gly, Phe, Pro, Ser, Thr, Trp, Tyr, Val no no no no no no
11 chr_7_d 7 841300 869700 899934 15.2 -0.157 3 Arg, Asn, Asp Arg/N-end Arg/N-end Arg/N-end no no no
12 chr_7_e 7 856120 870980 883830 107.5 0.409 5 His, Leu, Phe, Trp, Tyr Arg/N-end Arg/N-end Arg/N-end no no no
15 chr_8_a 8 50150 98050 127300 6 -0.101 1 Tyr Arg/N-end Arg/N-end Arg/N-end Tyr Tyr Tyr
14 chr_8_b 8 118875 148400 193675 8.4 0.104 2 Asp, Lys Arg/N-end Arg/N-end Arg/N-end no no no
17 chr_9_a 9 94550 130775 168000 5.1 0.099 2 His, Lys Arg/N-end Arg/N-end Arg/N-end no no no
16 chr_9_b 9 267641 291492 315417 23.5 0.195 6 Ala, Gly, Pro, Ser, Thr, Val Ac/N-end Ac/N-end Ac/N-end no no no
18 chr_10_a 10 323900 345482 371186 21.9 0.182 11 Ala, Arg, Cys, Gln, Gly, Lys, Phe, Pro, Ser, Trp, Tyr no no no no no no
19 chr_10_b 10 589770 615660 641600 25.1 0.192 5 Ala, Arg, Asn, Cys, Gln no no no no no no
21 chr_11_a 11 118750 156450 173600 16 0.152 1 His Arg/N-end Arg/N-end Arg/N-end His no no
20 chr_11_b 11 291030 339790 388750 8.4 -0.106 5 Ala, Arg, Cys, Phe, Thr no no no no no no
22 chr_12_a 12 157200 197050 226850 8.2 0.126 2 Ala, Pro Ac/N-end no no no no no
23 chr_12_b 12 644414 657241 674723 42.1 -0.25 11 Ala, Cys, Gly, Met, Phe, Pro, Ser, Thr, Trp, Tyr, Val no no no no no no
24 chr_12_c 12 639637 674125 702650 9.7 0.108 4 Asp, Gly, His, Ile Arg/N-end Arg/N-end Arg/N-end no no no
25 chr_13_a 13 0 24650 58950 9.6 -0.162 1 Ala Ac/N-end Ac/N-end no Ala no no
27 chr_13_b 13 31800 56525 77675 14.9 0.145 2 His, Phe Arg/N-end Arg/N-end Arg/N-end no no no
26 chr_13_c 13 265350 296167 331584 16.5 0.17 6 Arg, Asn, Asp, Glu, His, Lys Arg/N-end no no no no no
29 chr_14_a 14 450731 465394 482912 89.3 -0.342 8 His, Leu, Lys, Met, Phe, Pro, Trp, Tyr no no no no no no
28 chr_14_b 14 449725 470850 504025 86.7 0.336 2 Asp, Thr no no no no no no
31 chr_15_a 15 133625 163862 184831 20.2 -0.183 8 Arg, Asn, Asp, Glu, His, Lys, Met, Phe no no no no no no
30 chr_15_b 15 342850 388050 431300 10.1 0.116 3 Ala, Pro, Thr Ac/N-end no no no no no
32 chr_15_c 15 525200 561525 591775 10.6 0.128 2 Ser, Thr Ac/N-end Ac/N-end Ac/N-end no no no
33 chr_15_d 15 518550 569750 594650 8.8 -0.096 1 Cys Ac/N-end Ac/N-end Ac/N-end Cys Cys Cys
34 chr_16_a 16 166030 194830 222070 11.6 0.121 5 Ala, Gly, Pro, Thr, Val Ac/N-end Ac/N-end Ac/N-end no no no
35 chr_16_b 16 375050 403350 428950 24.5 -0.22 1 Pro Ac/N-end Ac/N-end Ac/N-end Pro Pro Pro

Abbreviations: "chr": chromosome, "CI": confidence interval, "RM_AFD": RM allele frequency difference (high – low UPS activity pools).

Generally, altering the significance threshold does not affect the conclusions that at least half of the detected QTLs are specific to an individual N-end Rule pathway and that approximately 10% of QTLs are specific to individual N-degrons. In the revised Results section (pg. 6, para. 3, line 18), the qualitative statement mentioned by the reviewer is replaced with the quantitative information in Author response table 1 for the 4.5 LOD threshold column.

We note that N-end Rule pathway-specificity for 3 QTL regions (IVb, IXa, and XVc) is subject to the following considerations that may not be immediately obvious from reviewing Supplementary file 2.

The Arg/N-end-specific chromosome IVb and IXa QTL regions overlap the leftmost shoulders of QTLs detected with Ac/N-degrons, creating the appearance that these QTLs are not Arg/N-degron-specific when plotted as in Supplementary file 2. However, the peaks of these Arg/N-degron and Ac/N-degron QTLs are not within 100 kb and their confidence intervals do not overlap. We therefore interpret their effects as pathway-specific.

The Ac/N-degron-specific chromosome XVc QTL contains peaks from both glutamate Arg/N-degron replicates. However, the direction of effect of these two peaks is not consistent between replicates (the only such instance in our dataset). Because we could not unambiguously assign a direction of effect for the glutamate N-degron at this region, we do not include it in our set of QTLs and the chromosome XVc region is deemed to be Ac/N-degron-specific. In the revised manuscript, we report the chromosome IVb, IXa, and XVc QTL regions as pathway-specific.

We also note that the pathway-specific QTL regions displayed on pages 7 and 8 (QTL regions containing UBR1) as well as 20 and 21 of Supplementary file 2 are highly overlapping. However, these overlapping regions have opposing directions of effect on the sets of reporters they affect. We, therefore, continue to report these QTL regions as distinct in the revised manuscript, an interpretation supported by our fine-mapping data (see e.g., Figure 3C).

2) You never quantify how much of the total variation your QTL explains.

Our bulk segregant analysis QTL mapping method is based on comparing allele frequencies obtained from whole-genome sequencing pools of cells with extreme phenotypes. The genotypes of individual segregants, which would be needed for calculating explained variance, are not ascertained using this approach. Thus, it is not readily possible to calculate the proportion of variance explained by our QTLs.

3) You never quantify how much of each QTL your identified causal SNPs explain.

Author response table 3 displays the variance in ubiquitin-proteasome system (UPS) activity explained by each tested causal gene allele and variant.

Author response table 3. Variance Explained by Causal Alleles and Variants.

Gene Allele N-degron Variance Explained
DOA10 K1012N Thr 0.04
DOA10 Q410E Thr 0.075
DOA10 RM_full Thr 0.911
DOA10 Y1186F Thr 0.597
DOA10 K1012N Gly 0.608
DOA10 Q410E Gly 0.74
DOA10 RM_full Gly 0.879
DOA10 Y1186F Gly 0.77
NTA1 D111E Asn 0.204
NTA1 E129G Asn 0.764
NTA1 RM_full Asn 0.789
NTA1 RM_pr Asn -0.014
UBC6 D229G Ala 0.717
UBC6 RM_full Ala 0.666
UBC6 RM_pr Ala -0.027
UBC6 RM_term Ala 0.183
UBC6 D229G Thr 0.762
UBC6 RM_full Thr 0.928
UBC6 RM_pr Thr 0.008
UBC6 RM_term Thr 0.006
UBR1 RM_causal Phe 0.9
UBR1 RM_full Phe 0.917
UBR1 RM_non Phe -0.019
UBR1 RM_pr Phe 0.818
UBR1 RM_causal Trp 0.937
UBR1 RM_full Trp 0.975
UBR1 RM_non Trp 0.059
UBR1 RM_pr Trp 0.934
UBR1 RM_full Asn 0.438
UBR1 RM_ORF Asn 0.134
UBR1 RM_pr Asn -0.012
UBR1 RM_term Asn 0.064
UBR1 RM_full Asp 0.403
UBR1 RM_ORF Asp 0.101
UBR1 RM_pr Asp 0.097
UBR1 RM_term Asp 0.048
UBR1 RM_full Phe 0.866
UBR1 RM_ORF Phe 0.343
UBR1 RM_pr Phe 0.316
UBR1 RM_term Phe -0.031
UBR1 RM_full Trp 0.964
UBR1 RM_ORF Trp 0.813
UBR1 RM_pr Trp 0.928
UBR1 RM_term Trp 0.028

We note that these estimates are obtained from experiments in near-isogenic strains that differ only at the tested causal gene allele or variant. The fraction of variance explained is thus inflated relative to what would be observed in the segregant populations used for QTL mapping and should not be used to estimate the variance explained by a QTL region. Therefore, we do not include these estimates in the revised manuscript.

4) When dissecting your UBR1 alleles, would be nice to pay homage to this classic: Laurie-Ahlberg, C. C., and L. F. Stam. 1987. "Use of P-Element-Mediated Transformation to Identify the Molecular Basis of Naturally Occurring Variants Affecting Adh Expression in Drosophila melanogaster." Genetics 115 (1): 129-40.

We thank the reviewer for the suggestion to include this important work on identifying causal variants for enzyme activity in highly polymorphic genomic regions. The work appears as reference 62 (pg. 8, para 3, line 218) in the revised manuscript.

Reviewer #2 (Recommendations for the authors): The only thing I would have appreciated is the validation of some of their findings using orthogonal approaches. There are a lot of readily established biochemical assays they can use to support their claims.

Previously published theoretical and empirical observations have demonstrated that tandem fluorescent timers (TFTs) provide precise and sensitive measures of protein degradation kinetics. In particular, the TFT system has been extensively used to measure differences in the degradation rate of UPS N-end Rule substrates (Khmelinskii et al., 2012, Khmelinskii and Knop, 2014, Kats et al., 2018). Given the well-established validity of the TFT system for measuring N-end Rule activity and the comparatively lower precision, sensitivity, and throughput of conventional biochemical measurements of protein degradation (in particular, pulse-chase Western blotting and cycloheximide chase analysis [Kong et al., 2021]), we argue that further experimentation is not needed to establish the claims made in our work.

Reviewer #3 (Recommendations for the authors):

1. Figure 1 was super helpful in explaining the question and the methods used in this paper.

Thank you!

2. In this sentence, I am not sure that the second portion follows from the first: "These results are consistent with our observation that RM had higher UPS activity for 15 of 20 N-degrons (Figure 1D, Supplementary Table 1), suggesting that the QTLs we have mapped underlie a substantial portion of the heritable UPS activity difference between BY and RM." Maybe you missed the majority of the genetic variation that influences UPS activity, but you got lucky and the portion you detected just happened to be consistent with the overall trend. I think the authors need a different type of analysis to draw a conclusion about whether they sampled "a substantial portion" of heritable differences or whether there is missing heritability here. Can the authors use their data to calculate the narrow sense heritability and then report how much of that heritability is explained by the detected variants?

As noted in our response to Reviewer 1, because of the pooled nature of our genetic mapping method, it is not readily possible to calculate heritability from our QTL mapping data. Author response table 3 presents the variance explained by the tested causal alleles and variants. However, as noted in our response to Reviewer 1, these estimates are inflated because they are derived from near-isogenic cell populations that differ only at the tested alleles / variants. Given these considerations, the second clause in the sentence mentioned by the reviewer has been removed from the revised manuscript.

3. Perhaps reconsider using the term N-degron in the abstract. It's a bit specific. The intro and figures do a great job at defining an N-degron.

The term “N-degron” has been removed from the abstract and replaced with more general terms.

4. I did not know that each of the 20 amino acids is dealt with by one of two UPS systems. This was eventually clear in figure 1B where it seems Trp and Met N-degrons are degraded by different systems. Perhaps make this clearer in the text where N-degrons are described and defined.

We thank the reviewer for this suggestion to improve the manuscript’s clarity. The revised introduction indicates that the N-end Rule contains two distinct targeting complexes when introducing the N-end Rule (pg. 3, para. 2, line 105) and references Figure 1, which illustrates the two pathways of the N-end rule.

5. I was mildly confused by the "italics vs plain" statement in figure 3D. Is it intended for only this panel or for the whole figure, or for the whole paper?

The “italics vs. plain” statement is intended only for panels 3D and 4B/F/J. To improve clarity, we have added additional annotation to these panels.

6. Some of the focus of the manuscript might be shifted away from general "straw man" questions, like whether this trait is mendelian and controlled by large effect variants (intro lines 55 – 66). The discovery of smaller effect variants is presented as a major finding, but it seems obvious that these exist. Perhaps instead focus on a more quantitative analysis of the power of the current study. How much heritability is explained by previously known genes or variants, and how much additional heritability is explained by the previously unknown genes/variants detected here? Even if a heritability analysis is not feasible, it think a shift in focus from a qualitative statement – we detect small effect variants – to a more quantitative or nuanced statement, would be appropriate.

The revised manuscript provides a more nuanced introduction to the genetics of UPS activity, in particular, emphasizing the expectation that UPS activity, like most traits, is genetically complex. We agree with the reviewer that the relative contributions of individual QTLs to the heritability of UPS activity would be an interesting and informative analysis. However, as noted in our response to Reviewer 1, it is not readily feasible to perform such an analysis. Instead, as suggested by the reviewer, several qualitative statements have been replaced by quantitative descriptions. In particular, the qualitative statement mentioned by the reviewer is replaced with a quantitative statement regarding QTL effect sizes (pg. 6, para. 1, line 164).

Reviewer #4 (Recommendations for the authors):

Specific comments:

I'm interested in the overall pattern that BY seems to have systematically lower UPS activity than RM. BY carries a rare allele at 3 out of the 4 examples in Figure S5, which hints that perhaps there has been an adaptation of BY for lower UPS. It would be interesting to explore this hypothesis further, or at least to discuss it. Has there been an overall adaptation for BY to be less transcriptionally active, coupled with lower degradation rates? Perhaps this might be explored using the previous data sets on mRNA in these lines – presumably, changes in transcription under this model could be either cis or trans. (That analysis is distinct from the analysis reported at L312, which focuses on specific trans-effects of the UPS variants.) And perhaps the authors may have other ideas as well to explore the evolutionary questions.

We thank the reviewer for these interesting suggestions, which we have addressed with a series of new analyses. In brief, we did not detect evidence for lineage-specific selection on UPS gene expression using eQTL data or on individual N-end reporters.

To explore whether the consistent differences in UPS activity between the BY and RM strains might reflect adaptive changes in these lineages, we performed several additional analyses. We first applied the sign test (Fraser et al., 2010; https://doi.org/10.1073/pnas.0912245107) to a recent comprehensive BY / RM eQTL mapping dataset, which became available after the original sign test publication. This newer eQTL dataset comprises 36,498 eQTLs mapped for 5,643 genes in a panel of 1,000 recombinant offspring from the BY / RM cross (Albert et al., 2018; https://doi.org/10.7554/eLife.35471). The results of this analysis are presented in Author response table 4. We performed the analysis at 11 different LOD thresholds (ranging from 2.5 to 50) to examine the influence of QTL effect size on the results. Across all genes, we do not find evidence for lineage-specific selection except at LOD thresholds of 40 and 45. Given the large fraction of eQTLs that are excluded at these high thresholds (94.6% and 95.3%), the large fraction of genes excluded (70.5% and 73.8%), and especially the marginally significant p-values (0.038 and 0.026) obtained at these two thresholds, we conclude that there is, at best, limited evidence from the sign test for lineage-specific selection on overall mRNA transcript abundance in the BY / RM cross. Future work is needed to reconcile discrepancies in the results of the sign test as obtained in different eQTL datasets from the same cross.

Author response table 4. Results of the Sign Test for Lineage-Specific Selection Applied to All BY / RM eQTLs.

LOD n_eQTLs n_genes n_pairs reinf_BY_up reinf_RM_up oppos_BY_up oppos_RM_up excess_reinforcing_pairs chi_sq_p
2.4 36498 5643 2845 627 757 952 509 –14 0.818
5 19694 5372 2081 449 577 698 357 19 0.701
10 9379 4514 1124 225 320 404 175 4.63 0.934
15 6125 3686 745 154 202 279 110 2.24 0.992
20 4452 3076 480 108 128 179 65 18.2 0.431
25 3458 2609 336 73 95 127 41 20.6 0.276
30 2788 2224 230 54 65 85 26 22.6 0.145
35 2310 1912 178 44 48 68 18 20 0.144
40 1963 1664 140 35 41 53 11 24.3 0.0375
45 1706 1474 110 28 33 42 7 22.9 0.0261
50 1475 1299 85 21 27 31 6 17.9 0.0569

Abbreviations: “LOD”: LOD score threshold for calling cis / trans eQTL pairs, "n_eQTLs": number of eQTLs, "n_genes": number of genes with an eQTL, "n_pairs": number of cis / trans eQTL pairs, "reinf_BY_up": number of cis / trans eQTLs where the BY allele of the cis and trans eQTLs increases expression (reinforcing pairs), "reinf_RM_up": number of cis / trans eQTLs where the RM allele of the cis and trans eQTLs increases expression (reinforcing pairs), "oppos_BY_up": number of cis / trans eQTLs where the BY allele of the cis eQTL increases expression and the RM allele of the trans eQTL increases expression (opposing pairs), "oppos_RM_up": number of cis / trans eQTLs where the RM allele of the cis eQTL increases expression and the BY allele of the trans eQTL increases expression (opposing pairs), "excess_reinforcing_pairs", the number of excess reinforcing eQTL pairs calculated as in Fraser et al., 2010, "chi_sq_p": p-value of the chi-square test for enrichment of reinforcing pairs.

As in the original manuscript sign test manuscript, we performed a gene ontology (GO) enrichment analysis on the sets of reinforcing cis / trans eQTL pairs from the set of eQTLs to determine whether such pairs are enriched for UPS genes. We were able to replicate the previously-described enrichment for genes of the ergosterol biosynthesis pathway (Fraser et al., 2010, Author response table 5). However, there was no enrichment for ubiquitin system or proteasome GO terms in the sets of reinforcing eQTL pairs at any of the tested LOD score thresholds (Author response table 5).

Author response table 5. Results of Gene Ontology Enrichment of All cis / trans eQTL pairs.

GOBPID p_value OddsRatio ExpCount Count Term category LOD
GO:0032197 2.8E-05 3 15.06 30 transposition; RNA-mediated all_reinforcing 2.426606803
GO:0015074 0.00046 3.1 9.87 20 DNA integration all_reinforcing 2.426606803
GO:0006487 0.00077 2.7 11.6 22 protein N-linked glycosylation all_reinforcing 2.426606803
GO:0046513 0.0012 10.7 2.22 7 ceramide biosynthetic process all_reinforcing 2.426606803
GO:0006458 0.0029 3.3 6.17 13 'de novo' protein folding all_reinforcing 2.426606803
GO:0018904 0.0037 Inf 0.99 4 ether metabolic process all_reinforcing 2.426606803
GO:0051131 0.0039 9.2 1.97 6 chaperone-mediated protein complex assembly all_reinforcing 2.426606803
GO:1901135 0.0045 1.4 92.08 114 carbohydrate derivative metabolic process all_reinforcing 2.426606803
GO:0070085 0.0063 1.8 21.23 32 glycosylation all_reinforcing 2.426606803
GO:0061077 0.0068 3.9 3.95 9 chaperone-mediated protein folding all_reinforcing 2.426606803
GO:0006672 0.007 5.4 2.72 7 ceramide metabolic process all_reinforcing 2.426606803
GO:0009101 0.007 1.8 19.75 30 glycoprotein biosynthetic process all_reinforcing 2.426606803
GO:0016226 0.0089 3.1 5.43 11 iron-sulfur cluster assembly all_reinforcing 2.426606803
GO:0031070 0.0093 6.1 2.22 6 intronic snoRNA processing all_reinforcing 2.426606803
GO:0034965 0.0093 6.1 2.22 6 intronic box C/D snoRNA processing all_reinforcing 2.426606803
GO:0032197 2.8E-05 3 15.06 30 transposition; RNA-mediated all_reinforcing 2.426606803
GO:0015074 0.00046 3.1 9.87 20 DNA integration all_reinforcing 2.426606803
GO:0006487 0.00077 2.7 11.6 22 protein N-linked glycosylation all_reinforcing 2.426606803
GO:0046513 0.0012 10.7 2.22 7 ceramide biosynthetic process all_reinforcing 2.426606803
GO:0006458 0.0029 3.3 6.17 13 'de novo' protein folding all_reinforcing 2.426606803
GO:0018904 0.0037 Inf 0.99 4 ether metabolic process all_reinforcing 2.426606803
GO:0051131 0.0039 9.2 1.97 6 chaperone-mediated protein complex assembly all_reinforcing 2.426606803
GO:1901135 0.0045 1.4 92.08 114 carbohydrate derivative metabolic process all_reinforcing 2.426606803
GO:0070085 0.0063 1.8 21.23 32 glycosylation all_reinforcing 2.426606803
GO:0061077 0.0068 3.9 3.95 9 chaperone-mediated protein folding all_reinforcing 2.426606803
GO:0006672 0.007 5.4 2.72 7 ceramide metabolic process all_reinforcing 2.426606803
GO:0009101 0.007 1.8 19.75 30 glycoprotein biosynthetic process all_reinforcing 2.426606803
GO:0016226 0.0089 3.1 5.43 11 iron-sulfur cluster assembly all_reinforcing 2.426606803
GO:0031070 0.0093 6.1 2.22 6 intronic snoRNA processing all_reinforcing 2.426606803
GO:0034965 0.0093 6.1 2.22 6 intronic box C/D snoRNA processing all_reinforcing 2.426606803
GO:0032197 4E-06 3.5 11.51 27 transposition; RNA-mediated all_reinforcing 5
GO:0015074 0.00011 3.7 7.48 18 DNA integration all_reinforcing 5
GO:0006458 0.001 3.9 4.79 12 'de novo' protein folding all_reinforcing 5
GO:0061077 0.0011 5.5 3.07 9 chaperone-mediated protein folding all_reinforcing 5
GO:0006487 0.0013 2.7 8.82 18 protein N-linked glycosylation all_reinforcing 5
GO:0018904 0.0013 Inf 0.77 4 ether metabolic process all_reinforcing 5
GO:0046513 0.0024 8.5 1.73 6 ceramide biosynthetic process all_reinforcing 5
GO:0070525 0.0024 8.5 1.73 6 tRNA threonylcarbamoyladenosine metabolic process all_reinforcing 5
GO:0051131 0.0038 10.6 1.34 5 chaperone-mediated protein complex assembly all_reinforcing 5
GO:0032259 0.0045 1.8 21.29 33 methylation all_reinforcing 5
GO:0046165 0.0052 2.2 11.32 20 alcohol biosynthetic process all_reinforcing 5
GO:0044281 0.0063 1.3 150.75 177 small molecule metabolic process all_reinforcing 5
GO:0006662 0.007 Inf 0.58 3 glycerol ether metabolic process all_reinforcing 5
GO:0002949 0.007 Inf 0.58 3 tRNA threonylcarbamoyladenosine modification all_reinforcing 5
GO:0033215 0.007 Inf 0.58 3 iron assimilation by reduction and transport all_reinforcing 5
GO:0046131 0.0083 3.8 3.26 8 pyrimidine ribonucleoside metabolic process all_reinforcing 5
GO:0019509 0.0086 7.1 1.53 5 L-methionine salvage from methylthioadenosine all_reinforcing 5
GO:0006694 0.0086 2.4 8.06 15 steroid biosynthetic process all_reinforcing 5
GO:0006696 0.0089 2.7 5.95 12 ergosterol biosynthetic process all_reinforcing 5
GO:0006672 0.0094 5.1 2.11 6 ceramide metabolic process all_reinforcing 5
GO:0019856 0.0094 5.1 2.11 6 pyrimidine nucleobase biosynthetic process all_reinforcing 5
GO:1902652 0.0098 2.5 6.71 13 secondary alcohol metabolic process all_reinforcing 5
GO:0032197 1.5E-06 4.5 6.44 20 transposition; RNA-mediated all_reinforcing 10
GO:0015074 7.7E-05 4.3 4.62 14 DNA integration all_reinforcing 10
GO:0018904 0.00022 Inf 0.49 4 ether metabolic process all_reinforcing 10
GO:0006662 0.0018 Inf 0.36 3 glycerol ether metabolic process all_reinforcing 10
GO:0006278 0.0021 2.7 6.8 15 RNA-dependent DNA biosynthetic process all_reinforcing 10
GO:1902047 0.0022 9.1 1.09 5 polyamine transmembrane transport all_reinforcing 10
GO:0046165 0.0041 2.6 6.56 14 alcohol biosynthetic process all_reinforcing 10
GO:0043605 0.0056 9.7 0.85 4 cellular amide catabolic process all_reinforcing 10
GO:0019509 0.0056 9.7 0.85 4 L-methionine salvage from methylthioadenosine all_reinforcing 10
GO:0072488 0.006 4.9 1.82 6 ammonium transmembrane transport all_reinforcing 10
GO:0015846 0.0064 6.1 1.34 5 polyamine transport all_reinforcing 10
GO:0006833 0.0065 21.8 0.49 3 water transport all_reinforcing 10
GO:0042044 0.0065 21.8 0.49 3 fluid transport all_reinforcing 10
GO:1902652 0.007 2.9 4.25 10 secondary alcohol metabolic process all_reinforcing 10
GO:0002098 0.007 3.9 2.43 7 tRNA wobble uridine modification all_reinforcing 10
GO:0006400 0.008 2.3 7.05 14 tRNA modification all_reinforcing 10
GO:0009067 0.0086 2.5 5.71 12 aspartate family amino acid biosynthetic process all_reinforcing 10
GO:0006696 0.0092 3 3.77 9 ergosterol biosynthetic process all_reinforcing 10
GO:0006694 0.0096 2.6 5.1 11 steroid biosynthetic process all_reinforcing 10
GO:0034220 0.0096 1.6 25.27 37 ion transmembrane transport all_reinforcing 10
GO:0032259 0.01 1.9 11.66 20 methylation all_reinforcing 10
GO:0015074 3.3E-05 5 3.71 13 DNA integration all_reinforcing 15
GO:0032197 6.2E-05 4 4.98 15 transposition; RNA-mediated all_reinforcing 15
GO:0046165 9.3E-05 4 4.59 14 alcohol biosynthetic process all_reinforcing 15
GO:0006278 0.00031 3.5 5.07 14 RNA-dependent DNA biosynthetic process all_reinforcing 15
GO:1902652 0.0011 3.9 3.32 10 secondary alcohol metabolic process all_reinforcing 15
GO:0019509 0.0011 18.7 0.59 4 L-methionine salvage from methylthioadenosine all_reinforcing 15
GO:0070525 0.0011 18.7 0.59 4 tRNA threonylcarbamoyladenosine metabolic process all_reinforcing 15
GO:0006696 0.0016 4 2.93 9 ergosterol biosynthetic process all_reinforcing 15
GO:0009086 0.0016 4 2.93 9 methionine biosynthetic process all_reinforcing 15
GO:0071267 0.0025 12.5 0.68 4 L-methionine salvage all_reinforcing 15
GO:0043102 0.0025 12.5 0.68 4 amino acid salvage all_reinforcing 15
GO:0043605 0.0025 12.5 0.68 4 cellular amide catabolic process all_reinforcing 15
GO:0016128 0.0033 3.5 3.22 9 phytosteroid metabolic process all_reinforcing 15
GO:0006811 0.0034 1.7 27.81 42 ion transport all_reinforcing 15
GO:0070900 0.0034 28 0.39 3 mitochondrial tRNA modification all_reinforcing 15
GO:1900864 0.0034 28 0.39 3 mitochondrial RNA modification all_reinforcing 15
GO:0006694 0.0049 3 4 10 steroid biosynthetic process all_reinforcing 15
GO:0009066 0.005 2.6 5.95 13 aspartate family amino acid metabolic process all_reinforcing 15
GO:0044107 0.0051 3.3 3.42 9 cellular alcohol metabolic process all_reinforcing 15
GO:0016125 0.0063 2.7 4.78 11 sterol metabolic process all_reinforcing 15
GO:1902047 0.0075 7.5 0.88 4 polyamine transmembrane transport all_reinforcing 15
GO:0006531 0.0079 14 0.49 3 aspartate metabolic process all_reinforcing 15
GO:0090646 0.0079 14 0.49 3 mitochondrial tRNA processing all_reinforcing 15
GO:0015804 0.0082 5.2 1.37 5 neutral amino acid transport all_reinforcing 15
GO:0016226 0.0082 5.2 1.37 5 iron-sulfur cluster assembly all_reinforcing 15
GO:0098655 0.0093 1.9 12.29 21 cation transmembrane transport all_reinforcing 15
GO:0015840 0.0095 Inf 0.2 2 urea transport all_reinforcing 15
GO:0034311 0.0095 Inf 0.2 2 diol metabolic process all_reinforcing 15
GO:0034312 0.0095 Inf 0.2 2 diol biosynthetic process all_reinforcing 15
GO:0019755 0.0095 Inf 0.2 2 one-carbon compound transport all_reinforcing 15
GO:0042883 0.0095 Inf 0.2 2 cysteine transport all_reinforcing 15
GO:0090502 0.0096 2.2 7.12 14 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 15
GO:0015074 1.8E-06 6.8 2.87 13 DNA integration all_reinforcing 20
GO:0032197 2.8E-06 5.4 3.88 15 transposition; RNA-mediated all_reinforcing 20
GO:0006278 4.4E-05 4.6 3.73 13 RNA-dependent DNA biosynthetic process all_reinforcing 20
GO:0070525 0.00047 24.2 0.47 4 tRNA threonylcarbamoyladenosine metabolic process all_reinforcing 20
GO:0090502 0.0006 3.3 4.73 13 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 20
GO:0034654 0.0017 1.6 41.44 59 nucleobase-containing compound biosynthetic process all_reinforcing 20
GO:0006551 0.0017 36.1 0.31 3 leucine metabolic process all_reinforcing 20
GO:0070900 0.0017 36.1 0.31 3 mitochondrial tRNA modification all_reinforcing 20
GO:1900864 0.0017 36.1 0.31 3 mitochondrial RNA modification all_reinforcing 20
GO:0015804 0.0021 7.6 1.01 5 neutral amino acid transport all_reinforcing 20
GO:0016226 0.0021 7.6 1.01 5 iron-sulfur cluster assembly all_reinforcing 20
GO:0001302 0.0029 3.9 2.56 8 replicative cell aging all_reinforcing 20
GO:0044249 0.0039 1.5 91.27 111 cellular biosynthetic process all_reinforcing 20
GO:1901576 0.004 1.5 92.28 112 organic substance biosynthetic process all_reinforcing 20
GO:0090646 0.0041 18 0.39 3 mitochondrial tRNA processing all_reinforcing 20
GO:0009083 0.0041 18 0.39 3 branched-chain amino acid catabolic process all_reinforcing 20
GO:0006310 0.005 2.2 8.69 17 DNA recombination all_reinforcing 20
GO:0015840 0.006 Inf 0.16 2 urea transport all_reinforcing 20
GO:0019755 0.006 Inf 0.16 2 one-carbon compound transport all_reinforcing 20
GO:0042883 0.006 Inf 0.16 2 cysteine transport all_reinforcing 20
GO:0043605 0.0077 12 0.47 3 cellular amide catabolic process all_reinforcing 20
GO:0019509 0.0077 12 0.47 3 L-methionine salvage from methylthioadenosine all_reinforcing 20
GO:0070880 0.0077 12 0.47 3 fungal-type cell wall beta-glucan biosynthetic process all_reinforcing 20
GO:0070879 0.0077 12 0.47 3 fungal-type cell wall beta-glucan metabolic process all_reinforcing 20
GO:0044283 0.0098 1.7 21.19 32 small molecule biosynthetic process all_reinforcing 20
GO:0015074 1.7E-06 7.4 2.4 12 DNA integration all_reinforcing 25
GO:0006278 1.7E-05 5.6 2.92 12 RNA-dependent DNA biosynthetic process all_reinforcing 25
GO:0032197 2.7E-05 5.2 3.05 12 transposition; RNA-mediated all_reinforcing 25
GO:0090502 9.7E-05 4.5 3.44 12 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 25
GO:0034654 0.0014 1.8 29.54 45 nucleobase-containing compound biosynthetic process all_reinforcing 25
GO:0001302 0.002 4.7 1.88 7 replicative cell aging all_reinforcing 25
GO:0044249 0.003 1.6 64.61 82 cellular biosynthetic process all_reinforcing 25
GO:0015804 0.0039 8.4 0.71 4 neutral amino acid transport all_reinforcing 25
GO:0006310 0.004 2.5 6.43 14 DNA recombination all_reinforcing 25
GO:1901576 0.004 1.6 65.19 82 organic substance biosynthetic process all_reinforcing 25
GO:0015840 0.0042 Inf 0.13 2 urea transport all_reinforcing 25
GO:0019755 0.0042 Inf 0.13 2 one-carbon compound transport all_reinforcing 25
GO:0070525 0.0046 14.6 0.39 3 tRNA threonylcarbamoyladenosine metabolic process all_reinforcing 25
GO:0007568 0.0063 3.3 2.86 8 aging all_reinforcing 25
GO:0090305 0.0087 2.3 6.3 13 nucleic acid phosphodiester bond hydrolysis all_reinforcing 25
GO:0006696 0.0093 3.9 1.88 6 ergosterol biosynthetic process all_reinforcing 25
GO:0015074 2E-07 9.3 1.99 12 DNA integration all_reinforcing 30
GO:0006278 1.6E-06 7.2 2.37 12 RNA-dependent DNA biosynthetic process all_reinforcing 30
GO:0032197 3.5E-06 6.6 2.53 12 transposition; RNA-mediated all_reinforcing 30
GO:0090502 4.5E-06 6.4 2.58 12 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 30
GO:0006310 0.00028 3.4 4.95 14 DNA recombination all_reinforcing 30
GO:0090305 0.0019 3 4.68 12 nucleic acid phosphodiester bond hydrolysis all_reinforcing 30
GO:1902047 0.0045 13.5 0.38 3 polyamine transmembrane transport all_reinforcing 30
GO:0044249 0.0046 1.7 45 59 cellular biosynthetic process all_reinforcing 30
GO:0034654 0.0047 1.8 20.59 32 nucleobase-containing compound biosynthetic process all_reinforcing 30
GO:1901576 0.0059 1.6 45.43 59 organic substance biosynthetic process all_reinforcing 30
GO:0015846 0.007 10.8 0.43 3 polyamine transport all_reinforcing 30
GO:1903008 0.0092 4.6 1.34 5 organelle disassembly all_reinforcing 30
GO:0015074 3.7E-08 11.2 1.74 12 DNA integration all_reinforcing 35
GO:0032197 4.5E-07 8.4 2.12 12 transposition; RNA-mediated all_reinforcing 35
GO:0006278 4.5E-07 8.4 2.12 12 RNA-dependent DNA biosynthetic process all_reinforcing 35
GO:0090502 5.9E-07 8.1 2.17 12 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 35
GO:0006310 0.00012 4 4.05 13 DNA recombination all_reinforcing 35
GO:0090305 0.00021 4 3.71 12 nucleic acid phosphodiester bond hydrolysis all_reinforcing 35
GO:0044249 0.00062 2.1 34.67 50 cellular biosynthetic process all_reinforcing 35
GO:0034654 0.00077 2.3 15.62 28 nucleobase-containing compound biosynthetic process all_reinforcing 35
GO:1901576 0.0008 2 35.01 50 organic substance biosynthetic process all_reinforcing 35
GO:1902047 0.0033 15.3 0.34 3 polyamine transmembrane transport all_reinforcing 35
GO:0015846 0.0051 12.2 0.39 3 polyamine transport all_reinforcing 35
GO:0043457 0.0067 40.3 0.14 2 regulation of cellular respiration all_reinforcing 35
GO:0006696 0.0095 4.5 1.35 5 ergosterol biosynthetic process all_reinforcing 35
GO:0015074 1.8E-08 12.2 1.64 12 DNA integration all_reinforcing 40
GO:0090502 1.7E-07 9.4 1.96 12 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 40
GO:0032197 2.2E-07 9.1 2.01 12 transposition; RNA-mediated all_reinforcing 40
GO:0006278 2.2E-07 9.1 2.01 12 RNA-dependent DNA biosynthetic process all_reinforcing 40
GO:0006310 3.1E-05 4.8 3.61 13 DNA recombination all_reinforcing 40
GO:0090305 6.7E-05 4.7 3.33 12 nucleic acid phosphodiester bond hydrolysis all_reinforcing 40
GO:0015846 0.0017 21.7 0.27 3 polyamine transport all_reinforcing 40
GO:1902047 0.0017 21.7 0.27 3 polyamine transmembrane transport all_reinforcing 40
GO:0044249 0.0027 2 28.87 41 cellular biosynthetic process all_reinforcing 40
GO:0034654 0.0053 2.1 12.88 22 nucleobase-containing compound biosynthetic process all_reinforcing 40
GO:0008610 0.0063 3.1 3.47 9 lipid biosynthetic process all_reinforcing 40
GO:0006696 0.0064 5 1.23 5 ergosterol biosynthetic process all_reinforcing 40
GO:0015804 0.0087 9.3 0.46 3 neutral amino acid transport all_reinforcing 40
GO:0015074 6.3E-08 12.3 1.5 11 DNA integration all_reinforcing 45
GO:0090502 3.7E-07 9.9 1.7 11 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 45
GO:0006278 3.7E-07 9.9 1.7 11 RNA-dependent DNA biosynthetic process all_reinforcing 45
GO:0032197 6.3E-07 9.3 1.8 11 transposition; RNA-mediated all_reinforcing 45
GO:0006310 2.8E-05 5.3 3.1 12 DNA recombination all_reinforcing 45
GO:0090305 7.6E-05 5.1 2.9 11 nucleic acid phosphodiester bond hydrolysis all_reinforcing 45
GO:0044249 0.00039 2.5 22.9 36 cellular biosynthetic process all_reinforcing 45
GO:1901576 0.0005 2.5 23.2 36 organic substance biosynthetic process all_reinforcing 45
GO:1901360 0.0018 2.2 20.7 32 organic cyclic compound metabolic process all_reinforcing 45
GO:0034654 0.0024 2.4 10 19 nucleobase-containing compound biosynthetic process all_reinforcing 45
GO:0006696 0.0029 6.2 1 5 ergosterol biosynthetic process all_reinforcing 45
GO:0016128 0.0048 5.4 1.2 5 phytosteroid metabolic process all_reinforcing 45
GO:0044107 0.0056 5.2 1.2 5 cellular alcohol metabolic process all_reinforcing 45
GO:1902652 0.0056 5.2 1.2 5 secondary alcohol metabolic process all_reinforcing 45
GO:0055114 0.0075 2.4 7.1 14 oxidation-reduction process all_reinforcing 45
GO:0006694 0.0075 4.8 1.3 5 steroid biosynthetic process all_reinforcing 45
GO:0046165 0.0087 4.6 1.3 5 alcohol biosynthetic process all_reinforcing 45
GO:0006644 0.0099 4.4 1.4 5 phospholipid metabolic process all_reinforcing 45
GO:0015074 2.5E-05 8.8 1.32 8 DNA integration all_reinforcing 50
GO:0006278 5.6E-05 7.7 1.47 8 RNA-dependent DNA biosynthetic process all_reinforcing 50
GO:0090502 6.8E-05 7.5 1.5 8 RNA phosphodiester bond hydrolysis; endonucleolytic all_reinforcing 50
GO:0032197 0.00012 6.8 1.61 8 transposition; RNA-mediated all_reinforcing 50
GO:0006310 0.00052 4.7 2.49 9 DNA recombination all_reinforcing 50
GO:0006696 0.0011 8.1 0.84 5 ergosterol biosynthetic process all_reinforcing 50
GO:0016128 0.002 6.9 0.95 5 phytosteroid metabolic process all_reinforcing 50
GO:0090305 0.002 4.2 2.42 8 nucleic acid phosphodiester bond hydrolysis all_reinforcing 50
GO:0044107 0.0023 6.6 0.99 5 cellular alcohol metabolic process all_reinforcing 50
GO:1902652 0.0023 6.6 0.99 5 secondary alcohol metabolic process all_reinforcing 50
GO:0006694 0.0033 6 1.06 5 steroid biosynthetic process all_reinforcing 50
GO:0046165 0.0038 5.8 1.1 5 alcohol biosynthetic process all_reinforcing 50
GO:0016125 0.0075 4.8 1.28 5 sterol metabolic process all_reinforcing 50
GO:1901362 0.0094 2.4 8.21 15 organic cyclic compound biosynthetic process all_reinforcing 50

Abbreviations: "GOBPID": gene ontology biological process ID, "ExpCount": expected number of genes for a given GOBPID, "LOD": LOD score threshold for including a gene with a cis / trans eQTL pair.

Because the results of our GO enrichment could be affected by the reference gene set (the list of all genes with at least one cis and one trans eQTL), we devised a complementary strategy to test for lineage-specific selection on UPS gene expression. Using the same set of BY / RM eQTLs, we applied the sign test to the sets of UPS genes, proteasome genes, ubiquitin system genes, E3 ligases, and proteasome chaperone genes. We did not detect lineage-specific selection in any of these gene sets (Author response table 6).

Author response table 6. Results of the Sign Test for Lineage-Specific Selection Applied to UPS Gene BY / RM eQTLs.

LOD n_eQTLs n_genes n_pairs reinf_by_up reinf_rm_up oppos_by_up oppos_rm_up excess_reinforcing_pairs chi_sq_p gene_set
2.4 1128 186 95 16 30 29 20 –4.21 0.815 all_UPS_genes
5 587 176 72 13 23 23 13 0 1 all_UPS_genes
10 250 142 29 2 12 8 7 –4.41 0.428 all_UPS_genes
15 162 105 19 1 8 7 3 –2.74 0.619 all_UPS_genes
20 115 85 10 0 4 4 2 –3.2 0.472 all_UPS_genes
25 90 72 8 0 3 3 2 -3 0.475 all_UPS_genes
30 63 56 5 0 2 2 1 –1.6 1 all_UPS_genes
35 52 47 4 0 1 2 1 -2 1 all_UPS_genes
40 37 34 3 0 1 2 0 0 NaN all_UPS_genes
45 32 30 2 0 1 1 0 0 NaN all_UPS_genes
2.4 242 33 14 3 3 2 6 –0.857 1 proteasome_genes
5 146 32 12 3 2 2 5 –1.33 1 proteasome_genes
10 62 30 4 0 0 0 4 0 NaN proteasome_genes
15 32 19 2 0 0 0 2 0 NaN proteasome_genes
20 18 14 1 0 0 0 1 0 NaN proteasome_genes
25 11 9 1 0 0 0 1 0 NaN proteasome_genes
2.4 829 145 76 13 23 25 15 -4 0.812 ubiquitin_system_genes
5 410 137 57 10 18 20 9 0 1 ubiquitin_system_genes
10 173 105 24 2 11 7 4 -1 1 ubiquitin_system_genes
15 118 80 17 1 7 7 2 –1.65 1 ubiquitin_system_genes
20 88 65 8 0 3 4 1 -2 1 ubiquitin_system_genes
25 73 58 6 0 2 3 1 -2 1 ubiquitin_system_genes
30 54 48 4 0 1 2 1 -2 1 ubiquitin_system_genes
35 44 40 3 0 0 2 1 –2.67 0.324 ubiquitin_system_genes
40 31 29 2 0 0 2 0 0 NaN ubiquitin_system_genes
45 26 25 1 0 0 1 0 0 NaN ubiquitin_system_genes
2.4 618 111 56 11 16 19 10 -1 1 E3_ligase_genes
5 300 103 39 8 12 14 5 2.67 0.909 E3_ligase_genes
10 129 79 15 2 7 4 2 1.6 1 E3_ligase_genes
15 83 56 9 1 4 4 0 1.78 1 E3_ligase_genes
20 64 46 5 0 2 3 0 0 NaN E3_ligase_genes
25 49 39 3 0 1 2 0 0 NaN E3_ligase_genes
30 37 33 3 0 1 2 0 0 NaN E3_ligase_genes
35 30 28 2 0 0 2 0 0 NaN E3_ligase_genes
40 22 20 2 0 0 2 0 0 NaN E3_ligase_genes
45 18 17 1 0 0 1 0 0 NaN E3_ligase_genes
2.4 57 9 6 0 4 2 0 0 NaN proteasome_chaperone_genes
5 31 8 4 0 3 1 0 0 NaN proteasome_chaperone_genes
10 15 8 2 0 1 1 0 0 NaN proteasome_chaperone_genes
15 12 7 1 0 1 0 0 0 NaN proteasome_chaperone_genes
20 9 6 1 0 1 0 0 0 NaN proteasome_chaperone_genes
25 6 5 1 0 1 0 0 0 NaN proteasome_chaperone_genes
30 5 4 1 0 1 0 0 0 NaN proteasome_chaperone_genes
35 4 3 1 0 1 0 0 0 NaN proteasome_chaperone_genes
40 4 3 1 0 1 0 0 0 NaN proteasome_chaperone_genes
45 4 3 1 0 1 0 0 0 NaN proteasome_chaperone_genes

Abbreviations: “LOD”: LOD score threshold for calling cis / trans eQTL pairs, "n_eQTLs": number of eQTLs, "n_genes": number of genes with an eQTL, "n_pairs": number of cis / trans eQTL pairs, "reinf_BY_up": number of cis / trans eQTLs where the BY allele of the cis and trans eQTLs increases expression (reinforcing pairs), "reinf_RM_up": number of cis / trans eQTLs where the RM allele of the cis and trans eQTLs increases expression (reinforcing pairs), "oppos_BY_up": number of cis / trans eQTLs where the BY allele of the cis eQTL increases expression and the RM allele of the trans eQTL increases expression (opposing pairs), "oppos_RM_up": number of cis / trans eQTLs where the RM allele of the cis eQTL increases expression and the BY allele of the trans eQTL increases expression (opposing pairs), "excess reinforcing pairs" the number of excess reinforcing pairs, calculated as in Fraser et al., 2010, "chi_sq_p": p-value of the chi-square test for enrichment of reinforcing pairs.

Although we did not detect lineage-specific selection on global or UPS gene mRNA abundance, we note that these analyses do not detect lineage-specific selection on factors other than transcript abundance. For example, lineage-specific selection may occur through causal missense variants, such as the ones we identified here. For example, the DOA10 and NTA1 genes each contain multiple causal missense variants that alter UPS activity, but neither gene's transcript abundance is affected by a local eQTL. Integration of the effects of missense variants with those that alter gene expression in selection tests remains an interesting avenue for future work.

We note that the significant excess of UPS QTLs at which the RM allele increases UPS activity (pg. 6, para. 2, line 170) provides evidence that N-end Rule activity could have been subject to lineage-specific selection. To test whether this result was due to a strong enrichment of RM alleles at certain individual reporters, we applied this analysis to the sets of QTLs obtained for the 20 individual N-degrons. We did not detect an enrichment of any of the individual N-degrons, suggesting that the enrichment of RM QTLs that increase UPS activity is a general effect (Author response table 7).

Author response table 7. Results of Binomial Enrichment Test Applied to QTLs for Individual N-degrons.

N-degron N_QTLs N_RM_up N_BY_up p_value
Ala 15 10 5 0.30
Arg 7 4 3 1.00
Asn 7 3 4 1.00
Asp 8 5 3 0.73
Cys 9 4 5 1.00
Gln 3 2 1 1.00
Glu 4 2 2 1.00
Gly 9 5 4 1.00
His 9 6 3 0.51
Ile 1 1 0 1.00
Leu 2 1 1 1.00
Lys 7 5 2 0.45
Met 4 1 3 0.63
Phe 10 6 4 0.75
Pro 13 8 5 0.58
Ser 9 6 3 0.51
Thr 12 8 4 0.39
Trp 6 4 2 0.69
Tyr 7 4 3 1.00
Val 7 4 3 1.00

Abbreviations: "N_QTLs": Number of QTLs for the indicated N-degron, "N_RM_up": Number of QTLs where the RM allele increases UPS activity for the indicated N-degron, "N_BY_up": Number of QTLs where the BY allele increases UPS activity for the indicated N-degron. "p_value": p-value of the binomial test for the set of QTLs for the indicated N-degron.

The paper reports 149 loci, but if I am reading this correctly it looks like this may double-count shared loci. It would be nice to discuss a bit more about likely sharing (ie pleiotropy) of the hits. (It's also a bit hard to see from 2B which of these are likely shared.)

The reviewer is correct. As described in our response to Reviewer 1, the revised manuscript now indicates that the 149 instances of QTL detection correspond to 35 distinct QTL regions (pg. 6, para. 3, line 185). Figure 2—figure supplement 2 provides detailed information on these regions, including which reporters each region affects. The revised manuscript includes an extended discussion of pleiotropy among the set of N-end Rule QTLs (pg. 6, para. 3).

L143: it should be possible to make this statement quantitative.

As noted in our response to Reviewer 1, it is not readily possible to calculate the amount of variance explained by each QTL due to the pooled nature of our bulk segregant analysis QTL mapping method. Accordingly, we have removed the statement mentioned by the reviewer from the revised Results section.

Para 303, 312: It seems likely that you may be looking for an effect that is small at individual genes but widespread. I would suggest looking more carefully at the overall distribution: eg in the protein data is there an upward shift in the mean? You could also use a more sophisticated method like ashR to study the distribution of changes.

We appreciate the reviewer’s suggestion, which we suspect may have been prompted by our wording of the following sentence from the Results section (page 23, paragraph 2, lines 307-310 of the original manuscript):

“This result is consistent with recent observations that suggest that altering UBR1 expression exerts broad effects on protein degradation or related processes controlling protein abundance and that protein sequences, rather than function or subcellular localization, are the primary determinants of degradation rates.”

Our intent was to convey that substrates of E3 ligases such as Ubr1 are more likely to share sequence features than they are to share a function or subcellular localization. In other words, E3 ligases typically target sets of proteins that are functionally diverse but that share common sequence features, e.g., an N-degron. The wording of this sentence may have unintentionally suggested that the causal UBR1 promoter variant should affect the abundance of many proteins. However, E3 ligases influence the abundance of distinct sets of dozens to several hundred proteins (Kong et al., 2021, Christiano et al., 2020). The moderate effect of the causal UBR1 promoter variant on UBR1 expression is, therefore, not expected to create widespread effects on the proteome but instead to affect a small subset of Ubr1 substrates and related proteins. We have revised this section to more clearly articulate these ideas. We have also re-analysed our data as described below.

The causal UBR1 variant significantly altered the abundance of 39 of 3,047 detected proteins at a 0.1 false discovery rate (FDR) threshold. Following the reviewer’s suggestion, we computed the overall median log2 fold change and found that it was -0.012 for all 3,047 detected proteins (that is, very close to zero) and 0.37 for the set of 39 differentially abundant proteins (that is, an average increase for proteins with significantly different abundance). The number of differentially abundant proteins, the average upward shift in log2 fold change for differentially abundant proteins, and the significant fraction of differentially abundant proteins exhibiting increased abundance (28 / 39, 72%, binomial p = 9.5e-3) are all consistent with the causal variant’s moderate effect on UBR1 expression.

The causal variant also did not have widespread effects on the transcriptome (median log2 fold change = -0.0024 [again, very close to zero]). As reported in our initial submission, 78 genes were differentially expressed at an FDR of 0.1. For these genes, the overall effect of the causal BY allele, which decreases UBR1 expression, was to decrease transcript abundance (median log2 fold change = -0.18, 60 / 78 decreased expression [77%, binomial p = 2e-6]).

Following the reviewer’s suggestion, we used ashr to explore how using the false sign rate (FSR, Stephens, 2017; https://doi.org/10.1093/biostatistics/kxw041) to call differentially expressed genes influenced our analysis of the causal UBR1 variant’s effects. Using an FSR threshold of 0.1, we detected 86 genes with altered mRNA transcript abundance. For these genes, the BY allele tended to decrease transcript abundance (median log2 fold change = -0.2, 65 / 86 [76%, binomial p = 2e-6]). All differentially expressed genes detected using the FDR were also detected using the FSR. Eight additional differentially expressed genes were detected using the FSR, but not the FDR. Each of these eight genes had lower absolute log2 fold changes than the 78 differentially expressed genes detected using the FDR. Therefore, we conclude that while the FSR is slightly more permissive, the FDR and FSR both capture the same global patterns of effects in our RNA-seq data.

In contrast to these consistent results in RNA-seq, ashr produced different results than the FDR when applied to our proteomics data. As reported in our initial submission, a 0.1 FDR threshold applied to abundance differences reported by the Proteome Discoverer software results in 39 differentially abundant proteins. The derived BY UBR1 allele increased the abundance of 28 of these 39 (72%, binomial p = 9.5e-3) differentially abundant proteins. At a 0.1 FSR threshold, we identified 56 differentially abundant proteins, with the BY allele increasing the abundance of 13 / 56 (23%, binomial p = 7.3e-5).This difference arises from the fact that there are considerable discrepancies among the sets of differentially abundant proteins called by FDR and FSR. We note that of the 39 differentially abundant proteins reported by Proteome Discoverer, 12 were estimated to have large absolute fold changes (>= 0.5, the 99th percentile observed in our data) but were not reported as significant by FSR (Author response image 1). Notably, these 12 genes included the known Ubr1-regulated proteins Tma10 and Adh2 (Kong et al., 2021, Christiano et al., 2020). For 8 of these 12 proteins, the BY allele increased protein abundance. These 12 proteins had relatively high standard errors and, as expected, their fold changes were considerably lower following the adaptive shrinkage that is part of ashr (Author response image 1). Second, ashr identified an additional 48 differentially abundant proteins. The BY allele decreases the abundance of 41 of these 48 proteins (Author response image 1).

Author response image 1.

Author response image 1.

Given that ashr strongly depends on correctly estimated standard errors, its application to mass-spectrometry data with a small sample size could be potentially problematic in a manner that does not seem to apply to RNA-seq, which, particularly at the very high sequencing depth used here, produces more accurate estimates of the standard error. We have therefore elected to continue to use the FDR applied to p-values reported by Proteome Discoverer to call differentially expressed genes at the protein and RNA levels. We note that our general conclusion that the causal UBR1 promoter variant affects the expression of dozens of genes at the protein and RNA levels remains unchanged irrespective of the method used to call differentially expressed genes.

For Figure 5 you could also consider adding information about global patterns in means and correlations, which are difficult to read from these plots.

The revised Figure 5 contains the median log2 fold changes for our proteomics and RNA-seq data (pg. 14). The revised Results section reports the correlation in log2 fold change for genes detected in both our proteomics (pg. 13, para. 4, line 362) and RNA-seq data (pg. 13, para. 5, line 377).

Associated Data

    This section collects any data citations, data availability statements, or supplementary materials included in this article.

    Data Citations

    1. Collins MA. 2022. Variation in ubiquitin system genes creates substrate-specific effects on proteasomal protein degradation. NCBI Gene Expression Omnibus. GSE213689 [DOI] [PMC free article] [PubMed]
    2. Collins MA. 2022. Variation in ubiquitin system genes creates substrate-specific effects on proteasomal protein degradation. NCBI BioProject. PRJNA881749 [DOI] [PMC free article] [PubMed]

    Supplementary Materials

    Figure 1—source data 1. Results of all between-strain comparisons for all N-degron TFTs.
    Figure 2—source data 1. All N-end rule QTLs.

    "chr" = chromosome, "LOD" = logarithm of the odds, "QTL_CI_left" = left index of QTL confidence interval, "QTL_peak" = peak position of QTL, "QTL_CI_right" = right index of QTL confidence interval.

    Figure 2—source data 2. All distinct N-end rule QTL regions.

    "chr" = chromosome, "QTL_CI_left" = left index of QTL confidence interval, "QTL_peak" = peak position of QTL, "QTL_right_CI" = right index of QTL confidence interval, "LOD" = logarithm of the odds, "RM_AFD" = RM allele frequency difference between high and low UPS activity pools.

    Figure 5—source data 1. Full proteomics results.
    elife-79570-fig5-data1.xlsx (153.2KB, xlsx)
    Figure 5—source data 2. Full RNA-seq results.
    elife-79570-fig5-data2.xlsx (317.8KB, xlsx)
    Supplementary file 1. Allele frequency difference and LOD score traces from QTL mapping experiments.

    The plots show the loess-smoothed RM allele frequency difference (high UPS activity pool minus low UPS activity pool) and LOD score traces for the 20 N-degrons. QTLs are marked with asterisks, which are colored by biological replicate.

    elife-79570-supp1.pdf (10MB, pdf)
    Supplementary file 2. Influence of LOD score significance threshold on QTL pathway specificity.

    The LOD score and RM allele frequency difference (QTL effect direction) traces for two independent biological replicates of each N-degron are shown for each of 23 pathway-specific QTL regions. Dashed lines at distinct LOD scores illustrate how changing the significance threshold changes the pathway-specificity of a given QTL region.

    elife-79570-supp2.pdf (31.7MB, pdf)
    Supplementary file 3. Oligonucleotides.

    Table listing oligonucleotides used in this study.

    elife-79570-supp3.xlsx (8.1KB, xlsx)
    Supplementary file 4. Plasmids.

    Table of plasmids used in this study.

    elife-79570-supp4.xlsx (6.2KB, xlsx)
    Supplementary file 5. Yeast strains.

    Table listing all yeast strains used in the study.

    elife-79570-supp5.xlsx (12.2KB, xlsx)
    MDAR checklist

    Data Availability Statement

    Raw sequencing reads from QTL mapping experiments are available from the NIH Sequence Read Archive under the Bioproject Accession PRJNA881749. Raw and processed RNA-seq data is available from the NIH Gene Expression Omnibus under the accession GSE213689. These datasets are fully available without restriction. Computational scripts used to process data, for statistical analysis, and to generate figures are available at: https://www.github.com/mac230/N-end_Rule_QTL_paper, (copy archived at swh:1:rev:24baa12af4e9c45691be2590ab30b2c1faf0c497).

    The following datasets were generated:

    Collins MA. 2022. Variation in ubiquitin system genes creates substrate-specific effects on proteasomal protein degradation. NCBI Gene Expression Omnibus. GSE213689

    Collins MA. 2022. Variation in ubiquitin system genes creates substrate-specific effects on proteasomal protein degradation. NCBI BioProject. PRJNA881749


    Articles from eLife are provided here courtesy of eLife Sciences Publications, Ltd

    RESOURCES