Compositions and methods for low glucosinolate brassica
By altering the MAM1 gene in Brassica napus using CRISPR-Cas9, glucosinolate levels are reduced, addressing the anti-nutritional issues in canola meal and enhancing its nutritional value for animal feed.
Patent Information
- Application Number
- PCT/US2025/010724
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2024-01-10
- Filing Date
- 2025-01-08
- Publication Date
- 2025-07-17
AI Technical Summary
The high levels of glucosinolates in Brassica napus meal, which are anti-nutritional factors affecting the palatability and animal health, have not been adequately reduced through conventional breeding, necessitating a need to modify the compositional properties of canola meal to enhance its nutritional value.
Targeted alterations of the MAM1 gene in Brassica napus using CRISPR-Cas9 to introduce mutations such as nonsense mutations, missense mutations, and deletions that reduce or eliminate MAM1 gene function, leading to decreased glucosinolate levels in the plant.
The method significantly decreases glucosinolate levels in Brassica napus plants, improving the nutritional value of the meal by enhancing palatability and reducing potential health risks to animals, thereby increasing its value as animal feed.
Smart Images

Figure US2025010724_17072025_PF_FP_ABST
Abstract
Description
COMPOSITIONS AND METHODS FOR LOW GLUCOSINOLATE BRASSICAREFERENCE TO A SEQUENCE LISTING SUBMITTTED ELECTRONICALLY
[0001] The official copy of the sequence listing is submitted electronically via Patent Center as an XML formatted sequence listing with a file named 211541-US-PRI-l.xml created on January 4, 2024, and having a size of 261,869 bytes and is filed concurrently with the specification. The sequence listing comprised in this XML formatted document is part of the specification and is herein incorporated by reference in its entirety.
[0002] The sequence descriptions (Table 1) and sequence listing attached hereto comply with the rules governing nucleotide and amino acid sequence disclosures in patent applications as set forth in 37 C.F.R. §§1.831-1.835.FIELD OF THE DISCLOSURE
[0003] The disclosure relates to glucosinolate metabolism in Brassica, variants of MAM1 gene that reduce glucosinolate in Brassica grain and meal made thereof.BACKGROUND OF THE INVENTION
[0004] Brassica napus (also referred to herein as canola or oilseed rape) is an allotetraploid (2n= 4x = 38, AACC) comprising two full genome sets, four component genomes, and total of 38 chromosomes. The A genome includes 10 chromosomes and is derived from B. rapa (2n = 2x = 20, AA). The C genome includes 9 chromosomes and is derived from B. oleracea (2n = 2x = 18, CC). B. napus is one of the most important vegetable oilseed crops in the world, especially in China, Canada, the European Union and Australia. Canola meal, the fraction of the seed remaining after crushing and oil extraction, is approximately 55% of the volume of canola seed.
[0005] While canola meal is rich in protein and capable of providing a substantial amount of energy when used in animal feed; it also includes levels of anti -nutritional factors such as glucosinolate, tannin, phytate (or phytic acid), and sinapine. Since meal comprises about half of the seed volume of canola, and the demand for canola / oilseed rape has risen and is expected to continue rising to meet demands for healthy cooking oils, biodiesel, and personal care products, there is a long-felt need to modify the compositional properties of canola meal and thereby increasing its nutritional value.
[0006] Glucosinolate constitute a large family of over 100 related molecules with a common sulfur containing core structure and with side chains of varying size and chemistry (Fahey et al.,2001; Halkier and Gershenzon, 2006). While glucosinolates are found in many plant structures (leaf, vascular tissue, stem, root, and flowers, to cite some examples), they are accumulated in high concentrations in the seed (Bellostas et al., 2004). This is particularly true for the oil seed brassicas. These compounds and their metabolites can impact the taste of the meal, reducing its palatability and in some cases (dependent on the type of glucosinolate and glucosinolate metabolites present) can also adversely impact the animal's health directly. For example, hydrolysis products of beta hydroxyalkenyl glucosinolates have been shown to possess goitrogenic activity in animal models (reviewed in Fahey et al., 2001). This is particularly an issue in monogastic animals such as swine, but poultry and cattle can be susceptible to varying degrees. Thus glucosinolate reduction in oil seed meal is an important and desirable objective and can have significant benefits in terms of meal value.
[0007] Arabidopsis MAM1 gene encodes a methylthioalkylmalate synthase which catalyzes the condensation reactions of the first two rounds of methionine chain elongation in the biosynthesis of methionine-derived aliphatic glucosinolates. Aliphatic glucosinolates are the major class of glucosinolates found in oilseed rape. Although the glucosinolate content in canola and winter oilseed rape hybrid has been significantly reduced through breeding since 1980s, efforts to further reduce glucosinolate continue. Brassica has several homologous genes of AtMAMl . Here, the CRISPR-Cas9 system was used to edit the BnMAMl genes for generating loss-of- function variants to reduce glucosinolate in Brassica seed.SUMMARY OF THE INVENTION
[0008] The disclosed compositions and methods are based, at least in part, on the surprising discovery that levels of glucosinolate in Brassica napus can be decreased by a targeted alteration of the genomic sequence of MAM 1 gene that encode for methylthioalkylmalate synthase. The evidence herein demonstrates that targeted alterations of MAM1 lead to significantly decreased levels of glucosinolate in Brassica napus plants. The term Brassica napus as used herein includes crop varieties of the species such as spring oilseed rape, winter oilseed rape, and low erucic cultivars of the foregoing which are called canola.
[0009] Provided herein is a method of decreasing glucosinolate in Brassica napus plant, cell, seed, tissue or germplasm thereof, that comprises introducing a targeted alteration to the sequence of methylthioalkylmalate synthase (MAM1) gene. Targeted alterations can be made to a MAM1 gene to thereby generate a Brassica napus plant, cell, seed, tissue or germplasm thereofthat comprises one or more MAM1 variants which can provide a decreased level of glucosinolate relative to the plant, seed, tissue or germplasm thereof prior to introducing the one or more MAM1 variants. In one aspect, the method comprises introducing a targeted alteration of one or more alleles of MAM 1 on chromosome A2, C2, A3, C7, or A4. The types of targeted alterations that can be used to create MAM1 variants disclosed herein include nonsense mutations, missense mutations, and deletions that eliminate or reduce MAM1 gene function. For example, the targeted alteration can be a premature termination codon that reduces or eliminates expression of a full-length protein encoded by the altered MAM1 variant. Particular examples of a targeted alteration that introduces a premature termination codon are shown in Table 3 for MAM1 on chromosome A2 (MAM1.A2), chromosome C2 (MAM1.C2), chromosome A3 (MAM1.A3), chromosome C7 (MAM1 ,C7a or MAM1 ,C7b), or chromosome A4 (MAM1. A4). An examples of a targeted alteration comprising a premature termination codon is shown in each of SEQ ID NO:13, SEQIDNO:14, SEQIDNO:15, SEQIDNO:16, SEQIDNO 17, SEQIDNO:18, SEQ ID NO: 19, SEQIDNO:20, SEQIDNO:21, SEQIDNO:22, SEQIDNO:23, SEQIDNO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQIDNO:31, SEQIDNO:32, SEQIDNO:33, SEQIDNO 34, SEQIDNO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ ID NO:38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQIDNO:41, SEQIDNO:41, SEQIDNO:41, SEQIDNO:41, SEQIDNO:41, SEQ ID NO:50, SEQIDNO:51, SEQIDNO:52, SEQIDNO:53, SEQIDNO 54, SEQIDNO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO: 67, SEQ ID NO: 68, SEQ ID NO: 69, or SEQ ID NO: 70.
[0010] Disclosed herein is a method of decreasing glucosinolate in Brassica napus plant, cell, seed, tissue or germplasm thereof, that comprises introducing a targeted alteration to the sequence of the methylthioalkylmalate synthase (MAM1) gene to generate & Brassica napus plant, cell, seed, tissue or germplasm thereof comprising a homozygous variant of MAM1. For example, targeted alterations can be introduced to make Brassica napus plant, cell, seed, tissue or germplasm thereof comprising homozygous knockouts of MAM1 (null for MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, MAM1.A4, or combinations thereof). Particular examples of such targeted alterations that can be used to generate MAM1 alleles are shown in Tables 4 and 5. Additionally examples of targeted alterations that can be used areshown in each of SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQ ID NO: 16, SEQ ID NO: 17, SEQ ID NO: 18, SEQ ID NO: 19, SEQ ID NO:20, SEQ ID NO 21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ ID NO 38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO 58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NO:70
[0011] In particular examples, the disclosed method comprises introducing a targeted alteration into MAM1 genes of Brassica napus plant, cell, seed, tissue or germplasm thereof, thereby introducing MAM1 variant alleles, which are disclosed by reference to the variant alleles disclosed in Table 3, Table 4, or Table 5 herein, such that the term “MAM1 ,A2 Variant” refers to MAM1 ,A2_vl or MAM1 ,A3_v2; the term “MAM1 ,C2 Variant” refers to MAM1. C2_vl, MAM1. C2_v2, MAM1. C2_v3, or MAMl. C2_v4; the term “MAM1.A3 Variant” refers to MAMl.A3_vl, MAMl .A3_v2, MAMl.A3_v3, MAMl.A3_v4, or MAMl.A3_v5; the term “MAMEC7a Variant” refers to MAMEC7a_vl, MAMl.C7a_v2, MAMEC7a_v3, MAMl.C7a_v4, MAMl.C7a_v5, or MAMl.C7a_v6; the term “MAMl.C7b Variant” refers to MAMl.C7b_vl, MAMl.C7b_v2, MAMl.C7b_v3, MAMl.C7b_v4, MAMl.C7b_v5, or MAMl.C7b_v6; the term “MAM1.A4 Variant” refers to MAMl.A4_vl, MAMl.A4_v2, MAM1 . A4_v3, MAM1 ,A4_v4, MAM1. A4_v5, or MAM1. A4_v6. For example, the method includes introducing (i) any homozygous MAM1 ,A2 Variant described herein combined with any heterozygous or homozygous MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, MAM1 ,C7b Variant, or MAM1 ,A4 Variant described herein; (ii) any homozygous MAM1 ,C2 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, MAMl .C7b Variant, or MAM1.A4 Variant described herein; (iii) any homozygous MAM1.A3 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAMl.C7a Variant, MAM1 ,C7b Variant, or MAM1 ,A4 Variant described herein; (iv) any homozygous MAM1 ,C7a Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant,MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (v) any homozygous MAMl.C7b Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAM1.A4 Variant described herein; or (vi) any homozygous MAM1.A4 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAMl.C7b Variant described herein.
[0012] In each of the foregoing methods of introducing targeted alterations to one or more MAM1 gene alleles, the Brassica napus plant, cell, seed, tissue or germplasm thereof that is altered can also further comprise an altered or variant gene that provides additional desirable meal quality. Thus, the method can comprise introducing targeted alterations that generate one or more MAM1 variants in Brassica napus plant, cell, seed, tissue or germplasm thereof that further includes a targeted alteration, mutation or variant of one or more of LPA1, TT2, TT8, DFR, F3H, ANR, LDOX, MYB28, MRP1, or MRP2. The desirable meal quality provided by each of LPA1, MRP1, and MRP2 is reduced phytate and / or increased inorganic phosphate content. Each of TT2, TT8, DFR, and F3H, ANR, and LDOX can reduce fiber content. MYB28 can reduce glucosinolate content. The method can also comprise introducing targeted alterations that generate one or more MAM1 variants in Brassica napus plant, cell, seed, tissue or germplasm thereof that provides increased Brassica napus meal protein and / or low fiber. Such germplasm, for example, are disclosed in US Patent Nos. 9,375,025 and 10,791,692; as well as International Patent Application Publication Nos. WO 2020 / 131600.
[0013] Also provided herein are Brassica napus plant materials produced by the methods disclosed herein. Accordingly, provided herein is a Brassica napus plant, cell, seed, tissue or germplasm thereof that comprises one or more MAM1 variants, wherein the variants comprise a targeted alteration of MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, or MAM1.A4. The MAM1 variants can be one or more alleles of MAM1 on chromosome A2, C2, A3, C7, or A4. The types of targeted alterations include nonsense mutations, missense mutations, and deletions that eliminate or reduce MAM1 gene function. For example, the targeted alteration can be a premature termination codon that reduces or eliminates expression of a full-length protein encoded by the altered MAM1 allele. Particular examples of a targeted alteration that introduces a premature termination codon are shown in Table 3 for MAM1 on chromosome A2(MAM1.A2) on chromosome C2 (MAM1.C2), on chromosome A3 (MAM1.A3), on chromosome C7 (MAM1 ,C7a or MAM1 ,C7b), or on chromosome A4 (MAM1 ,A4). An example of a targeted alteration comprising a premature termination codon is shown in each of SEQ ID NO:13, SEQIDNO:14, SEQIDNO:15, SEQIDNO:16, SEQZDNO 17, SEQIDNO:18, SEQ ID NO: 19, SEQIDNO:20, SEQIDNO:21, SEQIDNO:22, SEQIDNO:23, SEQIDNO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQIDNO:31, SEQIDNO:32, SEQIDNO:33, SEQIDNO34, SEQIDNO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ ID NO:38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQIDNO:4I, SEQIDNO:41, SEQIDNO:41, SEQIDNO:41, SEQIDNO:41, SEQ ID NO:50, SEQIDNO:51, SEQIDNO:52, SEQIDNO:53, SEQIDNO54, SEQIDNO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO: 67, SEQ ID NO: 68, SEQ ID NO: 69, or SEQ ID NO: 70.
[0014] In other examples, the disclosure provides Brassica napus plant, cell, seed, tissue or germplasm thereof comprising any combination of MAM1 variant alleles disclosed herein including (see definition above of terms “MAM1.A2 Variant”, “MAM1.C2 Variant”, “MAM1.A3 Variant”, “MAMl.C7a Variant”, “MAMl.C7b Variant”, or “MAM1.A4 Variant” and variants disclosed in Table 3, Table 4, or Table 5 herein): (i) any homozygous MAM1 ,A2 Variant described herein combined with any heterozygous or homozygous MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (ii) any homozygous MAM1 ,C2 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (iii) any homozygous MAM1.A3 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAMl.C7a Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (iv) any homozygous MAMl.C7a Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (v) any homozygous MAMl.C7b Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAM1.A4 Variant described herein; or (vi) any homozygous MAM1.A4 Variant described herein combined with anyheterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAMl.C7b Variant described herein.
[0015] In particular examples, each of the foregoing Brassica napus plant, cell, seed, tissue or germplasm thereof comprising one or more MAM1 variants can further include a targeted alteration, mutation or variant of one or more of the following genes: LPA1, TT2, TT8, DFR, F3H, ANR, LDOX, MYB28, MRP1 or MRP2, which can provide additional desirable meal quality. In additional examples, each of the foregoing Brassica napus plant, cell, seed, tissue or germplasm thereof comprising one or more MAM1 variants can further include increased Brassica napus meal protein and / or low fiber. Such germplasm, for example, are disclosed in US Patent Nos. 9,375,025 and 10,791,692; as well as International Patent Application Publication Nos. WO 2020 / 131600.
[0016] Each of the Brassica napus plant, seed, tissue or germplasm thereof comprising one or more MAM1 variants disclosed herein can be used to produce oilseed or grain which is milled to produce meal, e.g., canola meal. In preferred embodiments, the oilseed or grain comprises total glucosinolate content equal to or less than 10.0 pmole / g, 9.0 pmole / g, 8.0 pmole / g, or 7.0 pmole / g. In other preferred embodiments, the meal produced from the disclosed Brassica napus oilseed comprises decreased levels of glucosinolate relative to control seed or grain lacking the one or more MAM1 variants. As used herein control refers to seed or grain that lacks the one or more MAM1 variants but is otherwise isogenic or substantially isogenic to the Brassica napus disclosed herein.
[0017] Accordingly provided herein is a method of producing reduced glucosinolate Brassica napus meal, the method comprising providing or selecting any of the Brassica napus seed or grain disclosed herein that comprises one or more MAM1 variants, wherein the variants comprise a targeted alteration of MAM1 and the selected seed or grain comprise decreased glucosinolate relative to control seed or grain lacking the one or more MAM1 variants. The method then comprises milling the selected seed or grain to produce reduced glucosinolate Brassica napus meal. In any of the aspects, embodiments or examples of this method disclosed herein, the method can further include providing this reduced glucosinolate Brassica napus meal in feed to an animal, e.g., a monogastric animal. Preferred monogastric animals include swine and chickens (e.g., egg-laying hens, broilers, pullets, etc.) as well as other poultry such asturkeys, ducks, geese, guinea fowl, etc. Feed containing this reduced glucosinolate Brassica napus meal can also be fed to horses, rabbit, or any other commercially raised animal.
[0018] In one aspect, a method of producing reduced glucosinolate Brassica napus meal comprises milling Brassica napus seed or grain that comprises a targeted alteration to one or more alleles of MAM1 on chromosome A2, C2, A3, C7, or A4. The types of targeted alterations include nonsense mutations, missense mutations, and deletions that eliminate or reduce MAM1 gene function. For example, the targeted alteration can be a premature termination codon that reduces or eliminates expression of a full-length protein encoded by the altered MAM1 allele. Particular examples of a targeted alteration that introduces a premature termination codon are shown in Table 3 for MAM1 on chromosome A2, C2, A3, C7, or A4. An example of a targeted alteration comprising a premature termination codon is shown in each of SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQ ID NO: 16, SEQ ID NO:17, SEQ ID NO:18, SEQ ID NO: 19, SEQ ID NOTO, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NOT 1, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ ID NO:38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO 41, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO 66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NOTO. In other examples of this method of producing reduced glucosinolate meal, the method comprises milling seed or grain that comprises any combination of MAM1 variants disclosed herein (see definition above of terms “MAM1.A2 Variant”, “MAM1.C2 Variant”, “MAM1.A3 Variant”, “MAMl.C7a Variant”, “MAMl.C7b Variant”, or “MAM1.A4 Variant” and variants disclosed in Table 3, Table 4, or Table 5 herein): (i) any homozygous MAM1 ,A2 Variant described herein combined with any heterozygous or homozygous MAM1.C2 Variant, MAM1.A3 Variant, MAMl .C7a Variant, MAMl.C7b Variant, or MAM1. A4 Variant described herein; (ii) any homozygous MAM1 ,C2 Variant described herein combined with any heterozygous or homozygous MAM1 ,A2 Variant, MAM1 A3 Variant, MAM1 ,C7a Variant, MAM1 ,C7b Variant, or MAM1. A4 Variant described herein; (iii) any homozygous MAM1.A3 Variant described herein combined with any heterozygous orhomozygous MAM1.A2 Variant, MAM1.C2 Variant, MAMl.C7a Variant, MAMl.C7b Variant, or MAM1. A4 Variant described herein; (iv) any homozygous MAM1 ,C7a Variant described herein combined with any heterozygous or homozygous MAM1 ,A2 Variant, MAM1 C2 Variant, MAM1.A3 Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (v) any homozygous MAMl.C7b Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAM1.A4 Variant described herein; or (vi) any homozygous MAM1.A4 Variant described herein combined with any heterozygous or homozygous MAM1 ,A2 Variant, MAM1 C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAMl.C7b Variant described herein.
[0019] In another aspect, provided herein is a screening method for identifying Brassica napus plant, cell, seed, tissue or germplasm thereof comprising one or more MAM1 variants associated with decreased glucosinolate level. The method comprises providing a Brassica napus plant, cell, seed, tissue or germplasm thereof comprising one or more MAM1 variants disclosed herein (e.g., any of the foregoing disclosed examples of Brassica napus plant, cell, seed, tissue or germplasm thereof comprising one or more MAM1 variants), obtaining a sample comprising nucleic acid from the plant, cell, seed, tissue or germplasm thereof, then screening the sample for any of the following: a) the one or more MAM1 variants or b) one or more marker alleles that are genetically linked to the one or more of the MAM1 variants. The method can then further comprise detecting the one or more (i) MAM1 variants or (ii) marker alleles in the sample to thereby identifyBrassica napus plant, cell, seed, tissue or germplasm thereof as having an MAM1 variant associated with decreased glucosinolate level.
[0020] In still further embodiments, the method can additionally include selecting the Brassica napus plant, cell, seed, tissue or germplasm thereof identified as having an MAM1 variant associated with decreased glucosinolate level. The selected plant material can be used for breeding or trait introgression.
[0021] For example, the method of identifying can include screening the sample for the presence of a marker allele linked to the MAM1 variant, e.g., by 5 cM, 4 cM, 3 cM, 2 cM, 1 cM, 0.9 cM, 0.8 cM, 0.7 cM, 0.6 cM, 0.5 cM, 0.4 cM, 0.3 cM, 0.2 cM, 0.1 cM, or less on a single meiosis- based genetic map, and associated. In other examples, the method can include detecting one or more targeted alteration to one or more alleles of MAM1 on chromosome A2, C2, A3, C7, or A4. The types of targeted alterations include nonsense mutations, missense mutations, anddeletions that eliminate or reduce MAM1 gene function (e.g., a premature termination codon that reduces or eliminates expression of a full-length protein encoded by the altered MAM1 allele). Particular examples of a targeted alteration that can be detected in accordance with the method are shown in Table 3 for MAM1 on chromosome A2 (MAM1 ,A2), chromosome C2 (MAM1.C2), chromosome A3 (MAM1.A3), chromosome C7 (MAMl.C7a or MAMl .C7b), or chromosome A4 (MAM1. A4). The method can include detecting a premature termination codon is shown in each of SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQ ID NO: 16, SEQ ID NO: 17, SEQ ID NO: 18, SEQ ID NO: 19, SEQ ID NO:20, SEQ ID NO 21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ ID NO 38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NOTO. In other examples, this method can include detecting any combination of MAM1 variants disclosed herein (see definition above of terms “MAM1.A2 Variant”, “MAM1.C2 Variant,” “MAM1.A3 Variant”, ”MAMl.C7a Varaint”, “MAMl.C7b Variant”, or “MAM1.A4 Variant” and variants disclosed in Table 3, Table 4, or Table 5 herein): (i) any homozygous MAM1 ,A2 Variant described herein combined with any heterozygous or homozygous MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, MAMl .C7b Variant, or MAM1.A4 Variant described herein; (ii) any homozygous MAM1 ,C2 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.A3 Variant, MAMl .C7a Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (iii) any homozygous MAM1.A3 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1 ,C2 Variant, MAM1 ,C7a Variant, MAM1 ,C7b Variant, or MAM1 A4 Variant described herein; (iv) any homozygous MAMl.C7a Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (v) any homozygous MAMl.C7b Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant,MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAM1.A4 Variant described herein; or (vi) any homozygous MAM1.A4 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAMl.C7b Variant described herein.
[0022] In another aspect, disclosed herein is a method that includes crossing a Brassica napus plant disclosed herein comprising one or more MAM1 variants to a second plant that does not have the one or more MAM1 variants, thereby producing one or more progeny plants whose genome comprises the one or more MAM1 variants, i.e., any of the MAM1 variants or combinations disclosed herein and / or selected by the screening method disclosed herein. In a one example of this aspect, the second plant is one of a plant line (a “recurrent parent line”) and the method further includes crossing the progeny plant with another plant of the recurrent parent line to produce a second-generation progeny whose genome comprises the one or more MAM1 variants. Optionally, the second-generation progeny can be crossed with the recurrent parent line to produce a third-generation progeny whose genome comprises the one or more MAM1 variants. This process can be repeated three, four, five, six, seven, or more times, such that each subsequent generation progeny is crossed with the recurrent parent line, thereby introgressing the one or more MAM1 variants into the recurrent parent line.
[0023] In an alternative method, a plant having the one or more MAM1 variants disclosed herein is crossed with a second plant to produce progeny plants. The progeny plants are screened for the one or more MAM1 variants in accordance with the screening method disclosed herein. Generally, such screening includes obtaining a nucleic acid sample from each of the progeny plants and screening the sample for the presence of MAM1 variants; thereby identifying novel progeny plants comprising one of the MAM1 genes disclosed herein.
[0024] Provided herein is an isolated recombinant nucleic acid comprising one or more of SEQ ID NOs: 13-41, 50-70. In particular example, the isolated recombinant nucleic acid is a gene editing construct comprising a guide RNA comprising SEQ ID NO:42 or SEQ ID NO:45 for introducing an MAM1 variant disclosed herein. In another example, provided herein is a combination of primers for detecting an MAM1 variant disclosed herein, e.g., SEQ ID NO:43 and SEQ ID NO:44; or SEQ ID NO:46 and SEQ ID NO:47; or SEQ ID NO:48 and SEQ ID NO:49.
[0025] In another aspect, provided herein is a method of introducing an MAM1 variant into Brassica napus, the method comprising delivering to a cell or tissue of the Brassica napus a Cas endonuclease and a guide RNA targeting a sequence (target site) of a MAM1 gene on chromosome A2, C2, A3, C7, or A4. The endonuclease / guide RNA complex indues a targeted alteration at the target site which reduces or eliminates expression of the protein encoded by the targeted MAM1 gene, thereby introducing an MAM1 variant to the Brassica napus cell or tissue. In some examples, the method further comprises regenerating Brassica napus plant, cell, seed, tissue or germplasm thereof comprising the MAM1 variant. The guide RNA can be any one or more comprising SEQ ID NO:42 or SEQ ID NO:45. The MAM1 variant can be any MAM1 variant disclosed herein and shown in Table 3 for MAM1 on chromosome A2 (MAM1.A2), chromosome C2 (MAM1.C2), chromosome A3 (MAM1.A3), chromosome C7 (MAMl .C7a or MAM1 ,C7b), or chromosome A4 (MAM1 ,A4). Thus, the variant can include the premature termination codon shown in each of SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQ ID NO: 16, SEQ ID NO: 17, SEQ ID NO: 18, SEQ ID NO:19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO 37, SEQ ID NO:38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO 57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NO:70. Additionally the method can include introducing any combination of MAM1 variant alleles disclosed herein (see definition above of terms “MAM1.A2 Variant”, “MAM1.C2 Variant”, “MAM1.A3 Variant”, “MAMl.C7a Variant”, “MAMl.C7b Variant”, or “MAM1.A4 Variant” and variants disclosed in Table 3, Table 4, or Table 5 herein): (i) any homozygous MAM1.A2 Variant described herein combined with any heterozygous or homozygous MAM1.C2 Variant, MAM1.A3 Variant, MAMl .C7a Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (ii) any homozygous MAM1.C2 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, MAMl .C7b Variant, or MAM1.A4 Variant described herein; (iii) anyhomozygous MAM1.A3 Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAMl.C7a Variant, MAMl.C7b Variant, or MAM1. A4 Variant described herein; (iv) any homozygous MAM1 ,C7a Variant described herein combined with any heterozygous or homozygous MAM1 ,A2 Variant, MAM1 ,C2 Variant, MAM1.A3 Variant, MAMl.C7b Variant, or MAM1.A4 Variant described herein; (v) any homozygous MAMl.C7b Variant described herein combined with any heterozygous or homozygous MAM1.A2 Variant, MAM1.C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAM1.A4 Variant described herein; or (vi) any homozygous MAM1.A4 Variant described herein combined with any heterozygous or homozygous MAM1 ,A2 Variant, MAM1 ,C2 Variant, MAM1.A3 Variant, MAMl.C7a Variant, or MAMl.C7b Variant described herein.
[0026] Methods for introducing Cas endonucleases and guide RNAs are described in more detail herein.BRIEF DESCRIPTION OF DRAWINGS AND SEQUENCE LISTING
[0027] Figure 1 is a bar graph showing glucosinolate levels in seed of the plants grown in a field in 2023.
[0028] Nucleic acid sequences listed in the accompanying sequence listing and referenced herein are shown using standard letter abbreviations for nucleotide bases. While only one strand of each nucleic acid sequence is shown, the complementary strand is understood to be included in any reference to the displayed strand. Sequence listings are described in the following Table 1.Table 1DETAILED DESCRIPTION OF THE INVENTION
[0029] As used herein the singular forms “a”, “and”, and “the” include plural referents unless the context clearly dictates otherwise. Thus, for example, reference to “a cell” includes a plurality of such cells and reference to “the protein” includes reference to one or more proteins and equivalents thereof, and so forth. All technical and scientific terms used herein have the same meaning as commonly understood to one of ordinary skill in the art to which this disclosure belongs unless clearly indicated otherwise.
[0030] A gene or allele is “associated with” a trait when it is part of or linked to a DNA sequence or allele that affects the expression of the trait. The presence of the allele is an indicator of how the trait will be expressed.
[0031] “Brassica” refers to any one of Brassica napus (AACC, 2n=38), Brassica juncea (AABB, 2n=36), Brassica carinata (BBCC, 2n= 34), Brassica rapa (syn. B. campestris) (AA, 2n=20), Brassica oleracea (CC, 2n=18) ox Brassica nigra (BB, 2n= 16).
[0032] “Backcrossing” refers to the process whereby hybrid progeny are repeatedly crossed back to one of the parents. In a backcrossing scheme, the “donor” parent refers to the parental plant with the desired gene or locus to be introgressed. The “recipient” parent (used one or moretimes) or “recurrent” parent (used two or more times) refers to the parental plant into which the gene or locus is being introgressed.
[0033] “CRISPR” (Clustered Regularly Interspaced Short Palindromic Repeats) loci refers to certain genetic loci encoding components of DNA cleavage systems, for example, used by bacterial and archaeal cells to destroy foreign DNA (Horvath and Barrangou, 2010, Science 327: 167-170; International Application Publication W02007 / 025097, published 01 March 2007). A CRISPR locus can consist of a CRISPR array, comprising short direct repeats (CRISPR repeats) separated by short variable DNA sequences (called spacers), which can be flanked by diverse Cas (CRISPR-associated) genes.
[0034] The term “Cas protein” refers to a polypeptide encoded by a Cas (CRISPR-associated) gene. A Cas protein includes but is not limited to: a Cas9 protein, a Cpfl (Casl2) protein, a C2cl protein, a C2c2 protein, a C2c3 protein, Cas3, Cas3-HD, Cas 5, Cas7, Cas8, CaslO, or combinations or complexes of these. A Cas protein may be a “Cas endonuclease” or “Cas effector protein”, that when in complex with a suitable polynucleotide component, is capable of recognizing, binding to, and optionally nicking or cleaving all or part of a specific polynucleotide target sequence. A Cas endonuclease described herein comprises one or more nuclease domains. The endonucleases of the disclosure may include those having one or more RuvC nuclease domains. A Cas protein is further defined as a functional fragment or functional variant of a native Cas protein, or a protein that shares at least 50%, at least 55%, at least 60%, at least 65%, at least 70%, at least 75%, at least 80%, at least 85%, at least 90%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99%, or 100% sequence identity with a native Cas protein, and retains at least partial activity.
[0035] A “Cas endonuclease” may comprise domains that enable it to function as a doublestrand-break-inducing agent. A “Cas endonuclease” may also comprise one or more modifications or mutations that abolish or reduce its ability to cleave a double-strand polynucleotide (dCas). In some aspects, the Cas endonuclease molecule may retain the ability to nick a single-strand polynucleotide (for example, a D10A mutation in a Cas9 endonuclease molecule) (nCas9).
[0036] The term “crossed” or “cross” refers to a sexual cross and involved the fusion of two haploid gametes via pollination to produce diploid progeny (e.g., cells, seeds or plants). The termencompasses both the pollination of one plant by another and selfing (or self-pollination, e.g., when the pollen and ovule are from the same plant).
[0037] An “elite line” is any line that has resulted from breeding and selection for superior agronomic performance.
[0038] A “favorable allele” is the allele at a particular locus (a marker, a QTL, a gene etc.) that confers, or contributes to, an agronomically desirable phenotype, e.g., disease resistance, and that allows the identification of plants with that agronomically desirable phenotype. A favorable allele of a marker is a marker allele that segregates with the favorable phenotype.
[0039] “Gene” includes a nucleic acid fragment that expresses a functional molecule such as, but not limited to, a specific protein, including regulatory sequences preceding (5’ non-coding sequences) and following (3’ non-coding sequences) the coding sequence, as well as intervening intron sequences. “Native gene” refers to a gene as found in its natural endogenous location with its own regulatory sequences.
[0040] “Genetic markers” are nucleic acids that are polymorphic in a population and where the alleles of which can be detected and distinguished by one or more analytic methods, e.g., RFLP, AFLP, isozyme, SNP, SSR, and the like. The term also refers to nucleic acid sequences complementary to the genomic sequences, such as nucleic acids used as probes. Markers corresponding to genetic polymorphisms between members of a population can be detected by methods well-established in the art. These include, e.g., PCR-based sequence specific amplification methods, detection of restriction fragment length polymorphisms (RFLP), detection of isozyme markers, detection of polynucleotide polymorphisms by allele specific hybridization (ASH), detection of amplified variable sequences of the plant genome, detection of self-sustained sequence replication, detection of simple sequence repeats (SSRs), detection of single nucleotide polymorphisms (SNPs), or detection of amplified fragment length polymorphisms (AFLPs). Well established methods are also known for the detection of expressed sequence tags (ESTs) and SSR markers derived from EST sequences and randomly amplified polymorphic DNA (RAPD).
[0041] “Germplasm” refers to genetic material of or from an individual (e.g., a plant), a group of individuals (e.g., a plant line, variety or family), or a clone derived from a line, variety, species, or culture, or more generally, all individuals within a species or for several species (e.g., maize germplasm collection or Andean germplasm collection). The germplasm can be part of anorganism, cell, or can be separate from the organism or cell. In general, germplasm provides genetic material with a specific molecular makeup that provides a physical foundation for some or all of the hereditary qualities of an organism or cell culture. As used herein, germplasm includes cells, seed or tissues from which new plants may be grown, or plant parts, such as leaves, stems, pollen, or cells, that can be cultured into a whole plant.
[0042] The term “genome” as it applies to a prokaryotic and eukaryotic cell or organism cells encompasses not only chromosomal DNA found within the nucleus, but organelle DNA found within subcellular components (e.g., mitochondria, or plastid) of the cell.
[0043] As used herein, a “genomic sequence” or “genomic region” is a segment of a chromosome in the genome of a cell that is present on either side of the target site or, alternatively, also comprises the target site or a portion thereof. An “endogenous genomic sequence” refers to genomic sequence within a plant cell, (e.g. an endogenous genomic sequence of a MAM1 gene present within the genome of a Brassica plant cell).
[0044] A “genomic locus” as used herein refers to the genetic or physical location on a chromosome of a gene. As used herein, “gene” includes a nucleic acid fragment that expresses a functional molecule such as, but not limited to, a specific protein coding sequence and regulatory elements, such as those preceding (5’ non-coding sequences) and following (3’ non-coding sequences) the coding sequence.
[0045] As used herein, “genotype” is the actual nucleic acid sequence at one or more loci in an individual plant. As used herein, “phenotype” means the detectable characteristics (e.g. increased free phosphate) of a cell or organism which can be influenced by genotype.
[0046] “Glucosinolates” are P-thioglucoside N-hydroxysulfates with a variable side chain (R) and a sulfur-linked [3-d-glucopyranose moiety. They represent a large and heterogeneous family of naturally occurring compounds; more than 120 varieties are known to occur in nature (Fahey, et al., 2001). Glucosinolates are found in many species of plants, particularly those within the order Brassicales, but also among plants of the genus Drypetes and the genus Putranjiva (both genera of the Putranjivaceae family). They are accumulated to high levels in the seed, as well as other plant tissues. This is particularly true for the oil seed Brassicas. The glucosinolates found in meal samples of oilseeds include sinigrin, sinalbin, gluconapin, and gluconasturtin, among others, and their relative proportions can vary significantly depending on the species. These compounds and their metabolites can impact the taste of the meal, reducing its palatability and insome cases (dependent on the type of glucosinolate and glucosinolate metabolites present) can also adversely impact the health of an animal that has consumed plant material containing glucosinolates. Glucosinolate reduction in oil seed meal is an important and desirable objective and can have significant benefits in terms of meal value for animal feed.
[0047] As used herein, the term “guide polynucleotide”, relates to a polynucleotide sequence that can form a complex with a Cas endonuclease, including the Cas endonuclease described herein, and enables the Cas endonuclease to recognize, optionally bind to, and optionally cleave a DNA target site. The guide polynucleotide sequence can be a RNA sequence, a DNA sequence, or a combination thereof (a RNA-DNA combination sequence).
[0048] The terms “single guide RNA" and “sgRNA” are used interchangeably herein and relate to a synthetic fusion of two RNA molecules, a crRNA (CRISPR RNA) comprising a variable targeting domain (linked to a tracr mate sequence that hybridizes to a tracrRNA), fused to a tracrRNA (trans-activating CRISPR RNA). The single guide RNA can comprise a crRNA or crRNA fragment and a tracrRNA or tracrRNA fragment of the type II CRISPR / Cas system that can form a complex with a type II Cas endonuclease, wherein said guide RNA / Cas endonuclease complex can direct the Cas endonuclease to a DNA target site, enabling the Cas endonuclease to recognize, optionally bind to, and optionally nick or cleave (introduce a single or double-strand break) the DNA target site.
[0049] As used herein, the terms “guide polynucleotide / Cas endonuclease complex”, “guide polynucleotide / Cas endonuclease system”, “ guide polynucleotide / Cas complex”, “guide polynucleotide / Cas system”, “guided Cas system”, “Polynucleotide-guided endonuclease”, and “PGEN” are used interchangeably herein and refer to at least one guide polynucleotide and at least one Cas endonuclease, that are capable of forming a complex, wherein said guide polynucleotide / Cas endonuclease complex can direct the Cas endonuclease to a DNA target site, enabling the Cas endonuclease to recognize, bind to, and optionally nick or cleave (introduce a single or double-strand break) the DNA target site. A guide polynucleotide / Cas endonuclease complex herein can comprise Cas protein(s) and suitable polynucleotide component(s) of any of the known CRISPR systems (Horvath and Barrangou, 2010, Science 327: 167-170; Makarova et al. 2015, Nature Reviews Microbiology Vol. 13:1-15; Zetsche et al., 2015, Cell 163, 1-13;Shmakov et al., 2015, Molecular Cell 60, 1-13).
[0050] A “haplotype” is the genotype of an individual at a plurality of genetic loci, i.e. a combination of alleles. Typically, the genetic loci described by a haplotype are physically and genetically linked, i.e., on the same chromosome segment.
[0051] The term “heterogeneity” is used to indicate that individuals within the group differ in genotype at one or more specific loci.
[0052] The term “homogeneity” indicates that members of a group have the same genotype at one or more specific loci.
[0053] The term “hybrid” refers to the progeny obtained between the crossing of at least two genetically dissimilar parents.
[0054] The term “inbred” refers to a line that has been bred for genetic homogeneity.
[0055] The term “introgression” refers to the transmission of a desired allele of a genetic locus from one genetic background to another. For example, introgression of a desired R gene allele at a specified locus can be transmitted to at least one progeny via a sexual cross between two parents of the same species, where at least one of the parents has the desired allele in its genome. Alternatively, for example, transmission of an allele can occur by recombination between two donor genomes, e.g., in a fused protoplast, where at least one of the donor protoplasts has the desired allele in its genome. The desired allele can be, e.g., detected by a marker that is associated with a phenotype, at a QTL, a transgene, or the like. Offspring comprising the desired allele may be repeatedly backcrossed to a line having a desired genetic background and selected for the desired allele, to result in the allele becoming fixed in a selected genetic background.
[0056] The process of “introgressing” is often referred to as “backcrossing” when the process is repeated two or more times.
[0057] A “line” or “strain” is a group of individuals of identical parentage that are generally inbred to some degree and that are generally homozygous and homogeneous at most loci (isogenic or near isogenic). A “subline” refers to an inbred subset of descendants that are genetically distinct from other similarly inbred subsets descended from the same progenitor.
[0058] The term “plant material” includes whole plants, plant cells, plant protoplast, plant cell or tissue culture from which plants can be regenerated, plant calli, plant clumps and plant cells that are intact in plants, or parts of plants, such as seeds, flowers, cotyledons, leaves, stems, buds, roots, root tips and the like. As used herein, a “modified plant” means any plant that has a genetic change due to human intervention. A modified plant may have genetic changesintroduced through plant transformation, genome editing, mutagenesis, or conventional plant breeding.
[0059] A “marker” is a means of finding a position on a genetic or physical map, or else linkages among markers and trait loci (loci affecting traits). The position that the marker detects may be known via detection of polymorphic alleles and their genetic mapping, or else by hybridization, sequence match or amplification of a sequence that has been physically mapped. A marker can be a DNA marker (detects DNA polymorphisms), a protein (detects variation at an encoded polypeptide), or a simply inherited phenotype (such as a low-erucic acid oil profile). A DNA marker can be developed from genomic nucleotide sequence or from expressed nucleotide sequences (e.g., from a spliced RNA or a cDNA). Depending on the DNA marker technology, the marker may consist of primers complementary to sequence flanking the locus and / or probes that hybridize to polymorphic alleles at the locus. A DNA marker, or a genetic marker, may also be used to describe the gene, DNA sequence or nucleotide on the chromosome itself (rather than the components used to detect the gene or DNA sequence) and is often used when that DNA marker is associated with a particular trait in human genetics (e.g. a marker for breast cancer). The term marker locus is the locus (gene, sequence or nucleotide) that the marker detects.
[0060] Markers can be defined by the type of polymorphism that they detect and also the marker technology used to detect the polymorphism. Marker types include but are not limited to, e.g., detection of restriction fragment length polymorphisms (RFLP), detection of isozyme markers, randomly amplified polymorphic DNA (RAPD), amplified fragment length polymorphisms (AFLPs), detection of simple sequence repeats (SSRs), detection of amplified variable sequences of the plant genome, detection of self-sustained sequence replication, or detection of single nucleotide polymorphisms (SNPs). SNPs can be detected e.g. via DNA sequencing, PCR-based sequence specific amplification methods, detection of polynucleotide polymorphisms by allele specific hybridization (ASH), dynamic allele-specific hybridization (DASH), molecular beacons, microarray hybridization, oligonucleotide ligase assays, Flap endonucleases, 5’ endonucleases, primer extension, single strand conformation polymorphism (SSCP) or temperature gradient gel electrophoresis (TGGE). DNA sequencing, such as the pyrosequencing technology has the advantage of being able to detect a series of linked SNP alleles that constitute a haplotype. Haplotypes tend to be more informative (detect a higher level of polymorphism) than SNPs.
[0061] “Marker assisted selection” (of MAS) is a process by which individual plants are selected based on marker genotypes. “Marker assisted counter-selection” is a process by which marker genotypes are used to identify plants that will not be selected, allowing them to be removed from a breeding program or planting. A “marker haplotype” refers to a combination of alleles at a marker locus.
[0062] The term “molecular marker” may be used to refer to a genetic marker, as defined above, or an encoded product thereof (e.g., a protein) used as a point of reference when identifying a linked locus. A molecular marker can be derived from genomic nucleotide sequences or from expressed nucleotide sequences (e.g., from a spliced RNA, a cDNA, etc.), or from an encoded polypeptide. The term also refers to nucleic acid sequences complementary to or flanking the marker sequences, such as nucleic acids used as probes or primer pairs capable of amplifying the marker sequence. A “molecular marker probe” is a nucleic acid sequence or molecule that can be used to identify the presence of a marker locus, e.g., a nucleic acid probe that is complementary to a marker locus sequence. Alternatively, in some aspects, a marker probe refers to a probe of any type that is able to distinguish (i.e., genotype) the particular allele that is present at a marker locus. Nucleic acids are “complementary” when they specifically hybridize in solution. Some of the markers described herein are also referred to as hybridization markers when located on an indel region, such as the non-collinear region described herein. This is because the insertion region is, by definition, a polymorphism vis a vis a plant without the insertion. Thus, the marker need only indicate whether the indel region is present or absent. Any suitable marker detection technology may be used to identify such a hybridization marker, e.g. SNP technology is used in the examples provided herein.
[0063] “Meal” refers to the remaining fraction of the seed content after extraction of the oil and consists mainly of protein.
[0064] As used herein, a ‘nucleic acid molecule” is a polymeric form of nucleotides, which can include both sense and anti-sense strands of RNA, cDNA, genomic DNA, and synthetic forms and mixed polymers of the above. A nucleotide refers to a ribonucleotide, deoxynucleotide, or a modified form of either type of nucleotide. A "nucleic acid molecule" as used herein is synonymous with "nucleic acid", "nucleotide sequence", "nucleic acid sequence", and "polynucleotide." The term includes single- and double-stranded forms of DNA. A nucleic acidmolecule can include either or both naturally occurring and modified nucleotides linked together by naturally occurring and / or non-naturally occurring nucleotide linkages.
[0065] Nucleic acid molecules may be modified chemically or biochemically, or may contain non-natural or derivatized nucleotide bases, as will be readily appreciated by those of skill in the art. Such modifications include, for example, labels, methylation, substitution of one or more of the naturally occurring nucleotides with an analog, intemucleotide modifications, such as uncharged linkages (e.g., methyl phosphonates, phosphotriesters, phosphoramidates, carbamates, etc ), charged linkages (e.g., phosphorothioates, phosphorodithioates, etc.), pendent moieties (e.g., peptides), intercalators (e.g., acridine, psoralen, etc.), chelators, alkylators, and modified linkages (e.g., alpha anomeric nucleic acids, etc.). The term "nucleic acid molecule" also includes any topological conformation, including single-stranded, double-stranded, partially duplexed, triplexed, hairpinned, circular, and padlocked conformations. An "endogenous nucleic acid sequence" refers to a nucleic acid sequence within a plant cell, (e.g. an endogenous allele of an END gene present within the genome of a Brassica plant cell).
[0066] “Oilseed” refers to any crop species where oil is extracted from the seeds of these grains for food or industrial purposes, and includes Brassicaceae oilseeds such as canola, and non- Brassicaceae oilseeds, such as flaxseed, soybean, safflower, and sunflower. An example of a crop species that produces a seed used primarily for the production of edible oil is Brassica napus.
[0067] A “protospacer adjacent motif’ (PAM) herein refers to a short nucleotide sequence adjacent to a target sequence (protospacer) that is recognized (targeted) by a guide polynucleotide / Cas endonuclease system described herein. The Cas endonuclease may not successfully recognize a target DNA sequence if the target DNA sequence is not followed by a PAM sequence. The sequence and length of a PAM herein can differ depending on the Cas protein or Cas protein complex used. The PAM sequence can be of any length but is typically 1, 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, 13, 14, 15, 16, 17, 18, 19 or 20 nucleotides long.
[0068] As used herein, the term “plant material” refers to any processed or unprocessed material derived, in whole or in part, from a plant. For example, and without limitation, a plant material may be a plant part, a seed, a fruit, a leaf, a root, a plant tissue, a plant tissue culture, a plant explant, or a plant cell.
[0069] A “polymorphism” is a variation in the DNA between two or more individuals within a population. A polymorphism preferably has a frequency of at least 1% in a population. A useful polymorphism can include a single nucleotide polymorphism (SNP), a simple sequence repeat (SSR), or an insertion / deletion polymorphism, also referred to herein as an “indel”.
[0070] The term “quantitative trait locus” or “QTL” refers to a region of DNA that is associated with the differential expression of a quantitative phenotypic trait in at least one genetic background, e.g., in at least one breeding population. The region of the QTL encompasses or is closely linked to the gene or genes that affect the trait in question.
[0071] The terms “target site”, “target sequence”, “target site sequence, ’’target DNA”, “target locus”, “genomic target site”, “genomic target sequence”, “genomic target locus”, and “target polynucleotide”, can be used interchangeably herein and refer to a polynucleotide sequence such as, but not limited to, a nucleotide sequence on a chromosome, episome, a locus, or any other DNA molecule in the genome (including chromosomal, chloroplastic, mitochondrial DNA, plasmid DNA) of a cell, at which a guide polynucleotide / Cas endonuclease complex can recognize, bind to, and optionally nick or cleave. The target site can be an endogenous site in the genome of a cell, or alternatively, the target site can be heterologous to the cell and thereby not be naturally occurring in the genome of the cell, or the target site can be found in a heterologous genomic location compared to where it occurs in nature. As used herein, terms “endogenous target sequence” and “native target sequence” are used interchangeable herein to refer to a target sequence that is endogenous or native to the genome of a cell and is at the endogenous or native position of that target sequence in the genome of the cell.
[0072] A “targeted alteration” or “variant” is a gene (e.g., MAM1 gene) sequence that has been altered through human intervention. Such an “altered” or “modified” gene has a sequence that differs from the sequence of the corresponding native or non-altered gene by at least one nucleotide (i) insertion (i.e., addition of a nucleotide in a sequence), (ii) deletion, (iii) substitution (i.e., replacement of at least one nucleotide), or (iv) a combination of the foregoing alterations. An “altered” or “modified” plant is a plant comprising an altered gene sequence, e.g., a deletion. As used herein, a “targeted alteration” in a gene (referred to as the target gene) can be made by altering a target sequence within the target gene using any method known to one skilled in the art, including a method involving a guided Cas endonuclease system as disclosed herein.
[0073] A virus or vector “transforms” or “transduces” a cell when it transfers nucleic acid molecules into the cell. A cell is “transformed” by a nucleic acid molecule transduced into the cell when the nucleic acid molecule becomes stably replicated by the cell, either by incorporation of the nucleic acid molecule into the cellular genome, or by episomal replication. As used herein, the term “transformation” encompasses all techniques by which a nucleic acid molecule can be introduced into such a cell. Examples include, but are not limited to, transfection with viral vectors, transformation with plasmid vectors, electroporation (Fromm et al., 1986, Nature 319:791-3), lipofection (Feigner et al., 1987, Proc. Natl. Acad. Sci. USA 84:7413-7), microinjection (Mueller et al., 1978, Cell 15:579-85), Agrobacterium-mediated transfer (Fraley et al., 1983, Proc. Natl. Acad. Sci. USA 80:4803-7), direct DNA uptake, and microprojectile bombardment (Klein et al., 1987, Nature 327:70).
[0074] The term “variants” refer to substantially similar sequences. For polynucleotides, a variant comprises a deletion and / or addition of one or more nucleotides at one or more internal sites within the native polynucleotide and / or a substitution of one or more nucleotides at one or more sites in the native polynucleotide (e.g. a MAM1 variant disclosed herein). As used herein, a “native” polynucleotide or polypeptide comprises a naturally occurring nucleotide sequence or amino acid sequence, respectively.
[0075] The term “yield” refers to the productivity per unit area of a particular plant product of commercial value. Yield is affected by both genetic and environmental factors. “Agronomics,” “agronomic traits,” and “agronomic performance” refer to the traits (and underlying genetic elements) of a given plant variety that contribute to yield over the course of growing season. Individual agronomic traits include emergence vigor, vegetative vigor, stress tolerance, disease resistance or tolerance, herbicide resistance, branching, flowering, seed set, seed size, seed density, standability, threshability and the like. Yield can therefore be considered the final culmination of all agronomic traits.
[0076] The disclosed targeted alterations or variants of a MAM1 gene refer to human-made, intentionally produced and selected nucleic acid changes that can be generated by any known methods, including the use of targeted mutagenesis, Targeting Induced Local Lesions IN Genomes or TILLING (see e.g., McCallum et al., 2000, Nat Biotechnol 18:455-457), or the use of double-strand-break inducing agents (DSB Agents) or by chemical treatment, including theuse of ethyl methane sulfonate (EMS), methyl N-nitrosoguanidine (MNNG), ethidium bromide, diepoxybutane, or other mutagens known in the art.
[0077] Double- Strand-Break (DSB) Inducing Agents (DSB Agents). Double-strand breaks can be induced by agents such as endonucleases that cleave the phosphodiester bond within a polynucleotide chain, can result in the induction of DNA repair mechanisms, including the nonhom ologous end-joining pathway, and homologous recombination. Endonucleases include a range of different enzymes, including restriction endonucleases (see e.g. Roberts et al., 2003 Nucleic Acids Res 1 :418-20, Roberts et al., 2003, Nucleic Acids Res 31 : 1805-12, and Belfort et al., 2002 in Mobile DNA II, pp. 761-783, Eds. Craigie et al., (ASM Press, Washington, DC)), meganucleases (see e.g., International Application Publication WO 2009 / 114321; Gao et al., 2010, Plant loumal 1: 176-187), and TAL effector nucleases or TALENs (see e.g., US Application Publication US 20110145940 and Christian et al., 2010, Genetics 186(2): 757-61). Methods of targeting DNA double-strand breaks have been described for TALENs (Christian et al., 2010, Genetics 186(2): 757-61 and Boch et al., 2009, Science 326(5959): 1509-12), zinc finger nucleases (see e.g. Kim, et al., 1996, Proc. Nat’l Acad. Sci USA 93(3)1156-1160) and CRISPR-Cas endonucleases (see e.g. International Application Publication W02007 / 025097).
[0078] Any DSB or -nick or -modification inducing agent may be used for the methods described herein, including for example but not limited to: Cas endonucleases, recombinases, TALENs, zinc finger nucleases, restriction endonucleases, meganucleases, and deaminases.
[0079] Methods and compositions are provided for polynucleotide modification with a CRISPR Associated (Cas) endonuclease. Class I Cas endonucleases comprise multi-subunit effector complexes (Types I, III, and IV), while Class 2 systems comprise single protein effectors (Types II, V, and VI) (Makarova et al., 2015, Nature Reviews Microbiology 13:1-15; Zetsche et al., 2015, Cell 163: 1-13; Shmakov et al., 2015, Molecular Cell 60, 1-13; Haft et al., 2005, Computational Biology, PLoS Comput Biol 1(6): e60; and Koonin et al., 2017, Curr Opinion Microbiology 37:67-78). In Class 2 Type II systems, the Cas endonuclease acts in complex with a guide RNA (gRNA) that directs the Cas endonuclease to cleave the DNA target to enable target recognition, binding, and cleavage by the Cas endonuclease. The gRNA comprises a Cas endonuclease recognition (CER) domain that interacts with the Cas endonuclease, and a Variable Targeting (VT) domain that hybridizes to a nucleotide sequence in a target DNA. In some aspects, the gRNA comprises a CRISPR RNA (crRNA) and a trans-activating CRISPR RNA(tracrRNA) to guide the Cas endonuclease to its DNA target. The crRNA comprises a spacer region complementary to one strand of the double strand DNA target and a region that base pairs with the tracrRNA, forming an RNA duplex. In many systems, the Cas endonuclease-guide polynucleotide complex recognizes a short nucleotide sequence adjacent to the target sequence (protospacer), called a “protospacer adjacent motif’ (PAM).
[0080] Examples of a Cas endonuclease include but are not limited to Cas9, Casl2f, Casl2a or Cpfl, and variants thereof (See e.g., US Patent No. 10,934,536 and International Application Publication WO 2022 / 082179). Cas9 (formerly referred to as Cas5, Csnl, or Csxl2) is a Class 2 Type II Cas endonuclease (Makarova et al., 2015, Nature Reviews Microbiology 13: 1-15). For Cas9 and Casl2f, a Cas-gRNA complex recognizes a 3’ PAM sequence at the target site, permitting the spacer of the guide RNA to invade the double-stranded DNA target, and, if sufficient homology between the spacer and protospacer exists, generate a DSB cleavage. Cas9 endonucleases comprise RuvC and HNH domains that together produce DSBs, and separately can produce single strand breaks. For the S. pyogenes Cas9 endonuclease, the DSB leaves a blunt end. Cpfl is a Class 2 Type V Cas endonuclease and comprises nuclease RuvC domain but lacks an HNH domain (Yamane et al., 2016, Cell 165:949-962). Casl2f can generate 5’ staggered overhangs at DSB sites (Karvelis et al., Nucl Acids Res 48(12):5016-5023). Cpfl endonucleases create “sticky” overhang ends.
[0081] Some uses for Cas-gRNA systems at a genomic target site include but are not limited to insertions, deletions, substitutions, or modifications of one or more nucleotides at the target site; modifying or replacing nucleotide sequences of interest (such as a regulatory elements); insertion of polynucleotides of interest; gene dropout; gene knock-out; gene knock in; modification of splicing sites and / or introducing alternate splicing sites; modifications of nucleotide sequences encoding a protein of interest; amino acid and / or protein fusions; and gene silencing by expressing an inverted repeat into a gene of interest. Genome editing using DSB-inducing agents, such as Cas9-gRNA complexes, has been described, for example in U.S. Patent Application No. 2015 / 0082478, US Patent No. 10,934,536, International Application Publication WO2015 / 026886 Al, International Application Publication W02016007347, International Application Publication WO201625131, and International Application Publication WO 2022 / 082179 all of which are incorporated by reference herein.
[0082] In some aspects of the disclosure, a targeted genomic modification is introduced in a / f napus plant cell, wherein the targeted modification includes a targeted alteration of the genomic sequence of a MAM1 gene in the B. napus plant cell. In a further aspect, the targeted genomic modification is induced by a DSB Agent, such as a CRISPR-associated (Cas) nuclease. A Cas nuclease is introduced into the / i napus cell with a first and second guide RNAs as Cas-gRNA complexes that recognizes target sequences in the genome of the / f napus cell and is able to induce DSBs in the genomic sequence, e.g., thereby altering the endogenous target MAM1 gene.
[0083] Recombinant Constructs and Transformation of Cells. The disclosed guide polynucleotides can be introduced into a cell with the disclosed DSB agents e.g., CRISPR-Cas endonucleases. Cells include, but are not limited to, human, non-human, animal, bacterial, fungal, insect, yeast, non-conventional yeast, and plant cells as well as plants and seeds produced by the methods described herein. In a preferred aspect of the disclosure, the cells are B. napus cells.
[0084] Standard recombinant DNA and molecular cloning techniques used herein are known in the art and are described more fully in Sambrook et al., Molecular Cloning: A Laboratory Manual; Cold Spring Harbor Laboratory: Cold Spring Harbor, NY (1989). Transformation methods are well known to those skilled in the art and are described infra.
[0085] Vectors and constructs include circular plasmids, and linear polynucleotides, comprising a polynucleotide of interest and optionally other components including linkers, adapters, regulatory or analysis. In some examples a recognition site and / or target site can be comprised within an intron, coding sequence, 5' UTRs, 3' UTRs, and / or regulatory regions.
[0086] In one aspect, the constructs of the disclosure comprise a promoter operably linked to a nucleotide sequence encoding a DSB Agent, such as a CAS nuclease (e.g., gene encoding a Streptococcus pyrogenes Cas9 gene or Casl2f gene) and a promoter operably linked to a guide RNA of the present disclosure. The promoter is capable of driving expression of an operably linked nucleotide sequence in a prokaryotic or eukaryotic cell / organism. In some aspects, target specific guide RNAs are built as a fusion of CRISPR RNA (crRNA) fused to trans-activating CRISPR RNA (tracrRNA) of Streptococcus pyrogenes.
[0087] In accordance with the methods disclosed herein, a guide RNA comprising SEQ ID NO:42 or SEQ ID NO:45, can be used to alter endogenous genomic MAM1 sequence in the B. napus plant cell. The resulting targeted alterations can produce a MAM1 variant comprising apremature termination codon as shown in any of SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQ ID NO: 16, SEQ ID NO: 17, SEQ ID NO:18, SEQ ID NO 19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NOT 1, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO 36, SEQ ID NO:37, SEQ ID NO:38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO 56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NOTO.
[0088] The isolated polynucleotides, constructs and vectors disclosed herein (e.g., for expression of endonucleases or guide RNAs) can comprise a selectable marker to identify or select for or against a molecule or a cell that comprises the construct or vector. Examples of selectable markers that can be used in a construct or vector disclosed herein include DsRed and Glyphosate N-Acetyltransferase (GAT) gene variant 4621 for herbicide resistance.
[0089] Isolated Nucleic Acid Molecules and Variants and Fragments Thereof. Isolated or recombinant nucleic acid molecules comprising MAM1 variants disclosed herein as well as active portions thereof, as well as nucleic acid molecules sufficient for use as hybridization probes to identify MAM1 variants by sequence homology are provided. As used herein, the term “nucleic acid molecule” refers to DNA molecules (e.g., recombinant DNA, cDNA, genomic DNA, plastid DNA, mitochondrial DNA) and RNA molecules (e.g., mRNA) and analogs of the DNA or RNA generated using nucleotide analogs. In some examples, the nucleic acid molecule can be single- stranded. In some examples, the nucleic acid molecule can be double-stranded.
[0090] An “isolated” nucleic acid molecule (e.g., RNA or DNA) is used herein to refer to a nucleic acid sequence (e.g., RNA or DNA) that is no longer in its natural environment, for example in vitro. A “recombinant” nucleic acid molecule (e.g., RNA or DNA) is used herein to refer to a nucleic acid sequence (e.g., RNA or DNA) that is in a recombinant bacterial or plant host cell; has been edited from its native sequence; or is located in a different location than the native sequence. In some embodiments, an “isolated” or “recombinant” nucleic acid is free of sequences (preferably protein encoding sequences) that naturally flank the nucleic acid (i.e.,sequences located at the 5' and 3' ends of the nucleic acid) in the genomic DNA of the organism from which the nucleic acid is derived. For purposes of the disclosure, “isolated” or “recombinant” when used to refer to nucleic acid molecules excludes isolated chromosomes. For example, in various embodiments, the recombinant nucleic acid molecules can contain less than about 5 kb, 4 kb, 3 kb, 2 kb, 1 kb, 0.5 kb or 0.1 kb of nucleic acid sequences that naturally flank the MAM1 variant in the genome of the cell.
[0091] In some embodiments, an isolated nucleic acid molecule comprising a MAM 1 variant has one or more change in the nucleic acid sequence compared to the native or genomic nucleic acid sequence. In some embodiments, the change in the native or genomic nucleic acid sequence includes but is not limited to: changes in the nucleic acid sequence due to the degeneracy of the genetic code; changes in the nucleic acid sequence due to the amino acid substitution, insertion, deletion and / or addition compared to the native or genomic sequence; removal of one or more intron; deletion of one or more upstream or downstream regulatory regions; and deletion of the 5’ and / or 3’ untranslated region associated with the genomic nucleic acid sequence. In some embodiments, the nucleic acid molecule comprising one of SEQ ID NOs: 13-41, 50-70 is non- genomic sequence.
[0092] A variety of polynucleotides comprising MAM1 variants disclosed herein are contemplated. Such polynucleotides are useful for production of encoded polypeptides in host cells when operably linked to a suitable promoter, transcription termination and / or polyadenylation sequences. Such polynucleotides are also useful as probes for isolating homologous or substantially homologous polynucleotides that are MAM1 variants or related to MAM1 variants disclosed herein.
[0093] Provided herein are nucleic acid molecules comprising one or more of SEQ ID NOs: 13- 41, 50-70, and variants, fragments and complements thereof. “Complement” is used herein to refer to a nucleic acid sequence that is sufficiently complementary to a given nucleic acid sequence such that it can hybridize to the given nucleic acid sequence to thereby form a stable duplex. A reverse complement is a complement formed by exchanging each A with T, T with A, C with G, and G with C in a sequence and then reversing the 5’ to 3’ order of the exchanged sequence, such that the reverse complement of 5’-ACCTGAG-3’ is 5’-CTCAGGT-3’. “Polynucleotide sequence variants” is used herein to refer to a nucleic acid sequence that except for the degeneracy of the genetic code encodes the same polypeptide.
[0094] “Percent (%) sequence identity” with respect to a reference sequence (subject) is determined as the percentage of amino acid residues or nucleotides in a candidate sequence (query) that are identical with the respective amino acid residues or nucleotides in the reference sequence, after aligning the sequences and introducing gaps, if necessary, to achieve the maximum percent sequence identity, and not considering any amino acid conservative substitutions as part of the sequence identity. Alignment for purposes of determining percent sequence identity can be achieved in various ways, for instance, using publicly available computer software such as BLAST, BLAST-2. Those skilled in the art can determine appropriate parameters for aligning sequences, including any algorithms needed to achieve maximal alignment over the full length of the sequences being compared. The percent identity between the two sequences is a function of the number of identical positions shared by the sequences (e.g., percent identity of query sequence = number of identical positions between query and subject sequences / total number of positions of query sequence x 100).
[0095] Nucleotide Constructs, Expression Cassettes and Vectors. The use of the term “construct” in connection with isolated and / or heterologous polynucleotides herein is not intended to limit the disclosure to constructs comprising DNA. Polynucleotide constructs, particularly polynucleotides and oligonucleotides composed of ribonucleotides and combinations of ribonucleotides and deoxyribonucleotides, may also be employed in the methods disclosed herein. The isolated polynucleotide constructs, nucleic acids, and nucleotide sequences disclosed herein additionally encompass all complementary forms (e.g., the reverse complement) of each sequence disclosed for such a construct. Further, polynucleotide constructs and nucleotide sequences disclosed herein can encompass any such constructs, molecules, and sequences suitable for use in a method for transforming plant material disclosed herein. Such constructs can include naturally occurring molecules and / or synthetic analogues. The disclosed nucleotide constructs, nucleic acids, and nucleotide sequences also encompass all forms of nucleotide constructs including, but not limited to, single-stranded forms, double-stranded forms, hairpins, stem-and-loop structures and the like.
[0096] Transformed organisms disclosed herein include plant cells, bacteria, yeast, baculovirus, protozoa, nematodes and algae. The transformed organism comprises a disclosed sequence (e.g., as part of a construct, expression cassette, or vector comprising the nucleotide sequence disclosed herein which are associated with increased disease resistance.
[0097] The disclosed sequences can be used in constructs for expression in the organism of interest. Constructs can include 5’ and 3’; regulatory sequences operably linked to an R gene sequence, variant or fragment disclosed herein. The term “operably linked” as used herein refers to a functional linkage between a promoter and / or a regulatory sequence and a second sequence, wherein the promoter and / or regulatory sequence initiates, mediates, and / or affects transcription of the DNA sequence corresponding to the second sequence. Generally, operably linked means that the nucleic acid sequences being linked are contiguous and, where necessary, to join two protein coding regions in the same reading frame. The construct may additionally contain at least one additional gene to be cotransformed into the organism. Alternatively, the additional gene(s) can be provided on multiple DNA constructs.
[0098] Such a DNA construct is provided with a plurality of restriction sites for insertion of the polypeptide gene sequence of the disclosure to be under the transcriptional regulation of the regulatory regions. The DNA construct may additionally contain selectable marker genes.
[0099] The DNA construct will generally include in the 5' to 3' direction of transcription: a transcriptional and translational initiation region (e.g., a promoter), a DNA sequence of the embodiments, and a transcriptional and translational termination region (e.g., termination region) functional in the organism serving as a host. The transcriptional initiation region (e.g., the promoter) may be native, analogous, foreign or heterologous to the host organism and / or to the sequence of the embodiments. Additionally, the promoter or regulatory sequence may be the natural sequence or alternatively a synthetic sequence. The term “foreign” as used herein indicates that the promoter is not found in the native organism into which the promoter is introduced. As used herein, the term “heterologous” in reference to a sequence means a sequence that originates from a foreign species or, if from the same species, is substantially modified from its native form in composition and / or genomic locus by deliberate human intervention. As used herein, a chimeric gene comprises a coding sequence operably linked to a transcription initiation region that is heterologous to the coding sequence. Where the promoter is a native or natural sequence, the expression of the operably linked sequence is altered from the wild-type expression, which results in an alteration in phenotype.
[0100] In some embodiments the DNA construct comprises a polynucleotide comprising one or more of SEQ ID NOs: 13-41, 50-70 or a fragment or variant thereof.
[0101] A DNA construct may also include a transcriptional enhancer sequence. An “enhancer” refers to a DNA sequence which can stimulate promoter activity and may be an innate element of the promoter or a heterologous element inserted to enhance the level or tissue-specificity of a promoter. Various enhancers include, for example, introns with gene expression enhancing properties in plants (US Patent Application Publication Number 2009 / 0144863, the ubiquitin intron (i.e., the maize ubiquitin intron 1 (see, for example, NCBI sequence S94464)), the omega enhancer or the omega prime enhancer (Gallie et al. 1989 Molecular Biology of RNA ed. Cech (Liss, New York) 237-256 and Gallie et al. 1987 Gene 60: 217-25), the CaMV 35S enhancer (see, e.g., Benfey et al. 1990 EMBO J. 9:1685-96) and the enhancers of US Patent Number 7,803,992 may also be used. The above list of transcriptional enhancers is not meant to be limiting. Any appropriate transcriptional enhancer can be used in the embodiments.
[0102] The termination region may be native with the transcriptional initiation region, may be native with the operably linked DNA sequence of interest, may be native with the plant host or may be derived from another source (i.e., foreign or heterologous to the promoter, the sequence of interest, the plant host or any combination thereof).
[0103] Convenient termination regions are available from the Ti-plasmid of A. tumefciciens, such as the octopine synthase and nopaline synthase termination regions. See also, Guerineau et al. 1991 Mol. Gen. Genet. 262: 141-144; Proudfoot 1991 Cell 64:671-674; Sanfacon et al. 1991 Genes Dev. 5:141-149; Mogen et al. 1990 Plant Cell 2: 1261-1272; Munroe et al. 1990 Gene 91 :151-158; Ballas et al. 1989 Nucleic Acids Res. 17:7891-7903 and Joshi et al. 1987 Nucleic Acid Res. 15:9627-9639.
[0104] Where appropriate, a nucleic acid may be optimized for increased expression in the host organism. Thus, where the host organism is a plant, the synthetic nucleic acids can be synthesized using plant-preferred codons for improved expression. See, for example, Campbell and Gowri 1990 Plant Physiol. 92: 1-11 for a discussion of host-preferred usage. For example, although nucleic acid sequences of the embodiments may be expressed in both monocotyledonous and dicotyledonous plant species, sequences can be modified to account for the specific preferences and GC content preferences of monocotyledons or dicotyledons as these preferences have been shown to differ (Murray et al. 1989 Nucleic Acids Res. 17:477-498).Thus, the plant-preferred for a particular amino acid may be derived from known gene sequences from plants.
[0105] Additional sequence modifications are known to enhance gene expression in a cellular host. These include elimination of sequences encoding spurious polyadenylation signals, exonintron splice site signals, transposon-like repeats, and other well-characterized sequences that may be deleterious to gene expression. The GC content of the sequence may be adjusted to levels average for a given cellular host, as calculated by reference to known genes expressed in the host cell. The term “host cell” as used herein refers to a cell which contains a vector and supports the replication and / or expression of the expression vector is intended. Host cells may be prokaryotic cells such as E. coli or eukaryotic cells such as yeast, insect, amphibian or mammalian cells or monocotyledonous or dicotyledonous plant cells. An example of a monocotyledonous host cell is a maize host cell. When possible, the sequence is modified to avoid predicted hairpin secondary mRNA structures.
[0106] In preparing the expression cassette, the various DNA fragments may be manipulated so as to provide for the DNA sequences in the proper orientation and, as appropriate, in the proper reading frame. Toward this end, adapters or linkers may be employed to join the DNA fragments or other manipulations may be involved to provide for convenient restriction sites, removal of superfluous DNA, removal of restriction sites or the like. For this purpose, in vitro mutagenesis, primer repair, restriction, annealing, resubstitutions, e.g., transitions and transversions, may be involved.
[0107] A number of promoters can be used in the practice of the embodiments. The promoters can be selected based on the desired outcome. The nucleic acids can be combined with constitutive, tissue-preferred, inducible, or other promoters for expression in Brassica plants. “Tissue-preferred” promoters can be utilized to target enhanced expression of the gene of interest within a particular plant tissue. Such promoters can be modified, if necessary, for weak expression. “Seed-preferred” promoters include both “seed development” promoters (those promoters preferentially active during seed development such as promoters of seed storage proteins) as well as “seed-germinating” promoters (those promoters preferentially active during seed germination). An "inducible” promoter refers to a promoter that selectively express a coding sequence or functional RNA in response to the presence of an endogenous or exogenous stimulus, for example by chemical compounds (chemical inducers) or in response to environmental, hormonal, chemical, and / or developmental signals. Inducible or regulated promoters include, for example, promoters induced or regulated by light, heat, stress, flooding ordrought, salt stress, osmotic stress, phytohormones, wounding, or chemicals such as ethanol, abscisic acid (ABA), jasmonate, salicylic acid, or safeners.
[0108] Plant Transformation. Any suitable techniques known in the art for introduction of transgenes into plants may be used to produce a transformed plant or plant cell disclosed herein. Suitable methods for transformation of plants may include virtually any method by which DNA can be introduced into a cell, such as: by electroporation as illustrated in U.S. Patent No. 5,384,253; by microprojectile bombardment, as illustrated in U.S. Patent Nos. 5,015,580, 5,550,318, 5,538,880, 6,160,208, 6,399,861, and 6,403,865; by Agrobacterium-mediated transformation as illustrated in U.S. Patent Nos. 5,635,055, 5,824,877, 5,591,616; 5,981,840, and 6,384,301; and by protoplast transformation, as set forth in U.S. Patent No. 5,508,184, etc. These techniques can be used to transform plant cells and these cells may be developed into transgenic plants by techniques known to those of skill in the art. Techniques for transforming Brassica plants in particular are disclosed, for example, in U.S. Patent No. 5,750,871.
[0109] After effecting delivery of exogenous DNA to recipient cells, transformed cells are identified for further culturing and plant regeneration. In order to improve the ability to identify transformants, one may desire to employ a selectable marker gene with the transformation vector used to generate the transformant. In this case, the potentially transformed cell population can be assayed by exposing the cells to a selective agent or agents, or the cells can be screened for the desired marker.
[0110] Cells that survive the exposure to the selective agent, or cells that have been scored positive in a screening assay, may be cultured in media that supports regeneration of plants. In some embodiments, any suitable plant tissue culture media may be modified by including further substances, such as growth regulators. Tissue may be maintained on a basic media with growth regulators until sufficient tissue is available to begin plant regeneration efforts, or following repeated rounds of manual selection, until the morphology of the tissue is suitable for regeneration (e.g., at least 2 weeks), then transferred to media conducive to shoot formation. Cultures are transferred periodically until sufficient shoot formation has occurred. Once shoots are formed, they are transferred to media conducive to root formation. Once sufficient roots are formed, plants can be transferred to soil for further growth and maturity.[OHl] The alteration (e.g., introduction of a stop codon, mutation or deletion) of an endogenous gene (e.g., MAM1) in regenerating plants can be confirmed by one or more assays, for example,a molecular biological assay, such as Southern blotting, Northern blotting, or PCR; a biochemical assay, such as detecting the absence of a protein product by immunoassay (ELISA or Western blot) or by screening for reduced enzymatic function; plant part assays, such as leaf or root assays; and analysis of the phenotype of the whole regenerated plant.
[0112] Using the methods disclosed herein, MAM1 variant-containing plants are generated. For example, a plant comprising an altered endogenous genomic sequence containing one or more of the following MAM1 variants: SEQ ID N0: 13, SEQ ID N0: 14, SEQ ID N0:15, SEQ ID NO:16, SEQ ID NO: 17, SEQ ID NO: 18, SEQ ID NO: 19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NOT 1, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ ID NO:38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO 41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:41, SEQ ID NO:50, SEQ ID NO:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62, SEQ ID NO 63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NOTO. In some examples, the modified B. napus plant comprises multiple homozygous variants of MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, or MAM1.A4.
[0113] “Stable transformation” as used herein means that the nucleotide construct introduced into a plant integrates into the genome of the plant and is capable of being inherited by the progeny thereof. “Transient transformation” as used herein means that a polynucleotide is introduced into the plant and does not integrate into the genome of the plant or a polypeptide is introduced into a plant. “Plant” as used herein refers to whole plants, plant organs (e.g., leaves, stems, roots, etc.), seeds, plant cells, propagules, embryos and progeny of the same. Plant cells can be differentiated or undifferentiated (e.g. callus, suspension culture cells, protoplasts, leaf cells, root cells, phloem cells and pollen).
[0114] Transformation protocols as well as protocols for introducing nucleotide sequences into plants may vary depending on the type of plant or plant cell, i.e., monocot or dicot, targeted for transformation. Suitable methods of introducing nucleotide sequences into plant cells and subsequent insertion into the plant genome include microinjection (Crossway et al. (1986) Biotechniques 4:320-334), electroporation (Riggs et al. 1986 Proc. Natl. Acad. Sci. USA83:5602-5606), Agrobacterium-mediated transformation (US Patent Numbers 5,563,055 and 5,981,840), direct gene transfer (Paszkowski et al. 1984 EMBO J. 3:2717-2722) and ballistic particle acceleration (see, for example, US Patent Numbers 4,945,050; 5,879,918; 5,886,244 and 5,932,782; Tomes et al. 1995 in Plant Cell, Tissue, and Organ Culture: Fundamental Methods, ed. Gamborg and Phillips (Springer-Verlag, Berlin) and McCabe et al. 1988 Biotechnology 6:923-926) and Led transformation (WO 00 / 28058). For potato transformation see, Tu et al. 1998 Plant Molecular Biology 37:829-838 and Chong et al. 2000 Transgenic Research 9:71-78. Additional transformation procedures can be found in Weissinger et al. 1988 Ann. Rev. Genet. 22:421-477; Sanford et al. 1987 Particulate Science and Technology 5:27-37 (onion); Christou et al. 1988 Plant Physiol. 87:671-674 (soybean); McCabe et al. 1988 Bio / Technology 6:923-926 (soybean); Finer and McMullen 1991 In Vitro Cell Dev. Biol. 27P: 175-182 (soybean); Singh et al. 1998 Theor. Appl. Genet. 96:319-324 (soybean); Datta et al. 1990 Biotechnology 8:736-740 (rice); Klein et al. 1988 Proc. Natl. Acad. Sci. USA 85:4305-4309 (maize); Klein et al. 1988 Biotechnology 6:559-563 (maize); US Patent Numbers 5,240,855; 5,322,783 and 5,324,646; Klein et al. (1988) Plant Physiol. 91:440-444 (maize); Fromm et al. 1990 Biotechnology 8:833- 839 (maize); Hooykaas-Van Slogteren et al. 1984 Nature (London) 311 :763-764; US Patent Number 5,736,369 (cereals); Bytebier et al. 1987 Proc. Natl. Acad. Sci. USA 84:5345-5349 (Liliaceae); De Wet et al. 1985 in The Experimental Manipulation of Ovule Tissues, ed. Chapman et al. (Longman, New York), pp. 197-209 (pollen); Kaeppler et al. 1990 Plant Cell Reports 9:415-418 and Kaeppler et al. 1992 Theor. Appl. Genet. 84:560-566 (whisker-mediated transformation); D'Halluin et al. 1992 Plant Cell 4: 1495-1505 (electroporation); Li et al. (1993) Plant Cell Reports 12:250-255 and Christou and Ford 1995 Annals of Botany 75:407-413 (rice); Osjoda et al. 1996 Nature Biotechnology 14:745-750 (maize n Agrobacterium tumefaciens).
[0115] Marker assisted selection. Molecular markers can be used in a variety of plant breeding applications (e.g. see Staub et al. 1996 Hortscience 31:729-741; Tanksley 1983 Plant Molecular Biology Reporter. 1 :3-8). One of the main areas of interest is to increase the efficiency of backcrossing and introgressing genes using marker-assisted selection (MAS). Thus, MAS can be used for backcrossing and introgressing the MAM1 variants disclosed herein.
[0116] A molecular marker that demonstrates linkage with a locus affecting a desired phenotypic trait provides a useful tool for the selection of the trait in a plant population. This is particularly true where the phenotype is hard to assay. Since DNA marker assays are less laborious and takeup less physical space than field phenotyping, much larger populations can be assayed, increasing the chances of finding a recombinant with the target segment from the donor line moved to the recipient line. The closer the linkage, the more useful the marker, as recombination is less likely to occur between the marker and the gene causing the trait, which can result in false positives. Having flanking markers decreases the chances that false positive selection will occur as a double recombination event would be needed. In the most preferred case, a marker is located within the gene itself, so that recombination cannot occur between the marker and the gene. In some embodiments, the methods targeted alteration disclosed herein produce a marker that can be used to identify the MAM1 variant.
[0117] Gene introgression by MAS can be used to diminish linkage drag and yield drag. Gepts 2002 Crop Sci; 42: 1780-1790; Young et al. 1998 Genetics 120:579-585; Tanksley et al. 1989 Biotechnology 7: 257-264. Markers disclosed herein, as well as other marker types such as SSRs and FLPs, can be used in marker assisted selection protocols. SSRs can be defined as relatively short runs of tandemly repeated DNA with lengths of 6 bp or less, which are often highly suited to mapping and MAS. Tautz 1989 Nucleic Acid Research 17: 6463-6471; Wang et al. 1994 Theoretical and Applied Genetics, 88: 1-6; Levinson and Gutman 1987 Mol Biol Evol 4: 203- 221; Weber and May 1989 Am J Hum Genet. 44:388-396; Rafalski et al. 1996 Generating and using DNA markers in plants. In: Non-mammalian genomic analysis: a practical guide.Academic press, pp 75-135). FLP markers refer to fragment length polymorphisms that are in many ways similar to SSR markers, except that the region amplified by the primers is not typically a highly repetitive region. Bhattramakki et al. 2002 Plant Mol Biol 48, 539-547; Rafalski 2002b, supra.
[0118] SNP markers detect single base pair nucleotide substitutions, which can be assayed at an even higher level of throughput than SSRs, in so-called “ultra-high-throughput” fashion, as SNPs do not require large amounts of DNA and automation of the assay may be straight-forward.SNPs also have the promise of being relatively low-cost systems. These three factors together make SNPs highly attractive for use in MAS. Several methods are available for SNP genotyping, including but not limited to, hybridization, primer extension, oligonucleotide ligation, nuclease cleavage, mini sequencing, and coded spheres. Such methods have been reviewed in: Gut 2001 Hum Mutat 17: 475-492; Shi 2001 Clin Chem 47: 164-172; Kwok 2000 Pharmacogenomics 1: 95-100; and Bhattramakki and Rafalski 2001 Discovery and application of single nucleotidepolymorphism markers in plants. In: R. J. Henry, Ed, Plant Genotyping: The DNA Fingerprinting of Plants, CABI Publishing, Wallingford. A wide range of commercially available technologies utilize these and other methods to interrogate SNPs including Masscode. TM. (Qiagen), INVADER®. (Third Wave Technologies) and Invader PLUS®, SNAPSHOT®. (Applied Biosystems), TAQMAN®. (Applied Biosystems) and BEAD ARRAYS®. (Illumina).
[0119] In addition to SSR's, FLPs and SNPs, as described above, other types of molecular markers are also widely used, including but not limited to expressed sequence tags (ESTs), SSR markers derived from EST sequences, randomly amplified polymorphic DNA (RAPD), and other nucleic acid based markers. Isozyme profiles and linked morphological characteristics can, in some cases, also be indirectly used as markers. Even though they do not directly detect DNA differences, they are often influenced by specific genetic differences. However, markers that detect DNA variation are far more numerous and polymorphic than isozyme or morphological markers (Tanksley 1983 Plant Molecular Biology Reporter 1 : 3-8).
[0120] The following are examples of specific embodiments of some aspects of the invention. The examples are offered for illustrative purposes only and are not intended to limit the scope of the invention in any way.EXAMPLESExample 1
[0121] Six Brassica MAM1 genes, MAM1.A2, MAM1.C2, MAM1.A3, MAMl .C7a, MAMl.C7b and MAM1.A4 on chromosomes A2, C2, A3, C7 and A4, respectively, were selected for generating knockout variants in elite Brassica napus inbred line BC. Two guide RNAs, BNA-MAMl-CRla and BNA-MAMl-CRlb (Table 2), were designed to target the first exon ofMAMl .A2 and C2 and MAM1.A3, C7a, C7b and A4, respectively, and induce a double stranded cut in the MAM1 gene, leaving two free chromosomal ends. The plant’s repair mechanisms can attempt to repair the double-strand DNA break by non-homologous end joining (NHEJ), which can result in nucleotide insertions and deletions (Indels) at the cleavage site. Among these variants, those that result in a coding sequence frameshift are referred to as knockouts or KOs.Table 2
[0122] Briefly, epicotyl explants of canola inbred BC were incubated w h Agrobacterium carrying a CRISPR-Cas9 plasmid for plant transformation (Chu et al., 2020). Regenerated plantlets were screened using a NGS sequencing assay to obtain knockout variants, shown in Table 3 and combination of variants shown in Table 4 and Table 5. Positive TO plants were selfpollinated to produce T1 seeds. T1 seeds were planted in growth chamber and seedlings were genotyped with the same NGS sequencing assay to identify homozygous knockout plants of singles and combination of the MAM1 genes as well as wildtype (WT) segregants. T-DNA was incorporated into the genome in TO plants. Therefore, PCR assays were used to screen T1 plants to select null segregants that contain desired MAM1 variants but are free of gene-editing reagent plasmids. Selected clean T1 plants were self-pollinated to generate T2 seeds. In Table 3, a period indicates the double-stranded break point that was targeted; underlined residues represent nucleotide insertions; dashes represent nucleotide deletions.Table 3
[0123] Knocking out MAM1 was intended to reduce glucosinolate concentrations in Brassica seeds. The glucosinolate content was determined via high performance liquid chromatography (HPLC). In brief, 10 mg of hexane-defatted and dried canola meal was added to 1 ml of 70:30 (v:v) methanol: water (both HPLC grade) in screw capped tubes and heated at 70 °C under sonication for 30 minutes with grinding using a Geno / Grinder at the 15-minute mark (10 secondsat 12,000 rpm). After 30 minutes, samples were ground at same time and speed and then centrifuged at 3,700 rpm for 5 minutes at room temperature. 250 pl of supernatant was then transferred and dried using a SpeedVac concentrator and then reconstituted in 250 pl water and filtered with 0.2 pm filter (Thomson filter vials). 5 pl of this solution was injected onto HPLC flowing at 0.18 ml / minute featuring a Waters Acquity UPLC HSS T3 column (2.1 x 150 mm). Aqueous phase (A) contained 5 mM ammonium acetate and organic contained 0.1% (v%) formic acid in acetronitrile. Gradient elution: 98% A for 0.5 minutes then ramped to 55% A over the next 10.25 minutes, then to 50% A over next 0.45 minutes, then ramped back to 98% A over next 0.3 minutes and re-equilibrated for the next 4 minutes. UV detector set at 229 nm at 9 nm bandwidth at 20 Hz. Sinigrin standard curve was collected at concentrations at approximately 8, 16, 32 64, 128, 224, 320, 501 and 1001 nmoles / ml. Response factors used were 1.09 (progoitrin), 1.00 (sinigrin), 1.00 (glucoraphanin), 1.07 (glucoalyssin), 1.00 (gluconapoleiferin),I.11 (gluconapin), 0.28 (4-hydroxyglucobrassicin), 1.15 (glucobrassicanapin), 1.00 (glucoerucin), 0.29 (glucobrassicin), 0.95 (gluconasturtiin) and 0.2 (neoglucobrassicin).
[0124] In wildtype BC seeds produced in growth chamber, the glucosinolate concentration wasI I.5 pmoles g-1. T2 seeds of the MAM1 plants with 6 genes knocked out contained 7.4 pmoles of glucosinolates per gram seed (Table 3). The reduction in total glucosinolates was caused mainly by a reduction in aliphatic glucosinolates, from 4.24 pmoles g-1 in wildtype to 0.58 pmoles g-1 in knockouts. Knocking out MAM1 had negligible effects on indolic glucosinolates. The aromatic glucosinolate concentrations were very low in both wildtype BC and the MAM1 knockouts. Significant glucosinolate reduction was also observed in the seeds of the MAM1 plants with 5 or 4 genes knocked out (Table 4). Similar results were obtained in T3 seeds (Table 5).Table 4| WT | WT | WT | WT | WT | WT | 11,5 | 4,24 | 7,24 | 0,04 |Table 5Example 2
[0125] To determine if knocking out MAM1 affects seed germination, T2 seeds were placed in a row between two layers of filter papers wetted with deionized water. The filter papers were rolled up with a piece of waxed paper on the outside and set vertically in a beaker containing 1 inch (2.5 cm) of deionized water. The beaker was covered with plastic wrap to prevent evaporation and placed at 25°C in the dark. Germination was scored 5 days after seeding. The MAM1 knockouts had the same germination frequency as WT segregants and the WT inbred BC. To test germination under cold stress, the seeds rolled up in the filter papers were incubated at 8 °C in the dark for 7 days, or 4 °C for 5 days followed by 25 °C for 2 days. No significant difference in germination frequency was found between the homozygotes and WT under the cold- stressed conditions.Example 3
[0126] MAM1 knockouts were evaluated in the field for agronomic performance and glucosinolate content. The gene-edited inbred lines were grown in single-row plots in a randomized complete block design with three replications in field environments in Rockwood, ON, Canada. Fertilizer was applied to achieve maximum yields. Weeds and pests were controlled according to local practices. Number of emerged plants was counted 3 weeks after planting, plant vigor was rated on 1-9 scale (l=Low, 9=High) at 20 days after emergence, and flowering time was recorded as number of days after planting to the day when 50% of the plants in a plot have at least one flower in bloom. Seed samples were taken at maturity from plants in the middle of rows. Glucosinolate, protein, oil and ADF contents were analyzed using a nearinfrared (NIR) spectrometer. To reduce pollinators carrying pollen around the field, plants were covered up with white knitted shade cloth (20% density) during flowering time. However, wind and physical contact with plants in neighboring rows still can cause cross pollination; genotyping revealed that approximately 10% of the seeds from homozygous plants were heterozygous.
[0127] Glucosinolate reduction was detected in the MAM1 knockout variants relative to WT inbred BC controls under field conditions (Figure 1). Oil, protein and ADF contents were not significantly different between knockout variants and the WT inbred.
[0128] Figure l is a bar graph showing glucosinolate levels in seed of the following plants grown in a field in 2023: sextuple homozygous knockouts of MAM1.A2, C2, A3, C7a, C7b and A4; quintuple homozygous knockouts of MAM1.A2, C2, C7a, C7b and A4; quintuple homozygous knockouts of MAM1.A2, C2, A3, C7a and C7b; wild-type hybrid control (WT).Example 4: MAM1 Knockout Mutants generated by EMS mutagenesis
[0129] MAM1 knockout mutants (nonsense mutation and splicing site mutation) also were identified from an ethyl methanesulfonate (EMS) mutagenesis population for reduced glucosinolates in canola seed. To develop a canola mutagenesis population, dry seeds of the inbred BC were treated with 0.3%, 0.5% and 0.8% of EMS, rinsed with distilled water, and the treated seeds (Ml) were air-dried. Ml seeds were planted in the field and M2 seeds were harvested from 3,200 individual Ml plants. Three seeds from each M2 were planted out in greenhouse and 7,000 M2 plants were obtained. Leaf samples were taken from individual M2 plants for DNA extraction and M3 seeds were harvested. Of the 7,000 M2 plants, the genome of 550 lines were sequenced. Multiple lines carrying MAM1 knockout mutations were identified by sequence search. The MAM1 knockout mutants obtained are shown in Table 6.Table 6
Claims
CLAIMS1. A method of decreasing glucosinolate in Brassica napus plant, seed, tissue or germplasm thereof, the method comprising introducing a targeted alteration to the sequence of one or more methylthioalkylmalate synthase (MAM1) genes and thereby generating Brassica napus plant, seed, tissue or germplasm thereof comprising one or more MAM1 variants and decreased level of glucosinolate relative to the plant, seed, tissue or germplasm thereof prior to introducing the one ormoreMAMl variants.
2. The method of claim 1, comprising introducing a targeted alteration of one or more alleles of MAM1 on chromosome A2, C2, A3, C7, or A4 and thereby generating the Brassica napus plant, seed, tissue or germplasm thereof comprising one or more MAM1 variants.
3. The method of claim 1 or 2, wherein the targeted alteration is an insertion or deletion that reduces or eliminates expression of the full-length protein encoded by the targeted MAM1 genes or targeted alleles thereof.
4. The method of claim 3, wherein the targeted alteration introduces a premature termination codon in the one or more MAM1 variants.
5. The method of any one of claims 1-4, wherein the method comprises introducing a targeted alteration to one or more of MAM1 genes MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, orMAMl.A4.
6. The method of claim 1, wherein the targeted alteration introduces a premature termination codon in the position shown in Table 3 for MAM1 on chromosome A2, C2, A3, C7, or A4.
7. The method of claim 1, wherein each of the one or more MAM1 variants comprises a premature termination codons shown in SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQIDNO:16, SEQIDNO:17, SEQIDNO:18, SEQIDNO:19, SEQIDNO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQIDNO:27, SEQIDNO:28, SEQIDNO:29, SEQIDNO:30, SEQIDNO:31, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO:37, SEQIDNO:38, SEQIDNO:39, SEQIDNO:40, SEQIDNO:41, SEQIDNO:41, SEQ ID NO:41, SEQIDNO:41, SEQIDNO:41, SEQIDNO:41, SEQIDNO:50, SEQIDNO:51, SEQIDNO:52, SEQIDNO:53, SEQIDNO:54, SEQIDNO:55, SEQIDNO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID NO:61, SEQ ID NO:62,SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NO:70.
8. The method of any one of claims 1-7, wherein the method comprises introducing multiple homozygous variants of MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, or MAM1 . A4 and thereby generating Brassica napus plant, seed, tissue or germplasm thereof comprising multiple homozygous variants of MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, or MAMl.A4.
9. A method of producing decreased glucosinolate Brassica napus meal, the method comprising a) selecting Brassica napus seed or grain comprising one or more MAM1 variants, wherein the selected seed or grain comprise decreased glucosinolate relative to control seed or grain lacking the one or more MAM1 variants; and b) milling the selected seed or grain to produce decreased glucosinolate Brassica napus meal.
10. The method of claim 9, wherein the selected seed or grain comprises a targeted alteration to one or more alleles of MAM1 on chromosome A2, C2, A3, C7, or A4.
11. The method of claim 10, wherein the selected seed or grain comprises an insertion or deletion that reduces or eliminates expression of the full-length protein encoded by the targeted MAM1 genes.
12. The method of claim 11, wherein the insertion or deletion introduces a premature termination codon in the one or more MAM1 variants.
13. The method of claim 10, wherein the targeted alteration introduces a premature termination codon in the position shown in Table 3 for MAM1 on chromosome A2, C2, A3, C7, or A4.
14. The method of any one of claims 9-13, wherein the method comprises introducing a targeted alteration to one or more ofMAMl genes MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, orMAMl.A4.
15. The method of claim 9, wherein the one or more MAM1 variants comprise one or more premature stop codons shown in SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQ ID NO: 16, SEQ ID NO: 17, SEQ ID NO: 18, SEQ ID NO:19, SEQ ID NO:20, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NO:30, SEQ ID NO:31, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ IDNO:38, SEQ ID NO:39, SEQ ID NO:40, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID NO:50, SEQ ID N0:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID N0:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NOTO, including combinations thereof.
16. The method of any one of claims 9-15, wherein the selected seed or grain comprises multiple homozygous variants of MAM1.A2, MAM1.C2, MAM1.A3, MAMl.C7a, MAMl.C7b, or MAM1.A4.
17. The method of any one of claims 9-16, wherein the selected seed or grain further comprises a targeted alteration of one or more of the following genes: LPA1, TT2, TT8, DFR, F3H, ANR, LDOX, MYB28, MRP1 o MRP 2.
18. The Brassica napus plant, seed, tissue or germplasm thereof comprising one or more methylthioalkylmalate synthase (MAM1) gene variants which are single homozygous MAM1 variants on chromosome A2, C2, A3, C7, or A4 or multiple homozygous MAM1 variants on chromosome A2, C2, A3, C7, or A4, wherein the Brassica napus plant, seed, tissue or germplasm thereof comprising the one or more MAM1 variants have an decreased level of glucosinolate relative to a control lacking the one or more MAM1 variants.
19. The Brassica napus plant, seed, tissue or germplasm thereof of claim 18, which comprises a (i) single homozygous MAM1 variants on chromosome A2, C2, A3, C7, or A4, or (ii) multiple homozygous MAM1 variants on chromosome A2, C2, A3, C7, or A4.
20. The Brassica napus plant, seed, tissue or germplasm thereof of claim 18 or 19, wherein each of the MAM1 variants comprises a premature termination codon in the position shown in Table 3 for MAM1 on chromosome A2, C2, A3, C7, or A4.
21. The Brassica napus plant, seed, tissue or germplasm thereof of any one of claims 18-20, wherein each of the one or more MAM1 variants comprises a premature stop codon shown in SEQ ID NO: 13, SEQ ID NO: 14, SEQ ID NO: 15, SEQ ID NO: 16, SEQ ID NO: 17, SEQ ID NO: 18, SEQ ID NO: 19, SEQ ID NOTO, SEQ ID NO:21, SEQ ID NO:22, SEQ ID NO:23, SEQ ID NO:24, SEQ ID NO:25, SEQ ID NO:26, SEQ ID NO:27, SEQ ID NO:28, SEQ ID NO:29, SEQ ID NOTO, SEQ ID NO:31, SEQ ID NO:32, SEQ ID NO:33, SEQ ID NO:34, SEQ ID NO:35, SEQ ID NO:36, SEQ ID NO:37, SEQ ID NO:38, SEQ ID NO:39, SEQ IDNO:40, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID N0:41, SEQ ID NO:50, SEQ ID N0:51, SEQ ID NO:52, SEQ ID NO:53, SEQ ID NO:54, SEQ ID NO:55, SEQ ID NO:56, SEQ ID NO:57, SEQ ID NO:58, SEQ ID NO:59, SEQ ID NO:60, SEQ ID N0:61, SEQ ID NO:62, SEQ ID NO:63, SEQ ID NO:64, SEQ ID NO:65, SEQ ID NO:66, SEQ ID NO:67, SEQ ID NO:68, SEQ ID NO:69, or SEQ ID NOTO.
22. The Brassica napus seed of any one of claims 18-21, further comprising total glucosinolate content of less than 10.0 pmole / g.
23. The Brassica napus plant, seed, tissue or germplasm thereof of any one of claims 18-22, further comprising a targeted alteration of one or more of the following genes: LPA1, TT2, TT8, DFR, F3H, ANR, LDOX, MYB28, MRP1 xMRP2.
24. A method of identifying a plant, seed, tissue or germplasm thereof comprising an MAM1 variant associated with decreased glucosinolate level, the method comprising: a) providing plant, seed, tissue or germplasm thereof comprising one or more MAM1 variants in accordance any one of claims 18-23; b) obtaining a sample comprising nucleic acid from the plant, seed, tissue or germplasm thereof; c) screening the sample for the one or more MAM1 variants; and d) detecting the one or more (i) MAM1 variants or (ii) marker alleles in the sample that are genetically linked to the one or more MAM1 variants and thereby identifying the plant as having an MAM1 variant associated with decreased glucosinolate level.
25. A method of introducing an MAM1 variant into a Brassica napus plant comprising: a) crossing a first parent Brassica napus plant with a second parent Brassica napus plant to produce progeny plants, wherein the first parent plant is a plant comprising one or more MAM1 variants in accordance with any one of claims 18-23; and b) selecting at least one progeny plant comprising the one or more MAM1 variants.
26. The method of claim 25 further comprising: c) crossing the selected progeny plants with the second parent plant to produce backcross progeny plants; and d) selecting for backcross progeny plants comprising the one or more MAM1 variants.
27. A recombinant nucleic acid construct comprising one or more of SEQ ID NOs: 13-41, 50-70.
Citation Information
Patent Citations
Low Glucosinolate Pennycress Meal and Methods of Making
US20190225977A1