Potency assay for multiple coding nucleic acids
The use of MITD sequences enables simultaneous quantification of multiple nucleic acid sequences in T cell vaccines by correlating MITD abundance with expression levels, overcoming the limitations of traditional antibody-based assays, ensuring accurate and efficient expression level measurement.
Patent Information
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Filing Date
- 2025-08-28
- Publication Date
- 2026-03-05
AI Technical Summary
Existing assays are inadequate for quantifying the expression levels of multiple nucleic acid sequences, particularly in multivalent T cell vaccines, as they often require antibody-based approaches that are not applicable and cannot differentiate between various encoded antigens/epitopes, especially when sequences are personalized or de novo.
Utilizing different Major Histocompatibility Class I Trafficking Domain (MITD) sequences encoded by nucleic acids to differentiate between and quantify the expression levels of multiple nucleic acid sequences through methods like mass spectrometry, without requiring additional amino acid sequences, allowing for simultaneous analysis of multiple nucleic acid sequences.
Provides a rapid, cost-effective, and reliable method to measure and validate the expression levels of multiple nucleic acid sequences, including undetermined epitopes, without the need for antibody-based techniques, suitable for multivalent T cell vaccines and other nucleic acid-based therapeutics.
Smart Images

Figure IMGF000046_0001 
Figure IMGF000058_0001 
Figure IMGF000058_0002
Abstract
Description
[0001] POTENCY ASSAY FOR MULTIPLE CODING NUCLEIC ACIDS
[0002] Technical Field of the Invention
[0003] The invention provides methods for simultaneously analyzing at least two different nucleic acid sequences (such as RNA and / or DNA sequences), by using at least two different encoded MITD domains to differentiate between the different nucleic acid sequences. The methods of the present invention may be performed with nucleic acid sequences encoding amino acid sequences comprising antigens or epitopes. Nucleic acid sequences analysed by the methods of the present invention may therefore be used in downstream clinical applications, e.g., for eliciting an immune response against two or more antigens or epitopes encoded by the nucleic acid sequences, in a subject in which the immune response may be therapeutic or partially or fully protective. Thus, the nucleic acid sequences may be useful for vaccination. More particularly, the nucleic acid sequences may be useful as a multivalent T-cell-targeting vaccine.
[0004] Background to the Invention
[0005] Apart from their well-known ability to encode biologically active proteins, nucleic acids such as DNA and RNA have other remarkable properties that make them attractive therapeutic agents. Nucleic acid-based therapeutics are easy to manufacture and relatively inexpensive. The use of RNA to deliver foreign genetic information into target cells offers an attractive modality. The advantages of RNA include transient expression and non-transforming character. RNA does not require nucleus infiltration for expression and moreover cannot integrate into the host genome, thereby eliminating the risk of oncogenesis.
[0006] T cell vaccines are compositions comprising nucleic acid constructs (typically mRNA) designed to encode highly immunogenic regions or epitopes of target antigens concatenated together in a single polypeptide. These T cell vaccines are designed to raise an immunogenic response through recognition of a major histocompatibility complex (MHC)-presented epitope by a T cell. In other words, functional epitope sequences encoded by T cell vaccines are expressed and presented by MHCs, thereby raising an immunogenic response against the epitope(s) that are presented. As such, the epitopes encoded by T cell vaccines function in an entirely different way to more traditional means of eliciting an immune response, i.e. via the expression of a stable, folded antigen to be recognised by an antibody. T cell vaccines may utilize a MHC I trafficking domain (MITD) sequence fused to the functional epitope sequences, which improves epitope presentation to T cells and dendritic cells, and enhances the cellular immune response. Further information on MITD domains can be found in Kreiter et al. (J Immunol 180(1 ) (2008) 309-318). Furthermore, multivalent T cell vaccines may comprise multiple nucleic acids, such as multiple mRNA constructs, e.g. in the same composition, thereby encoding multiple different polypeptide chains encoding epitopes / antigens.
[0007] Quantitative tests are used to measure product attributes associated with product quality and manufacturing controls, and are performed to assure identity, purity, strength (potency), and stability of products used during all phases of clinical study. Similarly, such measurements are used to demonstrate that only product lots that meet defined specifications or acceptance criteria are administered during all phases of clinical investigation and following market approval.
[0008] Such assays involve the qualitative and / or quantitative measure of certain criteria that should describe the ability of a product to achieve a defined biological effect. The criteria measured should be closely related to the product's intended biological effect and ideally, it should be related to the product's clinical purpose. Measurement of the potency of a product is not the same as measuring clinical efficacy. Rather, it is a means to control product quality and provide appropriate release criteria, in particular under GMP. Normally, for each and every product which is to be administered to a subject, a separate assay must be developed. In the rapidly evolving field of nucleic acid-based therapeutics, with hundreds of different potential constructs, and where suitable antibodies are typically not available for detection, an assay which can be easily adapted to new products such as multivalent T cell vaccines would be of benefit. As part of the drug development process for T cell vaccines, an assay must therefore be developed to measure the expression levels of components of T cell vaccines. However, as mentioned above, T cell vaccines are not necessarily designed to raise immunogenic responses through expression of a stable, folded antigen to be recognized by an antibody. As such, until now, the development of a qualitative and / or quantitative assay for such vaccines could not have readily relied on typical antibodybased approaches such as flow cytometry, western blots, or ELISAs.
[0009] An additional consideration for the development of such an assay is that multivalent T cell vaccines may encode a number of different antigens / epitopes on at least two different nucleic acids, which may be present in the same composition / mixture. Thus, a qualitative and / or quantitative assay for a multivalent T cell vaccine needs to be able to determine the expression level of several different antigens / epitopes from different nucleic acids simultaneously. Further still, T cell vaccines may encode de novo epitope sequences. This is because, in some instances, T cell vaccines are personalized vaccines with sequences that vary from patient-to-patient. Accordingly, a T cell vaccine quantification assay must be developed which can take account of the need to assess the expression levels of de novo or undetermined epitope sequences. As an additional consideration, it would also be beneficial if the assay did not require the nucleic acid sequences to encode any extensive additional amino acid sequences which would not otherwise be present in the end product, e.g. in a multivalent T cell vaccine.
[0010] In short, there is a problem of providing an assay that can be used to determine the expression levels of nucleic acids encoding antigens / epitopes that raise immunogenic responses through MHC presentation, for which typical antibody-based approaches are not useful. There is also a problem of providing an assay that can determine the expression levels of multiple different expressed functional sequences (e.g. epitopes) from at least two different nucleic acids simultaneously, which may include novel expressed functional sequences. Although such an assay would be useful for analyzing the expression levels of nucleic acids in T cell vaccines, it would also be useful in any scenario where any of these limitations apply. This would include any scenario of simultaneously analysing the expression levels of at least two different nucleic acid sequences.
[0011] Summary of the Invention
[0012] The present invention provides methods of simultaneously analysing at least two different nucleic acid sequences, using different MITD sequences comprised in the amino acid sequences encoded by the at least two different nucleic acid sequences. In particular, when the at least two nucleic acid sequences each encode an amino acid sequence comprising a different MITD sequence that are expressed, the differences between the expressed MITD sequences can be used to differentiate between the encoding nucleic acid sequences. In one methodology, the abundance of each expressed MITD sequence can be measured (e.g. by mass spectrometry) and used as an indication for the expression level of the nucleic acid sequence that encodes the amino acid sequence comprising said MITD sequence. Also, because the abundance of each expressed MITD sequence can be measured and used as an indicator / differentiator of the expression level of the encoding nucleic acid, the expression level of other functional sequences encoded by the nucleic acids (e.g. epitope / antigen sequences) do not need to be directly measured. In other words, each MITD sequence acts as a molecular marker for the unique identification of the expression of the amino acid sequence comprising the MITD sequence, and the nucleic acid sequence encoding the amino acid sequence comprising the MITD sequence.
[0013] For this reason, the methods of the present invention work well for analysing multivalent nucleic acid compositions, and for analysing the expression levels of undetermined and / or de novo functional sequences, such as variable epitopes. Expressed MITD sequences can be quantified by means such as mass spectrometry, obviating the need to use antibody-based techniques for quantification. Based on these observations, a rapid, cost-effective, reliable, and easy to use and interpret assay to measure, determine, identify, quantify, confirm and / or validate the expression level and therefore the therapeutic potential of at least two different nucleic acids (such as RNA and / or DNA) each encoding at least two different functional sequences is provided.
[0014] A further benefit of the present invention is that the MITD sequences that are used for quantification may already advantageously be encoded by a multivalent nucleic acid composition, e.g. a multivalent T cell vaccine composition. As such, the present invention does not require any significant additional amino acid sequences to be present purely for the purpose of the assay. To yet further advantage, the present invention can exploit the wide variety of natural MITD polymorphisms to provide many different MITD sequences for analytical purposes, which will still be likely to have the same beneficial functions maintained, e.g. cellular trafficking, immune response and vaccination functions.
[0015] The present invention includes various aspects. It is understood that the various embodiments that are described herein as being applicable to any one aspect of the present invention will also generally be applicable to all the other aspects of the present invention.
[0016] In a first aspect, the present invention provides a method for simultaneously analysing at least two different nucleic acid sequences, wherein each different nucleic acid sequence encodes a different amino acid sequence, wherein each amino acid sequence comprises a different Major Histocompatibility (MHC) class I trafficking domain (MITD) sequence, wherein the method comprises the following steps: (i) providing the at least two different nucleic acid sequences;
[0017] (ii) expressing the at least two different amino acid sequences comprising different MITD sequences; and
[0018] (iii) using the at least two different expressed MITD sequences to differentiate between the at least two different nucleic acid sequences.
[0019] In a further aspect, the present invention provides a composition comprising at least two different nucleic acid sequences, wherein each of the nucleic acid sequences encodes a different amino acid sequence comprising: a) at least one antigen or epitope sequence; and b) an MITD sequence; wherein the MITD sequences in each of the different amino acid sequences are different.
[0020] In a further aspect, the present invention provides a composition comprising at least two different amino acid sequences, wherein each of the amino acid sequences comprises: a) at least one antigen or epitope sequence; and b) an MITD sequence; wherein the MITD sequences in each of the different amino acid sequences are different.
[0021] In a further aspect, the present invention provides a kit comprising at least two different nucleic acid sequences, wherein each nucleic acid sequence encodes an amino acid sequence comprising at least one antigen or epitope sequence and an MITD sequence, wherein the MITD sequences in each of the different amino acid sequences are different.
[0022] In a further aspect, the present invention provides a use of the composition or the kit according to the invention, for simultaneously analysing the expression level of at least two different nucleic acid sequences.
[0023] Brief Description of the Figures
[0024] Figure 1 illustrates two different nucleic acids (constructs) suitable for use in the present invention. In particular, the first nucleic acid “construct A” comprises a first MITD sequence “MITD-1” and the second nucleic acid “construct B” comprises a second, different MITD sequence “MITD-2”. Other optional / preferred sequences such as various epitope sequences may be present. Figure 2 lists twenty-two unique variations in MITD sequences suitable for use in the present invention, including a first identified naturally occurring sequence (first row) and twenty-one different variant sequences identified as naturally occurring polymorphisms which involve a single amino acid substitution relative to the naturally occurring sequence. The listed sequences form a full MITD sequence when they are present downstream (C-terminal) of an amino acid sequence comprising a transmembrane domain, preferably a MITD transmembrane domain, most preferably IVGIVAGLAVLAVWIGAWATVMCRRKSSGGK (SEQ ID NO: 1 ). The underlined part of SEQ ID NO: 1 is the transmembrane domain sequence per se. The variant sequences identified in Figure 2 have a high probability of unaltered MITD function due to their being identified in naturally occurring polymorphisms and only involving a single amino acid change relative to the first identified sequence.
[0025] Figure 3 lists further unique variations in MITD sequences suitable for use in the present invention and tested in Example 2 herein.
[0026] Figure 4 illustrates an overview of the experimental scheme of Example 2.
[0027] Figure 5 shows GFP expression kinetics of GFP containing various MITD point mutants, measured via HT microscopy (Incucyte). The WT (wild-type) construct is shown in black, with the variant constructs shown in grey. UTF is a negative control.
[0028] Figure 6 shows residuals of expression kinetics between WT and point mutant MITD sequences. GFP expression kinetics were measured via HT microscopy (Incucyte). The difference in fluorescent signal between the WT and indicated variant at each timepoint is show in the plot.
[0029] Figure 7 shows the mass spectrometry (MS)-measured expression for SP-GFP-MITD variants (normalized to WT). The relative abundance of GFP protein was quantified by measuring GFP tryptic peptides via targeted mass spectrometry 24 hours after transfection into Expi293 cells.
[0030] Figure 8 shows the MS-measured presentation of TPIGDGPVL on B*07:02 from SP- GFP-MITD variants. The relative abundance of the GFP-derived B*07:02 epitope TPIGDGPVL was quantified via targeted mass spectrometry 24 hours after transfection into Expi293 cells. HLA-I complexes were enriched by immunoprecipitation, and peptides were eluted and cleaned for MS analysis.
[0031] Figure 9 illustrates an overview of an experimental scheme for reciprocal testing of the in vivo immunogenicity of MITD sequence variant-decatope constructs. Figure 10 shows the initial characterization of MITD sequence variants. A: Charge state distribution, B: Summed fragment ion intensity as % of WT MITD sequence, C: Measured retention times.
[0032] Figure 11 shows MITD sequence spike-in experiments. A: Summed fragment ion intensity in the 40 fmol spike-in samples, as % of WT MITD sequence (with and without complex sample matrix background), B: Recovery (%) of the intensity values in 40 fmol spike-in samples, relative to the intensity measured in 4000 fmol spike-in samples. ▲ indicates MITD sequences which stood out due to higher signal intensities and accuracy / recovery.
[0033] Figure 12 shows multiplexed measurements of MITD sequence variants. A: MITD sequence set elution profiles in 10 nM spike-in, B: MITD sequence set intensities in spike-in samples.
[0034] Figure 13 illustrates an overview of an experimental scheme for in vitro immunogenicity testing of multiple selected MITD sequence variants.
[0035] Detailed Description of the Invention
[0036] Although the present disclosure is further described in more detail below, it is to be understood that this disclosure is not limited to the particular methodologies, protocols and reagents described herein as these may vary. It is also to be understood that the terminology used herein is for the purpose of describing particular embodiments only, and is not intended to limit the scope of the present disclosure which will be limited only by the appended claims. Unless defined otherwise, all technical and scientific terms used herein have the same meanings as commonly understood by one of ordinary skill in the art.
[0037] In the following, the elements of the present disclosure will be described in more detail. These elements are listed with specific embodiments, however, it should be understood that they may be combined in any manner and in any number to create additional embodiments. The variously described examples and preferred embodiments should not be construed to limit the present disclosure to only the explicitly described embodiments. This description should be understood to support and encompass embodiments which combine the explicitly described embodiments with any number of the disclosed and / or preferred elements. Furthermore, any permutations and combinations of all described elements in this application should be considered disclosed by the description of the present application unless the context indicates otherwise.
[0038] Methods
[0039] In an aspect, the present invention provides a method for simultaneously analysing at least two different nucleic acid sequences, wherein each different nucleic acid sequence encodes a different amino acid sequence, wherein each amino acid sequence comprises a different Major Histocompatibility (MHC) class I trafficking domain (MITD) sequence, wherein the method comprises the following steps:
[0040] (i) providing the at least two different nucleic acid sequences;
[0041] (ii) expressing the at least two different amino acid sequences comprising different MITD sequences; and
[0042] (iii) using the at least two different expressed MITD sequences to differentiate between the at least two different nucleic acid sequences.
[0043] In some embodiments, step (iii) comprises the following:
[0044] (a) determining the abundance of each of the at least two different expressed MITD sequences; and
[0045] (b) correlating the abundance of each expressed MITD sequence with the expression level of the nucleic acid sequence encoding said MITD sequence.
[0046] In some embodiments, step (a) comprises using mass spectrometry (MS), liquid chromatography MS (LC-MS), targeted LC-MS, a detection method using anti-MlTD antibodies (e.g. FACS or flow cytometry), ELISA or a detection method using anti- MlTD aptamers to determine the abundance of each of the at least two different expressed MITD sequences. It will be generally be understood by the skilled person that, in some embodiments where a detection method using anti-MlTD antibodies is employed in the present invention, a different antibody / aptamer as appropriate is used for each different MITD sequence, wherein each antibody / aptamer as appropriate specifically binds to a single MITD sequence used in the present invention, thus enabling the different MITD sequences to be differentiated. In embodiments, antibodybased detection can be used on the expressed MITD sequences per se, allowing the quantification of polypeptides present on the same amino acid sequence comprising the MITD sequences, including those which could not otherwise be quantified with antibody-based detection (e.g. variable or unknown epitopes). In some embodiments, step (b) comprises correlating the abundance of each expressed MITD sequence (peptide) with the expression level of the nucleic acid encoding the MITD sequence. In some embodiments, it will be generally understood that the expression level of the nucleic acid encoding the MITD sequence is the expression level of the protein (e.g. full-length or polyprotein) that is encoded by the nucleic acid. In embodiments wherein the nucleic ecid encodes an MITD sequence and antigen or eptitope sequences, the expression (level) of the MITD sequence can therefore be correlated with the expression (level) of the polyprotein encoding the antigen or epitope sequences (as well as the MITD sequence). In some embodiments, it will generally be understood that the correlation is between polypeptide (MITD seuqence) and polypeptide (polyprotein / antigen / epitope encoded by the same nucleic acid sequence as applicable) expression levels.
[0047] Method steps
[0048] In embodiments, it is generally understood that the steps of the method of the present invention are sequential and in the stated order, e.g. (i) then (ii) then (iii) etc.
[0049] In an embodiment, step (i) of the method of the present invention involves providing the at least two different nucleic acid sequences as a single composition or a single mixture or a single mixed composition, i.e. comprising the at least two different nucleic acids.
[0050] In an embodiment, step (ii) of the method of the present invention involves introducing, such as transfecting or transducing, the at least two different nucleic acids into a cell in vitro. In an embodiment, step (ii) of the method of the present invention involves introducing the at least two different nucleic acid sequences into a cell in vivo, such as by administration to a subject.
[0051] Step (ii) of the method of the present invention is understood to comprise expressing the at least two different amino acid sequences, e.g. in a cell. In an embodiment, step (ii) is instead defined as attempting to express the amino acid sequences.
[0052] In some such embodiments, the method of the present invention further comprises lysing the cells after step (ii) and / or prior to step (iii). In an embodiment, the method further comprises processing the cell lysate. In an embodiment, processing the cell lysate comprises one or more steps selected from the group consisting of proteolytic enzyme digestion, denaturation, reduction, alkylation, drying, reconstitution, and desalting. In an embodiment, processing the cell lysate comprises the proteolytic enzyme treatment of the present invention for excising the MITD sequences. In an embodiment of step (iii) of the method of the present invention, the MITD sequences are used to differentiate between (the expression levels of) the at least two different nucleic acid sequences. In an embodiment, step (iii) comprises differentiating between (the expression levels of) the at least two different nucleic acid sequences using the different MITD sequences.
[0053] In an embodiment, step (iii) comprises (a) determining the abundance of each of the at least two different expressed MITD sequences, and (b) correlating the abundance of each expressed MITD sequence with the expression level of the nucleic acid sequence encoding said MITD sequence. In an embodiment, determining the abundance means measuring the abundance, e.g. by mass spectrometry. In an embodiment, step (iii) is solely defined by these sub-steps (a) and (b). In such embodiments, it will be understood that the method does not need to recite the differentiation step mentioned above and may solely recite (a) and (b).
[0054] In embodiments, the abundance of each of the different MITD sequences is generally understood as the abundance of the MITD sequences that have been expressed in step (ii). In an embodiment of step (iii)(a) of the method of the present invention, the abundance of each of the different MITD sequences is the abundance determined using mass spectrometry (MS). In an embodiment, the abundance of each of the different MITD sequences is determined using liquid chromatography-mass spectrometry (LC-MS). In an embodiment, the abundance of each of the different MITD sequences is determined using targeted LC-MS. In some preferred embodiments using LC-MS the different MITD sequences comprise further sequences selected from SEQ ID NO: 7, SEQ ID NO: 8, SEQ ID NO: 13, SEQ ID NO: 15 and SEQ ID NO: 23, such as one, two, three, four or five sequences selected from SEQ ID NO: 7, SEQ ID NO: 8, SEQ ID NO: 13, SEQ ID NO: 15 and SEQ ID NO: 23. In an embodiment, the abundance of each of the different MITD sequences is determined using a detection method using anti-MlTD antibodies (e.g. FACS or flow cytometry). In an embodiment, the abundance of each of the different MITD sequences is determined using ELISA. In an embodiment, the abundance of each of the different MITD sequences is determined using a detection method using anti-MlTD aptamers. In an embodiment, e.g. when using a form of MS, the MITD sequences are cleaved or excised (e.g. proteolytically), or otherwise isolated from the amino acid sequences encoded by the at least two nucleic acid sequences. In an embodiment, this is prior to determining the expressed abundance of each of the MITD sequences. In some such embodiments, the method of the invention comprises isolating the expressed MITD sequences and determining the expressed amount of the MITD sequences. In an embodiment, the MITD sequences themselves are not (substantially) proteolytically cleaved. In an embodiment, the MITD sequences may undergo some proteolytic cleavage but only that which results in a different (e.g. analytically distinct) cleaved sequence being produced from each different MITD sequence.
[0055] In an aspect, the present invention relates to a method for analysing and / or quantifying the ability of at least two different nucleic acid sequences to express their encoded amino acid sequences in a biological system. In an embodiment, the biological system in the present invention is a biological system present in a cell. In an embodiment, the biological system in the present invention is a biological system present in a human cell. In an embodiment, the biological system in the present invention is a biological system present in a tissue, e.g. a human tissue, e.g. ex vivo. In an embodiment, the biological system in the present invention is a biological system present in an organism, e.g. a human, e.g. a human patient.
[0056] In embodiments, it will be generally understood that step (iii)(b) of the method of the present invention involves a positive correlation. In an embodiment, step (iii)(b) comprises determining the expression levels of the at least two nucleic acids using a positive correlation between the abundance of each expressed MITD sequence and the expression of the corresponding encoding nucleic acid. In an embodiment, step (iii)(b) comprises (positively) correlating the abundance of each MITD sequence with the expression level of the amino acid sequence comprising said MITD sequence. In an embodiment, step (iii)(b) comprises (positively) correlating the abundance of each MITD sequence with the expression strength of the nucleic acid sequence encoding said MITD sequence. In an embodiment, the abundance of each expressed MITD sequence is relative to abundances of the other MITD sequences. In an embodiment, the expression level of each nucleic acid is relative to the other nucleic acids. In other embodiments, the expression level of each nucleic acid is the absolute expression level. The absolute expression level of each nucleic acid may be determined, e.g. by including a control nucleic acid in the assay, wherein the control nucleic acid has a pre- determined / known / normalised / independently quantifiable expression level.
[0057] In an embodiment, the method is for determining the expression levels of at least two different nucleic acid sequences, e.g. in a biological system as described herein, such as in a human cell, human tissue or the human body. In an embodiment, the method is for determining the ability of at least two nucleic acid sequences to express their respective encoded amino acid sequences. In an embodiment, the method is for discriminating between the expression levels of at least two nucleic acid sequence. In an embodiment, the expression levels are the transcriptional and / or translational efficiencies of the nucleic acid sequences. In an embodiment, the method is quantitative.
[0058] Alternative measurement of mRNA abundance
[0059] In an alternative aspect / embodiments, the method of the present invention may be utilised as described herein, but for the measurement of nucleic acid abundance, in particular for measuring the abundance of at least two (different) transcribed nucleic acids that are transcribed from the at least two different nucleic acid sequences of the present invention.
[0060] In some such embodiments, the abundance of the transcribed nucleic acid sequences may be measured, e.g. by polymerase chain reaction (PCR), digital PCR (dPCR) or droplet dPCR (ddPCR). Such embodiments may be a method for measuring transcribed nucleic acid abundance. Such embodiments may be a method for measuring transcribed nucleic acid stability. In embodiments, the transcribed nucleic acids are typically RNA, preferably mRNA. In such embodiments, the skilled person will understand that each transcribed nucleic acid comprises a different transcribed MITD sequence.
[0061] In some such embodiments, the present invention provides a method for simultaneously analysing at least two different nucleic acid sequences, wherein each different nucleic acid sequence encodes a different amino acid sequence, wherein each amino acid sequence comprises a different Major Histocompatibility (MHC) class I trafficking domain (MITD) sequence, wherein the method comprises the following steps:
[0062] (i) providing the at least two different nucleic acid sequences;
[0063] (ii) transcribing the at least two different amino acid sequences comprising different MITD sequences; and
[0064] (iii) using the at least two different transcribed MITD sequences to differentiate between the at least two different nucleic acid sequences.
[0065] In some such embodiments, the method comprises (iii) using the at least two different transcribed MITD sequences to differentiate between the at least two different transcribed nucleic acid sequences. In some such embodiments, step (iii) comprises the following:
[0066] (a) determining the abundance of each of the at least two different transcribed MITD sequences; and (b1 ) or (b2):
[0067] (b1 ) correlating the abundance of each transcribed MITD sequence with the abundance and / or RNA stability of the transcribed nucleic acid comprising said transcribed MITD sequence; or
[0068] (b1 ) correlating the abundance of each transcribed MITD sequence with the abundance and / or RNA stability of the nucleic acid sequence encoding said MITD sequence.
[0069] At least two different nucleic acid sequences
[0070] The method of the present invention is carried out on at least two different nucleic acid sequences. In some embodiments, it will be understood by the skilled person that the at least two different nucleic acid sequences generally correspond to two or more different nucleic acid molecules, e.g. each nucleic acid sequences is present on a different nucleic acid molecule. In an embodiment, the at least two different nucleic acid sequences each reside on a different nucleic acid.
[0071] It will be understood that there is no strict practical limitation on the number of nucleic acids / nucleic acid sequences that can be simultaneously analysed in the present invention, as long as the different MITD sequence-based principles of the present invention are adhered to e.g. each expressed amino acid sequence comprises a unique MITD sequence within a given assay.
[0072] In an embodiment, the method of the present invention can be carried out on 2, 3, 4, 5, 6, 7, 8, 9, 10, 11 , 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, 25,26, 27, 28, 29, 30, 31 , 32, 33, 34, 35, 36, 37, 38, 39, 40, 41 , 42, 43, 44, 45, 46, 47, 48, 49, 50, 51 , 52,
[0073] 53, 54, 55, 56, 57, 58, 59, 60, 61 , 62, 63, 64, 65, 66, 67, 68, 69, 70, 71 , 72, 73, 74, 75,
[0074] 76, 77, 78, 79, 80, 81 , 82, 83, 84, 85, 86, 87, 88, 89, 90, 91 , 92, 93, 94, 95, 96, 97, 98,
[0075] 99, 100, 1000, 2000, 3000, 4000, 5000, 6000, 7000, 8000, 9000 or 10000 different nucleic acid sequences.
[0076] In some embodiments, the method is for simultaneously analysing two, three, four, five, six, seven, eight, nine, ten, eleven, twelve, thirteen, fourteen, fifteen, sixteen, seventeen, eighteen, nineteen, twenty, twenty-one, twenty-two, twenty-three, twenty- four, twenty-five, twenty-six, twenty-seven or twenty-eight different nucleic acid sequences. As generally used herein with reference to certain elements of the present invention, e.g. nucleic acids, amino acid sequences, antigen / epitope and linker sequences of the present invention, the term “different” means that the element in question does not consist of the same sequence as any other of the same element in that aspect of the present invention. For example, “at least two different amino acid sequences” means that the at least two amino acid sequences in question consist of different sequences to each other, and “at least two different MITD sequences” means that the at least two MITD sequences consist of different sequences to each other. In an embodiment, no two encoded MITD sequences share the same sequence.
[0077] In an embodiment, the nucleic acid sequences are RNA. In an embodiment, the nucleic acids comprising the nucleic acid sequences of the present invention are RNA. In an embodiment, the nucleic acid sequences are DNA. In an embodiment, the nucleic acids comprising the nucleic acid sequences of the present invention are DNA. In an embodiment, the nucleic acid sequences comprise at least one RNA sequence and at least one DNA sequence. In an embodiment, the nucleic acids comprising the nucleic acid sequences of the present invention comprise at least one RNA polynucleotide and at least one DNA polynucleotide.
[0078] In an embodiment, the at least two different nucleic acid sequences analysed in the method of the present invention are RNA sequences, DNA sequences, or comprise at least one RNA sequence and at least one DNA sequence. In some embodiments, the nucleic acid is DNA (e.g., one or more DNAs), RNA (e.g., one or more RNAs), or a mixture of DNA and RNA (e.g., one or more DNAs and one or more RNAs). In some embodiments, the DNA is present in the form of a vector, e.g., a vector comprising DNA encoding an amino acid sequence comprising the amino acid sequence of a peptide or polypeptide having biological activity. In some embodiments, the vector is a DNA vector. In some preferred embodiments, the at least two nucleic acid sequences are mRNA sequences. In some preferred embodiments, the nucleic acids comprising the at least two nucleic acid sequences are mRNAs, e.g. one mRNA per nucleic acid sequence.
[0079] References herein defining “a” nucleic acid sequence are also understood to apply to the at least two nucleic acid sequences in the present invention.
[0080] MITD sequences
[0081] The different MITD sequences for use in the present invention may be any different MITD sequences that provide the function (immune response) of an MITD domain, and which are also analytically distinct, i.e. their amino acid sequences are different. In some embodiments, each of the different MITD sequences is distinguishable by mass spectrometry. In some embodiments, each of the different MITD sequences has a different molecular weight. In some embodiments, each of the different MITD sequences is distinguishable by charge and / or molecular size / weight. In some embodiments, each of the different MITD sequences has a different chromatographic retention time.
[0082] In some embodiments, the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by any (e.g. homologous) amino acid substitution, e.g. one, two or three amino acid substitutions, relative to a naturally occurring MITD sequence.
[0083] In some embodiments, the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by a C-terminal extension comprising neutral amino acids (e.g. alanine) relative to a naturally occuring MITD sequence.
[0084] In some embodiments, the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by i) any (e.g. homologous) amino acid substitution, e.g. one, two or three amino acid substitutions; and / or ii) a C-terminal extension comprising neutral amino acids (e.g. alanine) relative to a naturally occuring MITD sequence.
[0085] In some embodiments, the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by one or more homologous exchange(s) of amino acids relative to a naturally occurring MITD sequence.
[0086] In some embodiments, the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by: i) any (e.g. homologous) amino acid substitution, e.g. one, two or three amino acid substitutions; and / or ii) one or more homologous exchange(s) of amino acids relative to a naturally occurring MITD sequence.
[0087] In some embodiments, the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by: i) a C-terminal extension comprising neutral amino acids (e.g. alanine); and / or ii) one or more homologous exchange(s) of amino acids relative to a naturally occurring MITD sequence.
[0088] In some embodiments, the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by: i) any (e.g. homologous) amino acid substitution, e.g. one, two or three amino acid substitutions; ii) a C-terminal extension comprising neutral amino acids (e.g. alanine); and / or iii) one or more homologous exchange(s) of amino acids relative to a naturally occurring MITD sequence.
[0089] In some embodiments, the MITD sequences differing by: i) any (e.g. homologous) amino acid substitution, e.g. one, two or three amino acid substitutions; ii) a C-terminal extension comprising neutral amino acids (e.g. alanine); and / or iii) one or more homologous exchange(s) of amino acids; relative to a naturally occurring MITD sequence has improved immunogenicity, e.g. relative to the naturally occuring MITD sequence.
[0090] In a preferred embodiment, the MITD sequences are of a HLA-B gene. In some embodiments, the MITD sequences each elicit the same, substantially the same or equivalent immune response, protective immune response, partially protective immune response and / or clinical immune response as a naturally occurring MITD sequence of a HLA gene, e.g. as a naturally occurring MITD sequence of a HLA gene in a human cell or patient.
[0091] In some embodiments, the MITD sequences each have the same, substantially the same or equivalent vaccine efficacy as a naturally occurring MITD sequence of a HLA gene. In some embodiments, the MITD sequences each have the same, substantially the same or equivalent epitope presentation function as a naturally occurring MITD sequence of a HLA gene. In some embodiments, the MITD sequences each have the same, substantially the same or equivalent cell trafficking function as a naturally occurring MITD sequence of a HLA gene. In some embodiments, the MITD sequences each have the same, substantially the same or equivalent translation efficacy as a naturally occurring MITD sequence of a HLA gene. In some embodiments, the MITD sequences each have the same, substantially the same or equivalent kinetics as a naturally occurring MITD sequence of a HLA gene. In some embodiments, the MITD sequences each have the same, substantially the same or equivalent half-life as a naturally occurring MITD sequence of a HLA gene. In embodiments, this can be the same as that of a naturally occurring MITD sequence of a HLA gene in a cell or organism, e.g. in a human cell or patient.
[0092] In some embodiments, the MITD sequences each elicit the same immune response, e.g. in a human, when expressed as a fusion protein with an epitope and / or antigen sequence. This may be determined as described in the examples section, e.g. by expressing the different MITD sequences as a fusion protein with the same epitope and / or antigen sequence, and determining that the same magnitude of beneficial immune response is elicited against said epitope by each fusion protein. In further detail, this is an immune response which is improved relative to the epitope / antigen when it is not fused to a functional MITD sequence. A MITD sequence with known functional efficacy can be used as a benchmark in this respect.
[0093] In some embodiments, one or more MITD sequence variants of the present invention (SEQ ID NOs: 3 to 29) has improved (greater) immunogenicity, e.g. when compared to a wild-type MITD sequence (e.g. SEQ ID NO: 2). Improved immunogenicity may comprises a greater, stronger, more specific or more neutralising immune response.
[0094] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation G1A relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 3, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0095] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation S3G relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 4, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0096] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation Y4N relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 5, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0097] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation Q6E relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 6, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0098] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation S9F relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 7, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0099] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation S19F relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 8, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0100] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A22D relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 9, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2. In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation L20V relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 10, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0101] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation D17E relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 11 , and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0102] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation D17N relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 12, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0103] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation D17V relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 13, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0104] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation G15S relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 14, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0105] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation G15V relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 15, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0106] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A13D relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 16, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0107] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A13S relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 17, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0108] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A13T relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 18, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0109] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A13V relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 19, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0110] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation D11 N relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 20, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0111] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation D11 Y relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 21 , and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0112] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A8V relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 22, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0113] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation S5Y relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 23, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0114] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation G2A relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 24, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0115] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A7G relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 25, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0116] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A8G relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 26, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2. In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A13G relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 27, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0117] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation G15A relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 28, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0118] In some embodiments, the present invention uses a variant MITD sequence comprising a sequence having the mutation A22G relative to SEQ ID NO: 2, optionally having 80%, 85%, 90%, 95%, 99% or 100% identity to SEQ ID NO: 29, and the MITD sequence has improved immunogenicity, e.g., relative to SEQ ID NO: 2.
[0119] In some embodiments, the MITD sequences each comprise an amino acid sequence comprising a transmembrane (TM) domain and a further sequence directly downstream of the last K / R residue in the amino acid sequence comprising a TM domain. In embodiments, the “further sequences” of the MITD domains are different / unique. In other words, the MITD sequences in the method of the present invention may share a common sequence comprising a TM domain, and still be different because the further sequences in each MITD domain are different. Moreover, the further sequences described herein may be employed regardless of whether the MITD sequences share a common sequence comprising a TM domain. In an embodiment, the further sequence comprises 22 amino acid residues.
[0120] In some embodiments, each MITD sequence comprises the same amino acid sequence comprising a TM domain. In an embodiment the amino acid sequence comprising a TM domain is IVGIVAGLAVLAWVIGAVVATVMCRRKSSGGK (SEQ ID NO: 1 ), or comprises or consists of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto.
[0121] In some embodiments, each of the further sequences differ by one, two or three amino acid substitutions. In some embodiments, the further sequences are not limited, insofar as they provide a functional MITD sequence, e.g. which has the same or equivalent function as a wild-type MITD sequence. Thus, in an embodiment, the further sequences are (any) different sequences that provide a functional MITD sequence. In some embodiments, the further sequences in each of the MITD sequences are different and selected from the following:
[0122] GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation G1A, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation G2A, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation S3G, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation Y4N, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation S5Y, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation Q6E, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A7G, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A8G, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A8V, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation S9F, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation D11 N, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation D11Y, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A13D, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A13G, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A13S, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A13T, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A13V, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation G15A, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation G15S, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation G15V, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation D17E, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation D17N, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation D17V, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation S19F, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation L20V, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A22G, a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity to GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2) and having the substitution mutation A22D,
[0123] AGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 3) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0124] GGGYSQAASSDSAQGSDVSLTA (SEQ ID NO: 4) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0125] GGSNSQAASSDSAQGSDVSLTA (SEQ ID NO: 5) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0126] GGSYSEAASSDSAQGSDVSLTA (SEQ ID NO: 6) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0127] GGSYSQAAFSDSAQGSDVSLTA (SEQ ID NO: 7) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0128] GGSYSQAASSDSAQGSDVFLTA (SEQ ID NO: 8) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0129] GGSYSQAASSDSAQGSDVSLTD (SEQ ID NO: 9) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0130] GGSYSQAASSDSAQGSDVSVTA (SEQ ID NO: 10) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0131] GGSYSQAASSDSAQGSEVSLTA (SEQ ID NO: 11 ) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0132] GGSYSQAASSDSAQGSNVSLTA (SEQ ID NO: 12) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0133] GGSYSQAASSDSAQGSVVSLTA (SEQ ID NO: 13) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0134] GGSYSQAASSDSAQSSDVSLTA (SEQ ID NO: 14) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0135] GGSYSQAASSDSAQVSDVSLTA (SEQ ID NO: 15) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0136] GGSYSQAASSDSDQGSDVSLTA (SEQ ID NO: 16) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto, GGSYSQAASSDSSQGSDVSLTA (SEQ ID NO: 17) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0137] GGSYSQAASSDSTQGSDVSLTA (SEQ ID NO: 18) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0138] GGSYSQAASSDSVQGSDVSLTA (SEQ ID NO: 19) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0139] GGSYSQAASSNSAQGSDVSLTA (SEQ ID NO: 20) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0140] GGSYSQAASSYSAQGSDVSLTA (SEQ ID NO: 21 ) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0141] GGSYSQAVSSDSAQGSDVSLTA (SEQ ID NO: 22) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0142] GGSYYQAASSDSAQGSDVSLTA (SEQ ID NO: 23) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0143] GASYSQAASSDSAQGSDVSLTA (SEQ ID NO: 24) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0144] GGSYSQGASSDSAQGSDVSLTA (SEQ ID NO: 25) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0145] GGSYSQAGSSDSAQGSDVSLTA (SEQ ID NO: 26) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0146] GGSYSQAASSDSGQGSDVSLTA (SEQ ID NO: 27) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto,
[0147] GGSYSQAASSDSAQASDVSLTA (SEQ ID NO: 28) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto, and
[0148] GGSYSQAASSDSAQGSDVSLTG (SEQ ID NO: 29) or a sequence comprising or consisting of a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto.
[0149] In an embodiment, the present invention uses at least two different further sequences selected from SEQ ID NOs: 2 to 29.
[0150] In an embodiment, the different further sequences are selected from the following: SEQ ID NO: 2 and SEQ ID NO: 3, SEQ ID NO: 2 and SEQ ID NO: 4, SEQ ID NO: 2 and SEQ ID NO: 5, SEQ ID NO: 2 and SEQ ID NO: 6, SEQ ID NO: 2 and SEQ ID NO: 7, SEQ ID NO: 2 and SEQ ID NO: 8, SEQ ID NO: 2 and SEQ ID NO: 9, SEQ ID NO: 2 and SEQ ID NO: 10, SEQ ID NO: 2 and SEQ ID NO: 11 , SEQ ID NO: 2 and SEQ ID NO: 12, SEQ ID NO: 2 and SEQ ID NO: 13, SEQ ID NO: 2 and SEQ ID NO: 14, SEQ ID NO: 2 and SEQ ID NO: 15, SEQ ID NO: 2 and SEQ ID NO: 16, SEQ ID NO: 2 and SEQ ID NO: 17, SEQ ID NO: 2 and SEQ ID NO: 18, SEQ ID NO: 2 and SEQ ID NO: 19, SEQ ID NO: 2 and SEQ ID NO: 20, SEQ ID NO: 2 and SEQ ID NO: 21 , SEQ ID NO: 2 and SEQ ID NO: 22, SEQ ID NO: 2 and SEQ ID NO: 23, SEQ ID NO: 2 and SEQ ID NO: 24, SEQ ID NO: 2 and SEQ ID NO: 25, SEQ ID NO: 2 and SEQ ID NO: 26, SEQ ID NO: 2 and SEQ ID NO: 27, SEQ ID NO: 2 and SEQ ID NO: 28, SEQ ID NO: 2 and SEQ ID NO: 29, SEQ ID NO: 3 and SEQ ID NO: 4, SEQ ID NO: 3 and SEQ ID NO: 5, SEQ ID NO: 3 and SEQ ID NO: 6, SEQ ID NO: 3 and SEQ ID NO: 7, SEQ ID NO: 3 and SEQ ID NO: 8, SEQ ID NO: 3 and SEQ ID NO: 9, SEQ ID NO: 3 and SEQ ID NO: 10, SEQ ID NO: 3 and SEQ ID NO: 11 , SEQ ID NO: 3 and SEQ ID NO:
[0151] 12, SEQ ID NO: 3 and SEQ ID NO: 13, SEQ ID NO: 3 and SEQ ID NO: 14, SEQ ID NO: 3 and SEQ ID NO: 15, SEQ ID NO: 3 and SEQ ID NO: 16, SEQ ID NO: 3 and SEQ ID NO: 17, SEQ ID NO: 3 and SEQ ID NO: 18, SEQ ID NO: 3 and SEQ ID NO:
[0152] 19, SEQ ID NO: 3 and SEQ ID NO: 20, SEQ ID NO: 3 and SEQ ID NO: 21 , SEQ ID NO: 3 and SEQ ID NO: 22, SEQ ID NO: 3 and SEQ ID NO: 23, SEQ ID NO: 3 and SEQ ID NO: 24, SEQ ID NO: 3 and SEQ ID NO: 25, SEQ ID NO: 3 and SEQ ID NO:
[0153] 26, SEQ ID NO: 3 and SEQ ID NO: 27, SEQ ID NO: 3 and SEQ ID NO: 28, SEQ ID NO: 3 and SEQ ID NO: 29, SEQ ID NO: 4 and SEQ ID NO: 5, SEQ ID NO: 4 and SEQ ID NO: 6, SEQ ID NO: 4 and SEQ ID NO: 7, SEQ ID NO: 4 and SEQ ID NO: 8, SEQ ID NO: 4 and SEQ ID NO: 9, SEQ ID NO: 4 and SEQ ID NO: 10, SEQ ID NO: 4 and SEQ ID NO: 11 , SEQ ID NO: 4 and SEQ ID NO: 12, SEQ ID NO: 4 and SEQ ID NO:
[0154] 13, SEQ ID NO: 4 and SEQ ID NO: 14, SEQ ID NO: 4 and SEQ ID NO: 15, SEQ ID NO: 4 and SEQ ID NO: 16, SEQ ID NO: 4 and SEQ ID NO: 17, SEQ ID NO: 4 and SEQ ID NO: 18, SEQ ID NO: 4 and SEQ ID NO: 19, SEQ ID NO: 4 and SEQ ID NO:
[0155] 20, SEQ ID NO: 4 and SEQ ID NO: 21 , SEQ ID NO: 4 and SEQ ID NO: 22, SEQ ID NO: 4 and SEQ ID NO: 23, SEQ ID NO: 4 and SEQ ID NO: 24, SEQ ID NO: 4 and SEQ ID NO: 25, SEQ ID NO: 4 and SEQ ID NO: 26, SEQ ID NO: 4 and SEQ ID NO:
[0156] 27, SEQ ID NO: 4 and SEQ ID NO: 28, SEQ ID NO: 4 and SEQ ID NO: 29, SEQ ID NO: 5 and SEQ ID NO: 6, SEQ ID NO: 5 and SEQ ID NO: 7, SEQ ID NO: 5 and SEQ ID NO: 8, SEQ ID NO: 5 and SEQ ID NO: 9, SEQ ID NO: 5 and SEQ ID NO: 10, SEQ ID NO: 5 and SEQ ID NO: 11 , SEQ ID NO: 5 and SEQ ID NO: 12, SEQ ID NO: 5 and SEQ ID NO: 13, SEQ ID NO: 5 and SEQ ID NO: 14, SEQ ID NO: 5 and SEQ ID NO: 15, SEQ ID NO: 5 and SEQ ID NO: 16, SEQ ID NO: 5 and SEQ ID NO: 17, SEQ ID NO: 5 and SEQ ID NO: 18, SEQ ID NO: 5 and SEQ ID NO: 19, SEQ ID NO: 5 and SEQ ID NO: 20, SEQ ID NO: 5 and SEQ ID NO: 21 , SEQ ID NO: 5 and SEQ ID NO: 22, SEQ ID NO: 5 and SEQ ID NO: 23, SEQ ID NO: 5 and SEQ ID NO: 24, SEQ ID NO: 5 and SEQ ID NO: 25, SEQ ID NO: 5 and SEQ ID NO: 26, SEQ ID NO: 5 and SEQ ID NO: 27, SEQ ID NO: 5 and SEQ ID NO: 28, SEQ ID NO: 5 and SEQ ID NO: 29, SEQ ID NO: 6 and SEQ ID NO: 7, SEQ ID NO: 6 and SEQ ID NO: 8, SEQ ID NO: 6 and SEQ ID NO: 9, SEQ ID NO: 6 and SEQ ID NO: 10, SEQ ID NO: 6 and SEQ ID NO: 11 , SEQ ID NO: 6 and SEQ ID NO: 12, SEQ ID NO: 6 and SEQ ID NO: 13, SEQ ID NO: 6 and SEQ ID NO: 14, SEQ ID NO: 6 and SEQ ID NO: 15, SEQ ID NO: 6 and SEQ ID NO: 16, SEQ ID NO: 6 and SEQ ID NO: 17, SEQ ID NO: 6 and SEQ ID NO:
[0157] 18, SEQ ID NO: 6 and SEQ ID NO: 19, SEQ ID NO: 6 and SEQ ID NO: 20, SEQ ID NO: 6 and SEQ ID NO: 21 , SEQ ID NO: 6 and SEQ ID NO: 22, SEQ ID NO: 6 and SEQ ID NO: 23, SEQ ID NO: 6 and SEQ ID NO: 24, SEQ ID NO: 6 and SEQ ID NO:
[0158] 25, SEQ ID NO: 6 and SEQ ID NO: 26, SEQ ID NO: 6 and SEQ ID NO: 27, SEQ ID NO: 6 and SEQ ID NO: 28, SEQ ID NO: 6 and SEQ ID NO: 29, SEQ ID NO: 7 and SEQ ID NO: 8, SEQ ID NO: 7 and SEQ ID NO: 9, SEQ ID NO: 7 and SEQ ID NO: 10, SEQ ID NO: 7 and SEQ ID NO: 11 , SEQ ID NO: 7 and SEQ ID NO: 12, SEQ ID NO: 7 and SEQ ID NO: 13, SEQ ID NO: 7 and SEQ ID NO: 14, SEQ ID NO: 7 and SEQ ID NO: 15, SEQ ID NO: 7 and SEQ ID NO: 16, SEQ ID NO: 7 and SEQ ID NO: 17, SEQ ID NO: 7 and SEQ ID NO: 18, SEQ ID NO: 7 and SEQ ID NO: 19, SEQ ID NO: 7 and SEQ ID NO: 20, SEQ ID NO: 7 and SEQ ID NO: 21 , SEQ ID NO: 7 and SEQ ID NO: 22, SEQ ID NO: 7 and SEQ ID NO: 23, SEQ ID NO: 7 and SEQ ID NO: 24, SEQ ID NO: 7 and SEQ ID NO: 25, SEQ ID NO: 7 and SEQ ID NO: 26, SEQ ID NO: 7 and SEQ ID NO: 27, SEQ ID NO: 7 and SEQ ID NO: 28, SEQ ID NO: 7 and SEQ ID NO: 29, SEQ ID NO: 8 and SEQ ID NO: 9, SEQ ID NO: 8 and SEQ ID NO: 10, SEQ ID NO: 8 and SEQ ID NO: 11 , SEQ ID NO: 8 and SEQ ID NO: 12, SEQ ID NO: 8 and SEQ ID NO: 13, SEQ ID NO: 8 and SEQ ID NO: 14, SEQ ID NO: 8 and SEQ ID NO: 15, SEQ ID NO: 8 and SEQ ID NO: 16, SEQ ID NO: 8 and SEQ ID NO: 17, SEQ ID NO: 8 and SEQ ID NO: 18, SEQ ID NO: 8 and SEQ ID NO: 19, SEQ ID NO: 8 and SEQ ID NO: 20, SEQ ID NO: 8 and SEQ ID NO: 21 , SEQ ID NO: 8 and SEQ ID NO: 22, SEQ ID NO: 8 and SEQ ID NO: 23, SEQ ID NO: 8 and SEQ ID NO: 8, SEQ ID NO: 8 and SEQ ID NO: 25, SEQ ID NO: 8 and SEQ ID NO: 8, SEQ ID NO: 8 and SEQ ID NO: 27, SEQ ID NO: 8 and SEQ ID NO: 28, SEQ ID NO: 8 and SEQ ID NO: 29, SEQ ID NO: 9 and SEQ ID NO: 10, SEQ ID NO: 9 and SEQ ID NO: 11 , SEQ ID NO: 9 and SEQ ID NO: 12, SEQ ID NO: 9 and SEQ ID NO: 13, SEQ ID NO: 9 and SEQ ID NO: 14, SEQ ID NO: 9 and SEQ ID NO: 15, SEQ ID NO: 9 and SEQ ID NO: 16, SEQ ID NO: 9 and SEQ ID NO: 17, SEQ ID NO: 9 and SEQ ID NO: 18, SEQ ID NO: 9 and SEQ ID NO:
[0159] 19, SEQ ID NO: 9 and SEQ ID NO: 20, SEQ ID NO: 9 and SEQ ID NO: 21 , SEQ ID NO: 9 and SEQ ID NO: 22, SEQ ID NO: 9 and SEQ ID NO: 23, SEQ ID NO: 9 and SEQ ID NO: 24, SEQ ID NO: 9 and SEQ ID NO: 25, SEQ ID NO: 9 and SEQ ID NO:
[0160] 26, SEQ ID NO: 9 and SEQ ID NO: 27, SEQ ID NO: 9 and SEQ ID NO: 28, SEQ ID NO: 9 and SEQ ID NO: 29, SEQ ID NO: 10 and SEQ ID NO: 11 , SEQ ID NO: 10 and SEQ ID NO: 12, SEQ ID NO: 10 and SEQ ID NO: 13, SEQ ID NO: 10 and SEQ ID NO: 14, SEQ ID NO: 10 and SEQ ID NO: 15, SEQ ID NO: 10 and SEQ ID NO: 16, SEQ ID NO: 10 and SEQ ID NO: 17, SEQ ID NO: 10 and SEQ ID NO: 18, SEQ ID NO: 10 and SEQ ID NO: 19, SEQ ID NO: 10 and SEQ ID NO: 20, SEQ ID NO: 10 and SEQ ID NO: 21 , SEQ ID NO: 10 and SEQ ID NO: 22, SEQ ID NO: 10 and SEQ ID NO: 23, SEQ ID NO: 10 and SEQ ID NO: 24, SEQ ID NO: 10 and SEQ ID NO: 25, SEQ ID NO: 10 and SEQ ID NO: 26, SEQ ID NO: 10 and SEQ ID NO: 27, SEQ ID NO: 10 and SEQ ID NO: 28, SEQ ID NO: 10 and SEQ ID NO: 29, SEQ ID NO: 11 and SEQ ID NO: 12, SEQ ID NO: 11 and SEQ ID NO: 13, SEQ ID NO: 11 and SEQ ID NO: 14, SEQ ID NO: 11 and SEQ ID NO: 15, SEQ ID NO: 11 and SEQ ID NO: 16, SEQ ID NO: 11 and SEQ ID NO:
[0161] 17, SEQ ID NO: 11 and SEQ ID NO: 18, SEQ ID NO: 11 and SEQ ID NO: 19, SEQ ID NO: 11 and SEQ ID NO: 20, SEQ ID NO: 11 and SEQ ID NO: 21 , SEQ ID NO: 11 and SEQ ID NO: 22, SEQ ID NO: 11 and SEQ ID NO: 23, SEQ ID NO: 11 and SEQ ID NO:
[0162] 24, SEQ ID NO: 11 and SEQ ID NO: 25, SEQ ID NO: 11 and SEQ ID NO: 26, SEQ ID NO: 11 and SEQ ID NO: 27, SEQ ID NO: 11 and SEQ ID NO: 28, SEQ ID NO: 11 and SEQ ID NO: 29, SEQ ID NO: 12 and SEQ ID NO: 13, SEQ ID NO: 12 and SEQ ID NO: 14, SEQ ID NO: 12 and SEQ ID NO: 15, SEQ ID NO: 12 and SEQ ID NO: 16, SEQ ID NO: 12 and SEQ ID NO: 17, SEQ ID NO: 12 and SEQ ID NO: 18, SEQ ID NO: 12 and SEQ ID NO: 19, SEQ ID NO: 12 and SEQ ID NO: 20, SEQ ID NO: 12 and SEQ ID NO: 21 , SEQ ID NO: 12 and SEQ ID NO: 22, SEQ ID NO: 12 and SEQ ID NO: 23, SEQ ID NO: 12 and SEQ ID NO: 24, SEQ ID NO: 12 and SEQ ID NO: 25, SEQ ID NO: 12 and SEQ ID NO: 26, SEQ ID NO: 12 and SEQ ID NO: 27, SEQ ID NO: 12 and SEQ ID NO: 28, SEQ ID NO: 12 and SEQ ID NO: 29, SEQ ID NO: 13 and SEQ ID NO: 14, SEQ ID NO: 13 and SEQ ID NO: 15, SEQ ID NO: 13 and SEQ ID NO: 16, SEQ ID NO: 13 and SEQ ID NO: 17, SEQ ID NO: 13 and SEQ ID NO: 18, SEQ ID NO: 13 and SEQ ID NO: 19, SEQ ID NO: 13 and SEQ ID NO: 20, SEQ ID NO: 13 and SEQ ID NO: 21 , SEQ ID NO: 13 and SEQ ID NO: 22, SEQ ID NO: 13 and SEQ ID NO: 23, SEQ ID NO: 13 and SEQ ID NO: 24, SEQ ID NO: 13 and SEQ ID NO: 25, SEQ ID NO: 13 and SEQ ID NO: 26, SEQ ID NO: 13 and SEQ ID NO: 27, SEQ ID NO: 13 and SEQ ID NO: 28, SEQ ID NO: 13 and SEQ ID NO: 29, SEQ ID NO: 14 and SEQ ID NO: 15, SEQ ID NO: 14 and SEQ ID NO: 16, SEQ ID NO: 14 and SEQ ID NO: 17, SEQ ID NO: 14 and SEQ ID NO:
[0163] 18, SEQ ID NO: 14 and SEQ ID NO: 19, SEQ ID NO: 14 and SEQ ID NO: 20, SEQ ID NO: 14 and SEQ ID NO: 21 , SEQ ID NO: 14 and SEQ ID NO: 22, SEQ ID NO: 14 and SEQ ID NO: 23, SEQ ID NO: 14 and SEQ ID NO: 24, SEQ ID NO: 14 and SEQ ID NO:
[0164] 25, SEQ ID NO: 14 and SEQ ID NO: 26, SEQ ID NO: 14 and SEQ ID NO: 27, SEQ ID NO: 14 and SEQ ID NO: 28, SEQ ID NO: 14 and SEQ ID NO: 29, SEQ ID NO: 15 and SEQ ID NO: 16, SEQ ID NO: 15 and SEQ ID NO: 17, SEQ ID NO: 15 and SEQ ID NO:
[0165] 18, SEQ ID NO: 15 and SEQ ID NO: 19, SEQ ID NO: 15 and SEQ ID NO: 20, SEQ ID NO: 15 and SEQ ID NO: 21 , SEQ ID NO: 15 and SEQ ID NO: 22, SEQ ID NO: 15 and SEQ ID NO: 23, SEQ ID NO: 15 and SEQ ID NO: 24, SEQ ID NO: 15 and SEQ ID NO:
[0166] 25, SEQ ID NO: 15 and SEQ ID NO: 26, SEQ ID NO: 15 and SEQ ID NO: 27, SEQ ID NO: 15 and SEQ ID NO: 28, SEQ ID NO: 15 and SEQ ID NO: 29, SEQ ID NO: 16 and SEQ ID NO: 17, SEQ ID NO: 16 and SEQ ID NO: 18, SEQ ID NO: 16 and SEQ ID NO:
[0167] 19, SEQ ID NO: 16 and SEQ ID NO: 20, SEQ ID NO: 16 and SEQ ID NO: 21 , SEQ ID NO: 16 and SEQ ID NO: 22, SEQ ID NO: 16 and SEQ ID NO: 23, SEQ ID NO: 16 and SEQ ID NO: 24, SEQ ID NO: 16 and SEQ ID NO: 25, SEQ ID NO: 16 and SEQ ID NO:
[0168] 26, SEQ ID NO: 16 and SEQ ID NO: 27, SEQ ID NO: 16 and SEQ ID NO: 28, SEQ ID NO: 16 and SEQ ID NO: 29, SEQ ID NO: 17 and SEQ ID NO: 18, SEQ ID NO: 17 and SEQ ID NO: 19, SEQ ID NO: 17 and SEQ ID NO: 20, SEQ ID NO: 17 and SEQ ID NO: 21 , SEQ ID NO: 17 and SEQ ID NO: 22, SEQ ID NO: 17 and SEQ ID NO: 23, SEQ ID NO: 17 and SEQ ID NO: 24, SEQ ID NO: 17 and SEQ ID NO: 25, SEQ ID NO: 17 and SEQ ID NO: 26, SEQ ID NO: 17 and SEQ ID NO: 27, SEQ ID NO: 17 and SEQ ID NO: 28, SEQ ID NO: 17 and SEQ ID NO: 29, SEQ ID NO: 18 and SEQ ID NO: 19, SEQ ID NO: W and SEQ ID NO: 20, SEQ ID NO: 18 and SEQ ID NO: 21 , SEQ ID NO: and SEQ ID NO: 22, SEQ ID NO: 18 and SEQ ID NO: 23, SEQ ID NO: 18 and SEQ ID NO:
[0169] 24, SEQ ID NO: 18 and SEQ ID NO: 25, SEQ ID NO: 18 and SEQ ID NO: 26, SEQ ID NO: 18 and SEQ ID NO: 27, SEQ ID NO: 18 and SEQ ID NO: 28, SEQ ID NO: 18 and SEQ ID NO: 29, SEQ ID NO: 19 and SEQ ID NO: 20, SEQ ID NO: 19 and SEQ ID NO: 21 , SEQ ID NO: 19 and SEQ ID NO: 22, SEQ ID NO: 19 and SEQ ID NO: 23, SEQ ID NO: 19 and SEQ ID NO: 24, SEQ ID NO: 19 and SEQ ID NO: 25, SEQ ID NO: 19 and SEQ ID NO: 26, SEQ ID NO: 19 and SEQ ID NO: 27, SEQ ID NO: 19 and SEQ ID NO: 28, SEQ ID NO: 19 and SEQ ID NO: 29, SEQ ID NO: 20 and SEQ ID NO: 21 , SEQ ID NO: 20 and SEQ ID NO: 22, SEQ ID NO: 20 and SEQ ID NO: 23, SEQ ID NO: 20 and SEQ ID NO: 24, SEQ ID NO: 20 and SEQ ID NO: 25, SEQ ID NO: 20 and SEQ ID NO: 26, SEQ ID NO: 20 and SEQ ID NO: 27, SEQ ID NO: 20 and SEQ ID NO: 28, SEQ ID NO: 20 and SEQ ID NO: 29, SEQ ID NO: 21 and SEQ ID NO: 22, SEQ ID NO: 21 and SEQ ID NO: 23, SEQ ID NO: 21 and SEQ ID NO: 24, SEQ ID NO: 21 and SEQ ID NO:
[0170] 25, SEQ ID NO: 21 and SEQ ID NO: 26, SEQ ID NO: 21 and SEQ ID NO: 27, SEQ ID NO: 21 and SEQ ID NO: 28, SEQ ID NO: 21 and SEQ ID NO: 29, SEQ ID NO: 22 and SEQ ID NO: 23, SEQ ID NO: 22 and SEQ ID NO: 24, SEQ ID NO: 22 and SEQ ID NO:
[0171] 25, SEQ ID NO: 22 and SEQ ID NO: 26, SEQ ID NO: 22 and SEQ ID NO: 27, SEQ ID NO: 22 and SEQ ID NO: 28, SEQ ID NO: 22 and SEQ ID NO: 29, SEQ ID NO: 23 and SEQ ID NO: 24, SEQ ID NO: 23 and SEQ ID NO: 25, SEQ ID NO: 23 and SEQ ID NO:
[0172] 26, SEQ ID NO: 23 and SEQ ID NO: 27, SEQ ID NO: 23 and SEQ ID NO: 28,- SEQ ID NO: 23 and SEQ ID NO: 29, SEQ ID NO: 24 and SEQ ID NO: 25, SEQ ID NO: 24 and SEQ ID NO: 26, SEQ ID NO: 24 and SEQ ID NO: 27, SEQ ID NO: 24 and SEQ ID NO: 28, SEQ ID NO: 24 and SEQ ID NO: 29, SEQ ID NO: 25 and SEQ ID NO: 26, SEQ ID NO: 25 and SEQ ID NO: 27, SEQ ID NO: 25 and SEQ ID NO: 28, SEQ ID NO: 25 and SEQ ID NO: 29, SEQ ID NO: 26 and SEQ ID NO: 27, SEQ ID NO: 26 and SEQ ID NO: 28, SEQ ID NO: 26 and SEQ ID NO: 29, SEQ ID NO: 27 and SEQ ID NO: 28, SEQ ID NO: 27 and SEQ ID NO: 29, and SEQ ID NO: 28 and SEQ ID NO: 29.
[0173] Unless otherwise specified, reference to each SEQ ID NO extends to sequences having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto.
[0174] In an embodiment, the present invention uses at least three different further sequences selected from SEQ ID NOs: 2 to 29. In an embodiment, the three different further sequences are selected from SEQ ID NO: 2 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 3 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 4 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 5 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 6 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 7 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 8 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 9 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 10 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 11 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 12 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 13 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 14 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 15 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 16 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 17 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 18 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 19 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 20 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 21 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 22 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 23 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 24 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 25 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 26 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 27 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 28 and any other combination listed above. In an embodiment, the three different further sequences are selected from SEQ ID NO: 29 and any other combination listed above. Unless otherwise specified, reference to each SEQ ID NO extends to sequences having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto.
[0175] In a preferred embodiment, the different further sequences are selected from SEQ ID NO: 7, SEQ ID NO: 8, SEQ ID NO: 13, SEQ ID NO: 15, SEQ ID NO: 17, SEQ ID NO: 18, SEQ ID NO: 22, SEQ ID NO: 25, SEQ ID NO: 26 and SEQ ID NO: 27. In a preferred embodiment, there are two different further sequences comprising SEQ ID NO: 7 and SEQ ID NO: 8, SEQ ID NO: 7 and SEQ ID NO: 13, SEQ ID NO: 7 and SEQ ID NO: 22, SEQ ID NO: 7 and SEQ ID NO: 25, SEQ ID NO: 7 and SEQ ID NO: 26, SEQ ID NO: 7 and SEQ ID NO: 27, SEQ ID NO: 8 and SEQ ID NO: 13, SEQ ID NO: 8 and SEQ ID NO: 22, SEQ ID NO: 8 and SEQ ID NO: 25, SEQ ID NO: 8 and SEQ ID NO: 26, SEQ ID NO: 8 and SEQ ID NO: 27, SEQ ID NO: 13 and SEQ ID NO: 22, SEQ ID NO: 13 and SEQ ID NO: 25, SEQ ID NO: 13 and SEQ ID NO: 26, SEQ ID NO: 13 and SEQ ID NO: 27, SEQ ID NO: 22 and SEQ ID NO: 25, SEQ ID NO: 22 and SEQ ID NO: 26, SEQ ID NO: 22 and SEQ ID NO: 27, SEQ ID NO: 25 and SEQ ID NO: 26, SEQ ID NO: 25 and SEQ ID NO: 27, and SEQ ID NO: 26 and SEQ ID NO: 27. In a preferred embodiment, there are three different further sequences comprising SEQ ID NO: 7 and any other combination listed above, SEQ ID NO: 8 and any other combination listed above, SEQ ID NO: 13 and any other combination listed above, SEQ ID NO: 15 and any other combination listed above, SEQ ID NO: 17 and any other combination listed above, SEQ ID NO: 18 and any other combination listed above, SEQ ID NO: 22 and any other combination listed above, SEQ ID NO: 25 and any other combination listed above, SEQ ID NO: 26 and any other combination listed above, or SEQ ID NO: 27 and any other combination listed above. Unless otherwise specified, reference to each SEQ ID NO extends to sequences having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto.
[0176] In some embodiments, the present invention uses one, two, three or four different MITD sequences comprising amino acid sequences selected from SEQ ID NO: 5, 2, 22 and 21. In some embodiment, the present invention uses one, two, three or four different MTID sequences comprising either a wild-type sequence or the mutation selected from Y4N, A8V and D11 Y. In some embodiments, the present invention uses one, two, three, four or five different MITD sequences comprising amino acid sequences selected from SEQ ID NO: 8, 15, 22, 13 and 7. In some embodiments, the present invention uses one, two, three, four or five different MITD sequences comprising the mutation selected from S F, G15V, A8V, D17V and S9F. In some embodiments, the present invention uses one, two, three, four or five different MITD sequences comprising amino acid sequences selected from SEQ ID NO: 23, 7, 15, 22 and 13. In some embodiments, the present invention uses one, two, three, four or five different MITD sequences comprising the mutation selected from S5Y, S9F, G15V, A8V and D17V. In some embodiments, the present invention uses one, two, three, four, five or six different MITD sequences comprising amino acid sequences selected from SEQ ID NO: 25, 26, 22, 27, 17 and 18. In some embodiments, the present invention uses one, two, three, four, five or six different MITD sequences comprising the mutation selected from A7G, A8G, A8V, A13G, A13S and A13T. References to SEQ ID NOs and mutations in such embodiments are understood to extend to other sequences that are defined in respect of those SEQ ID NOs and mutations herein, e.g. by % identity to SEQ ID NO and / or the mutation as being relative to SEQ ID NO: 2.
[0177] In some preferred embodiments, the combinations of further sequences listed above are used in combination with a wild-type MITD sequence. In some preferred embodiments, the combinations of further sequences listed above are used in combination with an MITD sequence comprising a further sequence of SEQ ID NO: 2 or a sequence having 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, 99% or 100% identity thereto.
[0178] In an embodiment, the different MITD sequences each have substantially the same function as a naturally occurring MITD sequence. In an embodiment, the function is the cellular trafficking function, the antigen / epitope presentation function, or the function as a vaccine when expressed as a fusion protein comprising an antigen / epitope sequence. In some embodiments, the different MITD sequences have substantially the same function as a naturally occurring MITD sequence but improved immunogenicity relative to the naturally occurring MITD sequence.
[0179] In an embodiment, the different MITD sequences comprise polymorphisms and / or mutants e.g. substitution mutants differing from naturally occurring sequences by at least one amino acid. In an embodiment, the different MITD sequences are naturally occurring polymorphisms. In an embodiment, the different MITD sequences are mutant variants. In an embodiment, each different MITD sequence comprises at least one amino acid difference (e.g. substitution) relative to each other different MITD sequence in the method / composition / kit / mixture / assay of the present invention. In an embodiment, “different” MITD sequences means that the MITD sequences can be distinguished from each other, e.g. by sequence. In embodiments, the abundance of each different MITD sequence can be differentiated from the abundance of the other MITD sequences and detected. In some embodiments, “different” MITD sequences means that the MITDs have different amino acid sequences but the same function, such as any or all of the MITD sequence functions described herein. In core embodiments of the present invention, the MITD sequences may be defined as functionally identical but analytically distinct, i.e. relative to each other MITD sequence used in the present invention.
[0180] In an embodiment, the MITD are functionally identical / equivalent to a naturally occurring MITD sequence, such as a polymorphism, such as any of SEQ ID NOs 2 to 29. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 3. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 4. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 5. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 6. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 7. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 8. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 9. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 10. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 11. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 12. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 13. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 14. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 15. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 16. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 17. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 18. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 19. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 20. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 21. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 22. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 23. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 24. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 25. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 26. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 27. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 28. In an embodiment, the MITD are functionally identical / equivalent to SEQ ID NO: 29. In an embodiment, functionally identical means that the MITD sequences elicit the same level of an immune response in vivo. In an embodiment, functionally identical means that the MITD sequences elicit at least the same level of an immune response in vivo. In an embodiment, functionally identical means that the MITD sequences elicit the same or greater level of an immune response in vivo.
[0181] Functional sequences
[0182] The present invention ultimately may analyse and quantify the expression of nucleic acids to produce amino acid sequences. In embodiments, each of the amino acid sequences of the present invention comprises a different MITD sequence, and may also comprise one or more functional sequences.
[0183] In some embodiments, each of the amino acid sequences further comprises one or more antigen or epitope sequences, such as one or more cancer antigen, neoantigen, infectious disease antigen, autoimmune disease antigen, T-cell epitope, fixed epitope or variable epitope sequences.
[0184] In some embodiments, the at least two different nucleic acid sequences comprise DNA and / or RNA. In some embodiments, the at least two different nucleic acid sequences comprise or consist of messenger RNA (mRNA).
[0185] Preferably, the functional sequences are “antigen sequences” or “epitope sequences”. In an embodiment, the functional sequences are peptides or polypeptides with therapeutic potential. In an embodiment, the functional sequences are antigens or epitopes, e.g. antigens or epitopes associated with cancer, infectious diseases or genetic diseases. In an embodiment, each functional sequence can comprise more than one antigen or epitope sequence. In an embodiment, the functional sequences are antigens. In an embodiment, the functional sequences are epitopes. In an embodiment, the functional sequences are T-cell epitopes. In an embodiment, the functional sequences are epitopes that are presented to a T-cell, e.g. as part of a fusion protein comprising an MITD sequence.
[0186] In an embodiment, the functional sequences are fixed antigens or variable epitopes. In an embodiment, the functional sequences are fixed antigens, such as known tumour antigens. In another embodiment, the functional sequences are variable epitopes or highly variable epitopes, such as patient-specific epitopes, or personalized epitopes, or de novo epitope sequences, or undetermined epitope sequences.
[0187] In an embodiment, the functional sequences are peptides or polypeptides having biological activity. In some embodiments, the peptides or polypeptides having biological activity are selected from the group consisting of vaccines (e.g., antigens, epitopes), proteins for replacement therapy, antibodies, antibody-like molecules, and cytokines. In some embodiments, the peptides or polypeptides having biological activity constitute a vaccine. In some embodiments, the vaccine is a vaccine against an infectious disease. In some embodiments, the vaccine is a vaccine against an autoimmune disease. In some embodiments, the vaccine is a vaccine against a genetic disease. In some embodiments, the vaccine is a cancer vaccine. In some embodiments, the vaccine is a T cell vaccine. In some embodiments, the functional sequences of the present invention are components of a multivalent T cell vaccine.
[0188] Positioning of MITD sequences
[0189] In some embodiments, each of the MITD sequences comprises the C or N terminus of its respective amino acid sequence, preferably the C terminus. In other words, each of the MITD sequences is at, or is located at, the C or N terminus of its respective amino acid sequence, preferably the C terminus. In some embodiments, for each MITD sequence, either the C-terminus of the MITD sequence is the C-terminus of the amino acid sequence in which the MITD sequence resides, or the N-terminus of the MITD sequence is the N-terminus of the amino acid sequence in which the MITD sequence resides. In some embodiments, for each MITD sequence, the C-terminus of the MITD sequence is the C-terminus of the amino acid sequence in which the MITD sequence resides. In some embodiments, it will be understood that each amino acid sequence of the present invention does not comprise further amino acid residues at either end (C and N termini) of the MITD sequence. In some embodiments, it will be understood that each amino acid sequence of the present invention does not comprise further amino acid residues at / adjacent to the C terminus of the MITD sequence.
[0190] Compositions
[0191] In an aspect, the present invention provides a composition comprising at least two different nucleic acid sequences, wherein each of the nucleic acid sequences encodes a different amino acid sequence comprising: a) at least one antigen or epitope sequence; and b) an MITD sequence; wherein the MITD sequences in each of the different amino acid sequences are different.
[0192] In an aspect, the present invention provides a composition comprising at least two different amino acid sequences, wherein each of the amino acid sequences comprises: a) at least one antigen or epitope sequence; and b) an MITD sequence; wherein the MITD sequences in each of the different amino acid sequences are different.
[0193] In some embodiments, the composition is a vaccine. In some embodiments, the vaccine is a vaccine against an infectious disease. In some embodiments, the vaccine is a vaccine against an autoimmune disease. In some embodiments, the vaccine is a vaccine against a genetic disease. In some embodiments, the vaccine is a cancer vaccine. In some embodiments, the vaccine is a T cell vaccine. In a preferred embodiment, particularly for the nucleic acid composition, the composition is a T cell vaccine composition. In embodiments, the composition is a multivalent T cell vaccine composition. In embodiments, one or more or each amino acid sequence encoded by the at least two nucleic acid sequences comprises an antigen and / or epitope sequence as described herein.
[0194] Kits and uses:
[0195] In an aspect, the present invention provides a kit comprising at least two different nucleic acid sequences, wherein each nucleic acid sequence encodes an amino acid sequence comprising at least one antigen or epitope sequence and an MITD sequence, wherein the MITD sequences in each of the different amino acid sequences are different.
[0196] In an aspect, the present invention provides a use of the composition or kit of the invention, for simultaneously analysing the expression level of at least two different nucleic acid sequences.
[0197] In an embodiment, the use of a composition or kit according to the present invention is for simultaneously analysing the expression level of at least two different (e.g. a first, second and one or more further) nucleic acid sequences.
[0198] Generally, in embodiments herein, “expression” refers to the ability of a nucleic acid sequence to produce an amino acid sequence, and “expression level” refers to the quantity of the amino acid sequence that is produced from a nucleic acid sequence. It is understood that, with other factors being controlled for, the skilled person may determine and compare the expression levels of two or more nucleic acid sequences as a characteristic of the nucleic acid sequences per se. Of particular interest for the present invention are expression levels in a cell, tissue or organism, such as a human or animal cell, tissue or organism, such as in a human cell, human tissue or human. Applications: the present invention can be broadly applied to any method of analysing at least two different nucleic acid sequences, particularly for analysing the expression level or potency of at least two different nucleic acid sequences. In an embodiment, the present invention is a method of analysing the expression levels of at least two different nucleic acid sequences for expressing at least two different antigen or epitope sequences. In an embodiment, the present invention is a method of analysing the expression levels of at least two nucleic acid sequences for expressing at least two antigen or epitope sequences in a biological system.
[0199] In an embodiment, the biological system is not particularly limited. In an embodiment, the biological system is a cellular assay. In another embodiment, the biological system is an ex vivo tissue sample. In an embodiment, the biological system is an ex vivo tissue sample from a rodent such as a mouse. In an embodiment, the biological system is an ex vivo tissue sample from a human. In an embodiment, the biological system is in vivo. In an embodiment, the biological system is in vivo in a rodent such as a mouse. In an embodiment, the biological system is in vivo in a human patient.
[0200] In an embodiment, the present invention relates to both in vitro, ex vivo, and in vivo methods and uses. In a different embodiment, the present invention relates to in vitro methods and uses. In a different embodiment, the present invention relates to ex vivo methods and uses. In a different embodiment, the present invention relates to in vivo methods and uses.
[0201] In an embodiment, the present invention is used for analysing a multivalent T cell vaccine.
[0202] In an embodiment, the present invention is used in a diagnostic method. In an embodiment, the present invention is used in a companion diagnostic method or in clinical follow-up studies. In an embodiment, the kit of the present invention is used as a companion diagnostic kit.
[0203] Definitions of general terms
[0204] The practice of the present disclosure will employ, unless otherwise indicated, conventional chemistry, biochemistry, cell biology, immunology, and recombinant DNA techniques which are explained in the literature in the field.
[0205] Throughout this specification and the claims which follow, unless the context requires otherwise, the word "comprise", and variations such as "comprises" and "comprising", will be understood to imply the inclusion of a stated feature, element, member, integer or step or group of features, elements, members, integers or steps but not the exclusion of any other feature, element, member, integer or step or group of features, elements, members, integers or steps. The term "consisting essentially of" limits the scope of a claim or disclosure to the specified features, elements, members, integers, or steps and those that do not materially affect the basic and novel characteristic(s) of the claim or disclosure. The term “consisting of’ limits the scope of a claim or disclosure to the specified features, elements, members, integers, or steps. The term "comprising" encompasses the term "consisting essentially of" which, in turn, encompasses the term "consisting of'. Thus, at each occurrence in the present application, the term "comprising" may be replaced with the term "consisting essentially of" or "consisting of". Likewise, at each occurrence in the present application, the term "consisting essentially of" may be replaced with the term "consisting of".
[0206] The terms "a", "an" and "the" and similar references used in the context of describing the present disclosure (especially in the context of the claims) are to be construed to cover both the singular and the plural, unless otherwise indicated herein or clearly contradicted by the context.
[0207] All methods described herein can be performed in any suitable order unless otherwise indicated herein or otherwise clearly contradicted by the context.
[0208] The use of any and all examples, or exemplary language (e.g., "such as"), provided herein is intended merely to better illustrate the present disclosure and does not pose a limitation on the scope of the present disclosure otherwise claimed. No language in the specification should be construed as indicating any non-claimed element essential to the practice of the present disclosure.
[0209] The term "optional" or "optionally" as used herein means that the subsequently described event, circumstance or condition may or may not occur, and that the description includes instances where said event, circumstance, or condition occurs and instances in which it does not occur.
[0210] Where used herein, "and / or" is to be taken as specific disclosure of each of the two specified features or components with or without the other. For example, "X and / or Y" is to be taken as specific disclosure of each of (i) X, (ii) Y, and (iii) X and Y, just as if each is set out individually herein.
[0211] In the context of the present disclosure, the term "about" denotes an interval of accuracy that the person of ordinary skill will understand to still ensure the technical effect of the feature in question. The term typically indicates deviation from the indicated numerical value by ±10%, ±5%, ±4%, ±3%, ±2%, ±1 %, ±0.9%, ±0.8%, ±0.7%, ±0.6%, ±0.5%, ±0.4%, ±0.3%, ±0.2%, ±0.1 %, ±0.05%, and for example ±0.01 %. In some embodiments, “about” indicates deviation from the indicated numerical value by ±10%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±5%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±4%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±3%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±2%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±1 %. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.9%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.8%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.7%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.6%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.5%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.4%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.3%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.2%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.1 %. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.05%. In some embodiments, “about” indicates deviation from the indicated numerical value by ±0.01 %. As will be appreciated by the person of ordinary skill, the specific such deviation for a numerical value for a given technical effect will depend on the nature of the technical effect. For example, a natural or biological technical effect may generally have a larger such deviation than one for a man-made or engineering technical effect.
[0212] Recitation of ranges of values herein is merely intended to serve as a shorthand method of referring individually to each separate value falling within the range. Unless otherwise indicated herein, each individual value is incorporated into the specification as if it were individually recited herein.
[0213] Each of the documents cited herein (including all patents, patent applications, scientific publications, manufacturer's specifications, instructions, etc.), whether supra or infra, are hereby incorporated by reference in their entirety. Nothing herein is to be construed as an admission that the invention is not entitled to antedate such disclosure by virtue of prior invention.
[0214] Specific definitions
[0215] In the following, definitions will be provided which apply to all aspects of the present disclosure. The following terms have the following meanings unless otherwise indicated. Any undefined terms have their art recognized meanings.
[0216] In some embodiments, “therapeutic potential” and / or “potency” are determined in place of or in addition to the expression levels of the at least two nucleic acids in the present invention. The “therapeutic potential” or "potency" of nucleic acid (such as RNA and / or DNA) refers to the therapeutic quality of the nucleic acid, the ability of the nucleic acid to provide a therapeutic benefit when administered to a subject. In particular embodiments, the therapeutic potential of nucleic acid can be measured, determined, identified, quantified, confirmed and / or validated by expression, in particular strong expression, e.g., expression above a threshold, of the peptide or polypeptide encoded by the nucleic acid that indicates the therapeutic potential of the nucleic acid. In one embodiment, therapeutic potential refers to an ability of a nucleic acid (such as an RNA and / or DNA) to express a pharmaceutically active peptide or polypeptide in vivo said pharmaceutically active peptide or polypeptide exerting its pharmaceutical, e.g., therapeutic, effect.
[0217] In some embodiments, nucleic acid (such as RNA and / or DNA) that shows strong expression, e.g., expression above a threshold, has "sufficient therapeutic potential". The therapeutic potential of the nucleic acid is sufficient if the nucleic acid has the ability in vivo to express an encoded pharmaceutically active peptide or polypeptide such that that a meaningful pharmaceutical, e.g., therapeutic, effect is achieved. In some embodiments, the present invention provides a method of determining the therapeutic potential of at least two nucleic acids.
[0218] As used herein, phrases such as “determining the expression level”, "determining the amount" or "determining expression" or similar phrases with reference to an amino acid sequence (peptide or polypeptide) refer to determining the quantity or presence of an amino acid sequence.
[0219] Terms such as "reduce" or "inhibit" as used herein means the ability to cause an overall decrease, for example, of about 5% or greater, about 10% or greater, about 15% or greater, about 20% or greater, about 25% or greater, about 30% or greater, about 40% or greater, about 50% or greater, or about 75% or greater, in the level. The term "inhibit" or similar phrases includes a complete or essentially complete inhibition, i.e. a reduction to zero or essentially to zero.
[0220] The term "enhance" as used herein means the ability to cause an overall increase, or enhancement, for example, by at least about 5% or greater, about 10% or greater, about 15% or greater, about 20% or greater, about 25% or greater, about 30% or greater, about 40% or greater, about 50% or greater, about 75% or greater, or about 100% or greater in the level.
[0221] "Physiological pH" as used herein refers to a pH of about 7.4. In some embodiments, physiological pH is from 7.3 to 7.5. In some embodiments, physiological pH is from 7.35 to 7.45. In some embodiments, physiological pH is 7.3, 7.35, 7.4, 7.45, or 7.5.
[0222] As used in the present disclosure, "% w / v" refers to weight by volume percent, which is a unit of concentration measuring the amount of solute in grams (g) expressed as a percent of the total volume of solution in milliliters (mL).
[0223] As used in the present disclosure, "% by weight" refers to weight percent, which is a unit of concentration measuring the amount of a substance in grams (g) expressed as a percent of the total weight of the total composition in grams (g).
[0224] As used in the present disclosure, "mol %" is defined as the ratio of the number of moles of one component to the total number of moles of all components, multiplied by 100.
[0225] As used in the present disclosure, "mol % of the total lipid" is defined as the ratio of the number of moles of one lipid component to the total number of moles of all lipids, multiplied by 100. In this context, in some embodiments, the term "total lipid" includes lipids and lipid-like material.
[0226] The term "ionic strength" refers to the mathematical relationship between the number of different kinds of ionic species in a particular solution and their respective charges. Thus, ionic strength I is represented mathematically by the formula: in which c is the molar concentration of a particular ionic species and z the absolute value of its charge. The sum Z is taken over all the different kinds of ions (i) in solution. According to the disclosure, the term "ionic strength" in some embodiments relates to the presence of monovalent ions. Regarding the presence of divalent ions, in particular divalent cations, their concentration or effective concentration (presence of free ions) due to the presence of chelating agents is, in some embodiments, sufficiently low so as to prevent degradation of the nucleic acid. In some embodiments, the concentration or effective concentration of divalent ions is below the catalytic level for hydrolysis of the phosphodiester bonds between nucleotides such as RNA nucleotides. In some embodiments, the concentration of free divalent ions is 20 pM or less. In some embodiments, there are no or essentially no free divalent ions.
[0227] "Osmolality" refers to the concentration of a particular solute expressed as the number of osmoles of solute per kilogram of solvent.
[0228] The term "lyophilizing" or "lyophilization" refers to the freeze-drying of a substance by freezing it and then reducing the surrounding pressure (e.g., below 15 Pa, such as below 10 Pa, below 5 Pa, or 1 Pa or less) to allow the frozen medium in the substance to sublimate directly from the solid phase to the gas phase. Thus, the terms "lyophilizing" and "freeze-drying" are used herein interchangeably.
[0229] The term "spray-drying" refers to spray-drying a substance by mixing (heated) gas with a fluid that is atomized (sprayed) within a vessel (spray dryer), where the solvent from the formed droplets evaporates, leading to a dry powder.
[0230] The term "reconstitute" relates to adding a solvent such as water to a dried product to return it to a liquid state such as its original liquid state.
[0231] The term "recombinant" in the context of the present disclosure means "made through genetic engineering". In one embodiment, a "recombinant object" in the context of the present disclosure is not occurring naturally.
[0232] The term "naturally occurring" as used herein refers to the fact that an object can be found in nature. For example, a peptide or nucleic acid that is present in an organism (including viruses) and can be isolated from a source in nature and which has not been intentionally modified by man in the laboratory is naturally occurring. The term "found in nature" means "present in nature" and includes known objects as well as objects that have not yet been discovered and / or isolated from nature, but that may be discovered and / or isolated in the future from a natural source. A naturally occurring MITD sequence, variant, mutant or polymorphism includes such sequences as have naturally arisen and can be detected in human populations. In an embodiment, all of the MITD sequences used in the present invention are naturally occurring polymorphisms, such as polymorphisms occurring naturally in humans.
[0233] The terms "room temperature" and "ambient temperature" are used interchangeably herein and refer to temperatures from at least about 15°C, e.g., from about 15°C to about 35°C, from about 15°C to about 30°C, from about 15°C to about 25°C, or from about 17°C to about 22°C. Such temperatures will include 15°C, 16°C, 17°C, 18°C, 19°C, 20°C, 21 °C and 22°C. In some embodiments, the temperature is from 15°C to about 25°C. In some embodiments, the temperature is from 17°C to about 25°C. In some embodiments, the temperature is about 15°C. In some embodiments, the temperature is about 16°C. In some embodiments, the temperature is about 17°C. In some embodiments, the temperature is about 18°C. In some embodiments, the temperature is about 19°C. In some embodiments, the temperature is about 20°C. In some embodiments, the temperature is about 21 °C. In some embodiments, the temperature is about 22°C.
[0234] The term “EDTA” refers to ethylenediaminetetraacetic acid disodium salt. All concentrations are given with respect to the EDTA disodium salt.
[0235] The term "cryoprotectant" relates to a substance that is added to a formulation in order to protect the active ingredients during the freezing stages.
[0236] The term "lyoprotectant" relates to a substance that is added to a formulation in order to protect the active ingredients during the drying stages.
[0237] According to the present disclosure, the term "peptide" refers to substances which comprise about two or more, about 3 or more, about 4 or more, about 6 or more, about 8 or more, about 10 or more, about 13 or more, about 16 or more, about 20 or more, and up to about 50, about 100 or about 150, consecutive amino acids linked to one another via peptide bonds. The term "polypeptide" refers to large peptides, in particular peptides having at least about 151 amino acids. "Peptides" and "polypeptides" are both protein molecules.
[0238] The term "biological activity" means the response of a biological system to a molecule. Such biological systems may be, for example, a cell or an organism. In some embodiments, such response is therapeutically or pharmaceutically useful. In some embodiments, a biological activity comprises a pharmaceutical activity.
[0239] The term "biological system", as used herein, refers to any system of interacting or potentially interacting biological constituents whose behavior can be characterized in whole or part by one or more biological processes or mechanisms. A biological system can include, for example, an individual cell, a collection of cells such as a cell culture, an organ, a tissue, and a multi-cellular organism such as an individual or subject, e.g., a human patient.
[0240] In some embodiments, a biological system is present in or is an individual or subject and a biological activity in such biological system is an activity which is therapeutically or pharmaceutically useful, i.e., the biological activity results in or contributes to a therapeutically or pharmaceutically useful effect.
[0241] According to various embodiments of the present disclosure, a nucleic acid (such as RNA and / or DNA) encoding a peptide or polypeptide is taken up by or introduced, i.e. transfected or transduced, into a cell which cell may be present in vitro or in a subject, resulting in expression of said peptide or polypeptide. The cell may, e.g., express the encoded peptide or polypeptide intracellularly (e.g. in the cytoplasm and / or in the nucleus), may secrete the encoded peptide or polypeptide, and / or may express it on the surface.
[0242] According to the present disclosure, terms such as "nucleic acid expressing" and "nucleic acid encoding" or similar terms are used interchangeably herein and with respect to a particular peptide or polypeptide mean that the nucleic acid, if present in the appropriate environment, e.g. within a cell, can be expressed to produce said peptide or polypeptide.
[0243] The term "portion" refers to a fraction. With respect to a particular structure such as an amino acid sequence or protein the term "portion" thereof may designate a continuous or a discontinuous fraction of said structure.
[0244] The terms "part" and "fragment" are used interchangeably herein and refer to a continuous element. For example, a part of a structure such as an amino acid sequence or protein refers to a continuous element of said structure. When used in context of a composition, the term "part" means a portion of the composition. For example, a part of a composition may be any portion from 0.1 % to 99.9% (such as 0.1 %, 0.5%, 1 %, 5%, 10%, 50%, 90%, or 99%) of said composition.
[0245] "Fragment", with reference to an amino acid sequence (peptide or polypeptide), relates to a part of an amino acid sequence, i.e. a sequence which represents the amino acid sequence shortened at the N-terminus and / or C-terminus. A fragment shortened at the C-terminus (N-terminal fragment) is obtainable, e.g., by translation of a truncated open reading frame that lacks the 3' -end of the open reading frame. A fragment shortened at the N-terminus (C-terminal fragment) is obtainable, e.g., by translation of a truncated open reading frame that lacks the 5' -end of the open reading frame, as long as the truncated open reading frame comprises a start codon that serves to initiate translation. A fragment of an amino acid sequence comprises, e.g., at least 50 %, at least 60 %, at least 70 %, at least 80%, at least 90% of the amino acid residues from an amino acid sequence. A fragment of an amino acid sequence comprises, e.g., at least 6, in particular at least 8, at least 10, at least 12, at least 15, at least 20, at least 30, at least 50, or at least 100 consecutive amino acids from an amino acid sequence. A fragment of an amino acid sequence comprises, e.g., a sequence of up to 8, in particular up to 10, up to 12, up to 15, up to 20, up to 30 or up to 55, consecutive amino acids of the amino acid sequence.
[0246] "Variant," as used herein and with reference to an amino acid sequence (peptide or polypeptide), is meant an amino acid sequence that differs from a parent amino acid sequence by virtue of at least one amino acid (e.g., a different amino acid, or a modification of the same amino acid). The parent amino acid sequence may be a naturally occurring or wild type (WT) amino acid sequence, or may be a modified version of a wild type amino acid sequence. In some embodiments, the variant amino acid sequence has at least one amino acid difference as compared to the parent amino acid sequence, e.g., from 1 to about 20 amino acid differences, such as from 1 to about 10 or from 1 to about 5 amino acid differences compared to the parent.
[0247] By "wild type" or "WT" or "native" herein is meant an amino acid sequence that is found in nature, including allelic variations. A wild type amino acid sequence, peptide or polypeptide has an amino acid sequence that has not been intentionally modified.
[0248] For the purposes of the present disclosure, "variants" of an amino acid sequence (peptide or polypeptide) may comprise amino acid insertion variants, amino acid addition variants, amino acid deletion variants and / or amino acid substitution variants. The term "variant" includes all mutants, splice variants, post-translationally modified variants, conformations, isoforms, allelic variants, species variants, and species homologs, in particular those which are naturally occurring. The term "variant" includes, in particular, fragments of an amino acid sequence.
[0249] Amino acid insertion variants comprise insertions of single or two or more amino acids in a particular amino acid sequence. In the case of amino acid sequence variants having an insertion, one or more amino acid residues are inserted into a particular site in an amino acid sequence, although random insertion with appropriate screening of the resulting product is also possible. Amino acid addition variants comprise amino- and / or carboxy-terminal fusions of one or more amino acids, such as 1 , 2, 3, 5, 10, 20, 30, 50, or more amino acids. Amino acid deletion variants are characterized by the removal of one or more amino acids from the sequence, such as by removal of 1 , 2, 3, 5, 10, 20, 30, 50, or more amino acids. The deletions may be in any position of the protein. Amino acid deletion variants that comprise the deletion at the N-terminal and / or C-terminal end of the protein are also called N-terminal and / or C-terminal truncation variants. Amino acid substitution variants are characterized by at least one residue in the sequence being removed and another residue being inserted in its place. Preference is given to the modifications being in positions in the amino acid sequence which are not conserved between homologous peptides or polypeptides and / or to replacing amino acids with other ones having similar properties. In some embodiments, amino acid changes in peptide and polypeptide variants are conservative amino acid changes, i.e., substitutions of similarly charged or uncharged amino acids. A conservative amino acid change involves substitution of one of a family of amino acids which are related in their side chains. Naturally occurring amino acids are generally divided into four families: acidic (aspartate, glutamate), basic (lysine, arginine, histidine), non-polar (alanine, valine, leucine, isoleucine, proline, phenylalanine, methionine, tryptophan), and uncharged polar (glycine, asparagine, glutamine, cysteine, serine, threonine, tyrosine) amino acids. Phenylalanine, tryptophan, and tyrosine are sometimes classified jointly as aromatic amino acids. In some embodiments, conservative amino acid substitutions include substitutions within the following groups:
[0250] - glycine, alanine;
[0251] - valine, isoleucine, leucine;
[0252] - aspartic acid, glutamic acid;
[0253] - asparagine, glutamine;
[0254] - serine, threonine;
[0255] - lysine, arginine; and
[0256] - phenylalanine, tyrosine. In some embodiments as may generally be applicable to the normal practice of the present invention, the “further sequence” of the MITD sequence as defined herein does not comprise the amino acid residues cysteine or methionine. In some embodiments the degree of similarity, such as identity between a given amino acid sequence and an amino acid sequence which is a variant of said given amino acid sequence, will be at least about 60%, 70%, 80%, 81 %, 82%, 83%, 84%, 85%, 86%, 87%, 88%, 89%, 90%, 91 %, 92%, 93%, 94%, 95%, 96%, 97%, 98%, or 99%. In some embodiments, the degree of similarity or identity is given for an amino acid region which is at least about 10%, at least about 20%, at least about 30%, at least about 40%, at least about 50%, at least about 60%, at least about 70%, at least about 80%, at least about 90% or about 100% of the entire length of the reference amino acid sequence. For example, if the reference amino acid sequence consists of 200 amino acids, the degree of similarity or identity is given, e.g., for at least about 20, at least about 40, at least about 60, at least about 80, at least about 100, at least about 120, at least about 140, at least about 160, at least about 180, or about 200 amino acids, in some embodiments continuous amino acids. In some embodiments, the degree of similarity or identity is given for the entire length of the reference amino acid sequence. The alignment for determining sequence similarity, such as sequence identity, can be done with art known tools, such as using the best sequence alignment, for example, using Align, using standard settings, preferably EMBOSS::needle, Matrix: Blosum62, Gap Open 10.0, Gap Extend 0.5.
[0257] "Sequence similarity" indicates the percentage of amino acids that either are identical or that represent conservative amino acid substitutions. "Sequence identity" between two amino acid sequences indicates the percentage of amino acids that are identical between the sequences. "Sequence identity" between two nucleic acid sequences indicates the percentage of nucleotides that are identical between the sequences.
[0258] The terms "% identical" and "% identity" or similar terms are intended to refer, in particular, to the percentage of nucleotides or amino acids which are identical in an optimal alignment between the sequences to be compared. Said percentage is purely statistical, and the differences between the two sequences may be but are not necessarily randomly distributed over the entire length of the sequences to be compared. Comparisons of two sequences are usually carried out by comparing the sequences, after optimal alignment, with respect to a segment or "window of comparison", in order to identify local regions of corresponding sequences. The optimal alignment for a comparison may be carried out manually or with the aid of the local homology algorithm by Smith and Waterman, 1981 , Ads App. Math. 2, 482, with the aid of the local homology algorithm by Neddleman and Wunsch, 1970, J. Mol. Biol. 48, 443, with the aid of the similarity search algorithm by Pearson and Lipman, 1988, Proc. Natl Acad. Sci. USA 88, 2444, or with the aid of computer programs using said algorithms (GAP, BESTFIT, FASTA, BLAST P, BLAST N and TFASTA in Wisconsin Genetics Software Package, Genetics Computer Group, 575 Science Drive, Madison, Wis.). In some embodiments, percent identity of two sequences is determined using the BLASTN or BLASTP algorithm, as available on the United States National Center for Biotechnology Information (NCBI) website (e.g., at blast. ncbi.nlm.nih.gov / Blast.cgi?PAGE_TYPE=BlastSearch&BLAST_SPEC=blast2se q&LINK_LOC=align2seq). In some embodiments, the algorithm parameters used for BLASTN algorithm on the NCBI website include: (i) Expect Threshold set to 10; (ii) Word Size set to 28; (iii) Max matches in a query range set to 0; (iv) Match / Mismatch Scores set to 1 , -2; (v) Gap Costs set to Linear; and (vi) the filter for low complexity regions being used. In some embodiments, the algorithm parameters used for BLASTP algorithm on the NCBI website include: (i) Expect Threshold set to 10; (ii) Word Size set to 3; (iii) Max matches in a query range set to 0; (iv) Matrix set to BLOSUM62; (v) Gap Costs set to Existence: 11 Extension: 1 ; and (vi) conditional compositional score matrix adjustment.
[0259] Percentage identity is obtained by determining the number of identical positions at which the sequences to be compared correspond, dividing this number by the number of positions compared (e.g., the number of positions in the reference sequence) and multiplying this result by 100.
[0260] In some embodiments, the degree of similarity or identity is given for a region which is at least about 50%, at least about 60%, at least about 70%, at least about 80%, at least about 90% or about 100% of the entire length of the reference sequence. For example, if the reference nucleic acid sequence consists of 200 nucleotides, the degree of identity is given for at least about 100, at least about 120, at least about 140, at least about 160, at least about 180, or about 200 nucleotides, in some embodiments continuous nucleotides. In some embodiments, the degree of similarity or identity is given for the entire length of the reference sequence.
[0261] Homologous amino acid sequences exhibit according to the disclosure at least 40%, in particular at least 50%, at least 60%, at least 70%, at least 80%, at least 90% and, e.g., at least 95%, at least 98 or at least 99% identity of the amino acid residues.
[0262] The amino acid sequence variants described herein may readily be prepared by the skilled person, for example, by recombinant DNA manipulation. The manipulation of DNA sequences for preparing peptides or polypeptides having substitutions, additions, insertions or deletions, is described in detail in Molecular Cloning: A Laboratory Manual, 4thEdition, M.R. Green and J. Sambrook eds., Cold Spring Harbor Laboratory Press, Cold Spring Harbor 2012, for example. Furthermore, the peptides, polypeptides and amino acid variants described herein may be readily prepared with the aid of known peptide synthesis techniques such as, for example, by solid phase synthesis and similar methods.
[0263] In some embodiments, a fragment or variant of an amino acid sequence (peptide or polypeptide) is a "functional fragment" or "functional variant". The term "functional fragment" or "functional variant" of an amino acid sequence relates to any fragment or variant exhibiting one or more functional properties identical or similar to those of the amino acid sequence from which it is derived, i.e., it is functionally equivalent. With respect to antigens or antigenic sequences, one particular function is one or more immunogenic activities displayed by the amino acid sequence from which the fragment or variant is derived. The term "functional fragment" or "functional variant", as used herein, in particular refers to a variant molecule or sequence that comprises an amino acid sequence that is altered by one or more amino acids compared to the amino acid sequence of the parent molecule or sequence and that is still capable of fulfilling one or more of the functions of the parent molecule or sequence, e.g., inducing an immune response. In some embodiments, the modifications in the amino acid sequence of the parent molecule or sequence do not significantly affect or alter the characteristics of the molecule or sequence. In different embodiments, the function of the functional fragment or functional variant may be reduced but still significantly present, e.g., function of the functional fragment or functional variant may be at least 50%, at least 60%, at least 70%, at least 80%, or at least 90% of the parent molecule or sequence. However, in other embodiments, function of the functional fragment or functional variant may be enhanced compared to the parent molecule or sequence.
[0264] An amino acid sequence (peptide or polypeptide) "derived from" a designated amino acid sequence (peptide or polypeptide) refers to the origin of the first amino acid sequence. In some embodiments, the amino acid sequence which is derived from a particular amino acid sequence has an amino acid sequence that is identical, essentially identical or homologous to that particular sequence or a fragment thereof. Amino acid sequences derived from a particular amino acid sequence may be variants of that particular sequence or a fragment thereof. For example, it will be understood by one of ordinary skill in the art that the antigens suitable for use herein may be altered such that they vary in sequence from the naturally occurring or native sequences from which they were derived, while retaining the desirable activity of the native sequences. In some embodiments, "isolated" means removed (e.g., purified) from the natural state or from an artificial composition, such as a composition from a production process. For example, a nucleic acid, peptide or polypeptide naturally present in a living animal is not "isolated", but the same nucleic acid, peptide or polypeptide partially or completely separated from the coexisting materials of its natural state is "isolated". An isolated nucleic acid, peptide or polypeptide can exist in substantially purified form, or can exist in a non-native environment such as, for example, a host cell.
[0265] The term "transfection" relates to the introduction of nucleic acids, in particular RNA, into a cell. For purposes of the present disclosure, the term "transfection" also includes the introduction of a nucleic acid into a cell or the uptake of a nucleic acid by such cell, wherein the cell may be present in a subject, e.g., a patient, or the cell may be in vitro, e.g., outside of a patient. Thus, according to the present disclosure, a cell for transfection of a nucleic acid described herein can be present in vitro or in vivo, e.g. the cell can form part of an organ, a tissue and / or the body of a patient. According to the disclosure, transfection can be transient or stable. For some applications of transfection, it is sufficient if the transfected genetic material is only transiently expressed. RNA can be transfected into cells to transiently express its coded protein. Since the nucleic acid introduced in the transfection process is usually not integrated into the nuclear genome, the foreign nucleic acid will be diluted through mitosis or degraded. Cells allowing episomal amplification of nucleic acids greatly reduce the rate of dilution. If it is desired that the transfected nucleic acid actually remains in the genome of the cell and its daughter cells, a stable transfection must occur. Such stable transfection can be achieved by using virus-based systems or transposon-based systems for transfection, for example. Generally, nucleic acid encoding antigen is transiently transfected into cells. RNA can be transfected into cells to transiently express its coded protein.
[0266] Cells which are useful for transfection in the methods described herein include, but are not limited to, cells from an animal cell line, such as Chinese hamster ovary (CHO), K562, HepG2, HEK293T, RAW, and C2C12 cells. In some embodiments, the cells are CHO, K562, HEK293T, RAW, and C2C12 cells. In some embodiments, the cells are Chinese hamster ovary (CHO) cells.
[0267] The disclosure includes analogs of a peptide or polypeptide. According to the present disclosure, an analog of a peptide or polypeptide is a modified form of said peptide or polypeptide from which it has been derived and has at least one functional property of said peptide or polypeptide. E.g., a pharmacological active analog of a peptide or polypeptide has at least one of the pharmacological activities of the peptide or polypeptide from which the analog has been derived. Such modifications include any chemical modification and comprise single or multiple substitutions, deletions and / or additions of any molecules associated with the peptide or polypeptide, such as carbohydrates, lipids and / or peptides or polypeptides. In some embodiments, "analogs" of peptides or polypeptides include those modified forms resulting from glycosylation, acetylation, phosphorylation, amidation, palmitoylation, myristoylation, isoprenylation, lipidation, alkylation, derivatization, introduction of protective / blocking groups, proteolytic cleavage or binding to an antibody or to another cellular ligand. The term "analog" also extends to all functional chemical equivalents of said peptides and polypeptides.
[0268] As used herein, the terms "linked", "fused", or "fusion" are used interchangeably. These terms refer to the joining together of two or more elements or components or domains.
[0269] As used herein "endogenous" refers to any material from or produced inside an organism, cell, tissue or system.
[0270] As used herein, the term "exogenous" refers to any material introduced from or produced outside an organism, cell, tissue or system.
[0271] The term "expression" as used herein is defined as the translation but may also include the transcription of a particular nucleotide sequence.
[0272] In the context of the present disclosure, the term "transcription" relates to a process, wherein the genetic code in a DNA sequence is transcribed into RNA (especially mRNA). Subsequently, the RNA may be translated into peptide or polypeptide.
[0273] With respect to RNA, the term "expression" or "translation" relates to the process in the ribosomes of a cell by which a strand of mRNA directs the assembly of a sequence of amino acids to make a peptide or polypeptide.
[0274] Prodrugs of a particular compound described herein are those compounds that upon administration to an individual undergo chemical conversion under physiological conditions to provide the particular compound. Additionally, prodrugs can be converted to the particular compound by chemical or biochemical methods in an ex vivo environment. For example, prodrugs can be slowly converted to the particular compound when, for example, placed in a transdermal patch reservoir with a suitable enzyme or chemical reagent. Exemplary prodrugs are esters (using an alcohol or a carboxy group contained in the particular compound) or amides (using an amino or a carboxy group contained in the particular compound) which are hydrolyzable in vivo. Specifically, any amino group which is contained in the particular compound and which bears at least one hydrogen atom can be converted into a prodrug form. Typical N- prodrug forms include carbamates, Mannich bases, enamines, and enaminones.
[0275] In the present specification, a structural formula of a compound may represent a certain isomer of said compound. It is to be understood, however, that the present invention includes all isomers such as geometrical isomers, optical isomers based on an asymmetrical carbon, stereoisomers, tautomers and the like which occur structurally and isomer mixtures and is not limited to the description of the formula.
[0276] "Isomers" are compounds having the same molecular formula but differ in structure ("structural isomers") or in the geometrical (spatial) positioning of the functional groups and / or atoms ("stereoisomers"). "Enantiomers" are a pair of stereoisomers which are non-superimposable mirror-images of each other. A "racemic mixture" or "racemate" contains a pair of enantiomers in equal amounts and is denoted by the prefix (±). "Diastereomers" are stereoisomers which are non-superimposable and which are not mirror-images of each other. "Tautomers" are structural isomers of the same chemical substance that spontaneously and reversibly interconvert into each other, even when pure, due to the migration of individual atoms or groups of atoms; i.e., the tautomers are in a dynamic chemical equilibrium with each other. An example of tautomers are the isomers of the keto-enol-tautomerism. "Conformers" are stereoisomers that can be interconverted just by rotations about formally single bonds, and include - in particular - those leading to different 3-dimentional forms of (hetero)cyclic rings, such as chair, half-chair, boat, and twist-boat forms of cyclohexane.
[0277] The term "average diameter" refers to the mean hydrodynamic diameter of particles as measured by dynamic light scattering (DLS) with data analysis using the so-called cumulant algorithm, which provides as results the so-called Zaverage with the dimension of a length, and the polydispersity index (PDI), which is dimensionless (Koppel, D., J. Chem. Phys. 57, 1972, pp 4814-4820, ISO 13321 ). Here "average diameter", "diameter" or "size" for particles is used synonymously with this value of the Zaverage.
[0278] In some embodiments, the "polydispersity index" is may be calculated based on dynamic light scattering measurements by the so-called cumulant analysis as mentioned in the definition of the "average diameter". Under certain prerequisites, it can be taken as a measure of the size distribution of an ensemble of nanoparticles. The "radius of gyration" (abbreviated herein as Rg) of a particle about an axis of rotation is the radial distance of a point from the axis of rotation at which, if the whole mass of the particle is assumed to be concentrated, its moment of inertia about the given axis would be the same as with its actual distribution of mass. Mathematically, Rgis the root mean square distance of the particle's components from either its center of mass or a given axis. For example, for a macromolecule composed of n mass elements, of masses ny ( / = 1 , 2, 3, ... , n), located at fixed distances s / from the center of mass, Rgis the square-root of the mass average of Si2over all mass elements and can be calculated as follows:
[0279] The radius of gyration can be determined or calculated experimentally, e.g., by using light scattering. In particular, for small scattering vectors q the structure function S is defined as follows: wherein N is the number of components (Guinier's law).
[0280] The "hydrodynamic radius" (which is sometimes called "Stokes radius" or "Stokes- Einstein radius") of a particle is the radius of a hypothetical hard sphere that diffuses at the same rate as said particle. The hydrodynamic radius is related to the mobility of the particle, taking into account not only size but also solvent effects. For example, a smaller charged particle with stronger hydration may have a greater hydrodynamic radius than a larger charged particle with weaker hydration. This is because the smaller particle drags a greater number of water molecules with it as it moves through the solution. Since the actual dimensions of the particle in a solvent are not directly measurable, the hydrodynamic radius may be defined by the Stokes-Einstein equation: kB- T
[0281] Rh~ 6 - n - q - D wherein ks is the Boltzmann constant; T is the temperature; is the viscosity of the solvent; and D is the diffusion coefficient. The diffusion coefficient can be determined experimentally, e.g., by using dynamic light scattering (DLS). Thus, one procedure to determine the hydrodynamic radius of a particle or a population of particles (such as the hydrodynamic radius of particles contained in a sample or control composition as disclosed herein or the hydrodynamic radius of a particle peak obtained from subjecting such a sample or control composition to field-flow fractionation) is to measure the DLS signal of said particle or population of particles (such as DLS signal of particles contained in a sample or control composition as disclosed herein or the DLS signal of a particle peak obtained from subjecting such a sample or control composition to field-flow fractionation).
[0282] The expression "light scattering" as used herein refers to the physical process where light is forced to deviate from a straight trajectory by one or more paths due to localized non-uniformities in the medium through which the light passes.
[0283] The term "UV" means ultraviolet and designates a band of the electromagnetic spectrum with a wavelength from 10 nm to 400 nm, i.e., shorter than that of visible light but longer than X-rays.
[0284] The expression "multi-angle light scattering" or "MALS" as used herein relates to a technique for measuring the light scattered by a sample into a plurality of angles. "Multiangle" means in this respect that scattered light can be detected at different discrete angles as measured, for example, by a single detector moved over a range including the specific angles selected or an array of detectors fixed at specific angular locations. In certain embodiments, the light source used in MALS is a laser source (MALLS: multiangle laser light scattering). Based on the MALS signal of a composition comprising particles and by using an appropriate formalism (e.g., Zimm plot, Berry plot, or Debye plot), it is possible to determine the radius of gyration (Rg) and, thus, the size of said particles. Preferably, the Zimm plot is a graphical presentation using the following equation: wherein c is the mass concentration of the particles in the solvent (g / mL); A2 is the second virial coefficient (mol mL / g2); P(6) is a form factor relating to the dependence of scattered light intensity on angle; Re is the excess Rayleigh ratio (cm-1); and K* is an optical constant that is equal to 4rr2r|o (dn / dc)2Ao’4 / \ / A’1, where q0is the refractive index of the solvent at the incident radiation (vacuum) wavelength, Ao is the incident radiation (vacuum) wavelength (nm), A / A is Avogadro’s number (mol-1), and dn / dc is the differential refractive index increment (mL / g) (cf., e.g., Buchholz et al. (Electrophoresis 22 (2001 ), 4118-4128); B.H. Zimm (J. Chem. Phys. 13 (1945), 141 ; P. Debye (J. Appl. Phys. 15 (1944): 338; and W. Burchard (Anal. Chem. 75 (2003), 4279-4291 ). Preferably, the Berry plot is calculated the following term: [R? K*Cwherein c, Re and K* are as defined above. Preferably, the Debye plot is calculated the following term:
[0285] K*c
[0286] ~R wherein c, Re and / are as defined above.
[0287] The expression "dynamic light scattering" or "DLS" as used herein refers to a technique to determine the size and size distribution profile of particles, in particular with respect to the hydrodynamic radius of the particles. A monochromatic light source, usually a laser, is shot through a polarizer and into a sample. The scattered light then goes through a second polarizer where it is detected and the resulting image is projected onto a screen. The particles in the solution are being hit with the light and diffract the light in all directions. The diffracted light from the particles can either interfere constructively (light regions) or destructively (dark regions). This process is repeated at short time intervals and the resulting set of speckle patterns are analyzed by an autocorrelator that compares the intensity of light at each spot over time.
[0288] The expression "static light scattering" or "SLS" as used herein refers to a technique to determine the size and size distribution profile of particles, in particular with respect to the radius of gyration of the particles, and / or the molar mass of particles. A high- intensity monochromatic light, usually a laser, is launched in a solution containing the particles. One or many detectors are used to measure the scattering intensity at one or many angles. The angular dependence is needed to obtain accurate measurements of both molar mass and size for all macromolecules of radius. Hence simultaneous measurements at several angles relative to the direction of incident light, known as multi-angle light scattering (MALS) or multi-angle laser light scattering (MALLS), is generally regarded as the standard implementation of static light scattering.
[0289] Nucleic acids
[0290] The term "nucleic acid" comprises deoxyribonucleic acid (DNA), ribonucleic acid (RNA), combinations thereof, and modified forms thereof. The term comprises genomic DNA, cDNA, mRNA, recombinantly produced and chemically synthesized molecules. In some embodiments, a nucleic acid is DNA. In some embodiments, a nucleic acid is RNA. In some embodiments, a nucleic acid is a mixture of DNA and RNA. In some embodiments, a nucleic acid is DNA. A nucleic acid may be present as a single-stranded or double-stranded and linear or covalently circularly closed molecule. A nucleic acid can be isolated. The term "isolated nucleic acid" means, according to the present disclosure, that the nucleic acid (i) was amplified in vitro, for example via polymerase chain reaction (PCR) for DNA or in vitro transcription (using, e.g., an RNA polymerase) for RNA, (ii) was produced recombinantly by cloning, (iii) was purified, for example, by cleavage and separation by gel electrophoresis, or (iv) was synthesized, for example, by chemical synthesis.
[0291] The term "nucleoside" (abbreviated herein as "N") relates to compounds which can be thought of as nucleotides without a phosphate group. While a nucleoside is a nucleobase linked to a sugar (e.g., ribose or deoxyribose), a nucleotide is composed of a nucleoside and one or more phosphate groups. Examples of nucleosides include cytidine, undine, pseudouridine, adenosine, and guanosine.
[0292] The five standard nucleosides which usually make up naturally occurring nucleic acids are uridine, adenosine, thymidine, cytidine and guanosine. The five nucleosides are commonly abbreviated to their one letter codes U, A, T, C and G, respectively. However, thymidine is more commonly written as "dT" ("d" represents "deoxy") as it contains a 2'-deoxyribofuranose moiety rather than the ribofuranose ring found in uridine. This is because thymidine is found in deoxyribonucleic acid (DNA) and not ribonucleic acid (RNA). Conversely, uridine is found in RNA and not DNA. The remaining three nucleosides may be found in both RNA and DNA. In RNA, they would be represented as A, C and G, whereas in DNA they would be represented as dA, dC and dG.
[0293] A modified purine (A or G) or pyrimidine (C, T, or U) base moiety is preferably modified by one or more alkyl groups, more preferably one or more C1-4 alkyl groups, even more preferably one or more methyl groups. Particular examples of modified purine or pyrimidine base moieties include N7-alkyl-guanine, N6-alkyl-adenine, 5-alkyl-cytosine, 5-alkyl-uracil, and N(1 )-alkyl-uracil, such as N7-CI-4 alkyl-guanine, N6-CI-4 alkyladenine, 5-CI-4 alkyl-cytosine, 5-CI-4 alkyl-uracil, and N(1 )-CI-4 alkyl-uracil, preferably N7-methyl-guanine, N6-methyl-adenine, 5-methyl-cytosine, 5-methyl-uracil, and N(1 )- methyl-uracil.
[0294] Herein, the term "DNA" relates to a nucleic acid molecule which includes deoxyribonucleotide residues. In preferred embodiments, the DNA contains all or a majority of deoxyribonucleotide residues. As used herein, "deoxyribonucleotide" refers to a nucleotide which lacks a hydroxyl group at the 2'-position of a [3-D-ribofuranosyl group. DNA encompasses without limitation, double stranded DNA, single stranded DNA, isolated DNA such as partially purified DNA, essentially pure DNA, synthetic DNA, recombinantly produced DNA, as well as modified DNA that differs from naturally occurring DNA by the addition, deletion, substitution and / or alteration of one or more nucleotides. Such alterations may refer to addition of non-nucleotide material to internal DNA nucleotides or to the end(s) of DNA. It is also contemplated herein that nucleotides in DNA may be non-standard nucleotides, such as chemically synthesized nucleotides or ribonucleotides. For the present disclosure, these altered DNAs are considered analogs of naturally-occurring DNA. A molecule contains "a majority of deoxyribonucleotide residues" if the content of deoxyribonucleotide residues in the molecule is more than 50% (such as at least 55%, at least 60%, at least 65%, at least 70%, at least 75%, at least 80%, at least 85%, at least 90%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99%), based on the total number of nucleotide residues in the molecule. The total number of nucleotide residues in a molecule is the sum of all nucleotide residues (irrespective of whether the nucleotide residues are standard ( / .e., naturally occurring) nucleotide residues or analogs thereof).
[0295] DNA may be recombinant DNA and may be obtained by cloning of a nucleic acid, in particular cDNA. The cDNA may be obtained by reverse transcription of RNA.
[0296] The term "RNA" relates to a nucleic acid molecule which includes ribonucleotide residues. In preferred embodiments, the RNA contains all or a majority of ribonucleotide residues. As used herein, "ribonucleotide" refers to a nucleotide with a hydroxyl group at the 2'-position of a [3-D-ribofuranosyl group. RNA encompasses without limitation, double stranded RNA, single stranded RNA, isolated RNA such as partially purified RNA, essentially pure RNA, synthetic RNA, recombinantly produced RNA, as well as modified RNA that differs from naturally occurring RNA by the addition, deletion, substitution and / or alteration of one or more nucleotides. Such alterations may refer to addition of non-nucleotide material to internal RNA nucleotides or to the end(s) of RNA. It is also contemplated herein that nucleotides in RNA may be nonstandard nucleotides, such as chemically synthesized nucleotides or deoxynucleotides. For the present disclosure, these altered / modified nucleotides can be referred to as analogs of naturally occurring nucleotides, and the corresponding RNAs containing such altered / modified nucleotides ( / .e., altered / modified RNAs) can be referred to as analogs of naturally occurring RNAs. A molecule contains "a majority of ribonucleotide residues" if the content of ribonucleotide residues in the molecule is more than 50% (such as at least 55%, at least 60%, at least 65%, at least 70%, at least 75%, at least 80%, at least 85%, at least 90%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99%), based on the total number of nucleotide residues in the molecule. The total number of nucleotide residues in a molecule is the sum of all nucleotide residues (irrespective of whether the nucleotide residues are standard ( / .e., naturally occurring) nucleotide residues or analogs thereof).
[0297] "RNA" includes mRNA, tRNA, ribosomal RNA (rRNA), small nuclear RNA (snRNA), self-amplifying RNA (saRNA), single-stranded RNA (ssRNA), dsRNA, inhibitory RNA (such as antisense ssRNA, small interfering RNA (siRNA), or microRNA (miRNA)), activating RNA (such as small activating RNA) and immunostimulatory RNA (isRNA). In some embodiments, "RNA" refers to mRNA.
[0298] The term "in vitro transcription" or "IVT" as used herein means that the transcription (i.e., the generation of RNA) is conducted in a cell-free manner. I.e., IVT does not use living / cultured cells but rather the transcription machinery extracted from cells (e.g., cell lysates or the isolated components thereof, including an RNA polymerase (preferably T7, T3 or SP6 polymerase)).
[0299] In some embodiments, the nucleic acids of the present invention, such as one, at least two or all of the nucleic acids of the present invention, are RNA.
[0300] In some embodiments, the RNA is single stranded RNA.
[0301] In some embodiments, the RNA is mRNA.
[0302] In some embodiments, the RNA is generated by RNA in vitro transcription.
[0303] In some embodiments, the RNA comprises a 5' cap structure.
[0304] In some embodiments, the RNA does not comprise modified ribonucleotides.
[0305] In some embodiments, the RNA comprises modified ribonucleotides. In some embodiments, the modified ribonucleotides comprise modified uridines. In some embodiments, the modified uridines comprise N1 -methyl-pseudouridine.
[0306] In some embodiments, the nucleic acids of the present invention, such as one, at least two or all of the nucleic acids of the present invention, are DNA.
[0307] In some embodiments, the DNA is present in the form of a vector.
[0308] In some embodiments, the vector comprises DNA encoding an amino acid sequence comprising the amino acid sequence of a peptide or polypeptide having biological activity.
[0309] In some embodiments, the vector is a DNA vector. In some embodiments, the nucleic acids of the present invention, such as one, at least two or all of the nucleic acids of the present invention comprise a mixture of RNA and DNA.
[0310] In some embodiments, the RNA in the mixture is single stranded RNA.
[0311] In some embodiments, the RNA in the mixture is mRNA.
[0312] In some embodiments, the RNA in the mixture is generated by RNA in vitro transcription.
[0313] In some embodiments, the RNA in the mixture comprises a 5' cap structure.
[0314] In some embodiments, the RNA in the mixture does not comprise modified ribonucleotides.
[0315] In some embodiments, the RNA in the mixture comprises modified ribonucleotides. In some embodiments, the modified ribonucleotides comprise modified uridines. In some embodiments, the modified uridines comprise N1 -methyl-pseudouridine.
[0316] In some embodiments, the DNA in the mixture is present in the form of a vector.
[0317] In some embodiments, the vector in the mixture comprises DNA encoding an amino acid sequence comprising the amino acid sequence of a peptide or polypeptide having biological activity.
[0318] In some embodiments, the vector in the mixture is a DNA vector.
[0319] In some embodiments, the nucleic acid (such as RNA and / or DNA) of the present invention, which can comprise one or at least two or more nucleic acid constructs, is formulated with a delivery vehicle.
[0320] In some embodiments, the nucleic acid (such as RNA and / or DNA) is formulated with one or more compounds complexing the nucleic acid (such as RNA and / or DNA).
[0321] In some embodiments, the nucleic acid (such as RNA and / or DNA) is formulated as particles.
[0322] In some embodiments, the nucleic acid (such as RNA and / or DNA) is formulated as lipoplex particles. In these embodiments, it is preferred that the cells are characterized by a macropinocytosis-mediated RNA uptake mechanism.
[0323] In some embodiments, the nucleic acid (such as RNA and / or DNA) is formulated as lipid nanoparticles.
[0324] In some embodiments, the nucleic acid (such as RNA and / or DNA) comprises a mixture of different nucleic acids (such as RNAs and / or DNAs, e.g., two or more RNAs, two or more DNAs, or one or more RNAs and one or more DNAs), wherein each nucleic acid (such as RNA and / or DNA) encodes an amino acid sequence comprising the amino acid sequence of a peptide or polypeptide having biological activity.
[0325] In some embodiments, the mixture of different nucleic acids (such as RNAs and / or DNAs, e.g., two or more RNAs, two or more DNAs, or one or more RNAs and one or more DNAs) comprises nucleic acids (such as RNAs and / or DNAs, e.g., two or more RNAs, two or more DNAs, or one or more RNAs and one or more DNAs) encoding different amino acid sequences comprising the amino acid sequence of a peptide or polypeptide having biological activity.
[0326] In some embodiments, the different amino acid sequences comprise the amino acid sequence of different peptides or polypeptides having biological activity.
[0327] In some embodiments, the different peptides or polypeptides having biological activity comprise different antigens.
[0328] In some embodiments, the nucleic acid (such as RNA and / or DNA) comprises a mixture of different nucleic acids (such as RNAs and / or DNAs, e.g., two or more RNAs, two or more DNAs, or one or more RNAs and one or more DNAs) encoding amino acid sequences comprising the amino acid sequence of different antigens.
[0329] In some embodiments, the RNA described herein is single-stranded RNA that may be translated into the respective protein upon entering cells, e.g., cells used in the assays described herein and cells of a recipient. In addition to wildtype or codon-optimized sequences encoding the amino acid sequence comprising the amino acid sequence of a peptide or polypeptide having biological activity, e.g., a pharmaceutically active peptide or polypeptide such as antigen sequence, the RNA may contain one or more structural elements optimized for maximal efficacy of the RNA with respect to stability and translational efficiency (5' cap, 5' UTR, 3' UTR, poly(A)-tail). In one embodiment, the RNA contains all of these elements. In one embodiment, beta-S-ARCA(DI ) (m27’2' °GppSpG) or m27’3’’0Gppp(mi2’’°)ApG may be utilized as specific capping structure at the 5'-end of the RNA drug substances. As 5'-UTR sequence, the 5'-UTR sequence of the human alpha-globin mRNA, optionally with an optimized ‘Kozak sequence’ to increase translational efficiency may be used. As 3'-UTR sequence, a combination of two sequence elements (Fl element) derived from the "amino terminal enhancer of split" (AES) mRNA (called F) and the mitochondrial encoded 12S ribosomal RNA (called I) placed between the coding sequence and the poly(A)-tail to assure higher maximum protein levels and prolonged persistence of the mRNA may be used. These were identified by an ex vivo selection process for sequences that confer RNA stability and augment total protein expression (see WO 2017 / 060314, herein incorporated by reference). Alternatively, the 3‘-UTR may be two re-iterated 3'-UTRs of the human beta-globin mRNA. Furthermore, a poly(A)-tail measuring 110 nucleotides in length, consisting of a stretch of 30 adenosine residues, followed by a 10 nucleotide linker sequence (of random nucleotides) and another 70 adenosine residues may be used. This poly(A)-tail sequence was designed to enhance RNA stability and translational efficiency.
[0330] The amino acid sequence comprising the amino acid sequence of a peptide or polypeptide having biological activity, e.g., a pharmaceutically active peptide or polypeptide such as antigen sequence, may comprise amino acid sequences other than the amino acid sequence of a peptide or polypeptide having biological activity. Such other amino acid sequences may support the function or activity of the peptide or polypeptide having biological activity. In some embodiments, such other amino acid sequences comprise an amino acid sequence enhancing antigen processing and / or presentation. Alternatively, or additionally, such other amino acid sequences comprise an amino acid sequence which breaks immunological tolerance. Alternatively, or additionally, such other amino acid sequences comprise an amino acid sequence which produces bioluminescence. Such other amino acid sequences may be useful for determining the amount of the amino acid sequence comprising the amino acid sequence of a peptide or polypeptide having biological activity or a fragment thereof in the assays described herein. In particular, such other amino acid sequences may be useful for quantification by LC-MS / MS analysis.
[0331] The nucleic acids (such as RNA and / or DNA) described herein may be complexed with polymers, proteins and / or lipids, preferably lipids, to generate nucleic acid-particles for administration. If a combination of different nucleic acids is used, the nucleic acids may be complexed together or complexed separately. mRNA
[0332] According to the present disclosure, the term "mRNA" means "messenger-RNA" and relates to a "transcript" which may be generated by using a DNA template and may encode a peptide or polypeptide. Typically, an mRNA comprises a 5'-UTR, a peptide / polypeptide coding region, and a 3'-UTR. In the context of the present disclosure, mRNA may be generated by in vitro transcription (IVT) from a DNA template. As set forth above, the in vitro transcription methodology is known to the skilled person, and a variety of in vitro transcription kits is commercially available. mRNA is single-stranded but may contain self-complementary sequences that allow parts of the mRNA to fold and pair with itself to form double helices.
[0333] According to the present disclosure, "dsRNA" means double-stranded RNA and is RNA with two partially or completely complementary strands.
[0334] In preferred embodiments of the present disclosure, the mRNA relates to an RNA transcript which encodes a peptide or polypeptide.
[0335] In some embodiments, the mRNA which preferably encodes a peptide or polypeptide has a length of at least 45 nucleotides (such as at least 60, at least 90, at least 100, at least 200, at least 300, at least 400, at least 500, at least 600, at least 700, at least 800, at least 900, at least 1 ,000, at least 1 ,500, at least 2,000, at least 2,500, at least 3,000, at least 3,500, at least 4,000, at least 4,500, at least 5,000, at least 6,000, at least 7,000, at least 8,000, at least 9,000 nucleotides), preferably up to 15,000, such as up to 14,000, up to 13,000, up to 12,000 nucleotides, up to 11 ,000 nucleotides or up to 10,000 nucleotides.
[0336] As established in the art, mRNA generally contains a 5' untranslated region (5'-UTR), a peptide / polypeptide coding region and a 3' untranslated region (3'-UTR). In some embodiments, the mRNA is produced by in vitro transcription or chemical synthesis. In some embodiments, the mRNA is produced by in vitro transcription using a DNA template. The in vitro transcription methodology is known to the skilled person; cf., e.g., Molecular Cloning: A Laboratory Manual, 4thEdition, M.R. Green and J. Sambrook eds., Cold Spring Harbor Laboratory Press, Cold Spring Harbor 2012. Furthermore, a variety of in vitro transcription kits is commercially available, e.g., from Thermo Fisher Scientific (such as TranscriptAid™ T7 kit, MEGAscript® T7 kit, MAXIscript®), New England BioLabs Inc. (such as HiScribe™ T7 kit, HiScribe™ T7 ARCA mRNA kit), Promega (such as RiboMAX™, HeLaScribe®, Riboprobe® systems), Jena Bioscience (such as SP6 or T7 transcription kits), and Epicentre (such as AmpliScribe™). For providing modified mRNA, correspondingly modified nucleotides, such as modified naturally occurring nucleotides, non-naturally occurring nucleotides and / or modified non-naturally occurring nucleotides, can be incorporated during synthesis (preferably in vitro transcription), or modifications can be effected in and / or added to the mRNA after transcription. In some embodiments, mRNA is in vitro transcribed mRNA (IVT-RNA) and may be obtained by in vitro transcription of an appropriate DNA template. The promoter for controlling transcription can be any promoter for any RNA polymerase. Particular examples of RNA polymerases are the T7, T3, and SP6 RNA polymerases. Preferably, the in vitro transcription is controlled by a T7 or SP6 promoter. A DNA template for in vitro transcription may be obtained by cloning of a nucleic acid, in particular cDNA, and introducing it into an appropriate vector for in vitro transcription. The cDNA may be obtained by reverse transcription of RNA.
[0337] In some embodiments of the present disclosure, the mRNA is "replicon mRNA" or simply a "replicon", in particular "self-replicating mRNA" or "self-amplifying mRNA". In certain embodiments, the replicon or self-replicating mRNA is derived from or comprises elements derived from an ssRNA virus, in particular a positive-stranded ssRNA virus such as an alphavirus. Alphaviruses are typical representatives of positive-stranded RNA viruses. Alphaviruses replicate in the cytoplasm of infected cells (for review of the alphaviral life cycle see Jose et al., Future Microbiol., 2009, vol. 4, pp. 837-856). The total genome length of many alphaviruses typically ranges between 11 ,000 and 12,000 nucleotides, and the genomic RNA typically has a 5’ -cap, and a 3’ poly(A) tail. The genome of alphaviruses encodes non-structural proteins (involved in transcription, modification and replication of viral RNA and in protein modification) and structural proteins (forming the virus particle). There are typically two open reading frames (ORFs) in the genome. The four non-structural proteins (nsP1-nsP4) are typically encoded together by a first ORF beginning near the 5' terminus of the genome, while alphavirus structural proteins are encoded together by a second ORF which is found downstream of the first ORF and extends near the 3’ terminus of the genome. Typically, the first ORF is larger than the second ORF, the ratio being roughly 2:1. In cells infected by an alphavirus, only the nucleic acid sequence encoding non-structural proteins is translated from the genomic RNA, while the genetic information encoding structural proteins is translatable from a subgenomic transcript, which is an RNA molecule that resembles eukaryotic messenger RNA (mRNA; Gould et al., 2010, Antiviral Res., vol. 87 pp. 111-124). Following infection, i.e. at early stages of the viral life cycle, the (+) stranded genomic RNA directly acts like a messenger RNA for the translation of the open reading frame encoding the non-structural poly-protein (nsP1234). Alphavirus-derived vectors have been proposed for delivery of foreign genetic information into target cells or target organisms. In simple approaches, the open reading frame encoding alphaviral structural proteins is replaced by an open reading frame encoding a protein of interest. Alphavirus-based trans-replication systems rely on alphavirus nucleotide sequence elements on two separate nucleic acid molecules: one nucleic acid molecule encodes a viral replicase, and the other nucleic acid molecule is capable of being replicated by said replicase in trans (hence the designation trans-replication system). Trans-replication requires the presence of both these nucleic acid molecules in a given host cell. The nucleic acid molecule capable of being replicated by the replicase in trans must comprise certain alphaviral sequence elements to allow recognition and RNA synthesis by the alphaviral replicase.
[0338] In some embodiments of the present disclosure, the mRNA contains one or more modifications, e.g., in order to increase its stability and / or increase translation efficiency and / or decrease immunogenicity and / or decrease cytotoxicity. For example, in order to increase expression of the mRNA, it may be modified within the coding region, i.e., the sequence encoding the expressed peptide or polypeptide, preferably without altering the sequence of the expressed peptide or polypeptide. Such modifications are described, for example, in WO 2007 / 036366 and PCT / EP2019 / 056502, and include the following: a 5'-cap structure; an extension or truncation of the naturally occurring poly(A) tail; an alteration of the 5'- and / or 3'- untranslated regions (UTR) such as introduction of a UTR which is not related to the coding region of said RNA; the replacement of one or more naturally occurring nucleotides with synthetic nucleotides; and codon optimization (e.g., to alter, preferably increase, the GC content of the RNA).
[0339] In some embodiments, the mRNA comprises a 5'-cap structure. In some embodiments, the mRNA does not have uncapped 5'-triphosphates. In some embodiments, the mRNA may comprise a conventional 5'-cap and / or a 5'-cap analog. The term "conventional 5'-cap" refers to a cap structure found on the 5'-end of an mRNA molecule and generally consists of a guanosine 5'-triphosphate (Gppp) which is connected via its triphosphate moiety to the 5'-end of the next nucleotide of the mRNA (i.e., the guanosine is connected via a 5' to 5' triphosphate linkage to the rest of the mRNA). The guanosine may be methylated at position N7(resulting in the cap structure m7Gppp). The term "5'-cap analog" includes a 5'-cap which is based on a conventional 5' -cap but which has been modified at either the 2'- or 3'-position of the m7guanosine structure in order to avoid an integration of the 5'-cap analog in the reverse orientation (such 5' -cap analogs are also called anti-reverse cap analogs (ARCAs)). Particularly preferred 5'-cap analogs are those having one or more substitutions at the bridging and non-bridging oxygen in the phosphate bridge, such as phosphorothioate modified 5' -cap analogs at the [3-phosphate (such as m27’2 OG(5')ppSp(5')G (referred to as beta- S-ARCA or [3-S-ARCA)), as described in PCT / EP2019 / 056502. Providing an mRNA with a 5' -cap structure as described herein may be achieved by in vitro transcription of a DNA template in presence of a corresponding 5' -cap compound, wherein said 5'-cap structure is co-transcriptionally incorporated into the generated mRNA strand, or the mRNA may be generated, for example, by in vitro transcription, and the 5' -cap structure may be attached to the mRNA post-transcriptionally using capping enzymes, for example, capping enzymes of vaccinia virus.
[0340] In some embodiments, the mRNA comprises a 5'-cap structure selected from the group consisting of m27’2 OG(5’)ppSp(5')G (in particular its D1 diastereomer), m27’3'°G(5')ppp(5')G, and m27’3’0Gppp(mi2'’°)ApG.
[0341] In some embodiments, the mRNA comprises a capO, cap1 , or cap2, preferably cap1 or cap2. According to the present disclosure, the term "capO" means the structure "m7GpppN", wherein N is any nucleoside bearing an OH moiety at position 2'. According to the present disclosure, the term "cap1" means the structure
[0342] "m7GpppNm", wherein Nm is any nucleoside bearing an OCH3 moiety at position 2'. According to the present disclosure, the term "cap2" means the structure
[0343] "m7GpppNmNm", wherein each Nm is independently any nucleoside bearing an OCH3 moiety at position 2'.
[0344] The D1 diastereomer of beta-S-ARCA ([3-S-ARCA) has the following structure:
[0345] The "D1 diastereomer of beta-S-ARCA" or "beta-S-ARCA(DI )" is the diastereomer of beta-S-ARCA which elutes first on an HPLC column compared to the D2 diastereomer of beta-S-ARCA (beta-S-ARCA(D2)) and thus exhibits a shorter retention time. The HPLC preferably is an analytical HPLC. In some embodiments, a Supelcosil LC-18-T RP column, preferably of the format: 5 pm, 4.6 x 250 mm is used for separation, whereby a flow rate of 1 .3 ml / min can be applied. In some embodiments, a gradient of methanol in ammonium acetate, for example, a 0-25% linear gradient of methanol in 0.05 M ammonium acetate, pH = 5.9, within 15 min is used. UV-detection (VWD) can be performed at 260 nm and fluorescence detection (FLD) can be performed with excitation at 280 nm and detection at 337 nm.
[0346] The 5'-cap analog m27’3’0Gppp(mi2'’°)ApG (also referred to as rri27’3'0G(5')ppp(5')m2'’ °ApG) which is a building block of a cap1 has the following structure:
[0347] An exemplary capO mRNA comprising [3-S-ARCA and mRNA has the following structure:
[0348] An exemplary capO mRNA comprising m27’3 OG(5')ppp(5')G and mRNA has the following structure:
[0349]
[0350] An exemplary cap1 mRNA comprising m27’3’0Gppp(mi2'’°)ApG and mRNA has the
[0351] As used herein, the term "poly-A tail" or "poly-A sequence" refers to an uninterrupted or interrupted sequence of adenylate residues which is typically located at the 3' -end of an mRNA molecule. Poly-A tails or poly-A sequences are known to those of skill in the art and may follow the 3’-UTR in the mRNAs described herein. An uninterrupted poly-A tail is characterized by consecutive adenylate residues. In nature, an uninterrupted poly-A tail is typical. mRNAs disclosed herein can have a poly-A tail attached to the free 3'-end of the mRNA by a template-independent RNA polymerase after transcription or a poly-A tail encoded by DNA and transcribed by a template- dependent RNA polymerase.
[0352] It has been demonstrated that a poly-A tail of about 120 A nucleotides has a beneficial influence on the levels of mRNA in transfected eukaryotic cells, as well as on the levels of protein that is translated from an open reading frame that is present upstream (5’) of the poly-A tail (Holtkamp et al., 2006, Blood, vol. 108, pp. 4009-4017).
[0353] The poly-A tail may be of any length. In some embodiments, a poly-A tail comprises, essentially consists of, or consists of at least 20, at least 30, at least 40, at least 80, or at least 100 and up to 500, up to 400, up to 300, up to 200, or up to 150 A nucleotides, and, in particular, about 120 A nucleotides. In this context, "essentially consists of" means that most nucleotides in the poly-A tail, typically at least 75%, at least 80%, at least 85%, at least 90%, at least 95%, at least 96%, at least 97%, at least 98%, or at least 99% by number of nucleotides in the poly-A tail are A nucleotides, but permits that remaining nucleotides are nucleotides other than A nucleotides, such as U nucleotides (uridylate), G nucleotides (guanylate), or C nucleotides (cytidylate). In this context, "consists of' means that all nucleotides in the poly-A tail, i.e., 100% by number of nucleotides in the poly-A tail, are A nucleotides. The term "A nucleotide" or "A" refers to adenylate.
[0354] In some embodiments, a poly-A tail is attached during RNA transcription, e.g., during preparation of in vitro transcribed RNA, based on a DNA template comprising repeated dT nucleotides (deoxythymidylate) in the strand complementary to the coding strand. The DNA sequence encoding a poly-A tail (coding strand) is referred to as poly(A) cassette.
[0355] In some embodiments, the poly(A) cassette present in the coding strand of DNA essentially consists of dA nucleotides, but is interrupted by a random sequence of the four nucleotides (dA, dC, dG, and dT). Such random sequence may be 5 to 50, 10 to 30, or 10 to 20 nucleotides in length. Such a cassette is disclosed in WO 2016 / 005324 A1 , hereby incorporated by reference. Any poly(A) cassette disclosed in WO 2016 / 005324 A1 may be used in the present disclosure. A poly(A) cassette that essentially consists of dA nucleotides, but is interrupted by a random sequence having an equal distribution of the four nucleotides (dA, dC, dG, dT) and having a length of e.g., 5 to 50 nucleotides shows, on DNA level, constant propagation of plasmid DNA in E. co / / and is still associated, on RNA level, with the beneficial properties with respect to supporting RNA stability and translational efficiency is encompassed. Consequently, in some embodiments, the poly-A tail contained in an mRNA molecule described herein essentially consists of A nucleotides, but is interrupted by a random sequence of the four nucleotides (A, C, G, U). Such random sequence may be 5 to 50, 10 to 30, or 10 to 20 nucleotides in length. In some embodiments, no nucleotides other than A nucleotides flank a poly-A tail at its 3' -end, i.e., the poly-A tail is not masked or followed at its 3' -end by a nucleotide other than A.
[0356] In some embodiments, a poly-A tail may comprise at least 20, at least 30, at least 40, at least 80, or at least 100 and up to 500, up to 400, up to 300, up to 200, or up to 150 nucleotides. In some embodiments, the poly-A tail may essentially consist of at least 20, at least 30, at least 40, at least 80, or at least 100 and up to 500, up to 400, up to 300, up to 200, or up to 150 nucleotides. In some embodiments, the poly-A tail may consist of at least 20, at least 30, at least 40, at least 80, or at least 100 and up to 500, up to 400, up to 300, up to 200, or up to 150 nucleotides. In some embodiments, the poly-A tail comprises at least 100 nucleotides. In some embodiments, the poly-A tail comprises about 150 nucleotides. In some embodiments, the poly-A tail comprises about 120 nucleotides.
[0357] In some embodiments, mRNA used in present disclosure comprises a 5'-UTR and / or a 3'-UTR. The term "untranslated region" or "UTR" relates to a region in a DNA molecule which is transcribed but is not translated into an amino acid sequence, or to the corresponding region in an RNA molecule, such as an mRNA molecule. An untranslated region (UTR) can be present 5' (upstream) of an open reading frame (5 - UTR) and / or 3' (downstream) of an open reading frame (3'-UTR). A 5'-UTR, if present, is located at the 5'-end, upstream of the start codon of a protein-encoding region. A 5'- UTR is downstream of the 5'-cap (if present), e.g., directly adjacent to the 5'-cap. A 3'- UTR, if present, is located at the 3' -end, downstream of the termination codon of a protein-encoding region, but the term "3'-UTR" does generally not include the poly-A sequence. Thus, the 3'-UTR is upstream of the poly-A sequence (if present), e.g., directly adjacent to the poly-A sequence. Incorporation of a 3'-UTR into the 3'-non translated region of an RNA (preferably mRNA) molecule can result in an enhancement in translation efficiency. A synergistic effect may be achieved by incorporating two or more of such 3'-UTRs (which are preferably arranged in a head- to-tail orientation; cf., e.g., Holtkamp etal., Blood 108, 4009-4017 (2006)). The 3'-UTRs may be autologous or heterologous to the RNA (e.g., mRNA) into which they are introduced. In certain embodiments, the 3'-UTR is derived from a globin gene or mRNA, such as a gene or mRNA of alpha2-globin, alphal -globin, or beta-globin, e.g., beta-globin, e.g., human beta-globin. For example, the RNA (e.g., mRNA) may be modified by the replacement of the existing 3'-UTR with or the insertion of one or more, e.g., two copies of a 3'-UTR derived from a globin gene, such as alpha2-globin, alphal - globin, beta-globin, e.g., beta-globin, e.g., human beta-globin.
[0358] The mRNA may have modified ribonucleotides in order to increase its stability and / or decrease immunogenicity and / or decrease cytotoxicity. For example, in some embodiments, undine in the mRNA described herein is replaced (partially or completely, preferably completely) by a modified nucleoside. In some embodiments, the modified nucleoside is a modified uridine.
[0359] In some embodiments, the modified undine replacing uridine is selected from the group consisting of pseudouridine (ip), N1 -methyl-pseudouridine (m1 ip), 5-methyl-uridine (m5U), and combinations thereof.
[0360] In some embodiments, the modified nucleoside replacing (partially or completely, preferably completely) uridine in the mRNA may be any one or more of 3-methyl- uridine (m3U), 5-methoxy-uridine (mo5U), 5-aza-uridine, 6-aza-uridine, 2-thio-5-aza- uridine, 2-thio-uridine (s2U), 4-thio-uridine (s4U), 4-thio-pseudouridine, 2-thio- pseudouridine, 5-hydroxy-uridine (ho5U), 5-aminoallyl-uridine, 5-halo-uridine (e.g., 5- iodo-uridineor 5-bromo-uridine), uridine 5-oxyacetic acid (cmo5U), uridine 5-oxyacetic acid methyl ester (mcmo5U), 5-carboxymethyl-uridine (cm5U), 1 -carboxymethyl- pseudouridine, 5-carboxyhydroxymethyl-uridine (chm5U), 5-carboxyhydroxymethyl- uridine methyl ester (mchm5U), 5-methoxycarbonylmethyl-uridine (mcm5U), 5- methoxycarbonylmethyl-2 -thio-uridine (mcm5s2U), 5-aminomethyl-2 -thio-uridine (nm5s2U), 5-methylaminomethyl-uridine (mnm5U), 1 -ethyl-pseudouridine, 5- methylaminomethyl-2-thio-uridine (mnm5s2U), 5-methylaminomethyl-2-seleno-uridine (mnm5se2U), 5-carbamoylmethyl-uridine (ncm5U), 5-carboxymethylaminomethyl- uridine (cmnm5U), 5-carboxymethylaminomethyl-2-thio-uridine (cmnm5s2U), 5- propynyl-uridine, 1 -propynyl-pseudouridine, 5-taurinomethyl-uridine (im5U), 1- taurinomethyl-pseudouridine, 5-taurinomethyl-2-thio-uridine(Tm5s2U), 1 - taurinomethyl-4-thio-pseudouridine), 5-methyl-2-thio-uridine (m5s2U), 1 -methyl-4- thio-pseudouridine (m1 s4ip), 4-thio-1 -methyl-pseudouridine, 3-methyl-pseudouridine (m3ip), 2-thio-1 -methyl-pseudouridine, 1 -methyl-1 -deaza-pseudouridine, 2-thio-1 - methyl-1 -deaza-pseudouridine, dihydrouridine (D), dihydropseudouridine, 5,6- dihydrouridine, 5-methyl-dihydrouridine (m5D), 2-thio-dihydrouridine, 2-thio- dihydropseudouridine, 2-methoxy-uridine, 2-methoxy-4-thio-uridine, 4-methoxy- pseudouridine, 4-methoxy-2-thio-pseudouridine, N1-methyl-pseudouridine, 3-(3- amino-3-carboxypropyl)uridine (acp3U), 1 -methyl-3-(3-amino-3- carboxypropyl)pseudouridine (acp3 ip), 5-(isopentenylaminomethyl)uridine (inm5U), 5- (isopentenylaminomethyl)-2-thio-uridine (inm5s2U), a-thio-uridine, 2'-O-methyl-uridine (Um), 5,2'-O-dimethyl-uridine (m5Um), 2'-O-methyl-pseudouridine (ipm), 2-thio-2'-O- methyl-uridine (s2Um), 5-methoxycarbonylmethyl-2'-O-methyl-uridine (mcm5Um), 5- carbamoylmethyl-2'-O-methyl-uridine (ncm5Um), 5-carboxymethylaminomethyl-2'-O- methyl-uridine (cmnm5Um), 3,2'-O-dimethyl-uridine (m3Um), 5-
[0361] (isopentenylaminomethyl)-2'-O-methyl-uridine (inm5Um), 1 -thio-uridine, deoxythymidine, 2'-F-ara-uridine, 2'-F-uridine, 2'-OH-ara-uridine, 5-(2- carbomethoxyvinyl) uridine, 5-[3-(1 -E-propenylamino)uridine, or any other modified uridine known in the art.
[0362] An RNA (preferably mRNA) which is modified by pseudouridine (replacing partially or completely, preferably completely, uridine) is referred to herein as "MJ-modified", whereas the term "ml ^P-modified" means that the RNA (preferably mRNA) contains N(1 )-methylpseudouridine (replacing partially or completely, preferably completely, uridine). Furthermore, the term "m5U-modified" means that the RNA (preferably mRNA) contains 5-methyluridine (replacing partially or completely, preferably completely, undine). Such ^P- or ml ^P- or m5U-modified RNAs usually exhibit decreased immunogenicity compared to their unmodified forms and, thus, are preferred in applications where the induction of an immune response is to be avoided or minimized. In some embodiments, the RNA (preferably mRNA) contains N(1 )- methylpseudouridine replacing completely uridine
[0363] The codons of the mRNA used in the present disclosure may further be optimized, e.g., to increase the GC content of the RNA and / or to replace codons which are rare in the cell (or subject) in which the peptide or polypeptide of interest is to be expressed by codons which are synonymous frequent codons in said cell (or subject). In some embodiments, the amino acid sequence encoded by the mRNA used in the present disclosure is encoded by a coding sequence which is codon-optimized and / or the G / C content of which is increased compared to wild type coding sequence. This also includes embodiments, wherein one or more sequence regions of the coding sequence are codon-optimized and / or increased in the G / C content compared to the corresponding sequence regions of the wild type coding sequence. In some embodiments, the codon-optimization and / or the increase in the G / C content preferably does not change the sequence of the encoded amino acid sequence. The term "codon-optimized" refers to the alteration of codons in the coding region of a nucleic acid molecule to reflect the typical codon usage of a host organism without preferably altering the amino acid sequence encoded by the nucleic acid molecule. Within the context of the present disclosure, coding regions may be codon-optimized for optimal expression in a subject to be treated using the mRNA described herein. Codon-optimization is based on the finding that the translation efficiency is also determined by a different frequency in the occurrence of tRNAs in cells. Thus, the sequence of mRNA may be modified such that codons for which frequently occurring tRNAs are available are inserted in place of "rare codons".
[0364] In some embodiments, the guanosine / cytosine (G / C) content of the coding region of the mRNA described herein is increased compared to the G / C content of the corresponding coding sequence of the wild type RNA, wherein the amino acid sequence encoded by the mRNA is preferably not modified compared to the amino acid sequence encoded by the wild type RNA. This modification of the mRNA sequence is based on the fact that the sequence of any RNA region to be translated is important for efficient translation of that mRNA. Sequences having an increased G (guanosine)ZC (cytosine) content are more stable than sequences having an increased A (adenosine)ZU (uracil) content. In respect to the fact that several codons code for one and the same amino acid (so-called degeneration of the genetic code), the most favorable codons for the stability can be determined (so-called alternative codon usage). Depending on the amino acid to be encoded by the mRNA, there are various possibilities for modification of the mRNA sequence, compared to its wild type sequence. In particular, codons which contain A and / or U nucleotides can be modified by substituting these codons by other codons, which code for the same amino acids but contain no A and / or U or contain a lower content of A and / or U nucleotides.
[0365] In various embodiments, the G / C content of the coding region of the mRNA described herein is increased by at least 10%, at least 20%, at least 30%, at least 40%, at least 50%, at least 55%, or even more compared to the G / C content of the coding region of the wild type RNA.
[0366] A combination of the above described modifications, i.e., incorporation of a 5'-cap structure, incorporation of a poly-A sequence, unmasking of a poly-A sequence, alteration of the 5'- and / or 3'-UTR (such as incorporation of one or more 3'-UTRs), replacing one or more naturally occurring nucleotides with synthetic nucleotides (e.g., 5-methylcytidine for cytidine and / or pseudouridine (^P) or N(1 )-methylpseudouridine (mI MJ) or 5-methyluridine (m5U) for uridine), and codon optimization, has a synergistic influence on the stability of RNA (preferably mRNA) and increase in translation efficiency. Thus, in some embodiments, the mRNA used in the present disclosure contains a combination of at least two, at least three, at least four or all five of the above-mentioned modifications, i.e., (i) incorporation of a 5'-cap structure, (ii) incorporation of a poly-A sequence, unmasking of a poly-A sequence; (iii) alteration of the 5'- and / or 3'-UTR (such as incorporation of one or more 3'-UTRs); (iv) replacing one or more naturally occurring nucleotides with synthetic nucleotides (e.g., 5- methylcytidine for cytidine and / or pseudouridine (^P) or N(1 )-methylpseudouridine (ml M-J) or 5-methyluridine (m5U) for uridine), and (v) codon optimization.
[0367] Some aspects of the disclosure involve the targeted delivery of the mRNA disclosed herein to certain cells or tissues. In some embodiments, the disclosure involves targeting the lymphatic system, in particular secondary lymphoid organs, more specifically spleen. Targeting the lymphatic system, in particular secondary lymphoid organs, more specifically spleen is in particular preferred if the mRNA administered is mRNA encoding an antigen or epitope for inducing an immune response. In some embodiments, the target cell is a spleen cell. In some embodiments, the target cell is an antigen presenting cell such as a professional antigen presenting cell in the spleen. In some embodiments, the target cell is a dendritic cell in the spleen. The "lymphatic system" is part of the circulatory system and an important part of the immune system, comprising a network of lymphatic vessels that carry lymph. The lymphatic system consists of lymphatic organs, a conducting network of lymphatic vessels, and the circulating lymph. The primary or central lymphoid organs generate lymphocytes from immature progenitor cells. The thymus and the bone marrow constitute the primary lymphoid organs. Secondary or peripheral lymphoid organs, which include lymph nodes and the spleen, maintain mature naive lymphocytes and initiate an adaptive immune response.
[0368] Lipid-based mRNA delivery systems have an inherent preference to the liver. Liver accumulation is caused by the discontinuous nature of the hepatic vasculature or the lipid metabolism (liposomes and lipid or cholesterol conjugates). In some embodiments, the target organ is liver and the target tissue is liver tissue. The delivery to such target tissue is preferred, in particular, if presence of mRNA or of the encoded peptide or polypeptide in this organ or tissue is desired and / or if it is desired to express large amounts of the encoded peptide or polypeptide and / or if systemic presence of the encoded peptide or polypeptide, in particular in significant amounts, is desired or required.
[0369] In some embodiments, after administration of the mRNA particles described herein, at least a portion of the mRNA is delivered to a target cell or target organ. In some embodiments, at least a portion of the mRNA is delivered to the cytosol of the target cell. In some embodiments, the mRNA is mRNA encoding a peptide or polypeptide and the mRNA is translated by the target cell to produce the peptide or polypeptide. In some embodiments, the target cell is a cell in the liver. In some embodiments, the target cell is a muscle cell. In some embodiments, the target cell is an endothelial cell. In some embodiments the target cell is a tumor cell or a cell in the tumor microenvironment. In some embodiments, the target cell is a blood cell. In some embodiments, the target cell is a cell in the lymph nodes. In some embodiments, the target cell is a cell in the lung. In some embodiments, the target cell is a blood cell. In some embodiments, the target cell is a cell in the skin. In some embodiments, the target cell is a spleen cell. In some embodiments, the target cell is an antigen presenting cell such as a professional antigen presenting cell in the spleen. In some embodiments, the target cell is a dendritic cell in the spleen. In some embodiments, the target cell is a T cell. In some embodiments, the target cell is a B cell. In some embodiments, the target cell is a NK cell. In some embodiments, the target cell is a monocyte. Thus, RNA particles described herein may be used for delivering mRNA to such target cell.
[0370] Pharmaceutically active peptides or polypeptides
[0371] "Encoding" refers to the inherent property of specific sequences of nucleotides in a polynucleotide, such as a gene, a cDNA, or an mRNA, to serve as templates for synthesis of other polymers and macromolecules in biological processes having either a defined sequence of nucleotides ( / .e., rRNA, tRNA and mRNA) or a defined sequence of amino acids and the biological properties resulting therefrom. Thus, a gene encodes a protein if transcription and translation of mRNA corresponding to that gene produces the protein in a cell or other biological system. Both the coding strand, the nucleotide sequence of which is identical to the mRNA sequence and is usually provided in sequence listings, and the non-coding strand, used as the template for transcription of a gene or cDNA, can be referred to as encoding the protein or other product of that gene or cDNA. In some embodiments, nucleic acid such as mRNA used in the present disclosure comprises a nucleic acid sequence encoding one or more functional sequences which can be peptides or polypeptides, preferably a pharmaceutically active peptide or polypeptide.
[0372] In a preferred embodiment, nucleic acid such as mRNA used in the present disclosure comprises a nucleic acid sequence encoding a peptide or polypeptide, preferably a pharmaceutically active peptide or polypeptide, and is capable of expressing said peptide or polypeptide, in particular if transferred into a cell or subject. Thus, in some embodiments, the nucleic acid used in the present disclosure contains a coding region (open reading frame (ORF)) encoding a peptide or polypeptide, e.g., encoding a pharmaceutically active peptide or polypeptide. In this respect, an "open reading frame" or "ORF" is a continuous stretch of codons beginning with a start codon and ending with a stop codon. Such nucleic acid encoding a pharmaceutically active peptide or polypeptide is also referred to herein as "pharmaceutically active nucleic acid". In particular, such mRNA encoding a pharmaceutically active peptide or polypeptide is also referred to herein as "pharmaceutically active mRNA".
[0373] According to the present disclosure, the term "pharmaceutically active peptide or polypeptide" means a peptide or polypeptide that can be used in the treatment of an individual where the expression of a peptide or polypeptide would be of benefit, e.g., in ameliorating the symptoms of a disease. Preferably, a pharmaceutically active peptide or polypeptide has curative or palliative properties and may be administered to ameliorate, relieve, alleviate, reverse, delay onset of or lessen the severity of one or more symptoms of a disease. In some embodiments, a pharmaceutically active peptide or polypeptide has a positive or advantageous effect on the condition or disease state of an individual when administered to the individual in a therapeutically effective amount. A pharmaceutically active peptide or polypeptide may have prophylactic properties and may be used to delay the onset of a disease or to lessen the severity of such disease. The term "pharmaceutically active peptide or polypeptide" includes entire peptides or polypeptides, and can also refer to pharmaceutically active fragments thereof. It can also include pharmaceutically active variants and / or analogs of a peptide or polypeptide.
[0374] Specific examples of pharmaceutically active peptides and polypeptides include, but are not limited to, cytokines, hormones, adhesion molecules, immunoglobulins, immunologically active compounds, growth factors, protease inhibitors, enzymes, receptors, apoptosis regulators, transcription factors, tumor suppressor proteins, structural proteins, reprogramming factors, genomic engineering proteins, and blood proteins.
[0375] The term "cytokines" relates to proteins which have a molecular weight of about 5 to 60 kDa and which participate in cell signaling (e.g., paracrine, endocrine, and / or autocrine signaling). In particular, when released, cytokines exert an effect on the behavior of cells around the place of their release. Examples of cytokines include lymphokines, interleukins, chemokines, interferons, and tumor necrosis factors (TNFs). According to the present disclosure, cytokines do not include hormones or growth factors. Cytokines differ from hormones in that (i) they usually act at much more variable concentrations than hormones and (ii) generally are made by a broad range of cells (nearly all nucleated cells can produce cytokines). Interferons are usually characterized by antiviral, antiproliferative and immunomodulatory activities. Interferons are proteins that alter and regulate the transcription of genes within a cell by binding to interferon receptors on the regulated cell's surface, thereby preventing viral replication within the cells. The interferons can be grouped into two types. IFN- gamma is the sole type II interferon; all others are type I interferons. Particular examples of cytokines include erythropoietin (EPO), colony stimulating factor (CSF), granulocyte colony stimulating factor (G-CSF), granulocyte-macrophage colony stimulating factor (GM-CSF), tumor necrosis factor (TNF), bone morphogenetic protein (BMP), interferon alfa (IFNa), interferon beta (IFN|3), interferon gamma (INFy), interleukin 2 (IL-2), interleukin 4 (IL-4), interleukin 10 (IL-10), interleukin 11 (IL-11 ), interleukin 12 (IL-12), interleukin 15 (IL-15), and interleukin 21 (IL-21 ), as well as variants and derivatives thereof.
[0376] In some embodiments, a pharmaceutically active peptide or polypeptide comprises a replacement protein. In these embodiments, the present disclosure provides a method for treatment of a subject having a disorder requiring protein replacement (e.g., protein deficiency disorders) comprising administering to the subject nucleic acid as described herein encoding a replacement protein. The term "protein replacement" refers to the introduction of a protein (including functional variants thereof) into a subject having a deficiency in such protein. The term also refers to the introduction of a protein into a subject otherwise requiring or benefiting from providing a protein, e.g., suffering from protein insufficiency. The term "disorder characterized by a protein deficiency" refers to any disorder that presents with a pathology caused by absent or insufficient amounts of a protein. This term encompasses protein folding disorders, i.e., conformational disorders, that result in a biologically inactive protein product. Protein insufficiency can be involved in infectious diseases, immunosuppression, organ failure, glandular problems, radiation illness, nutritional deficiency, poisoning, or other environmental or external insults.
[0377] The term "hormones" relates to a class of signaling molecules produced by glands, wherein signaling usually includes the following steps: (i) synthesis of a hormone in a particular tissue; (ii) storage and secretion; (iii) transport of the hormone to its target; (iv) binding of the hormone by a receptor; (v) relay and amplification of the signal; and (vi) breakdown of the hormone. Hormones differ from cytokines in that (1 ) hormones usually act in less variable concentrations and (2) generally are made by specific kinds of cells. In some embodiments, a "hormone" is a peptide or polypeptide hormone, such as insulin, vasopressin, prolactin, adrenocorticotropic hormone (ACTH), thyroid hormone, growth hormones (such as human grown hormone or bovine somatotropin), oxytocin, atrial-natriuretic peptide (ANP), glucagon, somatostatin, cholecystokinin, gastrin, and leptins.
[0378] The term "adhesion molecules" relates to proteins which are located on the surface of a cell and which are involved in binding of the cell with other cells or with the extracellular matrix (ECM). Adhesion molecules are typically transmembrane receptors and can be classified as calcium-independent (e.g., integrins, immunoglobulin superfamily, lymphocyte homing receptors) and calcium-dependent (cadherins and selectins). Particular examples of adhesion molecules are integrins, lymphocyte homing receptors, selectins (e.g., P-selectin), and addressins.
[0379] Integrins are also involved in signal transduction. In particular, upon ligand binding, integrins modulate cell signaling pathways, e.g., pathways of transmembrane protein kinases such as receptor tyrosine kinases (RTK). Such regulation can lead to cellular growth, division, survival, or differentiation or to apoptosis. Particular examples of integrins include: ai|3i , Q2|3I , cop-i, a4[3i , as|3i , ae|3i, a?[3i, ai_|32, OM[32, aiib^s, av|3i , av[33, avPs, av[36, av[38, and aef
[0380] The term "immunoglobulins" or "immunoglobulin superfamily" refers to molecules which are involved in the recognition, binding, and / or adhesion processes of cells. Molecules belonging to this superfamily share the feature that they contain a region known as immunoglobulin domain or fold. Members of the immunoglobulin superfamily include antibodies (e.g., IgG), T cell receptors (TCRs), major histocompatibility complex (MHC) molecules, co-receptors (e.g., CD4, CD8, CD19), antigen receptor accessory molecules (e.g., CD-3y, CD3-6, CD-3s, CD79a, CD79b), co-stimulatory or inhibitory molecules (e.g., CD28, CD80, CD86), and other.
[0381] The term "immunologically active compound" relates to any compound altering an immune response, e.g., by inducing and / or suppressing maturation of immune cells, inducing and / or suppressing cytokine biosynthesis, and / or altering humoral immunity by stimulating antibody production by B cells. Immunologically active compounds possess potent immunostimulating activity including, but not limited to, antiviral and antitumor activity, and can also down-regulate other aspects of the immune response, for example shifting the immune response away from a TH2 immune response, which is useful for treating a wide range of TH2 mediated diseases. Immunologically active compounds can be useful as vaccine adjuvants. Particular examples of immunologically active compounds include interleukins, colony stimulating factor (CSF), granulocyte colony stimulating factor (G-CSF), granulocyte-macrophage colony stimulating factor (GM-CSF), erythropoietin, tumor necrosis factor (TNF), interferons, integrins, addressins, selectins, homing receptors, and antigens, in particular tumor- associated antigens, pathogen-associated antigens (such as bacterial, parasitic, or viral antigens), allergens, and autoantigens. An immunologically active compound may be a vaccine antigen, i.e., an antigen whose inoculation into a subject induces an immune response.
[0382] An "antigen" according to the present disclosure covers any substance that will elicit an immune response and / or any substance against which an immune response or an immune mechanism such as a cellular response and / or humoral response is directed. This also includes situations wherein the antigen is processed into antigen peptides and an immune response or an immune mechanism is directed against one or more antigen peptides, in particular if presented in the context of MHC molecules. In particular, an "antigen" relates to any substance, such as a peptide or polypeptide, that reacts specifically with antibodies or T-lymphocytes (T-cells). The term "antigen" may comprise a molecule that comprises at least one epitope, such as a T cell epitope. In some embodiments, an antigen is a molecule which, optionally after processing, induces an immune reaction, which may be specific for the antigen (including cells expressing the antigen). In some embodiments, an antigen is a disease-associated antigen, such as a tumor antigen, a viral antigen, or a bacterial antigen, or an epitope derived from such antigen. The term "autoantigen" or "self-antigen" refers to an antigen which originates from within the body of a subject (i.e., the autoantigen can also be called "autologous antigen") and which produces an abnormally vigorous immune response against this normal part of the body. Such vigorous immune reactions against autoantigens may be the cause of "autoimmune diseases".
[0383] According to the present disclosure, any suitable antigen may be used, which is a candidate for an immune response, wherein the immune response may be both a humoral as well as a cellular immune response. In the context of some embodiments of the present disclosure, the antigen is presented by a cell, such as by an antigen presenting cell, in the context of MHC molecules, which results in an immune response against the antigen. An antigen may be a product which corresponds to or is derived from a naturally occurring antigen. Such naturally occurring antigens may include or may be derived from allergens, viruses, bacteria, fungi, parasites and other infectious agents and pathogens or an antigen may also be a tumor antigen. According to the present disclosure, an antigen may correspond to a naturally occurring product, for example, a viral protein, or a part thereof.
[0384] The term "disease-associated antigen" is used in its broadest sense to refer to any antigen associated with a disease. A disease-associated antigen is a molecule which contains epitopes that will stimulate a host's immune system to make a cellular antigenspecific immune response and / or a humoral antibody response against the disease. Disease-associated antigens include pathogen-associated antigens, i.e., antigens which are associated with infection by microbes, typically microbial antigens (such as bacterial or viral antigens), or antigens associated with cancer, typically tumors, such as tumor antigens.
[0385] In some embodiments, the antigen is a tumor antigen, i.e., a part of a tumor cell, in particular those which primarily occur intracellularly or as surface antigens of tumor cells. In another embodiment, the antigen is a pathogen-associated antigen, i.e., an antigen derived from a pathogen, e.g., from a virus, bacterium, unicellular organism, or parasite, for example a viral antigen such as viral ribonucleoprotein or coat protein. In some embodiments, the antigen should be presented by MHC molecules which results in modulation, in particular activation of cells of the immune system, such as CD4+ and CD8+ lymphocytes, in particular via the modulation of the activity of a T-cell receptor. The term "tumor antigen" refers to a constituent of cancer cells which may be derived from the cytoplasm, the cell surface or the cell nucleus. In particular, it refers to those antigens which are produced intracellularly or as surface antigens on tumor cells. For example, tumor antigens include the carcinoembryonal antigen, a1 -fetoprotein, isoferritin, and fetal sulphoglycoprotein, a2-H-ferroprotein and y-fetoprotein, as well as various virus tumor antigens. According to some embodiments of the present disclosure, a tumor antigen comprises any antigen which is characteristic for tumors or cancers as well as for tumor or cancer cells with respect to type and / or expression level.
[0386] The term "viral antigen" refers to any viral component having antigenic properties, i.e., being able to provoke an immune response in an individual. The viral antigen may be a viral ribonucleoprotein or an envelope protein.
[0387] The term "bacterial antigen" refers to any bacterial component having antigenic properties, i.e. being able to provoke an immune response in an individual. The bacterial antigen may be derived from the cell wall or cytoplasm membrane of the bacterium.
[0388] The term "epitope" refers to an antigenic determinant in a molecule such as an antigen, i.e., to a part in or fragment of the molecule that is recognized by the immune system, for example, that is recognized by antibodies, T cells or B cells, in particular when presented in the context of MHC molecules. An epitope of a protein may comprises a continuous or discontinuous portion of said protein and, e.g., may be between about 5 and about 100, between about 5 and about 50, between about 8 and about 30, or about 10 and about 25 amino acids in length, for example, the epitope may be preferably 9, 10, 11 , 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, or 25 amino acids in length. In some embodiments, the epitope in the context of the present disclosure is a T cell epitope.
[0389] Terms such as "epitope", "fragment of an antigen", "immunogenic peptide" and "antigen peptide" are used interchangeably herein and, e.g., may relate to an incomplete representation of an antigen which is, e.g., capable of eliciting an immune response against the antigen or a cell expressing or comprising and presenting the antigen. In some embodiments, the terms relate to an immunogenic portion of an antigen. In some embodiments, it is a portion of an antigen that is recognized (i.e., specifically bound) by a T cell receptor, in particular if presented in the context of MHC molecules. Certain preferred immunogenic portions bind to an MHC class I or class II molecule. The term "epitope" refers to a part or fragment of a molecule such as an antigen that is recognized by the immune system. For example, the epitope may be recognized by T cells, B cells or antibodies. An epitope of an antigen may include a continuous or discontinuous portion of the antigen and may be between about 5 and about 100, such as between about 5 and about 50, between about 8 and about 30, or between about 8 and about 25 amino acids in length, for example, the epitope may be 9, 10, 11 , 12, 13, 14, 15, 16, 17, 18, 19, 20, 21 , 22, 23, 24, or 25 amino acids in length. In some embodiments, an epitope is between about 10 and about 25 amino acids in length. The term "epitope" includes T cell epitopes.
[0390] The term "T cell epitope" refers to a part or fragment of a protein that is recognized by a T cell when presented in the context of MHC molecules. The term "major histocompatibility complex" and the abbreviation "MHC" includes MHC class I and MHC class II molecules and relates to a complex of genes which is present in all vertebrates. MHC proteins or molecules are important for signaling between lymphocytes and antigen presenting cells or diseased cells in immune reactions, wherein the MHC proteins or molecules bind peptide epitopes and present them for recognition by T cell receptors on T cells. The proteins encoded by the MHC are expressed on the surface of cells, and display both self-antigens (peptide fragments from the cell itself) and non-self-antigens (e.g., fragments of invading microorganisms) to a T cell. In the case of class I MHC / peptide complexes, the binding peptides are typically about 8 to about 10 amino acids long although longer or shorter peptides may be effective. In the case of class II MHC / peptide complexes, the binding peptides are typically about 10 to about 25 amino acids long and are in particular about 13 to about 18 amino acids long, whereas longer and shorter peptides may be effective.
[0391] The peptide and polypeptide antigen can be 2 to 100 amino acids, including for example, 5 amino acids, 10 amino acids, 15 amino acids, 20 amino acids, 25 amino acids, 30 amino acids, 35 amino acids, 40 amino acids, 45 amino acids, or 50 amino acids in length. In some embodiments, a peptide can be greater than 50 amino acids. In some embodiments, the peptide can be greater than 100 amino acids.
[0392] The peptide or polypeptide antigen can be any peptide or polypeptide that can induce or increase the ability of the immune system to develop antibodies and T cell responses to the peptide or polypeptide.
[0393] In some embodiments, vaccine antigen, i.e., an antigen whose inoculation into a subject induces an immune response, is recognized by an immune effector cell. In some embodiments, the vaccine antigen if recognized by an immune effector cell is able to induce in the presence of appropriate co-stimulatory signals, stimulation, priming and / or expansion of the immune effector cell carrying an antigen receptor recognizing the vaccine antigen. In the context of the embodiments of the present disclosure, the vaccine antigen may be, e.g., presented or present on the surface of a cell, such as an antigen presenting cell. In some embodiments, an antigen is presented by a diseased cell (such as tumor cell or an infected cell). In some embodiments, an antigen receptor is a TCR which binds to an epitope of an antigen presented in the context of MHC. In some embodiments, binding of a TCR when expressed by T cells and / or present on T cells to an antigen presented by cells such as antigen presenting cells results in stimulation, priming and / or expansion of said T cells. In some embodiments, binding of a TCR when expressed by T cells and / or present on T cells to an antigen presented on diseased cells results in cytolysis and / or apoptosis of the diseased cells, wherein said T cells release cytotoxic factors, e.g., perforins and granzymes.
[0394] According to some embodiments, an amino acid sequence enhancing antigen processing and / or presentation is fused, either directly or through the linker sequence, to an antigenic peptide or polypeptide. Accordingly, in some embodiments, the nucleic acid (such as RNA and / or DNA) described herein comprises at least one coding region encoding an antigenic peptide or polypeptide and an amino acid sequence enhancing antigen processing and / or presentation.
[0395] Such amino acid sequences enhancing antigen processing and / or presentation are preferably located at the C-terminus of the antigenic peptide or polypeptide and linker sequence (and optionally at the C-terminus of an amino acid sequence which breaks immunological tolerance), without being limited thereto. Amino acid sequences enhancing antigen processing and / or presentation as defined herein preferably improve antigen processing and presentation. In one embodiment, the amino acid sequence enhancing antigen processing and / or presentation as defined herein includes, without being limited thereto, sequences derived from the human MHC class l complex (HLA-B51 , haplotype A2, B27 / B51 , Cw2 / Cw3). Besides improving antigen processing and presentation such amino acid sequence enhancing antigen processing and / or presentation may also be used for determining expression of an amino acid sequence in the processes described herein. Accordingly, in particularly preferred embodiments, the RNA described herein comprises at least one coding region encoding an antigenic peptide or polypeptide and an amino acid sequence enhancing antigen processing and / or presentation, said amino acid sequence enhancing antigen processing and / or presentation preferably being fused to the antigenic peptide or polypeptide, more preferably to the C-terminus of the antigenic peptide or polypeptide as described herein.
[0396] Furthermore, a secretory sequence may be fused to the N-terminus of the antigenic peptide or polypeptide.
[0397] Amino acid sequences derived from tetanus toxoid of Clostridium tetani may be employed to overcome self-tolerance mechanisms in order to efficiently mount an immune response to self-antigens by providing T-cell help during priming.
[0398] It is known that tetanus toxoid heavy chain includes epitopes that can bind promiscuously to MHC class II alleles and induce CD4+memory T cells in almost all tetanus vaccinated individuals. In addition, the combination of tetanus toxoid (TT) helper epitopes with tumor-associated antigens is known to improve the immune stimulation compared to application of tumor-associated antigen alone by providing CD4+-mediated T-cell help during priming. To reduce the risk of stimulating CD8+T cells with the tetanus sequences which might compete with the intended induction of tumor antigen-specific T-cell response, not the whole fragment C of tetanus toxoid is used as it is known to contain CD8+T-cell epitopes.
[0399] According to some embodiments, an amino acid sequence which breaks immunological tolerance is fused, either directly or through a linker to the antigenic peptide or polypeptide.
[0400] Such amino acid sequences which break immunological tolerance are preferably located at the C-terminus of the antigenic peptide or polypeptide (and optionally at the N-terminus of the amino acid sequence enhancing antigen processing and / or presentation, wherein the amino acid sequence which breaks immunological tolerance and the amino acid sequence enhancing antigen processing and / or presentation may be fused either directly or through a linker. Amino acid sequences which break immunological tolerance as defined herein preferably improve T cell responses. In one embodiment, the amino acid sequence which breaks immunological tolerance as defined herein includes, without being limited thereto, sequences derived from tetanus toxoid-derived helper sequences p2 and p16 (P2P16). According to some embodiments, an amino acid sequence which produces bioluminescence is fused, either directly or through a linker to the antigenic peptide or polypeptide.
[0401] Such amino acid sequences which produces bioluminescence are preferably located at the C-terminus of the antigenic peptide or polypeptide (and optionally at the N- terminus of (i) the amino acid sequence enhancing antigen processing and / or presentation or (ii) the amino acid sequence which breaks immunological tolerance, wherein the amino acid sequence which produces bioluminescence and (i) the amino acid sequence enhancing antigen processing and / or presentation or (ii) the amino acid sequence which breaks immunological tolerance may be fused either directly or through a linker. Amino acid sequences which produce bioluminescence as defined herein preferably improve the determination of the amount of the antigenic peptide or polypeptide. In some embodiments, the amino acid sequence which produces bioluminescence as defined herein produces fluorescence. In some embodiments, the amino acid sequence which produces bioluminescence as defined herein includes, without being limited thereto, sequences derived from Green Fluorescent Protein (GFP), Yellow Fluorescent Protein (YFP), Red Fluorescent Protein (RFP), Blue Fluorescent Protein (EBFP), Cyan Fluorescent Protein (ECFP), their variants (such as enhanced GFP (EGFP), Superfolder GFP (sfGFP), and luciferase.
[0402] In the following, embodiments of vaccine RNAs are described, wherein certain terms used when describing elements thereof have the following meanings: hAg-Kozak: 5'-UTR sequence of the human alpha-globin mRNA with an optimized ‘Kozak sequence’ to increase translational efficiency. sec / MlTD: Fusion-protein tags derived from the sequence encoding the human MHC class l complex (HLA-B51 , haplotype A2, B27 / B51 , Cw2 / Cw3), which have been shown to improve antigen processing and presentation. Sec corresponds to the 78 bp fragment coding for the secretory signal peptide, which guides translocation of the nascent polypeptide chain into the endoplasmatic reticulum. MITD corresponds to the transmembrane and cytoplasmic domain of the MHC class I molecule, also called MHC class I trafficking domain.
[0403] Antigen: Sequences encoding the respective antigen / epitope.
[0404] Glycine-serine linker (GS): Sequences coding for linker sequences according to the present invention, which, in an embodiment, are glycine-serine linker sequences, short linker peptides predominantly consisting of the amino acids glycine (G) and serine (S), as commonly used for fusion proteins. In a specific embodiment of the present invention, the linker sequence is preceded at its N terminus by a lysine residue and can be represented as follows: GGSGGGGSGGR / K. Thus, part of the amino acid sequence comprising the linker sequence can be represented as follows: KAGGSGGGGSGGR / K (A indicates the proteolytic cleavage site). After cleavage, this results in the excising of the linker sequence as follows: GGSGGGGSGGR / K. In an embodiment, the linker sequences of the present invention are GS linkers each comprising at least one residue which is not G or S, wherein the amino acid residue forms the proteolytic cleavage site of a proteolytic enzyme.
[0405] P2P16: Sequence coding for tetanus toxoid-derived helper epitopes to break immunological tolerance.
[0406] Fl element: The 3'-UTR is a combination of two sequence elements derived from the “amino terminal enhancer of split” (AES) mRNA (called F) and the mitochondrial encoded 12S ribosomal RNA (called I). These were identified by an ex vivo selection process for sequences that confer RNA stability and augment total protein expression. A30L70: A poly(A)-tail measuring 110 nucleotides in length, consisting of a stretch of 30 adenosine residues, followed by a 10 nucleotide linker sequence and another 70 adenosine residues designed to enhance RNA stability and translational efficiency in dendritic cells.
[0407] In one embodiment, vaccine RNA described herein has the structure: beta-S-ARCA(D1 )-hAg-Kozak-sec-GS(1 )-Antigen-GS(2)-P2P16-GS(3)-MITD-FI- A30L70
[0408] In one embodiment, vaccine antigen described herein has the structure: sec-GS( 1 )-Antigen-GS(2)-P2P 16-GS(3)-M ITD
[0409] In one embodiment, there are multiple vaccine antigen RNA constructs (nucleic acids) as described herein comprised in one formulation, such as one particle (LNP, LPX, PLX etc.), wherein each vaccine RNA construct comprises a different linker sequence. In some embodiments, an antigen receptor is an antibody or B cell receptor which binds to an epitope in an antigen. In some embodiments, an antibody or B cell receptor binds to native epitopes of an antigen.
[0410] The term "expressed on the cell surface" or "associated with the cell surface" means that a molecule such as an antigen is associated with and located at the plasma membrane of a cell, wherein at least a part of the molecule faces the extracellular space of said cell and is accessible from the outside of said cell, e.g., by antibodies located outside the cell. In this context, a part may be, e.g., at least 4, at least 8, pat least 12, or at least 20 amino acids. The association may be direct or indirect. For example, the association may be by one or more transmembrane domains, one or more lipid anchors, or by the interaction with any other protein, lipid, saccharide, or other structure that can be found on the outer leaflet of the plasma membrane of a cell. For example, a molecule associated with the surface of a cell may be a transmembrane protein having an extracellular portion or may be a protein associated with the surface of a cell by interacting with another protein that is a transmembrane protein.
[0411] "Cell surface" or "surface of a cell" is used in accordance with its normal meaning in the art, and thus includes the outside of the cell which is accessible to binding by proteins and other molecules. An antigen is expressed on the surface of cells if it is located at the surface of said cells and is accessible to binding by, e.g. , antigen-specific antibodies added to the cells.
[0412] The term "extracellular portion" or "exodomain" in the context of the present disclosure refers to a part of a molecule such as a protein that is facing the extracellular space of a cell and preferably is accessible from the outside of said cell, e.g., by binding molecules such as antibodies located outside the cell. In some embodiments, the term refers to one or more extracellular loops or domains or a fragment thereof.
[0413] The terms "T cell" and "T lymphocyte" are used interchangeably herein and include T helper cells (CD4+ T cells) and cytotoxic T cells (CTLs, CD8+ T cells) which comprise cytolytic T cells. The term "antigen-specific T cell" or similar terms relate to a T cell which recognizes the antigen to which the T cell is targeted, in particular when presented on the surface of antigen presenting cells or diseased cells such as cancer cells in the context of MHC molecules and preferably exerts effector functions of T cells. T cells are considered to be specific for antigen if the cells kill target cells expressing an antigen. T cell specificity may be evaluated using any of a variety of standard techniques, for example, within a chromium release assay or proliferation assay. Alternatively, synthesis of lymphokines (such as interferon-y) can be measured. The term "target" shall mean an agent such as a cell or tissue which is a target for an immune response such as a cellular immune response. Targets include cells that present an antigen or an antigen epitope, i.e., a peptide fragment derived from an antigen. In some embodiments, the target cell is a cell expressing an antigen and presenting said antigen with class I MHC. "Antigen processing" refers to the degradation of an antigen into processing products which are fragments of said antigen (e.g., the degradation of a polypeptide into peptides) and the association of one or more of these fragments (e.g., via binding) with MHC molecules for presentation by cells, such as antigen-presenting cells to specific T-cells.
[0414] By "antigen-responsive CTL" is meant a CD8+T-cell that is responsive to an antigen or a peptide derived from said antigen, which is presented with class I MHC on the surface of antigen presenting cells.
[0415] According to the disclosure, CTL responsiveness may include sustained calcium flux, cell division, production of cytokines such as IFN-y and TNF-a, up-regulation of activation markers such as CD44 and CD69, and specific cytolytic killing of tumor antigen expressing target cells. CTL responsiveness may also be determined using an artificial reporter that accurately indicates CTL responsiveness.
[0416] "Activation" or "stimulation", as used herein, refers to the state of a cell that has been sufficiently stimulated to induce detectable cellular proliferation, such as an immune effector cell such as T cell. Activation can also be associated with initiation of signaling pathways, induced cytokine production, and detectable effector functions. The term "activated immune effector cells" refers to, among other things, immune effector cells that are undergoing cell division.
[0417] The term "priming" refers to a process wherein an immune effector cell such as a T cell has its first contact with its specific antigen and causes differentiation into effector cells such as effector T cells.
[0418] The term "expansion" refers to a process wherein a specific entity is multiplied. In some embodiments, the term is used in the context of an immunological response in which immune effector cells are stimulated by an antigen, proliferate, and the specific immune effector cell recognizing said antigen is amplified. In some embodiments, expansion leads to differentiation of the immune effector cells.
[0419] The terms "immune response" and "immune reaction" are used herein interchangeably in their conventional meaning and refer to an integrated bodily response to an antigen and may refer to a cellular immune response, a humoral immune response, or both. According to the disclosure, the term "immune response to" or "immune response against" with respect to an agent such as an antigen, cell or tissue, relates to an immune response such as a cellular response directed against the agent. An immune response may comprise one or more reactions selected from the group consisting of developing antibodies against one or more antigens and expansion of antigen-specific T-lymphocytes, such as CD4+and CD8+T-lymphocytes, e.g. CD8+T-lymphocytes, which may be detected in various proliferation or cytokine production tests in vitro.
[0420] The terms "inducing an immune response" and "eliciting an immune response" and similar terms in the context of the present disclosure refer to the induction of an immune response, such as the induction of a cellular immune response, a humoral immune response, or both. The immune response may be protective / preventive / prophylactic and / or therapeutic. The immune response may be directed against any immunogen or antigen or antigen peptide, such as against a tumor-associated antigen or a pathogen- associated antigen (e.g., an antigen of a virus (such as influenza virus (A, B, or C), CMV or RSV)). "Inducing" in this context may mean that there was no immune response against a particular antigen or pathogen before induction, but it may also mean that there was a certain level of immune response against a particular antigen or pathogen before induction and after induction said immune response is enhanced. Thus, "inducing the immune response" in this context also includes "enhancing the immune response". In some embodiments, after inducing an immune response in an individual, said individual is protected from developing a disease such as an infectious disease or a cancerous disease or the disease condition is ameliorated by inducing an immune response.
[0421] The terms "cellular immune response", "cellular response", "cell-mediated immunity" or similar terms are meant to include a cellular response directed to cells characterized by expression of an antigen and / or presentation of an antigen with class I or class II MHC. The cellular response relates to cells called T cells or T lymphocytes which act as either "helpers" or "killers". The helper T cells (also termed CD4+T cells) play a central role by regulating the immune response and the killer cells (also termed cytotoxic T cells, cytolytic T cells, CD8+T cells or CTLs) kill cells such as diseased cells.
[0422] The term "humoral immune response" refers to a process in living organisms wherein antibodies are produced in response to agents and organisms, which they ultimately neutralize and / or eliminate. The specificity of the antibody response is mediated by T and / or B cells through membrane-associated receptors that bind antigen of a single specificity. Following binding of an appropriate antigen and receipt of various other activating signals, B lymphocytes divide, which produces memory B cells as well as antibody secreting plasma cell clones, each producing antibodies that recognize the identical antigenic epitope as was recognized by its antigen receptor. Memory B lymphocytes remain dormant until they are subsequently activated by their specific antigen. These lymphocytes provide the cellular basis of memory and the resulting escalation in antibody response when re-exposed to a specific antigen.
[0423] The term "antibody" as used herein, refers to an immunoglobulin molecule, which is able to specifically bind to an epitope on an antigen. In particular, the term "antibody" refers to a glycoprotein comprising at least two heavy (H) chains and two light (L) chains inter-connected by disulfide bonds. The term "antibody" includes monoclonal antibodies, recombinant antibodies, human antibodies, humanized antibodies, chimeric antibodies and combinations of any of the foregoing. Each heavy chain is comprised of a heavy chain variable region (VH) and a heavy chain constant region (CH). Each light chain is comprised of a light chain variable region (VL) and a light chain constant region (CL). The variable regions and constant regions are also referred to herein as variable domains and constant domains, respectively. The VH and VL regions can be further subdivided into regions of hypervariability, termed complementarity determining regions (CDRs), interspersed with regions that are more conserved, termed framework regions (FRs). Each VH and VL is composed of three CDRs and four FRs, arranged from amino-terminus to carboxy-terminus in the following order: FR1 , CDR1 , FR2, CDR2, FR3, CDR3, FR4. The CDRs of a VH are termed HCDR1 , HCDR2 and HCDR3, the CDRs of a VL are termed LCDR1 , LCDR2 and LCDR3. The variable regions of the heavy and light chains contain a binding domain that interacts with an antigen. The constant regions of an antibody comprise the heavy chain constant region (CH) and the light chain constant region (CL), wherein CH can be further subdivided into constant domain CH1 , a hinge region, and constant domains CH2 and CH3 (arranged from amino-terminus to carboxy-terminus in the following order: CH1 , CH2, CH3). The constant regions of the antibodies may mediate the binding of the immunoglobulin to host tissues or factors, including various cells of the immune system (e.g., effector cells) and the first component (C1 q) of the classical complement system. Antibodies can be intact immunoglobulins derived from natural sources or from recombinant sources and can be immunoactive portions of intact immunoglobulins. Antibodies are typically tetramers of immunoglobulin molecules. Antibodies may exist in a variety of forms including, for example, polyclonal antibodies, monoclonal antibodies, Fv, Fab and F(ab)2, as well as single chain antibodies and humanized antibodies. The term "immunoglobulin" relates to proteins of the immunoglobulin superfamily, such as to antigen receptors such as antibodies or the B cell receptor (BCR). The immunoglobulins are characterized by a structural domain, i.e., the immunoglobulin domain, having a characteristic immunoglobulin (Ig) fold. The term encompasses membrane bound immunoglobulins as well as soluble immunoglobulins. Membrane bound immunoglobulins are also termed surface immunoglobulins or membrane immunoglobulins, which are generally part of the BCR. Soluble immunoglobulins are generally termed antibodies. Immunoglobulins generally comprise several chains, typically two identical heavy chains and two identical light chains which are linked via disulfide bonds. These chains are primarily composed of immunoglobulin domains, such as the VL (variable light chain) domain, CL (constant light chain) domain, VH (variable heavy chain) domain, and the CH (constant heavy chain) domains CH1 , CH2, CH3, and CH4. There are five types of mammalian immunoglobulin heavy chains, i.e., a, 6, s, y, and p which account for the different classes of antibodies, i.e., IgA, IgD, IgE, IgG, and IgM. As opposed to the heavy chains of soluble immunoglobulins, the heavy chains of membrane or surface immunoglobulins comprise a transmembrane domain and a short cytoplasmic domain at their carboxy-terminus. In mammals there are two types of light chains, i.e., lambda and kappa. The immunoglobulin chains comprise a variable region and a constant region. The constant region is essentially conserved within the different isotypes of the immunoglobulins, wherein the variable part is highly divers and accounts for antigen recognition.
[0424] The terms "vaccination" and "immunization" describe the process of treating an individual for therapeutic or prophylactic reasons and relate to the procedure of administering one or more immunogen(s) or antigen(s) or derivatives thereof, in particular in the form of RNA (especially mRNA) coding therefor, as described herein to an individual and stimulating an immune response against said one or more immunogen(s) or antigen(s) or cells characterized by presentation of said one or more immunogen(s) or antigen(s).
[0425] By "cell characterized by presentation of an antigen" or "cell presenting an antigen" or "MHO molecules which present an antigen on the surface of an antigen presenting cell" or similar expressions is meant a cell such as a diseased cell, in particular a tumor cell or an infected cell, or an antigen presenting cell presenting the antigen or an antigen peptide, either directly or following processing, in the context of MHC molecules, such as MHC class I and / or MHC class II molecules. In some embodiments, the MHC molecules are MHC class I molecules.
[0426] The term "allergen" refers to a kind of antigen which originates from outside the body of a subject ( / .e., the allergen can also be called "heterologous antigen") and which produces an abnormally vigorous immune response in which the immune system of the subject fights off a perceived threat that would otherwise be harmless to the subject. "Allergies" are the diseases caused by such vigorous immune reactions against allergens. An allergen usually is an antigen which is able to stimulate a type-1 hypersensitivity reaction in atopic individuals through immunoglobulin E (IgE) responses. Particular examples of allergens include allergens derived from peanut proteins (e.g., Ara h 2.02), ovalbumin, grass pollen proteins (e.g., Phi p 5), and proteins of dust mites (e.g., Der p 2).
[0427] The term "growth factors" refers to molecules which are able to stimulate cellular growth, proliferation, healing, and / or cellular differentiation. Typically, growth factors act as signaling molecules between cells. The term "growth factors" include particular cytokines and hormones which bind to specific receptors on the surface of their target cells. Examples of growth factors include bone morphogenetic proteins (BMPs), fibroblast growth factors (FGFs), vascular endothelial growth factors (VEGFs), such as VEGFA, epidermal growth factor (EGF), insulin-like growth factor, ephrins, macrophage colony-stimulating factor, granulocyte colony-stimulating factor, granulocyte macrophage colony-stimulating factor, neuregulins, neurotrophins (e.g., brain-derived neurotrophic factor (BDNF), nerve growth factor (NGF)), placental growth factor (PGF), platelet-derived growth factor (PDGF), renalase (RNLS) (anti- apoptotic survival factor), T-cell growth factor (TCGF), thrombopoietin (TPO), transforming growth factors (transforming growth factor alpha (TGF-a), transforming growth factor beta (TGF-|3)), and tumor necrosis factor-alpha (TNF-a). In some embodiments, a "growth factor" is a peptide or polypeptide growth factor.
[0428] The term "protease inhibitors" refers to molecules, in particular peptides or polypeptides, which inhibit the function of proteases. Protease inhibitors can be classified by the protease which is inhibited (e.g., aspartic protease inhibitors) or by their mechanism of action (e.g., suicide inhibitors, such as serpins). Particular examples of protease inhibitors include serpins, such as alpha 1 -antitrypsin, aprotinin, and bestatin. The term "enzymes" refers to macromolecular biological catalysts which accelerate chemical reactions. Like any catalyst, enzymes are not consumed in the reaction they catalyze and do not alter the equilibrium of said reaction. Unlike many other catalysts, enzymes are much more specific. In some embodiments, an enzyme is essential for homeostasis of a subject, e.g., any malfunction (in particular, decreased activity which may be caused by any of mutation, deletion or decreased production) of the enzyme results in a disease. Examples of enzymes include herpes simplex virus type 1 thymidine kinase (HSV1 -TK), hexosaminidase, phenylalanine hydroxylase, pseudocholinesterase, and lactase.
[0429] The term "receptors" refers to protein molecules which receive signals (in particular chemical signals called ligands) from outside a cell. The binding of a signal (e.g., ligand) to a receptor causes some kind of response of the cell, e.g., the intracellular activation of a kinase. Receptors include transmembrane receptors (such as ion channel-linked (ionotropic) receptors, G protein-linked (metabotropic) receptors, and enzyme-linked receptors) and intracellular receptors (such as cytoplasmic receptors and nuclear receptors). Particular examples of receptors include steroid hormone receptors, growth factor receptors, and peptide receptors ( / .e., receptors whose ligands are peptides), such as P-selectin glycoprotein ligand-1 (PSGL-1 ). The term "growth factor receptors" refers to receptors which bind to growth factors.
[0430] The term "apoptosis regulators" refers to molecules, in particular peptides or polypeptides, which modulate apoptosis, i.e., which either activate or inhibit apoptosis. Apoptosis regulators can be grouped into two broad classes: those which modulate mitochondrial function and those which regulate caspases. The first class includes proteins (e.g., BCL-2, BCL-xL) which act to preserve mitochondrial integrity by preventing loss of mitochondrial membrane potential and / or release of pro-apoptotic proteins such as cytochrome C into the cytosol. Also to this first class belong proapoptotic proteins (e.g., BAX, BAK, BIM) which promote release of cytochrome C. The second class includes proteins such as the inhibitors of apoptosis proteins (e.g., XIAP) or FLIP which block the activation of caspases.
[0431] The term "transcription factors" relates to proteins which regulate the rate of transcription of genetic information from DNA to messenger RNA, in particular by binding to a specific DNA sequence. Transcription factors may regulate cell division, cell growth, and cell death throughout life; cell migration and organization during embryonic development; and / or in response to signals from outside the cell, such as a hormone. Transcription factors contain at least one DNA-binding domain which binds to a specific DNA sequence, usually adjacent to the genes which are regulated by the transcription factors. Particular examples of transcription factors include MECP2, FOXP2, FOXP3, the STAT protein family, and the HOX protein family.
[0432] The term "tumor suppressor proteins" relates to molecules, in particular peptides or polypeptides, which protect a cell from one step on the path to cancer. Tumorsuppressor proteins (usually encoded by corresponding tumor-suppressor genes) exhibit a weakening or repressive effect on the regulation of the cell cycle and / or promote apoptosis. Their functions may be one or more of the following: repression of genes essential for the continuing of the cell cycle; coupling the cell cycle to DNA damage (as long as damaged DNA is present in a cell, no cell division should take place); initiation of apoptosis, if the damaged DNA cannot be repaired; metastasis suppression (e.g., preventing tumor cells from dispersing, blocking loss of contact inhibition, and inhibiting metastasis); and DNA repair. Particular examples of tumorsuppressor proteins include p53, phosphatase and tensin homolog (PTEN), SWI / SNF (SWItch / Sucrose Non-Fermentable), von Hippel-Lindau tumor suppressor (pVHL), adenomatous polyposis coli (APC), CD95, suppression of tumorigenicity 5 (ST5), suppression of tumorigenicity 5 (ST5), suppression of tumorigenicity 14 (STM), and Yippee-like 3 (YPEL3).
[0433] The term "structural proteins" refers to proteins which confer stiffness and rigidity to otherwise-fluid biological components. Structural proteins are mostly fibrous (such as collagen and elastin) but may also be globular (such as actin and tubulin). Usually, globular proteins are soluble as monomers, but polymerize to form long, fibers which, for example, may make up the cytoskeleton. Other structural proteins are motor proteins (such as myosin, kinesin, and dynein) which are capable of generating mechanical forces, and surfactant proteins. Particular examples of structural proteins include collagen, surfactant protein A, surfactant protein B, surfactant protein C, surfactant protein D, elastin, tubulin, actin, and myosin.
[0434] The term "reprogramming factors" or "reprogramming transcription factors" relates to molecules, in particular peptides or polypeptides, which, when expressed in somatic cells optionally together with further agents such as further reprogramming factors, lead to reprogramming or de-differentiation of said somatic cells to cells having stem cell characteristics, in particular pluripotency. Particular examples of reprogramming factors include 0CT4, S0X2, c-MYC, KLF4, LIN28, and NANOG. The term "genomic engineering proteins" relates to proteins which are able to insert, delete or replace DNA in the genome of a subject. Particular examples of genomic engineering proteins include meganucleases, zinc finger nucleases (ZFNs), transcription activator-like effector nucleases (TALENs), and clustered regularly spaced short palindromic repeat-CRISPR-associated protein 9 (CRISPR-Cas9).
[0435] The term "blood proteins" relates to peptides or polypeptides which are present in blood plasma of a subject, in particular blood plasma of a healthy subject. Blood proteins have diverse functions such as transport (e.g., albumin, transferrin), enzymatic activity (e.g., thrombin or ceruloplasmin), blood clotting (e.g., fibrinogen), defense against pathogens (e.g., complement components and immunoglobulins), protease inhibitors (e.g., alpha 1 -antitrypsin), etc. Particular examples of blood proteins include thrombin, serum albumin, Factor VII, Factor VIII, insulin, Factor IX, Factor X, tissue plasminogen activator, protein C, von Willebrand factor, antithrombin III, glucocerebrosidase, erythropoietin, granulocyte colony stimulating factor (G-CSF), modified Factor VIII, and anticoagulants.
[0436] Thus, in some embodiments, the pharmaceutically active peptide or polypeptide is (i) a cytokine, preferably selected from the group consisting of erythropoietin (EPO), interleukin 4 (IL-2), and interleukin 10 (IL-11 ), more preferably EPO; (ii) an adhesion molecule, in particular an integrin; (iii) an immunoglobulin, in particular an antibody; (iv) an immunologically active compound, in particular an antigen; (v) a hormone, in particular vasopressin, insulin or growth hormone; (vi) a growth factor, in particular VEGFA; (vii) a protease inhibitor, in particular alpha 1 -antitrypsin; (viii) an enzyme, preferably selected from the group consisting of herpes simplex virus type 1 thymidine kinase (HSV1 -TK), hexosaminidase, phenylalanine hydroxylase, pseudocholinesterase, pancreatic enzymes, and lactase; (ix) a receptor, in particular growth factor receptors; (x) an apoptosis regulator, in particular BAX; (xi) a transcription factor, in particular FOXP3; (xii) a tumor suppressor protein, in particular p53; (xiii) a structural protein, in particular surfactant protein B; (xiv) a reprogramming factor, e.g., selected from the group consisting of OCT4, SOX2, c-MYC, KLF4, LIN28 and NANOG; (xv) a genomic engineering protein, in particular clustered regularly spaced short palindromic repeat-CRISPR-associated protein 9 (CRISPR-Cas9); and (xvi) a blood protein, in particular fibrinogen.
[0437] In some embodiments, a pharmaceutically active peptide or polypeptide comprises one or more antigens or one or more epitopes, i.e., administration of the peptide or polypeptide to a subject elicits an immune response against the one or more antigens or one or more epitopes in a subject which may be therapeutic or partially or fully protective.
[0438] In some embodiments, the nucleic acid such as mRNA encodes at least one epitope. In some embodiments, the epitope is derived from a tumor antigen. The tumor antigen may be a "standard" antigen, which is generally known to be expressed in various cancers. The tumor antigen may also be a "neo-antigen", which is specific to an individual’s tumor and has not been previously recognized by the immune system. A neo-antigen or neo-epitope may result from one or more cancer-specific mutations in the genome of cancer cells resulting in amino acid changes. Examples of tumor antigens include, without limitation, p53, ART-4, BAGE, beta-catenin / m, Bcr-abL CAMEL, CAP-1 , CASP-8, CDC27 / m, CDK4 / m, CEA, the cell surface proteins of the claudin family, such as CLAUDIN-6, CLAUDIN-18.2 and CLAUDIN-12, c-MYC, CT, Cyp-B, DAM, ELF2M, ETV6-AML1 , G250, GAGE, GnT-V, Gap 100, HAGE, HER- 2 / neu, HPV-E7, HPV-E6, HAST-2, hTERT (or hTRT), LAGE, LDLR / FUT, MAGE-A, preferably MAGE-A1 , MAGE-A2, MAGE- A3, MAGE-A4, MAGE-A5, MAGE-A6, MAGE-A7, MAGE-A8, MAGE-A9, MAGE-A 10, MAGE-A 1 1 , or MAGE- A12, MAGE- B, MAGE-C, MART- 1 / Melan-A, MC1 R, Myosin / m, MUC1 , MUM-1 , MUM-2, MUM-3, NA88-A, NF1 , NY-ESO-1 , NY-BR-1 , pl90 minor BCR-abL, Pml / RARa, PRAME, proteinase 3, PSA, PSM, RAGE, RU1 or RU2, SAGE, SART-1 or SART-3, SCGB3A2, SCP1 , SCP2, SCP3, SSX, SURVIVIN, TEL / AML1 , TPI / m, TRP-1 , TRP-2, TRP- 2 / INT2, TPTE, WT, and WT-1.
[0439] Cancer mutations vary with each individual. Thus, cancer mutations that encode novel epitopes (neo-epitopes) represent attractive targets in the development of vaccine compositions and immunotherapies. The efficacy of tumor immunotherapy relies on the selection of cancer-specific antigens and epitopes capable of inducing a potent immune response within a host. RNA can be used to deliver patient-specific tumor epitopes to a patient. Dendritic cells (DCs) residing in the spleen represent antigen- presenting cells of particular interest for RNA expression of immunogenic epitopes or antigens such as tumor epitopes. The use of multiple epitopes has been shown to promote therapeutic efficacy in tumor vaccine compositions. Rapid sequencing of the tumor mutanome may provide multiple epitopes for individualized vaccines which can be encoded by mRNA described herein, e.g., as a single polypeptide wherein the epitopes are optionally separated by linkers. In some embodiments of the present disclosure, the mRNA encodes at least one epitope, at least two epitopes, at least three epitopes, at least four epitopes, at least five epitopes, at least six epitopes, at least seven epitopes, at least eight epitopes, at least nine epitopes, or at least ten epitopes. Exemplary embodiments include mRNA that encodes at least five epitopes (termed a "pentatope") and mRNA that encodes at least ten epitopes (termed a "decatope").
[0440] In some embodiments, the antigen or epitope is derived from a pathogen-associated antigen, in particular from a viral antigen. In some embodiments, the antigen or epitope is derived from a SARS-CoV-2 S protein, an immunogenic variant thereof, or an immunogenic fragment of the SARS-CoV-2 S protein or the immunogenic variant thereof. Thus, in some embodiments, the mRNA used in the present disclosure encodes an amino acid sequence comprising a SARS-CoV-2 S protein, an immunogenic variant thereof, or an immunogenic fragment of the SARS-CoV-2 S protein or the immunogenic variant thereof.
[0441] In some embodiments of the present disclosure the antigen (such as a tumor antigen or vaccine antigen) is preferably administered as single-stranded, 5' capped mRNA that is translated into the respective protein upon entering cells of a subject being administered the RNA. Preferably, the RNA contains structural elements optimized for maximal efficacy of the RNA with respect to stability and translational efficiency (5' cap, 5' UTR, 3' UTR, poly(A) sequence).
[0442] In some embodiments, beta-S-ARCA(DI ) is utilized as specific capping structure at the 5'-end of the mRNA. In some embodiments, m27’3’’°Gppp(mi2’’0) ApG is utilized as specific capping structure at the 5'-end of the mRNA. In some embodiments, the 5'- UTR sequence is derived from the human alpha-globin mRNA and optionally has an optimized 'Kozak sequence' to increase translational efficiency. In some embodiments, a combination of two sequence elements (Fl element) derived from the "amino terminal enhancer of split" (AES) mRNA (called F) and the mitochondrial encoded 12S ribosomal RNA (called I) are placed between the coding sequence and the poly(A) sequence to assure higher maximum protein levels and prolonged persistence of the mRNA. In some embodiments, two re-iterated 3'-UTRs derived from the human betaglobin mRNA are placed between the coding sequence and the poly(A) sequence to assure higher maximum protein levels and prolonged persistence of the mRNA. In some embodiments, a poly(A) sequence measuring 110 nucleotides in length, consisting of a stretch of 30 adenosine residues, followed by a 10 nucleotide linker sequence and another 70 adenosine residues is used. This poly(A) sequence was designed to enhance RNA stability and translational efficiency.
[0443] In some embodiments, mRNA encoding an antigen (such as a tumor antigen or a vaccine antigen) is expressed in cells of the subject treated to provide the antigen. In some embodiments, the mRNA is transiently expressed in cells of the subject. In some embodiments, the mRNA is in vitro transcribed. In some embodiments, expression of the antigen is at the cell surface. In some embodiments, the antigen is expressed and presented in the context of MHC. In some embodiments, expression of the antigen is into the extracellular space, i.e., the antigen is secreted.
[0444] The antigen molecule or a procession product thereof, e.g., a fragment thereof, may bind to an antigen receptor such as a BCR or TCR carried by immune effector cells, or to antibodies.
[0445] A peptide and polypeptide antigen which is provided to a subject according to the present disclosure by administering mRNA encoding a peptide and polypeptide antigen, wherein the antigen is a vaccine antigen, preferably results in the induction of an immune response, e.g., a humoral and / or cellular immune response in the subject being provided the peptide or polypeptide antigen. Said immune response is preferably directed against a target antigen. Thus, a vaccine antigen may comprise the target antigen, a variant thereof, or a fragment thereof. In some embodiments, such fragment or variant is immunologically equivalent to the target antigen. In the context of the present disclosure, the term "fragment of an antigen" or "variant of an antigen" means an agent which results in the induction of an immune response which immune response targets the antigen, i.e. a target antigen. Thus, the vaccine antigen may correspond to or may comprise the target antigen, may correspond to or may comprise a fragment of the target antigen or may correspond to or may comprise an antigen which is homologous to the target antigen or a fragment thereof. Thus, according to the present disclosure, a vaccine antigen may comprise an immunogenic fragment of a target antigen or an amino acid sequence being homologous to an immunogenic fragment of a target antigen. An "immunogenic fragment of an antigen" according to the disclosure preferably relates to a fragment of an antigen which is capable of inducing an immune response against the target antigen. The vaccine antigen may be a recombinant antigen.
[0446] The term "immunologically equivalent" means that the immunologically equivalent molecule such as the immunologically equivalent amino acid sequence exhibits the same or essentially the same immunological properties and / or exerts the same or essentially the same immunological effects, e.g., with respect to the type of the immunological effect. In the context of the present disclosure, the term "immunologically equivalent" is preferably used with respect to the immunological effects or properties of antigens or antigen variants used for immunization. For example, an amino acid sequence is immunologically equivalent to a reference amino acid sequence if said amino acid sequence when exposed to the immune system of a subject induces an immune reaction having a specificity of reacting with the reference amino acid sequence.
[0447] In some embodiments, the mRNA used in the present disclosure is non-immunogenic. RNA encoding an immunostimulant may be administered according to the present disclosure to provide an adjuvant effect. The RNA encoding an immunostimulant may be standard RNA or non-immunogenic RNA.
[0448] The term "non-immunogenic RNA" (such as "non-immunogenic mRNA") as used herein refers to RNA that does not induce a response by the immune system upon administration, e.g., to a mammal, or induces a weaker response than would have been induced by the same RNA that differs only in that it has not been subjected to the modifications and treatments that render the non-immunogenic RNA non- immunogenic, i.e., than would have been induced by standard RNA (stdRNA). In certain embodiments, non-immunogenic RNA, which is also termed modified RNA (modRNA) herein, is rendered non-immunogenic by incorporating modified nucleosides suppressing RNA-mediated activation of innate immune receptors into the RNA and / or removing double-stranded RNA (dsRNA).
[0449] For rendering the non-immunogenic RNA (especially mRNA) non-immunogenic by the incorporation of modified nucleosides, any modified nucleoside may be used as long as it lowers or suppresses immunogenicity of the RNA. Particularly preferred are modified nucleosides that suppress RNA-mediated activation of innate immune receptors. In some embodiments, the modified nucleosides comprise a replacement of one or more uridines with a nucleoside comprising a modified nucleobase. In some embodiments, the modified nucleobase is a modified uracil. In some embodiments, the nucleoside comprising a modified nucleobase is selected from the group consisting of 3-methyl-uridine (m3U), 5-methoxy-uridine (mo5U), 5-aza-uridine, 6-aza-uridine, 2-thio- 5-aza-uridine, 2-thio-uridine (s2U), 4-thio-uridine (s4U), 4-thio-pseudouridine, 2-thio- pseudouridine, 5-hydroxy-uridine (ho5U), 5-aminoallyl-uridine, 5-halo-uridine (e.g., 5- iodo-uridine or 5-bromo-uridine), uridine 5-oxyacetic acid (cmo5U), undine 5-oxyacetic acid methyl ester (mcmo5U), 5-carboxymethyl-uridine (cm5U), 1 -carboxymethyl- pseudouridine, 5-carboxyhydroxymethyl-uridine (chm5U), 5-carboxyhydroxymethyl- uridine methyl ester (mchm5U), 5-methoxycarbonylmethyl-uridine (mcm5U), 5- methoxycarbonylmethyl-2 -thio-uridine (mcm5s2U), 5-aminomethyl-2 -thio-uridine (nm5s2U), 5-methylaminomethyl-uridine (mnm5U), 1 -ethyl-pseudouridine, 5- methylaminomethyl-2-thio-uridine (mnm5s2U), 5-methylaminomethyl-2-seleno-uridine (mnm5se2U), 5-carbamoylmethyl-uridine (ncm5U), 5-carboxymethylaminomethyl- uridine (cmnm5U), 5-carboxymethylaminomethyl-2-thio-uridine (cmnm5s2U), 5- propynyl-uridine, 1-propynyl-pseudouridine, 5-taurinomethyl-uridine (im5U), 1 - taurinomethyl-pseudouridine, 5-taurinomethyl-2-thio-uridine(Tm5s2U), 1 - taurinomethyl-4-thio-pseudouridine), 5-methyl-2 -thio-uridine (m5s2U), 1 -methyl-4-thio- pseudouridine (m1s4ip), 4-thio-1 -methyl-pseudouridine, 3-methyl-pseudouridine (m3ip), 2-th io-1 -methyl-pseudouridine, 1 -methyl-1 -deaza-pseudouridine, 2-th io-1 -methyl-1 - deaza-pseudouridine, dihydrouridine (D), dihydropseudouridine, 5,6-dihydrouridine, 5- methyl-dihydrouridine (m5D), 2-thio-dihydrouridine, 2-thio-dihydropseudouridine, 2- methoxy-uridine, 2-methoxy-4-thio-uridine, 4-methoxy-pseudouridine, 4-methoxy-2- thio-pseudouridine, N1 -methyl-pseudouridine, 3-(3-amino-3-carboxypropyl)uridine (acp3U), 1-methyl-3-(3-amino-3-carboxypropyl)pseudouridine (acp3ip), 5- (isopentenylaminomethyl)uridine (inm5U), 5-(isopentenylaminomethyl)-2 -thio-uridine (inm5s2U), a-thio-uridine, 2'-O-methyl-uridine (Um), 5,2 -O-dimethyl-uridine (m5Um), 2'- O-methyl-pseudouridine (ipm), 2-thio-2'-O-methyl-uridine (s2Um), 5- methoxycarbonylmethyl-2'-O-methyl-uridine (mcm5Um), 5-carbamoylmethyl-2'-O- methyl-uridine (ncm5Um), 5-carboxymethylaminomethyl-2'-O-methyl-uridine (cmnm5Um), 3,2 -O-dimethyl-uridine (m3Um), 5-(isopentenylaminomethyl)-2'-O- methyl-uridine (inm5Um), 1 -thio-uridine, deoxythymidine, 2'-F-ara-uridine, 2'-F-uridine, 2'-OH-ara-uridine, 5-(2-carbomethoxyvinyl) undine, and 5-[3-(1 -E- propenylamino)uridine. In certain embodiments, the nucleoside comprising a modified nucleobase is pseudouridine (ip), N1 -methyl-pseudouridine (m1 ip) or 5-methyl-uridine (m5U), in particular N1 -methyl-pseudouridine.
[0450] In some embodiments, the replacement of one or more uridines with a nucleoside comprising a modified nucleobase comprises a replacement of at least 1 %, at least 2%, at least 3%, at least 4%, at least 5%, at least 10%, at least 25%, at least 50%, at least 75%, at least 90%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99% or 100% of the uridines.
[0451] During synthesis of mRNA by in vitro transcription (IVT) using T7 RNA polymerase significant amounts of aberrant products, including double-stranded RNA (dsRNA) are produced due to unconventional activity of the enzyme. dsRNA induces inflammatory cytokines and activates effector enzymes leading to protein synthesis inhibition. dsRNA can be removed from RNA such as IVT RNA, for example, by ion-pair reversed phase HPLC using a non-porous or porous C-18 polystyrene-divinylbenzene (PS- DVB) matrix. Alternatively, an enzymatic based method using E. coli RNaselll that specifically hydrolyzes dsRNA but not ssRNA, thereby eliminating dsRNA contaminants from IVT RNA preparations can be used. Furthermore, dsRNA can be separated from ssRNA by using a cellulose material. In some embodiments, an RNA preparation is contacted with a cellulose material and the ssRNA is separated from the cellulose material under conditions which allow binding of dsRNA to the cellulose material and do not allow binding of ssRNA to the cellulose material. Suitable methods for providing ssRNA are disclosed, for example, in WO 2017 / 182524.
[0452] As the term is used herein, "remove" or "removal" refers to the characteristic of a population of first substances, such as non-immunogenic RNA, being separated from the proximity of a population of second substances, such as dsRNA, wherein the population of first substances is not necessarily devoid of the second substance, and the population of second substances is not necessarily devoid of the first substance. However, a population of first substances characterized by the removal of a population of second substances has a measurably lower content of second substances as compared to the non-separated mixture of first and second substances.
[0453] In some embodiments, the removal of dsRNA (especially mRNA) from non- immunogenic RNA comprises a removal of dsRNA such that less than 10%, less than 5%, less than 4%, less than 3%, less than 2%, less than 1 %, less than 0.5%, less than 0.3%, or less than 0.1 % of the RNA in the non-immunogenic RNA composition is dsRNA. In some embodiments, the non-immunogenic RNA (especially mRNA) is free or essentially free of dsRNA. In some embodiments, the non-immunogenic RNA (especially mRNA) composition comprises a purified preparation of single-stranded nucleoside modified RNA. For example, in some embodiments, the purified preparation of single-stranded nucleoside modified RNA (especially mRNA) is substantially free of double stranded RNA (dsRNA). In some embodiments, the purified preparation is at least 90%, at least 91 %, at least 92%, at least 93 %, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99%, at least 99.5%, or at least 99.9% single stranded nucleoside modified RNA, relative to all other nucleic acid molecules (DNA, dsRNA, etc.).
[0454] In some embodiments, the non-immunogenic RNA (especially mRNA) is translated in a cell more efficiently than standard RNA with the same sequence. In some embodiments, translation is enhanced by a factor of 2-fold relative to its unmodified counterpart. In some embodiments, translation is enhanced by a 3-fold factor. In some embodiments, translation is enhanced by a 4-fold factor. In some embodiments, translation is enhanced by a 5-fold factor. In some embodiments, translation is enhanced by a 6-fold factor. In some embodiments, translation is enhanced by a 7-fold factor. In some embodiments, translation is enhanced by an 8-fold factor. In some embodiments, translation is enhanced by a 9-fold factor. In some embodiments, translation is enhanced by a 10-fold factor. In some embodiments, translation is enhanced by a 15-fold factor. In some embodiments, translation is enhanced by a 20- fold factor. In some embodiments, translation is enhanced by a 50-fold factor. In some embodiments, translation is enhanced by a 100-fold factor. In some embodiments, translation is enhanced by a 200-fold factor. In some embodiments, translation is enhanced by a 500-fold factor. In some embodiments, translation is enhanced by a 1000-fold factor. In some embodiments, translation is enhanced by a 2000-fold factor. In some embodiments, the factor is 10-1000-fold. In some embodiments, the factor is 10-100-fold. In some embodiments, the factor is 10-200-fold. In some embodiments, the factor is 10-300-fold. In some embodiments, the factor is 10-500-fold. In some embodiments, the factor is 20-1000-fold. In some embodiments, the factor is 30-1000- fold. In some embodiments, the factor is 50-1000-fold. In some embodiments, the factor is 100-1000-fold. In some embodiments, the factor is 200-1000-fold. In some embodiments, translation is enhanced by any other significant amount or range of amounts.
[0455] In some embodiments, the non-immunogenic RNA (especially mRNA) exhibits significantly less innate immunogenicity than standard RNA with the same sequence. In some embodiments, the non-immunogenic RNA (especially mRNA) exhibits an innate immune response that is 2-fold less than its unmodified counterpart. In some embodiments, innate immunogenicity is reduced by a 3-fold factor. In some embodiments, innate immunogenicity is reduced by a 4-fold factor. In some embodiments, innate immunogenicity is reduced by a 5-fold factor. In some embodiments, innate immunogenicity is reduced by a 6-fold factor. In some embodiments, innate immunogenicity is reduced by a 7-fold factor. In some embodiments, innate immunogenicity is reduced by a 8-fold factor. In some embodiments, innate immunogenicity is reduced by a 9-fold factor. In some embodiments, innate immunogenicity is reduced by a 10-fold factor. In some embodiments, innate immunogenicity is reduced by a 15-fold factor. In some embodiments, innate immunogenicity is reduced by a 20-fold factor. In some embodiments, innate immunogenicity is reduced by a 50-fold factor. In some embodiments, innate immunogenicity is reduced by a 100-fold factor. In some embodiments, innate immunogenicity is reduced by a 200-fold factor. In some embodiments, innate immunogenicity is reduced by a 500-fold factor. In some embodiments, innate immunogenicity is reduced by a 1000-fold factor. In some embodiments, innate immunogenicity is reduced by a 2000-fold factor.
[0456] The term "exhibits significantly less innate immunogenicity" refers to a detectable decrease in innate immunogenicity. In some embodiments, the term refers to a decrease such that an effective amount of the non-immunogenic RNA (especially mRNA) can be administered without triggering a detectable innate immune response. In some embodiments, the term refers to a decrease such that the non-immunogenic RNA (especially mRNA) can be repeatedly administered without eliciting an innate immune response sufficient to detectably reduce production of the protein encoded by the non-immunogenic RNA. In some embodiments, the decrease is such that the non- immunogenic RNA (especially mRNA) can be repeatedly administered without eliciting an innate immune response sufficient to eliminate detectable production of the protein encoded by the non-immunogenic RNA.
[0457] "Immunogenicity" is the ability of a foreign substance, such as RNA, to provoke an immune response in the body of a human or other animal. The innate immune system is the component of the immune system that is relatively unspecific and immediate. It is one of two main components of the vertebrate immune system, along with the adaptive immune system.
[0458] Particles
[0459] Nucleic acids (such as RNA and / or DNA, in particular mRNA) described herein may be present in particles comprising (i) the nucleic acid, and (ii) at least one cationic or cation ical ly ionizable compound such as a polymer or lipid complexing the nucleic acid. Electrostatic interactions between positively charged molecules such as polymers and lipids and negatively charged nucleic acid are involved in particle formation. This results in complexation and spontaneous formation of nucleic acid particles.
[0460] Different types of RNA containing particles have been described previously to be suitable for delivery of RNA in particulate form (cf., e.g., Kaczmarek, J. C. et al., 2017, Genome Medicine 9, 60). For non-viral RNA delivery vehicles, nanoparticle encapsulation of RNA physically protects RNA from degradation and, depending on the specific chemistry, can aid in cellular uptake and endosomal escape.
[0461] In the context of the present disclosure, the term "particle" relates to a structured entity formed by molecules or molecule complexes, in particular particle forming compounds. In some embodiments, the particle contains an envelope (e.g., one or more layers or lamellas) made of one or more types of amphiphilic substances (e.g., amphiphilic lipids). In this context, the expression "amphiphilic substance" means that the substance possesses both hydrophilic and lipophilic properties. The envelope may also comprise additional substances (e.g., additional lipids) which do not have to be amphiphilic. Thus, the particle may be a monolamellar or multilamellar structure, wherein the substances constituting the one or more layers or lamellas comprise one or more types of amphiphilic substances (in particular selected from the group consisting of amphiphilic lipids) optionally in combination with additional substances (e.g., additional lipids) which do not have to be amphiphilic. In some embodiments, the term "particle" relates to a micro- or nano-sized structure, such as a micro- or nanosized compact structure. According to the present disclosure, the term "particle" includes nanoparticles.
[0462] An "RNA particle" can be used to deliver RNA to a target site of interest (e.g., cell, tissue, organ, and the like). An RNA particle may be formed from lipids comprising at least one cationic or cationically ionizable lipid or lipid-like material. Without intending to be bound by any theory, it is believed that the cationic or cationically ionizable lipid or lipid-like material combines together with the RNA to form aggregates, and this aggregation results in colloidally stable particles.
[0463] Nucleic acid particles (such RNA particles, DNA particles or DNA / RNA particles) described herein include lipid nanoparticle (LNP)-based and lipoplex (LPX)-based formulations. In general, a lipoplex (LPX) is obtainable from mixing two aqueous phases, namely a phase comprising nucleic acid (such as RNA and / or DNA) and a phase comprising a dispersion of lipids. In some embodiments, the lipid phase comprises liposomes.
[0464] In some embodiments, liposomes are self-closed unilamellar or multilamellar vesicular particles wherein the lamellae comprise lipid bilayers and the encapsulated lumen comprises an aqueous phase. A prerequisite for using liposomes for nanoparticle formation is that the lipids in the mixture as required are able to form lamellar (bilayer) phases in the applied aqueous environment.
[0465] In some embodiments, liposomes comprise unilamellar or multilamellar phospholipid bilayers enclosing an aqueous core (also referred to herein as an aqueous lumen). They may be prepared from materials possessing polar head (hydrophilic) groups and nonpolar tail (hydrophobic) groups. In some embodiments, cationic lipids employed in formulating liposomes designed for the delivery of nucleic acids are amphiphilic in nature and consist of a positively charged (cationic) amine head group linked to a hydrocarbon chain or cholesterol derivative via glycerol.
[0466] In some embodiments, lipoplexes are multilamellar liposome-based formulations that form upon electrostatic interaction of cationic liposomes with nucleic acids (such as RNAs and / or DNAs). In some embodiments, formed lipoplexes possess distinct internal arrangements of molecules that arise due to the transformation from liposomal structure into compact nucleic acid-lipoplexes (such as RNA- and / or DNA-lipoplexes). In some embodiments, these formulations are characterized by their poor encapsulation of the nucleic acid (such as RNA) and incomplete entrapment of the nucleic acid (such as RNA).
[0467] In some embodiments, an LPX particle comprises an amphiphilic lipid, in particular cationic or cationically ionizable amphiphilic lipid, and nucleic acid (such as RNA and / or DNA, especially mRNA) as described herein. In some embodiments, electrostatic interactions between positively charged liposomes (made from one or more amphiphilic lipids, in particular cationic or cationically ionizable amphiphilic lipids) and negatively charged nucleic acid (especially mRNA) results in complexation and spontaneous formation of nucleic acid lipoplex particles. Positively charged liposomes may be generally synthesized using a cationic or cationically ionizable amphiphilic lipid, such as DOTMA and / or DODMA, and additional lipids, such as DOPE. In some embodiments, a nucleic acid (such as RNA and / or DNA, especially mRNA) lipoplex particle is a nanoparticle. In general, a lipid nanoparticle (LNP) is obtainable from direct mixing of nucleic acid (such as RNA and / or DNA) in an aqueous phase with lipids in a phase comprising an organic solvent, such as ethanol. In that case, lipids or lipid mixtures can be used for particle formation, which do not form lamellar (bilayer) phases in water.
[0468] In some embodiments, LNPs comprise or consist of a cationic / ionizable lipid and helper lipids such as phospholipids, cholesterol, and / or polyethylene glycol (PEG) lipids. In some embodiments, in the nucleic acid LNPs (such as RNA LNPs, e.g., mRNA LNPs) described herein the nucleic acid (such as RNA, e.g., mRNA) is bound by ionizable lipid that occupies the central core of the LNP. In some embodiments, PEG lipid forms the surface of the LNP, along with phospholipids. In some embodiments, the surface comprises a bilayer. In some embodiments, cholesterol and ionizable lipid in charged and uncharged forms can be distributed throughout the LNP. In some embodiments, nucleic acid (such as RNA and / or DNA, e.g., mRNA) may be noncovalently associated with a particle as described herein. In embodiments, the nucleic acid (such as RNA and / or DNA, especially mRNA) may be adhered to the outer surface of the particle (surface nucleic acid (such as surface RNA, especially surface mRNA)) and / or may be contained in the particle (encapsulated nucleic acid (such as encapsulated RNA, especially encapsulated mRNA)).
[0469] In some embodiments, the particles (e.g., LNPs and LPXs) described herein have a size (such as a diameter) in the range of about 10 to about 2000 nm, such as at least about 15 nm (e.g., at least about 20 nm, at least about 25 nm, at least about 30 nm, at least about 35 nm, at least about 40 nm, at least about 45 nm, at least about 50 nm, at least about 55 nm, at least about 60 nm, at least about 65 nm, at least about 70 nm, at least about 75 nm, at least about 80 nm, at least about 85 nm, at least about 90 nm, at least about 95 nm, or at least about 100 nm) and / or at most 1900 nm (e.g., at most about 1900 nm, at most about 1800 nm, at most about 1700 nm, at most about 1600 nm, at most about 1500 nm, at most about 1400 nm, at most about 1300 nm, at most about 1200 nm, at most about 1100 nm, at most about 1000 nm, at most about 950 nm, at most about 900 nm, at most about 850 nm, at most about 800 nm, at most about 750 nm, at most about 700 nm, at most about 650 nm, at most about 600 nm, at most about 550 nm, or at most about 500 nm), such as in the range of about 20 to about 1500 nm, such as about 30 to about 1200 nm, about 40 to about 1100 nm, about 50 to about 1000 nm, about 60 to about 900 nm, about 70 to 800 nm, about 80 to 700 nm, about 90 to 600 nm, or about 50 to 500 nm or about 100 to 500 nm, such as in the range of 10 to 1000 nm, 15 to 500 nm, 20 to 450 nm, 25 to 400 nm, 30 to 350 nm, 40 to 300 nm, 50 to 250 nm, 60 to 200 nm, or 70 to 150 nm.
[0470] In some embodiments, the particles (e.g., LNPs and LPXs) described herein have an average diameter that in some embodiments ranges from about 50 nm to about 1000 nm, from about 50 nm to about 800 nm, from about 50 nm to about 700 nm, from about 50 nm to about 600 nm, from about 50 nm to about 500 nm, from about 50 nm to about 450 nm, from about 50 nm to about 400 nm, from about 50 nm to about 350 nm, from about 50 nm to about 300 nm, from about 50 nm to about 250 nm, from about 50 nm to about 200 nm, from about 100 nm to about 1000 nm, from about 100 nm to about 800 nm, from about 100 nm to about 700 nm, from about 100 nm to about 600 nm, from about 100 nm to about 500 nm, from about 100 nm to about 450 nm, from about 100 nm to about 400 nm, from about 100 nm to about 350 nm, from about 100 nm to about 300 nm, from about 100 nm to about 250 nm, from about 100 nm to about 200 nm, from about 150 nm to about 1000 nm, from about 150 nm to about 800 nm, from about 150 nm to about 700 nm, from about 150 nm to about 600 nm, from about 150 nm to about 500 nm, from about 150 nm to about 450 nm, from about 150 nm to about 400 nm, from about 150 nm to about 350 nm, from about 150 nm to about 300 nm, from about 150 nm to about 250 nm, from about 150 nm to about 200 nm, from about 200 nm to about 1000 nm, from about 200 nm to about 800 nm, from about 200 nm to about 700 nm, from about 200 nm to about 600 nm, from about 200 nm to about 500 nm, from about 200 nm to about 450 nm, from about 200 nm to about 400 nm, from about 200 nm to about 350 nm, from about 200 nm to about 300 nm, or from about 200 nm to about 250 nm.
[0471] In some embodiments, the particles described herein are nanoparticles. The term "nanoparticle" relates to a nano-sized particle comprising nucleic acid (especially mRNA) as described herein and at least one cationic or cationically ionizable lipid, wherein all three external dimensions of the particle are in the nanoscale, i.e., at least about 1 nm and below about 1000 nm. Preferably, the size of a particle is its diameter. Nucleic acid particles described herein (especially mRNA particles) may exhibit a polydispersity index (PDI) less than about 0.5, less than about 0.4, less than about 0.3, less than about 0.2, less than about 0.1 , or less than about 0.05. By way of example, the nucleic acid particles can exhibit a polydispersity index in a range of about 0.01 to about 0.4 or about 0.1 to about 0.3. The N / P ratio gives the ratio of the nitrogen groups in the lipid to the number of phosphate groups in the nucleic acid. It is correlated to the charge ratio, as the nitrogen atoms (depending on the pH) are usually positively charged and the phosphate groups are negatively charged. The N / P ratio, where a charge equilibrium exists, depends on the pH. Lipid formulations are frequently formed at N / P ratios larger than four up to twelve, because positively charged nanoparticles are considered favorable for transfection. In that case, RNA is considered to be completely bound to nanoparticles. Nucleic acid particles (especially RNA particles such as mRNA particles) described herein can be prepared using a wide range of methods that may involve obtaining a colloid from at least one cationic or cationically ionizable lipid and mixing the colloid with nucleic acid to obtain nucleic acid particles.
[0472] The term "colloid" as used herein relates to a type of homogeneous mixture in which dispersed particles do not settle out. The insoluble particles in the mixture are microscopic, with particle sizes between 1 and 1000 nanometers. The mixture may be termed a colloid or a colloidal suspension. Sometimes the term "colloid" only refers to the particles in the mixture and not the entire suspension.
[0473] For the preparation of colloids comprising at least one cationic or cationically ionizable lipid methods are applicable herein that are conventionally used for preparing liposomal vesicles and are appropriately adapted. The most commonly used methods for preparing liposomal vesicles share the following fundamental stages: (i) lipids dissolution in organic solvents, (ii) drying of the resultant solution, and (iii) hydration of dried lipid (using various aqueous media).
[0474] In the film hydration method, lipids are firstly dissolved in a suitable organic solvent, and dried down to yield a thin film at the bottom of the flask. The obtained lipid film is hydrated using an appropriate aqueous medium to produce a liposomal dispersion. Furthermore, an additional downsizing step may be included.
[0475] Reverse phase evaporation is an alternative method to the film hydration for preparing liposomal vesicles that involves formation of a water-in-oil emulsion between an aqueous phase and an organic phase containing lipids. A brief sonication of this mixture is required for system homogenization. The removal of the organic phase under reduced pressure yields a milky gel that turns subsequently into a liposomal suspension.
[0476] The term "ethanol injection technique" refers to a process, in which an ethanol solution comprising lipids is rapidly injected into an aqueous solution through a needle. This action disperses the lipids throughout the solution and promotes lipid structure formation, for example lipid vesicle formation such as liposome formation. Generally, the nucleic acid (such as RNA and / or DNA, especially mRNA) lipoplex particles described herein are obtainable by adding nucleic acid (such as RNA and / or DNA, especially mRNA) to a colloidal liposome dispersion. Using the ethanol injection technique, such colloidal liposome dispersion is, in some embodiments, formed as follows: an ethanol solution comprising lipids, such as cationic or cationically ionizable lipids like DOTMA and / or DODMA and additional lipids, is injected into an aqueous solution under stirring. In some embodiments, the nucleic acid (such as RNA and / or DNA, especially mRNA) lipoplex particles described herein are obtainable without a step of extrusion.
[0477] The term "extruding" or "extrusion" refers to the creation of particles having a fixed, cross-sectional profile. In particular, it refers to the downsizing of a particle, whereby the particle is forced through filters with defined pores.
[0478] Other methods having organic solvent free characteristics may also be used according to the present disclosure for preparing a colloid.
[0479] In some embodiments, LNPs comprise four components: ionizable cationic lipids, neutral lipids such as phospholipids, a steroid such as cholesterol, and a polymer conjugated lipid. In some embodiments, LNPs may be prepared by mixing lipids dissolved in ethanol rapidly with nucleic acid (such as RNA and / or DNA) in an aqueous buffer. While nucleic acid (such as RNA and / or DNA) particles described herein may comprise polymer conjugated lipids such as PEG lipids, provided herein are also nucleic acid (such as RNA and / or DNA) particles which do not comprise polymer conjugated lipids such as PEG lipids.
[0480] In some embodiments, the LNPs comprising nucleic acid (such as RNA and / or DNA) and at least one cationic or cationically ionizable lipid described herein are prepared by (a) preparing a nucleic acid (such as RNA and / or DNA) solution containing water and a buffering system; (b) preparing an ethanolic solution comprising the cationic or cationically ionizable lipid and, if present, one or more additional lipids; and (c) mixing the nucleic acid (such as RNA and / or DNA) solution prepared under (a) with the ethanolic solution prepared under (b), thereby preparing the formulation comprising LNPs. After step (c) one or more steps selected from diluting and filtrating, such as tangential flow filtrating, can follow. In some embodiments, the LNPs comprising nucleic acid (such as RNA and / or DNA) and at least one cationic or cationically ionizable lipid described herein are prepared by (a’) preparing liposomes or a colloidal preparation of the cationic or cationically ionizable lipid and, if present, one or more additional lipids in an aqueous phase; and (b’) preparing a nucleic acid (such as RNA and / or DNA) solution containing water and a buffering system; and (c’) mixing the liposomes or colloidal preparation prepared under (a’) with the nucleic acid (such as RNA and / or DNA) solution prepared under (b’). After step (c’) one or more steps selected from diluting and filtrating, such as tangential flow filtrating, can follow.
[0481] The present disclosure describes particles comprising nucleic acid (such as RNA and / or DNA, especially mRNA) and at least one cationic or cationically ionizable lipid which associates with the nucleic acid (such as RNA and / or DNA) to form nucleic acid (such as RNA and / or DNA) particles and compositions comprising such particles. The nucleic acid (such as RNA and / or DNA) particles may comprise nucleic acid (such as RNA and / or DNA) which is complexed in different forms by non-covalent interactions to the particle. The particles described herein are not viral particles, in particular infectious viral particles, i.e., they are not able to virally infect cells.
[0482] Suitable cationic or cationically ionizable lipids are those that form nucleic acid particles and are included by the term "particle forming components" or "particle forming agents". The term "particle forming components" or "particle forming agents" relates to any components which associate with nucleic acid to form nucleic acid particles. Such components include any component which can be part of nucleic acid particles.
[0483] In some embodiments, nucleic acid particles (such as RNA and / or DNA particles, especially mRNA particles) comprise more than one type of nucleic acid (such as RNA and / or DNA) molecules, where the molecular parameters of the nucleic acid (such as RNA and / or DNA) molecules may be similar or different from each other, like with respect to molar mass or fundamental structural elements such as molecular architecture, capping (only RNA), coding regions or other features,
[0484] In particulate formulation, it is possible that each nucleic acid (such as RNA and / or DNA) species is separately formulated as an individual particulate formulation. In that case, each individual particulate formulation will comprise one nucleic acid (such as RNA and / or DNA) species. The individual particulate formulations may be present as separate entities, e.g. in separate containers. Such formulations are obtainable by providing each nucleic acid (such as RNA and / or DNA) species separately (typically each in the form of a nucleic acid (such as RNA and / or DNA)-containing solution) together with a particle-forming agent, thereby allowing the formation of particles. Respective particles will contain exclusively the specific nucleic acid (such as RNA and / or DNA) species that is being provided when the particles are formed (individual particulate formulations). In some embodiments, a composition such as a pharmaceutical composition comprises more than one individual particle formulation. Respective pharmaceutical compositions are referred to as mixed particulate formulations. Mixed particulate formulations according to the invention are obtainable by forming, separately, individual particulate formulations, followed by a step of mixing of the individual particulate formulations. By the step of mixing, a formulation comprising a mixed population of nucleic acid (such as RNA and / or DNA)-containing particles is obtainable. Individual particulate populations may be together in one container, comprising a mixed population of individual particulate formulations. Alternatively, it is possible that all nucleic acid (such as RNA and / or DNA) species of the pharmaceutical composition are formulated together as a combined particulate formulation. Such formulations are obtainable by providing a combined formulation (typically combined solution) of all nucleic acid (such as RNA and / or DNA) species together with a particle-forming agent, thereby allowing the formation of particles. As opposed to a mixed particulate formulation, a combined particulate formulation will typically comprise particles which comprise more than one nucleic acid (such as RNA and / or DNA) species. In a combined particulate composition different nucleic acid (such as RNA and / or DNA) species are typically present together in a single particle.
[0485] Polymers
[0486] Given their high degree of chemical flexibility, polymers are commonly used materials for nanoparticle-based delivery. Typically, cationic polymers are used to electrostatically condense the negatively charged nucleic acid into nanoparticles. These positively charged groups often consist of amines that change their state of protonation in the pH range between 5.5 and 7.5, thought to lead to an ion imbalance that results in endosomal rupture. Polymers such as poly-L-lysine, polyamidoamine, protamine and polyethyleneimine, as well as naturally occurring polymers such as chitosan have all been applied to nucleic acid delivery and are suitable as cationic polymers herein. In addition, some investigators have synthesized polymers specifically for nucleic acid delivery. Poly([3-amino esters), in particular, have gained widespread use in nucleic acid delivery owing to their ease of synthesis and biodegradability. Such synthetic polymers are also suitable as cationic polymers herein.
[0487] A "polymer," as used herein, is given its ordinary meaning, i.e. , a molecular structure comprising one or more repeat units (monomers), connected by covalent bonds. The repeat units can all be identical, or in some cases, there can be more than one type of repeat unit present within the polymer. In some cases, the polymer is biologically derived, i.e., a biopolymer such as a protein. In some cases, additional moieties can also be present in the polymer, for example targeting moieties.
[0488] If more than one type of repeat unit is present within the polymer, then the polymer is said to be a "copolymer." It is to be understood that the polymer being employed herein can be a copolymer. The repeat units forming the copolymer can be arranged in any fashion. For example, the repeat units can be arranged in a random order, in an alternating order, or as a "block" copolymer, i.e., comprising one or more regions each comprising a first repeat unit (e.g., a first block), and one or more regions each comprising a second repeat unit (e.g., a second block), etc. Block copolymers can have two (a diblock copolymer), three (a triblock copolymer), or more numbers of distinct blocks.
[0489] In certain embodiments, the polymer is biocompatible. Biocompatible polymers are polymers that typically do not result in significant cell death at moderate concentrations. In certain embodiments, the biocompatible polymer is biodegradable, i.e., the polymer is able to degrade, chemically and / or biologically, within a physiological environment, such as within the body.
[0490] In certain embodiments, polymer may be protamine or polyalkyleneimine.
[0491] The term "protamine" refers to any of various strongly basic proteins of relatively low molecular weight that are rich in arginine and are found associated especially with DNA in place of somatic histones in the sperm cells of various animals (as fish). In particular, the term "protamine" refers to proteins found in fish sperm that are strongly basic, are soluble in water, are not coagulated by heat, and yield chiefly arginine upon hydrolysis. In purified form, they are used in a long-acting formulation of insulin and to neutralize the anticoagulant effects of heparin.
[0492] According to the disclosure, the term "protamine" as used herein is meant to comprise any protamine amino acid sequence obtained or derived from natural or biological sources including fragments thereof and multimeric forms of said amino acid sequence or fragment thereof as well as (synthesized) polypeptides which are artificial and specifically designed for specific purposes and cannot be isolated from native or biological sources.
[0493] In one embodiment, the polyalkyleneimine comprises polyethylenimine and / or polypropylenimine, preferably polyethyleneimine. A preferred polyalkyleneimine is polyethyleneimine (PEI). The average molecular weight of PEI is preferably 0.75- 102to 107Da, preferably 1000 to 105Da, more preferably 10000 to 40000 Da, more preferably 15000 to 30000 Da, even more preferably 20000 to 25000 Da.
[0494] Preferred according to the disclosure is linear polyalkyleneimine such as linear polyethyleneimine (PEI).
[0495] Cationic polymers (including polycationic polymers) contemplated for use herein include any cationic polymers which are able to electrostatically bind nucleic acid. In one embodiment, cationic polymers contemplated for use herein include any cationic polymers with which nucleic acid can be associated, e.g. by forming complexes with the nucleic acid or forming vesicles in which the nucleic acid is enclosed or encapsulated.
[0496] Particles described herein may also comprise polymers other than cationic polymers, i.e. , non-cationic polymers and / or anionic polymers. Collectively, anionic and neutral polymers are referred to herein as non-cationic polymers.
[0497] Lipids
[0498] The terms "lipid" and "lipid-like material" are broadly defined herein as molecules which comprise one or more hydrophobic moieties or groups and optionally also one or more hydrophilic moieties or groups. Molecules comprising hydrophobic moieties and hydrophilic moieties are also frequently denoted as amphiphiles. Lipids are usually insoluble or poorly soluble in water, but soluble in many organic solvents. In an aqueous environment, the amphiphilic nature allows the molecules to self-assemble into organized structures and different phases. One of those phases consists of lipid bilayers, as they are present in vesicles, multilamellar / unilamellar liposomes, or membranes in an aqueous environment. Hydrophobicity can be conferred by the inclusion of apolar groups that include, but are not limited to, long-chain saturated and unsaturated aliphatic hydrocarbon groups and such groups substituted by one or more aromatic, cycloaliphatic, or heterocyclic group(s). The hydrophilic groups may comprise polar and / or charged groups and include carbohydrates, phosphate, carboxylic, sulfate, amino, sulfhydryl, nitro, hydroxyl, and other like groups.
[0499] As used herein, the term "hydrophobic" refers to any a molecule, moiety or group which is substantially immiscible or insoluble in aqueous solution. The term hydrophobic group includes hydrocarbons having at least 6 carbon atoms. The hydrophobic group can have functional groups (e.g., ether, ester, halide, etc.) and atoms other than carbon and hydrogen as long as the group satisfies the condition of being substantially immiscible or insoluble in aqueous solution.
[0500] The term "hydrocarbon" includes alkyl, alkenyl, or alkynyl as defined herein. It should be appreciated that one or more of the hydrogen in alkyl, alkenyl, or alkynyl may be substituted with other atoms, e.g., halogen, oxygen or sulfur. Unless stated otherwise, hydrocarbon groups can also include a cyclic (alkyl, alkenyl or alkynyl) group or an aryl group, provided that the overall polarity of the hydrocarbon remains relatively nonpolar. The term "alkyl" refers to a saturated linear or branched monovalent hydrocarbon moiety which may have six to thirty, typically six to twenty, often six to eighteen carbon atoms. Exemplary nonpolar alkyl groups include, but are not limited to, hexyl, decyl, dodecyl, tetradecyl, hexadecyl, octadecyl, and the like.
[0501] The term "alkenyl" refers to a linear or branched monovalent hydrocarbon moiety having at least one carbon carbon double bond in which the total carbon atoms may be six to thirty, typically six to twenty often six to eighteen.
[0502] The term "alkynyl" refers to a linear or branched monovalent hydrocarbon moiety having at least one carbon carbon triple bond in which the total carbon atoms may be six to thirty, typically six to twenty, often six to eighteen. Alkynyl groups can optionally have one or more carbon carbon double bonds.
[0503] As used herein, the term "amphiphilic" refers to a molecule having both a polar portion and a non-polar portion. Often, an amphiphilic compound has a polar head attached to a long hydrophobic tail. In some embodiments, the polar portion is soluble in water, while the non-polar portion is insoluble in water. In addition, the polar portion may have either a formal positive charge, or a formal negative charge. Alternatively, the polar portion may have both a formal positive and a negative charge, and be a zwitterion or inner salt. For purposes of the disclosure, the amphiphilic compound can be, but is not limited to, one or a plurality of natural or non-natural lipids and lipid-like compounds.
[0504] The term "lipid-like material", "lipid-like compound" or "lipid-like molecule" relates to substances, in particular amphiphilic substances, that structurally and / or functionally relate to lipids but may not be considered as lipids in a strict sense. For example, the term includes compounds that are able to form amphiphilic layers as they are present in vesicles, multilamellar / unilamellar liposomes, or membranes in an aqueous environment and includes surfactants, or synthesized compounds with both hydrophilic and hydrophobic moieties. Generally speaking, the term refers to molecules, which comprise hydrophilic and hydrophobic moieties with different structural organization, which may or may not be similar to that of lipids. Examples of lipid-like compounds capable of spontaneous integration into cell membranes include functional lipid constructs such as synthetic function-spacer-lipid constructs (FSL), synthetic function- spacer-sterol constructs (FSS) as well as artificial amphipathic molecules. Lipids are generally cylindrical. The area occupied by the two alkyl chains is similar to the area occupied by the polar head group. Lipids have low solubility as monomers and tend to aggregate into planar bilayers that are water insoluble. Traditional surfactant monomers are generally cone shaped. The hydrophilic head groups tend to occupy more molecular space than the linear alkyl chains. In some embodiments, surfactants tend to aggregate into spherical or elliptoid micelles that are water soluble. While lipids also have the same general structure as surfactants - a polar hydrophilic head group and a nonpolar hydrophobic tail - lipids differ from surfactants in the shape of the monomers, in the type of aggregates formed in solution, and in the concentration range required for aggregation. As used herein, the term "lipid" is to be construed to cover both lipids and lipid-like materials unless otherwise indicated herein or clearly contradicted by context.
[0505] Generally, lipids may be divided into eight categories: fatty acids, glycerolipids, glycerophospholipids, sphingolipids, saccharolipids, polyketides (derived from condensation of ketoacyl subunits), sterol lipids and prenol lipids (derived from condensation of isoprene subunits). Although the term "lipid" is sometimes used as a synonym for fats, fats are a subgroup of lipids called triglycerides. Lipids also encompass molecules such as fatty acids and their derivatives (including tri-, di-, monoglycerides, and phospholipids), as well as steroids, i.e., sterol-containing metabolites such as cholesterol or a derivative thereof. Examples of cholesterol derivatives include, but are not limited to, cholestanol, cholestanone, cholestenone, coprostanol, cholesteryl-2'-hydroxyethyl ether, cholesteryl-4'- hydroxybutyl ether, tocopherol and derivatives thereof, and mixtures thereof. Fatty acids, or fatty acid residues are a diverse group of molecules made of a hydrocarbon chain that terminates with a carboxylic acid group; this arrangement confers the molecule with a polar, hydrophilic end, and a nonpolar, hydrophobic end that is insoluble in water. The carbon chain, typically between four and 24 carbons long, may be saturated or unsaturated, and may be attached to functional groups containing oxygen, halogens, nitrogen, and sulfur. If a fatty acid contains a double bond, there is the possibility of either a cis or trans ge...
Claims
Claims1. A method for simultaneously analysing at least two different nucleic acid sequences, wherein each different nucleic acid sequence encodes a different amino acid sequence, wherein each amino acid sequence comprises a different Major Histocompatibility (MHC) class I trafficking domain (MITD) sequence, wherein the method comprises the following steps:(i) providing the at least two different nucleic acid sequences;(ii) expressing the at least two different amino acid sequences comprising different MITD sequences; and(iii) using the at least two different expressed MITD sequences to differentiate between the at least two different nucleic acid sequences.
2. The method of claim 1 , wherein step (iii) comprises the following:(a) determining the abundance of each of the at least two different expressed MITD sequences; and(b) correlating the abundance of each expressed MITD sequence with the expression level of the nucleic acid sequence encoding said MITD sequence.
3. The method of claim 2, wherein step (a) comprises using mass spectrometry (MS), liquid chromatography MS (LC-MS), targeted LC-MS, a detection method using anti-MlTD antibodies or ELISA to determine the abundance of each of the at least two different expressed MITD sequences.
4. The method of any one of claims 1 to 3, wherein the MITD sequences are: a) selected from naturally occurring MITD sequences of a human leukocyte antigen (HLA) gene, e.g. a HLA-A, HLA-B or HLA-C gene, preferably a HLA-B gene; and / or b) selected from MITD sequences differing by one, two or three amino acid substitutions relative to a naturally occurring MITD sequence.
5. The method of any one of claims 1 to 4, wherein the MITD sequences each elicit the same, substantially the same or equivalent immune response, protective immune response, partially protective immune response and / or clinical immune response as anaturally occurring MITD sequence of a HLA gene, e.g. as a naturally occurring MITD sequence of a HLA gene in a human cell or patient.
6. The method of any one of claims 1 to 5, wherein the MITD sequences each have the same, substantially the same or equivalent vaccine efficacy, epitope presentation and / or cellular trafficking function as a naturally occurring MITD sequence of a HLA gene, e.g. as a naturally occurring MITD sequence of a HLA gene in a human cell or patient.
7. The method of any one of claims 1 to 6, wherein the MITD sequences each comprise an amino acid sequence comprising a transmembrane (TM) domain and a further sequence directly downstream of the last K / R residue in the amino acid sequence comprising a TM domain, preferably wherein the further sequence comprises 22 amino acid residues.
8. The method of any one of claims 1 to 7, wherein each MITD sequence comprises the same amino acid sequence comprising a TM domain, preferably wherein the amino acid sequence comprising a TM domain is IVGIVAGLAVLAVWIGAWATVMCRRKSSGGK (SEQ ID NO: 1 ).
9. The method of claim 7 or 8, wherein each of the further sequences differ by one, two or three amino acid substitutions.
10. The method of any one of claims 7 to 9, wherein the further sequences in each of the MITD sequences are different and selected from the following:GGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 2), AGSYSQAASSDSAQGSDVSLTA (SEQ ID NO: 3), GGGYSQAASSDSAQGSDVSLTA (SEQ ID NO: 4), GGSNSQAASSDSAQGSDVSLTA (SEQ ID NO: 5), GGSYSEAASSDSAQGSDVSLTA (SEQ ID NO: 6), GGSYSQAAFSDSAQGSDVSLTA (SEQ ID NO: 7), GGSYSQAASSDSAQGSDVFLTA (SEQ ID NO: 8), GGSYSQAASSDSAQGSDVSLTD (SEQ ID NO: 9), GGSYSQAASSDSAQGSDVSVTA (SEQ ID NO: 10),GGSYSQAASSDSAQGSEVSLTA (SEQ ID NO: 11 ), GGSYSQAASSDSAQGSNVSLTA (SEQ ID NO: 12), GGSYSQAASSDSAQGSVVSLTA (SEQ ID NO: 13), GGSYSQAASSDSAQSSDVSLTA (SEQ ID NO: 14), GGSYSQAASSDSAQVSDVSLTA (SEQ ID NO: 15), GGSYSQAASSDSDQGSDVSLTA (SEQ ID NO: 16), GGSYSQAASSDSSQGSDVSLTA (SEQ ID NO: 17), GGSYSQAASSDSTQGSDVSLTA (SEQ ID NO: 18), GGSYSQAASSDSVQGSDVSLTA (SEQ ID NO: 19), GGSYSQAASSNSAQGSDVSLTA (SEQ ID NO: 20), GGSYSQAASSYSAQGSDVSLTA (SEQ ID NO: 21 ), GGSYSQAVSSDSAQGSDVSLTA (SEQ ID NO: 22), GGSYYQAASSDSAQGSDVSLTA (SEQ ID NO: 23), GASYSQAASSDSAQGSDVSLTA (SEQ ID NO: 24), GGSYSQGASSDSAQGSDVSLTA (SEQ ID NO: 25), GGSYSQAGSSDSAQGSDVSLTA (SEQ ID NO: 26), GGSYSQAASSDSGQGSDVSLTA (SEQ ID NO: 27), GGSYSQAASSDSAQASDVSLTA (SEQ ID NO: 28), and GGSYSQAASSDSAQGSDVSLTG (SEQ ID NO: 29).
11. The method of any one of claims 1 to 10, wherein the method is for simultaneously analysing two, three, four, five, six, seven, eight, nine, ten, eleven, twelve, thirteen, fourteen, fifteen, sixteen, seventeen, eighteen, nineteen, twenty, twenty-one, twenty-two, twenty-three, twenty-four, twenty-five, twenty-six, twentyseven or twenty-eight different nucleic acid sequences; and / or wherein each of the amino acid sequences further comprises one or more antigen or epitope sequences, such as one or more cancer antigen, neoantigen, infectious disease antigen, autoimmune disease antigen, T-cell epitope, fixed epitope or variable epitope sequences; and / or wherein the at least two different nucleic acid sequences comprise DNA and / or RNA, preferably wherein the at least two different nucleic acid sequences comprise messenger RNA (mRNA).17912. A composition comprising at least two different nucleic acid sequences, wherein each of the nucleic acid sequences encodes a different amino acid sequence comprising: a) at least one antigen or epitope sequence; and b) an MITD sequence; wherein the MITD sequences in each of the different amino acid sequences are different.
13. A composition comprising at least two different amino acid sequences, wherein each of the amino acid sequences comprises: a) at least one antigen or epitope sequence; and b) an MITD sequence; wherein the MITD sequences in each of the different amino acid sequences are different.
14. A kit comprising at least two different nucleic acid sequences, wherein each nucleic acid sequence encodes an amino acid sequence comprising at least one antigen or epitope sequence and an MITD sequence, wherein the MITD sequences in each of the different amino acid sequences are different.
15. Use of the composition according to claim 12 or claim 13, or the kit according to claim 14, for simultaneously analysing the expression level of at least two different nucleic acid sequences.
16. The method, composition, kit or use of any one of claims 1 to 15, wherein each of the MITD sequences comprises the C or N terminus of the amino acid sequence in which the MITD sequence resides.180
Citation Information
Patent Citations
Modification of RNA, producing an increased transcript stability and translation efficiency
WO2007036366A2
RNA formulation for immunotherapy
WO2013143683A1
Stabilization of poly(a) sequence encoding DNA sequences
WO2016005324A1
3' UTR sequences for stabilization of RNA
WO2017060314A2
Novel lipids and lipid nanoparticle formulations for delivery of nucleic acids
WO2017075531A1