Multisubunit RSV, HMPV and HPIV vaccines and therapeutics
A multisubunit nucleic acid and peptide vaccines targeting RSV, HMPV, and HPIV proteins are developed, providing effective immunization against these viruses through lipid nanoparticle delivery, overcoming the limitations of existing vaccines and therapies.
Patent Information
- Application Number
- PCT/IB2025/000330
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2024-06-26
- Filing Date
- 2025-06-24
- Publication Date
- 2026-01-02
AI Technical Summary
There are no effective vaccines or therapeutic options available for respiratory syncytial virus (RSV), human metapneumovirus (HMPV), and human parainfluenza virus (HPIV), particularly for infants and young children, with existing vaccines being inadequate or unaffordable, and current antibody therapies insufficient.
Development of a multisubunit nucleic acid comprising polynucleotide sequences derived from target sequences, linker sequences, and self-assembling sequences from RSV, HMPV, and HPIV proteins, which can be encoded into polypeptides and formulated in lipid nanoparticles for vaccine delivery.
The multisubunit nucleic acid and peptide vaccines provide a potent immunogenic response, offering protection against RSV, HMPV, and HPIV, addressing the lack of effective vaccines and therapies for these viruses.
Smart Images

Figure IB2025000330_02012026_PF_FP_ABST
Abstract
Description
[0001] Multisubunit RSV, HMPV and HPIV Vaccines and Therapeutics
[0002] RELATED APPLICATIONS
[0003] This application claims the benefit of priority to U.S. Provisional Patent Application serial numbers 63 / 664,456, 63 / 664,432 and 63 / 664,424 filed on June 26, 2024, the entire contents of which are hereby incorporated herein by reference in its entirety.
[0004] REFERENCE TO A SEQUENCE LISTING XML
[0005] This application contains a Sequence Listing which has been submitted electronically in XML format. The Sequence Listing XML is incorporated herein by reference. Said XML file, created on June 20, 2025, is named PVM-01125. xml and is 16,903,781 bytes in size.
[0006] BACKGROUND
[0007] Respiratory Syncytial Virus (RSV) is the major cause of severe respiratory illness globally, especially in young children, elderly, and immunocompromised individuals. It is one of the leading causes of infant hospitalization. A 2019 assessment of RSV impact worldwide estimated approximately 33 million RSV-associated lower respiratory episodes in young children, with nearly 95% of these seen in low- and middle-income countries. RSV causes significant morbidity and mortality, with more than 1 ,00,000 fatalities each year in children under 5 years, including 45,000 deaths in infants below 6 months.
[0008] Metapneumovirus is a genus within the family Paramyxoviridae comprising two species viz., human metapneumovirus (HMPV) and avian metapneumovirus (AMPV). HMPV has been responsible for mild respiratory illness to severe bronchiolitis and pneumonia in children below the age of 5 years. Infants, elderly, and immunocompromised individuals are at higher risk of acquiring HMPV infection. It is widely accepted that HMPV exposure occurs in more than 90% of the children by the age of 5 and reinfection remains a major concern.
[0009] Human Parainfluenza Virus (HPIV) is one of the leading causes of respiratory tract infections across all age groups with the highest incidence usually occurring in young children. HPIV causes a range of clinical manifestations from mild upper respiratory tract infections (URTI), croup, bronchitis, bronchiolitis, to severe pneumonia. Although most HPIV infections are transient and mild, they can lead to severe respiratory complications, especially in young children and immunocompromised individuals, or persons with underlying health conditions. Four serotypes of human parainfluenza virus have been identified viz., HPIV-1 to HPIV-4, with HPIV-4 further categorized into two subtypes - HPIV-4a and HPIV-4b. HPIVs are known to cause seasonal outbreaks of respiratory tract infections in spring and summer, with HPIV-1 causing the larger outbreaks, and HPIV-2 outbreaks usually follow the HPIV-1. HPIV-3 is the most virulent serotype among all the HPIVs.
[0010] Couple of vaccines against RSV have been recently approved for adults, but a potent vaccine against infants and young children is still elusive, approved antibody therapies against RSV are either inadequate or unaffordable to many. There are no approved vaccines or specific therapeutic against metapneumovirus and human parainfluenza virus prevention or treatment. Therefore, effective vaccines against respiratory syncytial virus (RSV), human metapneumovirus (HMPV) and human parainfluenza virus (HPIV) are an important necessity.
[0011] SUMMARY
[0012] Accordingly, the present disclosure relates to a multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein some or all polynucleotide sequences of the plurality comprises either a target sequence, a linker sequence, and a selfassembling sequence or a linker sequence, a target sequence, a linker sequence and a selfassembling sequence or a combination thereof, wherein the target sequence is obtained or derived from:
[0013] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0014] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0015] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0016] (d) a combination of (a), (b), and / or (c).
[0017] In some embodiments, each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence. In some embodiments, the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality. In some embodiments, the target sequence, the linker sequence, and the self-assembling sequence or the linker sequence, the target sequence, the linker sequence, and the self-assembling sequence are in 5' to 3' order.
[0018] In another aspect, provided herein is a multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a target sequence, a linker sequence, and a self-assembling sequence, wherein the target sequence is obtained or derived from:
[0019] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0020] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0021] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0022] (d) a combination of (a), (b), and / or (c).
[0023] In some embodiments, each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence. In some embodiments, the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality. In some embodiments, the target sequence, the linker sequence, and the self-assembling sequence are in 5' to 3' order.
[0024] In another aspect, provided herein is a multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a linker sequence, a target sequence, a linker sequence, and a selfassembling sequence, wherein the target sequence is obtained or derived from:
[0025] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0026] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0027] (d) a combination of (a), (b), and / or (c).
[0028] In some embodiments, each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence. In some embodiments, the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality. In some embodiments, the linker sequence, the target sequence, the linker sequence, and the self-assembling sequence are in 5' to 3' order.
[0029] In another aspect, provided herein is vaccine comprising a multisubunit nucleic acid, wherein the multisubunit nucleic acid comprises a plurality of polynucleotide sequences, wherein some or all polynucleotide sequences of the plurality comprises either a target sequence, a linker sequence, and a self-assembling sequence or a linker sequence, a target sequence, a linker sequence and a self-assembling sequence or a combination thereof, wherein the target sequence is obtained or derived from:
[0030] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0031] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0032] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0033] (d) a combination of (a), (b), and / or (c).
[0034] In some embodiments, each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence. In some embodiments, the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality. In some embodiments, the target sequence, the linker sequence, and the self-assembling sequence or the linker sequence, the target sequence, the linker sequence, and the self-assembling sequence are in 5' to 3' order.
[0035] In another aspect, provided herein a vaccine comprising a multisubunit nucleic acid, wherein the multisubunit nucleic acid comprises a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a target sequence, a linker sequence, and a self-assembling sequence, wherein the target sequence is obtained or derived from:
[0036] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0037] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0038] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0039] (d) a combination of (a), (b), and / or (c).
[0040] In some embodiments, each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence. In some embodiments, the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality. In some embodiments, the target sequence, the linker sequence, and the self-assembling sequence are in 5' to 3' order.
[0041] In another aspect, provided herein is a vaccine comprising a multisubunit nucleic acid, wherein the multisubunit nucleic acid comprises a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, wherein the target sequence is obtained or derived from:
[0042] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0043] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0044] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0045] (d) a combination of (a), (b), and / or (c).
[0046] In some embodiments, each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence. In some embodiments, the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality. In some embodiments, the linker sequence, the target sequence, the linker sequence, and the self-assembling sequence are in 5' to 3' order.
[0047] In another aspect, provided herein is a multisubunit nucleic acid encoding a plurality of polypeptides, wherein some or all polypeptides of the plurality comprises either a target peptide, a linker peptide, and a self-assembling peptide or a linker peptide, a target peptide, a linker peptide and a self-assembling peptide or a combination thereof, wherein the target peptide is obtained or derived from:
[0048] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0049] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0050] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0051] (d) a combination of (a), (b), and / or (c).
[0052] In some embodiments, each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide. In some embodiments, the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality. In some embodiments, the target peptide, the linker peptide, and the self-assembling peptide or the linker peptide, the target peptide, the linker peptide and the self- assembling peptide are in N-terminus to C-terminus order.
[0053] In another aspect, provided herein is a multisubunit nucleic acid encoding a plurality of polypeptides, wherein each polypeptide of the plurality comprises a target peptide, a linker peptide, and a self-assembling peptide, wherein the target peptide is obtained or derived from
[0054] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0055] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0056] (d) a combination of (a), (b), and / or (c).
[0057] In some embodiments, each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide. In some embodiments, the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality. In some embodiments, the target peptide, the linker peptide, and the self- assembling peptide are in N-terminus to C-terminus order.
[0058] In another aspect, provided herein is a multisubunit nucleic acid encoding a plurality of polypeptides, wherein each polypeptide of the plurality comprises a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, wherein the target peptide is obtained or derived from
[0059] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0060] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0061] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0062] (d) a combination of (a), (b), and / or (c).
[0063] In some embodiments, each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide. In some embodiments, the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality. In some embodiments, the linker peptide, the target peptide, the linker peptide, and the self-assembling peptide are in N-terminus to C-terminus order.
[0064] In another aspect, provided herein is a vaccine comprising a multisubunit nucleic acid encoding a plurality of polypeptides, wherein some or all polypeptides of the plurality comprises either a target peptide, a linker peptide, and a self-assembling peptide or a linker peptide, a target peptide, a linker peptide and a self-assembling peptide or a combination thereof, wherein the target peptide is obtained or derived from:
[0065] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0066] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0067] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0068] (d) a combination of (a), (b), and / or (c).
[0069] In some embodiments, each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide. In some embodiments, the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality. In some embodiments, the target peptide, the linker peptide, and the self-assembling peptide or the linker peptide, the target peptide, the linker peptide and the self- assembling peptide are in N-terminus to C-terminus order.
[0070] In another aspect, provided herein is a vaccine comprising a multisubunit nucleic acid encoding a plurality of polypeptides, wherein each polypeptide of the plurality comprises a target peptide, a linker peptide, and a self-assembling peptide, wherein the target peptide is obtained or derived from
[0071] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0072] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0073] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0074] (d) a combination of (a), (b), and / or (c).
[0075] In some embodiments, each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide. In some embodiments, the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality. In some embodiments, the target peptide, the linker peptide, and the self- assembling peptide are in N-terminus to C-terminus order.
[0076] In another aspect, provided herein is a vaccine comprising a multisubunit nucleic acid encoding a plurality of polypeptides, wherein each polypeptide of the plurality comprises a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, wherein the target peptide is obtained or derived from
[0077] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0078] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0079] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0080] (d) a combination of (a), (b), and / or (c).
[0081] In some embodiments, each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide. In some embodiments, the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality. In some embodiments, the linker peptide, the target peptide, the linker peptide, and the self-assembling peptide are in N-terminus to C-terminus order.
[0082] In some embodiments, total number of the polynucleotide sequences are not more than 100. In some embodiments, total number of the polynucleotide sequences are between 2-5, 2-10, 10-20, 20-30, 30-40, 40-50, 50-60, 60-70, 70-80, 80-90, or 90-99.
[0083] In some embodiments, the multisubunit nucleic acid is a DNA or an RNA. In some embodiments, the RNA is an mRNA. In some embodiments, the mRNA is obtained or synthesized through a single IVT process or step.
[0084] In some embodiments, the linker sequence encodes a linker peptide. In some embodiments, the linker peptide is an amino acid linker, a zipper motif, a foldon, a scaffold, or a combination thereof. In some embodiments, the linker peptide is an amino acid linker. In some embodiments, the linker peptide is a zipper motif. In some embodiments, the linker peptide is a foldon. In some embodiments, the linker peptide is a scaffold. In some embodiments, the linker peptide comprises an amino acid linker and a zipper motif. In some embodiments, the linker peptide comprises an amino acid linker and a foldon. In some embodiments, the linker peptide comprises an amino acid linker and a scaffold. In some embodiments, the linker peptide comprises a zipper motif and a scaffold. In some embodiments, the linker peptide comprises a foldon and a scaffold. In some embodiments, the linker peptide comprises an amino acid linker, a zipper motif, and a scaffold.
[0085] In some embodiments, the amino acid linker comprises 2 to 49 amino acids. In some embodiments, the amino acid linker is a glycine serine linker, a glycine proline linker, a glycine threonine linker, an alanine serine linker, any combination of two amino acids, or a combination thereof.
[0086] In some embodiments, the linker peptide has an amino acid sequence of any one of SEQ ID NOs: 12-51 or 12221-12225.
[0087] In some embodiments, the self-assembling sequence encodes a self-assembling peptide. In some embodiments, the self-assembling peptide is lumazine synthase, MS2 coat protein, hepatitis B surface antigen (HBsAg) from Hepatitis B Virus, hepatitis B core antigen (HBcAg) from Hepatitis B virus, human papillomavirus LI (HPV LI) protein, matrix protein Ml from influenza A virus, ferritin, riboflavin synthase, dihydrolipoyl acetyltransferase (E2p), or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents or functional analogs thereof.
[0088] In some embodiments, the ferritin comprises of ferritin subunit or ferritin peptide. In some embodiments, the ferritin peptide is obtained or derived from Listeria innocua or Helicobacter pylori. In some embodiments, the ferritin peptide is obtained or derived from Listeria innocua ferritin, or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the ferritin peptide is obtained or derived from Helicobacter pylori ferritin, or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the dihydrolipoyl acetyltransferase (E2p) is obtained or derived from Bacillus stearothermophilus or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the lumazine synthase is obtained or derived from Aquifex species (for example, Aquifex aeolicus) or Bacillus species (for example, Bacillus subtilis), or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the MS2 coat protein is obtained or derived from Emesvirus zinderi, or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the self-assembling peptide has an amino acid sequence of any one of SEQ ID NOs: 1-11 or 12217-12220.
[0089] In some embodiments, the cleavage sequence encodes a cleavage peptide. In some embodiments, the cleavage sequence encodes one or more cleavage peptides. In some embodiments, the cleavage peptide comprises two or more cleavage peptides, for example, cleavage peptide- 1, cleavage peptide-2, and so on. In some embodiments, the one or more cleavage peptides are optionally connected to each other by a linker peptide. In some embodiments, the cleavage peptide is a golgi specific cleavage peptide or self cleaving peptide. In some embodiments, the cleavage peptide has an amino acid sequence of any one of SEQ ID NOs: 52-66.
[0090] In some embodiments, the signal sequence encodes a signal peptide. In some embodiments, the signal peptide is present on the amino-terminus of the first polypeptide. In some embodiments, the multisubunit nucleic acid further encodes a second signal peptide on the amino-terminus of all or some polypeptides. In some embodiments, the signal peptide has an amino acid sequence of any one of SEQ ID NOs: 67-86.
[0091] In some embodiments, the target sequence encodes a target peptide. In some embodiments, the target peptide is encoded by a codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the target peptide is obtained or derived from:
[0092] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0093] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0094] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0095] (d) a combination of (a), (b), and / or (c).
[0096] In some embodiments, the target peptide is an envelope protein of a respiratory syncytial virus, a metapneumovirus, or a human parainfluenza virus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the envelope protein has an amino acid sequence of any one of SEQ ID NOs: 87-4794, 7372-7935, or 8628-10015, representing the respiratory syncytial virus, the metapneumovirus, or the human parainfluenza virus respectively, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0097] In some embodiments, the target peptide is a matrix protein of a respiratory syncytial virus, a metapneumovirus, a human parainfluenza virus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the matrix protein has an amino acid sequence of any one of SEQ ID NOs: 4795-4978, 7936-7976, or 10016-10214, representing the respiratory syncytial virus, the metapneumovirus, or the human parainfluenza virus respectively, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0098] In some embodiments, the target peptide is a nucleocapsid protein of a respiratory syncytial virus, a metapneumovirus, or a human parainfluenza virus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the nucleocapsid protein has an amino acid sequence of any one of SEQ ID NOs: 4979-6132, 7977-8483, or 10215-11485, representing the respiratory syncytial virus, the metapneumovirus, or the human parainfluenza virus respectively, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0099] In some embodiments, the target peptide is a M2-1 protein of a respiratory syncytial virus, or M2-1 protein of a metapneumovirus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the M2-1 protein has an amino acid sequence of any one of SEQ ID NOs: 6133-6381, or 8484-8545, representing the respiratory syncytial virus, or the metapneumovirus, respectively, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0100] In some embodiments, the target peptide is a M2-2 protein of a respiratory syncytial virus, M2-2 protein of a metapneumovirus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the M2-2 protein has an amino acid sequence of any one of SEQ ID NOs: 6382-6756, or 8546-8586, representing the respiratory syncytial virus, or the metapneumovirus, respectively, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0101] In some embodiments, the target peptide is an accessory protein of a human parainfluenza virus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the accessory protein has an amino acid sequence of any one of SEQ ID NOs: 11486-12188, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0102] In some embodiments, the target peptide is a non-structural protein of a respiratory syncytial virus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0103] In some embodiments, the target peptide is a B cell epitope of a respiratory syncytial virus, a metapneumovirus, or a human parainfluenza virus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the B cell epitope has an amino acid sequence of any one of SEQ ID NOs: 6757-7019 or 12189-12193, representing the respiratory syncytial virus or the human parainfluenza virus respectively, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0104] In some embodiments, the target peptide is a T cell epitope of a respiratory syncytial virus, a metapneumovirus, or a human parainfluenza virus, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof. In some embodiments, the T cell epitope has an amino acid sequence of any one of SEQ ID NOs: 7020-7371, 8587-8627, or 12194-12216, representing the respiratory syncytial virus, the metapneumovirus, or the human parainfluenza virus respectively, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0105] In some embodiments, the target peptide has an amino acid sequence of any one of SEQ ID NOs: 87-7371 or 7372-12216, including its codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0106] In another aspect, provided herein is a lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid according to any of the preceding embodiments or paragraphs. In some embodiments, the cationic lipid comprises an ionizable lipid. In some embodiments, the cationic lipid is present in an amount from 10 mol percent to 70 mol percent. In some embodiments, the phospholipid is present in an amount from 2 mol percent to 65 mol percent. In some embodiments, the sterol is present in an amount from 20 mol percent to 65 mol percent. In some embodiments, the PEG-lipid is present in an amount from 0.2 mol percent to 2.0 mol percent. In some embodiments, the lipid nanoparticle composition additionally comprises an ionizable polymer. In some embodiments, the ionizable polymer is present in an amount from 1 mol percent to 25 mol percent. In some embodiments, the ionizable polymer is selected from the group comprising a chitosan, chitosan derivatives, cellulose derivatives, a poly-L-lysine (PLL), a protamine, a polyethyleneimine, and / or their derivatives or a combination thereof. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or SM-102, or ALC-0315, or a combination thereof.
[0107] In some aspect, provided herein is a vaccine comprising a lipid nanoparticle, wherein the lipid nanoparticle comprises a cationic lipid, a phospholipid, a sterol, a PEG- lipid, and the multisubunit nucleic acid according to any of the preceding embodiments or paragraphs. In some embodiments, the cationic lipid comprises an ionizable lipid. In some embodiments, the cationic lipid is present in an amount from 10 mol percent to 70 mol percent. In some embodiments, the phospholipid is present in an amount from 2 mol percent to 65 mol percent. In some embodiments, the sterol is present in an amount from 20 mol percent to 65 mol percent. In some embodiments, the PEG-lipid is present in an amount from 0.2 mol percent to 2.0 mol percent. In some embodiments, the lipid nanoparticle composition additionally comprises an ionizable polymer. In some embodiments, the ionizable polymer is present in an amount from 1 mol percent to 25 mol percent. In some embodiments, the ionizable polymer is selected from the group comprising a chitosan, chitosan derivatives, cellulose derivatives, a poly-L-lysine (PLL), a protamine, a polyethyleneimine, and / or their derivatives or a combination thereof. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or SM-102, or ALC-0315, or a combination thereof. In another aspect, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof the multisubunit nucleic acid disclosed herein.
[0108] In another aspect, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof the multisubunit peptide disclosed herein.
[0109] In another aspect, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof the lipid nanoparticle or the lipid nanoparticle composition disclosed herein.
[0110] In another aspect, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof, the vaccine comprising the multisubunit nucleic acid disclosed herein.
[0111] In another aspect, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof, the vaccine comprising the multisubunit peptide disclosed herein.
[0112] In another aspect, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof, the vaccine comprising the lipid nanoparticle or the lipid nanoparticle composition disclosed herein.
[0113] In another aspect, provided herein is use of the multisubunit nucleic acid disclosed herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0114] In another aspect, provided herein is use of the vaccine comprising the multisubunit nucleic acid disclosed herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0115] In another aspect, provided herein is use of the lipid nanoparticle composition disclosed herein in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0116] In another aspect, provided herein is use of the vaccine comprising the lipid nanoparticle composition disclosed herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0117] In another aspect, provided herein is a multisubunit peptide encoded by the multisubunit nucleic acid disclosed herein. In another aspect, provided herein is a multisubunit peptide comprising two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, or a combination thereof, wherein one polypeptide is connected to another polypeptide by a cleavage peptide, wherein the multisubunit peptide includes a signal peptide upstream (amino-terminus of one or more of the polypeptides, wherein the target peptide is obtained or derived from:
[0118] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus.
[0119] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0120] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0121] (d) a combination of (a), (b), and / or (c).
[0122] In some embodiments, the signal peptide is present on the amino-terminus of the first polypeptide. In some embodiments, the signal peptide is present on the aminoterminus of some or each of the polypeptides.
[0123] In another aspect, provided herein is a polypeptide nanoparticle comprising at least 2 or up to 500 polypeptides disclosed herein. In some embodiments, the polypeptides are homologous polypeptides, heterologous polypeptides, oligomeric complexes, polypeptide clusters, or combination thereof. In some embodiments, the polypeptide nanoparticle is icosahedral, helical, spherical, rod-like or combination thereof.
[0124] In another aspect, provided herein is a multisubunit nucleic acid sequence comprising two or more polynucleotide sequences, wherein some or all polynucleotide sequences comprises either a target sequence, a linker sequence, and a self-assembling sequence, or a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, or a combination thereof, wherein one polynucleotide sequence is connected to another polynucleotide sequence by a cleavage sequence, wherein the multisubunit nucleic acid sequence includes a signal sequence upstream of one or more of the polynucleotide sequences, wherein the target sequence is obtained or derived from: (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus.
[0125] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0126] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0127] (d) a combination of (a), (b), and / or (c).
[0128] In some embodiments, the linker sequence connects the signal sequence with the first polynucleotide sequence. In some embodiments, the signal sequence is present upstream of all or some of the polynucleotide sequences. In some embodiments, the signal sequence is present upstream of the first polynucleotide sequence. In some embodiments, the signal sequence is present upstream of all polynucleotide sequences. In some embodiments, the linker sequence connects the target sequence with the self-assembling sequence in a polynucleotide sequence. In some embodiments, one linker sequence connects the cleavage sequence with the target sequence and another linker sequence connects the target sequence with the self-assembling sequence in a polynucleotide sequence. In some embodiments, the multisubunit nucleic acid sequence is a DNA or an RNA. In some embodiments, the multisubunit nucleic acid sequence is an mRNA. In some embodiments, the multisubunit nucleic acid sequence encodes a multisubunit peptide. In some embodiments, the multisubunit nucleic acid sequence is encapsulated or formulated in a lipid nanoparticle composition. In some embodiments, the multisubunit nucleic acid sequence is obtained or synthesized through one or more in vitro transcription (IVT) process. In some embodiments, the multisubunit nucleic acid sequence (for example mRNA) is synthesized or obtained through a single in vitro transcription (IVT) process or step.
[0129] In some embodiments, the disclosure relates to a multisubunit nucleic acid sequence encoding a multisubunit peptide described herein.
[0130] In some embodiments, the disclosure relates to a vaccine comprising a multisubunit nucleic acid sequence encoding a multisubunit peptide described herein.
[0131] In some embodiments, the disclosure relates to a multisubunit nucleic acid sequence encoding a multisubunit peptide comprising two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a selfassembling peptide, or a linker peptide, a target peptide, a linker peptide, and a selfassembling peptide, or a combination thereof, wherein one polypeptide is connected to another polypeptide by a cleavage peptide, wherein the multisubunit peptide includes a signal peptide upstream of one or more of the polypeptides, wherein the target peptide is obtained or derived from:
[0132] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0133] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0134] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0135] (d) a combination of (a), (b), and / or (c).
[0136] In some embodiments, a signal peptide is present upstream (amino-terminus) of all or some of the polypeptides. In some embodiments, a signal peptide is present upstream (amino-terminus) of the first polypeptide. In some embodiments, a signal peptide is present upstream (amino-terminus) of all polypeptides.
[0137] In some embodiments, the disclosure relates to a vaccine comprising a multisubunit nucleic acid sequence encoding a multisubunit peptide, wherein the multisubunit peptide comprises two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, or a combination thereof, wherein one polypeptide is connected to another polypeptide by a cleavage peptide, wherein the multisubunit peptide includes a signal peptide upstream of one or more of the polypeptides, wherein the target peptide is obtained or derived from:
[0138] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0139] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0140] (d) a combination of (a), (b), and / or (c).
[0141] In some embodiments, a signal peptide is present upstream (amino-terminus) of all or some of the polypeptides. In some embodiments, a signal peptide is present upstream (amino-terminus) of the first polypeptide. In some embodiments, a signal peptide is present upstream (amino-terminus) of all polypeptides.
[0142] In some embodiments, the disclosure also relates to a multisubunit peptide comprising two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, or a combination thereof, wherein one polypeptide is connected to another polypeptide by a cleavage peptide, wherein the multisubunit peptide includes a signal peptide upstream (amino-terminus) of one or more of the polypeptides, wherein the target peptide is obtained or derived from:
[0143] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0144] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0145] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0146] (d) a combination of (a), (b), and / or (c).
[0147] In some embodiments, a signal peptide is present upstream (amino-terminus) of all or some of the polypeptides. In some embodiments, a signal peptide is present upstream (amino-terminus) of the first polypeptide. In some embodiments, a signal peptide is present upstream (amino-terminus) of all the polypeptides.
[0148] In some embodiments, the disclosure also relates to a vaccine comprising a multisubunit peptide, wherein the multisubunit peptide comprises two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, or a combination thereof, wherein one polypeptide is connected to another polypeptide by a cleavage peptide, wherein the multisubunit peptide includes a signal peptide upstream (amino-terminus) of one or more of the polypeptides, wherein the target peptide is obtained or derived from:
[0149] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0150] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0151] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0152] (d) a combination of (a), (b), and / or (c).
[0153] In some embodiments, a signal peptide is present upstream (amino-terminus) of all or some of the polypeptides. In some embodiments, a signal peptide is present upstream (amino-terminus) of the first polypeptide. In some embodiments, a signal peptide is present upstream (amino-terminus) of all the polypeptides.
[0154] In some embodiments, the linker peptide connects the signal peptide with the first polypeptide in a multisubunit peptide. In some embodiments, the linker peptide connects the target peptide with the self-assembling peptide in a polypeptide. In some embodiments, one linker peptide connects the cleavage peptide with the target peptide and another linker peptide connects the target peptide with the self-assembling peptide in a polypeptide.
[0155] In some embodiments, the multisubunit peptide comprises homologous polypeptides. In some embodiments, the multisubunit peptide comprises heterologous polypeptides. In some embodiments, the multisubunit peptide comprises homologous polypeptides, or heterologous polypeptides. In some embodiments, the disclosure relates to a polypeptide nanoparticle comprising one or more homologous polypeptides, one or more heterologous polypeptides, one or more oligomeric complexes, one or more polypeptide clusters, or a combination thereof. In some embodiments, the homologous polypeptides, heterologous polypeptides, oligomeric complexes, or polypeptide clusters comprise either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, or a combination thereof, wherein the target peptide is obtained or derived from: (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus.
[0156] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0157] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0158] (d) a combination of (a), (b), and / or (c).
[0159] In some embodiments, the polypeptides in a polypeptide nanoparticle may also have some residues (amino acids) of the cleavage peptide.
[0160] In some embodiments, the disclosure relates to a polypeptide nanoparticle formed from the self-assembly of two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide or a combination thereof, wherein the target peptide is obtained or derived from:
[0161] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus
[0162] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0163] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0164] (d) a combination of (a), (b), and / or (c).
[0165] In some embodiments, the polypeptides in the polypeptide nanoparticle may also have some residues (amino acids) of the cleavage peptide. In some embodiments, the polypeptide nanoparticle comprises of homologous polypeptides, heterologous polypeptides, oligomeric complexes, polypeptide clusters, or combination thereof.
[0166] In some aspects, provided herein is the multisubunit nucleic acid sequence described herein, encapsulated or formulated in a lipid nanoparticle composition. In some aspects, the lipid nanoparticle composition comprises a cationic lipid, a phospholipid, a sterol, a PEG lipid and the multisubunit nucleic acid sequence described herein. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0167] In some aspects, provided herein is a vaccine comprising the multisubunit nucleic acid sequence described herein, encapsulated or formulated in a lipid nanoparticle composition. In some aspects, the lipid nanoparticle composition comprises a cationic lipid, a phospholipid, a sterol, a PEG lipid and the multisubunit nucleic acid sequence described herein. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0168] In some other aspects, the lipid nanoparticle composition comprises an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid and the multisubunit nucleic acid sequence described herein. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0169] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof the multisubunit nucleic acid sequence as described herein. In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a vaccine comprising the multisubunit nucleic acid sequence as described herein.
[0170] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid and the multisubunit nucleic acid sequence as described herein. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0171] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a lipid nanoparticle composition comprising an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid and the multisubunit nucleic acid sequence as described herein. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof. In some aspects, the disclosure relates to use of a lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid sequence as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0172] In some aspects, the disclosure relates to use of a vaccine comprising a lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid sequence as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0173] In some embodiments, the target sequence is obtained or derived from:
[0174] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus, including the codon optimized sequences, fragments, variants, mutants, comparable equivalent, or functional analog thereof.
[0175] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus, including the codon optimized sequences, fragments, variants, mutants, comparable equivalent, or functional analog thereof
[0176] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, including the codon optimized sequences, fragments, variants, mutants, comparable equivalent, or functional analog thereof, or
[0177] (d) a combination of (a), (b), and / or (c).
[0178] In some embodiments, the target sequence is modified or unmodified. In some embodiments, the target sequence encodes a target peptide obtained or derived from:
[0179] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus, including the codon optimized sequences, fragments, variants, mutants, comparable equivalent, or functional analog thereof,
[0180] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, including the codon optimized sequences, fragments, variants, mutants, comparable equivalent, or functional analog thereof, or
[0181] (d) a combination of (a), (b), and / or (c).
[0182] In some embodiments, the target sequence or target peptide is modified or unmodified. In some embodiments, the target peptide regulates or modulates cellular functions. In some embodiments, the target peptide has immunostimulatory or immunomodulatory effect.
[0183] In some embodiments, the self-assembling sequence encodes a self-assembling peptide. In some embodiments, the self-assembling peptide includes, but not limited to, lumazine synthase, MS2 coat protein, hepatitis B surface antigen (HBsAg) from Hepatitis B Virus, hepatitis B core antigen (HBcAg) from hepatitis B virus, human papillomavirus LI (HPV LI) protein, matrix protein Ml from influenza A virus, ferritin, riboflavin synthase, dihydrolipoyl acetyltransferase (E2p), or a combination thereof, including their fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0184] In some embodiments, the ferritin peptide is obtained or derived from Listeria innocua or Helicobacter pylori. In some embodiments, the ferritin peptide is a Listeria innocua ferritin or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the ferritin peptide is a Helicobacter pylori ferritin or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the dihydrolipoyl acetyltransferase (E2p) is obtained or derived from Bacillus stearothermophilus, or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the lumazine synthase is obtained or derived from Aquifex species (for example, Aquifex aeolicus) or Bacillus species (for example, Bacillus subtilis), or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the MS2 coat protein is obtained or derived from Emesvirus zinderi, or its fragment, mutant, variant, comparable equivalent, or functional analogs thereof. In some embodiments, the linker sequence encodes a linker peptide. In some embodiments, the linker peptide connects the target peptide with the self-assembling peptide in a polypeptide. In some embodiments, the linker peptide connects the signal peptide with the first polypeptide. In some embodiments, one linker peptide connects the cleavage peptide with the target peptide and another linker peptide connects the target peptide with the self-assembling peptide in a polypeptide. In some embodiments, the linker peptide connects two signal peptides. In some embodiments, the linker peptide connects two cleavage peptides. The linker peptide may be an amino acid linker, a zipper motif, a foldon, a scaffold or a combination thereof.
[0185] In some embodiments, the cleavage sequence encodes a cleavage peptide. In some embodiments, the cleavage peptide comprises one or more cleavage peptides. The cleavage peptide connects one polypeptide with another polypeptide, for example, adjacent polypeptide. The cleavage peptide carries a cleavage site. In some embodiments, the cleavage peptide carries one or more cleavage sites. In some embodiments, the cleavage peptide facilitates the action of cellular proteases to cleave the multisubunit peptide into individual polypeptides. In some embodiments, the cleavage peptide self cleaves into individual polypeptides. In some embodiments, the cleavage peptide comprises two or more cleavage peptides (for example, cleavage peptide- 1, cleavage peptide-2 and so on), optionally connected via a linker. In some embodiments, the cleavage peptide self cleaves into individual polypeptides or is cleaved by the action of cellular proteases. In some embodiments, the cleavage peptide is a substrate for cellular proteases. In some embodiments, the cleavage peptide is a substrate for golgi specific proteases. In some embodiments, the cleavage peptide is a self cleaving peptide. In some embodiments, the cleavage peptide comprises two or more cleavage peptides (for example, cleavage peptide - 1, cleavage peptide-2 and so on), optionally linked by a linker peptide, wherein one cleavage peptide is a substrate for cellular proteases and the other cleavage peptide is a self cleaving peptide.
[0186] In some embodiments, the signal sequence encodes a signal peptide. The signal peptide is present upstream (amino-terminus) of one or more polypeptides in a multisubunit peptide. In some embodiments, the signal peptide is present upstream (aminoterminus) of the first polypeptide. In some embodiments, the signal peptide is present upstream (amino-terminus) of some polypeptides. In some embodiments, the signal peptide is present upstream (amino-terminus) of all polypeptides. In some embodiments, the signal peptide transports the multisubunit peptide to cell organelles. In some embodiments, the signal peptide transports the multisubunit peptide to golgi body or golgi apparatus. In some embodiments, the signal peptide is a golgi targeting signal peptide.
[0187] In some aspects, the present disclosure also includes a method of transforming a cell with the multisubunit nucleic acid sequence as described herein.
[0188] BRIEF DESCRIPTION OF THE DRAWINGS
[0189] Figure 1 - shows representative schematic illustration of multisubunit nucleic acid sequence wherein each polynucleotide sequence (PS) comprises a target sequence (TS), a linker sequence (LS), and a self-assembling sequence (SAS) or a linker sequence (LS), a target sequence (TS), a linker sequence (LS), and a self-assembling sequence (SAS), wherein the target sequence is obtained or derived from:
[0190] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0191] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0192] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0193] (d) a combination of (a), (b), and / or (c).
[0194] Multiple polynucleotide sequences are connected through a cleavage sequence (CS) such that between any two polynucleotide sequences there is present a cleavage sequence. The multisubunit nucleic acid sequence has a signal sequence (SS) upstream of the first polynucleotide sequence. The letter ‘n’ in figure 1 represents any number between 1 to 98. The multisubunit nucleic acid sequence may additionally have 5’ cap and 3’ poly(A) tail. This multisubunit nucleic acid sequence encodes corresponding multisubunit peptide depicted in Figure 2.
[0195] Figure 2 - shows representative schematic illustration of multisubunit peptide wherein each polypeptide (PP) comprises a target peptide (TP), a linker peptide (LP), and a self-assembling peptide (SAP) or a linker peptide (LP), a target peptide (TP), a linker peptide (LP), and a self-assembling peptide (SAP), wherein the target peptide is obtained or derived from: (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0196] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0197] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0198] (d) a combination of (a), (b), and / or (c).
[0199] Multiple polypeptides are connected through a cleavage peptide (CP) such that between any two polypeptides there is present a cleavage peptide. The multisubunit peptide has a signal peptide (SP) on the N-terminus of the first polypeptide. The letter ‘n’ represents any number between 1 to 98.
[0200] DESCRIPTION
[0201] The present disclosure relates to multisubunit nucleic acid sequences, multisubunit peptides, polypeptide nanoparticle, and their compositions for vaccine and therapeutic purpose against respiratory syncytial virus (RSV), human metapneumovirus (HMPV) and human parainfluenza virus (HPIV).
[0202] Unless defined otherwise, technical, and scientific terms used herein have the same meaning as commonly understood by one of person skill in the art. Some of the terms are defined briefly here below; the definitions should not be construed in a limiting sense.
[0203] The singular forms “a”, “an” and “the” as used in the specification also include plural aspects unless the context dictates otherwise. Similarly, any singular term used in the specification also mean plural or vice versa unless the context dictates otherwise. As used herein in the claim(s), when used in conjunction with the word “comprising”, the words “a” or “an” may mean one or more than one. As used herein “another” may mean at least a second or more.
[0204] It must be noted that the words “comprising” or any of its form such as “comprise” or “comprises”, “having” or any of its forms such as “have” or “has”, “including” or any of its forms such as “include” or “includes”, or “containing” or any of its forms such as “contain” or “contains” are open-ended and do not exclude additional unrecited elements or method steps. Wherever any quantity or range is stated one skilled in the art will recognize that quantity or range within 10 or 20 percent of the stated values can also be expected to be appropriate and reasonable and included within the scope of the invention.
[0205] Unless otherwise defined herein, scientific, and technical terms used in connection with the present invention shall have the meanings that are commonly understood by those of ordinary skilled in the art. Generally, nomenclatures used in connection with, and techniques of, cell and tissue culture, molecular biology, immunology, microbiology, protein, adjuvant, pharmaceutical biotechnology, and biopharmaceutical manufacturing described herein are those well known and commonly used in the art. The methods and techniques of the present invention are generally performed according to conventional methods well known in the art and as described in various general and more specific references that are cited and discussed throughout the present specification.
[0206] The term “composition”, “formulation”, “lipid nanoparticle composition”, or “lipid nanoparticle” has been used interchangeably to mean a nanoparticle, nanostructure, vesicle, liposome, composition or formulation comprising one or more lipid components (for example, a cationic lipid, a phospholipid, a sterol, and a PEG-lipid), and / or an ionizable polymer component. In some embodiments, the lipid nanoparticle comprises a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and a multisubunit nucleic acid. In some embodiments, the lipid nanoparticle comprises one or more lipid components, an ionizable polymer component, and a multisubunit nucleic acid. In some embodiments, the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof. In some embodiments, the multisubunit nucleic acid associated with the lipid nanoparticle is a DNA, an mRNA, a micro RNA, a small interfering RNA, a small nucleolar RNA, a small nuclear RNA, a long non-coding RNA or a combination thereof. In some embodiments, the lipid nanoparticle composition may contain one or more pharmaceutical carriers or excipients, such as but not limited to, buffering agents, stabilizers, tonicity modifiers, surfactants, chelating agents, salts, anti-oxidants, diluents, and / or preservatives or a combination thereof. The term lipid nanoparticle also denotes lipid nanoparticles that are devoid of any encapsulated multisubunit nucleic acid (empty lipid nanoparticles or ghost lipid nanoparticles). In some embodiments, the lipid nanoparticle composition comprises lipid nanoparticles with encapsulated multisubunit nucleic acid as well as empty lipid nanoparticles. The term “therapeutic”, “therapeutic agent”, “prophylactic”, “prophylactic agent”, or “drug” has been used interchangeably to mean a compound (such as multisubunit nucleic acid sequence) or composition (such as a lipid nanoparticle composition described herein) having a biological effect or a combination of biological effects that prevents, inhibits, eliminates or prevents the progression of a disease or other aberrant biological processes in a subject, for example, an animal or human.
[0207] The term “preventing” is art-recognized, and when used in relation to a condition, such as an infection is well understood in the art, and includes administration of a composition, which reduces the frequency or severity, or delays the onset, of one or more symptoms of the medical condition in a subject relative to a subject who does not receive the composition. Thus, the prevention of a condition, such as an infection, includes, for example, the reduction of the frequency or severity of one or more symptoms of the medical condition in a population of patients receiving a therapy relative to a control population that did not receive the therapy, e.g., by a statistically and / or clinically significant amount. Similarly, the prevention of an infection includes reducing the likelihood that a patient receiving a therapy will develop the infection or related symptoms, relative to a patient who does not receive the therapy.
[0208] The term “molar percent”, “mol percent”, “molar %”, or “mol %” have been used interchangeably to mean number of moles of a component expressed as percentage relative to total moles of all lipid components present in the lipid nanoparticle compositions described herein. For example, 50 mol % cationic lipid means, 50 mol % of cationic lipid is present in the lipid nanoparticle composition and other lipid components together constitute remaining 50 mol % such that the total amount of all the lipid components constitute 100 mol %. In some embodiments, mol % also denotes to mean number of moles of a component expressed as percentage relative to total moles of all lipid components (such as cationic lipid, phospholipid, sterol and PEG-lipid) and ionizable polymer component(s) present in the lipid nanoparticle composition described herein. For example, 50 mol % of cationic lipid means, 50 mol % of cationic lipid is present in the lipid nanoparticle composition and other lipids components and ionizable polymer components together constitute the remaining 50 mol % such that the total amount of all the lipid components and ionizable polymer components constitute 100 mol %.
[0209] The term “N / P ratio”, “N:P ratio”, “lipid to nucleic acid ratio”, or “cationic lipid to nucleic acid ratio” have been used interchangeably herein and means the ratio (molar ratio) of the positive charges in the cationic lipid relative to the negative charges in the nucleic acid (such as the multisubunit nucleic acid sequence disclosed herein) in a lipid nanoparticle. In some embodiments, the N / P ratio refers to the ratio of protonable nitrogen present in the cationic lipid relative to the phosphate present in the nucleic acid in a lipid nanoparticle. In some embodiments, the N / P ratio is between 1 to 18 (i.e., 1 :1 to 18:1). For example, a N / P ratio of 18 refers to the presence of 18 protonable nitrogen of the cationic lipid relative to 1 phosphate of the nucleic acid in a lipid nanoparticle.
[0210] The terms “antibody” and “antibodies” have been used interchangeably herein and means any antibody or antibody fragment (whether produced naturally or recombinantly) which retains antigen binding activity. This includes a monoclonal or polyclonal antibody, a single chain antibody, a Fab fragment of a monoclonal or polyclonal antibody, a chimeric antibody, a humanized antibody, a human antibody, a bispecific antibody, a multispecific antibody, or a nanobody.
[0211] The term “buffer” as used herein means those agents that maintains the pH of a solution in a desired range.
[0212] The term “cell” as used herein means a single cell or a population of cells or plurality of cells.
[0213] The term “biologically effective amount” or “therapeutically effective amount” as used herein means an amount of an agent, for example, a therapeutic, drug, therapeutic agent, prophylactic agent, diagnostic agent, composition, etc., that is sufficient, when administered to a subject suffering from or susceptible to an infection, disease, disorder, and / or condition, to treat, prevent, diagnose, improve symptoms of, and / or delay the onset of the infection, disease, disorder, and / or condition. A therapeutically effective amount herein may vary according to factors such as the disease state, age, sex, and weight of the patient.
[0214] As used herein, the term “treating” or “treatment” includes reducing, arresting, or reversing the symptoms, clinical signs, or underlying pathology of a condition to stabilize or improve a subject’s condition or to reduce the likelihood that the subject’s condition will worsen as much as if the subject did not receive the treatment. Treatment may be administered to a subject who does not exhibit signs of a disease and / or exhibits only early signs of the disease for the purpose of decreasing the risk of developing pathology associated with the disease. The term “subject” as used herein refers to an animal or human, for example, a living mammal and may be interchangeably used with the term “patient”. Examples of mammals include, but are not limited to, any member of the mammalian class: humans, non-human primates such as chimpanzees, and other apes and monkey species; farm animals such as cattle, horses, sheep, goats, swine; domestic animals such as rabbits, dogs, and cats; laboratory animals including rodents, such as rats, mice, ferrets, and guinea pigs, and the like. The term does not denote a particular age or gender.
[0215] As used herein, an individual “at risk” of developing a particular disease, disorder, or condition may or may not have detectable disease or symptoms of disease, and may or may not have displayed detectable disease or symptoms of disease prior to the treatment methods described herein. “At risk” denotes that an individual has one or more risk factors, which are measurable parameters that correlate with development of a particular disease, disorder, or condition, as known in the art. An individual having one or more of these risk factors has a higher probability of developing a particular disease, disorder, or condition than an individual without one or more of these risk factors.
[0216] The term “disease” as used herein, means an interruption, cessation, or disorder of body function, system, or organ. Non limiting examples of disease include malignant diseases, autoimmune diseases, inherited diseases, metabolic disorders, or infectious diseases.
[0217] The term “vaccine” as used herein means a substance or composition comprising an antigen or immunogen for eliciting an immune response in a subject against the antigen or the immunogen. The term vaccine is also understood to mean a substance or composition comprising an antigen or immunogen that activates or stimulates an immune cell. In some cases, the antigen or immunogen is a peptide, a protein, a polysaccharide, or a combination thereof. In some cases, the antigen or immunogen is encoded by a nucleic acid, for example, a DNA, an RNA, or an mRNA. In some embodiments, the vaccine comprises a nucleic acid that encodes an antigen or an immunogen. In some embodiments, the vaccine comprises a multisubunit nucleic acid as described herein.
[0218] As used herein, administration “conjointly” with another compound or composition includes simultaneous administration and / or administration at different times. Conjoint administration also encompasses administration as a co-formulation or administration as separate compositions, including at different dosing frequencies or intervals, and using the same route of administration or different routes of administration. The term “multisubunit nucleic acid sequence” or “multisubunit nucleic acid” have been used interchangeably herein and means two or more polynucleotide sequences wherein some or all polynucleotide sequences comprises either a target sequence, a linker sequence, and a self-assembling sequence, or a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, or a combination thereof, wherein one polynucleotide sequence is connected to another polynucleotide sequence by a cleavage sequence, wherein the multisubunit nucleic acid sequence includes a signal sequence upstream of one or more polynucleotide sequences, wherein the target sequence is as described herein. In some embodiments, the signal sequence is present upstream of the first polynucleotide sequence. In some embodiments, the signal sequence is present upstream of some polynucleotide sequences. In some embodiments, the signal sequence is present upstream of each of the polynucleotide sequences. In some embodiments, the polynucleotide sequence comprises a target sequence, a linker sequence, and a selfassembling sequence, wherein the target sequence is as described herein. In some embodiments, the polynucleotide sequence comprises a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, wherein the target sequence is as described herein. Thus, in some embodiments, one linker sequence connects the cleavage sequence with the target sequence and another linker sequence connects the target sequence with the self-assembling sequence in a polynucleotide sequence. In some embodiments, the linker sequence connects the signal sequence with the polynucleotide sequence. As illustrated in figure 1 multisubunit nucleic acid sequence may comprise multiple repeats of polynucleotide sequences wherein each polynucleotide sequence may comprise either a target sequence, a linker sequence, and a self-assembling sequence, or a linker sequence, a target sequence, a linker sequence and a self-assembling sequence, or a combination thereof, wherein the target sequence is as described herein, such that total number of polynucleotide sequences in a multisubunit nucleic acid sequence are not more than 100. In some embodiments, the linker sequence connects the signal sequence with the first polynucleotide sequence. In some embodiments, the signal sequence is present upstream of each of some or all of the polynucleotide sequences. In some embodiments, the multisubunit nucleic acid sequence is obtained or synthesized through single in vitro transcription (IVT) process or step. The multisubunit nucleic acid sequence encodes multisubunit peptide. The terms “multisubunit peptide” as used herein means two or more polypeptides wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide and a selfassembling peptide, or a combination thereof, wherein one polypeptide is connected to another polypeptide by a cleavage peptide, wherein the multisubunit peptide includes a signal peptide upstream (amino-terminus) of one or more polypeptides, wherein the target peptide is as described herein. In some embodiments, the signal peptide is present on the amino-terminus of the first polypeptide. In some embodiments, the signal peptide is present on the amino-terminus of some polypeptides. In some embodiments, the signal peptide is present on the amino-terminus of each of the polypeptides. In some embodiments, the polypeptide comprises a target peptide, a linker peptide, and a selfassembling peptide, wherein the target peptide is as described herein. In some embodiments, the polypeptide comprises a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, wherein the target peptide is as described herein. Thus, in some embodiments, one linker peptide connects the cleavage peptide with the target peptide and another linker peptide connects the target peptide with the self-assembling peptide in a multisubunit peptide. In some embodiments, the multisubunit peptide either comprises a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, or a combination thereof, wherein the target peptide is as described herein, such that the total number of polypeptides in a multisubunit peptide are not more than 100. In some embodiments, the linker peptide connects the signal peptide with the polypeptide. In some embodiments, the signal peptide is present on the amino-terminus of each of some or all polypeptides. The multisubunit peptide may comprise homologous polypeptides or heterologous polypeptides.
[0219] The term “polynucleotide sequence” as used herein means a sequence of nucleotides that encodes a polypeptide.
[0220] The terms “protein” or “peptide” have been used interchangeably herein and mean a polymer of amino acids linked through peptide bonds, but do not imply any specific length. The term also includes fusion proteins, muteins, analogs or modified forms.
[0221] The term “polypeptide” as used herein means a sequence of amino acids that comprises either a target peptide, a linker peptide, and a self-assembling peptide or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, wherein the target peptide is as described herein. In some embodiments, the polypeptide comprises a target peptide, a linker peptide, and a self-assembling peptide, wherein the target peptide is as described herein. In some embodiments, the polypeptide comprises a linker peptide, a target peptide, a linker peptide, and a self- assembling peptide, wherein the target peptide is as described herein. In some embodiments, the polypeptide may have some residues (amino acids) of cleavage peptide. In some embodiments, the polypeptide comprises a signal peptide.
[0222] The term “target sequence” as used herein means a sequence of nucleotides that encodes a target peptide obtained or derived from:
[0223] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus.
[0224] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0225] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0226] (d) a combination of (a), (b), and / or (c).
[0227] The term “target peptide” as used herein means a sequence of amino acids obtained or derived from:
[0228] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0229] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0230] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0231] (d) a combination of (a), (b), and / or (c).
[0232] In some embodiments, the target peptides in two or more polypeptides are identical i.e., homologous polypeptides. In some other embodiments, the target peptides in two or more polypeptides are different i.e., heterologous polypeptides. The term “signal sequence” as used herein means a sequence of nucleotides that encodes a signal peptide.
[0233] The term “signal peptide” as used herein means a sequence of amino acids that transports the multisubunit peptide to specific cell organelles. In some embodiments the signal peptide transports the multisubunit peptide to golgi apparatus or golgi body. The signal peptide is present on the N-terminus (amino-terminus) of one or more polypeptides. In some embodiments, the signal peptide is present on the N-terminus of some or all polypeptides. In some embodiments, the signal peptide is present on the N-terminus of the first polypeptide. In some embodiments, the signal peptide is present on the N-terminus of some polypeptides. In some embodiments, the signal peptide is present on the N-terminus of all polypeptides. In some embodiments, the signal peptide is encoded by signal sequence. In some embodiments, the signal peptide is a golgi targeting signal peptide.
[0234] The term “cleavage sequence” as used herein means a sequence of nucleotides that encodes a cleavage peptide.
[0235] The term “cleavage peptide” as used herein means a sequence of amino acids that facilitates the action of cellular proteases to cleave the multisubunit peptide into individual polypeptides or self cleaves into individual polypeptides. The cleavage peptide is present between any two polypeptides. It connects one polypeptide with another polypeptide, for example, adjacent polypeptide. The cleavage peptide carries one or more cleavage sites. In some embodiments, the cleavage peptide is a substrate for proteases. In some embodiments, cleavage peptide undergoes self cleavage to result in individual polypeptides. In some embodiments, the cleavage peptide is a substrate for golgi specific proteases. In some embodiments, the cleavage peptide comprises one or more cleavage peptides, for example, cleavage peptide- 1, cleavage peptide-2 and so on. In some embodiments, the cleavage peptide optionally comprises a linker peptide between two cleavage peptides. In some embodiments, the cleavage peptide self cleaves into individual polypeptides or is cleaved by the action of cellular proteases. In some embodiments, the cleavage peptide is a substrate for cellular proteases. In some embodiments, the cleavage peptide is a substrate for golgi specific proteases. In some embodiments, the cleavage peptide is a self cleaving peptide. In some embodiments, the cleavage peptide comprises two or more cleavage peptides (for example, cleavage peptide- 1, cleavage peptide-2 and so on), optionally linked by a linker peptide, wherein one cleavage peptide is a substrate for cellular proteases and the other cleavage peptide is a self cleaving peptide. The term “linker sequence” as used herein means a sequence of nucleotides that encodes a linker peptide.
[0236] The term “linker peptide” or “peptide linker” have been used interchangeably to mean a sequence of amino acids that either connects the target peptide with the selfassembling peptide, connects the signal peptide with the target peptide, connects the cleavage peptide with the target peptide, connects the signal peptide with the polypeptide, or connects two cleavage peptides. In some embodiments, the linker peptide connects the signal peptide with the polypeptide. In some embodiments, the linker peptide connects the target peptide with the self-assembling peptide in a polypeptide. In some embodiments, one linker peptide connects the cleavage peptide with the target peptide and another linker peptide connects the target peptide with the self-assembling peptide in a polypeptide. In some embodiments, one linker peptide connects the cleavage peptide with the target peptide and another linker peptide connects the target peptide with the self-assembling peptide. In some embodiments, one linker peptide connects the cleavage peptide with the target peptide and another linker peptide connects the signal peptide with the target peptide. In some embodiments, the linker peptide connects two cleavage peptides. In some embodiments, the linker peptide is an amino acid linker, a zipper motif, a foldon, a scaffold or a combination thereof. In some embodiments, the linker peptide is an amino acid linker. In some embodiments, the linker peptide is a foldon. In some embodiments, the linker peptide is a zipper motif. In some embodiments, the linker peptide is a scaffold. In some embodiments, the linker peptide comprises an amino acid linker and a foldon. In some embodiments, the linker peptide comprises an amino acid linker and a zipper motif. In some embodiments, the linker peptide comprises an amino acid linker and a scaffold. In some embodiments, the linker peptide comprises a zipper motif and a scaffold. In some embodiments, the linker peptide comprises a foldon and a scaffold. In some embodiments, the linker peptide comprises an amino acid linker, a zipper motif, and a scaffold. In some embodiments, the linker peptide comprises an amino acid linker, a foldon, and a scaffold.
[0237] The term “amino acid linker sequence” as used herein means a sequence of nucleotides that encodes an amino acid linker.
[0238] The term “amino acid linker” as used herein means a sequence of amino acids that provides structural integrity to polypeptide such that the components of the polypeptide remain, as far as possible, in their native or stable conformation. In some embodiments, amino acid linker also helps in orientation of a polypeptide such that the domains or epitopes on the target peptide are exposed or displayed for interaction or communication with cells or biomolecules or immune system in the absence of foldon or scaffold. In some embodiments, the amino acid linker connects two cleavage peptides. Some of the nonlimiting examples of amino acid linkers includes, glycine serine linker, glycine proline linker, glycine threonine linker, alanine serine linker, any combination of two amino acids or a combination thereof. In some embodiments, amino acid linker is about 2-49 amino acid long.
[0239] The term “glycine serine linker sequence” as used herein means a sequence of nucleotides that encodes a glycine serine linker.
[0240] The term “glycine serine linker” as used herein means a sequence of amino acid comprising one or more glycine (G) and serine (S) in any combinations without any preference of order or limitation on number of appearances of either glycine or serine. In some embodiments, the glycine serine linker is few amino acids in length to several amino acids in length.
[0241] The term “zipper sequence” as used herein means a sequence of nucleotides that encodes a zipper motif.
[0242] The term “zipper motif’ or “zipper peptide” as used herein means a sequence of amino acids that facilitates homologous polypeptides or heterologous polypeptides to come together or associate to form a polypeptide cluster.
[0243] The term “foldon sequence” as used herein means a sequence of nucleotides that encodes a foldon.
[0244] The term “foldon” as used herein means a sequence of amino acids that enables two or more homologous polypeptides to organise to form an oligomeric complex. In some embodiments, the foldon also helps in orientation of a polypeptide such that the domains or epitopes on the target peptide are exposed or displayed for interaction or communication with cells or biomolecules or immune system.
[0245] The term “scaffold sequence” as used herein means a sequence of nucleotides that encodes a scaffold.
[0246] The term “scaffold” as used herein means a sequence of amino acids that provides structural and / or functional integrity or support to the target peptide and helps in orientation of target peptide such that the domains or epitopes of the target peptide are exposed or displayed for interaction or communication with cells or biomolecules or immune system. The term “oligomeric complex” as used herein means a complex formed by two or more homologous polypeptides. In some embodiments, the oligomeric complex has at least two homologous polypeptides, at least three homologous polypeptides, at least four homologous polypeptides, at least five homologous polypeptides, or at least six homologous polypeptides and so on.
[0247] The term “polypeptide cluster” as used herein means a complex formed by interaction of protein domains of two or more homologous polypeptides or two or more heterologous polypeptides orchestrated by the zipper motif. In some embodiments, the polypeptide cluster has at least two, at least three, at least four, at least five, at least six homologous polypeptides, heterologous polypeptides, or combination thereof.
[0248] The term “self- assembling sequence” as used herein means a sequence of nucleotides that encodes a self-assembling peptide.
[0249] The term “self-assembling peptide” as used herein means a sequence of amino acids that enables the polypeptides to self-assemble into a polypeptide nanoparticle.
[0250] The term “self-assembly” or “self-assemble” or “self-assembling” has been used interchangeably to means the ability of polypeptides to undergo multimerization to form a polypeptide nanoparticle. In some embodiments, the polypeptide nanoparticle may have at least two polypeptides (dimer or 2-mer), at least three polypeptides (trimer or 3-mer), at least four polypeptides (tetramer or 4-mer), at least five polypeptides (pentamer or 5-mer), at least six polypeptides (hexamer or 6-mer), at least seven polypeptides (heptamer or 7- mer), at least eight polypeptides (octamer or 8-mer), and so on. In some embodiments, the polypeptide nanoparticle is up to 500-mers. In some embodiments, hydrogen bonds, disulfide bonds, hydrophobic interactions, electrostatic interactions, and / or Van der Waals forces combine to maintain self-assembled structure.
[0251] The term “multimerization” as used herein means association of two or more units of homologous polypeptides, heterologous polypeptides, oligomeric complexes, polypeptide clusters, or their combination.
[0252] The term “polypeptide nanoparticle” as used herein means a nanoparticle formed by self-assembly of polypeptides. In some embodiments, the polypeptide nanoparticle comprises two or more homologous polypeptides, two or more heterologous polypeptides, one or more oligomeric complexes, one or more polypeptide clusters, or a combination thereof. The term “homologous polypeptides” as used herein means polypeptides in a multisubunit peptide that have identical target peptides. For example, if two polypeptides in the multisubunit peptide have identical target peptides they are considered to be homologous polypeptides.
[0253] The term “heterologous polypeptide” as used herein means polypeptides in a multisubunit peptide that have different target peptides. For example, if two polypeptides in the multisubunit peptide have different or non-identical target peptides, they are considered to be heterologous polypeptides.
[0254] The term “upstream”, “amino-terminus”, or “N-terminus” has been used interchangeably in the context of amino acid sequences (protein, peptide, polypeptide, or any other sequence composed of amino acids) or the nucleic acid sequences (DNA, RNA or any other sequence composed of nucleotides) to mean amino end of an amino acid sequence or the 5-prime end of a nucleic acid sequence respectively.
[0255] The term “fragment” as used herein, whether in the context of a nucleic acid, nucleotide, protein, polypeptide, or peptide, means any length of the nucleic acid, protein, polypeptide, or peptide sequence except the full length of the respective nucleic acid, protein, polypeptide, or peptide sequence. Fragment includes such portions of nucleic acid, protein, polypeptide, or peptide that are capable of treating, preventing, diagnosing, improving symptoms of, and / or delay the onset of an infection, disease, disorder, and / or condition. A fragment is also understood to mean an immunogenic fragment of the protein, polypeptide or peptide, or a fragment of nucleic acid encoding an immunogenic fragment of the protein, peptide, or polypeptide.
[0256] The term “variant” as used herein, whether in the context of a nucleic acid, nucleotide, protein, polypeptide, or peptide sequence, means homologs, orthologs, paralogs, mutants or analogs of respective nucleic acid, protein, polypeptide, or peptide sequence.
[0257] The term “mutant” as used herein, whether in the context of a nucleic acid, nucleotide, protein, polypeptide, or peptide sequence, means a sequence which is not a wild type sequence. A mutant is also understood to mean a nucleic acid, nucleotide, protein, polypeptide, or peptide sequence that carries a mutation.
[0258] The term “mutation” as used herein means, a change or modification in the sequence of nucleic acid or amino acid in comparison to a reference sequence and includes insertion, deletion, substitution, or a combination thereof. Mutations are introduced to impart desirable properties upon the nucleic acid, nucleotide, protein, polypeptide, or peptide sequence. This includes, for example, enabling the nucleic acid, nucleotide, protein, polypeptide, or peptide sequence to elicit an immune response while ensuring that any undesirable or deleterious effects are minimized or entirely removed.
[0259] The term “sequence” as used herein means nucleic acid sequences, nucleic acids, polynucleotides, amino acid sequences, proteins, polypeptides, or peptides, depending upon the context in which the term sequence is used, to mean a sequence of nucleotides or amino acids. In the context of nucleic acid, polynucleotide, or nucleotide sequence, the sequence is represented by a single letter code representing the nitrogenous base, for example A, T, G, C, or U. In the context of amino acid sequences, proteins, polypeptides, or peptides the sequence is represented by a single letter amino acid code as generally understood by persons skilled in the art. If the single letter amino acid code is represented by the letter “X”, it means the amino acid at that position is either absent or substituted by any other amino acid. In some embodiments, the sequences representing target sequence or target peptide, may contain a tag, for example, histidine tag, streptavidin tag etc, which may be deleted or removed from the respective sequence before employing the sequence in accordance with the present disclosure. In some embodiments, the sequences representing target sequence or target peptide may contain a signal sequence or signal peptide respectively, which may be deleted or removed from the respective sequence before employing the sequence in accordance with the present disclosure. The deleted or removed signal sequence or signal peptide may be employed in accordance with the present disclosure. The presence of tags, signal sequence or signal peptide, or other similar elements, may easily be recognized by those skilled in the art.
[0260] The term “percentage identity”, “percent identity”, “% age identity”, or “% identity” have been used interchangeably and means the extent of identity between two sequences (e.g. nucleic acid sequences or amino acid sequences). Percent identity can be determined by aligning two sequences, introducing gaps to maximize identity between the sequences. Percent identity should generally be calculated between the same types of sequences for example nucleic acids, i.e. for DNA sequences or RNA sequences or amino acid sequences. The alignment of sequences (nucleic acid or amino acid) can be performed with the appropriate pair wise sequence alignment programs. Identity can be calculated between two sequences by multiplying the number of matches in the pair by 100 and dividing by the length of the aligned region, including gaps. Gaps at the end of sequences are not included, and internal gaps are included in the length. In some embodiments, the nucleic acid sequence or the amino acid sequence, as the case may be, shares at least 50%, at least 55%, at least 60%, at least 65%, at least 70%, at least 75%, at least 80%, at least 81%, at least 82%, at least 83%, at least 84%, at least 85%, at least 86%, at least 87%, at least 88%, at least 89%, at least 90%, at least 91%, at least 92%, at least 93%, at least 94%, at least 95%, at least 96%, at least 97%, at least 98%, at least 99%, or 100% identity with the sequences disclosed herein. In some embodiments, the nucleic acid sequence or the amino acid sequence, as the case may be, shares at least 50% to 60%, at least 60% to 70%, at least 70% to 80%, at least 80% to 90%, or at least 90% to 100% identity with the sequences disclosed herein.
[0261] The term “functional analog” or “functional analogue” have been used interchangeably, whether in the context of nucleic acid sequence or amino acid (protein or peptide) sequence, and mean all sequences which essentially performs similar function compared to the sequence being referred. For example, functional analogs of a target peptide include all sequences, irrespective of their percentage identity, which essentially perform at least one function similar to the function the target peptide performs.
[0262] The term “comparable equivalent” in the context of nucleic acid sequences, amino acid sequences, proteins, polypeptides, or peptides, as disclosed herein, means structurally or functionally similar or identical nucleic acid sequences, amino acid sequences, proteins, polypeptides, or peptides compared to the nucleic acid sequences or amino acid sequences being referred.
[0263] The term “structural protein” as used herein means proteins that are components of the viral particle or structure, for example, capsid proteins, membrane or envelope proteins, proteins packaged within the virus particle etc.
[0264] The term “envelope protein” as used herein means a protein associated with the viral envelope, for example, anchored into, projecting from, or across the lipid layer of the viral envelope. Envelope proteins mediate attachment and fusion. In some embodiments, envelope proteins are glycosylated.
[0265] The term “matrix protein” as used herein means a protein that is found beneath the viral envelope, often acting as a bridge between nucleocapsid or core and the envelope.
[0266] The term “capsid protein” or “nucleocapsid protein” as used herein means a protein that encapsidates or packages the viral genetic material or genome. The term is also understood to mean those proteins which are involved or participate in one or more of the following functions or activities, such as virus assembly, budding or release of virus, mediating attachment to and penetration into the host cells (especially in case of nonenveloped viruses), packaging the genome, etc. The capsid protein or nucleocapsid proteins are the proteins associated with viral capsid or viral nucleocapsid or viral core.
[0267] The term “non-structural protein” as used herein means proteins that are encoded by the virus, but are not the component or part of the mature viral particle or structure, for example, enzymes, transcription factors etc.
[0268] The term “B cell epitope” as used herein means a protein determinant that is recognized by B cell receptors (BCR) and capable of specific binding to an antibody or immunoglobulin.
[0269] The term “T cell epitope” as used herein means a protein determinant derived from an antigen which is presented by an antigen presenting cell (APC) through major histocompatibility complex (MHC) for recognition by a T cell receptor (TCR).
[0270] The term “respiratory syncytial virus”, or “RSV” have been used interchangeably to mean enveloped, single stranded RNA virus belonging to the genus pneumovirus within the family paramyxoviridae. The term also includes subgroups, genotypes, or strains of RSV.
[0271] The term “RSV disease” as used herein means any illness, disease, or condition directly or indirectly caused by respiratory syncytial virus. Such illness, disease or condition include bronchiolitis, pneumonia, respiratory failure, or diseases or conditions that predisposes a person to RSV infection such as cystic fibrosis, congenital heart disease, cancer, age related immunosuppression and, generally, any condition that causes a state of immunosuppression or decreased function of the immune system such as post-operative organ transplantation regimens or premature birth.
[0272] The term “metapneumovirus”, “MPV”, “human metapneumovirus”, “HMPV”, or “hMPV” have been used interchangeably and means a virus belonging to the genus Metapneumovirus within the family Paramyxoviridae. The term also includes, human metapneumovirus, avian metapneumovirus, including their lineages, sub-lineages, and strains.
[0273] The term “human parainfluenza virus”, “HPIV”, or “hPIV” have been used interchangeably and means a virus belonging to the genera Respirovirus, Rubulavirus, or combination thereof, within the family Paramyxoviridae. The term also includes any of the serotypes, subtypes, or strains of human parainfluenza virus, such as but not limited to HPIV-1, HPIV-2, HPIV-3, HPIV-4a, or HPIV-4b.
[0274] Multisubunit nucleic acid sequence and multisubunit peptide
[0275] In the present disclosure, a multisubunit nucleic acid sequence includes two or more polynucleotide sequences wherein some or all polynucleotide sequences comprises either a target sequence, a linker sequence, and a self-assembling sequence, or a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, wherein one polynucleotide sequence is connected to the another polynucleotide sequence by a cleavage sequence, wherein the multisubunit nucleic acid sequence includes a signal sequence upstream of one or more polynucleotide sequences, wherein the target sequence is obtained or derived from:
[0276] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0277] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0278] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0279] (d) a combination of (a), (b), and / or (c).
[0280] In some embodiments, the signal sequence is present upstream of all or some polynucleotide sequences. In some embodiments, the signal sequence is present upstream of the first polynucleotide sequence. In some embodiments, the signal sequence is present upstream of some polynucleotide sequences. In some embodiments, the signal sequence is present upstream of each of the polynucleotide sequences. In some embodiments, the polynucleotide sequence comprises a target sequence, a linker sequence, and a selfassembling sequence, wherein the target sequence is obtained or derived from:
[0281] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0282] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0283] (d) a combination of (a), (b), and / or (c).
[0284] In some embodiments, the polynucleotide sequence comprises a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, wherein the target sequence is obtained or derived from:
[0285] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus.
[0286] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0287] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0288] (d) a combination of (a), (b), and / or (c).
[0289] In some embodiments, the multisubunit nucleic acid sequence comprises one polynucleotide sequence comprising a target sequence, a linker sequence and a selfassembling sequence, and another polynucleotide sequence comprising a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, wherein the target sequence is obtained or derived from:
[0290] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0291] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0292] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0293] (d) a combination of (a), (b), and / or (c). In some embodiments, the multisubunit nucleic acid sequence either comprises a target sequence, a linker sequence and a self-assembling sequence, or a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence or a combination thereof, wherein the target sequence is obtained or derived from:
[0294] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0295] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0296] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0297] (d) a combination of (a), (b), and / or (c). such that the total number of polynucleotide sequences in a multisubunit nucleic acid sequence are not more than 100. In some embodiments, the linker sequence connects the signal sequence with the polynucleotide sequence. In some embodiments, one linker sequence connects the cleavage sequence with the target sequence and another linker sequence connects the target sequence with the self-assembling sequence in a polynucleotide sequence. Some exemplary illustrations of the multisubunit nucleic acid sequences are provided in figures 1 and their representative encoded multisubunit peptides are provided in figures 2 respectively.
[0298] The term “nucleic acid” as used herein means a polymer comprising two or more nucleotides for example, deoxyribonucleotides or ribonucleotides, either in an unmodified or modified form. The nucleic acid may be either single stranded or double stranded, linear, or circular. The term nucleic acid also encompasses fragments, variants, mutants, or codon optimized sequences of deoxyribonucleotides, ribonucleotides, or functional analogs thereof.
[0299] The term “nucleotide” as used herein means a ribonucleotide or deoxyribonucleotide. If the term nucleotide is used in the context of RNA, it refers to ribonucleotide, and if it is used in the context of DNA, it refers to deoxyribonucleotide.
[0300] In some embodiments, the multisubunit nucleic acid sequence is a DNA, an RNA, or an mRNA. The multisubunit nucleic acid sequence may be few nucleotides long to several thousand nucleotides long. Deoxyribonucleic acid (DNA)
[0301] The term “deoxyribonucleic acid” or “DNA” has been used interchangeably herein and means a polymer of deoxyribonucleotides. The DNA may be either single stranded or double stranded, linear, or circular.
[0302] In some embodiments, the multisubunit nucleic acid sequence is a DNA. In some embodiments, the DNA encodes a multisubunit peptide described herein.
[0303] Ribonucleic acid (RNA)
[0304] The term “ribonucleic acid” or “RNA” has been used interchangeably herein and means a polymer of ribonucleotides. The RNA may be either single stranded or double stranded, linear, or circular. The term RNA also includes messenger RNA (mRNA). In some embodiments, the multisubunit nucleic acid sequence is an mRNA.
[0305] In some embodiments, the mRNA encodes a multisubunit peptide as described herein.
[0306] In some embodiments, the mRNA is unmodified or modified or a combination of both. The modification may be in the nucleobase of the nucleotide, or sugar moiety of the nucleotide, or the phosphate of the nucleotide.
[0307] In some embodiments, mRNA is produced using recombinant expression system, or chemically synthesized or obtained through in vitro transcription. In some embodiments, the mRNA is obtained through a single in vitro transcription (IVT) process or step. In vitro transcription (IVT) is a laboratory process used to synthesize RNA molecules (for example, mRNA) from a DNA template enzymatically outside of a cell or in a cell free system. A single IVT process or step is understood to mean one complete cycle of an IVT reaction which produces mRNA molecules, each comprising at least two polynucleotide sequences, as described herein, as against multiple IVT reactions that produces separate mRNA molecules, each comprising a single polynucleotide sequence.
[0308] In some embodiments, the mRNA is circular. In other embodiments, the mRNA is linear.
[0309] In some embodiments, the mRNA is self-amplifying or self-replicating. Selfamplifying or self-replicating mRNA as used herein means an mRNA that self-replicate upon delivery into the cells. Such mRNAs typically contain a replicase sequence, usually derived from an alphavirus, which enables amplification of the original strand of mRNA encoding the protein of interest upon delivery into the cells (Beissert, Tim et al. Molecular Therapy (2020) 28:119-128).
[0310] The present disclosure provides mRNAs which are few hundred nucleotides long to several thousand nucleotides long. State of the art discourages using long mRNAs for vaccines and therapeutics. Longer mRNA molecules are more susceptible to degradation, which can compromise their stability and reduce their effectiveness in experimental and therapeutic context. Besides, longer mRNAs are also harder to transcribe accurately as the RNA polymerase used during the IVT reaction is inherently vulnerable to introduce errors within the transcribed mRNA. Additionally, long mRNA molecules are more prone to form complex secondary and tertiary structures, which can interfere with their intended function, reduce their efficiency, complicates the production process, and may even lead to unintended outcomes. Therefore, longer mRNAs are avoided in the art owing to their inherent complexities and challenges.
[0311] In some embodiments, mRNA is few hundred nucleotides long to several thousand nucleotides long. In some embodiments, mRNA is about 0.5 kb, 1 kb, 1.5 kb, 2 kb, 2.5 kb, 3 kb, 3.5 kb, 4 kb, 4.5 kb, 5 kb, 5.5 kb, 6 kb, 6.5 kb, 7.0 kb, 7.5 kb, 8 kb, 8.5 kb, 9 kb, 9.5 kb, 10 kb, 10.5 kb, 11 kb, 11.5 kb, 12 kb, 12.5 kb, 13 kb, 13.5 kb, 14 kb, 14.5 kb, 15 kb, 16 kb, 17 kb, 18 kb, 19 kb, 20 kb, 21 kb, 22 kb, 23 kb, 24 kb, 25 kb, 26 kb, 27 kb, 28 kb, 29 kb, 30 kb in length, or a fraction thereof. In some embodiments, mRNA is about 0.5 to 30 kb, 0.5 to 25 kb, 0.5 to 20 kb in length, or any range therein. In some embodiments, mRNA is about 1 to 20 kb, 1 to 18 kb, 1 to 16 kb, 1 to 14 kb, 1 to 12 kb, 1 to 10 kb, 1 to 9 kb, 1 to 8 kb, 1 to 7 kb, 1 to 6 kb, 1 to 5 kb in length, or any range therein.
[0312] In some embodiments, mRNA is about 0.5 kb to about 1 kb, about 1 kb to about 2 kb, about 2 kb to about 3 kb, about 3 kb to about 4 kb, about 4 kb to about 5 kb, about 5 kb to about 6 kb, about 6 kb to about 7 kb, about 7 kb to about 8 kb, about 8 kb to about 9 kb, about 9 kb to about 10 kb, about 10 kb to about 11 kb, about 11 kb to about 12 kb, about 12 kb to about 13 kb, about 13 kb to about 14 kb, about 14 kb to about 15 kb, about 15 kb to about 16 kb, about 16 kb to about 17 kb, about 17 kb to about 18 kb, about 18 kb to about 19 kb, about 19 kb to about 20 kb, about 20 kb to about 21 kb, about 21 kb to about 22 kb, about 22 kb to about 23 kb, about 23 kb to about 24 kb, about 24 kb to about 25 kb, about 25 kb to about 26 kb, about 26 kb to about 27 kb, about 27 kb to about 28 kb, about 28 kb to about 29 kb, about 29 kb to about 30 kb in length, or any range therein. Target Sequence and target peptide
[0313] The multisubunit nucleic acid sequence and the multisubunit peptide includes target sequence and target peptide respectively, wherein the target sequence and target peptide is obtained or derived from:
[0314] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0315] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0316] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0317] (d) a combination of (a), (b), and / or (c).
[0318] In some embodiments, target sequence and target peptide contains recurring target sequences and target peptides respectively. In some embodiments, the target sequence is a DNA or an RNA. In another embodiment, the target sequence is an mRNA. In some embodiments, the target sequence is modified or unmodified. The target sequence includes codon optimized sequences, fragments, mutants, variants, comparable equivalents, functional analogs, or combination thereof.
[0319] In some embodiments, the target sequence is a sequence of nucleotides that encodes a target peptide obtained or derived from:
[0320] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,
[0321] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0322] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0323] (d) a combination of (a), (b), and / or (c).
[0324] In some embodiments, the target peptide is identical in two or more polypeptides, for example homologous polypeptides. In some embodiments, the target peptide is different in two or more polypeptides, for example as in heterologous polypeptides. In some embodiments, the target peptide is few amino acids long to several hundred amino acids long.
[0325] Respiratory Syncytial Virus
[0326] Respiratory syncytial virus belongs to the genus pneumovirus within the family paramyxoviridae. RSVs are enveloped, single stranded RNA viruses. RSV genome contains 10 genes encoding 11 proteins. RSV viral particle is composed of a lipid bilayer envelope with three proteins viz., glycoprotein G, glycoprotein F, and small hydrophobic SH protein, a nucleocapsid consisting of three proteins viz., nucleoprotein N, phosphoprotein P, and large polymerase protein L. An M protein or matrix protein surrounds the nucleocapsid and remains in close association with (underlies) the envelope. The M2 gene encodes two overlapping proteins viz., M2-1 protein (an elongation factor), and M2-2 protein (a transcription and replication regulator). There are two non- structural proteins viz., NS1 and NS2. Two antigenic subgroups have been identified in RSV viz., A and B and multiple genotypes within, on the basis of variability in the glycoprotein G. F protein is relatively conserved across the different RSV genotypes / strains.
[0327] In some embodiments, the target peptide is obtained or derived from envelope protein, matrix protein, nucleocapsid protein, M2- 1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus, including mutants, derivatives, variants, comparable equivalents, or functional analogs.
[0328] Envelope proteins
[0329] Respiratory syncytial virus contains three envelop proteins viz., glycoprotein G, glycoprotein F, and small hydrophobic protein SH. In some embodiments, the target peptide is obtained or derived from envelope protein of a respiratory syncytial virus. In some embodiments, the envelope protein is selected from the group comprising glycoprotein G, glycoprotein F, small hydrophobic protein SH, or a combination thereof of a respiratory syncytial virus.
[0330] Glycoprotein G (“G protein”, “attachment protein”, “G”, or “protein G”)
[0331] Glycoprotein G is one of the envelope proteins of RSV. G is a type II transmembrane glycoprotein, with a hydrophobic signal / anchor located near the N- terminus and the major part of the C-terminal oriented towards the outside. It is about 298 amino acid long with a molecular weight of approximately 32 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). Molecular weight of G protein significantly varies based upon the degree of glycosylation. The G protein contains a linear heparin binding domain essential for virus attachment to the host cells. G protein interacts with host CX3CR1, the receptor for the CX3C chemokine fractalkine, to modulate the immune response and facilitate infection.
[0332] In some embodiments, the target peptide is obtained or derived from glycoprotein G of a respiratory syncytial virus.
[0333] In some embodiments, the glycoprotein G is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants, or functional analogs thereof.
[0334] Exemplary glycoprotein G includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 87-2734 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0335] In some embodiments, the glycoprotein G shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0336] Given the amino acid sequences of the glycoprotein G, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above glycoprotein G. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the glycoprotein G is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0337] Glycoprotein F (“F protein”, “fusion protein”, “F”, or “protein F”)
[0338] Glycoprotein F is one of the envelope proteins of RSV. F is a type I transmembrane surface protein that has a cleaved signal peptide at the N-terminus and a membrane anchor near the C-terminus. It is about 574 amino acid long with a molecular weight of approximately 63 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). F is synthesized as an inactive F0 precursor which is cleaved at two sites by a furin-like protease resulting in the mature Fl and F2 fusion glycoproteins. The N-terminus of the Fl subunit contains a hydrophobic domain (the fusion peptide) that inserts directly into the target membrane to initiate fusion. The Fl subunit also contains two areas of heptad repeats that associate during fusion, driving a conformational shift that brings the viral and cellular membranes into proximity. In the later stages of infection, the F protein, present on the plasma membrane of infected cells, facilitates fusion with neighbouring cells, resulting in syncytia formation, a cytopathic effect that leads to tissue necrosis.
[0339] In some embodiments, the target peptide is obtained or derived from glycoprotein F of a respiratory syncytial virus.
[0340] In some embodiments, the glycoprotein F is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof. In some embodiments, the glycoprotein F is a prefusion stabilized F protein. In some embodiments, the F protein includes one or more mutations that stabilizes the protein in the prefusion state or conformation. In some embodiments, the prefusion stabilized F protein is as described in WO2014160463 and WO2017070622.
[0341] Exemplary glycoprotein F includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 2735-4559 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0342] In some embodiments, the glycoprotein F shares at least 50% identity with the sequences described herein, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0343] Given the amino acid sequences of the glycoprotein F, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above glycoprotein F. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the glycoprotein F is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0344] Small Hydrophobic Protein SH (SH Protein, Small Hydrophobic SH Protein, Protein SH, or SH)
[0345] SH Protein is one of the envelope proteins of RS V. SH is a transmembrane protein that has N-terminus anchored to the lipid bilayer membrane, with the C-terminus oriented extracellularly. It is about 64 amino acid long with a molecular weight of approximately 7.5 kDa (although variations in length and mol weight may be possible among different genotypes or strains of RSV). Although unglycosylated form of SH is common, but its glycosylated form varies in molecular weight from 13 kDa to over 30 kDa. SH is a viroporin (a class of small viral proteins that can modify membrane permeability) that forms a homopentameric ion channel displaying low ion selectivity. SH is known to play a role in virus morphogenesis and pathogenicity at various stages of the RSV life cycle. SH also appears to inhibit TNF-alpha, an antiviral cytokine, and may delay apoptosis, allowing time for the virus to replicate.
[0346] In some embodiments, the target peptide is obtained or derived from small hydrophobic protein SH of a respiratory syncytial virus.
[0347] In some embodiments, the small hydrophobic protein SH is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0348] Exemplary small hydrophobic SH protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 4560-4794 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0349] In some embodiments, the small hydrophobic SH protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0350] Given the amino acid sequences of the small hydrophobic SH protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above small hydrophobic SH protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the small hydrophobic SH protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0351] Matrix Protein (M Protein, Matrix Protein M, Protein M, or M)
[0352] Matrix protein underlies the lipid bilayer envelope. It surrounds the nucleocapsid. It is about 256 amino acid long with a molecular weight of approximately 28.7 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). Plays a crucial role in virus assembly into filaments and budding. Early in infection, M protein localizes in the nucleus where it inhibits host cell transcription through direct binding to host chromatin. Later in infection, M protein traffics to the cytoplasm through the action of host nuclear transport protein, CRM1, to associate with inclusion bodies, which is the site of viral transcription and replication. During virus assembly and budding, M protein acts as a bridge between the nucleocapsid and the lipid bilayer.
[0353] In some embodiments, the target peptide is obtained or derived from matrix protein of a respiratory syncytial virus.
[0354] In some embodiments, the target protein is a full length matrix protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0355] Exemplary matrix protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 4795-4978 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0356] In some embodiments, the matrix protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0357] Given the amino acid sequences of the matrix protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above matrix protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the matrix protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0358] Nucleocapsid Proteins
[0359] The nucleocapsid of respiratory syncytial virus contains three proteins viz., nucleoprotein N, phosphoprotein P, and large polymerase protein L. In some embodiments, the target peptide is obtained or derived from nucleocapsid protein of a respiratory syncytial virus. In some embodiments, the nucleocapsid protein is selected from the group comprising nucleoprotein N, phosphoprotein P, large polymerase protein L, or a combination thereof of a respiratory syncytial virus. Nucleoprotein N (N Protein, Protein N, or N)
[0360] Nucleoprotein N encapsidates the viral RNA genome. It is about 391 amino acid long with a molecular weight of approximately 43.4 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). N protein has a tendency to multimerize. Each N monomer consists of N-terminal and C-terminal domains separated by a hinge. The N protein binds along the entire length of genomic RNA protecting RNA from nucleases and the N protein genomic RNA complex together serves as a template for RNA synthesis. N protein interacts with several host factors, such as EIF2AK2 / PKR, EIF1AX, NF-kappa-B, TAX1BP1 etc., facilitating viral replication and growth.
[0361] In some embodiments, the target peptide is obtained or derived from nucleoprotein N of a respiratory syncytial virus.
[0362] In some embodiments, the nucleoprotein N is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0363] Exemplary nucleoprotein N includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 4979-5255 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0364] In some embodiments, the nucleoprotein N shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0365] Given the amino acid sequences of the nucleoprotein N, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above nucleoprotein N. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the nucleoprotein N is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0366] Phosphoprotein P (P protein, Protein P, or P)
[0367] Phosphoprotein P is an essential polymerase co-factor. It is about 241 amino acid long with a molecular weight of approximately 27 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). P protein exists as a homotetramer formed through a multimerization domain present in the P protein. P protein plays critical roles in regulating RNA replication and transcription through its interactions with multiple proteins. P protein binds to free N protein monomers and delivers them to nascent genomes / antigenomes, thus preventing N protein from selfaggregating or binding to nonviral RNA. It tethers the RNA-directed RNA polymerase L to the nucleoprotein-RNA complex. P protein recruits the M2- 1 protein required for efficient transcription of viral RNA.
[0368] In some embodiments, the target peptide is obtained or derived from phosphoprotein P of a respiratory syncytial virus.
[0369] In some embodiments, the phosphoprotein P is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0370] Exemplary phosphoprotein P includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 5256-5668 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0371] In some embodiments, the phosphoprotein P shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0372] Given the amino acid sequences of the phosphoprotein P, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above phosphoprotein P. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the phosphoprotein P is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0373] Large Polymerase Protein L (RNA-Directed RNA Polymerase L, RNA- Dependent RNA Polymerase L, L Protein, Protein L, or L)
[0374] Large polymerase protein L is an RNA-dependent RNA polymerase of RS V. It is about 2166 amino acid long with a molecular weight of approximately 249.7 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). L protein contains an RNA-dependent RNA polymerase (RdRp) domain, a polyribonucleotidyl transferase (PRNTase or capping) domain and a methyltransferase (MTase) domain. L protein is responsible for RNA synthesis, cap addition, and cap methylation. It also performs polyadenylation of subgenomic mRNAs. In some embodiments, the target peptide is obtained or derived from large polymerase protein L of a respiratory syncytial virus.
[0375] In some embodiments, the large polymerase protein L is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants thereof.
[0376] Exemplary large polymerase protein L includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 5669-6132 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0377] In some embodiments, the large polymerase protein L shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0378] Given the amino acid sequences of the large polymerase protein L, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above large polymerase protein L. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the large polymerase protein L is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0379] M2-1 Protein (Envelope associated 22 kDa protein, Transcription antitermination factor M2-1, or M2-1)
[0380] M2-1 protein is an essential transcription process! vity factor. It is about 194 amino acid long with a molecular weight of approximately 22.1 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). It forms a layer between the matrix protein and nucleocapsid. It exists as a homotetramer that binds to RNA and phosphoprotein to prevent premature termination of transcription. It facilitates association of matrix protein with nucleocapsid contributing to viral assembly and budding.
[0381] In some embodiments, the target peptide is obtained or derived from M2-1 protein of a respiratory syncytial virus.
[0382] In some embodiments, the M2-1 protein is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof. Exemplary M2-1 protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 6133-6381 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0383] In some embodiments, the M2-1 protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0384] Given the amino acid sequences of the M2-1 protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above M2-1 protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the M2-1 protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0385] M2-2 Protein (Transcription-replication factor M2-2, or M2-2)
[0386] M2-2 is a small protein which is expressed in low levels with no clear indication of it being packaged in the virus particle. It is about 90 amino acid long with a molecular weight of approximately 10.6 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). It is believed that M2-2 might play a role in shifting RNA synthesis from transcription to RNA replication, that M2-2 can be inhibitory to RNA synthesis, and that the inhibitory activity occurs with increased M2-2 expression.
[0387] In some embodiments, the target peptide is obtained or derived from M2-2 protein of a respiratory syncytial virus.
[0388] In some embodiments, the M2-2 protein is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants, or functional analogs thereof.
[0389] Exemplary M2-2 protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 6382-6756 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0390] In some embodiments, the M2-2 protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof. Given the amino acid sequences of the M2-2 protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above M2-2 protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the M2-2 protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0391] Non-structural proteins
[0392] Respiratory syncytial virus genome encodes two non-structural proteins viz., NS1 Protein and NS2 Protein. In some embodiments, the target peptide is obtained or derived from non-structural proteins of a respiratory syncytial virus. In some embodiments, the non-structural protein is selected from the group comprising NS1, NS2, or a combination thereof of a respiratory syncytial virus.
[0393] NS1 Protein (Non-structural protein 1 or NS1)
[0394] NS1 is the first gene to be transcribed in RSV. It is about 139 amino acid long with a molecular weight of approximately 15.5 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). NS1 predominantly exists in monomeric form, although multimeric forms have also been observed. It can also forms heteromers with NS2. NS1 interferes with induction and signalling of type I IFN and type III IFN in human epithelial cells, macrophages, and dendritic cells leading to suppression of host innate defense.
[0395] In some embodiments, the target peptide is obtained or derived from NS1 protein of a respiratory syncytial virus.
[0396] In some embodiments, the NS1 protein is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants thereof.
[0397] NS2 Protein (Non-structural protein 2 or NS2)
[0398] NS2 is the second gene to be transcribed after NS1 in RSV. It is slightly shorter than NS1, about 124 amino acid long with a molecular weight of approximately 14.7 kDa (although variations in length and molecular weight may be possible among different genotypes or strains of RSV). NS 2 typically exists as a homomultimer as monomer is unstable. It also exists as hetermultimer with NS1 as NS1-NS2 complex. Most of the NS2 resides in the mitochondria as NS1-NS2 heteromer complex. NS2 interacts with RIGI preventing host signalling pathway involved in interferon production. NS2 suppresses premature apoptosis by an NF-kappa-B-dependent, interferon-independent mechanism promoting continued viral replication.
[0399] In some embodiments, the target peptide is obtained or derived from NS2 protein of a respiratory syncytial virus.
[0400] In some embodiments, the NS2 protein is a full length protein or a fragment thereof of a respiratory syncytial virus or comparable equivalents, including mutants, derivatives, variants thereof.
[0401] B cell epitope
[0402] In some embodiments, the target peptide is obtained or derived from B cell epitope of a respiratory syncytial virus, including a fragment, mutant, derivative or variant thereof.
[0403] Exemplary B cell epitopes include, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 6757-7019 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0404] In some embodiments, the B cell epitopes share at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0405] Given the amino acid sequences of the B cell epitopes, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above B cell epitopes. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the B cell epitopes are encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0406] T cell epitope
[0407] In some embodiments, the target peptide is obtained or derived from T cell epitope of a respiratory syncytial virus, including a fragment, mutant, derivative or variant thereof.
[0408] Exemplary T cell epitopes include, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 7020-7371 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof. In some embodiments, the T cell epitopes share at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0409] Given the amino acid sequences of the T cell epitopes, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above T cell epitopes. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the T cell epitopes are encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0410] Human Metapneumovirus (HMPY)
[0411] Human metapneumovirus (HMPV) belongs to the genus Metapneumovirus within the family Paramyxoviridae. HMPVs are enveloped, single stranded RNA viruses. HMPV genome contains 8 genes encoding 9 proteins. HMPV viral particle is composed of a lipid bilayer envelope with three proteins viz., glycoprotein G, glycoprotein F, and small hydrophobic protein SH, a nucleocapsid consisting of three proteins viz., nucleoprotein N, phosphoprotein P, and large polymerase protein L. An M protein or matrix protein surrounds the nucleocapsid and remains in close association with (underlines) the envelope. The M2 gene encodes two overlapping proteins viz., M2-1 protein and M2-2 protein (a transcription and replication regulator). Metapneumovirus lacks the non- structural proteins. Two antigenic lineages have been identified based upon the variability in the glycoprotein G, and are designated as A and B, with each divided into at least two sub-lineages designated Al, Al, and Bl, B2 respectively.
[0412] In some embodiments, the target peptide is obtained or derived from envelope protein, matrix protein, nucleocapsid protein, M2- 1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus, including mutants, derivatives, variants, comparable equivalents, or functional analogs.
[0413] Envelope Proteins
[0414] HMPV contains three envelope proteins - glycoprotein G, glycoprotein F, and small hydrophobic protein SH. In some embodiments, the target peptide is obtained or derived from envelope protein of a metapneumovirus. In some embodiments, the envelope protein is selected from the group comprising glycoprotein G, glycoprotein F, small hydrophobic protein SH, or a combination thereof of a metapneumo virus.
[0415] Glycoprotein G (G Protein, Attachment Protein, Protein G or G)
[0416] Glycoprotein G is one of the envelope proteins of HMPV. G is a membrane glycoprotein, with a hydrophobic signal / anchor located near the N-terminus and the major part of the C-terminal oriented towards the outside. It is about 219 amino acid long with a molecular weight of approximately 23.6 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). Molecular weight of G protein significantly varies based upon the degree of glycosylation, for example, in some cases around 97 kDa due to the presence of N-linked and O-linked sugars. G protein facilitates attachment of virion to host cell membrane by interacting with glycosaminoglycans, initiating the infection. In addition to its role in attachment, glycoprotein G interacts with host RIGI and inhibits RIGI-mediated signalling pathway in order to prevent the establishment of the antiviral state.
[0417] In some embodiments, the target peptide is obtained or derived from glycoprotein G of a metapneumovirus.
[0418] In some embodiments, the glycoprotein G is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants, or functional analogs thereof.
[0419] Exemplary glycoprotein G includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 7372-7681 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0420] In some embodiments, the glycoprotein G shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0421] Given the amino acid sequences of the glycoprotein G, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above glycoprotein G. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the glycoprotein G is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA. Glycoprotein F (“F protein”, “fusion protein”, “F”, or “protein F”)
[0422] Glycoprotein F is one of the envelope proteins of HMPV. F is a membrane glycoprotein, that has a cleaved signal peptide at the N-terminus and a membrane anchor near the C-terminus. It is about 539 amino acid long with a molecular weight of approximately 58.4 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). F is synthesized as an inactive FO precursor which is cleaved into mature Fl and F2 fusion glycoproteins. The N-terminus of the Fl subunit contains a hydrophobic domain (the fusion peptide) that inserts directly into the target membrane to initiate fusion. The Fl subunit also contains two areas of heptad repeats that associate during fusion, driving a conformational shift that brings the viral and cellular membranes into proximity. Fl and F2 forms heterodimer with disulfide linkage and interacts with host heparan sulfate.
[0423] In some embodiments, the target peptide is obtained or derived from glycoprotein F of a metapneumovirus.
[0424] In some embodiments, the glycoprotein F is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0425] Exemplary glycoprotein F includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 7682-7793 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0426] In some embodiments, the glycoprotein F shares at least 50% identity with the sequences described herein, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0427] Given the amino acid sequences of the glycoprotein F, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above glycoprotein F. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the glycoprotein F is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA. Small Hydrophobic SH Protein (SH Protein, Small Hydrophobic Protein SH, Protein SH, or SH)
[0428] SH protein is one of the envelope proteins of HMPV. It is about 179 amino acid long with a molecular weight of approximately 20.6 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). SH is a transmembrane protein that has N-terminus anchored to the lipid bilayer membrane, with the C-terminus oriented extracellularly. It exists in both glycosylated and non-glycosylated forms. SH is a viroporin (a class of small viral proteins that can modify membrane permeability) that forms ion channel displaying low ion selectivity.
[0429] In some embodiments, the target peptide is obtained or derived from small hydrophobic protein SH of a metapneumovirus.
[0430] In some embodiments, the small hydrophobic protein SH is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0431] Exemplary small hydrophobic SH protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 7794-7935 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0432] In some embodiments, the small hydrophobic SH protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0433] Given the amino acid sequences of the small hydrophobic SH protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above small hydrophobic SH protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the small hydrophobic SH protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0434] Matrix Protein (M Protein, Matrix Protein M, Protein M, or M)
[0435] Matrix protein underlies the lipid bilayer envelope. It surrounds the nucleocapsid. It is about 254 amino acid long with a molecular weight of approximately 27.6 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). Plays a crucial role in virus assembly into filaments and budding. Early in infection, M protein localizes in the nucleus where it inhibits host cell transcription through direct binding to host chromatin. Later in infection, M protein traffics to the cytoplasm through the action of host nuclear transport protein, CRM1, to associate with inclusion bodies, which is the site of viral transcription and replication. During virus assembly and budding, M protein acts as a bridge between the nucleocapsid and the lipid bilayer.
[0436] In some embodiments, the target peptide is obtained or derived from matrix protein of a metapneumovirus.
[0437] In some embodiments, the matrix protein is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0438] Exemplary matrix protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 7936-7976 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0439] In some embodiments, the matrix protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0440] Given the amino acid sequences of the matrix protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above matrix protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the matrix protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0441] Nucleocapsid Proteins
[0442] The nucleocapsid of metapneumovirus contains three proteins viz., nucleoprotein N, phosphoprotein P, and large polymerase protein L. In some embodiments, the target peptide is obtained or derived from nucleocapsid protein of a metapneumovirus. In some embodiments, the nucleocapsid protein is selected from the group comprising nucleoprotein N, phosphoprotein P, large polymerase protein L, or a combination thereof of a metapneumovirus. Nucleoprotein N (N Protein, Protein N, or N)
[0443] Nucleoprotein N encapsidates the viral RNA genome. It is about 394 amino acid long with a molecular weight of approximately 43.5 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). N protein has a tendency to multimerize. The N protein binds to RNA and protect it from the action of nucleases. N protein interacts with several host factors facilitating viral replication and growth.
[0444] In some embodiments, the target peptide is obtained or derived from nucleoprotein N of a metapneumovirus.
[0445] In some embodiments, the nucleoprotein N is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0446] Exemplary nucleoprotein N includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 7977-8042 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0447] In some embodiments, the nucleoprotein N shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0448] Given the amino acid sequences of the nucleoprotein N, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above nucleoprotein N. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the nucleoprotein N is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0449] Phosphoprotein P (P protein, Protein P, or P)
[0450] Phosphoprotein P is an essential polymerase co-factor. It is about 294 amino acid long with a molecular weight of approximately 32.7 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). P protein exists as a homotetramer formed through a multimerization domain present in the P protein. P protein plays critical roles in regulating RNA replication and transcription through its interactions with multiple proteins. It tethers the RNA-directed RNA polymerase L to the nucleoprotein-RNA complex. P protein recruits the M2-1 protein required for efficient transcription of viral RNA.
[0451] In some embodiments, the target peptide is obtained or derived from phosphoprotein P of a metapneumovirus.
[0452] In some embodiments, the phosphoprotein P is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0453] Exemplary phosphoprotein P includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 8043-8211 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0454] In some embodiments, the phosphoprotein P shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0455] Given the amino acid sequences of the phosphoprotein P, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above phosphoprotein P. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the phosphoprotein P is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0456] Large Polymerase Protein L (RNA-Directed RNA Polymerase L, RNA- Dependent RNA Polymerase L, L Protein, Protein L, or L)
[0457] Large polymerase protein L is an RNA-dependent RNA polymerase of HMPV. It is about 2005 amino acid long with a molecular weight of approximately 230.7 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). L protein contains an RNA-dependent RNA polymerase (RdRp) domain, a polyribonucleotidyl transferase (PRNTase or capping) domain and a methyltransferase (MTase) domain. L protein is responsible for RNA synthesis, cap addition, and cap methylation. It also performs polyadenylation of sub genomic mRNAs.
[0458] In some embodiments, the target peptide is obtained or derived from large polymerase protein L of a metapneumovirus. In some embodiments, the large polymerase protein L is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0459] Exemplary large polymerase protein L includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 8212-8483 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0460] In some embodiments, the large polymerase protein L shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0461] Given the amino acid sequences of the large polymerase protein L, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above large polymerase protein L. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the large polymerase protein L is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0462] M2-1 Protein (Envelope associated 22 kDa protein, Transcription antitermination factor M2-1, or M2-1)
[0463] M2-1 protein is an essential transcription process! vity factor. It is about 187 amino acid long with a molecular weight of approximately 21.2 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). It forms a layer between the matrix protein and nucleocapsid. It exists as a homotetramer that binds to RNA and phosphoprotein to prevent premature termination of transcription. It facilitates association of matrix protein with nucleocapsid contributing to viral assembly and budding.
[0464] In some embodiments, the target peptide is obtained or derived from M2-1 protein of a metapneumovirus.
[0465] In some embodiments, the M2-1 protein is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0466] Exemplary M2-1 protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 8484-8545 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0467] In some embodiments, the M2-1 protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0468] Given the amino acid sequences of the M2-1 protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above M2-1 protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the M2-1 protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0469] M2-2 Protein (Transcription-replication factor M2-2, or M2-2)
[0470] M2-2 is a small protein which is expressed in low levels by HMPV. It is about 71 amino acid long with a molecular weight of approximately 8.1 kDa (although variations in length and molecular weight may be possible among different lineages or strains of MPV). It is believed that M2-2 might play a role in shifting RNA synthesis from transcription to RNA replication, that M2-2 can be inhibitory to RNA synthesis, and that the inhibitory activity occurs with increased M2-2 expression.
[0471] In some embodiments, the target peptide is obtained or derived from M2-2 protein of a metapneumovirus.
[0472] In some embodiments, the M2-2 protein is a full length protein or a fragment thereof of a metapneumovirus or comparable equivalents, including mutants, derivatives, variants, or functional analogs thereof.
[0473] Exemplary M2-2 protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 8546-8586 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0474] In some embodiments, the M2-2 protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0475] Given the amino acid sequences of the M2-2 protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above M2-2 protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the M2-2 protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0476] B cell epitope
[0477] In some embodiments, the target peptide is obtained or derived from B cell epitope of a metapneumovirus, including a fragment, mutant, derivative or variant thereof.
[0478] T cell epitope
[0479] In some embodiments, the target peptide is obtained or derived from T cell epitope of a metapneumovirus, including a fragment, mutant, derivative or variant thereof.
[0480] Exemplary T cell epitopes include, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 8587-8627 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0481] In some embodiments, the T cell epitopes share at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0482] Given the amino acid sequences of the T cell epitopes, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above T cell epitopes. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the T cell epitopes are encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0483] Human Parainfluenza Virus (HPIV or hPIV)
[0484] Human parainfluenza viruses belong to two different genera within the family Paramyxoviridae viz., Respirovirus genus comprising serotypes HPIV-1 and HPIV-3, and Rubulavirus genus comprising serotypes HPIV-2 and HPIV-4, HPIV-4 further comprising two subtypes viz., HPIV-4a and HPIV-4b. HPIVs are enveloped, single stranded RNA viruses measuring about 150-300 in diameter. HPIV genome is about 15-16 kb long and encodes 6 common proteins viz., glycoprotein F and hemagglutinin neuraminidase HN remain associated with the lipid bilayer envelope, nucleoprotein N, phosphoprotein P, and large polymerase protein L forms the nucleocapsid, and the matrix protein or M protein underline the lipid bilayer. In addition, the P gene produces some small non-structural proteins (accessory proteins) from multiple overlapping reading frames which differs in different HPIVs. For example, the P gene encodes C protein in HPIV-1 and HPIV-3, D protein in HPIV-3, V and I proteins in HPIV-2 and HPIV4.
[0485] In some embodiments, the target peptide is obtained or derived from envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, including mutants, derivatives, variants, comparable equivalents, or functional analogs.
[0486] Envelope Proteins
[0487] Human parainfluenza viruses contain two envelope proteins - glycoprotein F and hemagglutinin neuraminidase HN. In some embodiments, the target peptide is obtained or derived from envelope protein of a human parainfluenza virus. In some embodiments, the envelope protein is selected from the group comprising glycoprotein F, hemagglutinin neuraminidase HN, or a combination thereof of a human parainfluenza virus.
[0488] Glycoprotein F (“F protein”, “fusion protein”, “F”, or “protein F”)
[0489] Glycoprotein F is one of the envelope proteins of HPIV. F is a type I transmembrane surface protein that has a cleaved signal peptide at the N-terminus and a membrane anchor near the C-terminus. It is about 555 amino acid long with a molecular weight of approximately 60.7 kDa (although variations in length and molecular weight may be possible among different serotypes, subtypes, or strains of HPIV, for example, in HPIV-2 it is about 551 amino acid long with a molecular weight of approximately 59.6 kDa). F is synthesized as an inactive F0 precursor which is cleaved into mature Fl and F2 fusion glycoproteins. Fl and F2 forms heterodimer with disulfide linkage. F protein directs fusion of viral and cellular membranes leading to delivery of the viral genome into the cytoplasm.
[0490] In some embodiments, the target peptide is obtained or derived from glycoprotein F of a human parainfluenza virus.
[0491] In some embodiments, the glycoprotein F is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0492] Exemplary glycoprotein F includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 8628-9088 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0493] In some embodiments, the glycoprotein F shares at least 50% identity with the sequences described herein, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0494] Given the amino acid sequences of the glycoprotein F, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above glycoprotein F. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the glycoprotein F is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0495] Hemagglutinin Neuraminidase (HN Protein, HN Glycoprotein, or HN)
[0496] Hemagglutinin is one of the envelope proteins of HPIV. HN exists as tetramer. It is about 575 amino acid long with a molecular weight of approximately 64 kDa (although variations in length and molecular weight may be possible among different serotypes, subtypes, or strains of HPIV, for example, in HPIV-2 it is about 571 amino acid long with a molecular weight of approximately 63.5 kDa). HN protein binds to host cell sialic acid, which allows the virus to attach to the cell membrane. This activity is responsible for the agglutination of erythrocytes (hemagglutination activity). Late into the infection, HN protein cleaves sialic acid to facilitate the release of progeny virions (neuraminidase activity). Neuraminidase activity contributes to the spread of the virus.
[0497] In some embodiments, the target peptide is obtained or derived from hemagglutinin neuraminidase of a human parainfluenza virus.
[0498] In some embodiments, the hemagglutinin neuraminidase is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0499] Exemplary hemagglutinin neuraminidase includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 9089-10015 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0500] In some embodiments, the hemagglutinin neuraminidase shares at least 50% identity with the sequences described herein, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0501] Given the amino acid sequences of the hemagglutinin neuraminidase, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above hemagglutinin neuraminidase. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the hemagglutinin neuraminidase is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0502] Matrix Protein (M Protein, Matrix Protein M, Protein M, or M)
[0503] Matrix protein underlies the lipid bilayer envelope. It surrounds the nucleocapsid. It is about 348 amino acid long with a molecular weight of approximately 38.4 kDa (although variations in length and molecular weight may be possible among different serotypes, subtypes, or strains of HPIV, for example, in HPIV-2 it is about 377 amino acid long). Matrix protein plays a crucial role in virus assembly, budding, and release.
[0504] In some embodiments, the target peptide is obtained or derived from matrix protein of a human parainfluenza virus.
[0505] In some embodiments, the matrix protein is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0506] Exemplary matrix protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 10016-10214 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0507] In some embodiments, the matrix protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0508] Given the amino acid sequences of the matrix protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above matrix protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the matrix protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA. Nucleocapsid Proteins
[0509] The nucleocapsid of parainfluenza viruses contains three proteins viz., nucleoprotein N, phosphoprotein P, and large polymerase protein L. In some embodiments, the target peptide is obtained or derived from nucleocapsid protein of a human parainfluenza virus. In some embodiments, the nucleocapsid protein is selected from the group comprising nucleoprotein N, phosphoprotein P, large polymerase protein L, or a combination thereof of a human parainfluenza virus.
[0510] Nucleoprotein N (N Protein, Protein N, or N)
[0511] Nucleoprotein N is one of the nucleocapsid proteins of human parainfluenza virus. It is about 524 amino acid long with a molecular weight of approximately 57.7 kDa (although variations in length and molecular weight may be possible among different serotypes, subtypes, or strains of HPIV, for example, in HPIV-2 it is about 542 amino acid long). N protein carries domains for self-assembly and RNA-binding. It binds to viral genome in a ratio of 1:6 (1 N per 6 ribonucleotides), protecting the genome from nucleases. The encapsidated RNA then serves as a template for transcription and replication. It also forms complex with phosphoprotein P.
[0512] In some embodiments, the target peptide is obtained or derived from nucleoprotein N of a human parainfluenza virus.
[0513] In some embodiments, the nucleoprotein N is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0514] Exemplary nucleoprotein N includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 10215-10468 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0515] In some embodiments, the nucleoprotein N shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0516] Given the amino acid sequences of the nucleoprotein N, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above nucleoprotein N. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the nucleoprotein N is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0517] Phosphoprotein P (P protein, Protein P, or P)
[0518] Phosphoprotein P is an essential polymerase co-factor. It is about 568 amino acid long with a molecular weight of approximately 64.6 kDa (although variations in length and molecular weight may be possible among different serotypes, subtypes, or strains of HPIV, for example, in HPIV-2 it is about 395 amino acid long). It exists as a homotetramer formed through a multimerization domain present in the P protein. N-terminal region of P protein binds to free N protein monomers. C-terminal region contains the polymerase cofactor domain necessary for transcription. P protein tethers the RNA-directed RNA polymerase L to the nucleoprotein-RNA complex.
[0519] In some embodiments, the target peptide is obtained or derived from phosphoprotein P of a human parainfluenza virus.
[0520] In some embodiments, the phosphoprotein P is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0521] Exemplary phosphoprotein P includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 10469-11128 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0522] In some embodiments, the phosphoprotein P shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0523] Given the amino acid sequences of the phosphoprotein P, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above phosphoprotein P. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the phosphoprotein P is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA. Large Polymerase Protein L (RNA-Directed RNA Polymerase L, RNA- Dependent RNA Polymerase L, L Protein, Protein L, or L)
[0524] Large polymerase protein L is an RNA-dependent RNA polymerase of HPIV. It is about 2223 amino acid long with a molecular weight of approximately 253.7 kDa (although variations in length and molecular weight may be possible among different serotypes, subtypes, or strains of HPIV, for example, in HPIV-2 it is about 2262 amino acid long). L protein contains an RNA-dependent RNA polymerase (RdRp) domain, a polyribonucleotidyl transferase (PRNTase or capping) domain and a methyltransferase (MTase) domain. L protein is responsible for RNA synthesis, cap addition, and cap methylation.
[0525] In some embodiments, the target peptide is obtained or derived from large polymerase protein L of a human parainfluenza virus.
[0526] In some embodiments, the large polymerase protein L is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0527] Exemplary large polymerase protein L includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 11129-11485 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0528] In some embodiments, the large polymerase protein L shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0529] Given the amino acid sequences of the large polymerase protein L, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above large polymerase protein L. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the large polymerase protein L is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0530] Accessory Proteins
[0531] The P gene of parainfluenza viruses encodes one or more accessory protein depending on the virus. The accessory proteins include C protein, D protein, I protein, V protein, and W protein. These proteins are termed accessory proteins as they are not essential for virus replication. In some of the embodiments, the target peptide is obtained or derived from accessory protein of a human parainfluenza virus. In some embodiments, the accessory protein is selected from the group comprising C protein, D protein, I protein, V protein, W protein, or a combination thereof of a human parainfluenza virus.
[0532] C Protein (Protein C, or C)
[0533] C protein is one of the accessory proteins present in some of the HPIVs. The virions of HPIV-1 and HPIV-3 (both belonging to the genus Respirovirus) contains a small protein called C protein. C protein has been reported to inhibit acti vation of the transcription factors IRF02 and NF-kB that leads to induction of IFN-beta. C protein also appears to downregulate production of viral RNA at the level of transcription and RNA replication .
[0534] In some embodiments, the target peptide is obtained or derived from C protein of a human parainfluenza virus.
[0535] In some embodiments, the C protein is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0536] Exemplary C protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 11486-11714 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0537] In some embodiments, the C protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0538] Given the amino acid sequences of the C protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above C protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the C protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA. D protein (Protein D, or D)
[0539] D protein is one of the accessory proteins present in some of the HPIVs. HPIV-3 possess D protein. It is about 373 amino acid long. It has been shown to accumulate in the nucleus of HPIV-3 infected cells.
[0540] In some embodiments, the target peptide is obtained or derived from D protein of a human parainfluenza virus.
[0541] In some embodiments, the D protein is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0542] Exemplary D protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 11715-12119 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0543] In some embodiments, the D protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0544] Given the amino acid sequences of the D protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above D protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the D protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0545] I Protein (Protein I, or I)
[0546] I protein is one of the accessory proteins present in some of the HPIVs. HPIV-2 and HPIV-4 possess I protein. It is about 165 and 155 amino acid long protein in HIPV-2 and HIPV-4 respectively.
[0547] In some embodiments, the target peptide is obtained or derived from I protein of a human parainfluenza virus.
[0548] In some embodiments, the I protein is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof. V Protein (Protein V, or V)
[0549] V protein is one of the accessory proteins present in some of the HPIVs. HPIV-2 and HPTV-4 possess V protein. It is about 225 and 229 amino acid long protein in HPIV-2 and HP1V-4 respectively. V protein contains cystine rich domain. V protein has been shown to bind to MDA-5 and inhibit induction of IFN-beta. V protein appears to delay apoptosis during viral infection and downregulates viral transcription and RNA replication. V protein is important for members of Rubulavirus (HPIV-2 and HPIV-4) given their lack of C protein which performs some of these functions.
[0550] In some embodiments, the target peptide is obtained or derived from V protein of a human parainfluenza virus.
[0551] In some embodiments, the V protein is a full length protein or a fragment thereof of a human parainfluenza virus or comparable equivalents, including mutants, derivatives, variants or functional analogs thereof.
[0552] Exemplary V protein includes, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 12120-12188 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0553] In some embodiments, the V protein shares at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0554] Given the amino acid sequences of the V protein, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above V protein. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the V protein is encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0555] B cell epitope
[0556] In some embodiments, the target peptide is obtained or derived from B cell epitope of a human parainfluenza virus, including a fragment, mutant, derivative or variant thereof.
[0557] Exemplary B cell epitopes include, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 12189-12193 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof. In some embodiments, the B cell epitopes share at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0558] Given the amino acid sequences of the B cell epitopes, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above B cell epitopes. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the B cell epitopes are encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0559] T cell epitope
[0560] In some embodiments, the target peptide is obtained or derived from T cell epitope of a human parainfluenza virus, including a fragment, mutant, derivative or variant thereof.
[0561] Exemplary T cell epitopes include, but not limited to, the one represented by the amino acid sequences having SEQ ID NOs: 12194-12216 or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0562] In some embodiments, the T cell epitopes share at least 50% identity with the sequences described herein or comparable equivalents, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0563] Given the amino acid sequences of the T cell epitopes, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above T cell epitopes. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the T cell epitopes are encoded by a target sequence which may be either a DNA, an RNA, or an mRNA.
[0564] Self-assembling sequence and self-assembling peptide
[0565] The multisubunit nucleic acid sequence and multisubunit peptide includes selfassembling sequence and self-assembling peptide respectively. The self-assembling sequence comprises of a sequence of nucleotides, either deoxyribonucleotides or ribonucleotides, that encodes a self-assembling peptide. The self-assembling sequence includes codon optimized sequences, fragments, mutants, variants, comparable equivalents, functional analogs, or a combination thereof. In some embodiments, the selfassembling sequence is a DNA, an RNA or an mRNA. Any self-assembling peptide that is capable of self-assembling into a polypeptide nanoparticle can be employed in accordance with the present disclosure.
[0566] In some embodiments self-assembling peptide is a full-length protein or its fragment, mutants, or variant thereof.
[0567] In some embodiments, the self-assembling peptide includes, but not limited to, lumazine synthase, MS2 coat protein, hepatitis B surface antigen (HBsAg) from Hepatitis B Virus, hepatitis B core antigen (HbcAg) from Hepatitis B virus, human papillomavirus LI (HPV LI) protein, matrix protein Ml from influenza A virus, ferritin, riboflavin synthase, dihydrolipoyl acetyltransferase (E2p), or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalent, or functional analogs thereof.
[0568] In some embodiments, the self-assembling peptide is a ferritin peptide.
[0569] Ferritin is one of the ubiquitous proteins found in nature. It is produced by all living organisms including archaea, bacteria, algae, higher plants, and animals. Each ferritin protein is generally composed of 12 or 24 subunits or peptides which selfassembles into a ferritin nanoparticle.
[0570] In some aspects, the multisubunit nucleic acid sequence and multisubunit peptide includes ferritin sequence and ferritin peptide respectively. The ferritin sequence comprises of a sequence of nucleotides, either deoxyribonucleotides or ribonucleotides, that encodes a ferritin peptide. In some embodiments, the ferritin sequence is a DNA, an RNA or an mRNA.
[0571] Any ferritin peptide that is capable of self-assembling into a nanoparticle can be employed in accordance with the present disclosure. In some embodiments, the ferritin peptide is obtained or derived from Helicobacter pylori ferritin, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents or functional analogs thereof. In some embodiments, the ferritin peptide is obtained or derived from Listeria innocua ferritin, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents or functional analogs thereof.
[0572] In some embodiments, the self-assembling peptide is lumazine synthase. In some embodiments, the lumazine synthase is obtained or derived from Aquifex species (for example, Aquifex aeolicus) or Bacillus species (for example, Bacillus subtil is), including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0573] In some embodiments, the self-assembling peptide is dihydrolipoyl acetyltransferase (E2p). In some embodiments, the dihydrolipoyl acetyltransferase (E2p) is obtained or derived from Bacillus stearothermophilus, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0574] In some embodiments, the self-assembling peptide is MS2 coat protein. In some embodiments, the self-assembling peptide is MS2 coat protein obtained or derived from Emesvirus zinderi, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
[0575] Exemplary self-assembling peptides includes, but not limited to, the ones represented by the following amino acid sequences or comparable equivalents, or a combination thereof, including codon optimized sequences, fragments, mutants, variants, or functional analogs, thereof: LSKDIIKLLNEQVNKEMNSSNLYMSMSSWCYTHSLDGAGLFLFDHAAEEYEHAK KLIIFLNENNVPVQLTSISAPEHKFEGLTQIFQKAYEHEQHISESINNIVDHAIKSKDH ATFNFLQWYVAEQHEEEVLFKDILDKIELIGNENHGLYLADQYVKGI (SEQ ID NO: 1);
[0576] LSKDIIKLLNEQVNKEMNSSNLYMSMSSWCYTHSLDGAGLFLFDHAAEEYEHAK KLIIFLNENNVPVQLTSISAPEHKFEGLTQIFQKAYEHEQHISESINNIVDHAIKSKDH ATFNFLQWYVAEQHEEEVLFKDILDKIELIGNENHGLYLADQYVKGIRRKR (SEQ ID NO: 2);
[0577] LSKDIIKLLNEQVNKEMNSSNLYMSMSSWCYTHSLDGAGLFLFDHAAEEYEHAK KLIIFLNENNVPVQLTSISAPEHKFEGLTQIFQKAYEHEQHISESINNIVDHAIKSKDH ATFNFLQWYVAEQHEEEVLFKDILDKIELIGNENHGLYLADQYVKGIAKSRKS (SEQ ID NO: 3);
[0578] MQIYEGKLTAEGLRFGIVASRFNHALVDRLVEGAIDCIVRHGGREEDITLVRVPGS WEIPVAAGELARKEDIDAVIAIGVLIRGATPHFDYIASEVSKGLANLSLELRKPITFG VITADTLEQAIERAGTKHGNKGWEAALSAIEMANLFKSLR (SEQ ID NO: 4);
[0579] MTKKVGIVDTTFARVDMASIAIKKLKELSPNIKIIRKTVPGIKDLPVACKKLLEEEG CDIVMALGMPGKAEKDKVCAHEASLGLMLAQLMTNKHIIEVFVHEDEAKDDKEL DWLAKRRAEEHAENVYYLLFKPEYLTRMAGKGLRQGFEDAGPARE (SEQ ID NO: 5);
[0580] MTEKEKMLAEKWYDANFDQYLINERARAKDICFELNHTRPSATNKRKELIDQLFQ TTTDNVSISIPFDTDYGWNVKLGKNVYVNTNCYFMDGGQITIGDNVFIGPNCGFY TATHPLNFHHRNEGFEKAGPIHIGSNTWFGGHVAVLPGVTIGEGSVIGAGSVVTKD IPPHSLAVGNPCKVVRKIDNDLPSETLNDETIK (SEQ ID NO: 6);
[0581] MENTTSGFLGPLLVLQAGFFLLTRILTIPQSLDSWWTSLNFQGGAPTCPGQNSQSPT SNHSPTSCPPICPGYRWMCLRRFIIFLFILLLCLIFLLVLLDYQGMLPVCPLLPGTSTT GTGPCRTCTIPAQGTSMFPSCCCTKPSDGNCTCIPIPSSWAFARFLWEWASVRFSW LSLLVPFVQWFAGLSPTVWLSVIWMMWYRGPSLYNTLSPFLPLLPISFCLWVYI (SEQ ID NO: 7);
[0582] MIFVLGGCRHKLVCSPAPCNFFHLCLIISCSCPTVHASKLCLGWLWGMHIDPYKEF GASVELLSFLPSDFFPSIRDLLDTASALYREALESPEHCSPHHTALRQAILCWGELM NLATWVGSNLEDPASRELVVSYVNVNMGLKIRQLLWFHISCLTFGRETVLEYLVS FGVWIRTPPAYRPPNAPILSTLPETTVVRRRGRSPRRRTPSPRRRRSQSPRRRRSQSR ESQC (SEQ ID NO: 8);
[0583] MSLWLPSEATVYLPPVPVSKVVSTDEYVARTNIYYHAGTSRLLAVGHPYFPIKKP NNNKILVPKVSGLQYRVFRIHLPDPNKFGFPDTSFYNPDTQRLVWACVGVEVGRG QPLGVGISGHPLLNKLDDTENASAYAANAGVDNRECISMDYKQTQLCLIGCKPPI GEHWGKGSPCTNVAVNPGDCPPLELINTVIQDGDMVDTGFGAMDFTTLQANKSE VPLDICTSICKYPDYIKMVSEPYGDSLFFYLRREQMFVRHLFNRAGAVGENVPDDL YIKGSGSTANLASSNYFPTPSGSMVTSDAQIFNKPYWLQRAQGHNNGICWGNQLF VTVVDTTRSTNMSLCAAISTSETTYKNTNFKEYLRHGEEYDLQFIFQLCKITLTAD VMTYIHSMNSTILEDWNFGLQPPPGGTLEDTYRFVTSQAIACQKHTPPAPKEDPLK KYTFWEVNLKEKFSADLDQFPLGRKFLLQAGLKAKPKFTLGKRKATPTTSSTSTT AKRKKRKL (SEQ ID NO: 9);
[0584] MSLLTEVETYVLSIVPSGPLKAEIAQRLEDVFAGKNTDLEALMEWLKTRPILSPLT KGILGFVFTLTVPSERGLQRRRFVQNALNGNGDPNNMDRAVKLYRKLKREITFHG AKEIALSYSAGALASCMGLIYNRMGAVTTEVAFGLVCATCEQIADSQHRSHRQM VTTTNPLIRHENRMVLASTTAKAMEQMAGSSEQAAEAMEVASQARQMVQAMRA IGTHPRSSAGLKDDLLENLQAYQKRMGVQMQRFK (SEQ ID NO: 10):
[0585] AAAKPATTEGEFPETREKMSGIRRAIAKAMVHSKHTAPHVTLMDEADVTKLVAH RKKFKAIAAEKGIKL;rFI.JPYVVKALVSALREYPVLN’IAlDDET'EEllQKHYYNIGIA ADTDRGLLVPVIKHADRKPIFALAQEINELAEKARDGKLTPGEMKGASCTITNIGS AGGQWFTPVINHPEVAILGIGRIAEKPIVRDGEIV AAPML ALSI ..SFDHRMIDGAT AQ KAJ.,NHIKRI..I,SDPELJ.d..M (SEQ ID NO: 1 1 );
[0586] ASNFTQFVLVDNGGTGDVTVAPSNFANGVAEWISSNSRSQAYKVTCSVRQSSAQ NRK YTIKV E VPKVATQTVGGVELPVA A WR S YLNMELTIPIFATN SDCELIV KAMQ GLLKDGNPIPSAIAANSGIY (SEQ ID NO: 12217);
[0587] QIYEGKLTAEGLRFGIVASRFNHALVDRLVEGAIDCIVRHGGREEDITLVRVPGSW EIPVAAGELARKEDIDAVIAIGVLIRGATPHFDYIASEVSKGLAQLSLELRKPITFGV ITADTLEQA1ERAGTKHGNKGWEAALSAIEMANLFKSLR (SEQ ID NO: 12218);
[0588] NIIQGNLVGTGLKIGIVVGRFNDFITSKLLSGAEDALLRHGVDTNDIDVAWVPGAF EIPFAAKKMAETKKYDAIITLGTVIRGATTHYDYVCNEAAKGIAQAAQTTGVPVIF GIVTTENIEQAIETAGTKAGNKGVDCAVSAIEMANLQRSFE (SEQ ID NO: 12219); KTINSVDTKEFLNHQVANLNVFTVKIHQTHWYMRGHNFTTLHEKMDDLYSEFGE QMDEVAERLLAIGGSPFSTLKEFLENASVEEAPYTKPKTMDQLMEDLVGTLELLR DEYKQGIELTDKEGDDVTNDMLIAFKASIDKHIWMFKAFLGKAPLE (SEQ ID NO: 12220).
[0589] In some embodiments, the self-assembling peptide shares at least 50% identity with the sequences described herein above or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs, thereof.
[0590] Given the disclosed amino acid sequences of the self- assembling peptides, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above self- assembling peptides. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the self- assembling peptide is encoded by the self-assembling sequence which may be either a DNA, an RNA, or an mRNA.
[0591] Linker sequence and linker peptide
[0592] The multisubunit nucleic acid sequence and multisubunit peptide includes linker sequence and linker peptide respectively. The linker sequence comprises a sequence of nucleotides, either deoxyribonucleotides or ribonucleotides, that encodes a linker peptide. The linker sequence includes codon optimized sequences, fragments, mutants, variants, comparable equivalents, functional analogs, or a combination thereof. In some embodiments, the linker sequence is a DNA, an RNA, or an mRNA.
[0593] In some embodiments, the linker peptide connects the target peptide with the selfassembling peptide in a polypeptide. In some embodiments, the linker peptide connects the signal peptide with a polypeptide. In some other embodiments, the linker peptide connects the signal peptide with the first polypeptide. In some embodiments, one linker peptide connects the cleavage peptide with the target peptide and another linker peptide connects the target peptide with the self-assembling peptide in a polypeptide. In some embodiments, the linker peptide connects two cleavage peptides. Any suitable linker peptides can be employed in accordance with the present disclosure.
[0594] In some embodiments, the linker peptide is an amino acid linker, a zipper motif, a foldon, a scaffold, or a combination thereof. In some embodiments, the linker peptide is an amino acid linker. In some embodiments, the linker peptide is a zipper motif. In some embodiments, the linker peptide is a foldon. In some embodiments, the linker peptide is a scaffold. In some embodiments, the linker peptide comprises a combination of an amino acid linker and a zipper motif. In some embodiments, the linker peptide comprises a combination of an amino acid linker and a foldon. In some embodiments, the linker peptide comprises a combination of an amino acid linker and a scaffold. In some embodiments, the linker peptide comprises a combination of a zipper motif and a foldon. In some embodiments, the linker peptide comprises a combination of a zipper motif and a scaffold. In some embodiments, the linker peptide comprises a combination of a foldon and a scaffold. In some embodiments, the linker peptide comprises a combination of an amino acid linker, a zipper motif, and a foldon. In some embodiments, the linker peptide comprises a combination of an amino acid linker, a zipper motif, and a scaffold. In some embodiments, the linker peptide comprises a combination of an amino acid linker, a foldon, and a scaffold. In some embodiments, the linker peptide comprises a combination of a zipper motif, a foldon, and a scaffold. In some embodiments, the linker peptide comprises a combination of an amino acid linker, a zipper motif, a foldon, and a scaffold.
[0595] In some embodiments, the amino acid linker comprises of about 2-49 amino acids, 2-40 amino acids, 2-30 amino acids, 2-20 amino acids, 2-15 amino acids, or 2-10 amino acids. In some embodiments, the amino acid linker comprises a glycine serine linker, a glycine proline linker, a glycine threonine linker, an alanine serine linker, any combination of two amino acids, or a combination thereof. The glycine proline linker comprises of glycine (G) and proline (P) amino acids consecutively without any preference of order of appearance of either glycine or proline. In some embodiments, the glycine proline linker is 2-49 amino acids in length.
[0596] The glycine threonine linker comprises of glycine (G) and threonine (T) amino acids consecutively without any preference of order of appearance of either glycine or threonine. In some embodiments, the glycine threonine linker is 2-49 amino acids in length.
[0597] The alanine serine linker comprises of alanine (A) and serine (S) amino acids consecutively without any preference of order of appearance of either alanine or serine. In some embodiments, the alanine serine linker is 2-49 amino acids in length.
[0598] The glycine serine linker comprises of glycine (G) and serine (S) amino acids consecutively without any preference of order of appearance of either glycine or serine. In some embodiments, the glycine serine linker is 2-49 amino acids in length.
[0599] Exemplary amino acid linkers include, but not limited to, the ones represented by the following amino acid sequences or comparable equivalents, or their combinations, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof: GSG (SEQ ID NO: 12);
[0600] GSGG (SEQ ID NO: 13);
[0601] GGSGG (SEQ ID NO: 14);
[0602] GGSGGGGSGG (SEQ ID NO: 15);
[0603] GGSGGGGSGGGGSGG (SEQ ID NO: 16);
[0604] SGGSGG (SEQ ID NO: 17);
[0605] GGGGSGGGGS (SEQ ID NO: 18);
[0606] GGGGSGGGGSGGGGS (SEQ ID NO: 19);
[0607] PGG (SEQ ID NO: 12221);
[0608] In some embodiments, the amino acid linkers share at least 50% identity with the sequences described herein above or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0609] Given the disclosed amino acid sequences of the amino acid linkers, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above amino acid linkers. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the amino acid linker is encoded by an amino acid linker sequence which may be either a DNA, an RNA, or an mRNA.
[0610] In some embodiments, linker peptide is a zipper motif. A zipper motif comprises of a sequence of amino acids encoded by a zipper sequence. Zipper motifs are generally a class of protein-protein interaction domains that facilitates formation of a complex i.e., enables two, three, four, five, or six homologous or heterologous polypeptides to associate themselves into a polypeptide cluster. In some embodiments, the zipper motif is a leucine zipper, an isoleucine zipper, or any synthetic zipper.
[0611] Exemplary zipper motifs include, but not limited to, the ones represented by the following amino acid sequences or comparable equivalents, or a combination thereof, including codon optimized sequences, fragments, mutants, variants, or functional analogs thereof:
[0612] RIARLEEKVKTLKAQNSELASTANMLREQVAQLK QKVMNY (SEQ ID NO: 12222); LTDTLQAETDQLEDKKSALQTEIANLLKEKEKLEFILAAY (SEQ ID NO: 12223);
[0613] RNAYLRKKIARLKKDNLQLERDEQNLEKIIANLRDEIARLENEVA (SEQ ID NO: 12224);
[0614] LVAQLENEVASLENENETLKKKNLHKKDLIAYLEKEIANLRKKIE (SEQ ID NO: 12225);
[0615] In some embodiments, the zipper motif shares at least 50% identity with the sequences described herein above or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0616] Given the disclosed amino acid sequences of the zipper motifs, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above zipper motifs. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the zipper motif is encoded by a zipper sequence which may be either a DNA, an RNA, or an mRNA.
[0617] In some embodiments, the linker peptide is a foldon. A foldon comprises of a sequence of amino acids encoded by a foldon sequence. The foldon sequence includes codon optimized sequences, fragments, mutants, variants, comparable equivalent, functional analogs or a combination thereof. A foldon enables two or more homologous polypeptides to organise to form an oligomeric complex. In some embodiments, foldon also helps in orientation of a polypeptide such that the domains or epitopes on the target peptide are exposed or displayed for interaction or communication with cells or biomolecules or immune system.
[0618] Exemplary foldons includes, but not limited to, the ones represented by the following amino acid sequences or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof:
[0619] YIPEAPRDGQAYVRKDGEWVLLSTFL (SEQ ID NO: 20);
[0620] HENEISHHAKEIERLQKEIERHKQSIKKLKQSE (SEQ ID NO: 21);
[0621] In some embodiments, the foldon shares at least 50% identity with the sequences described herein above or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0622] Given the disclosed amino acid sequences of the foldons, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above foldons. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the foldon is encoded by a foldon sequence which may be either a DNA, an RNA or an mRNA.
[0623] In some embodiments, the linker peptide is a scaffold. A scaffold comprises of a sequence of amino acids encoded by a scaffold sequence. The scaffold sequence includes codon optimized sequences, fragments, mutants, variants, comparable equivalent, functional analogs, or a combination thereof. A scaffold provides structural and / or functional integrity or support to the target peptide and may also help in orientation of target peptide such that the domains or epitopes of the target peptide are exposed or displayed for interaction or communication with cells or biomolecules or immune system.
[0624] Exemplary scaffolds include, but not limited to, the ones represented by the following amino acid sequences or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs, thereof: VDNKFNKEMRNAYWEIALLPNLNNQQKRAFIRSLYDDPSQSANLLAEAKKLNDA QAPK (SEQ ID NO: 22);
[0625] QDSTSDLIPAPPLSKVPLQQNFQDNQFHGKWYVVGKAGNHDLREDKDPRKMQAT IYELKEDKSYNVTNVRFVHKKCNYRIWTFVPGSQPGEFTLGNIKSWPGLTSWLVR VVSTNYNQHAMVFFKRVYQNRELFEITLYGRTKELTNELKENFIRFSKSLGLPENH IVFPVPIDQCIDGSAWSHPQFEK (SEQ ID NO: 23);
[0626] VSDVPRDLEVVAATPTSLLISWDAPAVTVRYYRITYGETGGNSPVQEFTVPGSKST ATISGLKPGVDYTITVYAVTGRGDSPASSKPISINYRT (SEQ ID NO: 24);
[0627] GCPRILMRCKQDSDCLAGCVCGPNGFCG (SEQ ID NO: 25);
[0628] MRGSHHHHHHGSDLGKKLLEAARAGQDDEVRILMANGADVNATDNDGYTPLHL AASNGHLEIVEVLLKNGADVNASDLTGITPLHLAAATGHLEIVEVLLKHGADVNA YDNDGHTPLHLAAKYGHLEIVEVLLKHGADVNAQDKFGKTAFDISIDNGNEDLA EILQ (SEQ ID NO: 26);
[0629] MRGSHHHHHHGSVKVKFFWNGEEKEVDTSKIVWVKRAGKSVLFIYDDNGKNGY GDVTEKDAPKELLDMLARAEREKKL (SEQ ID NO: 27);
[0630] MLPAPKNLVVSEVTEDSARLSWDDPAAFYESFLIQYQESEKVGEAIVLTVPGSERS YDLTGLKPGTEYTVSIYGVHNVYKDTNMRGLPLSAIFTTGGHHHHHH (SEQ ID NO: 28);
[0631] ETDICKLPKDEGTCRDFILKWYYDPNTKSCARFWYGGCGGNENKFGSQKECEKV CAPV (SEQ ID NO: 29);
[0632] MIPGGLSEAKPATPEIQEIVDKVKPQLEEKTNETYGKLEAVQYKTQVVAGTNYYI KVRAGDNKYMHLKVFKSLPGQNEDLVLTGYQVDKNKDDELTGF (SEQ ID NO: 30);
[0633] PCSAFEFHCLSGECIHSSWRCDGGPDCKDKSDEENCA (SEQ ID NO: 31);
[0634] MQIFVKTLTGKTITLEVEPSDTIENVKAKIQDKEGIPPDQQRLIFAGKQLEDGRTLS DYNIQKESTLHLVLRLRGG (SEQ ID NO: 32);
[0635] MGSIIFLEDRAFQGRIYGCTTDCPNLQPYFSRCNSIVVQSGCWMIYERPNYQGHQY FLRRGEYPDYQQWMGLSDSIRSCCLIPPHSGAYRMKIYDRDELRGQMSELTDDCL SVQDRFHLTEIHSLNVLEGSWILYEMPNYRGRQYLLRPGEYRRFLDWGAPNAKV GSLRRVMDLYLEHHHHHH (SEQ ID NO: 33);
[0636] AGHRIAWLLMMGHPRQQLAIIFGIGVSTLYRYFPA (SEQ ID NO: 34);
[0637] AFSKSEEARHSSLERECIEEICDHAEAWDIMM (SEQ ID NO: 35);
[0638] RECDYCGTDIEPGTGGMAVHGDGATTHFCSHRCAWDAMMGAEARNLEWTDTA R (SEQ ID NO: 36);
[0639] CSQNEYFDSLLHACIPCQLRCSGAPHRCAWDCMM (SEQ ID NO: 37);
[0640] EHIPGTLAARLSHRAAWDLMMHSLDASQGTATGPRGIFTAEDALKLVQLKQTGK TFPTYKCGHRFAWDCMMGSGLNGAACFAVKIADLPVYSCECAIGFMGQRCEYKE (SEQ ID NO: 38);
[0641] ACYGHRCAWDCMMLGFSSGKCINSKCKCYK (SEQ ID NO: 39);
[0642] GEYVVEKVLDKRVVKGKVEYLLKWKGFSDEDNTWEPDENLDGHRLAWDFMM ADVYEVEAILADRVNKNGINEYYIKWAGYDWYDNTWEPEQNLFGAGHRLAWW MMR (SEQ ID NO: 40);
[0643] AGTIKITQTRSAIGRLPAHKATLLGLGLRRIGHTVEREDGHRIAWDIMMVSFMVKV EG (SEQ ID NO: 41);
[0644] GIPCGESCGSPCISSAIGCSCKLINTNGSWHIVCYRN (SEQ ID NO: 42);
[0645] GKCPETFDAWYCLNDAHCFAVLINTNGSWHIVYSCECAIGFMGQRCEYKE (SEQ ID NO: 43);
[0646] QEEADRTVFVGNLEARVREEILYELFLQAGPLTKVTICKDREGKPKSFGFVCFKHP ESVSYAIALAGLINLNGSWIIVSGPSSG (SEQ ID NO: 44);
[0647] NEEDAGKMFVGGLSWDTSKKDLKDYFTKFGEVVDCTIKMDPNTGRSRGFGFILF KDAASVEKVLDAGLHNLNGSWIIPKKA (SEQ ID NO: 45);
[0648] SGNIFIKNLDKSIDNKALYDTFSAFGNILSCKVVCDEQGSKGYGFVHFETQEAAER AIAKMGLMNLNGSWVIVGRFKSRKE (SEQ ID NO: 46);
[0649] PSRVVYLGSIPYDQTEEQILDLCSNVGPVINLKMMFDPQTGRSKGYAFIEFRDLESS ASAVGALGLYNLNGSWLICGYSSNSDISGVSLEHHHH (SEQ ID NO: 47);
[0650] LAILVFGYPETMANQVIAYFQEFGTILEDFEVLRKPQAMTVGLQDRQFVPIFSGNS WTKITYDNPASAVDALAEGLANFNGSWLLVIPYTKDAVERLQ (SEQ ID NO: 48); RLVNCNGSWLIGLDRPPYPGAKGEDIYNNVSRKAWDEWQKHQTMLINERRLNM MNAEDRKFLQQEMDKFLSGEDY (SEQ ID NO: 49);
[0651] FAVESIEKLRNRNGSWEILVKWRGWSPKYNTWEPEENIG (SEQ ID NO: 50);
[0652] MRDFFVITNSLYNFNGSWYIKGAVLHVSPTQKRAFWVIADQENFIKQVNKNIEYV EKQASPAFLQRIVEIYQVKFEGKNVG (SEQ ID NO: 51).
[0653] In some embodiments, the scaffold shares at least 50% identity with the sequences described herein above or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0654] Given the disclosed amino acid sequences of the scaffold, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above scaffold. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the scaffold is encoded by the scaffold sequence which may be either a DNA, an RNA, or an mRNA.
[0655] In some embodiments, the linker peptide comprises an amino acid linker and a zipper motif. In some embodiments, the linker peptide comprises an amino acid linker followed by a zipper motif. In some embodiments, the linker peptide comprises a zipper motif followed by an amino acid linker. In some embodiments, the linker peptide comprises an amino acid linker followed by a zipper motif and another an amino acid linker. In some embodiments, the linker peptide comprises a zipper motif followed by an amino acid linker, and another zipper motif.
[0656] In some embodiments, the linker peptide comprises an amino acid linker and a foldon. In some embodiments, the linker peptide comprises an amino acid linker followed by a foldon. In another embodiment, the linker peptide comprises a foldon followed by an amino acid linker. In some embodiments, the linker peptide comprises an amino acid followed by a foldon and another amino acid linker. In some embodiments, the linker peptide comprises a foldon followed by an amino acid linker and another foldon.
[0657] In some embodiments, the linker peptide comprises an amino acid linker and a scaffold. In some embodiments, the linker peptide comprises an amino acid linker followed by a scaffold. In some embodiments, the linker peptide comprises a scaffold followed by an amino acid linker. In some embodiments, the linker peptide comprises an amino acid linker followed by a scaffold and another amino acid linker. In some embodiments, the linker peptide comprises a scaffold followed by an amino acid linker and another scaffold.
[0658] In some embodiments, the linker peptide comprises a zipper motif and a scaffold. In some embodiments, the linker peptide comprises a zipper motif followed by a scaffold. In some embodiments, the linker peptide comprises a scaffold followed by a zipper motif. In some embodiments, the linker peptide comprises a scaffold followed by a zipper motif and another scaffold. In some embodiments, the linker peptide comprises a zipper motif followed by a scaffold and another zipper motif.
[0659] In some embodiments, the linker peptide comprises a foldon and a scaffold. In some embodiments, the linker peptide comprises a foldon followed by a scaffold. In some embodiments, the linker peptide comprises a scaffold followed by a foldon. In some embodiments, the linker peptide comprises a foldon followed by a scaffold and another foldon. In some embodiments, linker peptide comprises a scaffold followed by a foldon and another scaffold.
[0660] In some embodiments, the linker peptide comprises an amino acid linker, a zipper motif, and a scaffold. In some embodiments, the linker peptide comprises an amino acid linker followed by a zipper motif, and a scaffold. In some embodiments, the linker peptide comprises a zipper motif followed by an amino acid linker, and a scaffold. In some embodiments, the linker peptide comprises a scaffold followed by an amino acid linker, and a zipper motif. In some embodiments, the linker peptide comprises a first amino acid linker followed by a zipper motif, a second amino acid linker followed by a scaffold, and a third amino acid linker. In some embodiments, the linker peptide comprises a first amino acid linker followed by a scaffold, a second amino acid linker followed by a zipper motif, and a third amino acid linker. In some embodiments, the linker peptide comprises a scaffold followed by a first amino acid linker, and a zipper motif followed by a second amino acid linker. In some embodiments, the linker peptide comprises a first amino acid linker followed by a scaffold, and a second amino acid linker followed by a zipper motif.
[0661] In some embodiments, the linker peptide comprises an amino acid linker, a foldon and a scaffold. In some embodiments, the linker peptide comprises an amino acid linker followed by a foldon, and a scaffold. In some embodiments, the linker peptide comprises a foldon followed by an amino acid linker, and a scaffold. In some embodiments, the linker peptide comprises a scaffold followed by an amino acid linker, and a foldon. In some embodiments, the linker peptide comprises a first amino acid linker followed by a foldon, a second amino acid linker followed by scaffold, and a third amino acid linker. In some embodiments, the linker peptide comprises a first amino acid linker followed by a scaffold, a second amino acid linker followed by a foldon, and a third amino acid linker. In some embodiments, the linker peptide comprises a scaffold followed by a first amino acid linker, and a foldon followed by a second amino acid linker. In some embodiments, the linker peptide comprises a first amino acid linker followed by a scaffold, and a second amino acid linker followed by a foldon.
[0662] Cleavage sequence and cleavage peptide
[0663] The multisubunit nucleic acid sequence and multisubunit peptide includes cleavage sequence and cleavage peptide respectively. The cleavage sequence comprises of a sequence of nucleotides, either deoxyribonucleotides or ribonucleotides, that encodes a cleavage peptide. The cleavage sequence includes codon optimized sequences, fragments, mutants, variants, comparable equivalents, functional analogs, or a combination thereof. In some embodiments, the cleavage sequence is a DNA, an RNA, or an mRNA.
[0664] The cleavage peptide connects one polypeptide with another polypeptide, for example, the adjacent polypeptide. The cleavage peptide carries one or more cleavage sites. In some embodiments, the cleavage peptide comprises one or more cleavage peptides, for example cleavage peptide- 1, cleavage peptide-2 and so on. In some embodiments, the cleavage peptide optionally comprises a linker peptide between two cleavage peptides.
[0665] In some embodiments, the cleavage peptide facilitates the action of cellular proteases to cleave the multisubunit polypeptide into individual polypeptides or self cleaves into individual polypeptides. In some embodiments, the resulting polypeptides comprise either a target peptide, a linker peptide and a self-assembling peptide or a linker peptide, target peptide, linker peptide and a self-assembling peptide or a combination thereof. In some embodiments, the polypeptide, in addition to these peptides, may also have some residues (amino acids) of cleavage peptide. Any cleavage peptide that is susceptible to the action of cellular proteases or a cleavage peptide that has the ability to undergo self cleavage can be employed in accordance with the present disclosure.
[0666] In some embodiments, the cleavage peptide is a substrate for cellular proteases. In some embodiments, the cleavage peptide is a substrate for golgi specific proteases. In some embodiments, the cleavage peptide is a self cleaving peptide. In some embodiments, the cleavage peptide comprises two or more cleavage peptides (for example, cleavage peptide- 1, cleavage peptide-2 and so on), optionally linked by a linker peptide, wherein one cleavage peptide is a substrate for cellular proteases and the other cleavage peptide is a self cleaving peptide. In some embodiments, the cleavage peptide is a golgi specific cleavage peptide i.e., susceptible to action of golgi specific proteases.
[0667] Exemplary cleavage peptide includes, but not limited to, the one represented by the following amino acid sequence or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof:
[0668] RRKRSVS (SEQ ID NO: 52);
[0669] GIRRKRSVSH (SEQ ID NO: 53);
[0670] VQREKRAVGI (SEQ ID NO: 54); SIRHKREPSV (SEQ ID NO: 55);
[0671] KRRQRRRPPQ (SEQ ID NO: 56);
[0672] KIRRRRDVVD (SEQ ID NO: 57);
[0673] HNRTKRSTDG (SEQ ID NO: 58);
[0674] RKRRKRELET (SEQ ID NO: 59);
[0675] THRTRRSTSD (SEQ ID NO: 60);
[0676] SRRKRRSAST (SEQ ID NO: 61);
[0677] NLRRRRDLVD (SEQ ID NO: 62);
[0678] LRRRRRDAGN (SEQ ID NO: 63);
[0679] ATNFSLLKQAGDVEENPGP (SEQ ID NO: 64);
[0680] EGRGSLLTCGDVEENPGP (SEQ ID NO: 65);
[0681] QCTNYALLKLAGDVESNPGP (SEQ ID NO: 66).
[0682] In some embodiments, the cleavage peptide shares at least 50% identity with the sequence described herein above or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0683] Given the disclosed amino acid sequences of the cleavage peptide, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above cleavage peptide. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the cleavage peptide is encoded by a cleavage sequence which may be either a DNA, an RNA, or an mRNA.
[0684] Signal sequence and signal peptide
[0685] The multisubunit nucleic acid sequence and multisubunit peptide includes a signal sequence and a signal peptide respectively. The signal sequence comprises of a sequence of nucleotides, either deoxyribonucleotides or ribonucleotides, that encodes a signal peptide. The signal sequence includes codon optimized sequences, fragments, mutants, variants, comparable equivalents, functional analogs, or a combination thereof. In some embodiments, the signal sequence is a DNA, an RNA, or an mRNA.
[0686] The signal peptide is present upstream (N-terminus or amino-terminus) of one or more polypeptides. In some embodiments, the signal peptide is present on the N-terminus of all or some polypeptides. In some embodiments, the signal peptide is present on the N- terminus of the first polypeptide. In some embodiments, the signal peptide is present upstream (N-terminus) of some polypeptides. In some embodiments, the signal peptide is present upstream (N-terminus) of each of the polypeptides.
[0687] In some embodiments, the signal peptide transports the multisubunit peptide to cell organelles. In some embodiments, the signal peptide transports the multisubunit peptide to golgi body or golgi apparatus / complex. Any signal peptide that transports the multisubunit peptide to golgi bodies can be employed in accordance with the present disclosure. In some embodiments, the signal peptide is a golgi targeting signal peptide i.e., directs the multisubunit peptide to golgi complex.
[0688] Exemplary signal peptide includes, but not limited to, the one represented by the following amino acid sequence or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof:
[0689] MPSSVSWGILLLAGLCCLVPVSLAEDPQGDAA (SEQ ID NO: 67); MGSSVSWGILLLAGLCCLVPVSLAEDPQGDAA (SEQ ID NO: 68); MASSVSWGILLLAGLCCLVPVSLAEDPQGDAA (SEQ ID NO: 69); MDMRAPAGIFGFLLVLFPGYRS (SEQ ID NO: 70);
[0690] MKWVTFISLLFLFSSAYS (SEQ ID NO: 71);
[0691] MDWTWRVFCLLAVTPGAHP (SEQ ID NO: 72);
[0692] MAWSPLFLTLITHCAGSWA (SEQ ID NO: 73);
[0693] MTRLTVLALLAGLLASSRA (SEQ ID NO: 74);
[0694] MARPLCTLLLLMATLAGALA (SEQ ID NO: 75);
[0695] MRSLVFVLLIGAAFA (SEQ ID NO: 76);
[0696] MSRLFVFILIALFLSAIIDVMS (SEQ ID NO: 77);
[0697] MGMRMMFIMFMLVVLATTVVS (SEQ ID NO: 78);
[0698] MRAFLFLTACISLPGVFG (SEQ ID NO: 79);
[0699] MKFQSTLLLAAAAGSALA (SEQ ID NO: 80);
[0700] MASSLYSFLLALSIVYIFVAPTHS (SEQ ID NO: 81);
[0701] MKTHYSSAILPILTLFVFLSINPSHG (SEQ ID NO: 82);
[0702] MESVSSLFNIFSTIMVNYKSLVLALLSVSNLKYARG (SEQ ID NO: 83);
[0703] MKAAQILTASIVSLLPIYTSA (SEQ ID NO: 84);
[0704] MIKLKFGVFFTVLLSSAYA (SEQ ID NO: 85);
[0705] MGVKVLFALICIAVAEA (SEQ ID NO: 86). In some embodiments, the signal peptide shares at least 50% identity with the sequence described herein above or comparable equivalents, or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, or functional analogs thereof.
[0706] Given the disclosed amino acid sequences of the signal peptide, a person skilled in the art would be able to deduce all possible DNA or RNA sequences that encodes the above signal peptide. Such DNA or RNA sequences are deemed to be incorporated in this disclosure. In some embodiments, the signal peptide is encoded by a signal sequence which may be either a DNA, an RNA, or an mRNA.
[0707] Synthesis of multisubunit nucleic acid sequences
[0708] Multisubunit nucleic acid sequence according to the present disclosure can be either a DNA, an RNA or an mRNA. The multisubunit nucleic acid sequence as described herein can be synthesized by molecular biology or genetic engineering techniques well known in the art, for example, using recombinant expression system, chemical synthesis, or in vitro transcription (IVT).
[0709] In some embodiments, the multisubunit nucleic acid sequence is obtained through a single IVT process or step. In some embodiments, the multisubunit nucleic acid sequence obtained through single IVT process or step is an mRNA. In some embodiments, the multisubunit nucleic acid sequence is obtained or synthesized through a single in vitro transcription (IVT) process or step.
[0710] In some embodiments, the multisubunit nucleic acid sequence is a messenger RNA (mRNA). The mRNA encodes a multisubunit peptide as described herein.
[0711] Typically, an mRNA includes at least a coding region (which encodes the multisubunit peptide), a 5’ UTR, a 3’ UTR, a 5’ cap and a 3’ poly(A) tail. UTR (untranslated regions) flanks the coding region or open reading frame (ORF). The 5’ UTR and the 3’ UTR are sections of the mRNA before the start codon and after the stop codon respectively. The 5’ UTR has a cap (5’ cap) consisting of altered nucleotides. mRNA also contains a poly adenylated region at its 3’ end having adenine nucleotides called poly(A) tail.
[0712] In some embodiments, the mRNA is unmodified or modified or a combination of both. The modification may be in the nucleobase of the nucleotide, or sugar moiety of the nucleotide, or the phosphate of the nucleotide. In some embodiments, unmodified mRNA comprises naturally occurring nucleosides, for example, adenosine, guanosine, cytidine, and uridine. mRNA comprises one or more modified nucleosides, for example, adenosine analog, guanosine analog, cytidine analog, or uridine analog.
[0713] In some embodiments, the one or more modified nucleosides is a nucleoside analog selected from 2-aminoadenosine, 3-methyl adenosine, 7-deazaadenosine, 7- deazaguanosine, 8-oxoadenosine, or 8-oxoguanosine or a combination thereof.
[0714] In some embodiments, the one or more modified nucleosides is a uridine analog selected from propynyl-uridine, pseudouridine, C5 -bromouridine, C5-fluorouridine, C5- iodouridine, C5-propynyl-uridine, 5-aza-uridine, 2-thio-5-aza-uridine, 2-thio-uridine, 4- thio-pseudouridine, 2-thio-pseudouridine, 5-hydroxy-uridine, 3-methyl-uridine, 5- carboxymethyl-uridine, 1 -carboxymethyl-pseudouridine, l-methyl-3-(3-amino-3- carboxypropyl)pseudouridine, 2-thio-2 ’ -O-methyl-uridine, 5 -methoxycarbonylmethyl-2 ’ - O-methyl-uridine, 5-carboxymethylaminomethyl-2 ’ -O-methyl-uridine, 3 ,2 ’ -O-dimethyl- uridine, 5-propynyl-uridine, 1-propynyl-pseudouridine, 5-taurinomethyl-uridine, 1- taurinomethyl-pseudouridine, 5 -taurinomethyl-2-thio-uridine, 1 -taurino-4-thio- pseudouridine, 1-methyl-pseudouridine, 4-thio-l-methyl-pseudouridine, 2-thio- 1-methyl- pseudouridine, 1 -methyl- 1 deaza-pseudouridine, 2-thio- 1 -methyl- 1 -deaza-pseudouridine, dihydro-uridine, dihydro-pseudouridine, 2-thio-dihydro-uridine, 2-thio-dihydro- pseudouridine, 2-methoxy-uridine, 2-methoxy-4-thio-uridine, 4-methoxy-pseudouridine, or 4-methoxy-2-thio-pseudouridine, or a combination thereof.
[0715] In some embodiments, the one or more modified nucleosides is a cytidine analog selected from 5-methylcytidine, C5-propynyl-cytidine, C5-methylcytidine, pseudoisocytidine, 1-methyl-pseudoisocytidine, pyrrolo-pseudoisocytidine, 4-thio- pseudoisocytidine, 4-thio- 1 -methyl-pseudoisocytidine, 4-thio- 1 -methyl- 1 -deaza- pseudoisocytidine, 1 -methyl- 1-1 deaza-pseudoisocytidine, 4-methoxy- 1 -methyl- pseudoisocytidine, or a combination thereof.
[0716] Methods for making modified nucleosides are well known in the art.
[0717] In some embodiments, the modified nucleoside is pseudouridine, for example, 1- methyl-pseudouridine, 1-propynyl-pseudouridine, 1 -carboxymethyl-pseudouridine, 1- methyl-3-(3-amino-3-carboxypropyl)pseudouridine, 4-methoxy-pseudouridine, or 4- methoxy-2-thio-pseudouridine 4-thio-pseudouridine, 2-thio-pseudouridine, 4-thio-l- methyl-pseudouridine, 2-thio- 1-methyl-pseudouridine, dihydro-pseudouridine, or a combination thereof. In some embodiments, mRNA is produced using recombinant expression system, chemically synthesized, or obtained through in vitro transcription.
[0718] In some embodiments, the multisubunit nucleic acid sequence is obtained or synthesized through a single IVT process or step. mRNAs according to the present disclosure may be synthesized via in vitro transcription (IVT). Briefly, IVT is typically performed with a DNA template containing a promoter, a pool of ribonucleotide triphosphates, a buffer system that may include DTT and magnesium ions, and an appropriate RNA polymerase (e.g., T3, T7, or SP6 RNA polymerase), DNase I, pyrophosphatase, and / or RNase inhibitor. The exact conditions may vary according to the specific application. Methods of making mRNA through IVT reaction is well known in the art (see for example, Beckert, Bertrand and Masquida, Benoit Methods in Molecular Biology (2011) 703, 29-41 ; Brunelle, Julie L. and Green Rachel Methods in Enzymology (2013) 530, 101-114; Kamakaka, Rohinton T. and Kraus W. Lee Current Protocols in Cell Biology (1999) 11.6.1-11.6.17; Kanwal, Fariha et al. Cellular Physiology and Biochemistry (2018) 48:1915-1927; WO2018157153;
[0719] WO202Q185811; W02022082001).
[0720] In some embodiments, the in vitro transcription occurs in a single batch. In some embodiments, IVT reaction includes capping and tailing reactions either co- transcriptionally or separately. A cap analog is added to the in vitro transcription reaction and will be incorporated at the 5’ end of the mRNA during the reaction. Alternative method of capping involves adding the cap post-transcriptionally through an enzymatic reaction. The poly (A) tail can be incorporated into the DNA template sequence, and thus the poly (A) tail will be incorporated into the mRNA by T7 RNA polymerase during the in vitro transcription. Alternative method of tailing involves adding the poly (A) tail post- transcriptionally through an enzymatic reaction. In some embodiments, capping and tailing reactions are performed co-transcriptionally i.e., during the IVT reaction. In some embodiments, capping and tailing reactions are performed separately from IVT reaction i.e., post transcriptionally. mRNA produced as a result of IVT reaction may be purified using techniques well known in the art, such as, centrifugation, filtration and / or chromatographic techniques. The purification of mRNA may be accomplished before capping and tailing steps are performed or after capping and tailing. The synthesized mRNA may be purified by ethanol precipitation or filtration or chromatography methods. In some embodiments, tangential flow filtration is used to purify mRNA. In some embodiments, mRNA is purified by chromatographic step. In other embodiments, mRNA is purified by a combination of filtration and chromatography steps.
[0721] In some embodiments, a suitable mRNA sequence is an mRNA sequence encoding a protein, peptide, polypeptide. In some embodiments, a suitable mRNA sequence is codon optimized for efficient expression in a host cell or organism. Codon optimization typically includes modifying a naturally-occurring or wild-type nucleic acid sequence encoding a peptide, polypeptide, or protein to achieve the highest possible expression of peptide, polypeptide, protein, or an antibody without altering the amino acid sequence.
[0722] In some embodiments, the mRNA is circular. In other embodiments, the mRNA is linear.
[0723] In some embodiments, the mRNA is self-amplifying or self-replicating.
[0724] In some embodiments, mRNA is few hundred nucleotides long to several thousand nucleotides long. In some embodiments, mRNA is about 0.5 kb, 1 kb, 1.5 kb, 2 kb, 2.5 kb, 3 kb, 3.5 kb, 4 kb, 4.5 kb, 5 kb, 5.5 kb, 6 kb, 6.5 kb, 7.0 kb, 7.5 kb, 8 kb, 8.5 kb, 9 kb, 9.5 kb, 10 kb, 10.5 kb, 11 kb, 11.5 kb, 12 kb, 12.5 kb, 13 kb, 13.5 kb, 14 kb, 14.5 kb, 15 kb, 16 kb, 17 kb, 18 kb, 19 kb, 20 kb, 21 kb, 22 kb, 23 kb, 24 kb, 25 kb, 26 kb, 27 kb, 28 kb, 29 kb, 30 kb in length, or a fraction thereof. In some embodiments, mRNA is about 0.5 to 30 kb, 0.5 to 25 kb, 0.5 to 20 kb in length, or any range therein. In some embodiments, mRNA is about 1 to 20 kb, 1 to 18 kb, 1 to 16 kb, 1 to 14 kb, 1 to 12 kb, 1 to 10 kb, 1 to 9 kb, 1 to 8 kb, 1 to 7 kb, 1 to 6 kb, 1 to 5 kb in length or any range therein.
[0725] In some embodiments, mRNA is about 0.5 kb to about 1 kb, about 1 kb to about 2 kb, about 2 kb to about 3 kb, about 3 kb to about 4 kb, about 4 kb to about 5 kb, about 5 kb to about 6 kb, about 6 kb to about 7 kb, about 7 kb to about 8 kb, about 8 kb to about 9 kb, about 9 kb to about 10 kb, about 10 kb to about 11 kb, about 11 kb to about 12 kb, about 12 kb to about 13 kb, about 13 kb to about 14 kb, about 14 kb to about 15 kb, about 15 kb to about 16 kb, about 16 kb to about 17 kb, about 17 kb to about 18 kb, about 18 kb to about 19 kb, about 19 kb to about 20 kb, about 20 kb to about 21 kb, about 21 kb to about 22 kb, about 22 kb to about 23 kb, about 23 kb to about 24 kb, about 24 kb to about 25 kb, about 25 kb to about 26 kb, about 26 kb to about 27 kb, about 27 kb to about 28 kb, about 28 kb to about 29 kb, about 29 kb to about 30 kb in length, or any range therein.
[0726] The multisubunit nucleic acid sequence as described herein, express multisubunit peptide. Polypeptide nanoparticle
[0727] The multisubunit peptide, encoded by the multisubunit nucleic acid, comprises multiple repeats of polypeptide comprising either a target peptide, a linker peptide, and a self-assembling peptide or a linker peptide, target peptide, a linker peptide, and a selfassembling peptide, or a combination thereof, interspersed with cleavage peptide (see illustration in figures), wherein the target peptide is obtained or derived from:
[0728] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or combination thereof of respiratory syncytial virus.
[0729] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0730] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0731] (d) a combination of (a), (b), and / or (c).
[0732] In some embodiments, the total number of polypeptides present in a multisubunit peptide are up to 100 polypeptides. In some embodiments, one or more polypeptides in the multisubunit peptide has identical target peptides (homologous polypeptides). In some embodiments, one or more polypeptides in the multisubunit peptide has different target peptides (heterologous polypeptides).
[0733] A multisubunit peptide as described herein is encoded by the multisubunit nucleic acid sequence as described herein. Each multisubunit peptide comprises two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide or a combination thereof, wherein the target peptide is obtained or derived from:
[0734] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or combination thereof of respiratory syncytial virus.
[0735] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0736] (d) a combination of (a), (b), and / or (c).
[0737] The polypeptides are connected with each other through a cleavage peptide. The multisubunit peptide includes a signal peptide upstream (N-terminus) of one or more polypeptides. In some embodiments, the multisubunit peptide includes a signal peptide upstream (N-terminus) of each of all or some polypeptides. In some embodiments, the multisubunit peptide optionally includes a signal peptide upstream (N-terminus) of each polypeptide. In some embodiments, the multisubunit peptide includes a signal peptide upstream (N-terminus) of some polypeptides. In some embodiments, the multisubunit peptide includes a signal peptide upstream (N-terminus) of all polypeptides.
[0738] The signal peptide transports the multisubunit peptide to golgi body or golgi apparatus. The cellular proteases act on the cleavage sites present in the cleavage peptides or the cleavage peptide undergoes self cleavage and cleaves the multisubunit peptide into individual polypeptides comprising either a target peptide, a linker peptide, and a selfassembling peptide, or a linker peptide, a target peptide, a linker peptide, and a selfassembling peptide, or a signal peptide, a target peptide, a linker peptide, and a selfassembling peptide, or a signal peptide, a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, or a combination thereof, wherein the target peptide is obtained or derived from:
[0739] (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or combination thereof of respiratory syncytial virus.
[0740] (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,
[0741] (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or
[0742] (d) a combination of (a), (b), and / or (c).
[0743] The polypeptides may additionally also have some residues (amino acids) of the cleavage peptide.
[0744] In some embodiments, the linker peptide is an amino acid linker, a zipper motif, a foldon, a scaffold, or a combination thereof. In some embodiments, the polypeptides are homologous polypeptides.
[0745] In some embodiments, two or more homologous polypeptides organise to form an oligomeric complex. In some embodiments, the oligomeric complex comprises at least two homologous polypeptides, at least three homologous polypeptides, at least four homologous polypeptides, at least five homologous polypeptides, or at least six homologous polypeptides and so on.
[0746] In some embodiments, the polypeptides are heterologous polypeptides.
[0747] In some embodiments, the homologous polypeptides or the heterologous polypeptides organize to form a polypeptide cluster.
[0748] A polypeptide nanoparticle is formed by self-assembly of two or more homologous polypeptides, two or more heterologous polypeptides, one or more oligomeric complexes, one or more polypeptide cluster, or their combination.
[0749] In some embodiments, a polypeptide nanoparticle comprises homologous polypeptides, heterologous polypeptides, oligomeric complexes, polypeptide clusters, or a combination thereof.
[0750] In some embodiments, the polypeptide nanoparticles are symmetrical, non- symmetrical, asymmetrical, or a combination thereof.
[0751] In some embodiments, the polypeptide nanoparticles are icosahedral, helical, spherical, rod-like, or a combination thereof.
[0752] In some embodiments, the polypeptide nanoparticles are enveloped or nonenveloped or a combination thereof.
[0753] In some embodiments, the polypeptide nanoparticles are single layered or multilayered or a combination thereof.
[0754] In some of the embodiments, the polypeptide nanoparticle comprises at least 2 or up to 500 polypeptides.
[0755] In some embodiments, the polypeptide nanoparticle comprises polypeptides between 2-5, 2-10, 2-20, 20-40, 40-60, 60-80, 80-100, 100-120, 120-140, 140-160, 160- 180, 180-200, 200-220, 220-240, 240-260, 260-280, 280-300, 300-320, 320-340, 340-360, 360-380, 380-400, 400-420, 420-440, 440-460, 460-480, or 480-500.
[0756] In some embodiments, the polypeptide nanoparticle comprises polypeptides between 2-5, 2-10, 10-20, 20-30, 30-40, 40-50, 50-60, 60-70, 70-80, 80-90, or 90-99.
[0757] In one of the embodiments, the polypeptide nanoparticle comprises at least 2 or up to 500 homologous polypeptides. In some embodiments, the polypeptide nanoparticle comprises at least 2 or up to 500 heterologous polypeptides.
[0758] In some embodiments, the polypeptide nanoparticle may comprise two or more oligomeric complexes such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0759] In some embodiments, the polypeptide nanoparticle comprises two or more polypeptide clusters such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0760] In some embodiments, the polypeptide nanoparticle comprises some homologous polypeptides and some heterologous polypeptides such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0761] In some embodiments, the polypeptide nanoparticle comprises some homologous polypeptides and some oligomeric complexes such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0762] In some embodiments, the polypeptide nanoparticle comprises some heterologous polypeptides and some oligomeric complexes such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0763] In some embodiments, the polypeptide nanoparticle comprises some homologous polypeptides and some polypeptide clusters such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0764] In some embodiments, the polypeptide nanoparticle comprises some heterologous polypeptides and some polypeptide clusters such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0765] In some embodiments, the polypeptide nanoparticle comprises some polypeptide clusters and some oligomeric complexes such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0766] In some embodiments, the polypeptide nanoparticle comprises some homologous polypeptides, some heterologous polypeptides, some oligomeric complexes, or their combination such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0767] In some embodiments, the polypeptide nanoparticle comprises some homologous polypeptides, some heterologous polypeptides, some polypeptide clusters, or their combination such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0768] In some embodiments, the polypeptide nanoparticle comprises some homologous polypeptides, some polypeptide clusters, some oligomeric complexes, or their combination such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0769] In some embodiments, the polypeptide nanoparticle comprises some heterologous polypeptides, some polypeptide clusters, some oligomeric complexes, or their combination such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0770] In some embodiments, the polypeptide nanoparticle comprises some heterologous polypeptides, some homologous polypeptides, some polypeptide clusters, some oligomeric complexes, or their combination such that the total number of polypeptides in the polypeptide nanoparticle are not more than 500.
[0771] Lipid nanoparticles (LNP) composition
[0772] The multisubunit nucleic acid sequences as described herein may be encapsulated or formulated in a lipid nanoparticle composition.
[0773] In some embodiments, the lipid nanoparticle composition comprises lipid components, ionizable polymer, or a combination thereof and a multisubunit nucleic acid sequence as described herein.
[0774] In some embodiments, the lipid nanoparticle composition comprises lipid components such as a cationic lipid, a phospholipid, a sterol, a PEG-lipid and a multisubunit nucleic acid sequence as described herein.
[0775] In another embodiment, the lipid nanoparticle composition comprises an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid and a multisubunit nucleic acid sequence as described herein.
[0776] In some embodiments, a vaccine comprising the lipid nanoparticle composition is provided herein.
[0777] Lipid Components
[0778] Lipid components of the lipid nanoparticle compositions may include one or more lipids, such as a cationic lipid, a phospholipid, a sterol, and a PEG-lipid. Cationic lipid
[0779] Cationic lipid refers to a lipid that has a net positive charge at a selected pH. Cationic lipids generally comprise a hydrophilic head group that carries the charge and a hydrophobic tail.
[0780] In some embodiments, the cationic lipid is a cationic lipid with an amine head group. The amine head group can be primary, secondary, tertiary or quaternary. The cationic lipid may comprise one (monoamine) or more (polyamine) such amine groups.
[0781] In some embodiments, the cationic lipids are positively charged at pH below the pKa of the cationic lipid. In certain embodiments, the cationic lipids are neutral i.e., when pH is same or above the pKa of the cationic lipid. In some embodiments, the cationic lipids are positively charged at acidic pH i.e., pH 1.0 to pH 6.9. In certain embodiments, the cationic lipids are neutral at certain pH i.e., around physiological pH (pH 7.0 to pH 7.5). A cationic lipid that can exist in a positively charged or neutral form depending on the pH is commonly referred to as ionizable lipid. In some embodiments, the cationic lipids are ionizable such that they can exist in a positively charged or neutral form depending on the pH. In some embodiments, the cationic lipids are positively charged irrespective of the pH.
[0782] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid disclosed in US provisional applications viz., 63 / 575930; 63 / 575934; 63 / 575938; 63 / 575939; and 63 / 575942.
[0783] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (I) formula (I) or isomer, or salt thereof, wherein:
[0784] Ri and R2 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, C5-10 heteroaryl containing 1-4 heteroatoms, -O-L7-R5, and -NReR?, or
[0785] Ri and R2 may combine together to form a saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms;
[0786] R3 and R4 are independently chosen from branched or unbranched C1-26 alkyl, C2-26 alkenyl, C2-26 alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26;
[0787] Rs, Re, and R7 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, and C5-10 heteroaryl containing 1-4 heteroatoms; each of Li, L2, L3, L4, L5, Le, and L7 is either absent or independently chosen from C1-10 alkylene, C2-10 alkenylene, and C2-10 alkynylene;
[0788] Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and - OP(O)(O-)O-;
[0789] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms; C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, S, or N provided that when Ai is -CH2-, -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, -C(S)O-, - OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, or S, one of Ri or R2is absent; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, or heteroaryl is independently optionally substituted with one or more substituent.
[0790] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (II) formula (II) or isomer, or salt thereof, wherein:
[0791] == represents either a single bond or a double bond;
[0792] Ri and R2 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, C5-10 heteroaryl containing 1-4 heteroatoms, -O-L7-R5, and -NReR?, or
[0793] Ri and R2 may combine together to form a saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms;
[0794] R3 and R4 are independently chosen from, branched or unbranched, C1-26 alkyl, C2-26 alkenyl, C2-26 alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26;
[0795] Rs, Re, and R7 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, and C5-10 heteroaryl containing 1-4 heteroatoms; each of Li, L2, L3, L4, L5, Le, and L7 is either absent or independently chosen from C1-10 alkylene, C2-10 alkenylene, and C2-10 alkynylene;
[0796] Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and - OP(O)(O-)O-;
[0797] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms;
[0798] — Hcf
[0799] Ai is \ , -CH2-, -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, - C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, S, or N provided that when Al is -CH2-, -C(0)0-, -0C(0)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, -C(S)O-, - OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, or S, one of Ri or R2is absent; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, or heteroaryl is independently optionally substituted with one or more substituent.
[0800] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (III) formula (III) or isomer, or salt thereof, wherein:
[0801] Ri and R2are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, C5-10 heteroaryl containing 1-4 heteroatoms, -O-L7-R5, and -NReR?, or
[0802] Ri and R2may combine together to form a saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms;
[0803] R3 and R4 are independently chosen from branched or unbranched C1-26 alkyl, C2.26 alkenyl, C2.26alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26;
[0804] Rs, Re, and R7 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, and C5-10 heteroaryl containing 1-4 heteroatoms; each of Li, L2, L3, L4, Ls Le, and L7 is either absent or independently chosen from C1-10 alkylene, C2-10 alkenylene, and C2-10 alkynylene; Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and - OP(O)(O-)O-;
[0805] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatom, or C5-10 heteroaryl containing 1-4 heteroatom; C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, S, or N provided that when Ai is -CH2-, -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, -C(S)O-, - OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, or S, one of Ri or R2is absent; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, or heteroaryl is independently optionally substituted with one or more substituent.
[0806] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (IV) formula (IV) or isomer, or salt thereof, wherein:
[0807] == represents either a single bond or a double bond;
[0808] Ri and R2 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, C5-10 heteroaryl containing 1-4 heteroatoms, -O-L7-R5, and -NReR?, or
[0809] Ri and R2 may combine together to form a saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms; R3and R4 are independently chosen from branched or unbranched C1-26 alkyl, C2-26 alkenyl, C2-26 alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26;
[0810] Rs, Re, and R7 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, and C5-10 heteroaryl containing 1-4 heteroatoms; each of Li, L2, L3, L4, Ls, Le, and L7 is either absent or independently chosen from C1-10 alkylene, C2-10 alkenylene, and C2-10 alkynylene;
[0811] Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and - OP(O)(O-)O-;
[0812] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatom, or C5-10 heteroaryl containing 1-4 heteroatom;
[0813] — Hc
[0814] Ai is \ , -CH2-, -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, - C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, S, or N provided that when Ai is -CH2-, -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, -C(S)O-, - OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, or S, one of Ri or R2is absent; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, or heteroaryl is independently optionally substituted with one or more substituent.
[0815] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (V) formula (V) or isomer, or salt thereof, wherein: R3and R4 are independently chosen from branched or unbranched C1-26 alkyl, C2-26 alkenyl, C2-26 alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26;
[0816] (Rio)qis chosen from H, -OH, optionally substituted C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, C5-10 heteroaryl containing 1-4 heteroatoms, -O-L7-R5, and -NReR?, wherein q is an integer ranging from 0 to 5;
[0817] Rs, Re, and R7 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, and C5-10 heteroaryl containing 1-4 heteroatoms; each of Li, L2, L3, L4, and L7 is either absent or independently chosen from C1-10 alkylene, C2-10 alkenylene, and C2-10 alkynylene;
[0818] Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and - OP(O)(O-)O-;
[0819] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, heteroaryl is independently optionally substituted with one or more substituent.
[0820] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (VI) formula (VI) or isomer, or salt thereof, wherein: R3and R4 are independently chosen from branched or unbranched C1-26 alkyl, C2-26 alkenyl, C2-26 alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26; each of Li, L2, L3, and L4 is either absent or independently chosen from C1-10 alkylene, C2- 10 alkenylene, and C2-io alkynylene;
[0821] Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and OP(O)(O-)O-;
[0822] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, heteroaryl is independently optionally substituted with one or more substituent.
[0823] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (VII) formula (VII) or isomer, or salt thereof, wherein:
[0824] == represents either a single bond or a double bond;
[0825] R3and R4 are independently chosen from, branched or unbranched, C1-26 alkyl, C2-26 alkenyl, C2-26 alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26;
[0826] (Rio)qis chosen from H, -OH, -CH2OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, C5-10 heteroaryl containing 1-4 heteroatoms, -O-Ls-Rs, and - NReR?, wherein q is an integer ranging from 0 to 5;
[0827] Rs, Re, and R7 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, and C5-10 heteroaryl containing 1-4 heteroatoms; each of Li, L2, L3, L4, Ls, L7, and Ls is either absent or independently chosen from C1-10 alkylene, C2-10 alkenylene, and C2-10 alkynylene;
[0828] Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and - OP(O)(O-)O-;
[0829] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms;
[0830] Z is -CH2-, O, N, or S;
[0831] Ai is -CH2-, -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, -C(S)O-, - OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, S, or N; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, or heteroaryl is independently optionally substituted with one or more substituent.
[0832] In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid represented by formula (VIII) formula (VIII) or isomer, or salt thereof, wherein:
[0833] R3 and R4 are independently chosen from branched or unbranched C1-26 alkyl, C2-26 alkenyl, C2-26 alkynyl, -(CH2)m-A-(CH2)n-(CH3)y, and -CH((CH2)m-A-(CH2)n-(CH3)y)2, wherein each of m, n, and y is independently an integer ranging from 0 to 26;
[0834] (Rio)qis chosen from H, -OH, -CH2OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, C5-10 heteroaryl containing 1-4 heteroatoms, -O-Ls-Rs, and - NReR?, wherein q is an integer ranging from 0 to 4; Rs, Re, and R7 are independently chosen from H, -OH, C1-10 alkyl, C2-10 alkenyl, C2-10 alkynyl, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, and C5-10 heteroaryl containing 1-4 heteroatoms; each of Li, L2, L3, L4, Ls, L7, and Ls is either absent or independently chosen from C1-10 alkylene, C2-10 alkenylene, or C2-10 alkynylene;
[0835] Xi and X2 are independently chosen from -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, - C(S)NH-, -NHC(S)-, -C(S)O-, -OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, and - OP(O)(O-)O-;
[0836] A is H, a bond, saturated or unsaturated C3-10 cycloalkyl, C6-10 aryl, saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, or C5-10 heteroaryl containing 1-4 heteroatoms;
[0837] Z is -CH2-, O, N, or S;
[0838] Ai is -CH2-, -C(O)O-, -OC(O)-, -C(O)NH-, -NHC(O)-, -C(S)NH-, -NHC(S)-, -C(S)O-, - OC(S)-, -OC(O)NH-, -NHC(O)O-, -OP(O)(OH)O-, S, or N; and wherein each alkyl, alkenyl, alkynyl, alkylene, alkenylene, alkynylene, cycloalkyl, aryl, heterocycloalkyl, or heteroaryl is independently optionally substituted with one or more substituent.
[0839] In some embodiments, the cationic lipid represented by formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), or formula (VIII) are substituted with substituents independently chosen from H, OH, Cl, Br, I, O, S, N, P, optionally substituted C1-6 alkoxy, optionally substituted C1-6 alkyl, optionally substituted C2-6 alkenyl, optionally substituted C2-6 alkynyl, optionally substituted saturated or unsaturated C3-10 cycloalkyl, optionally substituted C6-10 aryl, optionally substituted saturated or unsaturated C3-10 heterocycloalkyl containing 1-4 heteroatoms, optionally substituted C5-10 heteroaryl containing 1-4 heteroatoms or combination thereof. In some embodiments, substituent may further be substituted with H, -OH, Cl, Br, I, or Ci- 6 hydroxyalkyl.
[0840] In some embodiments, N:P ratio or cationic lipid to nucleic acid ratio in LNP formulation is between 1 to 18, between 1 to 17, between 1 to 16, between 1 to 15, between 1 to 14, between 1 to 13, between 1 to 12, between 1 to 11, between 1 to 10, between 1 to 9, between 1 to 8, between 1 to 7, between 1 to 6, between 1 to 5, between 1 to 4, between 1 to 3, between 1 to 2, or any range therein. In some embodiments, N:P ratio or cationic lipid to nucleic acid ratio in LNP formulation is about 18, about 17, about 16, about 15, about 14, about 13, about 12, about 11, about 10, about 9, about 8, about 7, about 6, about 5, about 4, about 3, about 2, about 1, or any portion or fraction thereof.
[0841] In some embodiments, other exemplary cationic lipid for use in the lipid nanoparticle compositions include, but are not limited to, N,N-dioleyl-N,N- dimethylammonium chloride (DODAC); N-(2,3-dioleyloxy)propyl)-N,N,N- trimethylammonium chloride (DOTMA); N,N-distearyl-N,N-dimethylammonium bromide(DDAB); N-(2,3dioleoyloxy)propyl)-N,N,N-trimethylammonium chloride (DOTAP); 3-(N — (N’,N’-dimethylaminoethane)-carbamoyl)cholesterol (DC-Chol), N-(l- (2,3-dioleoyloxy)propyl)N-2-(sperminecarboxamido)ethyl)-N,N- dimethylammoniumtrifluoracetate (DOSPA), dioctadecylamidoglycyl carboxyspermine (DOGS), 1,2-dioleoyl- 3 -dimethylammonium propane (DODAP), N,N-dimethyl-2,3-dioleoyloxy)propylamine (DODMA), N-(l ,2-dimyristyloxyprop-3-yl)-N,N-dimethyl-N-hydroxyethyl ammonium bromide (DMRIE), l,2-dilinoleyloxy-N,N-dimethylaminopropane (Dlin-DMA), 3- dimethylamino-2-(cholest-5-en-3-beta-oxybutan-4-oxy)-l-(cis,cis-9,12-oc- tadecadienoxy)propane (Clin-DMA), 2-[5’-(cholest-5-en-3-beta-oxy)-3’-oxapentoxy)-3- dimethyl-l-(cis,cis-9’,12’-octadecadienoxy)propane (CpLin-DMA), 2,3-Dilinoleoyloxy- N,N-dimethylpropylamine (Dlin-DAP), l,2-N,N’-Dilinoleylcarbamyl-3- dimethylaminopropane (Dlincarb-DAP), l,2-Dilinoleoylcarbamyl-3- dimethylaminopropane (Dlin-CD AP) , 2 ,2-dilinoleyl-4-dimethylaminomethyl- [1,3]- dioxolane (Dlin-K-DMA), heptatriaconta-6,9,28,31-tetraen-19-yl 4- (dimethylamino)butanoate (Dlin-MC3-DMA), heptadecane-9-yl 8-[2-hydroxyethyl-(6- oxo-6-undecoxyhexyl)amino]octanoate (SM-102), 6-[6-(2-hexyldecanoyloxy)hexyl-(4- hydroxybutyl)amino]hexyl 2-hexyldecanoate (ALC-0315), nonyl 8-[(8-heptadecan-9- yloxy-8-oxooctyl)-(2-hydroxyethyl)amino]octanoate (SLP-0001), or a combination thereof.
[0842] Methods of making cationic lipid and / or ionizable lipid or imparting the cationic lipid the ability to behave as an ionizable lipid are well known in the art (WO2005121348; W02009127060; W02009086558; W02010042877; W02010144740; WO2011075656; WO2017049245; WO2017075531; WO2018118102; WO2015199952; Reynier P. et al. Journal of Drug Targeting (2004) 12: 25-38; Sabnis, Staci et al. Molecular Therapy (2018) 26: 1509-1519). In some embodiments, the cationic lipid present in the lipid nanoparticle composition comprises a cationic lipid disclosed in published patent application viz., WO2019 / 152557; WO2019 / 232095; WO2021 / 077067; WO2019 / 089828;
[0843] US2019 / 0240354; US2010 / 0130588; US2021 / 0087135; US2021 / 0128488; US2020 / 0121809; US2013 / 0108685; US2013 / 0195920; US2015 / 0005363; US2014 / 0308304; US 2017 / 0210697; and US2013 / 0053572.
[0844] The proportion of cationic lipid present in the lipid nanoparticle compositions is from about 10 mol % to about 70 mol % or any range therein.
[0845] In some embodiments, the proportion of cationic lipid present in the lipid nanoparticle compositions is from about 10 mol % to about 70 mol %, from about 10 mol % to about 65 mol %, from about 10 mol % to about 60 mol %, from about 10 mol % to about 55 mol %, from about 10 mol % to about 50 mol %, or any range therein.
[0846] In some embodiments, the proportion of cationic lipid present in the lipid nanoparticle compositions is about 10 mol %, about 11 mol %, about 12 mol %, about 13 mol %, about 14 mol %, about 15 mol %, about 16 mol %, about 17 mol %, about 18 mol %, about 19 mol %, about 20 mol %, about 21 mol %, about 22 mol %, about 23 mol %, about 24 mol %, about 25 mol %, about 26 mol %, about 27 mol %, about 28 mol %, about 29 mol %, about 30 mol %, about 31 mol %, about 32 mol %, about 33 mol %, about 34 mol %, about 35 mol %, about 36 mol %, about 37 mol %, about 38 mol %, about 39 mol %, about 40 mol %, about 41 mol %, about 42 mol %, about 43 mol %, about 44 mol %, about 45 mol %, about 46 mol %, about 47 mol %, about 48 mol %, about 49 mol %, about 50 mol %, about 51 mol %, about 52 mol %, about 53 mol %, about 54 mol %, about 55 mol %, about 56 mol %, about 57 mol %, about 58 mol %, about 59 mol %, about 60 mol %, about 61 mol %, about 62 mol %, about 63 mol %, about 64 mol %, about 65 mol %, about 66 mol %, about 67 mol %, about 68 mol %, about 69 mol %, about 70 mol %, or any portion or fraction thereof. In some embodiments, the proportion of cationic lipid present in the lipid nanoparticle compositions is about 10 mol% to about 20 mol%, about 20 mol% to about 30 mol%, about 30 mol% to about 40 mol%, about 40 mol% to about 50 mol%, about 50 mol% to about 60 mol%, about 60 mol% to about 70 mol%, or any range therein. Phospholipids
[0847] Phospholipid includes a lipid containing a hydrophilic head with a phosphate group and a hydrophobic tail composed of fatty acid chains attached to a glycerol or sphingosine backbone.
[0848] Exemplary phospholipids for use in the lipid nanoparticle compositions include, but are not limited to, l,2-dilinoleoyl-sn-glycero-3-phosphocholine (DLPC), 1,2- dimyristoyl-sn-glycero-phosphocholine (DMPC), 1 ,2-dioleoyl-sn-glycero-3- phosphocholine (DOPC), l,2-dipalmitoyl-sn-glycero-3-phosphocholine (DPPC), 1,2- distearoyl-sn-glycero-3-phosphocholine (DSPC), 1 ,2-diundecanoyl-sn-glycero- phosphocholine (DUPC), l-palmitoyl-2-oleoyl-sn-glycero-3-phosphocholine (POPC), 1,2- di-O-octadecenyl-sn-glycero-3-phosphocholine (18:0 Diether PC), l-oleoyl-2- cholesterylhemisuccinoyl-sn-glycero-3-phosphocholine (OchemsPC), 1 -hexadecyl-sn- glycero-3-phosphocholine (C16 Lyso PC), l,2-dilinolenoyl-sn-glycero-3-phosphocholine, 1 ,2-diarachidonoyl-sn-glycero-3-phosphocholine, 1 ,2-didocosahexaenoyl-sn-glycero-3- phosphocholine, l,2-dioleoyl-sn-glycero-3-phosphoethanolamine (DOPE), 1,2- diphytanoyl-sn-glycero-3-phosphoethanolamine (ME 16.0 PE), 1,2-distearoyl-sn-glycero- 3-phosphoethanolamine, l,2-dilinoleoyl-sn-glycero-3-phosphoethanolamine, 1,2- dilinolenoyl-sn-glycero-3-phosphoethanolamine, l,2-diarachidonoyl-sn-glycero-3- phosphoethanolamine, 1 ,2-didocosahexaenoyl-sn-glycero-3-phosphoethanolamine, 1 ,2- dioleoyl-sn-glycero-3-phospho-rac-(l -glycerol) sodium salt (DOPG), l-myristoyl-2- stearoyl-sn-glycero-3-phosphocholine (MSPC), l-palmitoyl-2-myristoyl-sn-glycero-3- phosphocholine (PMPC), l-palmitoyl-2-stearoyl-sn-glycero-3-phosphocholine (PSPC), 1- stearoyl-2-myristoyl-sn-glycero-3-Phosphocholine (SMPC), l-Stearoyl-2-palmitoyl-sn- glycero-3-phosphocholine (SPPC), l-stearoyl-2-oleoyl-sn-glycero-3-phosphocholine (SOPC), l-stearoyl-2-docosahexaenoyl-sn-glycero-3-phosphocholine (SDPC), sphingomyelin, or a combination thereof.
[0849] The proportion of phospholipid present in the lipid nanoparticle compositions is from about 2 mol % to about 65 mol %, from about 5 mol % to about 65 mol %, from about 10 mol % to about 65 mol %, from about 10 mol % to about 55 mol %, from about 10 mol % to about 50 mol %, or any range therein.
[0850] In some embodiments, the proportion of phospholipid present in the lipid nanoparticle compositions is about 65 mol %, about 60 mol %, about 55 mol %, about 50 mol %, about 45 mol %, about 44 mol %, about 43 mol %, about 42 mol %, about 41 mol %, about 40 mol %, about 39 mol %, about 38 mol %, about 37 mol %, about 36 mol %, about 35 mol %, about 34 mol %, about 33 mol %, about 32 mol %, about 31 mol %, about 30 mol %, about 29 mol %, about 28 mol %, about 27 mol %, about 26 mol %, about 25 mol %, about 24 mol %, about 23 mol %, about 22 mol %, about 21 mol %, about 20 mol %, about 19 mol %, about 18 mol %, about 17 mol %, about 16 mol %, about 15 mol %, about 14 mol %, about 13 mol %, about 12 mol %, about 11 mol %, about 10 mol %, about 9 mol %, about 8 mol %, about 7 mol %, about 6 mol %, about 5 mol %, about 4 mol %, about 3 mol %, about 2 mol %, or any portion or fraction thereof.
[0851] Sterol
[0852] Lipid nanoparticle composition disclosed herein may include sterol and / or sterol derivatives. The term “sterol” as used herein include, but not limited to, cholesterol, sitosterol, fecosterol, ergosterol, campesterol, stigmasterol or their derivatives. In some embodiments, lipid nanoparticle composition comprises cholesterol and / or cholesterol derivatives. Non- limiting examples of cholesterol and cholesterol derivatives include 5 a- cholestanol, 5P-coprostanol, cholesteryl-(2’-hydroxy)-ethyl ether, cholesteryl- (4’- hydroxy)-butyl ether, 6-ketocholestanol, 5a-cholestane, cholestenone, 5a-cholestanone, 5P-cholestanone, cholesteryl decanoate, or mixtures thereof. Methods of making cholesterol and cholesterol derivatives are well known in the art.
[0853] The proportion of sterol present in the lipid nanoparticle compositions may be from about 20 mol % to about 65 mol % or any range therein.
[0854] In some embodiments, the proportion of sterol present in the lipid nanoparticle compositions is from about 20 mol % to about 65 mol %, from about 25 mol % to about 65 mol %, from about 30 mol % to about 65 mol %, from about 31 mol % to about 60 mol %, from about 32 mol % to about 60 mol %, from about 33 mol % to about 60 mol %, from about 34 mol % to about 60 mol %, from about 35 mol % to about 60 mol %, or any range therein.
[0855] In some embodiments, the proportion of sterol present in the lipid nanoparticle compositions is about 65 mol %, about 60 mol %, about 55 mol %, about 50 mol %, about 45 mol %, about 44 mol %, about 43 mol %, about 42 mol %, about 41 mol %, about 40 mol %, about 39 mol %, about 38 mol %, about 37 mol %, about 36 mol %, about 35 mol %, about 34 mol %, about 33 mol %, about 32 mol %, about 31 mol %, about 30 mol %, about 29 mol %, about 28 mol %, about 27 mol %, about 26 mol %, about 25 mol %, about 24 mol %, about 23 mol %, about 22 mol %, about 21 mol %, about 20 mol %, or any portion or fraction thereof.
[0856] PEG-lipid
[0857] The term PEG-lipid, pegylated lipid, PEG linked lipid, PEG conjugated lipid, PEG- lipid conjugate, PEG modified lipid have been used interchangeably to mean polyethylene glycol linked to a lipid moiety. The lipid moiety may be linked directly to the PEG molecule or through a linker. In some embodiments, a PEG-lipid comprises a PEG- modified phosphatidylethanolamines, PEG-modified phosphatidic acids, PEG-modified ceramides, PEG-modified dialkylamines, PEG-modified diacylglycerols, PEG-modified dialkylglycerols, and / or PEG-modified cholesterol, and / or mixtures thereof. The methods of making PEG-lipid are well known to persons skilled in the art.
[0858] In some embodiments, PEG-lipid is selected from mPEG-Dimyristoyl glycerol (mPEG-DMG), mPEG-N,N-Ditetradecylacetamide (mPEG-DTA or ALC0159), mPEG- Cholesterol (mPEG-CLS), mPEG-DSPE, mPEG-DMPE, mPEG-DPPE, mPEG-DLPE, mPEG-DOPE, mPEG-DPPC, mPEG-DSPC, l,2-Distearoyl-sn-Glycero-3- Phosphoethanolamine with conjugated methoxyl poly(ethylene glycol) (mPEG-DSPE), l,2-dimyristoyl-rac-glycero-3-methoxypolyethylene glycol-2000 (mPEG2000-DMG), a- (3 ’ - { [ 1 ,2-di(myristyloxy)propanoxy]carbonylamino }propyl)-co-methoxy, polyoxyethylene (mPEG2000C-DMG), or mixtures thereof.
[0859] The PEG moiety of the PEG-lipid may comprise an average molecular weight ranging from 0.5 kDa to 10 kDa. In some embodiments, the PEG-lipid has an average molecular weight of about 0.5 kDa to 5 kDa, about 0.5 kDa to 4 kDa, 0.5 kDa to 3 kDa, 0.5 kDa to 2 kDa. In preferred embodiments, the PEG-lipid has an average molecular weight of about 0.5 kDa to about 2 kDa.
[0860] The proportion of PEG-lipid present in the lipid nanoparticle compositions may be from about 0.2 mol % to about 2.0 mol % or any range therein.
[0861] In some embodiments, the proportion of PEG-lipid present in the lipid nanoparticle compositions is from about 0.2 mol % to about 2.0 mol %, from about 0.2 mol % to about 1.9 mol %, from about 0.2 mol % to about 1.8 mol %, from about 0.2 mol % to about 1.7 mol %, from about 0.2 mol % to about 1.6 mol %, from about 0.2 mol % to about 1.5 mol %, or any range therein. In some embodiments, the proportion of PEG-lipid present in the lipid nanoparticle compositions is about 0.2 mol %, about 0.3 mol %, about 0. 4 mol %, about 0.5 mol %, about 0.6 mol %, about 0.7 mol %, about 0.8 mol %, about 0.9 mol %, about 1.0 mol %, about 1.1 mol %, about 1.2 mol %, about 1.3 mol %, about 1.4 mol %, about 1.5 mol %, about 1.6 mol %, about 1.7 mol %, about 1.8 mol %, about 1.9 mol %, about 2.0 mol %, or any portion or fraction thereof.
[0862] In some embodiments, the lipid nanoparticle composition additionally contains an ionizable polymer.
[0863] Ionizable polymer
[0864] As used herein the term “polymer” means a compound formed from a plurality of repeating units called monomers. Polymers are produced through a process called polymerization wherein two or more monomers are linked through chemical bonds to form the polymer. In some embodiments, the polymer is branched or unbranched. In some embodiments, the polymer is homopolymer, i.e., comprising same type of repeat units or monomers, or heteropolymer, i.e., comprising more than one type of repeat units or monomers. The terms heteropolymer and copolymer have been used interchangeably herein.
[0865] The term “Ionizable polymer” as used herein means, a polymer that can exist in a positively charged or neutral form depending on the pH of the solution or environment, for example, ionizable polymer will be cationic (positively charged) when pH of the solution is below the pKa of the ionizable polymer and neutral (no charge) when pH of the solution is same or above the pKa of the ionizable polymer. In some embodiments, ionizable polymer is positively charge in acidic pH i.e., pH 1.0 to pH 6.9. In some embodiments, ionizable polymer is neutral (no charge) around physiological pH (pH 7.0 to pH 7.5).
[0866] In some embodiments, the ionizable polymer is a biocompatible polymer or biodegradable polymer. The term “biocompatible polymer” and “biodegradable polymer” have been used interchangeably to mean a polymer that is substantially free from any deleterious effects when introduced into a living or biological system. Such polymers are capable of undergoing degradation when introduced into the living or biological systems and are not expected to produce significant toxicity or immunological response.
[0867] In some embodiments, the lipid nanoparticle compositions comprise an ionizable polymer. The ionizable polymer may be selected from a chitosan, chitosan derivatives, cellulose derivatives, a poly-L-lysine (PLL), a protamine, a polyethyleneimine, their derivatives, or a combination thereof.
[0868] In some embodiments, the ionizable polymer is positively charged at acidic pH i.e., pH 1.0 to 6.9 and is neutral around physiological pH (pH 7.0 to 7.5).
[0869] The proportion of ionizable polymer present in the lipid nanoparticle compositions may be from about 1 mol % to about 25 mol %.
[0870] In some embodiments, the proportion of ionizable polymer present in the lipid nanoparticle compositions is from about 1 mol % to about 25 mol %, from about 1 mol % to about 24 mol %, from about 1 mol % to about 23 mol %, from about 1 mol % to about 22 mol %, from about 1 mol % to about 21 mol %, from about 1 mol % to about 20 mol %, from about 1 mol % to about 19 mol %, from about 1 mol % to about 18 mol %, from about 1 mol % to about 17 mol %, from about 1 mol % to about 16 mol %, from about 1 mol % to about 15 mol %, or any range therein.
[0871] In some embodiments, the proportion of ionizable polymer present in the lipid nanoparticle compositions is about 1 mol %, about 2 mol %, about 3 mol %, about 4 mol %, about 5 mol %, about 6 mol %, about 7 mol %, about 8 mol %, about 9 mol %, about 10 mol %, about 11 mol %, about 12 mol %, about 13 mol %, about 14 mol %, about 15 mol %, about 16 mol %, about 17 mol %, about 18 mol %, about 19 mol %, about 20 mol %, about 21 mol %, about 22 mol %, about 23 mol %, about 24 mol %, about 25 mol %, or any portion or fraction thereof.
[0872] In some embodiments, the preferred ionizable polymer comprises a chitosan, chitosan derivatives, cellulose derivatives, a poly-L-lysine (PLL), a protamine, a polyethyleneimine, and / or their derivatives, or a combination thereof.
[0873] Method of treatment
[0874] In some aspects, provided herein is a method of beating or preventing a disease, comprising administering to a subject in need thereof the multisubunit nucleic acid sequence as described herein.
[0875] In some aspects, provided herein is a method of heating or preventing a disease, comprising administering to a subject in need thereof the lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid and the multisubunit nucleic acid sequence as described herein. In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof the lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid and the multisubunit nucleic acid sequence as described herein, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0876] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof the lipid nanoparticle composition comprising an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid and the multisubunit nucleic acid sequence as described herein.
[0877] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof the lipid nanoparticle composition comprising an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid and the multisubunit nucleic acid sequence as described herein, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0878] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a vaccine comprising the multisubunit nucleic acid as described herein.
[0879] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a vaccine comprising the lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid as described herein.
[0880] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a vaccine comprising the lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid as described herein, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0881] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a vaccine comprising the lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid described herein.
[0882] In some aspects, provided herein is a method of treating or preventing a disease, comprising administering to a subject in need thereof a vaccine comprising the lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid described herein, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0883] In some aspects, the disclosure relates to use of the multisubunit nucleic acid as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0884] In some aspects, the disclosure relates to use of a lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid sequence as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0885] In some aspects, the disclosure relates to use of a lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid sequence as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0886] In some aspects, the disclosure relates to use of a lipid nanoparticle composition comprising an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid sequence as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0887] In some aspects, the disclosure relates to use of a lipid nanoparticle composition comprising an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid sequence as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof. In some embodiments, the disclosure relates to use of a vaccine comprising the multisubunit nucleic acid as described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0888] In some embodiments, the disclosure relates to use of a vaccine comprising a lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0889] In some embodiments, the disclosure relates to use of a vaccine comprising a lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0890] In some embodiments, the disclosure relates to use of a vaccine comprising a lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
[0891] In some embodiments, the disclosure relates to use of a vaccine comprising a lipid nanoparticle composition, wherein the lipid nanoparticle composition comprises an ionizable polymer, a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid described herein, in the manufacture of a medicament for the treatment or prevention of a disease in a subject, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or a combination thereof.
[0892] In some embodiments, the diseases are RSV, HMPV, HPIV diseases.
[0893] In some embodiments, the multisubunit nucleic acid sequence is present in biologically effective amount or therapeutically effective amount. In some embodiments, the biologically effective amount of the multisubunit nucleic acid sequence is between 0.1 pg to 2000 pg, 0.1 pg to 1800 pg, 0.1 pg to 1600 pg, 0.1 pg to 1400 pg, 0.1 pg to 1200 pg, 0.1 pg to 1000 pg, 0.1 pg to 950 pg, 0.1 pg to 900 pg, 0.1 pg to 850 pg, 0.1 pg to 800 pg, o.i ng to 750 pg, o.i ng to 700 ng, o.i ng to 650 ng, o.i ng to 600 ng, o.i ng to 550 ng, o.i ng to 500 ng, o.i ng to 450 ng, o.i ng to 400 ng, o.i ng to 350 ng, o.i ng to 300 ng, o.i ng to 250 ng, o.i to 200 ng, 0.1 to 175 ng, 0.1 to 150 ng, 0.1 to 125 ng, 0.1 to 100 ng, 0.1 ng to 90 ng, 0.1 ng to so ng, 0.1 ng to 70 ng, 0.1 ng to 60 ng, 0.1 ng to 50 ng, 0.1 ng to 40 ng, 0.1 ng to 30 ng, O.I ng to 20 ng, O.I ng to 10 ng, O.I ng to 5 ng, or any range therein.
[0894] In some embodiments, the biologically effective amount of the multisubunit nucleic acid sequence is from about 0.1 ng to 1000 ng, 0.1 ng to 950 ng, 0.1 ng to 900 ng, o.i ng to 850 ng, o.i ng to soo ng, o.i ng to 750 ng, 0.1 ng to 700 ng, 0.1 ng to 650 ng, 0.1 ng to 600 ng, 0.1 ng to 550 ng, O.I ng to 500 ng, or any range therein.
[0895] In some embodiments, the biologically effective amount of the multisubunit nucleic acid sequence is 0.1 ng, 0.2 ng, 0.3 ng, 0.4 ng, 0.5 ng, 0.6 ng, 0.7 ng, 0.8 ng, 0.9 ng, 1 ng,2Mg, 3 ng,4Mg, 5 Mg, 6 Mg,7Mg, 8 Mg, 9 Mg, 10 Mg, 15 Mg, 20 Mg, 25 Mg, 30 ng, 35 ng, 40 ng, 45 ng, 50 ng, 55 ng, 60 ng, 65 ng, 70 ng, 75 ng, 80 ng, 85 ng, 90 ng, 95 ng, IOO ng, no ng, 120 ng, 130 ng, 140 ng, 150 ng, 160 ng, no ng, 180 ng, 190 ng, 200 ng, 220 ng, 240 ng, 260 ng, 280 ng, 300 ng, 350 ng, 400 ng, 450 ng, 500 ng, 600 ng, 700 ng, soo ng, 900 ng, 1000 ng, 1100 ng, 1200 ng, BOO ng, 1400 ng, 1500 ng, 1600 ng, noo ng, 1800 ng, 1900 ng, 2000 ng, or any portion or fraction thereof.
[0896] In one aspect, provided herein is a nucleic acid comprising a plurality of polynucleotide sequences, wherein some or all polynucleotide sequences of the plurality comprises either a target sequence, a linker sequence, and a self- assembling sequence or a linker sequence, a target sequence, a linker sequence and a self-assembling sequence or a combination thereof, wherein each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality, wherein the target sequence is obtained or derived from (a) envelope protein, matrix protein, nucleocapsid protein, M2- 1 protein, M2- 2 protein, non- structural protein, B cell epitope, T cell epitope, or a combination thereof of respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2- 1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumo virus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or (d) a combination thereof. In another aspect, provided herein is a multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a target sequence, a linker sequence, and a self- assembling sequence or a second linker sequence, a second target sequence, a third linker sequence, and a second self-assembling sequence, wherein each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality, wherein the target sequence or the second target sequence is obtained or derived from (a) envelope protein, matrix protein, nucleocapsid protein, M2- 1 protein, M2- 2 protein, non- structural protein, B cell epitope, T cell epitope, or a combination thereof of respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2- 1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumo virus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or (d) a combination thereof. In another aspect, provided herein a multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a target sequence, a linker sequence, and a self-assembling sequence, wherein each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality, wherein the target sequence is obtained or derived from (a) envelope protein, matrix protein, nucleocapsid protein, M2- 1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or (d) a combination thereof. In yet another aspect, provided herein is a multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, wherein each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality, wherein the target sequence is obtained or derived from (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non- structural protein, B cell epitope, T cell epitope, or a combination thereof of respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumo virus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or (d) a combination thereof. In one aspect, provided herein is a multisubunit nucleic acid comprising a first plurality of polynucleotide sequences and a second plurality of polynucleotide sequences, each polynucleotide sequence of the first plurality comprises a first target sequence, a first linker sequence, and a first self-assembling sequence, wherein each polynucleotide sequence of the second plurality comprises a second linker sequence, a second target sequence, a third linker sequence, and a second self-assembling sequence, wherein each polynucleotide sequence of the first plurality and each polynucleotide sequence of the second plurality is connected to an adjacent polynucleotide sequence of the first plurality or an adjacent polynucleotide sequence of the second plurality by a cleavage sequence, and wherein the first polynucleotide sequence in the multisubunit nucleic acid is a polynucleotide sequence of the first plurality or a polynucleotide sequence of the second plurality, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of the first polynucleotide sequence of the first plurality or the second polynucleotide sequence of the second plurality, or a combination thereof, wherein the first target sequence and the second target sequence is obtained or derived from (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non- structural protein, B cell epitope, T cell epitope, or a combination thereof of respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumo virus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or (d) a combination thereof. In another aspect, provided herein is a multisubunit nucleic acid encoding a plurality of polypeptides, wherein some or all polypeptides of the plurality comprises either a target peptide, a linker peptide, and a self-assembling peptide or a linker peptide, a target peptide, a linker peptide and a self-assembling peptide or a combination thereof, wherein each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide, and wherein the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality, wherein the target peptide is obtained or derived from (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non- structural protein, B cell epitope, T cell epitope, or a combination thereof of respiratory syncytial virus, (b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumo virus, (c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or (d) a combination thereof. In another aspect, provided herein is a multisubunit nucleic acid encoding a plurality of polypeptides, wherein some or all polypeptides of the plurality comprises either a target peptide, a linker peptide, and a self-assembling peptide or a second linker peptide, a second target peptide, a third linker peptide and a second selfassembling peptide or combination thereof, wherein each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide, and wherein the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of the first polypeptide of the plurality or the second polypeptide of the plurality or a combination thereof, wherein the target peptide and the second target peptide is obtained or derived from (a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of respiratory syncytial virus, (b) envelope protein, m...
Claims
What is claimed:
1. A multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein some or all polynucleotide sequences of the plurality comprises either a target sequence, a linker sequence, and a self-assembling sequence or a linker sequence, a target sequence, a linker sequence and a self-assembling sequence or a combination thereof, wherein each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality, wherein the target sequence is obtained or derived from:(a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,(b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,(c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
2. A multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a target sequence, a linker sequence, and a self-assembling sequence, wherein each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality, wherein the target sequence is obtained or derived from:(a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,(b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,(c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
3. A multisubunit nucleic acid comprising a plurality of polynucleotide sequences, wherein each polynucleotide sequence of the plurality comprises a linker sequence, a target sequence, a linker sequence, and a self-assembling sequence, wherein each polynucleotide sequence of the plurality is connected to an adjacent polynucleotide sequence of the plurality by a cleavage sequence, and wherein the multisubunit nucleic acid further comprises a signal sequence upstream of one or more of the polynucleotide sequences of the plurality, wherein the target sequence is obtained or derived from:(a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,(b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,(c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
4. The multisubunit nucleic acid according to any one of the preceding claims, wherein the target sequence, the linker sequence, and the self-assembling sequence or the linker sequence, the target sequence, the linker sequence, and the self-assembling sequence are in 5’ to 3’ order.
5. A multisubunit nucleic acid encoding a plurality of polypeptides, wherein some or all polypeptides of the plurality comprises either a target peptide, a linker peptide, and aself-assembling peptide or a linker peptide, a target peptide, a linker peptide and a selfassembling peptide or a combination thereof, wherein each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide, and wherein the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality, wherein the target peptide is obtained or derived from:(a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,(b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,(c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
6. A multisubunit nucleic acid encoding a plurality of polypeptides, wherein each polypeptide of the plurality comprises a target peptide, a linker peptide, and a selfassembling peptide, wherein each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide, and wherein the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality, wherein the target peptide is obtained or derived from:(a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,(b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,(c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
7. A multisubunit nucleic acid encoding a plurality of polypeptides, wherein each polypeptide of the plurality comprises a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide, wherein each polypeptide of the plurality is connected to an adjacent polypeptide of the plurality by a cleavage peptide, and wherein the multisubunit nucleic acid further encodes a signal peptide on the amino-terminus of one or more of the polypeptides of the plurality, wherein the target peptide is obtained or derived from:(a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination thereof of a respiratory syncytial virus,(b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,(c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
8. The multisubunit nucleic acid according to any one of the claims 5-7, wherein the target peptide, the linker peptide, and the self-assembling peptide or the linker peptide, the target peptide, the linker peptide, and the self-assembling peptide are in N-terminus to C- terminus order.
9. The multisubunit nucleic acid according to any one of the preceding claims, wherein total number of the polynucleotide sequences or the polypeptides are not more than 100.
10. The multisubunit nucleic acid according to claim 9, wherein total number of the polynucleotide sequences or the polypeptides are between 2-5, 2-10, 10-20, 20-30, 30-40, 40-50, 50-60, 60-70, 70-80, 80-90, or 90-99.
11. The multisubunit nucleic acid according to any one of the preceding claims, wherein the multisubunit nucleic acid is a DNA or an RNA.
12. The multisubunit nucleic acid according to claim 11, wherein the RNA is an mRNA.
13. The multisubunit nucleic acid according to claim 12, wherein the mRNA is 1 to 20 kb, 1 to 18 kb, 1 to 16 kb, 1 to 14 kb, 1 to 12 kb, 1 to 10 kb, 1 to 9 kb, 1 to 8 kb, 1 to 7 kb, 1 to 6 kb, 1 to 5 kb in length, or any range therein.
14. The multisubunit nucleic acid according to any one of the claims 12-13, wherein the mRNA is obtained through a single IVT process or step.
15. The multisubunit nucleic acid according to any one of the claims 1-4 and 9-14, wherein the linker sequence encodes a linker peptide.
16. The multisubunit nucleic acid according to any one of the claims 5-15, wherein the linker peptide is an amino acid linker, a zipper motif, a foldon, a scaffold, or a combination thereof.
17. The multisubunit nucleic acid according to claim 16, wherein the linker peptide is the amino acid linker.
18. The multisubunit nucleic acid according to claim 17, wherein the amino acid linker comprises 2 to 49 amino acids.
19. The multisubunit nucleic acid according to any one of the claims 17-18, wherein the amino acid linker is a glycine serine linker, a glycine proline linker, a glycine threonine linker, an alanine serine linker, any combination of two amino acids, or a combination thereof.
20. The multisubunit nucleic acid according to claim 16, wherein the linker peptide is the zipper motif.
21. The multisubunit nucleic acid according to claim 16, wherein the linker peptide is the foldon.
22. The multisubunit nucleic acid according to claim 16, wherein the linker peptide is the scaffold.
23. The multisubunit nucleic acid according to any one of the claims 16-19, wherein the linker peptide comprises the amino acid linker and the zipper motif.
24. The multisubunit nucleic acid according to any one of the claims 16-19, wherein the linker peptide comprises the amino acid linker and the foldon.
25. The multisubunit nucleic acid according to any one of the claims 16-19, wherein the linker peptide comprises the amino acid linker and the scaffold.
26. The multisubunit nucleic acid according to any one of the claims 16-19, wherein the linker peptide comprises the amino acid linker, the zipper motif, and the scaffold.
27. The multisubunit nucleic acid according to any one of the claims 16-19, wherein the linker peptide comprises the amino acid linker, the foldon, and the scaffold.
28. The multisubunit nucleic acid according to claim 16, wherein the linker peptide comprises the foldon and the scaffold.
29. The multisubunit nucleic acid according to claim 16, wherein the linker peptide comprises the zipper motif and the scaffold.
30. The multisubunit nucleic acid according to any one of the claims 5-29, wherein the linker peptide has an amino acid sequence of any one of SEQ ID NOs: 12-51 or 12221- 12225.
31. The multisubunit nucleic acid according to any one of the claims 1-4 and 9-30, wherein the self-assembling sequence encodes a self-assembling peptide.
32. The multisubunit nucleic acid according to any one of the claims 5-31, wherein the self-assembling peptide is a lumazine synthase, an MS2 coat protein, , a hepatitis B surface antigen (HBsAg) from Hepatitis B Virus, a hepatitis B core antigen (HbcAg) from Hepatitis B virus, a human papillomavirus LI (HPV LI) protein, a matrix protein Ml from influenza A virus, a ferritin, a riboflavin synthase, a dihydrolipoyl acetyltransferase (E2p), or a combination thereof, including their codon optimized nucleic acid sequences, fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
33. The multisubunit nucleic acid according to any one of the claims 31-32, wherein the self-assembling peptide is a ferritin comprising a ferritin subunit or a ferritin peptide, a dihydrolipoyl acetyltransferase (E2p), a lumazine synthase, an MS2 coat protein, or a combination thereof.
34. The multisubunit nucleic acid according to claim 33, wherein the ferritin peptide is obtained or derived from Helicobacter pylori ferritin or Listeria innocua ferritin, the lumazine synthase is obtained or derived from Aquifex aeolicus or Bacillus subtilis, the MS2 coat protein is obtained or derived from Emesvirus zinderi, and the dihydrolipoyl acetyltransferase (E2p) obtained or derived from Bacillus stearothermophilus.
35. The multisubunit nucleic acid according to any one of the claims 5-34, wherein the self-assembling peptide has an amino acid sequence of any one of SEQ ID NOs: 1-11 or 12217-12220.
36. The multisubunit nucleic acid according to any one of the claims 1-4 and 9-35, wherein the cleavage sequence encodes one or more cleavage peptides.
37. The multisubunit nucleic acid according to any one of the claims 5-36, wherein the one or more cleavage peptides are optionally connected to each other by a linker peptide.
38. The multisubunit nucleic acid according to claim 37, wherein the cleavage peptide is a golgi specific cleavage peptide, a self cleaving peptide, or a combination thereof.
39. The multisubunit nucleic acid according to any one of the claims 5-38, wherein the cleavage peptide has an amino acid sequence of any one of SEQ ID NOs: 52-66.
40. The multisubunit nucleic acid according to any one of the claims 1-4 and 9-39, wherein the signal sequence encodes a signal peptide.
41. The multisubunit nucleic acid according to any one of the claims 5-40, wherein the signal peptide is present on the amino-terminus of one or more of the polypeptides of the plurality.
42. The multisubunit nucleic acid according to claim 41, wherein the multisubunit nucleic acid further encodes a second signal peptide on the amino-terminus of all or some polypeptides of the plurality.
43. The multisubunit nucleic acid according to any one of the claims 5-42, wherein the signal peptide has an amino acid sequence of any one of SEQ ID NOs: 67-86.
44. The multisubunit nucleic acid according to any one of the claims 1-4 and 9-43, wherein the target sequence encodes a target peptide.
45. The multisubunit nucleic acid according to any one of the claims 5-44, wherein the target peptide is encoded by a codon optimized nucleic acid sequence, or fragments, mutants, variants, comparable equivalents, or functional analogs thereof.
46. The multisubunit nucleic acid according to claim 45, wherein the target peptide is obtained or derived from:(a) envelope protein of a respiratory syncytial virus,(b) envelope protein of a metapneumovirus,(c) envelope protein of a human parainfluenza virus, or(d) a combination thereof.
47. The multisubunit nucleic acid according to claim 46, wherein(a) the envelope protein is selected from the group comprising glycoprotein G, glycoprotein F, small hydrophobic SH protein, or a combination thereof of a respiratory syncytial virus,(b) the envelope protein is selected from the group comprising glycoprotein G, glycoprotein F, small hydrophobic SH protein, or a combination thereof of a metapneumovirus,(c) the envelope protein is selected from the group comprising glycoprotein F, hemagglutinin neuraminidase, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
48. The multisubunit nucleic acid according to claim 47, wherein(a) the envelope protein is selected from the group comprising glycoprotein G having one of SEQ ID NOs: 87-2734, glycoprotein F having one of SEQ ID NOs: 2735-4559, small hydrophobic SH protein having one of SEQ ID NOs: 4560-4794, or a combination thereof of a respiratory syncytial virus,(b) the envelope protein is selected from the group comprising glycoprotein G having one of SEQ ID NOs: 7372-7681, glycoprotein F having one of SEQ ID NOs: 7682-7793, small hydrophobic SH protein having one of SEQ ID NOs: 7794- 7935, or a combination thereof of a metapneumovirus,(c) the envelope protein is selected from the group comprising glycoprotein F having one of SEQ ID NOs: 8628-9088, hemagglutinin neuraminidase having one of SEQ ID NOs: 9089-10015, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
49. The multisubunit nucleic acid according to claim 45, wherein the target peptide is obtained or derived from:(a) matrix protein of a respiratory syncytial virus,(b) matrix protein of a metapneumovirus,(c) matrix protein of a human parainfluenza virus, or(d) a combination thereof.
50. The multisubunit nucleic acid according to claim 49, wherein the target peptide is obtained or derived from:(a) the matrix protein having one of SEQ ID NOs: 4795-4978 of a respiratory syncytial virus,(b) the matrix protein having one of SEQ ID NOs: 7936-7976 of a metapneumovirus,(c) the matrix protein having one of SEQ ID NOs: 10016-10214 of a human parainfluenza virus, or(d) a combination thereof.
51. The multisubunit nucleic acid according to claim 45, wherein the target peptide is obtained or derived from:(a) nucleocapsid protein of a respiratory syncytial virus,(b) nucleocapsid protein of a metapneumovirus,(c) nucleocapsid protein of a human parainfluenza virus, or(d) a combination thereof.
52. The multisubunit nucleic acid according to claim 51, wherein(a) the nucleocapsid protein is selected from the group comprising nucleoprotein N, phosphoprotein P, large polymerase protein L, or a combination thereof of a respiratory syncytial virus,(b) the nucleocapsid protein is selected from the group comprising nucleoprotein N, phosphoprotein P, large polymerase protein L, or a combination thereof of a metapneumovirus,(c) the nucleocapsid protein is selected from the group comprising nucleoprotein N, phosphoprotein P, large polymerase protein L, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
53. The multisubunit nucleic acid according to claim 52, wherein(a) the nucleocapsid protein is selected from the group comprising nucleoprotein N having one of SEQ ID NOs: 4979-5255, phosphoprotein P havingone of SEQ ID NOs: 5256-5668, large polymerase protein L having one of SEQ ID NOs: 5669-6132, or a combination thereof of a respiratory syncytial virus,(b) the nucleocapsid protein is selected from the group comprising nucleoprotein N having one of SEQ ID NOs: 7977-8042, phosphoprotein P having one of SEQ ID NOs: 8043-8211, large polymerase protein L having one of SEQ ID NOs: 8212-8483, or a combination thereof of a metapneumovirus,(c) the nucleocapsid protein is selected from the group comprising nucleoprotein N having one of SEQ ID NOs: 10215-10468, phosphoprotein P having one of SEQ ID NOs: 10469-11128, large polymerase protein L having one of SEQ ID NOs: 11129-11485, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
54. The multisubunit nucleic acid according to claim 45, wherein the target peptide is obtained or derived from:(a) M2-1 protein, M2-2 protein, non-structural protein, or a combination thereof of a respiratory syncytial virus,(b) M2-1 protein, M2-2 protein, or a combination thereof of a metapneumovirus,(c) accessory protein of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
55. The multisubunit nucleic acid according to claim 54, wherein the target peptide is obtained or derived from:(a) the M2-1 protein having one of SEQ ID NOs: 6133-6381, M2-2 protein having one of SEQ ID NOs: 6382-6756, or a combination thereof of a respiratory syncytial virus,(b) the M2-1 protein having one of SEQ ID NOs: 8484-8545, M2-2 protein having one of SEQ ID NOs: 8546-8586, or a combination thereof of a metapneumovirus,(c) the accessory protein having one of SEQ ID NOs: 11486-12188 of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
56. The multisubunit nucleic acid according to claim 45, wherein the target peptide is obtained or derived from:(a) B cell epitope of a respiratory syncytial virus.(b) B cell epitope of a metapneumovirus,(c) B cell epitope of a human parainfluenza virus, or(d) a combination thereof.
57. The multisubunit nucleic acid according to claim 56, wherein the target peptide is obtained or derived from:(a) the B cell epitope having one of SEQ ID NOs: 6757-7019 of a respiratory syncytial virus.(b) the B cell epitope having one of SEQ ID NOs: 12189-12193 of a human parainfluenza virus, or(c) a combination thereof.
58. The multisubunit nucleic acid according to claim 45, wherein the target peptide is obtained or derived from:(a) T cell epitope of a respiratory syncytial virus.(b) T cell epitope of a metapneumovirus,(c) T cell epitope of a human parainfluenza virus, or(d) a combination thereof.
59. The multisubunit nucleic acid according to claim 58, wherein the target peptide is obtained or derived from:(a) the T cell epitope having one of SEQ ID NOs: 7020-7371 of a respiratory syncytial virus.(b) the T cell epitope having one of SEQ ID NOs: 8587-8627 of a metapneumovirus,(c) the T cell epitope having one of SEQ ID NOs: 12194-12216 of a human parainfluenza virus, or(d) a combination thereof.
60. The multisubunit nucleic acid according to any one of the claims 5-59, wherein the target peptide has an amino acid sequence of any one of SEQ ID NOs: 87-12216.
61. A lipid nanoparticle composition comprising a cationic lipid, a phospholipid, a sterol, a PEG-lipid, and the multisubunit nucleic acid according to any one of the preceding claims.
62. The lipid nanoparticle composition according to claim 61, wherein the cationic lipid is present in an amount from 10 mol percent to 70 mol percent.
63. The lipid nanoparticle composition according to claim 61, wherein the phospholipid is present in an amount from 2 mol percent to 65 mol percent.
64. The lipid nanoparticle composition according to claim 61, wherein the sterol is present in an amount from 20 mol percent to 65 mol percent.
65. The lipid nanoparticle composition according to claim 61, wherein the PEG-lipid is present in an amount from 0.2 mol percent to 2.0 mol percent.
66. The lipid nanoparticle composition according to claim 61, additionally comprising an ionizable polymer.
67. The lipid nanoparticle composition according to claim 66, wherein the ionizable polymer is present in an amount from 1 mol percent to 25 mol percent.
68. The lipid nanoparticle composition according to any one of the claims 61-67, wherein the cationic lipid is represented by any one of formula (I), formula (II), formula (III), formula (IV), formula (V), formula (VI), formula (VII), formula (VIII), or SM-102, or ALC-0315, or a combination thereof.
69. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (I).
70. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (II).
71. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (III).
72. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (IV).
73. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (V).
74. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (VI).
75. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (VII).
76. The lipid nanoparticle composition according to any one of the claims 61-68, wherein the cationic lipid is represented by formula (VIII).
77. A multisubunit peptide encoded by the multisubunit nucleic acid according to any one of the claims 1-60.
78. A multisubunit peptide comprising two or more polypeptides, wherein some or all polypeptides comprises either a target peptide, a linker peptide, and a self-assembling peptide, or a linker peptide, a target peptide, a linker peptide, and a self-assembling peptide or a combination thereof, wherein one polypeptide is connected to another polypeptide by a cleavage peptide, wherein the multisubunit peptide includes a signal peptide on the amino-terminus of one or more of the polypeptides, wherein the target peptide is obtained or derived from:(a) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, non-structural protein, B cell epitope, T cell epitope, or a combination of a respiratory syncytial virus,(b) envelope protein, matrix protein, nucleocapsid protein, M2-1 protein, M2-2 protein, B cell epitope, T cell epitope, or a combination thereof of a metapneumovirus,(c) envelope protein, nucleocapsid protein, matrix protein, accessory protein, B cell epitope, T cell epitope, or a combination thereof of a human parainfluenza virus, or(d) a combination of (a), (b), and / or (c).
79. A polypeptide nanoparticle comprising at least 2 or up to 500 polypeptides according to any one of the claims 5-60 or 78.
80. The polypeptide nanoparticle according to claim 79, comprising a homologous polypeptide, a heterologous polypeptide, an oligomeric complex, a polypeptide cluster, or a combination thereof.
81. The polypeptide nanoparticle according to any one of the claims 79-80, wherein the polypeptide nanoparticle is icosahedral, helical, spherical, rod-like, or a combination thereof.
82. A vaccine comprising the multisubunit nucleic acid according to any one of the claims 1-60.
83. A vaccine comprising the lipid nanoparticle composition according to any one of the claims 61-76.
84. A vaccine comprising the polypeptide nanoparticle according to any one the claims 79-81.
85. A method of treating or preventing a disease, comprising administering to a subject in need thereof the multisubunit nucleic acid according to any one of the claims 1-60 or the vaccine according to any one the claims 82 or 84.
86. A method of treating or preventing a disease, comprising administering to a subject in need thereof the lipid nanoparticle composition according to any one of the claims 61- 76 or the vaccine according to claim 83.
87. The method according to any one of the claims 85-86, wherein the disease is(a) an RSV disease caused by respiratory syncytial virus,(b) a respiratory tract infection, bronchiolitis, or pneumonia caused by metapneumovirus, or(c) a respiratory tract infection, croup, bronchitis, bronchiolitis, or pneumonia caused by human parainfluenza virus.
88. Use of the multisubunit nucleic acid according to any one of the claims 1-60 or the vaccine according to any one the claims 82 or 84, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
89. Use of the lipid nanoparticle composition according to any one of the claims 61-76 or the vaccine according to claim 83, in the manufacture of a medicament for the treatment or prevention of a disease in a subject.
90. The use according to any one of the claims 88-89, wherein the disease is(a) an RSV disease caused by respiratory syncytial virus,(b) a respiratory tract infection, bronchiolitis, or pneumonia caused by metapneumovirus, or(c) a respiratory tract infection, croup, bronchitis, bronchiolitis, or pneumonia caused by human parainfluenza virus.
Citation Information
Patent Citations
Recombinant parainfluenza virus expression systems and vaccines comprising heterologous antigens derived from metapneumovirus
US20050142148A1
Vaccine compositions and methods for enhanced antigen-specific vaccination
US20210106667A1
Cited By
Human metapneumovirus virus-like particle, preparation method, application and vaccine
CN116554341A
A human metapneumovirus virus-like particle, preparation method, application and vaccine
CN116554341B