US20030207406A1 - Method of transferring at least two saccharide units with a polyglycosyltransferase - Google Patents
Method of transferring at least two saccharide units with a polyglycosyltransferase Download PDFInfo
- Publication number
- US20030207406A1 US20030207406A1 US10/096,129 US9612902A US2003207406A1 US 20030207406 A1 US20030207406 A1 US 20030207406A1 US 9612902 A US9612902 A US 9612902A US 2003207406 A1 US2003207406 A1 US 2003207406A1
- Authority
- US
- United States
- Prior art keywords
- ala
- leu
- arg
- asp
- polyglycosyltransferase
- Prior art date
- Legal status (The legal status is an assumption and is not a legal conclusion. Google has not performed a legal analysis and makes no representation as to the accuracy of the status listed.)
- Abandoned
Links
- 238000000034 method Methods 0.000 title claims abstract description 29
- 125000000837 carbohydrate group Chemical group 0.000 title abstract 2
- 150000001720 carbohydrates Chemical group 0.000 claims description 23
- 241000588652 Neisseria gonorrhoeae Species 0.000 claims description 19
- 235000000346 sugar Nutrition 0.000 claims description 16
- 229930182830 galactose Natural products 0.000 claims description 14
- MBLBDJOUHNCFQT-UHFFFAOYSA-N N-acetyl-D-galactosamine Natural products CC(=O)NC(C=O)C(O)C(O)C(O)CO MBLBDJOUHNCFQT-UHFFFAOYSA-N 0.000 claims description 10
- OVRNDRQMDRJTHS-UHFFFAOYSA-N N-acelyl-D-glucosamine Natural products CC(=O)NC1C(O)OC(CO)C(O)C1O OVRNDRQMDRJTHS-UHFFFAOYSA-N 0.000 claims description 9
- MBLBDJOUHNCFQT-LXGUWJNJSA-N N-acetylglucosamine Natural products CC(=O)N[C@@H](C=O)[C@@H](O)[C@H](O)[C@H](O)CO MBLBDJOUHNCFQT-LXGUWJNJSA-N 0.000 claims description 9
- 102000004190 Enzymes Human genes 0.000 claims description 8
- 108090000790 Enzymes Proteins 0.000 claims description 8
- 241000588653 Neisseria Species 0.000 claims description 8
- OVRNDRQMDRJTHS-CBQIKETKSA-N N-Acetyl-D-Galactosamine Chemical compound CC(=O)N[C@H]1[C@@H](O)O[C@H](CO)[C@H](O)[C@@H]1O OVRNDRQMDRJTHS-CBQIKETKSA-N 0.000 claims description 5
- 102100023315 N-acetyllactosaminide beta-1,6-N-acetylglucosaminyl-transferase Human genes 0.000 claims description 4
- 108010056664 N-acetyllactosaminide beta-1,6-N-acetylglucosaminyltransferase Proteins 0.000 claims description 4
- OVRNDRQMDRJTHS-FMDGEEDCSA-N N-acetyl-beta-D-glucosamine Chemical compound CC(=O)N[C@H]1[C@H](O)O[C@H](CO)[C@@H](O)[C@@H]1O OVRNDRQMDRJTHS-FMDGEEDCSA-N 0.000 claims description 3
- 229950006780 n-acetylglucosamine Drugs 0.000 claims description 3
- 108090000623 proteins and genes Proteins 0.000 abstract description 67
- 108020004414 DNA Proteins 0.000 description 31
- 108700023372 Glycosyltransferases Proteins 0.000 description 25
- 102000051366 Glycosyltransferases Human genes 0.000 description 25
- 229920001542 oligosaccharide Polymers 0.000 description 23
- 230000000694 effects Effects 0.000 description 22
- 150000002482 oligosaccharides Chemical class 0.000 description 21
- 239000013598 vector Substances 0.000 description 18
- 230000015572 biosynthetic process Effects 0.000 description 17
- 239000012634 fragment Substances 0.000 description 17
- 235000018102 proteins Nutrition 0.000 description 17
- 102000004169 proteins and genes Human genes 0.000 description 17
- 238000012546 transfer Methods 0.000 description 15
- 150000007523 nucleic acids Chemical class 0.000 description 14
- 230000014509 gene expression Effects 0.000 description 13
- 235000001014 amino acid Nutrition 0.000 description 12
- 229940024606 amino acid Drugs 0.000 description 11
- 150000001413 amino acids Chemical class 0.000 description 10
- 238000003786 synthesis reaction Methods 0.000 description 10
- 108020004707 nucleic acids Proteins 0.000 description 9
- 102000039446 nucleic acids Human genes 0.000 description 9
- WQZGKKKJIJFFOK-SVZMEOIVSA-N (+)-Galactose Chemical compound OC[C@H]1OC(O)[C@H](O)[C@@H](O)[C@H]1O WQZGKKKJIJFFOK-SVZMEOIVSA-N 0.000 description 8
- 108091028043 Nucleic acid sequence Proteins 0.000 description 8
- 239000013604 expression vector Substances 0.000 description 8
- 108700014210 glycosyltransferase activity proteins Proteins 0.000 description 8
- 238000000338 in vitro Methods 0.000 description 8
- 108091008146 restriction endonucleases Proteins 0.000 description 8
- OVRNDRQMDRJTHS-RTRLPJTCSA-N N-acetyl-D-glucosamine Chemical compound CC(=O)N[C@H]1C(O)O[C@H](CO)[C@@H](O)[C@@H]1O OVRNDRQMDRJTHS-RTRLPJTCSA-N 0.000 description 7
- 125000003275 alpha amino acid group Chemical group 0.000 description 7
- 210000004027 cell Anatomy 0.000 description 7
- 229940088598 enzyme Drugs 0.000 description 7
- 239000002773 nucleotide Substances 0.000 description 7
- 125000003729 nucleotide group Chemical group 0.000 description 7
- WQZGKKKJIJFFOK-GASJEMHNSA-N Glucose Natural products OC[C@H]1OC(O)[C@H](O)[C@@H](O)[C@@H]1O WQZGKKKJIJFFOK-GASJEMHNSA-N 0.000 description 6
- 241000880493 Leptailurus serval Species 0.000 description 6
- OVRNDRQMDRJTHS-KEWYIRBNSA-N N-acetyl-D-galactosamine Chemical compound CC(=O)N[C@H]1C(O)O[C@H](CO)[C@H](O)[C@@H]1O OVRNDRQMDRJTHS-KEWYIRBNSA-N 0.000 description 6
- 108020004511 Recombinant DNA Proteins 0.000 description 6
- 238000013459 approach Methods 0.000 description 6
- -1 carbohydrate monosaccharide Chemical class 0.000 description 6
- 238000010367 cloning Methods 0.000 description 6
- 102000037865 fusion proteins Human genes 0.000 description 6
- 108020001507 fusion proteins Proteins 0.000 description 6
- VPZXBVLAVMBEQI-UHFFFAOYSA-N glycyl-DL-alpha-alanine Natural products OC(=O)C(C)NC(=O)CN VPZXBVLAVMBEQI-UHFFFAOYSA-N 0.000 description 6
- 239000000047 product Substances 0.000 description 6
- 239000000523 sample Substances 0.000 description 6
- 241000894006 Bacteria Species 0.000 description 5
- WHUUTDBJXJRKMK-UHFFFAOYSA-N Glutamic acid Natural products OC(=O)C(N)CCC(O)=O WHUUTDBJXJRKMK-UHFFFAOYSA-N 0.000 description 5
- 108010047857 aspartylglycine Proteins 0.000 description 5
- 230000001580 bacterial effect Effects 0.000 description 5
- 239000013599 cloning vector Substances 0.000 description 5
- 239000008103 glucose Substances 0.000 description 5
- 108010003700 lysyl aspartic acid Proteins 0.000 description 5
- 239000013612 plasmid Substances 0.000 description 5
- FKQITMVNILRUCQ-IHRRRGAJSA-N Arg-Phe-Asp Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(O)=O)C(O)=O FKQITMVNILRUCQ-IHRRRGAJSA-N 0.000 description 4
- KTTCQQNRRLCIBC-GHCJXIJMSA-N Asp-Ile-Ala Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](C)C(O)=O KTTCQQNRRLCIBC-GHCJXIJMSA-N 0.000 description 4
- 241000588724 Escherichia coli Species 0.000 description 4
- RCFDOSNHHZGBOY-UHFFFAOYSA-N L-isoleucyl-L-alanine Natural products CCC(C)C(N)C(=O)NC(C)C(O)=O RCFDOSNHHZGBOY-UHFFFAOYSA-N 0.000 description 4
- CQGSYZCULZMEDE-UHFFFAOYSA-N Leu-Gln-Pro Natural products CC(C)CC(N)C(=O)NC(CCC(N)=O)C(=O)N1CCCC1C(O)=O CQGSYZCULZMEDE-UHFFFAOYSA-N 0.000 description 4
- 108010079364 N-glycylalanine Proteins 0.000 description 4
- USYGMBIIUDLYHJ-GVARAGBVSA-N Tyr-Ile-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 USYGMBIIUDLYHJ-GVARAGBVSA-N 0.000 description 4
- LFTYTUAZOPRMMI-UHFFFAOYSA-N UNPD164450 Natural products O1C(CO)C(O)C(O)C(NC(=O)C)C1OP(O)(=O)OP(O)(=O)OCC1C(O)C(O)C(N2C(NC(=O)C=C2)=O)O1 LFTYTUAZOPRMMI-UHFFFAOYSA-N 0.000 description 4
- 241000700605 Viruses Species 0.000 description 4
- 108010047495 alanylglycine Proteins 0.000 description 4
- 125000000539 amino acid group Chemical group 0.000 description 4
- 108010068380 arginylarginine Proteins 0.000 description 4
- 238000012217 deletion Methods 0.000 description 4
- 230000037430 deletion Effects 0.000 description 4
- 238000004519 manufacturing process Methods 0.000 description 4
- 239000003550 marker Substances 0.000 description 4
- 108091033319 polynucleotide Proteins 0.000 description 4
- 102000040430 polynucleotide Human genes 0.000 description 4
- 239000002157 polynucleotide Substances 0.000 description 4
- 230000007306 turnover Effects 0.000 description 4
- JKMHFZQWWAIEOD-UHFFFAOYSA-N 2-[4-(2-hydroxyethyl)piperazin-1-yl]ethanesulfonic acid Chemical compound OCC[NH+]1CCN(CCS([O-])(=O)=O)CC1 JKMHFZQWWAIEOD-UHFFFAOYSA-N 0.000 description 3
- DVWVZSJAYIJZFI-FXQIFTODSA-N Ala-Arg-Asn Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(N)=O)C(O)=O DVWVZSJAYIJZFI-FXQIFTODSA-N 0.000 description 3
- AOAKQKVICDWCLB-UWJYBYFXSA-N Ala-Tyr-Asn Chemical compound C[C@@H](C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)N[C@@H](CC(=O)N)C(=O)O)N AOAKQKVICDWCLB-UWJYBYFXSA-N 0.000 description 3
- LKDHUGLXOHYINY-XUXIUFHCSA-N Arg-Ile-Lys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N LKDHUGLXOHYINY-XUXIUFHCSA-N 0.000 description 3
- FANQWNCPNFEPGZ-WHFBIAKZSA-N Asp-Asp-Gly Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)NCC(O)=O FANQWNCPNFEPGZ-WHFBIAKZSA-N 0.000 description 3
- LTCKTLYKRMCFOC-KKUMJFAQSA-N Asp-Phe-Leu Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(O)=O LTCKTLYKRMCFOC-KKUMJFAQSA-N 0.000 description 3
- 101100512078 Caenorhabditis elegans lys-1 gene Proteins 0.000 description 3
- 108091026890 Coding region Proteins 0.000 description 3
- 108060003306 Galactosyltransferase Proteins 0.000 description 3
- 102000030902 Galactosyltransferase Human genes 0.000 description 3
- LPIKVBWNNVFHCQ-GUBZILKMSA-N Gln-Ser-Leu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(O)=O LPIKVBWNNVFHCQ-GUBZILKMSA-N 0.000 description 3
- XMWNHGKDDIFXQJ-NWLDYVSISA-N Gln-Thr-Trp Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)O)NC(=O)[C@H](CCC(=O)N)N)O XMWNHGKDDIFXQJ-NWLDYVSISA-N 0.000 description 3
- HVYWQYLBVXMXSV-GUBZILKMSA-N Glu-Leu-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(O)=O HVYWQYLBVXMXSV-GUBZILKMSA-N 0.000 description 3
- GDOZQTNZPCUARW-YFKPBYRVSA-N Gly-Gly-Glu Chemical compound NCC(=O)NCC(=O)N[C@H](C(O)=O)CCC(O)=O GDOZQTNZPCUARW-YFKPBYRVSA-N 0.000 description 3
- 239000007995 HEPES buffer Substances 0.000 description 3
- NYEYYMLUABXDMC-NHCYSSNCSA-N Ile-Gly-Leu Chemical compound CC[C@H](C)[C@@H](C(=O)NCC(=O)N[C@@H](CC(C)C)C(=O)O)N NYEYYMLUABXDMC-NHCYSSNCSA-N 0.000 description 3
- PMGDADKJMCOXHX-UHFFFAOYSA-N L-Arginyl-L-glutamin-acetat Natural products NC(=N)NCCCC(N)C(=O)NC(CCC(N)=O)C(O)=O PMGDADKJMCOXHX-UHFFFAOYSA-N 0.000 description 3
- FADYJNXDPBKVCA-UHFFFAOYSA-N L-Phenylalanyl-L-lysin Natural products NCCCCC(C(O)=O)NC(=O)C(N)CC1=CC=CC=C1 FADYJNXDPBKVCA-UHFFFAOYSA-N 0.000 description 3
- GUBGYTABKSRVRQ-QKKXKWKRSA-N Lactose Natural products OC[C@H]1O[C@@H](O[C@H]2[C@H](O)[C@@H](O)C(O)O[C@@H]2CO)[C@H](O)[C@@H](O)[C@H]1O GUBGYTABKSRVRQ-QKKXKWKRSA-N 0.000 description 3
- KTFHTMHHKXUYPW-ZPFDUUQYSA-N Leu-Asp-Ile Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O KTFHTMHHKXUYPW-ZPFDUUQYSA-N 0.000 description 3
- CQGSYZCULZMEDE-SRVKXCTJSA-N Leu-Gln-Pro Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N1CCC[C@H]1C(O)=O CQGSYZCULZMEDE-SRVKXCTJSA-N 0.000 description 3
- ORWTWZXGDBYVCP-BJDJZHNGSA-N Leu-Ile-Cys Chemical compound SC[C@@H](C(O)=O)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CC(C)C ORWTWZXGDBYVCP-BJDJZHNGSA-N 0.000 description 3
- JKSIBWITFMQTOA-XUXIUFHCSA-N Leu-Ile-Val Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](C(C)C)C(O)=O JKSIBWITFMQTOA-XUXIUFHCSA-N 0.000 description 3
- YQFZRHYZLARWDY-IHRRRGAJSA-N Leu-Val-Lys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CCCCN YQFZRHYZLARWDY-IHRRRGAJSA-N 0.000 description 3
- VKVDRTGWLVZJOM-DCAQKATOSA-N Leu-Val-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O VKVDRTGWLVZJOM-DCAQKATOSA-N 0.000 description 3
- SBQDRNOLGSYHQA-YUMQZZPRSA-N Lys-Ser-Gly Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CO)C(=O)NCC(O)=O SBQDRNOLGSYHQA-YUMQZZPRSA-N 0.000 description 3
- KCNSGAMPBPYUAI-CIUDSAMLSA-N Ser-Leu-Asn Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(N)=O)C(O)=O KCNSGAMPBPYUAI-CIUDSAMLSA-N 0.000 description 3
- WUXCHQZLUHBSDJ-LKXGYXEUSA-N Ser-Thr-Asp Chemical compound OC[C@H](N)C(=O)N[C@@H]([C@H](O)C)C(=O)N[C@@H](CC(O)=O)C(O)=O WUXCHQZLUHBSDJ-LKXGYXEUSA-N 0.000 description 3
- WTTRJMAZPDHPGS-KKXDTOCCSA-N Tyr-Phe-Ala Chemical compound C[C@H](NC(=O)[C@H](Cc1ccccc1)NC(=O)[C@@H](N)Cc1ccc(O)cc1)C(O)=O WTTRJMAZPDHPGS-KKXDTOCCSA-N 0.000 description 3
- HSCJRCZFDFQWRP-ABVWGUQPSA-N UDP-alpha-D-galactose Chemical compound O[C@@H]1[C@@H](O)[C@@H](O)[C@@H](CO)O[C@@H]1OP(O)(=O)OP(O)(=O)OC[C@@H]1[C@@H](O)[C@@H](O)[C@H](N2C(NC(=O)C=C2)=O)O1 HSCJRCZFDFQWRP-ABVWGUQPSA-N 0.000 description 3
- HSCJRCZFDFQWRP-UHFFFAOYSA-N Uridindiphosphoglukose Natural products OC1C(O)C(O)C(CO)OC1OP(O)(=O)OP(O)(=O)OCC1C(O)C(O)C(N2C(NC(=O)C=C2)=O)O1 HSCJRCZFDFQWRP-UHFFFAOYSA-N 0.000 description 3
- ZXAGTABZUOMUDO-GVXVVHGQSA-N Val-Glu-Lys Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)N[C@@H](CCCCN)C(=O)O)N ZXAGTABZUOMUDO-GVXVVHGQSA-N 0.000 description 3
- 238000007792 addition Methods 0.000 description 3
- 108010008685 alanyl-glutamyl-aspartic acid Proteins 0.000 description 3
- 108010044940 alanylglutamine Proteins 0.000 description 3
- WQZGKKKJIJFFOK-PHYPRBDBSA-N alpha-D-galactose Chemical compound OC[C@H]1O[C@H](O)[C@H](O)[C@@H](O)[C@H]1O WQZGKKKJIJFFOK-PHYPRBDBSA-N 0.000 description 3
- 239000007864 aqueous solution Substances 0.000 description 3
- 108010069205 aspartyl-phenylalanine Proteins 0.000 description 3
- WQZGKKKJIJFFOK-VFUOTHLCSA-N beta-D-glucose Chemical compound OC[C@H]1O[C@@H](O)[C@H](O)[C@@H](O)[C@@H]1O WQZGKKKJIJFFOK-VFUOTHLCSA-N 0.000 description 3
- 238000006243 chemical reaction Methods 0.000 description 3
- 230000006870 function Effects 0.000 description 3
- 108010050848 glycylleucine Proteins 0.000 description 3
- UYTPUPDQBNUYGX-UHFFFAOYSA-N guanine Chemical group O=C1NC(N)=NC2=C1N=CN2 UYTPUPDQBNUYGX-UHFFFAOYSA-N 0.000 description 3
- 238000003780 insertion Methods 0.000 description 3
- 230000037431 insertion Effects 0.000 description 3
- 239000008101 lactose Substances 0.000 description 3
- 108010025153 lysyl-alanyl-alanine Proteins 0.000 description 3
- 238000007899 nucleic acid hybridization Methods 0.000 description 3
- 108090000765 processed proteins & peptides Proteins 0.000 description 3
- 108010090894 prolylleucine Proteins 0.000 description 3
- 239000011541 reaction mixture Substances 0.000 description 3
- 230000001105 regulatory effect Effects 0.000 description 3
- 238000010561 standard procedure Methods 0.000 description 3
- 238000006467 substitution reaction Methods 0.000 description 3
- 238000013519 translation Methods 0.000 description 3
- 108010015666 tryptophyl-leucyl-glutamic acid Proteins 0.000 description 3
- HPYLHFWTUAGUNX-BGZSDMPXSA-N (3s)-4-[[(1s)-1-carboxy-2-hydroxyethyl]amino]-4-oxo-3-[[(2s,6s)-2,6,10-triamino-4-[(diaminomethylideneamino)methyl]-5-oxodecanoyl]amino]butanoic acid Chemical compound NCCCC[C@H](N)C(=O)C(CN=C(N)N)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O HPYLHFWTUAGUNX-BGZSDMPXSA-N 0.000 description 2
- AAQGRPOPTAUUBM-ZLUOBGJFSA-N Ala-Ala-Asn Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(N)=O)C(O)=O AAQGRPOPTAUUBM-ZLUOBGJFSA-N 0.000 description 2
- TTXMOJWKNRJWQJ-FXQIFTODSA-N Ala-Arg-Ser Chemical compound OC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)C)CCCN=C(N)N TTXMOJWKNRJWQJ-FXQIFTODSA-N 0.000 description 2
- KXEVYGKATAMXJJ-ACZMJKKPSA-N Ala-Glu-Asp Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O KXEVYGKATAMXJJ-ACZMJKKPSA-N 0.000 description 2
- NIZKGBJVCMRDKO-KWQFWETISA-N Ala-Gly-Tyr Chemical compound C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 NIZKGBJVCMRDKO-KWQFWETISA-N 0.000 description 2
- MNZHHDPWDWQJCQ-YUMQZZPRSA-N Ala-Leu-Gly Chemical compound C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O MNZHHDPWDWQJCQ-YUMQZZPRSA-N 0.000 description 2
- OMDNCNKNEGFOMM-BQBZGAKWSA-N Ala-Met-Gly Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCSC)C(=O)NCC(O)=O OMDNCNKNEGFOMM-BQBZGAKWSA-N 0.000 description 2
- YXXPVUOMPSZURS-ZLIFDBKOSA-N Ala-Trp-Leu Chemical compound C1=CC=C2C(C[C@@H](C(=O)N[C@@H](CC(C)C)C(O)=O)NC(=O)[C@H](C)N)=CNC2=C1 YXXPVUOMPSZURS-ZLIFDBKOSA-N 0.000 description 2
- BGGAIXWIZCIFSG-XDTLVQLUSA-N Ala-Tyr-Glu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCC(O)=O)C(O)=O BGGAIXWIZCIFSG-XDTLVQLUSA-N 0.000 description 2
- MUGAESARFRGOTQ-IGNZVWTISA-N Ala-Tyr-Tyr Chemical compound C[C@@H](C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)N[C@@H](CC2=CC=C(C=C2)O)C(=O)O)N MUGAESARFRGOTQ-IGNZVWTISA-N 0.000 description 2
- SGYSTDWPNPKJPP-GUBZILKMSA-N Arg-Ala-Arg Chemical compound NC(=N)NCCC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O SGYSTDWPNPKJPP-GUBZILKMSA-N 0.000 description 2
- MTANSHNQTWPZKP-KKUMJFAQSA-N Arg-Gln-Tyr Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CCCN=C(N)N)N)O MTANSHNQTWPZKP-KKUMJFAQSA-N 0.000 description 2
- JTZUZBADHGISJD-SRVKXCTJSA-N Arg-His-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCC(O)=O)C(O)=O JTZUZBADHGISJD-SRVKXCTJSA-N 0.000 description 2
- OQPAZKMGCWPERI-GUBZILKMSA-N Arg-Ser-Val Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(O)=O OQPAZKMGCWPERI-GUBZILKMSA-N 0.000 description 2
- GMRGSBAMMMVDGG-GUBZILKMSA-N Asn-Arg-Arg Chemical compound C(C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](CC(=O)N)N)CN=C(N)N GMRGSBAMMMVDGG-GUBZILKMSA-N 0.000 description 2
- MFFOYNGMOYFPBD-DCAQKATOSA-N Asn-Arg-Leu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(O)=O MFFOYNGMOYFPBD-DCAQKATOSA-N 0.000 description 2
- VKCOHFFSTKCXEQ-OLHMAJIHSA-N Asn-Asn-Thr Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O VKCOHFFSTKCXEQ-OLHMAJIHSA-N 0.000 description 2
- HPBNLFLSSQDFQW-WHFBIAKZSA-N Asn-Ser-Gly Chemical compound NC(=O)C[C@H](N)C(=O)N[C@@H](CO)C(=O)NCC(O)=O HPBNLFLSSQDFQW-WHFBIAKZSA-N 0.000 description 2
- WSWYMRLTJVKRCE-ZLUOBGJFSA-N Asp-Ala-Asp Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(O)=O WSWYMRLTJVKRCE-ZLUOBGJFSA-N 0.000 description 2
- YNCHFVRXEQFPBY-BQBZGAKWSA-N Asp-Gly-Arg Chemical compound OC(=O)C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCCN=C(N)N YNCHFVRXEQFPBY-BQBZGAKWSA-N 0.000 description 2
- JUWZKMBALYLZCK-WHFBIAKZSA-N Asp-Gly-Asn Chemical compound OC(=O)C[C@H](N)C(=O)NCC(=O)N[C@@H](CC(N)=O)C(O)=O JUWZKMBALYLZCK-WHFBIAKZSA-N 0.000 description 2
- YRZIYQGXTSBRLT-AVGNSLFASA-N Asp-Phe-Gln Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(N)=O)C(O)=O YRZIYQGXTSBRLT-AVGNSLFASA-N 0.000 description 2
- GWIJZUVQVDJHDI-AVGNSLFASA-N Asp-Phe-Glu Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(O)=O)C(O)=O GWIJZUVQVDJHDI-AVGNSLFASA-N 0.000 description 2
- WMLFFCRUSPNENW-ZLUOBGJFSA-N Asp-Ser-Ala Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O WMLFFCRUSPNENW-ZLUOBGJFSA-N 0.000 description 2
- PDIYGFYAMZZFCW-JIOCBJNQSA-N Asp-Thr-Pro Chemical compound C[C@H]([C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CC(=O)O)N)O PDIYGFYAMZZFCW-JIOCBJNQSA-N 0.000 description 2
- 241000283690 Bos taurus Species 0.000 description 2
- 102000053602 DNA Human genes 0.000 description 2
- 102000004533 Endonucleases Human genes 0.000 description 2
- NNQHEEQNPQYPGL-FXQIFTODSA-N Gln-Ala-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CCC(N)=O)C(O)=O NNQHEEQNPQYPGL-FXQIFTODSA-N 0.000 description 2
- LPJVZYMINRLCQA-AVGNSLFASA-N Gln-Cys-Phe Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)O)NC(=O)[C@H](CS)NC(=O)[C@H](CCC(=O)N)N LPJVZYMINRLCQA-AVGNSLFASA-N 0.000 description 2
- KDXKFBSNIJYNNR-YVNDNENWSA-N Gln-Glu-Ile Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O KDXKFBSNIJYNNR-YVNDNENWSA-N 0.000 description 2
- CELXWPDNIGWCJN-WDCWCFNPSA-N Gln-Lys-Thr Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)O)C(O)=O CELXWPDNIGWCJN-WDCWCFNPSA-N 0.000 description 2
- OKQLXOYFUPVEHI-CIUDSAMLSA-N Gln-Ser-Met Chemical compound CSCC[C@@H](C(=O)O)NC(=O)[C@H](CO)NC(=O)[C@H](CCC(=O)N)N OKQLXOYFUPVEHI-CIUDSAMLSA-N 0.000 description 2
- DIXKFOPPGWKZLY-CIUDSAMLSA-N Glu-Arg-Asp Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(O)=O)C(O)=O DIXKFOPPGWKZLY-CIUDSAMLSA-N 0.000 description 2
- OXEMJGCAJFFREE-FXQIFTODSA-N Glu-Gln-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(O)=O OXEMJGCAJFFREE-FXQIFTODSA-N 0.000 description 2
- LGYZYFFDELZWRS-DCAQKATOSA-N Glu-Glu-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CCC(O)=O LGYZYFFDELZWRS-DCAQKATOSA-N 0.000 description 2
- QIQABBIDHGQXGA-ZPFDUUQYSA-N Glu-Ile-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O QIQABBIDHGQXGA-ZPFDUUQYSA-N 0.000 description 2
- JHSRJMUJOGLIHK-GUBZILKMSA-N Glu-Met-Glu Chemical compound CSCC[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CCC(=O)O)N JHSRJMUJOGLIHK-GUBZILKMSA-N 0.000 description 2
- NSTUFLGQJCOCDL-UWVGGRQHSA-N Gly-Leu-Arg Chemical compound NCC(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N NSTUFLGQJCOCDL-UWVGGRQHSA-N 0.000 description 2
- GMTXWRIDLGTVFC-IUCAKERBSA-N Gly-Lys-Glu Chemical compound [H]NCC(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(O)=O)C(O)=O GMTXWRIDLGTVFC-IUCAKERBSA-N 0.000 description 2
- YLEIWGJJBFBFHC-KBPBESRZSA-N Gly-Phe-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)CN)CC1=CC=CC=C1 YLEIWGJJBFBFHC-KBPBESRZSA-N 0.000 description 2
- ZZWUYQXMIFTIIY-WEDXCCLWSA-N Gly-Thr-Leu Chemical compound [H]NCC(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(C)C)C(O)=O ZZWUYQXMIFTIIY-WEDXCCLWSA-N 0.000 description 2
- DHMQDGOQFOQNFH-UHFFFAOYSA-N Glycine Chemical compound NCC(O)=O DHMQDGOQFOQNFH-UHFFFAOYSA-N 0.000 description 2
- MJNWEIMBXKKCSF-XVYDVKMFSA-N His-Ala-Asn Chemical compound C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CC1=CN=CN1)N MJNWEIMBXKKCSF-XVYDVKMFSA-N 0.000 description 2
- NQKRILCJYCASDV-QWRGUYRKSA-N His-Gly-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CC1=CN=CN1 NQKRILCJYCASDV-QWRGUYRKSA-N 0.000 description 2
- AKAPKBNIVNPIPO-KKUMJFAQSA-N His-His-Lys Chemical compound C([C@@H](C(=O)N[C@@H](CCCCN)C(O)=O)NC(=O)[C@@H](N)CC=1NC=NC=1)C1=CN=CN1 AKAPKBNIVNPIPO-KKUMJFAQSA-N 0.000 description 2
- RWIKBYVJQAJYDP-BJDJZHNGSA-N Ile-Ala-Lys Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCCCN RWIKBYVJQAJYDP-BJDJZHNGSA-N 0.000 description 2
- NKRJALPCDNXULF-BYULHYEWSA-N Ile-Asp-Gly Chemical compound [H]N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(O)=O)C(=O)NCC(O)=O NKRJALPCDNXULF-BYULHYEWSA-N 0.000 description 2
- PFPUFNLHBXKPHY-HTFCKZLJSA-N Ile-Ile-Ser Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CO)C(=O)O)N PFPUFNLHBXKPHY-HTFCKZLJSA-N 0.000 description 2
- KLBVGHCGHUNHEA-BJDJZHNGSA-N Ile-Leu-Ala Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(=O)O)N KLBVGHCGHUNHEA-BJDJZHNGSA-N 0.000 description 2
- VBGCPJBKUXRYDA-DSYPUSFNSA-N Ile-Trp-Lys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)N[C@@H](CCCCN)C(=O)O)N VBGCPJBKUXRYDA-DSYPUSFNSA-N 0.000 description 2
- 108010079091 KRDS peptide Proteins 0.000 description 2
- AGPKZVBTJJNPAG-WHFBIAKZSA-N L-isoleucine Chemical compound CC[C@H](C)[C@H](N)C(O)=O AGPKZVBTJJNPAG-WHFBIAKZSA-N 0.000 description 2
- OUYCCCASQSFEME-QMMMGPOBSA-N L-tyrosine Chemical compound OC(=O)[C@@H](N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-QMMMGPOBSA-N 0.000 description 2
- KAFOIVJDVSZUMD-UHFFFAOYSA-N Leu-Gln-Gln Natural products CC(C)CC(N)C(=O)NC(CCC(N)=O)C(=O)NC(CCC(N)=O)C(O)=O KAFOIVJDVSZUMD-UHFFFAOYSA-N 0.000 description 2
- HQUXQAMSWFIRET-AVGNSLFASA-N Leu-Glu-Lys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@H](C(O)=O)CCCCN HQUXQAMSWFIRET-AVGNSLFASA-N 0.000 description 2
- BABSVXFGKFLIGW-UWVGGRQHSA-N Leu-Gly-Arg Chemical compound CC(C)C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCCNC(N)=N BABSVXFGKFLIGW-UWVGGRQHSA-N 0.000 description 2
- KGCLIYGPQXUNLO-IUCAKERBSA-N Leu-Gly-Glu Chemical compound CC(C)C[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCC(O)=O KGCLIYGPQXUNLO-IUCAKERBSA-N 0.000 description 2
- HNDWYLYAYNBWMP-AJNGGQMLSA-N Leu-Ile-Lys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CC(C)C)N HNDWYLYAYNBWMP-AJNGGQMLSA-N 0.000 description 2
- OMHLATXVNQSALM-FQUUOJAGSA-N Leu-Ile-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CC(C)C)N OMHLATXVNQSALM-FQUUOJAGSA-N 0.000 description 2
- YWKNKRAKOCLOLH-OEAJRASXSA-N Leu-Phe-Thr Chemical compound CC(C)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H]([C@@H](C)O)C(O)=O)CC1=CC=CC=C1 YWKNKRAKOCLOLH-OEAJRASXSA-N 0.000 description 2
- VULJUQZPSOASBZ-SRVKXCTJSA-N Leu-Pro-Glu Chemical compound [H]N[C@@H](CC(C)C)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(O)=O VULJUQZPSOASBZ-SRVKXCTJSA-N 0.000 description 2
- RGUXWMDNCPMQFB-YUMQZZPRSA-N Leu-Ser-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)NCC(O)=O RGUXWMDNCPMQFB-YUMQZZPRSA-N 0.000 description 2
- UCRJTSIIAYHOHE-ULQDDVLXSA-N Leu-Tyr-Arg Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)N UCRJTSIIAYHOHE-ULQDDVLXSA-N 0.000 description 2
- SWWCDAGDQHTKIE-RHYQMDGZSA-N Lys-Arg-Thr Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SWWCDAGDQHTKIE-RHYQMDGZSA-N 0.000 description 2
- FLCMXEFCTLXBTL-DCAQKATOSA-N Lys-Asp-Arg Chemical compound C(CCN)C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)N FLCMXEFCTLXBTL-DCAQKATOSA-N 0.000 description 2
- IWWMPCPLFXFBAF-SRVKXCTJSA-N Lys-Asp-Leu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O IWWMPCPLFXFBAF-SRVKXCTJSA-N 0.000 description 2
- PGLGNCVOWIORQE-SRVKXCTJSA-N Lys-His-Ser Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CO)C(O)=O PGLGNCVOWIORQE-SRVKXCTJSA-N 0.000 description 2
- WAIHHELKYSFIQN-XUXIUFHCSA-N Lys-Ile-Val Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](C(C)C)C(O)=O WAIHHELKYSFIQN-XUXIUFHCSA-N 0.000 description 2
- LOGFVTREOLYCPF-RHYQMDGZSA-N Lys-Pro-Thr Chemical compound C[C@@H](O)[C@@H](C(O)=O)NC(=O)[C@@H]1CCCN1C(=O)[C@@H](N)CCCCN LOGFVTREOLYCPF-RHYQMDGZSA-N 0.000 description 2
- KDXKERNSBIXSRK-UHFFFAOYSA-N Lysine Natural products NCCCCC(N)C(O)=O KDXKERNSBIXSRK-UHFFFAOYSA-N 0.000 description 2
- JQEBITVYKUCBMC-SRVKXCTJSA-N Met-Arg-Arg Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O JQEBITVYKUCBMC-SRVKXCTJSA-N 0.000 description 2
- FTQOFRPGLYXRFM-CYDGBPFRSA-N Met-Ile-Met Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCSC)C(=O)O)NC(=O)[C@H](CCSC)N FTQOFRPGLYXRFM-CYDGBPFRSA-N 0.000 description 2
- YBAFDPFAUTYYRW-UHFFFAOYSA-N N-L-alpha-glutamyl-L-leucine Natural products CC(C)CC(C(O)=O)NC(=O)C(N)CCC(O)=O YBAFDPFAUTYYRW-UHFFFAOYSA-N 0.000 description 2
- AJHCSUXXECOXOY-UHFFFAOYSA-N N-glycyl-L-tryptophan Natural products C1=CC=C2C(CC(NC(=O)CN)C(O)=O)=CNC2=C1 AJHCSUXXECOXOY-UHFFFAOYSA-N 0.000 description 2
- 241000588673 Neisseria elongata Species 0.000 description 2
- 241000588659 Neisseria mucosa Species 0.000 description 2
- BJEYSVHMGIJORT-NHCYSSNCSA-N Phe-Ala-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C)NC(=O)[C@@H](N)CC1=CC=CC=C1 BJEYSVHMGIJORT-NHCYSSNCSA-N 0.000 description 2
- WPTYDQPGBMDUBI-QWRGUYRKSA-N Phe-Gly-Asn Chemical compound N[C@@H](Cc1ccccc1)C(=O)NCC(=O)N[C@@H](CC(N)=O)C(O)=O WPTYDQPGBMDUBI-QWRGUYRKSA-N 0.000 description 2
- HGNGAMWHGGANAU-WHOFXGATSA-N Phe-Gly-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CC1=CC=CC=C1 HGNGAMWHGGANAU-WHOFXGATSA-N 0.000 description 2
- KNYPNEYICHHLQL-ACRUOGEOSA-N Phe-Leu-Tyr Chemical compound C([C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)C1=CC=CC=C1 KNYPNEYICHHLQL-ACRUOGEOSA-N 0.000 description 2
- KAJLHCWRWDSROH-BZSNNMDCSA-N Phe-Phe-Asp Chemical compound C([C@H](N)C(=O)N[C@@H](CC=1C=CC=CC=1)C(=O)N[C@@H](CC(O)=O)C(O)=O)C1=CC=CC=C1 KAJLHCWRWDSROH-BZSNNMDCSA-N 0.000 description 2
- WKLMCMXFMQEKCX-SLFFLAALSA-N Phe-Phe-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CC=CC=C2)NC(=O)[C@H](CC3=CC=CC=C3)N)C(=O)O WKLMCMXFMQEKCX-SLFFLAALSA-N 0.000 description 2
- LHALYDBUDCWMDY-CIUDSAMLSA-N Pro-Glu-Ala Chemical compound C[C@H](NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H]1CCCN1)C(O)=O LHALYDBUDCWMDY-CIUDSAMLSA-N 0.000 description 2
- FRKBNXCFJBPJOL-GUBZILKMSA-N Pro-Glu-Glu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O FRKBNXCFJBPJOL-GUBZILKMSA-N 0.000 description 2
- FJLODLCIOJUDRG-PYJNHQTQSA-N Pro-Ile-His Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)NC(=O)[C@@H]2CCCN2 FJLODLCIOJUDRG-PYJNHQTQSA-N 0.000 description 2
- BGWKULMLUIUPKY-BQBZGAKWSA-N Pro-Ser-Gly Chemical compound OC(=O)CNC(=O)[C@H](CO)NC(=O)[C@@H]1CCCN1 BGWKULMLUIUPKY-BQBZGAKWSA-N 0.000 description 2
- 240000004808 Saccharomyces cerevisiae Species 0.000 description 2
- HBTCFCHYALPXME-HTFCKZLJSA-N Ser-Ile-Ile Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O HBTCFCHYALPXME-HTFCKZLJSA-N 0.000 description 2
- ZIFYDQAFEMIZII-GUBZILKMSA-N Ser-Leu-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O ZIFYDQAFEMIZII-GUBZILKMSA-N 0.000 description 2
- RHAPJNVNWDBFQI-BQBZGAKWSA-N Ser-Pro-Gly Chemical compound OC[C@H](N)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O RHAPJNVNWDBFQI-BQBZGAKWSA-N 0.000 description 2
- KRDSCBLRHORMRK-JXUBOQSCSA-N Thr-Lys-Ala Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C)C(O)=O KRDSCBLRHORMRK-JXUBOQSCSA-N 0.000 description 2
- 102000006601 Thymidine Kinase Human genes 0.000 description 2
- 108020004440 Thymidine kinase Proteins 0.000 description 2
- YTCNLMSUXPCFBW-SXNHZJKMSA-N Trp-Ile-Glu Chemical compound [H]N[C@@H](CC1=CNC2=C1C=CC=C2)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCC(O)=O)C(O)=O YTCNLMSUXPCFBW-SXNHZJKMSA-N 0.000 description 2
- OGZRZMJASKKMJZ-XIRDDKMYSA-N Trp-Leu-Asp Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](CC1=CNC2=CC=CC=C21)N OGZRZMJASKKMJZ-XIRDDKMYSA-N 0.000 description 2
- FHHYVSCGOMPLLO-IHPCNDPISA-N Trp-Tyr-Asp Chemical compound C([C@H](NC(=O)[C@H](CC=1C2=CC=CC=C2NC=1)N)C(=O)N[C@@H](CC(O)=O)C(O)=O)C1=CC=C(O)C=C1 FHHYVSCGOMPLLO-IHPCNDPISA-N 0.000 description 2
- HKIUVWMZYFBIHG-KKUMJFAQSA-N Tyr-Arg-Gln Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N)O HKIUVWMZYFBIHG-KKUMJFAQSA-N 0.000 description 2
- ADBDQGBDNUTRDB-ULQDDVLXSA-N Tyr-Arg-Leu Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(O)=O ADBDQGBDNUTRDB-ULQDDVLXSA-N 0.000 description 2
- MNMYOSZWCKYEDI-JRQIVUDYSA-N Tyr-Asp-Thr Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O MNMYOSZWCKYEDI-JRQIVUDYSA-N 0.000 description 2
- FXYOYUMPUJONGW-FHWLQOOXSA-N Tyr-Gln-Phe Chemical compound C([C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CC=C(O)C=C1 FXYOYUMPUJONGW-FHWLQOOXSA-N 0.000 description 2
- LFTYTUAZOPRMMI-NESSUJCYSA-N UDP-N-acetyl-alpha-D-galactosamine Chemical compound O1[C@H](CO)[C@H](O)[C@H](O)[C@@H](NC(=O)C)[C@H]1O[P@](O)(=O)O[P@](O)(=O)OC[C@@H]1[C@@H](O)[C@@H](O)[C@H](N2C(NC(=O)C=C2)=O)O1 LFTYTUAZOPRMMI-NESSUJCYSA-N 0.000 description 2
- LFTYTUAZOPRMMI-CFRASDGPSA-N UDP-N-acetyl-alpha-D-glucosamine Chemical compound O1[C@H](CO)[C@@H](O)[C@H](O)[C@@H](NC(=O)C)[C@H]1OP(O)(=O)OP(O)(=O)OC[C@@H]1[C@@H](O)[C@@H](O)[C@H](N2C(NC(=O)C=C2)=O)O1 LFTYTUAZOPRMMI-CFRASDGPSA-N 0.000 description 2
- BTWMICVCQLKKNR-DCAQKATOSA-N Val-Leu-Ser Chemical compound CC(C)[C@H]([NH3+])C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C([O-])=O BTWMICVCQLKKNR-DCAQKATOSA-N 0.000 description 2
- QTPQHINADBYBNA-DCAQKATOSA-N Val-Ser-Lys Chemical compound CC(C)[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCCCN QTPQHINADBYBNA-DCAQKATOSA-N 0.000 description 2
- PZTZYZUTCPZWJH-FXQIFTODSA-N Val-Ser-Ser Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(=O)O)N PZTZYZUTCPZWJH-FXQIFTODSA-N 0.000 description 2
- DFQZDQPLWBSFEJ-LSJOCFKGSA-N Val-Val-Asn Chemical compound CC(C)[C@@H](C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(=O)N)C(=O)O)N DFQZDQPLWBSFEJ-LSJOCFKGSA-N 0.000 description 2
- 108010005233 alanylglutamic acid Proteins 0.000 description 2
- 108010011559 alanylphenylalanine Proteins 0.000 description 2
- 108010089442 arginyl-leucyl-alanyl-arginine Proteins 0.000 description 2
- 108010084758 arginyl-tyrosyl-aspartic acid Proteins 0.000 description 2
- 108010077245 asparaginyl-proline Proteins 0.000 description 2
- 238000003556 assay Methods 0.000 description 2
- 230000008901 benefit Effects 0.000 description 2
- SQVRNKJHWKZAKO-UHFFFAOYSA-N beta-N-Acetyl-D-neuraminic acid Natural products CC(=O)NC1C(O)CC(O)(C(O)=O)OC1C(O)C(O)CO SQVRNKJHWKZAKO-UHFFFAOYSA-N 0.000 description 2
- 210000004899 c-terminal region Anatomy 0.000 description 2
- 235000014633 carbohydrates Nutrition 0.000 description 2
- 230000000295 complement effect Effects 0.000 description 2
- IERHLVCPSMICTF-XVFCMESISA-N cytidine 5'-monophosphate Chemical compound O=C1N=C(N)C=CN1[C@H]1[C@H](O)[C@H](O)[C@@H](COP(O)(O)=O)O1 IERHLVCPSMICTF-XVFCMESISA-N 0.000 description 2
- 239000002158 endotoxin Substances 0.000 description 2
- 238000002474 experimental method Methods 0.000 description 2
- 230000005714 functional activity Effects 0.000 description 2
- 108010063718 gamma-glutamylaspartic acid Proteins 0.000 description 2
- 108010049041 glutamylalanine Proteins 0.000 description 2
- 150000004676 glycans Chemical class 0.000 description 2
- 125000003630 glycyl group Chemical group [H]N([H])C([H])([H])C(*)=O 0.000 description 2
- 108010067216 glycyl-glycyl-glycine Proteins 0.000 description 2
- XKUKSGPZAADMRA-UHFFFAOYSA-N glycyl-glycyl-glycine Natural products NCC(=O)NCC(=O)NCC(O)=O XKUKSGPZAADMRA-UHFFFAOYSA-N 0.000 description 2
- 108010001064 glycyl-glycyl-glycyl-glycine Proteins 0.000 description 2
- 108010010147 glycylglutamine Proteins 0.000 description 2
- 108010084389 glycyltryptophan Proteins 0.000 description 2
- 108010037850 glycylvaline Proteins 0.000 description 2
- 238000001727 in vivo Methods 0.000 description 2
- 238000002955 isolation Methods 0.000 description 2
- 108010078274 isoleucylvaline Proteins 0.000 description 2
- 108010034529 leucyl-lysine Proteins 0.000 description 2
- 108010057821 leucylproline Proteins 0.000 description 2
- 150000002632 lipids Chemical class 0.000 description 2
- 229920006008 lipopolysaccharide Polymers 0.000 description 2
- 108010054155 lysyllysine Proteins 0.000 description 2
- 235000013336 milk Nutrition 0.000 description 2
- 239000008267 milk Substances 0.000 description 2
- 210000004080 milk Anatomy 0.000 description 2
- 230000004048 modification Effects 0.000 description 2
- 238000012986 modification Methods 0.000 description 2
- 238000010369 molecular cloning Methods 0.000 description 2
- 108010083476 phenylalanyltryptophan Proteins 0.000 description 2
- 229920001184 polypeptide Polymers 0.000 description 2
- 229920001282 polysaccharide Polymers 0.000 description 2
- 239000005017 polysaccharide Substances 0.000 description 2
- 239000013615 primer Substances 0.000 description 2
- 102000004196 processed proteins & peptides Human genes 0.000 description 2
- 108010077112 prolyl-proline Proteins 0.000 description 2
- 108010031719 prolyl-serine Proteins 0.000 description 2
- 230000002797 proteolythic effect Effects 0.000 description 2
- 238000003259 recombinant expression Methods 0.000 description 2
- 108010048397 seryl-lysyl-leucine Proteins 0.000 description 2
- 238000002741 site-directed mutagenesis Methods 0.000 description 2
- 239000007787 solid Substances 0.000 description 2
- 239000007790 solid phase Substances 0.000 description 2
- 239000000243 solution Substances 0.000 description 2
- 230000002194 synthesizing effect Effects 0.000 description 2
- 238000013518 transcription Methods 0.000 description 2
- 230000035897 transcription Effects 0.000 description 2
- 230000002103 transcriptional effect Effects 0.000 description 2
- 230000009466 transformation Effects 0.000 description 2
- 150000004043 trisaccharides Chemical class 0.000 description 2
- 108010020532 tyrosyl-proline Proteins 0.000 description 2
- 241000701447 unidentified baculovirus Species 0.000 description 2
- 241001515965 unidentified phage Species 0.000 description 2
- MTCFGRXMJLQNBG-REOHCLBHSA-N (2S)-2-Amino-3-hydroxypropansäure Chemical compound OC[C@H](N)C(O)=O MTCFGRXMJLQNBG-REOHCLBHSA-N 0.000 description 1
- 229920000936 Agarose Polymers 0.000 description 1
- BUANFPRKJKJSRR-ACZMJKKPSA-N Ala-Ala-Gln Chemical compound C[C@H]([NH3+])C(=O)N[C@@H](C)C(=O)N[C@H](C([O-])=O)CCC(N)=O BUANFPRKJKJSRR-ACZMJKKPSA-N 0.000 description 1
- WQVFQXXBNHHPLX-ZKWXMUAHSA-N Ala-Ala-His Chemical compound C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](Cc1cnc[nH]1)C(O)=O WQVFQXXBNHHPLX-ZKWXMUAHSA-N 0.000 description 1
- LGQPPBQRUBVTIF-JBDRJPRFSA-N Ala-Ala-Ile Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](C)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O LGQPPBQRUBVTIF-JBDRJPRFSA-N 0.000 description 1
- FJVAQLJNTSUQPY-CIUDSAMLSA-N Ala-Ala-Lys Chemical compound C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCCCN FJVAQLJNTSUQPY-CIUDSAMLSA-N 0.000 description 1
- CXRCVCURMBFFOL-FXQIFTODSA-N Ala-Ala-Pro Chemical compound C[C@H](N)C(=O)N[C@@H](C)C(=O)N1CCC[C@H]1C(O)=O CXRCVCURMBFFOL-FXQIFTODSA-N 0.000 description 1
- UGLPMYSCWHTZQU-AUTRQRHGSA-N Ala-Ala-Tyr Chemical compound C[C@H]([NH3+])C(=O)N[C@@H](C)C(=O)N[C@H](C([O-])=O)CC1=CC=C(O)C=C1 UGLPMYSCWHTZQU-AUTRQRHGSA-N 0.000 description 1
- KVWLTGNCJYDJET-LSJOCFKGSA-N Ala-Arg-His Chemical compound C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N KVWLTGNCJYDJET-LSJOCFKGSA-N 0.000 description 1
- FXKNPWNXPQZLES-ZLUOBGJFSA-N Ala-Asn-Ser Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CO)C(O)=O FXKNPWNXPQZLES-ZLUOBGJFSA-N 0.000 description 1
- NHCPCLJZRSIDHS-ZLUOBGJFSA-N Ala-Asp-Ala Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(O)=O NHCPCLJZRSIDHS-ZLUOBGJFSA-N 0.000 description 1
- KIUYPHAMDKDICO-WHFBIAKZSA-N Ala-Asp-Gly Chemical compound C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)NCC(O)=O KIUYPHAMDKDICO-WHFBIAKZSA-N 0.000 description 1
- GWFSQQNGMPGBEF-GHCJXIJMSA-N Ala-Asp-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](C)N GWFSQQNGMPGBEF-GHCJXIJMSA-N 0.000 description 1
- IFTVANMRTIHKML-WDSKDSINSA-N Ala-Gln-Gly Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)NCC(O)=O IFTVANMRTIHKML-WDSKDSINSA-N 0.000 description 1
- NWVVKQZOVSTDBQ-CIUDSAMLSA-N Ala-Glu-Arg Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O NWVVKQZOVSTDBQ-CIUDSAMLSA-N 0.000 description 1
- BGNLUHXLSAQYRQ-FXQIFTODSA-N Ala-Glu-Gln Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O BGNLUHXLSAQYRQ-FXQIFTODSA-N 0.000 description 1
- HXNNRBHASOSVPG-GUBZILKMSA-N Ala-Glu-Leu Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O HXNNRBHASOSVPG-GUBZILKMSA-N 0.000 description 1
- HMRWQTHUDVXMGH-GUBZILKMSA-N Ala-Glu-Lys Chemical compound C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@H](C(O)=O)CCCCN HMRWQTHUDVXMGH-GUBZILKMSA-N 0.000 description 1
- LMFXXZPPZDCPTA-ZKWXMUAHSA-N Ala-Gly-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)CNC(=O)[C@H](C)N LMFXXZPPZDCPTA-ZKWXMUAHSA-N 0.000 description 1
- QJABSQFUHKHTNP-SYWGBEHUSA-N Ala-Ile-Trp Chemical compound [H]N[C@@H](C)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC1=CNC2=C1C=CC=C2)C(O)=O QJABSQFUHKHTNP-SYWGBEHUSA-N 0.000 description 1
- RGQCNKIDEQJEBT-CQDKDKBSSA-N Ala-Leu-Tyr Chemical compound C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 RGQCNKIDEQJEBT-CQDKDKBSSA-N 0.000 description 1
- BLTRAARCJYVJKV-QEJZJMRPSA-N Ala-Lys-Phe Chemical compound C[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](Cc1ccccc1)C(O)=O BLTRAARCJYVJKV-QEJZJMRPSA-N 0.000 description 1
- MDNAVFBZPROEHO-DCAQKATOSA-N Ala-Lys-Val Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(O)=O MDNAVFBZPROEHO-DCAQKATOSA-N 0.000 description 1
- MDNAVFBZPROEHO-UHFFFAOYSA-N Ala-Lys-Val Natural products CC(C)C(C(O)=O)NC(=O)C(NC(=O)C(C)N)CCCCN MDNAVFBZPROEHO-UHFFFAOYSA-N 0.000 description 1
- RUXQNKVQSKOOBS-JURCDPSOSA-N Ala-Phe-Ile Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O RUXQNKVQSKOOBS-JURCDPSOSA-N 0.000 description 1
- VJVQKGYHIZPSNS-FXQIFTODSA-N Ala-Ser-Arg Chemical compound C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCCN=C(N)N VJVQKGYHIZPSNS-FXQIFTODSA-N 0.000 description 1
- NZGRHTKZFSVPAN-BIIVOSGPSA-N Ala-Ser-Pro Chemical compound C[C@@H](C(=O)N[C@@H](CO)C(=O)N1CCC[C@@H]1C(=O)O)N NZGRHTKZFSVPAN-BIIVOSGPSA-N 0.000 description 1
- BVLPIIBTWIYOML-ZKWXMUAHSA-N Ala-Val-Asp Chemical compound [H]N[C@@H](C)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(O)=O)C(O)=O BVLPIIBTWIYOML-ZKWXMUAHSA-N 0.000 description 1
- VHAQSYHSDKERBS-XPUUQOCRSA-N Ala-Val-Gly Chemical compound C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)NCC(O)=O VHAQSYHSDKERBS-XPUUQOCRSA-N 0.000 description 1
- ANNKVZSFQJGVDY-XUXIUFHCSA-N Ala-Val-Pro-Pro Chemical compound C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 ANNKVZSFQJGVDY-XUXIUFHCSA-N 0.000 description 1
- GUBGYTABKSRVRQ-XLOQQCSPSA-N Alpha-Lactose Chemical compound O[C@@H]1[C@@H](O)[C@@H](O)[C@@H](CO)O[C@H]1O[C@@H]1[C@@H](CO)O[C@H](O)[C@H](O)[C@H]1O GUBGYTABKSRVRQ-XLOQQCSPSA-N 0.000 description 1
- JGDGLDNAQJJGJI-AVGNSLFASA-N Arg-Arg-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCCN=C(N)N)NC(=O)[C@H](CCCN=C(N)N)N JGDGLDNAQJJGJI-AVGNSLFASA-N 0.000 description 1
- UISQLSIBJKEJSS-GUBZILKMSA-N Arg-Arg-Ser Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CO)C(O)=O UISQLSIBJKEJSS-GUBZILKMSA-N 0.000 description 1
- RCAUJZASOAFTAJ-FXQIFTODSA-N Arg-Asp-Cys Chemical compound C(C[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CS)C(=O)O)N)CN=C(N)N RCAUJZASOAFTAJ-FXQIFTODSA-N 0.000 description 1
- OGUPCHKBOKJFMA-SRVKXCTJSA-N Arg-Glu-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CCCN=C(N)N OGUPCHKBOKJFMA-SRVKXCTJSA-N 0.000 description 1
- YBIAYFFIVAZXPK-AVGNSLFASA-N Arg-His-Arg Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O YBIAYFFIVAZXPK-AVGNSLFASA-N 0.000 description 1
- NVUIWHJLPSZZQC-CYDGBPFRSA-N Arg-Ile-Arg Chemical compound NC(N)=NCCC[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O NVUIWHJLPSZZQC-CYDGBPFRSA-N 0.000 description 1
- OTZMRMHZCMZOJZ-SRVKXCTJSA-N Arg-Leu-Glu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O OTZMRMHZCMZOJZ-SRVKXCTJSA-N 0.000 description 1
- UZGFHWIJWPUPOH-IHRRRGAJSA-N Arg-Leu-Lys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N UZGFHWIJWPUPOH-IHRRRGAJSA-N 0.000 description 1
- IIAXFBUTKIDDIP-ULQDDVLXSA-N Arg-Leu-Phe Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O IIAXFBUTKIDDIP-ULQDDVLXSA-N 0.000 description 1
- YTMKMRSYXHBGER-IHRRRGAJSA-N Arg-Phe-Asn Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N YTMKMRSYXHBGER-IHRRRGAJSA-N 0.000 description 1
- IGFJVXOATGZTHD-UHFFFAOYSA-N Arg-Phe-His Natural products NC(CCNC(=N)N)C(=O)NC(Cc1ccccc1)C(=O)NC(Cc2c[nH]cn2)C(=O)O IGFJVXOATGZTHD-UHFFFAOYSA-N 0.000 description 1
- MNBHKGYCLBUIBC-UFYCRDLUSA-N Arg-Phe-Phe Chemical compound C([C@H](NC(=O)[C@H](CCCNC(N)=N)N)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CC=CC=C1 MNBHKGYCLBUIBC-UFYCRDLUSA-N 0.000 description 1
- QHVRVUNEAIFTEK-SZMVWBNQSA-N Arg-Pro-Trp Chemical compound N[C@@H](CCCNC(N)=N)C(=O)N1CCC[C@H]1C(=O)N[C@@H](Cc1c[nH]c2ccccc12)C(O)=O QHVRVUNEAIFTEK-SZMVWBNQSA-N 0.000 description 1
- URAUIUGLHBRPMF-NAKRPEOUSA-N Arg-Ser-Ile Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O URAUIUGLHBRPMF-NAKRPEOUSA-N 0.000 description 1
- KMFPQTITXUKJOV-DCAQKATOSA-N Arg-Ser-Leu Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(O)=O KMFPQTITXUKJOV-DCAQKATOSA-N 0.000 description 1
- ASQKVGRCKOFKIU-KZVJFYERSA-N Arg-Thr-Ala Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](C)C(=O)O)NC(=O)[C@H](CCCN=C(N)N)N)O ASQKVGRCKOFKIU-KZVJFYERSA-N 0.000 description 1
- WCZXPVPHUMYLMS-VEVYYDQMSA-N Arg-Thr-Asp Chemical compound [H]N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(O)=O)C(O)=O WCZXPVPHUMYLMS-VEVYYDQMSA-N 0.000 description 1
- QUBKBPZGMZWOKQ-SZMVWBNQSA-N Arg-Trp-Arg Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@H](CCCN=C(N)N)N)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O)=CNC2=C1 QUBKBPZGMZWOKQ-SZMVWBNQSA-N 0.000 description 1
- 239000004475 Arginine Substances 0.000 description 1
- QEYJFBMTSMLPKZ-ZKWXMUAHSA-N Asn-Ala-Val Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](C(C)C)C(O)=O QEYJFBMTSMLPKZ-ZKWXMUAHSA-N 0.000 description 1
- APHUDFFMXFYRKP-CIUDSAMLSA-N Asn-Asn-Lys Chemical compound C(CCN)C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)N)NC(=O)[C@H](CC(=O)N)N APHUDFFMXFYRKP-CIUDSAMLSA-N 0.000 description 1
- JQSWHKKUZMTOIH-QWRGUYRKSA-N Asn-Gly-Phe Chemical compound C1=CC=C(C=C1)C[C@@H](C(=O)O)NC(=O)CNC(=O)[C@H](CC(=O)N)N JQSWHKKUZMTOIH-QWRGUYRKSA-N 0.000 description 1
- OLISTMZJGQUOGS-GMOBBJLQSA-N Asn-Ile-Arg Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)NC(=O)[C@H](CC(=O)N)N OLISTMZJGQUOGS-GMOBBJLQSA-N 0.000 description 1
- HFPXZWPUVFVNLL-GUBZILKMSA-N Asn-Leu-Gln Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(O)=O HFPXZWPUVFVNLL-GUBZILKMSA-N 0.000 description 1
- WIDVAWAQBRAKTI-YUMQZZPRSA-N Asn-Leu-Gly Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O WIDVAWAQBRAKTI-YUMQZZPRSA-N 0.000 description 1
- GLWFAWNYGWBMOC-SRVKXCTJSA-N Asn-Leu-Leu Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O GLWFAWNYGWBMOC-SRVKXCTJSA-N 0.000 description 1
- QXOPPIDJKPEKCW-GUBZILKMSA-N Asn-Pro-Arg Chemical compound C1C[C@H](N(C1)C(=O)[C@H](CC(=O)N)N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)O QXOPPIDJKPEKCW-GUBZILKMSA-N 0.000 description 1
- NJSNXIOKBHPFMB-GMOBBJLQSA-N Asn-Pro-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@@H]1CCCN1C(=O)[C@H](CC(=O)N)N NJSNXIOKBHPFMB-GMOBBJLQSA-N 0.000 description 1
- ZNYKKCADEQAZKA-FXQIFTODSA-N Asn-Ser-Met Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCSC)C(O)=O ZNYKKCADEQAZKA-FXQIFTODSA-N 0.000 description 1
- WLVLIYYBPPONRJ-GCJQMDKQSA-N Asn-Thr-Ala Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C)C(O)=O WLVLIYYBPPONRJ-GCJQMDKQSA-N 0.000 description 1
- YQPSDMUGFKJZHR-QRTARXTBSA-N Asn-Trp-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CC1=CNC2=CC=CC=C21)NC(=O)[C@H](CC(=O)N)N YQPSDMUGFKJZHR-QRTARXTBSA-N 0.000 description 1
- JPPLRQVZMZFOSX-UWJYBYFXSA-N Asn-Tyr-Ala Chemical compound NC(=O)C[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C)C(O)=O)CC1=CC=C(O)C=C1 JPPLRQVZMZFOSX-UWJYBYFXSA-N 0.000 description 1
- YSYTWUMRHSFODC-QWRGUYRKSA-N Asn-Tyr-Gly Chemical compound [H]N[C@@H](CC(N)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)NCC(O)=O YSYTWUMRHSFODC-QWRGUYRKSA-N 0.000 description 1
- KRXIWXCXOARFNT-ZLUOBGJFSA-N Asp-Ala-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C)NC(=O)[C@@H](N)CC(O)=O KRXIWXCXOARFNT-ZLUOBGJFSA-N 0.000 description 1
- PBVLJOIPOGUQQP-CIUDSAMLSA-N Asp-Ala-Leu Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CC(C)C)C(O)=O PBVLJOIPOGUQQP-CIUDSAMLSA-N 0.000 description 1
- WSOKZUVWBXVJHX-CIUDSAMLSA-N Asp-Arg-Glu Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(O)=O)C(O)=O WSOKZUVWBXVJHX-CIUDSAMLSA-N 0.000 description 1
- NYLBGYLHBDFRHL-VEVYYDQMSA-N Asp-Arg-Thr Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H]([C@@H](C)O)C(O)=O NYLBGYLHBDFRHL-VEVYYDQMSA-N 0.000 description 1
- QOVWVLLHMMCFFY-ZLUOBGJFSA-N Asp-Asp-Asn Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O QOVWVLLHMMCFFY-ZLUOBGJFSA-N 0.000 description 1
- TVVYVAUGRHNTGT-UGYAYLCHSA-N Asp-Asp-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@@H](N)CC(O)=O TVVYVAUGRHNTGT-UGYAYLCHSA-N 0.000 description 1
- MJKBOVWWADWLHV-ZLUOBGJFSA-N Asp-Cys-Asp Chemical compound C([C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CC(=O)O)C(=O)O)N)C(=O)O MJKBOVWWADWLHV-ZLUOBGJFSA-N 0.000 description 1
- FTNVLGCFIJEMQT-CIUDSAMLSA-N Asp-Cys-Leu Chemical compound CC(C)C[C@@H](C(=O)O)NC(=O)[C@H](CS)NC(=O)[C@H](CC(=O)O)N FTNVLGCFIJEMQT-CIUDSAMLSA-N 0.000 description 1
- GHODABZPVZMWCE-FXQIFTODSA-N Asp-Glu-Glu Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O GHODABZPVZMWCE-FXQIFTODSA-N 0.000 description 1
- QCVXMEHGFUMKCO-YUMQZZPRSA-N Asp-Gly-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)CNC(=O)[C@@H](N)CC(O)=O QCVXMEHGFUMKCO-YUMQZZPRSA-N 0.000 description 1
- HOBNTSHITVVNBN-ZPFDUUQYSA-N Asp-Ile-Leu Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)O)NC(=O)[C@H](CC(=O)O)N HOBNTSHITVVNBN-ZPFDUUQYSA-N 0.000 description 1
- PYXXJFRXIYAESU-PCBIJLKTSA-N Asp-Ile-Phe Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O PYXXJFRXIYAESU-PCBIJLKTSA-N 0.000 description 1
- CJUKAWUWBZCTDQ-SRVKXCTJSA-N Asp-Leu-Lys Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(O)=O CJUKAWUWBZCTDQ-SRVKXCTJSA-N 0.000 description 1
- QBJCJWAZOPCNIX-JPLJXNOCSA-N Asp-Leu-Phe-Val Chemical compound OC(=O)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(=O)N[C@@H](C(C)C)C(O)=O)CC1=CC=CC=C1 QBJCJWAZOPCNIX-JPLJXNOCSA-N 0.000 description 1
- NZWDWXSWUQCNMG-GARJFASQSA-N Asp-Lys-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CCCCN)NC(=O)[C@H](CC(=O)O)N)C(=O)O NZWDWXSWUQCNMG-GARJFASQSA-N 0.000 description 1
- MYLZFUMPZCPJCJ-NHCYSSNCSA-N Asp-Lys-Val Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(O)=O MYLZFUMPZCPJCJ-NHCYSSNCSA-N 0.000 description 1
- MGSVBZIBCCKGCY-ZLUOBGJFSA-N Asp-Ser-Ser Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CO)C(O)=O MGSVBZIBCCKGCY-ZLUOBGJFSA-N 0.000 description 1
- IWLZBRTUIVXZJD-OLHMAJIHSA-N Asp-Thr-Asp Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(O)=O)C(O)=O IWLZBRTUIVXZJD-OLHMAJIHSA-N 0.000 description 1
- GXHDGYOXPNQCKM-XVSYOHENSA-N Asp-Thr-Phe Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)NC(=O)[C@H](CC(=O)O)N)O GXHDGYOXPNQCKM-XVSYOHENSA-N 0.000 description 1
- BJDHEININLSZOT-KKUMJFAQSA-N Asp-Tyr-Lys Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCCN)C(O)=O BJDHEININLSZOT-KKUMJFAQSA-N 0.000 description 1
- GIKOVDMXBAFXDF-NHCYSSNCSA-N Asp-Val-Leu Chemical compound [H]N[C@@H](CC(O)=O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O GIKOVDMXBAFXDF-NHCYSSNCSA-N 0.000 description 1
- DCXYFEDJOCDNAF-UHFFFAOYSA-N Asparagine Natural products OC(=O)C(N)CC(N)=O DCXYFEDJOCDNAF-UHFFFAOYSA-N 0.000 description 1
- 244000063299 Bacillus subtilis Species 0.000 description 1
- 208000035143 Bacterial infection Diseases 0.000 description 1
- 241000588676 Bergeriella denitrificans Species 0.000 description 1
- CPELXLSAUQHCOX-UHFFFAOYSA-M Bromide Chemical compound [Br-] CPELXLSAUQHCOX-UHFFFAOYSA-M 0.000 description 1
- 241000589513 Burkholderia cepacia Species 0.000 description 1
- TXCIAUNLDRJGJZ-BILDWYJOSA-N CMP-N-acetyl-beta-neuraminic acid Chemical compound O1[C@@H]([C@H](O)[C@H](O)CO)[C@H](NC(=O)C)[C@@H](O)C[C@]1(C(O)=O)OP(O)(=O)OC[C@@H]1[C@@H](O)[C@@H](O)[C@H](N2C(N=C(N)C=C2)=O)O1 TXCIAUNLDRJGJZ-BILDWYJOSA-N 0.000 description 1
- OKTJSMMVPCPJKN-UHFFFAOYSA-N Carbon Chemical group [C] OKTJSMMVPCPJKN-UHFFFAOYSA-N 0.000 description 1
- 108020004705 Codon Proteins 0.000 description 1
- GOKFTBDYUJCCSN-QEJZJMRPSA-N Cys-Glu-Trp Chemical compound C1=CC=C2C(=C1)C(=CN2)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CS)N GOKFTBDYUJCCSN-QEJZJMRPSA-N 0.000 description 1
- OXOQBEVULIBOSH-ZDLURKLDSA-N Cys-Gly-Thr Chemical compound [H]N[C@@H](CS)C(=O)NCC(=O)N[C@@H]([C@@H](C)O)C(O)=O OXOQBEVULIBOSH-ZDLURKLDSA-N 0.000 description 1
- BSGXXYRIDXUEOM-IHRRRGAJSA-N Cys-Phe-Met Chemical compound CSCC[C@@H](C(=O)O)NC(=O)[C@H](CC1=CC=CC=C1)NC(=O)[C@H](CS)N BSGXXYRIDXUEOM-IHRRRGAJSA-N 0.000 description 1
- MWVDDZUTWXFYHL-XKBZYTNZSA-N Cys-Thr-Gln Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)O)NC(=O)[C@H](CS)N)O MWVDDZUTWXFYHL-XKBZYTNZSA-N 0.000 description 1
- WQZGKKKJIJFFOK-QTVWNMPRSA-N D-mannopyranose Chemical compound OC[C@H]1OC(O)[C@@H](O)[C@@H](O)[C@@H]1O WQZGKKKJIJFFOK-QTVWNMPRSA-N 0.000 description 1
- 230000004544 DNA amplification Effects 0.000 description 1
- 239000003155 DNA primer Substances 0.000 description 1
- 241000588722 Escherichia Species 0.000 description 1
- 108010074860 Factor Xa Proteins 0.000 description 1
- QGWNDRXFNXRZMB-UUOKFMHZSA-N GDP Chemical compound C1=2NC(N)=NC(=O)C=2N=CN1[C@@H]1O[C@H](COP(O)(=O)OP(O)(O)=O)[C@@H](O)[C@H]1O QGWNDRXFNXRZMB-UUOKFMHZSA-N 0.000 description 1
- 108700028146 Genetic Enhancer Elements Proteins 0.000 description 1
- 108700039691 Genetic Promoter Regions Proteins 0.000 description 1
- KWUSGAIFNHQCBY-DCAQKATOSA-N Gln-Arg-Arg Chemical compound NC(=O)CC[C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O KWUSGAIFNHQCBY-DCAQKATOSA-N 0.000 description 1
- WOACHWLUOFZLGJ-GUBZILKMSA-N Gln-Arg-Gln Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(O)=O WOACHWLUOFZLGJ-GUBZILKMSA-N 0.000 description 1
- AAOBFSKXAVIORT-GUBZILKMSA-N Gln-Asn-Leu Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@@H](CC(C)C)C(O)=O AAOBFSKXAVIORT-GUBZILKMSA-N 0.000 description 1
- RKAQZCDMSUQTSS-FXQIFTODSA-N Gln-Asp-Gln Chemical compound C(CC(=O)N)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N RKAQZCDMSUQTSS-FXQIFTODSA-N 0.000 description 1
- WQWMZOIPXWSZNE-WDSKDSINSA-N Gln-Asp-Gly Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(=O)NCC(O)=O WQWMZOIPXWSZNE-WDSKDSINSA-N 0.000 description 1
- CRRFJBGUGNNOCS-PEFMBERDSA-N Gln-Asp-Ile Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O CRRFJBGUGNNOCS-PEFMBERDSA-N 0.000 description 1
- IKFZXRLDMYWNBU-YUMQZZPRSA-N Gln-Gly-Arg Chemical compound NC(=O)CC[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCCN=C(N)N IKFZXRLDMYWNBU-YUMQZZPRSA-N 0.000 description 1
- GLEGHWQNGPMKHO-DCAQKATOSA-N Gln-His-Glu Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](CCC(=O)N)N GLEGHWQNGPMKHO-DCAQKATOSA-N 0.000 description 1
- ZBKUIQNCRIYVGH-SDDRHHMPSA-N Gln-Leu-Pro Chemical compound CC(C)C[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)[C@H](CCC(=O)N)N ZBKUIQNCRIYVGH-SDDRHHMPSA-N 0.000 description 1
- ARYKRXHBIPLULY-XKBZYTNZSA-N Gln-Thr-Ser Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CO)C(O)=O ARYKRXHBIPLULY-XKBZYTNZSA-N 0.000 description 1
- FITIQFSXXBKFFM-NRPADANISA-N Gln-Val-Ser Chemical compound [H]N[C@@H](CCC(N)=O)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CO)C(O)=O FITIQFSXXBKFFM-NRPADANISA-N 0.000 description 1
- 102100039847 Globoside alpha-1,3-N-acetylgalactosaminyltransferase 1 Human genes 0.000 description 1
- RUFHOVYUYSNDNY-ACZMJKKPSA-N Glu-Ala-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](C)NC(=O)[C@@H](N)CCC(O)=O RUFHOVYUYSNDNY-ACZMJKKPSA-N 0.000 description 1
- UTKUTMJSWKKHEM-WDSKDSINSA-N Glu-Ala-Gly Chemical compound OC(=O)CNC(=O)[C@H](C)NC(=O)[C@@H](N)CCC(O)=O UTKUTMJSWKKHEM-WDSKDSINSA-N 0.000 description 1
- HUWSBFYAGXCXKC-CIUDSAMLSA-N Glu-Ala-Met Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](C)C(=O)N[C@@H](CCSC)C(O)=O HUWSBFYAGXCXKC-CIUDSAMLSA-N 0.000 description 1
- WOMUDRVDJMHTCV-DCAQKATOSA-N Glu-Arg-Arg Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O WOMUDRVDJMHTCV-DCAQKATOSA-N 0.000 description 1
- RCCDHXSRMWCOOY-GUBZILKMSA-N Glu-Arg-Gln Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CCC(N)=O)C(O)=O RCCDHXSRMWCOOY-GUBZILKMSA-N 0.000 description 1
- KKCUFHUTMKQQCF-SRVKXCTJSA-N Glu-Arg-Leu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](CC(C)C)C(O)=O KKCUFHUTMKQQCF-SRVKXCTJSA-N 0.000 description 1
- FLLRAEJOLZPSMN-CIUDSAMLSA-N Glu-Asn-Arg Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N FLLRAEJOLZPSMN-CIUDSAMLSA-N 0.000 description 1
- IESFZVCAVACGPH-PEFMBERDSA-N Glu-Asp-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@@H](N)CCC(O)=O IESFZVCAVACGPH-PEFMBERDSA-N 0.000 description 1
- SBCYJMOOHUDWDA-NUMRIWBASA-N Glu-Asp-Thr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SBCYJMOOHUDWDA-NUMRIWBASA-N 0.000 description 1
- GYCPQVFKCPPRQB-GUBZILKMSA-N Glu-Gln-Met Chemical compound CSCC[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)N)NC(=O)[C@H](CCC(=O)O)N GYCPQVFKCPPRQB-GUBZILKMSA-N 0.000 description 1
- HUFCEIHAFNVSNR-IHRRRGAJSA-N Glu-Gln-Tyr Chemical compound OC(=O)CC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 HUFCEIHAFNVSNR-IHRRRGAJSA-N 0.000 description 1
- QQLBPVKLJBAXBS-FXQIFTODSA-N Glu-Glu-Asn Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O QQLBPVKLJBAXBS-FXQIFTODSA-N 0.000 description 1
- VOORMNJKNBGYGK-YUMQZZPRSA-N Glu-Gly-Met Chemical compound CSCC[C@@H](C(=O)O)NC(=O)CNC(=O)[C@H](CCC(=O)O)N VOORMNJKNBGYGK-YUMQZZPRSA-N 0.000 description 1
- XMPAXPSENRSOSV-RYUDHWBXSA-N Glu-Gly-Tyr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)NCC(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O XMPAXPSENRSOSV-RYUDHWBXSA-N 0.000 description 1
- OQXDUSZKISQQSS-GUBZILKMSA-N Glu-Lys-Ala Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C)C(O)=O OQXDUSZKISQQSS-GUBZILKMSA-N 0.000 description 1
- CUPSDFQZTVVTSK-GUBZILKMSA-N Glu-Lys-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CCC(O)=O CUPSDFQZTVVTSK-GUBZILKMSA-N 0.000 description 1
- YKBUCXNNBYZYAY-MNXVOIDGSA-N Glu-Lys-Ile Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O YKBUCXNNBYZYAY-MNXVOIDGSA-N 0.000 description 1
- SUIAHERNFYRBDZ-GVXVVHGQSA-N Glu-Lys-Val Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](C(C)C)C(O)=O SUIAHERNFYRBDZ-GVXVVHGQSA-N 0.000 description 1
- IDEODOAVGCMUQV-GUBZILKMSA-N Glu-Ser-Leu Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(O)=O IDEODOAVGCMUQV-GUBZILKMSA-N 0.000 description 1
- JDAYMLXPUJRSDJ-XIRDDKMYSA-N Glu-Trp-Arg Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@H](CCC(O)=O)N)C(=O)N[C@@H](CCCN=C(N)N)C(O)=O)=CNC2=C1 JDAYMLXPUJRSDJ-XIRDDKMYSA-N 0.000 description 1
- LSYFGBRDBIQYAQ-FHWLQOOXSA-N Glu-Tyr-Tyr Chemical compound [H]N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O LSYFGBRDBIQYAQ-FHWLQOOXSA-N 0.000 description 1
- UZWUBBRJWFTHTD-LAEOZQHASA-N Glu-Val-Asn Chemical compound NC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CCC(O)=O UZWUBBRJWFTHTD-LAEOZQHASA-N 0.000 description 1
- 108010055629 Glucosyltransferases Proteins 0.000 description 1
- 102000000340 Glucosyltransferases Human genes 0.000 description 1
- GQGAFTPXAPKSCF-WHFBIAKZSA-N Gly-Ala-Cys Chemical compound NCC(=O)N[C@@H](C)C(=O)N[C@@H](CS)C(=O)O GQGAFTPXAPKSCF-WHFBIAKZSA-N 0.000 description 1
- LERGJIVJIIODPZ-ZANVPECISA-N Gly-Ala-Trp Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@@H](NC(=O)CN)C)C(O)=O)=CNC2=C1 LERGJIVJIIODPZ-ZANVPECISA-N 0.000 description 1
- UPOJUWHGMDJUQZ-IUCAKERBSA-N Gly-Arg-Arg Chemical compound NC(=N)NCCC[C@H](NC(=O)CN)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O UPOJUWHGMDJUQZ-IUCAKERBSA-N 0.000 description 1
- JVACNFOPSUPDTK-QWRGUYRKSA-N Gly-Asn-Phe Chemical compound NCC(=O)N[C@@H](CC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 JVACNFOPSUPDTK-QWRGUYRKSA-N 0.000 description 1
- IXKRSKPKSLXIHN-YUMQZZPRSA-N Gly-Cys-Leu Chemical compound [H]NCC(=O)N[C@@H](CS)C(=O)N[C@@H](CC(C)C)C(O)=O IXKRSKPKSLXIHN-YUMQZZPRSA-N 0.000 description 1
- KMSGYZQRXPUKGI-BYPYZUCNSA-N Gly-Gly-Asn Chemical compound NCC(=O)NCC(=O)N[C@H](C(O)=O)CC(N)=O KMSGYZQRXPUKGI-BYPYZUCNSA-N 0.000 description 1
- XPJBQTCXPJNIFE-ZETCQYMHSA-N Gly-Gly-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)CNC(=O)CN XPJBQTCXPJNIFE-ZETCQYMHSA-N 0.000 description 1
- SWQALSGKVLYKDT-UHFFFAOYSA-N Gly-Ile-Ala Natural products NCC(=O)NC(C(C)CC)C(=O)NC(C)C(O)=O SWQALSGKVLYKDT-UHFFFAOYSA-N 0.000 description 1
- BHPQOIPBLYJNAW-NGZCFLSTSA-N Gly-Ile-Pro Chemical compound CC[C@H](C)[C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)CN BHPQOIPBLYJNAW-NGZCFLSTSA-N 0.000 description 1
- HAXARWKYFIIHKD-ZKWXMUAHSA-N Gly-Ile-Ser Chemical compound NCC(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CO)C(O)=O HAXARWKYFIIHKD-ZKWXMUAHSA-N 0.000 description 1
- TWTPDFFBLQEBOE-IUCAKERBSA-N Gly-Leu-Gln Chemical compound [H]NCC(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(O)=O TWTPDFFBLQEBOE-IUCAKERBSA-N 0.000 description 1
- TVUWMSBGMVAHSJ-KBPBESRZSA-N Gly-Leu-Phe Chemical compound NCC(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CC1=CC=CC=C1 TVUWMSBGMVAHSJ-KBPBESRZSA-N 0.000 description 1
- MHXKHKWHPNETGG-QWRGUYRKSA-N Gly-Lys-Leu Chemical compound [H]NCC(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O MHXKHKWHPNETGG-QWRGUYRKSA-N 0.000 description 1
- YOBGUCWZPXJHTN-BQBZGAKWSA-N Gly-Ser-Arg Chemical compound NCC(=O)N[C@@H](CO)C(=O)N[C@H](C(O)=O)CCCN=C(N)N YOBGUCWZPXJHTN-BQBZGAKWSA-N 0.000 description 1
- MYXNLWDWWOTERK-BHNWBGBOSA-N Gly-Thr-Pro Chemical compound C[C@H]([C@@H](C(=O)N1CCC[C@@H]1C(=O)O)NC(=O)CN)O MYXNLWDWWOTERK-BHNWBGBOSA-N 0.000 description 1
- WTUSRDZLLWGYAT-KCTSRDHCSA-N Gly-Trp-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CC1=CNC2=CC=CC=C21)NC(=O)CN WTUSRDZLLWGYAT-KCTSRDHCSA-N 0.000 description 1
- BAYQNCWLXIDLHX-ONGXEEELSA-N Gly-Val-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)CN BAYQNCWLXIDLHX-ONGXEEELSA-N 0.000 description 1
- MUGLKCQHTUFLGF-WPRPVWTQSA-N Gly-Val-Met Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCSC)C(=O)O)NC(=O)CN MUGLKCQHTUFLGF-WPRPVWTQSA-N 0.000 description 1
- 239000004471 Glycine Substances 0.000 description 1
- JBCLFWXMTIKCCB-UHFFFAOYSA-N H-Gly-Phe-OH Natural products NCC(=O)NC(C(O)=O)CC1=CC=CC=C1 JBCLFWXMTIKCCB-UHFFFAOYSA-N 0.000 description 1
- 241000606768 Haemophilus influenzae Species 0.000 description 1
- 241000238631 Hexapoda Species 0.000 description 1
- AVQOSMRPITVTRB-CIUDSAMLSA-N His-Asn-Asn Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CC(=O)N)C(=O)N[C@@H](CC(=O)N)C(=O)O)N AVQOSMRPITVTRB-CIUDSAMLSA-N 0.000 description 1
- FYVHHKMHFPMBBG-GUBZILKMSA-N His-Gln-Asp Chemical compound C1=C(NC=N1)C[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](CC(=O)O)C(=O)O)N FYVHHKMHFPMBBG-GUBZILKMSA-N 0.000 description 1
- HQKADFMLECZIQJ-HVTMNAMFSA-N His-Glu-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CC1=CN=CN1)N HQKADFMLECZIQJ-HVTMNAMFSA-N 0.000 description 1
- FSOXZQBMPBQKGJ-QSFUFRPTSA-N His-Ile-Ala Chemical compound [O-]C(=O)[C@H](C)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H]([NH3+])CC1=CN=CN1 FSOXZQBMPBQKGJ-QSFUFRPTSA-N 0.000 description 1
- ORERHHPZDDEMSC-VGDYDELISA-N His-Ile-Ser Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CC1=CN=CN1)N ORERHHPZDDEMSC-VGDYDELISA-N 0.000 description 1
- CHIAUHSHDARFBD-ULQDDVLXSA-N His-Pro-Tyr Chemical compound C([C@H](N)C(=O)N1[C@@H](CCC1)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)C1=CN=CN1 CHIAUHSHDARFBD-ULQDDVLXSA-N 0.000 description 1
- QLBXWYXMLHAREM-PYJNHQTQSA-N His-Val-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](C(C)C)NC(=O)[C@H](CC1=CN=CN1)N QLBXWYXMLHAREM-PYJNHQTQSA-N 0.000 description 1
- 101000887519 Homo sapiens Globoside alpha-1,3-N-acetylgalactosaminyltransferase 1 Proteins 0.000 description 1
- XQFRJNBWHJMXHO-RRKCRQDMSA-N IDUR Chemical compound C1[C@H](O)[C@@H](CO)O[C@H]1N1C(=O)NC(=O)C(I)=C1 XQFRJNBWHJMXHO-RRKCRQDMSA-N 0.000 description 1
- YKRYHWJRQUSTKG-KBIXCLLPSA-N Ile-Ala-Gln Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](C)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N YKRYHWJRQUSTKG-KBIXCLLPSA-N 0.000 description 1
- WUEIUSDAECDLQO-NAKRPEOUSA-N Ile-Ala-Met Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](C)C(=O)N[C@@H](CCSC)C(=O)O)N WUEIUSDAECDLQO-NAKRPEOUSA-N 0.000 description 1
- MKWSZEHGHSLNPF-NAKRPEOUSA-N Ile-Ala-Val Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](C)C(=O)N[C@@H](C(C)C)C(=O)O)N MKWSZEHGHSLNPF-NAKRPEOUSA-N 0.000 description 1
- TZCGZYWNIDZZMR-UHFFFAOYSA-N Ile-Arg-Ala Natural products CCC(C)C(N)C(=O)NC(C(=O)NC(C)C(O)=O)CCCN=C(N)N TZCGZYWNIDZZMR-UHFFFAOYSA-N 0.000 description 1
- VZIFYHYNQDIPLI-HJWJTTGWSA-N Ile-Arg-Phe Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)N VZIFYHYNQDIPLI-HJWJTTGWSA-N 0.000 description 1
- LLZLRXBTOOFODM-QSFUFRPTSA-N Ile-Asp-Val Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N[C@@H](C(C)C)C(=O)O)N LLZLRXBTOOFODM-QSFUFRPTSA-N 0.000 description 1
- LWWILHPVAKKLQS-QXEWZRGKSA-N Ile-Gly-Met Chemical compound CC[C@H](C)[C@@H](C(=O)NCC(=O)N[C@@H](CCSC)C(=O)O)N LWWILHPVAKKLQS-QXEWZRGKSA-N 0.000 description 1
- JLWLMGADIQFKRD-QSFUFRPTSA-N Ile-His-Ala Chemical compound CC[C@H](C)[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C)C(O)=O)CC1=CN=CN1 JLWLMGADIQFKRD-QSFUFRPTSA-N 0.000 description 1
- TVYWVSJGSHQWMT-AJNGGQMLSA-N Ile-Leu-Lys Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(=O)O)N TVYWVSJGSHQWMT-AJNGGQMLSA-N 0.000 description 1
- MUFXDFWAJSPHIQ-XDTLVQLUSA-N Ile-Tyr Chemical compound CC[C@H](C)[C@H]([NH3+])C(=O)N[C@H](C([O-])=O)CC1=CC=C(O)C=C1 MUFXDFWAJSPHIQ-XDTLVQLUSA-N 0.000 description 1
- 108010065920 Insulin Lispro Proteins 0.000 description 1
- XUJNEKJLAYXESH-REOHCLBHSA-N L-Cysteine Chemical compound SC[C@H](N)C(O)=O XUJNEKJLAYXESH-REOHCLBHSA-N 0.000 description 1
- IBMVEYRWAWIOTN-UHFFFAOYSA-N L-Leucyl-L-Arginyl-L-Proline Natural products CC(C)CC(N)C(=O)NC(CCCN=C(N)N)C(=O)N1CCCC1C(O)=O IBMVEYRWAWIOTN-UHFFFAOYSA-N 0.000 description 1
- HGCNKOLVKRAVHD-UHFFFAOYSA-N L-Met-L-Phe Natural products CSCCC(N)C(=O)NC(C(O)=O)CC1=CC=CC=C1 HGCNKOLVKRAVHD-UHFFFAOYSA-N 0.000 description 1
- ONIBWKKTOPOVIA-BYPYZUCNSA-N L-Proline Chemical compound OC(=O)[C@@H]1CCCN1 ONIBWKKTOPOVIA-BYPYZUCNSA-N 0.000 description 1
- QNAYBMKLOCPYGJ-REOHCLBHSA-N L-alanine Chemical compound C[C@H](N)C(O)=O QNAYBMKLOCPYGJ-REOHCLBHSA-N 0.000 description 1
- ODKSFYDXXFIFQN-BYPYZUCNSA-P L-argininium(2+) Chemical compound NC(=[NH2+])NCCC[C@H]([NH3+])C(O)=O ODKSFYDXXFIFQN-BYPYZUCNSA-P 0.000 description 1
- DCXYFEDJOCDNAF-REOHCLBHSA-N L-asparagine Chemical compound OC(=O)[C@@H](N)CC(N)=O DCXYFEDJOCDNAF-REOHCLBHSA-N 0.000 description 1
- CKLJMWTZIZZHCS-REOHCLBHSA-N L-aspartic acid Chemical compound OC(=O)[C@@H](N)CC(O)=O CKLJMWTZIZZHCS-REOHCLBHSA-N 0.000 description 1
- WHUUTDBJXJRKMK-VKHMYHEASA-N L-glutamic acid Chemical compound OC(=O)[C@@H](N)CCC(O)=O WHUUTDBJXJRKMK-VKHMYHEASA-N 0.000 description 1
- ZDXPYRJPNDTMRX-VKHMYHEASA-N L-glutamine Chemical compound OC(=O)[C@@H](N)CCC(N)=O ZDXPYRJPNDTMRX-VKHMYHEASA-N 0.000 description 1
- HNDVDQJCIGZPNO-YFKPBYRVSA-N L-histidine Chemical compound OC(=O)[C@@H](N)CC1=CN=CN1 HNDVDQJCIGZPNO-YFKPBYRVSA-N 0.000 description 1
- ROHFNLRQFUQHCH-YFKPBYRVSA-N L-leucine Chemical compound CC(C)C[C@H](N)C(O)=O ROHFNLRQFUQHCH-YFKPBYRVSA-N 0.000 description 1
- LHSGPCFBGJHPCY-UHFFFAOYSA-N L-leucine-L-tyrosine Natural products CC(C)CC(N)C(=O)NC(C(O)=O)CC1=CC=C(O)C=C1 LHSGPCFBGJHPCY-UHFFFAOYSA-N 0.000 description 1
- KDXKERNSBIXSRK-YFKPBYRVSA-N L-lysine Chemical compound NCCCC[C@H](N)C(O)=O KDXKERNSBIXSRK-YFKPBYRVSA-N 0.000 description 1
- FFEARJCKVFRZRR-BYPYZUCNSA-N L-methionine Chemical compound CSCC[C@H](N)C(O)=O FFEARJCKVFRZRR-BYPYZUCNSA-N 0.000 description 1
- COLNVLDHVKWLRT-QMMMGPOBSA-N L-phenylalanine Chemical compound OC(=O)[C@@H](N)CC1=CC=CC=C1 COLNVLDHVKWLRT-QMMMGPOBSA-N 0.000 description 1
- AYFVYJQAPQTCCC-GBXIJSLDSA-N L-threonine Chemical compound C[C@@H](O)[C@H](N)C(O)=O AYFVYJQAPQTCCC-GBXIJSLDSA-N 0.000 description 1
- QIVBCDIJIAJPQS-VIFPVBQESA-N L-tryptophane Chemical compound C1=CC=C2C(C[C@H](N)C(O)=O)=CNC2=C1 QIVBCDIJIAJPQS-VIFPVBQESA-N 0.000 description 1
- KZSNJWFQEVHDMF-BYPYZUCNSA-N L-valine Chemical compound CC(C)[C@H](N)C(O)=O KZSNJWFQEVHDMF-BYPYZUCNSA-N 0.000 description 1
- 102000004407 Lactalbumin Human genes 0.000 description 1
- 108090000942 Lactalbumin Proteins 0.000 description 1
- PNIWLNAGKUGXDO-UHFFFAOYSA-N Lactosamine Natural products OC1C(N)C(O)OC(CO)C1OC1C(O)C(O)C(O)C(CO)O1 PNIWLNAGKUGXDO-UHFFFAOYSA-N 0.000 description 1
- WNGVUZWBXZKQES-YUMQZZPRSA-N Leu-Ala-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C)C(=O)NCC(O)=O WNGVUZWBXZKQES-YUMQZZPRSA-N 0.000 description 1
- BQSLGJHIAGOZCD-CIUDSAMLSA-N Leu-Ala-Ser Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@@H](CO)C(O)=O BQSLGJHIAGOZCD-CIUDSAMLSA-N 0.000 description 1
- KSZCCRIGNVSHFH-UWVGGRQHSA-N Leu-Arg-Gly Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCCNC(N)=N)C(=O)NCC(O)=O KSZCCRIGNVSHFH-UWVGGRQHSA-N 0.000 description 1
- YOZCKMXHBYKOMQ-IHRRRGAJSA-N Leu-Arg-Lys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCCN)C(=O)O)N YOZCKMXHBYKOMQ-IHRRRGAJSA-N 0.000 description 1
- WXHFZJFZWNCDNB-KKUMJFAQSA-N Leu-Asn-Tyr Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(N)=O)C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 WXHFZJFZWNCDNB-KKUMJFAQSA-N 0.000 description 1
- BPANDPNDMJHFEV-CIUDSAMLSA-N Leu-Asp-Ala Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C)C(O)=O BPANDPNDMJHFEV-CIUDSAMLSA-N 0.000 description 1
- YKNBJXOJTURHCU-DCAQKATOSA-N Leu-Asp-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@H](C(O)=O)CCCN=C(N)N YKNBJXOJTURHCU-DCAQKATOSA-N 0.000 description 1
- ILJREDZFPHTUIE-GUBZILKMSA-N Leu-Asp-Glu Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O ILJREDZFPHTUIE-GUBZILKMSA-N 0.000 description 1
- VQPPIMUZCZCOIL-GUBZILKMSA-N Leu-Gln-Ala Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](C)C(O)=O VQPPIMUZCZCOIL-GUBZILKMSA-N 0.000 description 1
- KAFOIVJDVSZUMD-DCAQKATOSA-N Leu-Gln-Gln Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CCC(N)=O)C(O)=O KAFOIVJDVSZUMD-DCAQKATOSA-N 0.000 description 1
- WIDZHJTYKYBLSR-DCAQKATOSA-N Leu-Glu-Glu Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O WIDZHJTYKYBLSR-DCAQKATOSA-N 0.000 description 1
- HVJVUYQWFYMGJS-GVXVVHGQSA-N Leu-Glu-Val Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O HVJVUYQWFYMGJS-GVXVVHGQSA-N 0.000 description 1
- VWHGTYCRDRBSFI-ZETCQYMHSA-N Leu-Gly-Gly Chemical compound CC(C)C[C@H](N)C(=O)NCC(=O)NCC(O)=O VWHGTYCRDRBSFI-ZETCQYMHSA-N 0.000 description 1
- VGPCJSXPPOQPBK-YUMQZZPRSA-N Leu-Gly-Ser Chemical compound CC(C)C[C@H](N)C(=O)NCC(=O)N[C@@H](CO)C(O)=O VGPCJSXPPOQPBK-YUMQZZPRSA-N 0.000 description 1
- HMDDEJADNKQTBR-BZSNNMDCSA-N Leu-His-Tyr Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O HMDDEJADNKQTBR-BZSNNMDCSA-N 0.000 description 1
- USLNHQZCDQJBOV-ZPFDUUQYSA-N Leu-Ile-Asn Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(N)=O)C(O)=O USLNHQZCDQJBOV-ZPFDUUQYSA-N 0.000 description 1
- QNBVTHNJGCOVFA-AVGNSLFASA-N Leu-Leu-Glu Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)N[C@H](C(O)=O)CCC(O)=O QNBVTHNJGCOVFA-AVGNSLFASA-N 0.000 description 1
- YOKVEHGYYQEQOP-QWRGUYRKSA-N Leu-Leu-Gly Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CC(C)C)C(=O)NCC(O)=O YOKVEHGYYQEQOP-QWRGUYRKSA-N 0.000 description 1
- LXKNSJLSGPNHSK-KKUMJFAQSA-N Leu-Leu-Lys Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCCCN)C(=O)O)N LXKNSJLSGPNHSK-KKUMJFAQSA-N 0.000 description 1
- KPYAOIVPJKPIOU-KKUMJFAQSA-N Leu-Lys-Lys Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(O)=O KPYAOIVPJKPIOU-KKUMJFAQSA-N 0.000 description 1
- FZMNAYBEFGZEIF-AVGNSLFASA-N Leu-Met-Met Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CCSC)C(=O)O)N FZMNAYBEFGZEIF-AVGNSLFASA-N 0.000 description 1
- DPURXCQCHSQPAN-AVGNSLFASA-N Leu-Pro-Pro Chemical compound CC(C)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N1[C@H](C(O)=O)CCC1 DPURXCQCHSQPAN-AVGNSLFASA-N 0.000 description 1
- UCXQIIIFOOGYEM-ULQDDVLXSA-N Leu-Pro-Tyr Chemical compound CC(C)C[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@H](C(O)=O)CC1=CC=C(O)C=C1 UCXQIIIFOOGYEM-ULQDDVLXSA-N 0.000 description 1
- IRMLZWSRWSGTOP-CIUDSAMLSA-N Leu-Ser-Ala Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](CO)C(=O)N[C@@H](C)C(O)=O IRMLZWSRWSGTOP-CIUDSAMLSA-N 0.000 description 1
- AKVBOOKXVAMKSS-GUBZILKMSA-N Leu-Ser-Gln Chemical compound [H]N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCC(N)=O)C(O)=O AKVBOOKXVAMKSS-GUBZILKMSA-N 0.000 description 1
- JGKHAFUAPZCCDU-BZSNNMDCSA-N Leu-Tyr-Leu Chemical compound CC(C)C[C@H]([NH3+])C(=O)N[C@H](C(=O)N[C@@H](CC(C)C)C([O-])=O)CC1=CC=C(O)C=C1 JGKHAFUAPZCCDU-BZSNNMDCSA-N 0.000 description 1
- FBNPMTNBFFAMMH-AVGNSLFASA-N Leu-Val-Arg Chemical compound CC(C)C[C@H](N)C(=O)N[C@@H](C(C)C)C(=O)N[C@H](C(O)=O)CCCN=C(N)N FBNPMTNBFFAMMH-AVGNSLFASA-N 0.000 description 1
- FBNPMTNBFFAMMH-UHFFFAOYSA-N Leu-Val-Arg Natural products CC(C)CC(N)C(=O)NC(C(C)C)C(=O)NC(C(O)=O)CCCN=C(N)N FBNPMTNBFFAMMH-UHFFFAOYSA-N 0.000 description 1
- ROHFNLRQFUQHCH-UHFFFAOYSA-N Leucine Natural products CC(C)CC(N)C(O)=O ROHFNLRQFUQHCH-UHFFFAOYSA-N 0.000 description 1
- NTSPQIONFJUMJV-AVGNSLFASA-N Lys-Arg-Val Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCNC(N)=N)C(=O)N[C@@H](C(C)C)C(O)=O NTSPQIONFJUMJV-AVGNSLFASA-N 0.000 description 1
- NRQRKMYZONPCTM-CIUDSAMLSA-N Lys-Asp-Ser Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CO)C(O)=O NRQRKMYZONPCTM-CIUDSAMLSA-N 0.000 description 1
- GKFNXYMAMKJSKD-NHCYSSNCSA-N Lys-Asp-Val Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O GKFNXYMAMKJSKD-NHCYSSNCSA-N 0.000 description 1
- QQUJSUFWEDZQQY-AVGNSLFASA-N Lys-Gln-Lys Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@H](C(O)=O)CCCCN QQUJSUFWEDZQQY-AVGNSLFASA-N 0.000 description 1
- ISHNZELVUVPCHY-ZETCQYMHSA-N Lys-Gly-Gly Chemical compound NCCCC[C@H](N)C(=O)NCC(=O)NCC(O)=O ISHNZELVUVPCHY-ZETCQYMHSA-N 0.000 description 1
- QOJDBRUCOXQSSK-AJNGGQMLSA-N Lys-Ile-Lys Chemical compound NCCCC[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CCCCN)C(O)=O QOJDBRUCOXQSSK-AJNGGQMLSA-N 0.000 description 1
- VMTYLUGCXIEDMV-QWRGUYRKSA-N Lys-Leu-Gly Chemical compound OC(=O)CNC(=O)[C@H](CC(C)C)NC(=O)[C@@H](N)CCCCN VMTYLUGCXIEDMV-QWRGUYRKSA-N 0.000 description 1
- XIZQPFCRXLUNMK-BZSNNMDCSA-N Lys-Leu-Phe Chemical compound CC(C)C[C@@H](C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)NC(=O)[C@H](CCCCN)N XIZQPFCRXLUNMK-BZSNNMDCSA-N 0.000 description 1
- ZJWIXBZTAAJERF-IHRRRGAJSA-N Lys-Lys-Arg Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](CCCCN)C(=O)N[C@H](C(O)=O)CCCN=C(N)N ZJWIXBZTAAJERF-IHRRRGAJSA-N 0.000 description 1
- HVAUKHLDSDDROB-KKUMJFAQSA-N Lys-Lys-Leu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC(C)C)C(O)=O HVAUKHLDSDDROB-KKUMJFAQSA-N 0.000 description 1
- BEGQVWUZFXLNHZ-IHPCNDPISA-N Lys-Lys-Trp Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@H](CCCCN)NC(=O)[C@@H](N)CCCCN)C(O)=O)=CNC2=C1 BEGQVWUZFXLNHZ-IHPCNDPISA-N 0.000 description 1
- VSTNAUBHKQPVJX-IHRRRGAJSA-N Lys-Met-Leu Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC(C)C)C(O)=O VSTNAUBHKQPVJX-IHRRRGAJSA-N 0.000 description 1
- JYVCOTWSRGFABJ-DCAQKATOSA-N Lys-Met-Ser Chemical compound CSCC[C@@H](C(=O)N[C@@H](CO)C(=O)O)NC(=O)[C@H](CCCCN)N JYVCOTWSRGFABJ-DCAQKATOSA-N 0.000 description 1
- MSSJJDVQTFTLIF-KBPBESRZSA-N Lys-Phe-Gly Chemical compound NCCCC[C@H](N)C(=O)N[C@@H](Cc1ccccc1)C(=O)NCC(O)=O MSSJJDVQTFTLIF-KBPBESRZSA-N 0.000 description 1
- SVSQSPICRKBMSZ-SRVKXCTJSA-N Lys-Pro-Gln Chemical compound [H]N[C@@H](CCCCN)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(N)=O)C(O)=O SVSQSPICRKBMSZ-SRVKXCTJSA-N 0.000 description 1
- ZVZRQKJOQQAFCF-ULQDDVLXSA-N Lys-Tyr-Arg Chemical compound [H]N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O ZVZRQKJOQQAFCF-ULQDDVLXSA-N 0.000 description 1
- DRRXXZBXDMLGFC-IHRRRGAJSA-N Lys-Val-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](C(C)C)NC(=O)[C@@H](N)CCCCN DRRXXZBXDMLGFC-IHRRRGAJSA-N 0.000 description 1
- 239000004472 Lysine Substances 0.000 description 1
- PWHULOQIROXLJO-UHFFFAOYSA-N Manganese Chemical compound [Mn] PWHULOQIROXLJO-UHFFFAOYSA-N 0.000 description 1
- ONGCSGVHCSAATF-CIUDSAMLSA-N Met-Ala-Glu Chemical compound CSCC[C@H](N)C(=O)N[C@@H](C)C(=O)N[C@H](C(O)=O)CCC(O)=O ONGCSGVHCSAATF-CIUDSAMLSA-N 0.000 description 1
- XMMWDTUFTZMQFD-GMOBBJLQSA-N Met-Asp-Ile Chemical compound CC[C@H](C)[C@@H](C(O)=O)NC(=O)[C@H](CC(O)=O)NC(=O)[C@@H](N)CCSC XMMWDTUFTZMQFD-GMOBBJLQSA-N 0.000 description 1
- MYKLINMAGAIRPJ-CIUDSAMLSA-N Met-Gln-Asn Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC(N)=O)C(O)=O MYKLINMAGAIRPJ-CIUDSAMLSA-N 0.000 description 1
- QMIXOTQHYHOUJP-KKUMJFAQSA-N Met-Gln-Tyr Chemical compound CSCC[C@@H](C(=O)N[C@@H](CCC(=O)N)C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)O)N QMIXOTQHYHOUJP-KKUMJFAQSA-N 0.000 description 1
- KQBJYJXPZBNEIK-DCAQKATOSA-N Met-Glu-Arg Chemical compound CSCC[C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@H](C(O)=O)CCCNC(N)=N KQBJYJXPZBNEIK-DCAQKATOSA-N 0.000 description 1
- ZIIMORLEZLVRIP-SRVKXCTJSA-N Met-Leu-Gln Chemical compound [H]N[C@@H](CCSC)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CCC(N)=O)C(O)=O ZIIMORLEZLVRIP-SRVKXCTJSA-N 0.000 description 1
- NHXXGBXJTLRGJI-GUBZILKMSA-N Met-Pro-Ser Chemical compound [H]N[C@@H](CCSC)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CO)C(O)=O NHXXGBXJTLRGJI-GUBZILKMSA-N 0.000 description 1
- LUYURUYVNYGKGM-RCWTZXSCSA-N Met-Pro-Thr Chemical compound CSCC[C@H](N)C(=O)N1CCC[C@H]1C(=O)N[C@@H]([C@@H](C)O)C(O)=O LUYURUYVNYGKGM-RCWTZXSCSA-N 0.000 description 1
- BJPQKNHZHUCQNQ-SRVKXCTJSA-N Met-Pro-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@@H]1CCCN1C(=O)[C@H](CCSC)N BJPQKNHZHUCQNQ-SRVKXCTJSA-N 0.000 description 1
- 108090000157 Metallothionein Proteins 0.000 description 1
- 241000588655 Moraxella catarrhalis Species 0.000 description 1
- 241001478293 Moraxella cuniculi Species 0.000 description 1
- 108010046220 N-Acetylgalactosaminyltransferases Proteins 0.000 description 1
- 102000007524 N-Acetylgalactosaminyltransferases Human genes 0.000 description 1
- 125000003047 N-acetyl group Chemical group 0.000 description 1
- SQVRNKJHWKZAKO-PFQGKNLYSA-N N-acetyl-beta-neuraminic acid Chemical compound CC(=O)N[C@@H]1[C@@H](O)C[C@@](O)(C(O)=O)O[C@H]1[C@H](O)[C@H](O)CO SQVRNKJHWKZAKO-PFQGKNLYSA-N 0.000 description 1
- 241000588678 Neisseria animalis Species 0.000 description 1
- 241000588675 Neisseria canis Species 0.000 description 1
- 241000588654 Neisseria cinerea Species 0.000 description 1
- 241000293864 Neisseria elongata subsp. nitroreducens Species 0.000 description 1
- 241000588651 Neisseria flavescens Species 0.000 description 1
- 241000588649 Neisseria lactamica Species 0.000 description 1
- 241000588674 Neisseria macacae Species 0.000 description 1
- 241000588650 Neisseria meningitidis Species 0.000 description 1
- 241000588660 Neisseria polysaccharea Species 0.000 description 1
- 241000588645 Neisseria sicca Species 0.000 description 1
- 241001136170 Neisseria subflava Species 0.000 description 1
- 101100342977 Neurospora crassa (strain ATCC 24698 / 74-OR23-1A / CBS 708.71 / DSM 1257 / FGSC 987) leu-1 gene Proteins 0.000 description 1
- 108091034117 Oligonucleotide Proteins 0.000 description 1
- 238000012408 PCR amplification Methods 0.000 description 1
- 108091005804 Peptidases Proteins 0.000 description 1
- 108010033276 Peptide Fragments Proteins 0.000 description 1
- 102000007079 Peptide Fragments Human genes 0.000 description 1
- FPTXMUIBLMGTQH-ONGXEEELSA-N Phe-Ala-Gly Chemical compound OC(=O)CNC(=O)[C@H](C)NC(=O)[C@@H](N)CC1=CC=CC=C1 FPTXMUIBLMGTQH-ONGXEEELSA-N 0.000 description 1
- SEPNOAFMZLLCEW-UBHSHLNASA-N Phe-Ala-Val Chemical compound N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](C)C(=O)N[C@@H](C(C)C)C(=O)O SEPNOAFMZLLCEW-UBHSHLNASA-N 0.000 description 1
- QCHNRQQVLJYDSI-DLOVCJGASA-N Phe-Asn-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H](CC(N)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 QCHNRQQVLJYDSI-DLOVCJGASA-N 0.000 description 1
- UNLYPPYNDXHGDG-IHRRRGAJSA-N Phe-Gln-Glu Chemical compound OC(=O)CC[C@@H](C(O)=O)NC(=O)[C@H](CCC(N)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 UNLYPPYNDXHGDG-IHRRRGAJSA-N 0.000 description 1
- GDBOREPXIRKSEQ-FHWLQOOXSA-N Phe-Gln-Phe Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCC(N)=O)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O GDBOREPXIRKSEQ-FHWLQOOXSA-N 0.000 description 1
- MPFGIYLYWUCSJG-AVGNSLFASA-N Phe-Glu-Asp Chemical compound OC(=O)C[C@@H](C(O)=O)NC(=O)[C@H](CCC(O)=O)NC(=O)[C@@H](N)CC1=CC=CC=C1 MPFGIYLYWUCSJG-AVGNSLFASA-N 0.000 description 1
- SPXWRYVHOZVYBU-ULQDDVLXSA-N Phe-His-Val Chemical compound CC(C)[C@@H](C(=O)O)NC(=O)[C@H](CC1=CN=CN1)NC(=O)[C@H](CC2=CC=CC=C2)N SPXWRYVHOZVYBU-ULQDDVLXSA-N 0.000 description 1
- VZFPYFRVHMSSNA-JURCDPSOSA-N Phe-Ile-Ala Chemical compound OC(=O)[C@H](C)NC(=O)[C@H]([C@@H](C)CC)NC(=O)[C@@H](N)CC1=CC=CC=C1 VZFPYFRVHMSSNA-JURCDPSOSA-N 0.000 description 1
- KXUZHWXENMYOHC-QEJZJMRPSA-N Phe-Leu-Ala Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(O)=O KXUZHWXENMYOHC-QEJZJMRPSA-N 0.000 description 1
- YCCUXNNKXDGMAM-KKUMJFAQSA-N Phe-Leu-Ser Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CO)C(O)=O YCCUXNNKXDGMAM-KKUMJFAQSA-N 0.000 description 1
- SZYBZVANEAOIPE-UBHSHLNASA-N Phe-Met-Ala Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](C)C(O)=O SZYBZVANEAOIPE-UBHSHLNASA-N 0.000 description 1
- WWPAHTZOWURIMR-ULQDDVLXSA-N Phe-Pro-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@@H]1CCCN1C(=O)[C@@H](N)CC1=CC=CC=C1 WWPAHTZOWURIMR-ULQDDVLXSA-N 0.000 description 1
- XNQMZHLAYFWSGJ-HTUGSXCWSA-N Phe-Thr-Glu Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCC(O)=O)C(O)=O XNQMZHLAYFWSGJ-HTUGSXCWSA-N 0.000 description 1
- VGTJSEYTVMAASM-RPTUDFQQSA-N Phe-Thr-Tyr Chemical compound [H]N[C@@H](CC1=CC=CC=C1)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O VGTJSEYTVMAASM-RPTUDFQQSA-N 0.000 description 1
- APECKGGXAXNFLL-RNXOBYDBSA-N Phe-Trp-Tyr Chemical compound C([C@H](N)C(=O)N[C@@H](CC=1C2=CC=CC=C2NC=1)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)C1=CC=CC=C1 APECKGGXAXNFLL-RNXOBYDBSA-N 0.000 description 1
- IFMDQWDAJUMMJC-DCAQKATOSA-N Pro-Ala-Leu Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](C)C(=O)N[C@@H](CC(C)C)C(O)=O IFMDQWDAJUMMJC-DCAQKATOSA-N 0.000 description 1
- LCRSGSIRKLXZMZ-BPNCWPANSA-N Pro-Ala-Tyr Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O LCRSGSIRKLXZMZ-BPNCWPANSA-N 0.000 description 1
- SFECXGVELZFBFJ-VEVYYDQMSA-N Pro-Asp-Thr Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H]([C@@H](C)O)C(O)=O SFECXGVELZFBFJ-VEVYYDQMSA-N 0.000 description 1
- MGDFPGCFVJFITQ-CIUDSAMLSA-N Pro-Glu-Asp Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC(O)=O)C(O)=O MGDFPGCFVJFITQ-CIUDSAMLSA-N 0.000 description 1
- SRBFGSGDNNQABI-FHWLQOOXSA-N Pro-Leu-Trp Chemical compound N([C@@H](CC(C)C)C(=O)N[C@@H](CC=1C2=CC=CC=C2NC=1)C(O)=O)C(=O)[C@@H]1CCCN1 SRBFGSGDNNQABI-FHWLQOOXSA-N 0.000 description 1
- VWHJZETTZDAGOM-XUXIUFHCSA-N Pro-Lys-Ile Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](CCCCN)C(=O)N[C@@H]([C@@H](C)CC)C(O)=O VWHJZETTZDAGOM-XUXIUFHCSA-N 0.000 description 1
- AJBQTGZIZQXBLT-STQMWFEESA-N Pro-Phe-Gly Chemical compound C([C@@H](C(=O)NCC(=O)O)NC(=O)[C@H]1NCCC1)C1=CC=CC=C1 AJBQTGZIZQXBLT-STQMWFEESA-N 0.000 description 1
- RJTUIDFUUHPJMP-FHWLQOOXSA-N Pro-Trp-His Chemical compound C1C[C@H](NC1)C(=O)N[C@@H](CC2=CNC3=CC=CC=C32)C(=O)N[C@@H](CC4=CN=CN4)C(=O)O RJTUIDFUUHPJMP-FHWLQOOXSA-N 0.000 description 1
- PGSWNLRYYONGPE-JYJNAYRXSA-N Pro-Val-Tyr Chemical compound [H]N1CCC[C@H]1C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O PGSWNLRYYONGPE-JYJNAYRXSA-N 0.000 description 1
- ONIBWKKTOPOVIA-UHFFFAOYSA-N Proline Natural products OC(=O)C1CCCN1 ONIBWKKTOPOVIA-UHFFFAOYSA-N 0.000 description 1
- 239000004365 Protease Substances 0.000 description 1
- 108010076504 Protein Sorting Signals Proteins 0.000 description 1
- 241000589516 Pseudomonas Species 0.000 description 1
- 241000589517 Pseudomonas aeruginosa Species 0.000 description 1
- 108020005091 Replication Origin Proteins 0.000 description 1
- 102100037486 Reverse transcriptase/ribonuclease H Human genes 0.000 description 1
- 241000714474 Rous sarcoma virus Species 0.000 description 1
- 229920005654 Sephadex Polymers 0.000 description 1
- 239000012507 Sephadex™ Substances 0.000 description 1
- 229920002684 Sepharose Polymers 0.000 description 1
- ZUGXSSFMTXKHJS-ZLUOBGJFSA-N Ser-Ala-Ala Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C)C(=O)N[C@@H](C)C(O)=O ZUGXSSFMTXKHJS-ZLUOBGJFSA-N 0.000 description 1
- FIXILCYTSAUERA-FXQIFTODSA-N Ser-Ala-Arg Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O FIXILCYTSAUERA-FXQIFTODSA-N 0.000 description 1
- WTUJZHKANPDPIN-CIUDSAMLSA-N Ser-Ala-Lys Chemical compound C[C@@H](C(=O)N[C@@H](CCCCN)C(=O)O)NC(=O)[C@H](CO)N WTUJZHKANPDPIN-CIUDSAMLSA-N 0.000 description 1
- UOLGINIHBRIECN-FXQIFTODSA-N Ser-Glu-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CCC(O)=O)C(O)=O UOLGINIHBRIECN-FXQIFTODSA-N 0.000 description 1
- UICKAKRRRBTILH-GUBZILKMSA-N Ser-Glu-His Chemical compound C1=C(NC=N1)C[C@@H](C(=O)O)NC(=O)[C@H](CCC(=O)O)NC(=O)[C@H](CO)N UICKAKRRRBTILH-GUBZILKMSA-N 0.000 description 1
- FYUIFUJFNCLUIX-XVYDVKMFSA-N Ser-His-Ala Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC1=CNC=N1)C(=O)N[C@@H](C)C(O)=O FYUIFUJFNCLUIX-XVYDVKMFSA-N 0.000 description 1
- FUMGHWDRRFCKEP-CIUDSAMLSA-N Ser-Leu-Ala Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](C)C(O)=O FUMGHWDRRFCKEP-CIUDSAMLSA-N 0.000 description 1
- MUJQWSAWLLRJCE-KATARQTJSA-N Ser-Leu-Thr Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H]([C@@H](C)O)C(O)=O MUJQWSAWLLRJCE-KATARQTJSA-N 0.000 description 1
- QJKPECIAWNNKIT-KKUMJFAQSA-N Ser-Lys-Tyr Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O QJKPECIAWNNKIT-KKUMJFAQSA-N 0.000 description 1
- RWDVVSKYZBNDCO-MELADBBJSA-N Ser-Phe-Pro Chemical compound C1C[C@@H](N(C1)C(=O)[C@H](CC2=CC=CC=C2)NC(=O)[C@H](CO)N)C(=O)O RWDVVSKYZBNDCO-MELADBBJSA-N 0.000 description 1
- XJDMUQCLVSCRSJ-VZFHVOOUSA-N Ser-Thr-Ala Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](C)C(O)=O XJDMUQCLVSCRSJ-VZFHVOOUSA-N 0.000 description 1
- BEBVVQPDSHHWQL-NRPADANISA-N Ser-Val-Glu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CCC(O)=O)C(O)=O BEBVVQPDSHHWQL-NRPADANISA-N 0.000 description 1
- YEDSOSIKVUMIJE-DCAQKATOSA-N Ser-Val-Leu Chemical compound [H]N[C@@H](CO)C(=O)N[C@@H](C(C)C)C(=O)N[C@@H](CC(C)C)C(O)=O YEDSOSIKVUMIJE-DCAQKATOSA-N 0.000 description 1
- MTCFGRXMJLQNBG-UHFFFAOYSA-N Serine Natural products OCC(N)C(O)=O MTCFGRXMJLQNBG-UHFFFAOYSA-N 0.000 description 1
- 102000003838 Sialyltransferases Human genes 0.000 description 1
- 108090000141 Sialyltransferases Proteins 0.000 description 1
- NFMPFBCXABPALN-OWLDWWDNSA-N Thr-Ala-Trp Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](C)C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)O)N)O NFMPFBCXABPALN-OWLDWWDNSA-N 0.000 description 1
- PKXHGEXFMIZSER-QTKMDUPCSA-N Thr-Arg-His Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N)O PKXHGEXFMIZSER-QTKMDUPCSA-N 0.000 description 1
- VXMHQKHDKCATDV-VEVYYDQMSA-N Thr-Asp-Arg Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O VXMHQKHDKCATDV-VEVYYDQMSA-N 0.000 description 1
- NLSNVZAREYQMGR-HJGDQZAQSA-N Thr-Asp-Leu Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](CC(C)C)C(O)=O NLSNVZAREYQMGR-HJGDQZAQSA-N 0.000 description 1
- JXKMXEBNZCKSDY-JIOCBJNQSA-N Thr-Asp-Pro Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)N1CCC[C@@H]1C(=O)O)N)O JXKMXEBNZCKSDY-JIOCBJNQSA-N 0.000 description 1
- ZUUDNCOCILSYAM-KKHAAJSZSA-N Thr-Asp-Val Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC(O)=O)C(=O)N[C@@H](C(C)C)C(O)=O ZUUDNCOCILSYAM-KKHAAJSZSA-N 0.000 description 1
- HPQHHRLWSAMMKG-KATARQTJSA-N Thr-Lys-Cys Chemical compound C[C@H]([C@@H](C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CS)C(=O)O)N)O HPQHHRLWSAMMKG-KATARQTJSA-N 0.000 description 1
- QFEYTTHKPSOFLV-OSUNSFLBSA-N Thr-Met-Ile Chemical compound CC[C@H](C)[C@@H](C(=O)O)NC(=O)[C@H](CCSC)NC(=O)[C@H]([C@@H](C)O)N QFEYTTHKPSOFLV-OSUNSFLBSA-N 0.000 description 1
- CGCMNOIQVAXYMA-UNQGMJICSA-N Thr-Met-Phe Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC1=CC=CC=C1)C(O)=O CGCMNOIQVAXYMA-UNQGMJICSA-N 0.000 description 1
- GRIUMVXCJDKVPI-IZPVPAKOSA-N Thr-Thr-Tyr Chemical compound [H]N[C@@H]([C@@H](C)O)C(=O)N[C@@H]([C@@H](C)O)C(=O)N[C@@H](CC1=CC=C(O)C=C1)C(O)=O GRIUMVXCJDKVPI-IZPVPAKOSA-N 0.000 description 1
- AYFVYJQAPQTCCC-UHFFFAOYSA-N Threonine Natural products CC(O)C(N)C(O)=O AYFVYJQAPQTCCC-UHFFFAOYSA-N 0.000 description 1
- 239000004473 Threonine Substances 0.000 description 1
- 108090000190 Thrombin Proteins 0.000 description 1
- NLLARHRWSFNEMH-NUTKFTJISA-N Trp-Lys-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CCCCN)NC(=O)[C@H](CC1=CNC2=CC=CC=C21)N NLLARHRWSFNEMH-NUTKFTJISA-N 0.000 description 1
- QIVBCDIJIAJPQS-UHFFFAOYSA-N Tryptophan Natural products C1=CC=C2C(CC(N)C(O)=O)=CNC2=C1 QIVBCDIJIAJPQS-UHFFFAOYSA-N 0.000 description 1
- TVOGEPLDNYTAHD-CQDKDKBSSA-N Tyr-Ala-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](C)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 TVOGEPLDNYTAHD-CQDKDKBSSA-N 0.000 description 1
- GFZQWWDXJVGEMW-ULQDDVLXSA-N Tyr-Arg-Lys Chemical compound C1=CC(=CC=C1C[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCCCN)C(=O)O)N)O GFZQWWDXJVGEMW-ULQDDVLXSA-N 0.000 description 1
- XHALUUQSNXSPLP-UFYCRDLUSA-N Tyr-Arg-Phe Chemical compound C([C@H](N)C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CC=1C=CC=CC=1)C(O)=O)C1=CC=C(O)C=C1 XHALUUQSNXSPLP-UFYCRDLUSA-N 0.000 description 1
- FQNUWOHNGJWNLM-QWRGUYRKSA-N Tyr-Cys-Gly Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CS)C(=O)NCC(O)=O FQNUWOHNGJWNLM-QWRGUYRKSA-N 0.000 description 1
- NJLQMKZSXYQRTO-FHWLQOOXSA-N Tyr-Glu-Tyr Chemical compound C([C@H](N)C(=O)N[C@@H](CCC(O)=O)C(=O)N[C@@H](CC=1C=CC(O)=CC=1)C(O)=O)C1=CC=C(O)C=C1 NJLQMKZSXYQRTO-FHWLQOOXSA-N 0.000 description 1
- BSCBBPKDVOZICB-KKUMJFAQSA-N Tyr-Leu-Asp Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CC(C)C)C(=O)N[C@@H](CC(O)=O)C(O)=O BSCBBPKDVOZICB-KKUMJFAQSA-N 0.000 description 1
- QKXAEWMHAAVVGS-KKUMJFAQSA-N Tyr-Pro-Glu Chemical compound N[C@@H](Cc1ccc(O)cc1)C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(O)=O)C(O)=O QKXAEWMHAAVVGS-KKUMJFAQSA-N 0.000 description 1
- RWOKVQUCENPXGE-IHRRRGAJSA-N Tyr-Ser-Arg Chemical compound [H]N[C@@H](CC1=CC=C(O)C=C1)C(=O)N[C@@H](CO)C(=O)N[C@@H](CCCNC(N)=N)C(O)=O RWOKVQUCENPXGE-IHRRRGAJSA-N 0.000 description 1
- MQGGXGKQSVEQHR-KKUMJFAQSA-N Tyr-Ser-Leu Chemical compound CC(C)C[C@@H](C(O)=O)NC(=O)[C@H](CO)NC(=O)[C@@H](N)CC1=CC=C(O)C=C1 MQGGXGKQSVEQHR-KKUMJFAQSA-N 0.000 description 1
- 108010090473 UDP-N-acetylglucosamine-peptide beta-N-acetylglucosaminyltransferase Proteins 0.000 description 1
- DRTQHJPVMGBUCF-XVFCMESISA-N Uridine Chemical compound O[C@@H]1[C@H](O)[C@@H](CO)O[C@H]1N1C(=O)NC(=O)C=C1 DRTQHJPVMGBUCF-XVFCMESISA-N 0.000 description 1
- DJJCXFVJDGTHFX-UHFFFAOYSA-N Uridinemonophosphate Natural products OC1C(O)C(COP(O)(O)=O)OC1N1C(=O)NC(=O)C=C1 DJJCXFVJDGTHFX-UHFFFAOYSA-N 0.000 description 1
- YFOCMOVJBQDBCE-NRPADANISA-N Val-Ala-Glu Chemical compound C[C@@H](C(=O)N[C@@H](CCC(=O)O)C(=O)O)NC(=O)[C@H](C(C)C)N YFOCMOVJBQDBCE-NRPADANISA-N 0.000 description 1
- JIODCDXKCJRMEH-NHCYSSNCSA-N Val-Arg-Gln Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N[C@@H](CCC(=O)N)C(=O)O)N JIODCDXKCJRMEH-NHCYSSNCSA-N 0.000 description 1
- CVUDMNSZAIZFAE-TUAOUCFPSA-N Val-Arg-Pro Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCCN=C(N)N)C(=O)N1CCC[C@@H]1C(=O)O)N CVUDMNSZAIZFAE-TUAOUCFPSA-N 0.000 description 1
- CVUDMNSZAIZFAE-UHFFFAOYSA-N Val-Arg-Pro Natural products NC(N)=NCCCC(NC(=O)C(N)C(C)C)C(=O)N1CCCC1C(O)=O CVUDMNSZAIZFAE-UHFFFAOYSA-N 0.000 description 1
- ISERLACIZUGCDX-ZKWXMUAHSA-N Val-Asp-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CC(=O)O)NC(=O)[C@H](C(C)C)N ISERLACIZUGCDX-ZKWXMUAHSA-N 0.000 description 1
- CJDZKZFMAXGUOJ-IHRRRGAJSA-N Val-Cys-Tyr Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CS)C(=O)N[C@@H](CC1=CC=C(C=C1)O)C(=O)O)N CJDZKZFMAXGUOJ-IHRRRGAJSA-N 0.000 description 1
- WFENBJPLZMPVAX-XVKPBYJWSA-N Val-Gly-Glu Chemical compound CC(C)[C@H](N)C(=O)NCC(=O)N[C@H](C(O)=O)CCC(O)=O WFENBJPLZMPVAX-XVKPBYJWSA-N 0.000 description 1
- LKUDRJSNRWVGMS-QSFUFRPTSA-N Val-Ile-Asp Chemical compound CC[C@H](C)[C@@H](C(=O)N[C@@H](CC(=O)O)C(=O)O)NC(=O)[C@H](C(C)C)N LKUDRJSNRWVGMS-QSFUFRPTSA-N 0.000 description 1
- FTKXYXACXYOHND-XUXIUFHCSA-N Val-Ile-Leu Chemical compound CC(C)[C@H](N)C(=O)N[C@@H]([C@@H](C)CC)C(=O)N[C@@H](CC(C)C)C(O)=O FTKXYXACXYOHND-XUXIUFHCSA-N 0.000 description 1
- RFKJNTRMXGCKFE-FHWLQOOXSA-N Val-Leu-Trp Chemical compound C1=CC=C2C(C[C@H](NC(=O)[C@@H](NC(=O)[C@@H](N)C(C)C)CC(C)C)C(O)=O)=CNC2=C1 RFKJNTRMXGCKFE-FHWLQOOXSA-N 0.000 description 1
- WBAJDGWKRIHOAC-GVXVVHGQSA-N Val-Lys-Gln Chemical compound [H]N[C@@H](C(C)C)C(=O)N[C@@H](CCCCN)C(=O)N[C@@H](CCC(N)=O)C(O)=O WBAJDGWKRIHOAC-GVXVVHGQSA-N 0.000 description 1
- MBGFDZDWMDLXHQ-GUBZILKMSA-N Val-Met-Ala Chemical compound C[C@@H](C(=O)O)NC(=O)[C@H](CCSC)NC(=O)[C@H](C(C)C)N MBGFDZDWMDLXHQ-GUBZILKMSA-N 0.000 description 1
- OJPRSVJGNCAKQX-SRVKXCTJSA-N Val-Met-Arg Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CCCN=C(N)N)C(=O)O)N OJPRSVJGNCAKQX-SRVKXCTJSA-N 0.000 description 1
- RSGHLMMKXJGCMK-JYJNAYRXSA-N Val-Met-Phe Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CCSC)C(=O)N[C@@H](CC1=CC=CC=C1)C(=O)O)N RSGHLMMKXJGCMK-JYJNAYRXSA-N 0.000 description 1
- LJSZPMSUYKKKCP-UBHSHLNASA-N Val-Phe-Ala Chemical compound CC(C)[C@H](N)C(=O)N[C@H](C(=O)N[C@@H](C)C(O)=O)CC1=CC=CC=C1 LJSZPMSUYKKKCP-UBHSHLNASA-N 0.000 description 1
- MHHAWNPHDLCPLF-ULQDDVLXSA-N Val-Phe-Lys Chemical compound NCCCC[C@@H](C(O)=O)NC(=O)[C@@H](NC(=O)[C@@H](N)C(C)C)CC1=CC=CC=C1 MHHAWNPHDLCPLF-ULQDDVLXSA-N 0.000 description 1
- RYQUMYBMOJYYDK-NHCYSSNCSA-N Val-Pro-Glu Chemical compound CC(C)[C@@H](C(=O)N1CCC[C@H]1C(=O)N[C@@H](CCC(=O)O)C(=O)O)N RYQUMYBMOJYYDK-NHCYSSNCSA-N 0.000 description 1
- SJRUJQFQVLMZFW-WPRPVWTQSA-N Val-Pro-Gly Chemical compound CC(C)[C@H](N)C(=O)N1CCC[C@H]1C(=O)NCC(O)=O SJRUJQFQVLMZFW-WPRPVWTQSA-N 0.000 description 1
- KRAHMIJVUPUOTQ-DCAQKATOSA-N Val-Ser-His Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CO)C(=O)N[C@@H](CC1=CN=CN1)C(=O)O)N KRAHMIJVUPUOTQ-DCAQKATOSA-N 0.000 description 1
- GBIUHAYJGWVNLN-UHFFFAOYSA-N Val-Ser-Pro Natural products CC(C)C(N)C(=O)NC(CO)C(=O)N1CCCC1C(O)=O GBIUHAYJGWVNLN-UHFFFAOYSA-N 0.000 description 1
- QHSSPPHOHJSTML-HOCLYGCPSA-N Val-Trp-Gly Chemical compound CC(C)[C@@H](C(=O)N[C@@H](CC1=CNC2=CC=CC=C21)C(=O)NCC(=O)O)N QHSSPPHOHJSTML-HOCLYGCPSA-N 0.000 description 1
- AEFJNECXZCODJM-UWVGGRQHSA-N Val-Val-Gly Chemical compound CC(C)[C@H]([NH3+])C(=O)N[C@@H](C(C)C)C(=O)NCC([O-])=O AEFJNECXZCODJM-UWVGGRQHSA-N 0.000 description 1
- KZSNJWFQEVHDMF-UHFFFAOYSA-N Valine Natural products CC(C)C(N)C(O)=O KZSNJWFQEVHDMF-UHFFFAOYSA-N 0.000 description 1
- JLCPHMBAVCMARE-UHFFFAOYSA-N [3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[3-[[3-[[3-[[3-[[3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-[[5-(2-amino-6-oxo-1H-purin-9-yl)-3-hydroxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxyoxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(5-methyl-2,4-dioxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(6-aminopurin-9-yl)oxolan-2-yl]methoxy-hydroxyphosphoryl]oxy-5-(4-amino-2-oxopyrimidin-1-yl)oxolan-2-yl]methyl [5-(6-aminopurin-9-yl)-2-(hydroxymethyl)oxolan-3-yl] hydrogen phosphate Polymers Cc1cn(C2CC(OP(O)(=O)OCC3OC(CC3OP(O)(=O)OCC3OC(CC3O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c3nc(N)[nH]c4=O)C(COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3COP(O)(=O)OC3CC(OC3CO)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3ccc(N)nc3=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cc(C)c(=O)[nH]c3=O)n3cc(C)c(=O)[nH]c3=O)n3ccc(N)nc3=O)n3cc(C)c(=O)[nH]c3=O)n3cnc4c3nc(N)[nH]c4=O)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)n3cnc4c(N)ncnc34)O2)c(=O)[nH]c1=O JLCPHMBAVCMARE-UHFFFAOYSA-N 0.000 description 1
- 230000002378 acidificating effect Effects 0.000 description 1
- 238000000246 agarose gel electrophoresis Methods 0.000 description 1
- 235000004279 alanine Nutrition 0.000 description 1
- 108010024078 alanyl-glycyl-serine Proteins 0.000 description 1
- 108010045350 alanyl-tyrosyl-alanine Proteins 0.000 description 1
- 108010041407 alanylaspartic acid Proteins 0.000 description 1
- 108010070944 alanylhistidine Proteins 0.000 description 1
- KOSRFJWDECSPRO-UHFFFAOYSA-N alpha-L-glutamyl-L-glutamic acid Natural products OC(=O)CCC(N)C(=O)NC(CCC(O)=O)C(O)=O KOSRFJWDECSPRO-UHFFFAOYSA-N 0.000 description 1
- 230000004075 alteration Effects 0.000 description 1
- 239000003242 anti bacterial agent Substances 0.000 description 1
- 229940088710 antibiotic agent Drugs 0.000 description 1
- 239000000427 antigen Substances 0.000 description 1
- 230000000890 antigenic effect Effects 0.000 description 1
- 102000036639 antigens Human genes 0.000 description 1
- 108091007433 antigens Proteins 0.000 description 1
- ODKSFYDXXFIFQN-UHFFFAOYSA-N arginine Natural products OC(=O)C(N)CCCNC(N)=N ODKSFYDXXFIFQN-UHFFFAOYSA-N 0.000 description 1
- 108010008355 arginyl-glutamine Proteins 0.000 description 1
- 108010094001 arginyl-tryptophyl-arginine Proteins 0.000 description 1
- 108010062796 arginyllysine Proteins 0.000 description 1
- 235000009582 asparagine Nutrition 0.000 description 1
- 229960001230 asparagine Drugs 0.000 description 1
- 235000003704 aspartic acid Nutrition 0.000 description 1
- 108010040443 aspartyl-aspartic acid Proteins 0.000 description 1
- 108010093581 aspartyl-proline Proteins 0.000 description 1
- 108010092854 aspartyllysine Proteins 0.000 description 1
- 208000022362 bacterial infectious disease Diseases 0.000 description 1
- OQFSQFPPLPISGP-UHFFFAOYSA-N beta-carboxyaspartic acid Natural products OC(=O)C(N)C(C(O)=O)C(O)=O OQFSQFPPLPISGP-UHFFFAOYSA-N 0.000 description 1
- 230000001851 biosynthetic effect Effects 0.000 description 1
- 239000008280 blood Substances 0.000 description 1
- 210000004369 blood Anatomy 0.000 description 1
- 230000003197 catalytic effect Effects 0.000 description 1
- 150000001768 cations Chemical class 0.000 description 1
- 230000008859 change Effects 0.000 description 1
- 239000007795 chemical reaction product Substances 0.000 description 1
- 239000003153 chemical reaction reagent Substances 0.000 description 1
- 238000003776 cleavage reaction Methods 0.000 description 1
- 238000004440 column chromatography Methods 0.000 description 1
- 239000002299 complementary DNA Substances 0.000 description 1
- 230000021615 conjugation Effects 0.000 description 1
- 238000007796 conventional method Methods 0.000 description 1
- 230000008878 coupling Effects 0.000 description 1
- 238000010168 coupling process Methods 0.000 description 1
- 238000005859 coupling reaction Methods 0.000 description 1
- XUJNEKJLAYXESH-UHFFFAOYSA-N cysteine Natural products SCC(N)C(O)=O XUJNEKJLAYXESH-UHFFFAOYSA-N 0.000 description 1
- 235000018417 cysteine Nutrition 0.000 description 1
- IERHLVCPSMICTF-UHFFFAOYSA-N cytidine monophosphate Natural products O=C1N=C(N)C=CN1C1C(O)C(O)C(COP(O)(O)=O)O1 IERHLVCPSMICTF-UHFFFAOYSA-N 0.000 description 1
- 230000003247 decreasing effect Effects 0.000 description 1
- 238000010511 deprotection reaction Methods 0.000 description 1
- 238000001514 detection method Methods 0.000 description 1
- 230000029087 digestion Effects 0.000 description 1
- 239000001177 diphosphate Substances 0.000 description 1
- XPPKVPWEQAFLFU-UHFFFAOYSA-J diphosphate(4-) Chemical compound [O-]P([O-])(=O)OP([O-])([O-])=O XPPKVPWEQAFLFU-UHFFFAOYSA-J 0.000 description 1
- 235000011180 diphosphates Nutrition 0.000 description 1
- 150000002016 disaccharides Chemical class 0.000 description 1
- 238000004520 electroporation Methods 0.000 description 1
- 238000010828 elution Methods 0.000 description 1
- 238000005516 engineering process Methods 0.000 description 1
- 230000002255 enzymatic effect Effects 0.000 description 1
- 230000009144 enzymatic modification Effects 0.000 description 1
- 210000003527 eukaryotic cell Anatomy 0.000 description 1
- 230000017188 evasion or tolerance of host immune response Effects 0.000 description 1
- 150000002243 furanoses Chemical class 0.000 description 1
- 150000002256 galaktoses Chemical class 0.000 description 1
- 235000013922 glutamic acid Nutrition 0.000 description 1
- 239000004220 glutamic acid Substances 0.000 description 1
- ZDXPYRJPNDTMRX-UHFFFAOYSA-N glutamine Natural products OC(=O)C(N)CCC(N)=O ZDXPYRJPNDTMRX-UHFFFAOYSA-N 0.000 description 1
- 230000013595 glycosylation Effects 0.000 description 1
- 238000006206 glycosylation reaction Methods 0.000 description 1
- 108010066198 glycyl-leucyl-phenylalanine Proteins 0.000 description 1
- 108010059898 glycyl-tyrosyl-lysine Proteins 0.000 description 1
- 108010015792 glycyllysine Proteins 0.000 description 1
- QGWNDRXFNXRZMB-UHFFFAOYSA-N guanidine diphosphate Natural products C1=2NC(N)=NC(=O)C=2N=CN1C1OC(COP(O)(=O)OP(O)(O)=O)C(O)C1O QGWNDRXFNXRZMB-UHFFFAOYSA-N 0.000 description 1
- 229940047650 haemophilus influenzae Drugs 0.000 description 1
- HNDVDQJCIGZPNO-UHFFFAOYSA-N histidine Natural products OC(=O)C(N)CC1=CN=CN1 HNDVDQJCIGZPNO-UHFFFAOYSA-N 0.000 description 1
- 108010092114 histidylphenylalanine Proteins 0.000 description 1
- 230000002209 hydrophobic effect Effects 0.000 description 1
- 230000014726 immortalization of host cell Effects 0.000 description 1
- 230000001900 immune effect Effects 0.000 description 1
- 238000000099 in vitro assay Methods 0.000 description 1
- 238000010348 incorporation Methods 0.000 description 1
- 208000015181 infectious disease Diseases 0.000 description 1
- 230000000977 initiatory effect Effects 0.000 description 1
- 238000001155 isoelectric focusing Methods 0.000 description 1
- AGPKZVBTJJNPAG-UHFFFAOYSA-N isoleucine Natural products CCC(C)C(N)C(O)=O AGPKZVBTJJNPAG-UHFFFAOYSA-N 0.000 description 1
- 229960000310 isoleucine Drugs 0.000 description 1
- 108010044374 isoleucyl-tyrosine Proteins 0.000 description 1
- IEQCXFNWPAHHQR-UHFFFAOYSA-N lacto-N-neotetraose Natural products OCC1OC(OC2C(C(OC3C(OC(O)C(O)C3O)CO)OC(CO)C2O)O)C(NC(=O)C)C(O)C1OC1OC(CO)C(O)C(O)C1O IEQCXFNWPAHHQR-UHFFFAOYSA-N 0.000 description 1
- 229940062780 lacto-n-neotetraose Drugs 0.000 description 1
- DOVBXGDYENZJBJ-ONMPCKGSSA-N lactosamine Chemical compound O=C[C@H](N)[C@@H](O)[C@H](O)[C@@H](CO)O[C@@H]1O[C@H](CO)[C@H](O)[C@H](O)[C@H]1O DOVBXGDYENZJBJ-ONMPCKGSSA-N 0.000 description 1
- 108010044311 leucyl-glycyl-glycine Proteins 0.000 description 1
- 108010051673 leucyl-glycyl-phenylalanine Proteins 0.000 description 1
- 108010073472 leucyl-prolyl-proline Proteins 0.000 description 1
- 108010087810 leucyl-seryl-glutamyl-leucine Proteins 0.000 description 1
- 108010012058 leucyltyrosine Proteins 0.000 description 1
- 101150018810 lgtB gene Proteins 0.000 description 1
- 101150071729 lgtE gene Proteins 0.000 description 1
- 108010009298 lysylglutamic acid Proteins 0.000 description 1
- 108010038320 lysylphenylalanine Proteins 0.000 description 1
- 210000004962 mammalian cell Anatomy 0.000 description 1
- 229910052748 manganese Inorganic materials 0.000 description 1
- 239000011572 manganese Substances 0.000 description 1
- 108020004999 messenger RNA Proteins 0.000 description 1
- 229930182817 methionine Natural products 0.000 description 1
- 108010005942 methionylglycine Proteins 0.000 description 1
- 108010068488 methionylphenylalanine Proteins 0.000 description 1
- 244000005700 microbiome Species 0.000 description 1
- 238000013508 migration Methods 0.000 description 1
- 230000005012 migration Effects 0.000 description 1
- 238000002703 mutagenesis Methods 0.000 description 1
- 231100000350 mutagenesis Toxicity 0.000 description 1
- RBMYDHMFFAVMMM-PLQWBNBWSA-N neolactotetraose Chemical compound O([C@H]1[C@H](O)[C@H]([C@@H](O[C@@H]1CO)O[C@@H]1[C@H]([C@H](O[C@H]([C@H](O)CO)[C@H](O)[C@@H](O)C=O)O[C@H](CO)[C@@H]1O)O)NC(=O)C)[C@@H]1O[C@H](CO)[C@H](O)[C@H](O)[C@H]1O RBMYDHMFFAVMMM-PLQWBNBWSA-N 0.000 description 1
- 230000007935 neutral effect Effects 0.000 description 1
- 239000002777 nucleoside Substances 0.000 description 1
- 238000010647 peptide synthesis reaction Methods 0.000 description 1
- 210000001322 periplasm Anatomy 0.000 description 1
- COLNVLDHVKWLRT-UHFFFAOYSA-N phenylalanine Natural products OC(=O)C(N)CC1=CC=CC=C1 COLNVLDHVKWLRT-UHFFFAOYSA-N 0.000 description 1
- 108010073025 phenylalanylphenylalanine Proteins 0.000 description 1
- 230000004962 physiological condition Effects 0.000 description 1
- 238000002264 polyacrylamide gel electrophoresis Methods 0.000 description 1
- 229920000642 polymer Polymers 0.000 description 1
- 125000002924 primary amino group Chemical group [H]N([H])* 0.000 description 1
- 230000009465 prokaryotic expression Effects 0.000 description 1
- 230000000644 propagated effect Effects 0.000 description 1
- 230000004952 protein activity Effects 0.000 description 1
- 230000004853 protein function Effects 0.000 description 1
- 150000003214 pyranose derivatives Chemical class 0.000 description 1
- 230000006798 recombination Effects 0.000 description 1
- 239000011347 resin Substances 0.000 description 1
- 229920005989 resin Polymers 0.000 description 1
- CVHZOJJKTDOEJC-UHFFFAOYSA-N saccharin Chemical group C1=CC=C2C(=O)NS(=O)(=O)C2=C1 CVHZOJJKTDOEJC-UHFFFAOYSA-N 0.000 description 1
- 230000007017 scission Effects 0.000 description 1
- 238000000926 separation method Methods 0.000 description 1
- SQVRNKJHWKZAKO-OQPLDHBCSA-N sialic acid Chemical compound CC(=O)N[C@@H]1[C@@H](O)C[C@@](O)(C(O)=O)OC1[C@H](O)[C@H](O)CO SQVRNKJHWKZAKO-OQPLDHBCSA-N 0.000 description 1
- 238000000527 sonication Methods 0.000 description 1
- 241000894007 species Species 0.000 description 1
- 230000000707 stereoselective effect Effects 0.000 description 1
- 239000000126 substance Substances 0.000 description 1
- 150000008163 sugars Chemical class 0.000 description 1
- 238000012360 testing method Methods 0.000 description 1
- 150000004044 tetrasaccharides Chemical class 0.000 description 1
- 229960004072 thrombin Drugs 0.000 description 1
- 210000001519 tissue Anatomy 0.000 description 1
- 238000001890 transfection Methods 0.000 description 1
- 230000010474 transient expression Effects 0.000 description 1
- OUYCCCASQSFEME-UHFFFAOYSA-N tyrosine Natural products OC(=O)C(N)CC1=CC=C(O)C=C1 OUYCCCASQSFEME-UHFFFAOYSA-N 0.000 description 1
- 108010051110 tyrosyl-lysine Proteins 0.000 description 1
- 108010077037 tyrosyl-tyrosyl-phenylalanine Proteins 0.000 description 1
- 241000701161 unidentified adenovirus Species 0.000 description 1
- 229940045145 uridine Drugs 0.000 description 1
- 229960005486 vaccine Drugs 0.000 description 1
- 239000004474 valine Substances 0.000 description 1
- 230000003612 virological effect Effects 0.000 description 1
- 210000005253 yeast cell Anatomy 0.000 description 1
- 235000021241 α-lactalbumin Nutrition 0.000 description 1
Images
Classifications
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12P—FERMENTATION OR ENZYME-USING PROCESSES TO SYNTHESISE A DESIRED CHEMICAL COMPOUND OR COMPOSITION OR TO SEPARATE OPTICAL ISOMERS FROM A RACEMIC MIXTURE
- C12P19/00—Preparation of compounds containing saccharide radicals
- C12P19/18—Preparation of compounds containing saccharide radicals produced by the action of a glycosyl transferase, e.g. alpha-, beta- or gamma-cyclodextrins
-
- C—CHEMISTRY; METALLURGY
- C12—BIOCHEMISTRY; BEER; SPIRITS; WINE; VINEGAR; MICROBIOLOGY; ENZYMOLOGY; MUTATION OR GENETIC ENGINEERING
- C12N—MICROORGANISMS OR ENZYMES; COMPOSITIONS THEREOF; PROPAGATING, PRESERVING, OR MAINTAINING MICROORGANISMS; MUTATION OR GENETIC ENGINEERING; CULTURE MEDIA
- C12N9/00—Enzymes; Proenzymes; Compositions thereof; Processes for preparing, activating, inhibiting, separating or purifying enzymes
- C12N9/10—Transferases (2.)
- C12N9/1048—Glycosyltransferases (2.4)
-
- Y—GENERAL TAGGING OF NEW TECHNOLOGICAL DEVELOPMENTS; GENERAL TAGGING OF CROSS-SECTIONAL TECHNOLOGIES SPANNING OVER SEVERAL SECTIONS OF THE IPC; TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
- Y10—TECHNICAL SUBJECTS COVERED BY FORMER USPC
- Y10S—TECHNICAL SUBJECTS COVERED BY FORMER USPC CROSS-REFERENCE ART COLLECTIONS [XRACs] AND DIGESTS
- Y10S435/00—Chemistry: molecular biology and microbiology
- Y10S435/8215—Microorganisms
- Y10S435/822—Microorganisms using bacteria or actinomycetales
- Y10S435/871—Neisseria
Definitions
- the present invention relates to a method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and a gene encoding such a polyglycosyltransferase.
- Oligosaccharides are polymers of varying number of residues, linkages, and subunits.
- the basic subunit is a carbohydrate monosaccharide or sugar, such as mannose, glucose, galactose, N-acetylglucosamine, N-acetylgalactosamine, and the like.
- the number of different possible stereoisomeric oligosaccharide chains is enormous.
- Oligosaccharides and polysaccharides play an important role in protein function and activity, by serving as half-life modulators, and, in some instances, by providing structure. As pointed out above, oligosaccharides are critical to the antigenic variability, and hence immune evasion, of Neisseria, especially gonococcus.
- glycosyltransferases have only provided for the transfer of a single saccharide unit, specific for the glycosyltransferases.
- a galactosyltransferase would transfer only galactose
- a glucosyltransferase would transfer only glucose
- an N-acetylglucosamine and a sialyl transferase would transfer only sialic acid.
- glycosyltransferase which transferred at least two different sugar donors would be advantageous in synthesizing two glycosidic bonds of at least a trisaccharide, using the same glycosyltransferase.
- a locus involved in the biosynthesis of gonococcal lipooligosaccharide (LOS) has been reported as being cloned from the gonococcal strain F62 (Gotschlich J. Exp. Med (1994) 180, 2181-2190).
- Five genes lgtA, lgtB, lgtC, lgtD and lgtE are reported, and based on deletion experiments, activities are postulated, as encoding for glycosyltransferases. Due to the uncertainty caused by the nature of the deletion experiments, the exact activity of the proteins encoded by each of the genes was not ascertained and some of the genes are only suggested as being responsible for one or another activity, in the alternative.
- the gene lgtA is suggested as most likely to code for a GlcNAc transferase.
- the present invention is directed to a method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and nucleic acids encoding a polyglycosyltransferase.
- the invention is directed to a method of transferring at least two saccharide units with a polyglycosyltransferase.
- another aspect of the invention is directed to a method of transferring at least two saccharide units with a polyglycosyltransferase, which transfers both GlcNAc and GalNAc, from the corresponding sugar nucleotides, to a sugar acceptor.
- a polyglycosyltransferase is obtained from a bacteria of the genus Neisseria, Escherichia or Pseudomonas.
- [0018] is directed to a method of making at least two oligosaccharide compounds, from the same acceptor, with a polyglycosyltransferase.
- another aspect of the invention is directed to a method of making at least two oligosaccharide compounds, from the same acceptor, with a polyglycosyltransferase, which transfers both GlcNAc and Ga1NAc, from the corresponding sugar nucleotides, to the sugar acceptor.
- [0020] is a method of transferring an N-acetylgalactosamine using a glycosyltransferase of SEQ ID NO:3
- the invention relates to a nucleic acid that has a nucleotide sequence which encodes for polypeptide sequence shown in SEQ ID NO: 3.
- the functionally active polyglycosyltransferase of the invention is characterized by catalyzing both the addition of GalNAc ⁇ 1-3 to Gal and the addition of G1cNAc ⁇ 1-3 to Gal.
- FIG. 1 provides the amino acid sequence of a polyglycosyltransferase of SEQ ID NO. 3.
- FIG. 2 provides the polynucleotide sequence of a LOS encoding gene isolated from N. gonorrhoeae, of which nucleotides 445-1488 encode for a polyglycosyltransferase
- the present invention provides for a method of transferring at least two saccharide units with a polyglycosyltransferase, a gene encoding for a polyglycosyltransferase, and a polyglycosyltransferase.
- the polyglycosyltransferases of the invention can be used for in vitro biosynthesis of various oligosaccharides, such as the core oligosaccharide of the human blood group antigens, i.e., lacto-N-neotetraose.
- Cloning and expression of a polyglycosyltransferase of the invention can be accomplished using standard techniques, as disclosed herein. Such a polyglycosyltransferase is useful for biosynthesis of oligosaccharides in vitro, or alternatively genes encoding such a polyglycosyltransferase can be transfected into cells, e.g., yeast cells or eukaryotic cells, to provide for alternative glycosylation of proteins and lipids.
- cells e.g., yeast cells or eukaryotic cells
- the instant invention is based, in part, on the discovery that a polyglycosyltransferase isolated from Neisseria gonorrhoeae is capable of transferring both GlcNAc ⁇ 1-3 to Gal and GalNAc ⁇ 1-3 to Gal, from the corresponding sugar nucleotides.
- a protein, glycosyltransferase activity is reported by Gotschlich et al U.S. Ser. No. 08/312,387 filed on Sep. 26, 1994, by cloning of a locus involved in the biosynthesis of gonococcal LOS, strain F62.
- the protein sequence identified as SEQ ID NO: 3, a 348 amino acid protein, has now been discovered to have a polyglycosyltransferase activity. More specifically, the protein sequence identified as SEQ ID NO: 3 has been discovered to transfer both GlcNAc ⁇ 1-3 to Gal and Ga1NAc ⁇ 1-3 to Gal, from the corresponding sugar nucleotides.
- a polynucleotide sequence encoding for a polyglycosyltransferase is similar to the sequence of nucleic acids no 445 to 1488 of an LOS isolated from N. gonorrhoeae (see FIG. 2) in which three or six of the guanine units occurring between nucleic acids no 700 to 715 have been deleted.
- Another polynucleotide sequence is similar to the sequence of nucleic acids no 445 to 1488 of an LOS isolated from N. gonorrhoeae (see FIG. 2), in which nucleic acids sufficient to encode the amino acid sequence -Tyr-Ser-Arg-Asp-Ser-Ser, can be appended to nucleic acid 1488.
- Lipopolysaccharide LPS
- Lipooligosaccharide LOS
- N-Acetyl-neuraminic acid cytidine mono phosphate CMP-NANA
- wild type wt
- Gal galactose
- Glc glucose
- NAc N-acetyl (e.g., GalNAc or GlcNAc).
- any Neisseria bacterial cell can potentially serve as the nucleic acid source for the molecular cloning of a polyglycosyltransferase gene.
- the genes are isolated from Neisseria gonorrhoeae.
- the DNA may be obtained by standard procedures known in the art from cloned DNA (e.g., a DNA “library'), by chemical synthesis, by cDNA cloning, or by the cloning of genomic DNA, or fragments thereof, purified from the desired cell (See, for example, Sambrook et al., 1989, supra Glover, D. M.
- a N. gonorrhoeae genomic DNA can be digested with a restriction endonuclease or endonucleases, e.g., Sau3A, into a phage vector digested with a restriction endonuclease or endonucleases, e.g., BamHI/EcoRI, for creation of a phage genomic library.
- a restriction endonuclease or endonucleases e.g., Sau3A
- a phage vector digested with a restriction endonuclease or endonucleases e.g., BamHI/EcoRI
- the gene should be molecularly cloned into a suitable vector for propagation of the gene.
- DNA fragments are generated, some of which will encode the desired gene.
- the DNA may be cleaved at specific sites using; various restriction enzymes.
- DNAse the presence of manganese to fragment the DNA, or the DNA can be physically sheared, as for example, by sonication.
- the linear DNA fragments can then be separated according to size by standard techniques, including but not limited to, agarose and polyacrylamide gel electrophoresis and column chromatography.
- the generated DNA fragments may be screened by nucleic acid hybridization to the labeled probe synthesized with a sequence as disclosed herein (Benton and Davis, 1977, Science 196:180; Grunstein and Hogness, 1975, Proc. Natl. Acad. Sci. U.S.A. 72:3961). Those DNA fragments with substantial homology to the probe will hybridize.
- Suitable probes can be generated by PCR using random primers.
- a probe which will hybridize to the polynucleotide sequence encoding for a four or five glycine residue i.e. a twelve or fifteen guanine residue
- the presence of the gene may be detected by assays based on the physical, chemical, or immunological properties of its expressed product. For example DNA clones that produce a protein that, e.g., has similar or identical electrophoretic migration, isoelectric focusing behavior, proteolytic digestion maps, proteolytic activity, or functional properties, in particular polyglycosyltransferase activity, the ability of a polyglycosyltransferase protein to mediate transfer of two different saccharide units to an acceptor molecule.
- DNA for a polyglycosyltransferase gene can be isolated by PCR using oligonucleotide primers designed from the nucleotide sequences disclosed herein. Other methods are possible and within the scope of the invention.
- the identified and isolated gene can then be inserted into an appropriate cloning vector.
- vector-host systems known in the art may be used. Possible vectors include, but are not limited to, plasmids or modified viruses, but the vector system must be compatible with the host cell used. In a specific aspect of the invention, the polyglycosyltransferase coding sequence is inserted in an E. coli cloning vector.
- vectors include, but are not limited to, bacteriophages such as lambda derivatives, or plasmids such as pBR322 derivatives or pUC plasmid derivatives, e.g., pGEX vectors, pmal-c, pFLAG, etc.
- the insertion into a cloning vector can, for example, be accomplished by ligating the DNA fragment into a cloning vector which has complementary cohesive termini.
- the ends of the DNA molecules may be enzymatically modified.
- any site desired may be produced by ligating nucleotide sequences (linkers) onto the DNA termini, these ligated linkers may comprise specific chemically synthesized oligonucleotides encoding restriction endonuclease recognition sequences.
- PCR primers containing such linker sites can be used to amplify the DNA for cloning.
- Recombinant molecules can be introduced into host cells via transformation, transfection, infection, electroporation, etc., so that many copies of the gene sequence are generated.
- Transformation of host cells with recombinant DNA molecules that incorporate the isolated polyglycosyltransferase gene or synthesized DNA sequence enables generation of multiple copies of the gene.
- the gene may be obtained in large quantities by growing transformants, isolating the recombinant DNA molecules from the transformants and, when necessary, retrieving the inserted gene from the isolated recombinant DNA.
- the present invention also relates to vectors containing genes encoding truncated forms of the enzyme (fragments) and derivatives of polyglycosyltransferases that have the same functional activity as a polyglycosyltransferase.
- fragments and derivatives related to polyglycosyltransferases are within the scope of the present invention.
- the fragment or derivative is functionally active, i.e., capable of mediating transfer of two sugar donors to a acceptor moieties.
- Truncated fragments of the polyglycosyltransferases can be prepared by eliminating N-terminal, C-terminal, or internal regions of the protein that are not required for functional activity. Usually, such portions that are eliminated will include only a few, e.g., between 1 and 5, amino acid residues, but larger segments may be removed.
- Chimeric molecules e.g., fusion proteins, containing all or a functionally active portion of a polyglycosyltransferase of the invention joined to another protein are also envisioned.
- a polyglycosyltransferase fusion protein comprises at least a functionally active portion of a non-glycosyltransferase protein joined via a peptide bond to at least a functionally active portion of a polyglycosyltransferase polypeptide.
- the non-glycosyltransferase sequences can be amino- or carboxy-terminal to the polyglycosyltransferase sequences.
- a recombinant DNA molecule encoding such a fusion protein comprises a sequence encoding at least a functionally active portion of a non-glycosyltransferase protein joined in-frame to the polyglycosyltransferase coding sequence, and preferably encodes a cleavage site for a specific protease, e.g., thrombin or Factor Xa, preferably at the polyglycosyltransferase-non-glycosyltransferase juncture.
- the fusion protein may be expressed in Escherichia coli.
- polyglycosyltransferase derivatives can be made by altering encoding nucleic acid sequences by substitutions, additions or deletions that provide for functionally equivalent molecules. Due to the degeneracy of nucleotide coding sequences, other DNA sequences which encode substantially the same amino acid sequence as an polyglycosyltransferase gene may be used in the practice of the present invention. These include but are not limited to nucleotide sequences comprising all or portions of polyglycosyltransferase genes that are altered by the substitution of different codons that encode the same amino acid residue within the sequence, thus producing a silent change.
- polyglycosyltransferase derivatives of the invention include, but are not limited to, those containing, as a primary amino acid sequence, all or part of the amino acid sequence of a polyglycosyltransferase including altered sequences in which functionally equivalent amino acid residues are substituted for residues within the sequence resulting in a conservative amino acid substitution.
- one or more amino acid residues within the sequence can be substituted by another amino acid of a similar polarity, which acts as a functional equivalent, resulting in a silent alteration.
- Substitutes for an amino acid within the sequence may be selected from other members of the class to which the amino acid belongs.
- the nonpolar (hydrophobic) amino acids include alanine, leucine, isoleucine, valine, proline, phenylalanine, tryptophan and methionine.
- the polar neutral amino acids include glycine, serine, threonine, cysteine, tyrosine, asparagine, and glutamine.
- the positively charged (basic) amino acids include arginine, lysine and histidine.
- the negatively charged (acidic) amino acids include aspartic acid and glutamic acid.
- genes encoding polyglycosyltransferase derivatives and analogs of the invention can be produced by various methods known in the art (e.g., Sambrook et al., 1989, supra). The sequence can be cleaved at appropriate sites with restriction endonuclease(s), followed by further enzymatic modification if desired, isolated, and ligated in vitro. In the production of the gene encoding a derivative or analog of polyglycosyltransferase, care should be taken to ensure that the modified gene remains within the same translational reading frame as the polyglycosyltransferase gene, uninterrupted by translational stop signals, in the gene region where the desired activity is encoded.
- polyglycosyltransferase nucleic acid sequence can be mutated in vitro or in vivo, to create and/or destroy translation, initiation, and/or termination sequences, or to create variations in coding regions and/or form new restriction endonuclease sites or destroy preexisting ones, to facilitate further in vitro modification.
- Any technique for mutagenesis known in the art can be used, including but not limited to, in vitro site-directed mutagenesis (Hutchinson, C., et al., 1978, J. Biol. Chem.
- PCR techniques are preferred for site directed mutagenesis (see Higuchi, 1989, “Using PCR to Engineer DNA”, in PCR Technology: Principles and Applications for “DNA Amplification, H. Erlich, ed., Stockton Press, Chapter 6, pp. 61-70).
- polyglycosyltransferases can also be isolated from other bacterial species of Neisseria.
- Exemplary Neisseria bacterial sources include N. animalis (ATCC 19573), N. canis (ATCC 14687), N. cinerea (ATCC 14685), N. cuniculi (ATCC 14688), N. denitrificans (ATCC 14686), N. elongata (ATCC 25295), N. elongata subsp glycolytica (ATCC 29315), N. elongata subsp. nitroreducens (ATCC 49377), N.
- polyglycosyltransferases can be isolated from Branhamella catarrhalis, Haemophilus influenzae, Escherichia coli, Pseudomonas aeruginosa and Pseudomonas cepacia.
- the gene coding for a polyglycosyltransferase, or a functionally active fragment or other derivative thereof can be inserted into an appropriate expression vector, i.e., a vector which contains the necessary elements for the transcription and translation of the inserted protein-coding sequence.
- An expression vector also preferably includes a replication origin.
- the necessary transcriptional and translational signals can also be supplied by the native polyglycosyltransferase gene and/or its flanking regions.
- host-vector systems may be utilized to express the protein-coding sequence.
- a bacterial expression system is used to provide for high level expression of the protein with a higher probability of the native conformation.
- Potential host-vector systems include but are not limited to mammalian cell systems infected with virus (e.g., vaccine virus, adenovirus, etc.); insect cell systems infected with virus (e.g., baculovirus); microorganisms such as yeast containing yeast vectors, or bacteria transformed with bacteriophage, DNA, plasmid DNA, or cosmid DNA.
- virus e.g., vaccine virus, adenovirus, etc.
- insect cell systems infected with virus e.g., baculovirus
- microorganisms such as yeast containing yeast vectors, or bacteria transformed with bacteriophage, DNA, plasmid DNA, or cosmid DNA.
- the expression elements of vectors vary in their strengths and specificities. Depending on the host-vector system utilized, any one of a number of suitable transcription and translation elements may be used.
- the periplasmic form of the polyglycosyltransferase (containing a signal sequence) is produced for export of the protein to the Escherichia coli periplasm or in an expression system based on Bacillus subtillis.
- any of the methods previously described for the insertion of DNA fragments into a vector may be used to construct expression vectors containing a chimeric gene consisting of appropriate transcriptional/translational control signals and the protein coding sequences. These methods may include in vitro recombinant DNA and synthetic techniques and in vivo recombinants (genetic recombination).
- Expression of nucleic acid sequence encoding a polyglycosyltransferase or peptide fragment may be regulated by a second nucleic acid sequence so that the polyglycosyltransferase or peptide is expressed in a host transformed with the recombinant DNA molecule.
- expression of a polyglycosyltransferase may be controlled by any promoter/enhancer element known in the art, but these regulatory elements must be functional in the host selected for expression.
- bacterial promoters are required for expression in bacteria.
- Eukaryotic viral or eukaryotic promoters are preferred when a vector containing a polyglycosyltransferase gene is injected directly into a subject for transient expression, resulting in heterologous protection against bacterial infection, as described in detail below.
- Promoters which may be used to control polyglycosyltransferase gene expression include, but are not limited to, the SV40 early promoter region (Benoist and Chambon, 198 1, Nature 290:304-310), the promoter contained in the 3′ long terminal repeat of Rous sarcoma virus (Yamamoto, et al., 1980, Cell 22:787-797), the herpes thymidine kinase promoter (Wagner et al., 1981, Proc. Natl. Acad. Sci. U.S.A.
- Expression vectors containing polyglycosyltransferase gene inserts can be identified by four general approaches: (a) PCR amplification of the desired plasmid DNA or specific mRNA, (b) nucleic acid hybridization, (c) presence or absence of “marker” gene functions, and (d) expression of inserted sequences.
- the nucleic acids can be amplified by PCR with incorporation of radionucleotides or stained with ethidiuin bromide to provide for detection of the amplified product.
- the presence of a foreign gene inserted in an expression vector can be detected by nucleic acid hybridization using probes comprising sequences that are homologous to an inserted polyglycosyltransferase gene.
- the recombinant vector/host system can be identified and selected based upon the presence or absence of certain “marker” gene functions (e.g., 3-galactosidase activity, PhoA activity, thymidine kinase activity, resistance to antibiotics, transformation phenotype, occlusion body formation in baculovirus, etc.) caused by the insertion of foreign genes in the vector.
- certain “marker” gene functions e.g., 3-galactosidase activity, PhoA activity, thymidine kinase activity, resistance to antibiotics, transformation phenotype, occlusion body formation in baculovirus, etc.
- recombinants containing the polyglycosyltransferase insert can be identified by the absence of the marker gene function.
- recombinant expression vectors can be identified by assaying for the activity of the polyglycosyltransferase gene product expressed by the recombinant. Such assays can be based, for example, on the physical or functional properties of the polyglycosyltransferase gene product in in vitro assay systems, e.g., polyglycosyltransferase activity. Once a suitable host system and growth conditions are established, recombinant expression vectors can be propagated and prepared in quantity.
- the polyglycosyltransferase of the present invention can be used in the biosynthesis of oligosaccharides.
- the polyglycosyltransferases of the invention are capable of stereospecific conjugation of two specific activated saccharide unit to specific acceptor molecules.
- Such activated saccharides generally consist of uridine or guanosine diphosphate and cytidine monophosphate derivatives of the saccharides, in which the nucleoside mono- and diphosphate serves as a leaving group.
- the activated saccharide may be a saccharide-UDP, a saccharide-GDP, or a saccharide-CMP.
- the activated saccharide is UDP-G1cNAc, UDP-Ga1NAc, or UDP-Gal.
- two different saccharide units means saccharides which differ in structure and/or stereochemistry at a position other than C 1 and accordingly the pyranose and furanose of the same carbon backbone are considered to be the same saccharide unit, while glucose and galactose (i.e. C 4 isomers) are considered different.
- a glycosyltransferase typically has a catalytic activity of from about 1 to 250 turnovers/sec in order to be considered to possess a specific glycosyltransferase activity.
- each individual glycosyltransferase activity of the polyglycosyltransferase of the present invention is within the range of from 1 to 250 turnovers/sec, preferably from 5 to 100 turnovers/sec, more preferably from 10 to 30 turnovers/sec.
- the relative rates of each of the identified glycosyltransferase activities, in the polyglycosyltransferase has relative activities of from 0.1 to 10 times, preferably from 0.2 to 5 times, more preferably from 0.5 to 2 times and most preferably from 0.8 to 1.5, the rate of any one of the other glycosyltransferase activities.
- acceptor moiety refers to the molecules to which the polyglycosyltransferase transfers activated sugars.
- a polyglycosyltransferase is contacted with an appropriate activated saccharide and an appropriate acceptor moiety under conditions effective to transfer and covalently bond the saccharide to the acceptor molecule.
- Conditions of time, temperature, and pH appropriate and optimal for a particular saccharine unit transfer can be determined through routine testing; generally, physiological conditions will be acceptable.
- Certain co-reagents may also be desirable; for example, it may be more effective to contact the polyglycosyltransferase with the activated saccharide and the acceptor moiety in the presence of a divalent cation.
- the polyglycosyltransferase enzymes can be covalently or non-covalently immobilized on a solid phase support such as SEPHADEX, SEPHAROSE, or poly(acrylamide-co-N-acryloxysucciimide) (PAN) resin.
- a solid phase support such as SEPHADEX, SEPHAROSE, or poly(acrylamide-co-N-acryloxysucciimide) (PAN) resin.
- a specific reaction can be performed in an isolated reaction solution, with facile separation of the solid phase enzyme from the reaction products. Immobilization of the enzyme also allows for a continuous biosynthetic stream, with the specific polyglycosyltransferases attached to a solid support, with the supports arranged randomly or in distinct zones in the specified order in a column, with passage of the reaction solution through the column and elution of the desired oligosaccharide at the end.
- An oligosaccharide, e.g., a disaccharide, prepared using a polyglycosyltransferase of the present invention can serve as an acceptor moiety for further synthesis, either using other polyglycosyltransferases of the invention, or glycosyltransferases known in the art (see, e.g., Roth, U.S. Pat. No. 5,180,674).
- the polyglycosyltransferases of the present invention can be used to prepare GalNAc ⁇ 1-3-Gal ⁇ 1-4-GlcNAc ⁇ 1-3-Gal ⁇ 1-4-Glc or GalNAc ⁇ 1-3-Gal ⁇ 1-4-GlcNAc ⁇ 1-3-Gal ⁇ 1-4-GlcNAc from lactose or lactosamine respectively, in which a polyglycosyltransferase is used to synthesize both the GlcNAc ⁇ 1-3-Gal and GalNAc ⁇ 1-3 Gal linkages.
- a method for preparing an oligosaccharide having the structure GalNAc ⁇ 1-3-Gal ⁇ 1-4-G1cNAc ⁇ 1-3-Gal ⁇ 1-4-Glc comprises sequentially performing the steps of:
- an activated GlcNAc such as UDP-GlcNAc
- step a) contacting a reaction mixture comprising an activated GalNAc (i.e. UDP-GalNAc) to the acceptor moiety comprising a Gal ⁇ 1-4-GlcNAc ⁇ 1-3-Gal ⁇ 1-4-Glc residue in the presence of the polyglycosyltransferase of step a).
- an activated GalNAc i.e. UDP-GalNAc
- a suitable ⁇ 1-4 galactosyltransferase can be isolated from bovine milk.
- Oligosaccharide synthesis, using a polyglycosyltransferase is generally conducted at a temperature of from 15 to 38° C., preferably from 20 to 25° C. While enzymatic activity is comparable at 25° C. and 37° C., the polyglycosyltransferase stability is greater at 25° C.
- polyglycosyltransferase activity is observed in the absence of ⁇ -lactalbumin.
- polyglycosyltransferase activity is observed at the same pH, more preferably at pH 6.5 to 7.5.
- polyglycosyltransferase activity is observed at the same temperature.
- Lactose was contacted with UDP-N-acetylglucosamine and a ⁇ -galactoside ⁇ 1-3 N-acetylglucosaminyl transferase of SEQ ID NO: 3, in a 0.5 M HEPES buffered aqueous solution at 25° C.
- the product trisaccharide was then contacted with UDP-Gal and a ⁇ -N-acetylglucosaminoside ⁇ 1-4 Galactosyltransferase isolated from bovine milk, in a 0.05 M HEPES buffered aqueous solution at 37° C.
- the product tetrasaccharide was then contacted with UDP-N-acetylgalactosamine and a ⁇ -galactoside ⁇ 1-3 N-acetylgalactosaminyl transferase of SEQ ID NO: 3, in a 0.05 M HEPES buffered aqueous solution at 25° C.
- the title pentasaccharide was isolated by conventional methods.
Landscapes
- Chemical & Material Sciences (AREA)
- Organic Chemistry (AREA)
- Life Sciences & Earth Sciences (AREA)
- Engineering & Computer Science (AREA)
- Health & Medical Sciences (AREA)
- Zoology (AREA)
- Wood Science & Technology (AREA)
- Genetics & Genomics (AREA)
- Bioinformatics & Cheminformatics (AREA)
- General Health & Medical Sciences (AREA)
- Biochemistry (AREA)
- General Engineering & Computer Science (AREA)
- Microbiology (AREA)
- Biotechnology (AREA)
- Chemical Kinetics & Catalysis (AREA)
- General Chemical & Material Sciences (AREA)
- Medicinal Chemistry (AREA)
- Molecular Biology (AREA)
- Biomedical Technology (AREA)
- Micro-Organisms Or Cultivation Processes Thereof (AREA)
- Enzymes And Modification Thereof (AREA)
- Preparation Of Compounds By Using Micro-Organisms (AREA)
Abstract
The present invention relates to a method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and a gene encoding such a polyglycosyltransferase.
Description
- 1. Field of the Invention
- The present invention relates to a method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and a gene encoding such a polyglycosyltransferase.
- 2. Discussion of the Background:
- Biosynthesis of Oligosaccharides
- Oligosaccharides are polymers of varying number of residues, linkages, and subunits. The basic subunit is a carbohydrate monosaccharide or sugar, such as mannose, glucose, galactose, N-acetylglucosamine, N-acetylgalactosamine, and the like. The number of different possible stereoisomeric oligosaccharide chains is enormous.
- Oligosaccharides and polysaccharides play an important role in protein function and activity, by serving as half-life modulators, and, in some instances, by providing structure. As pointed out above, oligosaccharides are critical to the antigenic variability, and hence immune evasion, of Neisseria, especially gonococcus.
- Numerous classical techniques for the synthesis of carbohydrates have been developed, but these techniques suffer the difficulty of requiring selective protection and deprotection. Organic synthesis of oligosaccharides is further hampered by the liability of many glycosidic bonds, difficulties in achieving regioselective sugar coupling, and generally low synthetic yields. In short, unlike the experience with peptide synthesis, traditional synthetic organic chemistry cannot provide for quantitative, reliable synthesis of even fairly simple oligosaccharides.
- Recent advances in oligosaccharide synthesis have occurred with the isolation of glycosyltransferases from natural sources. These enzymes can be used in vitro to prepare oligosaccharides and polysaccharides (see, e.g., Roth, U.S. Pat. No. 5,180,674). The advantage of biosynthesis with glycosyltransferases is that the glycosidic linkages formed by enzymes are highly stereo and regiospecific. However, each enzyme catalyzes linkage of specific sugar donor residues to other specific acceptor molecules, e.g., an oligosaccharide or lipid. Thus, synthesis of a desired oligosaccharide has required the use of a different glycosyltransferase for each different saccharide unit being transferred.
- More specifically, such glycosyltransferases have only provided for the transfer of a single saccharide unit, specific for the glycosyltransferases. For example a galactosyltransferase would transfer only galactose, a glucosyltransferase would transfer only glucose, an N-acetylglucosamine and a sialyl transferase would transfer only sialic acid.
- However, the lack of generality of glycosyltransferase, makes it necessary to use a different glycosyltransferase for every different sugar donor being transferred. As the usefulness of oligosaccharide compounds expands, the ability to transfer more than one sugar donor would provide a tremendous advantage, by decreasing the number of glycosyltransferases necessary to form necessary glycosidic bonds.
- In addition, a glycosyltransferase which transferred at least two different sugar donors would be advantageous in synthesizing two glycosidic bonds of at least a trisaccharide, using the same glycosyltransferase.
- A locus involved in the biosynthesis of gonococcal lipooligosaccharide (LOS) has been reported as being cloned from the gonococcal strain F62 (Gotschlich J. Exp. Med (1994) 180, 2181-2190). Five genes lgtA, lgtB, lgtC, lgtD and lgtE are reported, and based on deletion experiments, activities are postulated, as encoding for glycosyltransferases. Due to the uncertainty caused by the nature of the deletion experiments, the exact activity of the proteins encoded by each of the genes was not ascertained and some of the genes are only suggested as being responsible for one or another activity, in the alternative. The gene lgtA is suggested as most likely to code for a GlcNAc transferase.
- The transfer of more than one different saccharide moietie, by a polyglycosyltransferase has heretofore been unreported.
- The present invention is directed to a method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and nucleic acids encoding a polyglycosyltransferase.
- Accordingly, in one aspect, the invention is directed to a method of transferring at least two saccharide units with a polyglycosyltransferase.
- Accordingly, another aspect of the invention is directed to a method of transferring at least two saccharide units with a polyglycosyltransferase, which transfers both GlcNAc and GalNAc, from the corresponding sugar nucleotides, to a sugar acceptor.
- According to another aspect of the invention, a polyglycosyltransferase is obtained from a bacteria of the genus Neisseria, Escherichia or Pseudomonas.
- According to another aspect of the invention, is directed to a method of making at least two oligosaccharide compounds, from the same acceptor, with a polyglycosyltransferase.
- Accordingly, another aspect of the invention is directed to a method of making at least two oligosaccharide compounds, from the same acceptor, with a polyglycosyltransferase, which transfers both GlcNAc and Ga1NAc, from the corresponding sugar nucleotides, to the sugar acceptor.
- According to another embodiment of the present invention is a method of transferring an N-acetylgalactosamine using a glycosyltransferase of SEQ ID NO:3
- In specific embodiments, the invention relates to a nucleic acid that has a nucleotide sequence which encodes for polypeptide sequence shown in SEQ ID NO: 3.
- The functionally active polyglycosyltransferase of the invention is characterized by catalyzing both the addition of GalNAc β1-3 to Gal and the addition of G1cNAc β1-3 to Gal.
- FIG. 1: provides the amino acid sequence of a polyglycosyltransferase of SEQ ID NO. 3.
- FIG. 2: provides the polynucleotide sequence of a LOS encoding gene isolated from N. gonorrhoeae, of which nucleotides 445-1488 encode for a polyglycosyltransferase
- As disclosed above, the present invention provides for a method of transferring at least two saccharide units with a polyglycosyltransferase, a gene encoding for a polyglycosyltransferase, and a polyglycosyltransferase. The polyglycosyltransferases of the invention can be used for in vitro biosynthesis of various oligosaccharides, such as the core oligosaccharide of the human blood group antigens, i.e., lacto-N-neotetraose.
- Cloning and expression of a polyglycosyltransferase of the invention can be accomplished using standard techniques, as disclosed herein. Such a polyglycosyltransferase is useful for biosynthesis of oligosaccharides in vitro, or alternatively genes encoding such a polyglycosyltransferase can be transfected into cells, e.g., yeast cells or eukaryotic cells, to provide for alternative glycosylation of proteins and lipids.
- The instant invention is based, in part, on the discovery that a polyglycosyltransferase isolated from Neisseria gonorrhoeae is capable of transferring both GlcNAc β1-3 to Gal and GalNAc β1-3 to Gal, from the corresponding sugar nucleotides.
- A protein, glycosyltransferase activity, is reported by Gotschlich et al U.S. Ser. No. 08/312,387 filed on Sep. 26, 1994, by cloning of a locus involved in the biosynthesis of gonococcal LOS, strain F62. The protein sequence identified as SEQ ID NO: 3, a 348 amino acid protein, has now been discovered to have a polyglycosyltransferase activity. More specifically, the protein sequence identified as SEQ ID NO: 3 has been discovered to transfer both GlcNAc β1-3 to Gal and Ga1NAc β1-3 to Gal, from the corresponding sugar nucleotides.
- In addition to the protein sequence SEQ ID NO: 3 and nucleic acid sequences for encoding them reported in U.S. Ser. No. 08/312,387, new polyglycosyltransferases have been discovered which transfer two different sugar units. This protein is similar to the protein of SEQ ID: 3 with the deletion of one or two of the five glycine units occurring between amino acid nos 86-90 of lgtA. In addition, an additional amino acid sequence -Tyr-Ser-Arg-Asp-Ser-Ser can be appended to the carboxy terminus of Ile (amino acid no 348) of SEQ ID NO: 3, while retaining the polyglycosyltransferase activity.
- A polynucleotide sequence encoding for a polyglycosyltransferase is similar to the sequence of nucleic acids no 445 to 1488 of an LOS isolated from N. gonorrhoeae (see FIG. 2) in which three or six of the guanine units occurring between nucleic acids no 700 to 715 have been deleted.
- Another polynucleotide sequence is similar to the sequence of nucleic acids no 445 to 1488 of an LOS isolated from N. gonorrhoeae (see FIG. 2), in which nucleic acids sufficient to encode the amino acid sequence -Tyr-Ser-Arg-Asp-Ser-Ser, can be appended to nucleic acid 1488.
- Abbreviations used throughout this specification include: Lipopolysaccharide, LPS; Lipooligosaccharide, LOS; N-Acetyl-neuraminic acid cytidine mono phosphate, CMP-NANA; wild type, wt; Gal, galactose; Glc, glucose; NAc, N-acetyl (e.g., GalNAc or GlcNAc).
- Any Neisseria bacterial cell can potentially serve as the nucleic acid source for the molecular cloning of a polyglycosyltransferase gene. In a specific embodiment, infra, the genes are isolated from Neisseria gonorrhoeae. The DNA may be obtained by standard procedures known in the art from cloned DNA (e.g., a DNA “library'), by chemical synthesis, by cDNA cloning, or by the cloning of genomic DNA, or fragments thereof, purified from the desired cell (See, for example, Sambrook et al., 1989, supra Glover, D. M. (ed.), 1985, DNA Cloning: A Practical Approach, MRL Press, Ltd., Oxford, U.K. Vol. 1, II). For example, a N. gonorrhoeae genomic DNA can be digested with a restriction endonuclease or endonucleases, e.g., Sau3A, into a phage vector digested with a restriction endonuclease or endonucleases, e.g., BamHI/EcoRI, for creation of a phage genomic library. Whatever the source, the gene should be molecularly cloned into a suitable vector for propagation of the gene.
- In the molecular cloning of the gene from genomic DNA, DNA fragments are generated, some of which will encode the desired gene. The DNA may be cleaved at specific sites using; various restriction enzymes. Alternatively, one may use DNAse the presence of manganese to fragment the DNA, or the DNA can be physically sheared, as for example, by sonication. The linear DNA fragments can then be separated according to size by standard techniques, including but not limited to, agarose and polyacrylamide gel electrophoresis and column chromatography.
- Once the DNA fragments are generated, identification of the specific DNA fragment containing the desired polyglycosyltransferase gene may be accomplished in a number of ways. For example, the generated DNA fragments may be screened by nucleic acid hybridization to the labeled probe synthesized with a sequence as disclosed herein (Benton and Davis, 1977, Science 196:180; Grunstein and Hogness, 1975, Proc. Natl. Acad. Sci. U.S.A. 72:3961). Those DNA fragments with substantial homology to the probe will hybridize.
- Suitable probes can be generated by PCR using random primers. In particular a probe which will hybridize to the polynucleotide sequence encoding for a four or five glycine residue (i.e. a twelve or fifteen guanine residue) would be a suitable probe for a polyglycosyltransferase.
- As described above, the presence of the gene may be detected by assays based on the physical, chemical, or immunological properties of its expressed product. For example DNA clones that produce a protein that, e.g., has similar or identical electrophoretic migration, isoelectric focusing behavior, proteolytic digestion maps, proteolytic activity, or functional properties, in particular polyglycosyltransferase activity, the ability of a polyglycosyltransferase protein to mediate transfer of two different saccharide units to an acceptor molecule.
- Alternatives to isolating a polyglycosyltransferase genomic DNA include, but are not limited to, chemically synthesizing the gene sequence itself from a known sequence that encodes a polyglycosyltransferase. In another embodiment, DNA for a polyglycosyltransferase gene can be isolated by PCR using oligonucleotide primers designed from the nucleotide sequences disclosed herein. Other methods are possible and within the scope of the invention.
- The identified and isolated gene can then be inserted into an appropriate cloning vector. A large number of vector-host systems known in the art may be used. Possible vectors include, but are not limited to, plasmids or modified viruses, but the vector system must be compatible with the host cell used. In a specific aspect of the invention, the polyglycosyltransferase coding sequence is inserted in an E. coli cloning vector. Other examples of vectors include, but are not limited to, bacteriophages such as lambda derivatives, or plasmids such as pBR322 derivatives or pUC plasmid derivatives, e.g., pGEX vectors, pmal-c, pFLAG, etc. The insertion into a cloning vector can, for example, be accomplished by ligating the DNA fragment into a cloning vector which has complementary cohesive termini. However, if the complementary restriction sites used to fragment the DNA are not present in the cloning vector, the ends of the DNA molecules may be enzymatically modified. Alternatively, any site desired may be produced by ligating nucleotide sequences (linkers) onto the DNA termini, these ligated linkers may comprise specific chemically synthesized oligonucleotides encoding restriction endonuclease recognition sequences. In specific embodiment, PCR primers containing such linker sites can be used to amplify the DNA for cloning. Recombinant molecules can be introduced into host cells via transformation, transfection, infection, electroporation, etc., so that many copies of the gene sequence are generated.
- Transformation of host cells with recombinant DNA molecules that incorporate the isolated polyglycosyltransferase gene or synthesized DNA sequence enables generation of multiple copies of the gene. Thus, the gene may be obtained in large quantities by growing transformants, isolating the recombinant DNA molecules from the transformants and, when necessary, retrieving the inserted gene from the isolated recombinant DNA.
- The present invention also relates to vectors containing genes encoding truncated forms of the enzyme (fragments) and derivatives of polyglycosyltransferases that have the same functional activity as a polyglycosyltransferase. The production and use of fragments and derivatives related to polyglycosyltransferases are within the scope of the present invention. In a specific embodiment, the fragment or derivative is functionally active, i.e., capable of mediating transfer of two sugar donors to a acceptor moieties.
- Truncated fragments of the polyglycosyltransferases can be prepared by eliminating N-terminal, C-terminal, or internal regions of the protein that are not required for functional activity. Usually, such portions that are eliminated will include only a few, e.g., between 1 and 5, amino acid residues, but larger segments may be removed.
- Chimeric molecules, e.g., fusion proteins, containing all or a functionally active portion of a polyglycosyltransferase of the invention joined to another protein are also envisioned. A polyglycosyltransferase fusion protein comprises at least a functionally active portion of a non-glycosyltransferase protein joined via a peptide bond to at least a functionally active portion of a polyglycosyltransferase polypeptide. The non-glycosyltransferase sequences can be amino- or carboxy-terminal to the polyglycosyltransferase sequences. Expression of a fusion protein can result in an enzymatically inactive polyglycosyltransferase fusion protein. A recombinant DNA molecule encoding such a fusion protein comprises a sequence encoding at least a functionally active portion of a non-glycosyltransferase protein joined in-frame to the polyglycosyltransferase coding sequence, and preferably encodes a cleavage site for a specific protease, e.g., thrombin or Factor Xa, preferably at the polyglycosyltransferase-non-glycosyltransferase juncture. In a specific embodiment, the fusion protein may be expressed in Escherichia coli.
- In particular, polyglycosyltransferase derivatives can be made by altering encoding nucleic acid sequences by substitutions, additions or deletions that provide for functionally equivalent molecules. Due to the degeneracy of nucleotide coding sequences, other DNA sequences which encode substantially the same amino acid sequence as an polyglycosyltransferase gene may be used in the practice of the present invention. These include but are not limited to nucleotide sequences comprising all or portions of polyglycosyltransferase genes that are altered by the substitution of different codons that encode the same amino acid residue within the sequence, thus producing a silent change. Likewise, the polyglycosyltransferase derivatives of the invention include, but are not limited to, those containing, as a primary amino acid sequence, all or part of the amino acid sequence of a polyglycosyltransferase including altered sequences in which functionally equivalent amino acid residues are substituted for residues within the sequence resulting in a conservative amino acid substitution. For example, one or more amino acid residues within the sequence can be substituted by another amino acid of a similar polarity, which acts as a functional equivalent, resulting in a silent alteration. Substitutes for an amino acid within the sequence may be selected from other members of the class to which the amino acid belongs. For example, the nonpolar (hydrophobic) amino acids include alanine, leucine, isoleucine, valine, proline, phenylalanine, tryptophan and methionine. The polar neutral amino acids include glycine, serine, threonine, cysteine, tyrosine, asparagine, and glutamine. The positively charged (basic) amino acids include arginine, lysine and histidine. The negatively charged (acidic) amino acids include aspartic acid and glutamic acid.
- The genes encoding polyglycosyltransferase derivatives and analogs of the invention can be produced by various methods known in the art (e.g., Sambrook et al., 1989, supra). The sequence can be cleaved at appropriate sites with restriction endonuclease(s), followed by further enzymatic modification if desired, isolated, and ligated in vitro. In the production of the gene encoding a derivative or analog of polyglycosyltransferase, care should be taken to ensure that the modified gene remains within the same translational reading frame as the polyglycosyltransferase gene, uninterrupted by translational stop signals, in the gene region where the desired activity is encoded.
- Additionally, the polyglycosyltransferase nucleic acid sequence can be mutated in vitro or in vivo, to create and/or destroy translation, initiation, and/or termination sequences, or to create variations in coding regions and/or form new restriction endonuclease sites or destroy preexisting ones, to facilitate further in vitro modification. Any technique for mutagenesis known in the art can be used, including but not limited to, in vitro site-directed mutagenesis (Hutchinson, C., et al., 1978, J. Biol. Chem. 253:6551; Zoller and Smith, 1984, DNA 3:479-488; Oliphant et al., 1986, Gene 44:177; Hutchinson et al., 1986, Proc. Natl. Acad. Sci. U.S.A. 83:710), use of TABO linkers (Pharmacia), etc. PCR techniques are preferred for site directed mutagenesis (see Higuchi, 1989, “Using PCR to Engineer DNA”, in PCR Technology: Principles and Applications for “DNA Amplification, H. Erlich, ed., Stockton Press, Chapter 6, pp. 61-70).
- While a polyglycosyltransferase has been isolated from a bacteria of Neisseria gonorrhoeae, polyglycosyltransferases can also be isolated from other bacterial species of Neisseria. Exemplary Neisseria bacterial sources include N. animalis (ATCC 19573), N. canis (ATCC 14687), N. cinerea (ATCC 14685), N. cuniculi (ATCC 14688), N. denitrificans (ATCC 14686), N. elongata (ATCC 25295), N. elongata subsp glycolytica (ATCC 29315), N. elongata subsp. nitroreducens (ATCC 49377), N. flavescens (ATCC 13115), N. gonorrhoeae (ATCC 33084), N. lactamica (ATCC 23970), N. macaca (ATCC 33926), N. meningitidis, N. mucosa (ATCC 19695), N. mucosa subsp. eidelbergensis (ATCC 25998), N. polysaccharea (ATCC 43768), N. sicca (ATCC 29256) and N. subflava (ATCC 49275). In addition polyglycosyltransferases can be isolated from Branhamella catarrhalis, Haemophilus influenzae, Escherichia coli, Pseudomonas aeruginosa and Pseudomonas cepacia.
- The gene coding for a polyglycosyltransferase, or a functionally active fragment or other derivative thereof, can be inserted into an appropriate expression vector, i.e., a vector which contains the necessary elements for the transcription and translation of the inserted protein-coding sequence. An expression vector also preferably includes a replication origin. The necessary transcriptional and translational signals can also be supplied by the native polyglycosyltransferase gene and/or its flanking regions. A variety of host-vector systems may be utilized to express the protein-coding sequence. Preferably, however, a bacterial expression system is used to provide for high level expression of the protein with a higher probability of the native conformation. Potential host-vector systems include but are not limited to mammalian cell systems infected with virus (e.g., vaccine virus, adenovirus, etc.); insect cell systems infected with virus (e.g., baculovirus); microorganisms such as yeast containing yeast vectors, or bacteria transformed with bacteriophage, DNA, plasmid DNA, or cosmid DNA. The expression elements of vectors vary in their strengths and specificities. Depending on the host-vector system utilized, any one of a number of suitable transcription and translation elements may be used.
- Preferably, the periplasmic form of the polyglycosyltransferase (containing a signal sequence) is produced for export of the protein to the Escherichia coli periplasm or in an expression system based on Bacillus subtillis.
- Any of the methods previously described for the insertion of DNA fragments into a vector may be used to construct expression vectors containing a chimeric gene consisting of appropriate transcriptional/translational control signals and the protein coding sequences. These methods may include in vitro recombinant DNA and synthetic techniques and in vivo recombinants (genetic recombination).
- Expression of nucleic acid sequence encoding a polyglycosyltransferase or peptide fragment may be regulated by a second nucleic acid sequence so that the polyglycosyltransferase or peptide is expressed in a host transformed with the recombinant DNA molecule. For example, expression of a polyglycosyltransferase may be controlled by any promoter/enhancer element known in the art, but these regulatory elements must be functional in the host selected for expression. For expression in bacteria, bacterial promoters are required. Eukaryotic viral or eukaryotic promoters, including tissue specific promoters, are preferred when a vector containing a polyglycosyltransferase gene is injected directly into a subject for transient expression, resulting in heterologous protection against bacterial infection, as described in detail below. Promoters which may be used to control polyglycosyltransferase gene expression include, but are not limited to, the SV40 early promoter region (Benoist and Chambon, 198 1, Nature 290:304-310), the promoter contained in the 3′ long terminal repeat of Rous sarcoma virus (Yamamoto, et al., 1980, Cell 22:787-797), the herpes thymidine kinase promoter (Wagner et al., 1981, Proc. Natl. Acad. Sci. U.S.A. 78:1441-1445), the regulatory sequences of the metallothionein gene (Brinster et al., 1982, Nature 296:39-42); prokaryotic expression vectors such as the O-lactamase promoter (Villa-Kamaroff, et al., 1978, Proc. Natl. Acad. Sci. U.S.A. 75:3727-3731), or the tac promoter (DeBoer, et al., 1983, Proc. Natl. Acad. Sci. U.S.A. 80:21-25); see also “Useful proteins from recombinant bacteria” in Scientific American, 1980, 242:74-94; and the like
- Expression vectors containing polyglycosyltransferase gene inserts can be identified by four general approaches: (a) PCR amplification of the desired plasmid DNA or specific mRNA, (b) nucleic acid hybridization, (c) presence or absence of “marker” gene functions, and (d) expression of inserted sequences. In the first approach, the nucleic acids can be amplified by PCR with incorporation of radionucleotides or stained with ethidiuin bromide to provide for detection of the amplified product. In the second approach, the presence of a foreign gene inserted in an expression vector can be detected by nucleic acid hybridization using probes comprising sequences that are homologous to an inserted polyglycosyltransferase gene. In the third approach, the recombinant vector/host system can be identified and selected based upon the presence or absence of certain “marker” gene functions (e.g., 3-galactosidase activity, PhoA activity, thymidine kinase activity, resistance to antibiotics, transformation phenotype, occlusion body formation in baculovirus, etc.) caused by the insertion of foreign genes in the vector. If the polyglycosyltransferase gene is inserted within the marker gene sequence of the vector, recombinants containing the polyglycosyltransferase insert can be identified by the absence of the marker gene function. In the fourth approach, recombinant expression vectors can be identified by assaying for the activity of the polyglycosyltransferase gene product expressed by the recombinant. Such assays can be based, for example, on the physical or functional properties of the polyglycosyltransferase gene product in in vitro assay systems, e.g., polyglycosyltransferase activity. Once a suitable host system and growth conditions are established, recombinant expression vectors can be propagated and prepared in quantity.
- Biosynthesis of Oligosaccharides
- The polyglycosyltransferase of the present invention can be used in the biosynthesis of oligosaccharides. The polyglycosyltransferases of the invention are capable of stereospecific conjugation of two specific activated saccharide unit to specific acceptor molecules. Such activated saccharides generally consist of uridine or guanosine diphosphate and cytidine monophosphate derivatives of the saccharides, in which the nucleoside mono- and diphosphate serves as a leaving group. Thus, the activated saccharide may be a saccharide-UDP, a saccharide-GDP, or a saccharide-CMP. In specific embodiments, the activated saccharide is UDP-G1cNAc, UDP-Ga1NAc, or UDP-Gal.
- Within the context of the claimed invention, two different saccharide units means saccharides which differ in structure and/or stereochemistry at a position other than C 1 and accordingly the pyranose and furanose of the same carbon backbone are considered to be the same saccharide unit, while glucose and galactose (i.e. C4 isomers) are considered different.
- A glycosyltransferase typically has a catalytic activity of from about 1 to 250 turnovers/sec in order to be considered to possess a specific glycosyltransferase activity.
- Accordingly each individual glycosyltransferase activity of the polyglycosyltransferase of the present invention, is within the range of from 1 to 250 turnovers/sec, preferably from 5 to 100 turnovers/sec, more preferably from 10 to 30 turnovers/sec.
- In addition to absolute glycosyltransferase activity, the relative rates of each of the identified glycosyltransferase activities, in the polyglycosyltransferase, has relative activities of from 0.1 to 10 times, preferably from 0.2 to 5 times, more preferably from 0.5 to 2 times and most preferably from 0.8 to 1.5, the rate of any one of the other glycosyltransferase activities.
- The term “acceptor moiety” as used herein refers to the molecules to which the polyglycosyltransferase transfers activated sugars.
- For the synthesis of an oligosaccharide, a polyglycosyltransferase is contacted with an appropriate activated saccharide and an appropriate acceptor moiety under conditions effective to transfer and covalently bond the saccharide to the acceptor molecule. Conditions of time, temperature, and pH appropriate and optimal for a particular saccharine unit transfer can be determined through routine testing; generally, physiological conditions will be acceptable. Certain co-reagents may also be desirable; for example, it may be more effective to contact the polyglycosyltransferase with the activated saccharide and the acceptor moiety in the presence of a divalent cation.
- According to the invention, the polyglycosyltransferase enzymes can be covalently or non-covalently immobilized on a solid phase support such as SEPHADEX, SEPHAROSE, or poly(acrylamide-co-N-acryloxysucciimide) (PAN) resin. A specific reaction can be performed in an isolated reaction solution, with facile separation of the solid phase enzyme from the reaction products. Immobilization of the enzyme also allows for a continuous biosynthetic stream, with the specific polyglycosyltransferases attached to a solid support, with the supports arranged randomly or in distinct zones in the specified order in a column, with passage of the reaction solution through the column and elution of the desired oligosaccharide at the end. An efficient method for attaching the polyglycosyltransferase to a solid support and using such immobilized polyglycosyltransferases is described in U.S. Pat. No. 5,180,674, issued Jan. 19, 1993 to Roth, which is specifically incorporated herein by reference in its entirety.
- An oligosaccharide, e.g., a disaccharide, prepared using a polyglycosyltransferase of the present invention can serve as an acceptor moiety for further synthesis, either using other polyglycosyltransferases of the invention, or glycosyltransferases known in the art (see, e.g., Roth, U.S. Pat. No. 5,180,674).
- Alternatively, the polyglycosyltransferases of the present invention can be used to prepare GalNAcβ1-3-Galβ1-4-GlcNAcβ1-3-Galβ1-4-Glc or GalNAcβ1-3-Galβ1-4-GlcNAcβ1-3-Galβ1-4-GlcNAc from lactose or lactosamine respectively, in which a polyglycosyltransferase is used to synthesize both the GlcNAc β1-3-Gal and GalNAc β1-3 Gal linkages.
- Accordingly, a method for preparing an oligosaccharide having the structure GalNAcβ1-3-Galβ1-4-G1cNAc β1-3-Galβ1-4-Glc comprises sequentially performing the steps of:
- a) contacting a reaction mixture comprising an activated GlcNAc (such as UDP-GlcNAc) to lactose with a polyglycosyltransferase having an amino acid sequence of SEQ ID NO:3, or a functionally active fragment thereof;
- b) contacting a reaction mixture comprising an activated Gal (i.e. UDP-Gal) to the acceptor moiety comprising a GlcNAcβ1-3-Galβ1-4-Glc residue in the presence of a β1-4-galactosyltransferase; and
- c) contacting a reaction mixture comprising an activated GalNAc (i.e. UDP-GalNAc) to the acceptor moiety comprising a Galβ1-4-GlcNAcβ1-3-Galβ1-4-Glc residue in the presence of the polyglycosyltransferase of step a).
- A suitable β1-4 galactosyltransferase can be isolated from bovine milk.
- Oligosaccharide synthesis, using a polyglycosyltransferase is generally conducted at a temperature of from 15 to 38° C., preferably from 20 to 25° C. While enzymatic activity is comparable at 25° C. and 37° C., the polyglycosyltransferase stability is greater at 25° C.
- In a preferred embodiment polyglycosyltransferase activity is observed in the absence of α-lactalbumin.
- In a preferred embodiment polyglycosyltransferase activity is observed at the same pH, more preferably at pH 6.5 to 7.5.
- In a preferred embodiment polyglycosyltransferase activity is observed at the same temperature.
- Other features of the invention will become apparent in the course of the following descriptions of exemplary embodiments which are given for illustration of the invention and are not intended to be limiting thereof.
- Synthesis of GalNAcβ1-3-Galβ1-4-G1cNAcβ1-3-Ga1β1-4-Glc:
- Lactose was contacted with UDP-N-acetylglucosamine and a β-galactoside β1-3 N-acetylglucosaminyl transferase of SEQ ID NO: 3, in a 0.5 M HEPES buffered aqueous solution at 25° C. The product trisaccharide was then contacted with UDP-Gal and a β-N-acetylglucosaminoside β1-4 Galactosyltransferase isolated from bovine milk, in a 0.05 M HEPES buffered aqueous solution at 37° C. The product tetrasaccharide was then contacted with UDP-N-acetylgalactosamine and a β-galactoside β1-3 N-acetylgalactosaminyl transferase of SEQ ID NO: 3, in a 0.05 M HEPES buffered aqueous solution at 25° C. The title pentasaccharide was isolated by conventional methods.
- Obviously, numerous modifications and variations of the present invention are possible in light of the above teachings. It is therefore to be understood that within the scope of the appended claims, the invention may be practiced otherwise than as specifically described therein.
-
1 8 1 5859 DNA Neisseria gonorrhoeae 1 ctgcaggccg tcgccgtatt caaacaactg cccgaagccg ccgcgctcgc cgccgccaac 60 aaacgcgtgc aaaacctgct gaaaaaagcc gatgccgcgt tgggcgaagt caatgaaagc 120 ctgctgcaac aggacgaaga aaaagccctg tacgctgccg cgcaaggttt gcagccgaaa 180 attgccgccg ccgtcgccga aggcaatttc cgaaccgcct tgtccgaact ggcttccgtc 240 aagccgcagg ttgatgcctt cttcgacggc gtgatggtga tggcggaaga tgccgccgta 300 aaacaaaacc gcctgaacct gctgaaccgc ttggcagagc agatgaacgc ggtggccgac 360 atcgcgcttt tgggcgagta accgttgtac agtccaaatg ccgtctgaag ccttcaggcg 420 gcatcaaatt atcgggagag taaattgcag cctttagtca gcgtattgat ttgcgcctac 480 aacgtagaaa aatattttgc ccaatcatta gccgccgtcg tgaatcagac ttggcgcaac 540 ttggatattt tgattgtcga tgacggctcg acagacggca cacttgccat tgccaaggat 600 tttcaaaagc gggacagccg tatcaaaatc cttgcacaag ctcaaaattc cggcctgatt 660 ccctctttaa acatcgggct ggacgaattg gcaaagtcgg gggggggggg gggggaatat 720 attgcgcgca ccgatgccga cgatattgcc tcccccggct ggattgagaa aatcgtgggc 780 gagatggaaa aagaccgcag catcattgcg atgggcgcgt ggctggaagt tttgtcggaa 840 gaaaaggacg gcaaccggct ggcgcggcac cacaaacacg gcaaaatttg gaaaaagccg 900 acccggcacg aagacatcgc cgcctttttc cctttcggca accccataca caacaacacg 960 atgattatgc ggcgcagcgt cattgacggc ggtttgcgtt acgacaccga gcgggattgg 1020 gcggaagatt accaattttg gtacgatgtc agcaaattgg gcaggctggc ttattatccc 1080 gaagccttgg tcaaataccg ccttcacgcc aatcaggttt catccaaaca cagcgtccgc 1140 caacacgaaa tcgcgcaagg catccaaaaa accgccagaa acgatttttt gcagtctatg 1200 ggttttaaaa cccggttcga cagcctagaa taccgccaaa caaaagcagc ggcgtatgaa 1260 ctgccggaga aggatttgcc ggaagaagat tttgaacgcg cccgccggtt tttgtaccaa 1320 tgcttcaaac ggacggacac gccgccctcc ggcgcgtggc tggatttcgc ggcagacggc 1380 aggatgaggc ggctgtttac cttgaggcaa tacttcggca ttttgtaccg gctgattaaa 1440 aaccgccggc aggcgcggtc ggattcggca gggaaagaac aggagattta atgcaaaacc 1500 acgttatcag cttggcttcc gccgcagaac gcagggcgca cattgccgca accttcggca 1560 gtcgcggcat cccgttccag tttttcgacg cactgatgcc gtctgaaagg ctggaacggg 1620 caatggcgga actcgtcccc ggcttgtcgg cgcaccccta tttgagcgga gtggaaaaag 1680 cctgctttat gagccacgcc gtattgtggg aacaggcatt ggacgaaggc gtaccgtata 1740 tcgccgtatt tgaagatgat gtcttactcg gcgaaggcgc ggagcagttc cttgccgaag 1800 atacttggct gcaagaacgc tttgaccccg attccgcctt tgtcgtccgc ttggaaacga 1860 tgtttatgca cgtcctgacc tcgccctccg gcgtggcgga ctacggcggg cgcgcctttc 1920 cgcttttgga aagcgaacac tgcgggacgg cgggctatat tatttcccga aaggcgatgc 1980 gttttttctt ggacaggttt gccgttttgc cgcccgaacg cctgcaccct gtcgatttga 2040 tgatgttcgg caaccctgac gacagggaag gaatgccggt ttgccagctc aatcccgcct 2100 tgtgcgccca agagctgcat tatgccaagt ttcacgacca aaacagcgca ttgggcagcc 2160 tgatcgaaca tgaccgccgc ctgaaccgca aacagcaatg gcgcgattcc cccgccaaca 2220 cattcaaaca ccgcctgatc cgcgccttga ccaaaatcgg cagggaaagg gaaaaacgcc 2280 ggcaaaggcg cgaacagtta atcggcaaga ttattgtgcc tttccaataa aaggagaaaa 2340 gatggacatc gtatttgcgg cagacgacaa ctatgccgcc tacctttgcg ttgcggcaaa 2400 aagcgtggaa gcggcccatc ccgatacgga aatcaggttc cacgtcctcg atgccggcat 2460 cagtgaggaa aaccgggcgg cggttgccgc caatttgcgg ggggggggta atatccgctt 2520 tatagacgta aaccccgaag atttcgccgg cttcccctta aacatcaggc acatttccat 2580 tacgacttat gcccgcctga aattgggcga atacattgcc gattgcgaca aagtcctgta 2640 tctggatacg gacgtattgg tcagggacgg cctgaagccc ttatgggata ccgatttggg 2700 cggtaactgg gtcggcgcgt gcatcgattt gtttgtcgaa aggcaggaag gatacaaaca 2760 aaaaatcggt atggcggacg gagaatatta tttcaatgcc ggcgtattgc tgatcaacct 2820 gaaaaagtgg cggcggcacg atattttcaa aatgtcctgc gaatgggtgg aacaatacaa 2880 ggacgtgatg caatatcagg atcaggacat tttgaacggg ctgtttaaag gcggggtgtg 2940 ttatgcgaac agccgtttca actttatgcc gaccaattat gcctttatgg cgaacgggtt 3000 tgcgtcccgc cataccgacc cgctttacct cgaccgtacc aatacggcga tgcccgtcgc 3060 cgtcagccat tattgcggct cggcaaagcc gtggcacagg gactgcaccg tttggggtgc 3120 ggaacgtttc acagagttgg ccggcagcct gacgaccgtt cccgaagaat ggcgcggcaa 3180 acttgccgtc ccgccgacaa agtgtatgct tcaaagatgg cgcaaaaagc tgtctgccag 3240 attcttacgc aagatttatt gacggggcag gccgtctgaa gccttcagac ggcatcggac 3300 gtatcggaaa ggagaaacgg attgcagcct ttagtcagcg tattgatttg cgcctacaac 3360 gcagaaaaat attttgccca atcattggcc gccgtagtgg ggcagacttg gcgcaacttg 3420 gatattttga ttgtcgatga cggctcgacg gacggcacgc ccgccattgc ccggcatttc 3480 caagaacagg acggcaggat caggataatt tccaatcccc gcaatttggg ctttatcgcc 3540 tctttaaaca tcgggctgga cgaattggca aagtcggggg ggggggaata tattgcgcgc 3600 accgatgccg acgatattgc ctcccccggc tggattgaga aaatcgtggg cgagatggaa 3660 aaagaccgca gcatcattgc gatgggcgcg tggttggaag ttttgtcgga agaaaacaat 3720 aaaagcgtgc ttgccgccat tgcccgaaac ggcgcaattt gggacaaacc gacccggcat 3780 gaagacattg tcgccgtttt ccctttcggc aaccccatac acaacaacac gatgattatg 3840 aggcgcagcg tcattgacgg cggtttgcgg ttcgatccag cctatatcca cgccgaagac 3900 tataagtttt ggtacgaagc cggcaaactg ggcaggctgg cttattatcc cgaagccttg 3960 gtcaaatacc gcttccatca agaccagact tcttccaaat acaacctgca acagcgcagg 4020 acggcgtgga aaatcaaaga agaaatcagg gcggggtatt ggaaggcggc aggcatagcc 4080 gtcggggcgg actgcctgaa ttacgggctt ttgaaatcaa cggcatatgc gttgtacgaa 4140 aaagccttgt ccggacagga tatcggatgc ctccgcctgt tcctgtacga atatttcttg 4200 tcgttggaaa agtattcttt gaccgatttg ctggatttct tgacagaccg cgtgatgagg 4260 aagctgtttg ccgcaccgca atataggaaa atcctgaaaa aaatgttacg cccttggaaa 4320 taccgcagct attgaaaccg aacaggataa atcatgcaaa accacgttat cagcttggct 4380 tccgccgcag agcgcagggc gcacattgcc gataccttcg gcagtcgcgg catcccgttc 4440 cagtttttcg acgcactgat gccgtctgaa aggctggaac aggcgatggc ggaactcgtc 4500 cccggcttgt cggcgcaccc ctatttgagc ggagtggaaa aagcctgctt tatgagccac 4560 gccgtattgt gggaacaggc gttggatgaa ggtctgccgt atatcgccgt atttgaggac 4620 gacgttttac tcggcgaagg cgcggagcag ttccttgccg aagatacttg gttggaagag 4680 cgttttgaca aggattccgc ctttatcgtc cgtttggaaa cgatgtttgc gaaagttatt 4740 gtcagaccgg ataaagtcct gaattatgaa aaccggtcat ttcctttgct ggagagcgaa 4800 cattgtggga cggctggcta tatcatttcg cgtgaggcga tgcggttttt cttggacagg 4860 tttgccgttt tgccgccaga gcggattaaa gcggtagatt tgatgatgtt tacttatttc 4920 tttgataagg aggggatgcc tgtttatcag gttagtcccg ccttatgtac ccaagaattg 4980 cattatgcca agtttctcag tcaaaacagt atgttgggta gcgatttgga aaaagatagg 5040 gaacaaggaa gaagacaccg ccgttcgttg aaggtgatgt ttgacttgaa gcgtgctttg 5100 ggtaaattcg gtagggaaaa gaagaaaaga atggagcgtc aaaggcaggc ggagcttgag 5160 aaagtttacg gcaggcgggt catattgttc aaatagtttg tgtaaaatat aggggattaa 5220 aatcagaaat ggacacactg tcattcccgc gcaggcggga atctaggtct ttaaacttcg 5280 gttttttccg ataaattctt gccgcattaa aattccagat tcccgctttc gcggggatga 5340 cggcgggggg attgttgctt tttcggataa aatcccgtgt tttttcatct gctaggtaaa 5400 atcgccccaa agcgtctgca tcgcggcgat ggcggcgagt ggggcggttt ctgtgcgtaa 5460 aatccgtttt ccgagtgtaa ccgcctgaaa gccggcttca aatgcctgtt gttcttcctg 5520 ttctgtccag ccgccttcgg gcccgaccat aaagacgatt gcgccggacg ggtggcggat 5580 gtcgccgagt ttgcaggcgc ggttgatgct cataatcagc ttggtgtttt cagacggcat 5640 tttgtcgagt gcttcacggt agccgatgat gggcagtacg gggggaacgg tgttcctgcc 5700 gctttgttcg cacgcggaga tgacgatttc ctgccagcgt gcgaggcgtt tggcggcgcg 5760 ttctccgtcg aggcggacga tgcagcgttc gctgatgacg ggctgtatgg cggttacgcc 5820 gagttcgacg cttttttgca gggtgaaatc catgcgatc 5859 2 126 PRT Neisseria gonorrhoeae 2 Leu Gln Ala Val Ala Val Phe Lys Gln Leu Pro Glu Ala Ala Ala Leu 1 5 10 15 Ala Ala Ala Asn Lys Arg Val Gln Asn Leu Leu Lys Lys Ala Asp Ala 20 25 30 Ala Leu Gly Glu Val Asn Glu Ser Leu Leu Gln Gln Asp Glu Glu Lys 35 40 45 Ala Leu Tyr Ala Ala Ala Gln Gly Leu Gln Pro Lys Ile Ala Ala Ala 50 55 60 Val Ala Glu Gly Asn Phe Arg Thr Ala Leu Ser Glu Leu Ala Ser Val 65 70 75 80 Lys Pro Gln Val Asp Ala Phe Phe Asp Gly Val Met Val Met Ala Glu 85 90 95 Asp Ala Ala Val Lys Gln Asn Arg Leu Asn Leu Leu Asn Arg Leu Ala 100 105 110 Glu Gln Met Asn Ala Val Ala Asp Ile Ala Leu Leu Gly Glu 115 120 125 3 348 PRT Neisseria gonorrhoeae 3 Leu Gln Pro Leu Val Ser Val Leu Ile Cys Ala Tyr Asn Val Glu Lys 1 5 10 15 Tyr Phe Ala Gln Ser Leu Ala Ala Val Val Asn Gln Thr Trp Arg Asn 20 25 30 Leu Asp Ile Leu Ile Val Asp Asp Gly Ser Thr Asp Gly Thr Leu Ala 35 40 45 Ile Ala Lys Asp Phe Gln Lys Arg Asp Ser Arg Ile Lys Ile Leu Ala 50 55 60 Gln Ala Gln Asn Ser Gly Leu Ile Pro Ser Leu Asn Ile Gly Leu Asp 65 70 75 80 Glu Leu Ala Lys Ser Gly Gly Gly Gly Gly Glu Tyr Ile Ala Arg Thr 85 90 95 Asp Ala Asp Asp Ile Ala Ser Pro Gly Trp Ile Glu Lys Ile Val Gly 100 105 110 Glu Met Glu Lys Asp Arg Ser Ile Ile Ala Met Gly Ala Trp Leu Glu 115 120 125 Val Leu Ser Glu Glu Lys Asp Gly Asn Arg Leu Ala Arg His His Lys 130 135 140 His Gly Lys Ile Trp Lys Lys Pro Thr Arg His Glu Asp Ile Ala Ala 145 150 155 160 Phe Phe Pro Phe Gly Asn Pro Ile His Asn Asn Thr Met Ile Met Arg 165 170 175 Arg Ser Val Ile Asp Gly Gly Leu Arg Tyr Asp Thr Glu Arg Asp Trp 180 185 190 Ala Glu Asp Tyr Gln Phe Trp Tyr Asp Val Ser Lys Leu Gly Arg Leu 195 200 205 Ala Tyr Tyr Pro Glu Ala Leu Val Lys Tyr Arg Leu His Ala Asn Gln 210 215 220 Val Ser Ser Lys His Ser Val Arg Gln His Glu Ile Ala Gln Gly Ile 225 230 235 240 Gln Lys Thr Ala Arg Asn Asp Phe Leu Gln Ser Met Gly Phe Lys Thr 245 250 255 Arg Phe Asp Ser Leu Glu Tyr Arg Gln Thr Lys Ala Ala Ala Tyr Glu 260 265 270 Leu Pro Glu Lys Asp Leu Pro Glu Glu Asp Phe Glu Arg Ala Arg Arg 275 280 285 Phe Leu Tyr Gln Cys Phe Lys Arg Thr Asp Thr Pro Pro Ser Gly Ala 290 295 300 Trp Leu Asp Phe Ala Ala Asp Gly Arg Met Arg Arg Leu Phe Thr Leu 305 310 315 320 Arg Gln Tyr Phe Gly Ile Leu Tyr Arg Leu Ile Lys Asn Arg Arg Gln 325 330 335 Ala Arg Ser Asp Ser Ala Gly Lys Glu Gln Glu Ile 340 345 4 306 PRT Neisseria gonorrhoeae 4 Met Asp Ile Val Phe Ala Ala Asp Asp Asn Tyr Ala Ala Tyr Leu Cys 1 5 10 15 Val Ala Ala Lys Ser Val Glu Ala Ala His Pro Asp Thr Glu Ile Arg 20 25 30 Phe His Val Leu Asp Ala Gly Ile Ser Glu Glu Asn Arg Ala Ala Val 35 40 45 Ala Ala Asn Leu Arg Gly Gly Gly Asn Ile Arg Phe Ile Asp Val Asn 50 55 60 Pro Glu Asp Phe Ala Gly Phe Pro Leu Asn Ile Arg His Ile Ser Ile 65 70 75 80 Thr Thr Tyr Ala Arg Leu Lys Leu Gly Glu Tyr Ile Ala Asp Cys Asp 85 90 95 Lys Val Leu Tyr Leu Asp Thr Asp Val Leu Val Arg Asp Gly Leu Lys 100 105 110 Pro Leu Trp Asp Thr Asp Leu Gly Gly Asn Trp Val Gly Ala Cys Ile 115 120 125 Asp Leu Phe Val Glu Arg Gln Glu Gly Tyr Lys Gln Lys Ile Gly Met 130 135 140 Ala Asp Gly Glu Tyr Tyr Phe Asn Ala Gly Val Leu Leu Ile Asn Leu 145 150 155 160 Lys Lys Trp Arg Arg His Asp Ile Phe Lys Met Ser Cys Glu Trp Val 165 170 175 Glu Gln Tyr Lys Asp Val Met Gln Tyr Gln Asp Gln Asp Ile Leu Asn 180 185 190 Gly Leu Phe Lys Gly Gly Val Cys Tyr Ala Asn Ser Arg Phe Asn Phe 195 200 205 Met Pro Thr Asn Tyr Ala Phe Met Ala Asn Gly Phe Ala Ser Arg His 210 215 220 Thr Asp Pro Leu Tyr Leu Asp Arg Thr Asn Thr Ala Met Pro Val Ala 225 230 235 240 Val Ser His Tyr Cys Gly Ser Ala Lys Pro Trp His Arg Asp Cys Thr 245 250 255 Val Trp Gly Ala Glu Arg Phe Thr Glu Leu Ala Gly Ser Leu Thr Thr 260 265 270 Val Pro Glu Glu Trp Arg Gly Lys Leu Ala Val Pro Pro Thr Lys Cys 275 280 285 Met Leu Gln Arg Trp Arg Lys Lys Leu Ser Ala Arg Phe Leu Arg Lys 290 295 300 Ile Tyr 305 5 337 PRT Neisseria gonorrhoeae 5 Leu Gln Pro Leu Val Ser Val Leu Ile Cys Ala Tyr Asn Ala Glu Lys 1 5 10 15 Tyr Phe Ala Gln Ser Leu Ala Ala Val Val Gly Gln Thr Trp Arg Asn 20 25 30 Leu Asp Ile Leu Ile Val Asp Asp Gly Ser Thr Asp Gly Thr Pro Ala 35 40 45 Ile Ala Arg His Phe Gln Glu Gln Asp Gly Arg Ile Arg Ile Ile Ser 50 55 60 Asn Pro Arg Asn Leu Gly Phe Ile Ala Ser Leu Asn Ile Gly Leu Asp 65 70 75 80 Glu Leu Ala Lys Ser Gly Gly Gly Glu Tyr Ile Ala Arg Thr Asp Ala 85 90 95 Asp Asp Ile Ala Ser Pro Gly Trp Ile Glu Lys Ile Val Gly Glu Met 100 105 110 Glu Lys Asp Arg Ser Ile Ile Ala Met Gly Ala Trp Leu Glu Val Leu 115 120 125 Ser Glu Glu Asn Asn Lys Ser Val Leu Ala Ala Ile Ala Arg Asn Gly 130 135 140 Ala Ile Trp Asp Lys Pro Thr Arg His Glu Asp Ile Val Ala Val Phe 145 150 155 160 Pro Phe Gly Asn Pro Ile His Asn Asn Thr Met Ile Met Arg Arg Ser 165 170 175 Val Ile Asp Gly Gly Leu Arg Phe Asp Pro Ala Tyr Ile His Ala Glu 180 185 190 Asp Tyr Lys Phe Trp Tyr Glu Ala Gly Lys Leu Gly Arg Leu Ala Tyr 195 200 205 Tyr Pro Glu Ala Leu Val Lys Tyr Arg Phe His Gln Asp Gln Thr Ser 210 215 220 Ser Lys Tyr Asn Leu Gln Gln Arg Arg Thr Ala Trp Lys Ile Lys Glu 225 230 235 240 Glu Ile Arg Ala Gly Tyr Trp Lys Ala Ala Gly Ile Ala Val Gly Ala 245 250 255 Asp Cys Leu Asn Tyr Gly Leu Leu Lys Ser Thr Ala Tyr Ala Leu Tyr 260 265 270 Glu Lys Ala Leu Ser Gly Gln Asp Ile Gly Cys Leu Arg Leu Phe Leu 275 280 285 Tyr Glu Tyr Phe Leu Ser Leu Glu Lys Tyr Ser Leu Thr Asp Leu Leu 290 295 300 Asp Phe Leu Thr Asp Arg Val Met Arg Lys Leu Phe Ala Ala Pro Gln 305 310 315 320 Tyr Arg Lys Ile Leu Lys Lys Met Leu Arg Pro Trp Lys Tyr Arg Ser 325 330 335 Tyr 6 280 PRT Neisseria gonorrhoeae 6 Met Gln Asn His Val Ile Ser Leu Ala Ser Ala Ala Glu Arg Arg Ala 1 5 10 15 His Ile Ala Asp Thr Phe Gly Ser Arg Gly Ile Pro Phe Gln Phe Phe 20 25 30 Asp Ala Leu Met Pro Ser Glu Arg Leu Glu Gln Ala Met Ala Glu Leu 35 40 45 Val Pro Gly Leu Ser Ala His Pro Tyr Leu Ser Gly Val Glu Lys Ala 50 55 60 Cys Phe Met Ser His Ala Val Leu Trp Glu Gln Ala Leu Asp Glu Gly 65 70 75 80 Leu Pro Tyr Ile Ala Val Phe Glu Asp Asp Val Leu Leu Gly Glu Gly 85 90 95 Ala Glu Gln Phe Leu Ala Glu Asp Thr Trp Leu Glu Glu Arg Phe Asp 100 105 110 Lys Asp Ser Ala Phe Ile Val Arg Leu Glu Thr Met Phe Ala Lys Val 115 120 125 Ile Val Arg Pro Asp Lys Val Leu Asn Tyr Glu Asn Arg Ser Phe Pro 130 135 140 Leu Leu Glu Ser Glu His Cys Gly Thr Ala Gly Tyr Ile Ile Ser Arg 145 150 155 160 Glu Ala Met Arg Phe Phe Leu Asp Arg Phe Ala Val Leu Pro Pro Glu 165 170 175 Arg Ile Lys Ala Val Asp Leu Met Met Phe Thr Tyr Phe Phe Asp Lys 180 185 190 Glu Gly Met Pro Val Tyr Gln Val Ser Pro Ala Leu Cys Thr Gln Glu 195 200 205 Leu His Tyr Ala Lys Phe Leu Ser Gln Asn Ser Met Leu Gly Ser Asp 210 215 220 Leu Glu Lys Asp Arg Glu Gln Gly Arg Arg His Arg Arg Ser Leu Lys 225 230 235 240 Val Met Phe Asp Leu Lys Arg Ala Leu Gly Lys Phe Gly Arg Glu Lys 245 250 255 Lys Lys Arg Met Glu Arg Gln Arg Gln Ala Glu Leu Glu Lys Val Tyr 260 265 270 Gly Arg Arg Val Ile Leu Phe Lys 275 280 7 6 PRT Neisseria gonorrhoeae 7 Tyr Ser Arg Asp Ser Ser 1 5 8 348 PRT Neisseria gonorrhoeae 8 Leu Gln Pro Leu Val Ser Val Leu Ile Cys Ala Tyr Asn Val Glu Lys 1 5 10 15 Tyr Phe Ala Gln Ser Leu Ala Ala Val Val Asn Gln Thr Trp Arg Asn 20 25 30 Leu Asp Ile Leu Ile Val Asp Asp Gly Ser Thr Asp Gly Thr Leu Ala 35 40 45 Ile Ala Lys Asp Phe Gln Lys Arg Asp Ser Arg Ile Lys Ile Leu Ala 50 55 60 Gln Ala Gln Asn Ser Gly Leu Ile Pro Ser Leu Asn Ile Gly Leu Asp 65 70 75 80 Glu Leu Ala Lys Ser Gly Gly Gly Gly Gly Glu Tyr Ile Ala Arg Thr 85 90 95 Asp Ala Asp Asp Ile Ala Ser Pro Gly Trp Ile Glu Lys Ile Val Gly 100 105 110 Glu Met Glu Lys Asp Arg Ser Ile Ile Ala Met Gly Ala Trp Leu Glu 115 120 125 Val Leu Ser Glu Glu Lys Asp Gly Asn Arg Leu Ala Arg His His Lys 130 135 140 His Gly Lys Ile Trp Lys Lys Pro Thr Arg His Glu Asp Ile Ala Ala 145 150 155 160 Phe Phe Pro Phe Gly Asn Pro Ile His Asn Asn Thr Met Ile Met Arg 165 170 175 Arg Ser Val Ile Asp Gly Gly Leu Arg Tyr Asp Thr Glu Arg Asp Trp 180 185 190 Ala Glu Asp Tyr Gln Phe Trp Tyr Asp Val Ser Lys Leu Gly Arg Leu 195 200 205 Ala Tyr Tyr Pro Glu Ala Leu Val Lys Tyr Arg Leu His Ala Asn Gln 210 215 220 Val Ser Ser Lys His Ser Val Arg Gln His Glu Ile Ala Gln Gly Ile 225 230 235 240 Gln Lys Thr Ala Arg Asn Asp Phe Leu Gln Ser Met Gly Phe Lys Thr 245 250 255 Arg Phe Asp Ser Leu Glu Tyr Arg Gln Thr Lys Ala Ala Ala Tyr Glu 260 265 270 Leu Pro Glu Lys Asp Leu Pro Glu Glu Asp Phe Glu Arg Ala Arg Arg 275 280 285 Phe Leu Tyr Gln Cys Phe Lys Arg Thr Asp Thr Pro Pro Ser Gly Ala 290 295 300 Trp Leu Asp Phe Ala Ala Asp Gly Arg Met Arg Arg Leu Phe Thr Leu 305 310 315 320 Arg Gln Tyr Phe Gly Ile Leu Tyr Arg Leu Ile Lys Asn Arg Arg Gln 325 330 335 Ala Arg Ser Asp Ser Ala Gly Lys Glu Gln Glu Ile 340 345
Claims (10)
1. A method of transferring at least two saccharide units with a single enzyme comprising contacting an acceptor moiety and two different sugar donors with a polyglycosyltransferase.
2. The method of claim 1 , wherein said two saccharide units are N-acetylglucosamine and N-acetylgalactosamine.
3. The method of claim 1 , wherein said acceptor moiety has a galactose at a non-reducing end.
4. The method of claim 1 , wherein said polyglycosyltransferase is isolated from Neisseria.
5. The method of claim 1 , wherein said polyglycosyltransferase is isolated from Neisseria gonorrhoeae.
6. A method of transferring N-acetylgalactosamine, to an acceptor moiety, comprising contacting an N-acetylgalactosamine donor and an acceptor moiety with an N-acetylglucosaminyl transferase.
7. The method of claim 6 , wherein said N-acetylglucosaminyl transferase is isolated from Neisseria.
8. The method of claim 6 , wherein said N-acetylglucosaminyl transferase is isolated from Neisseria gonorrhoeae.
9. The method of claim 6 , wherein said acceptor moiety has a galactose at a non-reducing end.
10. The method of claim 8 , wherein said Neisseria gonorrhoeae is ATCC 33084.
Priority Applications (1)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US10/096,129 US20030207406A1 (en) | 1995-06-07 | 2002-03-07 | Method of transferring at least two saccharide units with a polyglycosyltransferase |
Applications Claiming Priority (3)
| Application Number | Priority Date | Filing Date | Title |
|---|---|---|---|
| US08/478,140 US6127153A (en) | 1995-06-07 | 1995-06-07 | Method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and gene encoding a polyglycosyltransferase |
| US09/338,943 US6379933B1 (en) | 1995-06-07 | 1999-06-24 | Method of transferring at least two saccharide units with a polyglycosyltransferase |
| US10/096,129 US20030207406A1 (en) | 1995-06-07 | 2002-03-07 | Method of transferring at least two saccharide units with a polyglycosyltransferase |
Related Parent Applications (1)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| US09/338,943 Continuation US6379933B1 (en) | 1995-06-07 | 1999-06-24 | Method of transferring at least two saccharide units with a polyglycosyltransferase |
Publications (1)
| Publication Number | Publication Date |
|---|---|
| US20030207406A1 true US20030207406A1 (en) | 2003-11-06 |
Family
ID=23898699
Family Applications (3)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| US08/478,140 Expired - Fee Related US6127153A (en) | 1995-06-07 | 1995-06-07 | Method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and gene encoding a polyglycosyltransferase |
| US09/338,943 Expired - Fee Related US6379933B1 (en) | 1995-06-07 | 1999-06-24 | Method of transferring at least two saccharide units with a polyglycosyltransferase |
| US10/096,129 Abandoned US20030207406A1 (en) | 1995-06-07 | 2002-03-07 | Method of transferring at least two saccharide units with a polyglycosyltransferase |
Family Applications Before (2)
| Application Number | Title | Priority Date | Filing Date |
|---|---|---|---|
| US08/478,140 Expired - Fee Related US6127153A (en) | 1995-06-07 | 1995-06-07 | Method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and gene encoding a polyglycosyltransferase |
| US09/338,943 Expired - Fee Related US6379933B1 (en) | 1995-06-07 | 1999-06-24 | Method of transferring at least two saccharide units with a polyglycosyltransferase |
Country Status (6)
| Country | Link |
|---|---|
| US (3) | US6127153A (en) |
| EP (1) | EP0832277A4 (en) |
| JP (1) | JPH11506927A (en) |
| AU (1) | AU709942B2 (en) |
| CA (1) | CA2224144A1 (en) |
| WO (1) | WO1996040971A1 (en) |
Cited By (31)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| WO2006050247A2 (en) | 2004-10-29 | 2006-05-11 | Neose Technologies, Inc. | Remodeling and glycopegylation of fibroblast growth factor (fgf) |
| US20070154992A1 (en) * | 2005-04-08 | 2007-07-05 | Neose Technologies, Inc. | Compositions and methods for the preparation of protease resistant human growth hormone glycosylation mutants |
| US7842661B2 (en) | 2003-11-24 | 2010-11-30 | Novo Nordisk A/S | Glycopegylated erythropoietin formulations |
| US7932364B2 (en) | 2003-05-09 | 2011-04-26 | Novo Nordisk A/S | Compositions and methods for the preparation of human growth hormone glycosylation mutants |
| US7956032B2 (en) | 2003-12-03 | 2011-06-07 | Novo Nordisk A/S | Glycopegylated granulocyte colony stimulating factor |
| US8008252B2 (en) | 2001-10-10 | 2011-08-30 | Novo Nordisk A/S | Factor VII: remodeling and glycoconjugation of Factor VII |
| US8053410B2 (en) | 2002-06-21 | 2011-11-08 | Novo Nordisk Health Care A/G | Pegylated factor VII glycoforms |
| US8063015B2 (en) | 2003-04-09 | 2011-11-22 | Novo Nordisk A/S | Glycopegylation methods and proteins/peptides produced by the methods |
| US8076292B2 (en) | 2001-10-10 | 2011-12-13 | Novo Nordisk A/S | Factor VIII: remodeling and glycoconjugation of factor VIII |
| US8207112B2 (en) | 2007-08-29 | 2012-06-26 | Biogenerix Ag | Liquid formulation of G-CSF conjugate |
| US8247381B2 (en) | 2003-03-14 | 2012-08-21 | Biogenerix Ag | Branched water-soluble polymers and their conjugates |
| US8268967B2 (en) | 2004-09-10 | 2012-09-18 | Novo Nordisk A/S | Glycopegylated interferon α |
| US8361961B2 (en) | 2004-01-08 | 2013-01-29 | Biogenerix Ag | O-linked glycosylation of peptides |
| US8404809B2 (en) | 2005-05-25 | 2013-03-26 | Novo Nordisk A/S | Glycopegylated factor IX |
| US8633157B2 (en) | 2003-11-24 | 2014-01-21 | Novo Nordisk A/S | Glycopegylated erythropoietin |
| US8632770B2 (en) | 2003-12-03 | 2014-01-21 | Novo Nordisk A/S | Glycopegylated factor IX |
| US8716240B2 (en) | 2001-10-10 | 2014-05-06 | Novo Nordisk A/S | Erythropoietin: remodeling and glycoconjugation of erythropoietin |
| US8716239B2 (en) | 2001-10-10 | 2014-05-06 | Novo Nordisk A/S | Granulocyte colony stimulating factor: remodeling and glycoconjugation G-CSF |
| US8791066B2 (en) | 2004-07-13 | 2014-07-29 | Novo Nordisk A/S | Branched PEG remodeling and glycosylation of glucagon-like peptide-1 [GLP-1] |
| US8791070B2 (en) | 2003-04-09 | 2014-07-29 | Novo Nordisk A/S | Glycopegylated factor IX |
| US8841439B2 (en) | 2005-11-03 | 2014-09-23 | Novo Nordisk A/S | Nucleotide sugar purification using membranes |
| US8911967B2 (en) | 2005-08-19 | 2014-12-16 | Novo Nordisk A/S | One pot desialylation and glycopegylation of therapeutic peptides |
| US8916360B2 (en) | 2003-11-24 | 2014-12-23 | Novo Nordisk A/S | Glycopegylated erythropoietin |
| US8969532B2 (en) | 2006-10-03 | 2015-03-03 | Novo Nordisk A/S | Methods for the purification of polypeptide conjugates comprising polyalkylene oxide using hydrophobic interaction chromatography |
| US9005625B2 (en) | 2003-07-25 | 2015-04-14 | Novo Nordisk A/S | Antibody toxin conjugates |
| US9029331B2 (en) | 2005-01-10 | 2015-05-12 | Novo Nordisk A/S | Glycopegylated granulocyte colony stimulating factor |
| US9050304B2 (en) | 2007-04-03 | 2015-06-09 | Ratiopharm Gmbh | Methods of treatment using glycopegylated G-CSF |
| US9150848B2 (en) | 2008-02-27 | 2015-10-06 | Novo Nordisk A/S | Conjugated factor VIII molecules |
| US9187532B2 (en) | 2006-07-21 | 2015-11-17 | Novo Nordisk A/S | Glycosylation of peptides via O-linked glycosylation sequences |
| US9493499B2 (en) | 2007-06-12 | 2016-11-15 | Novo Nordisk A/S | Process for the production of purified cytidinemonophosphate-sialic acid-polyalkylene oxide (CMP-SA-PEG) as modified nucleotide sugars via anion exchange chromatography |
| WO2019043457A2 (en) | 2017-09-04 | 2019-03-07 | 89Bio Ltd. | Mutant fgf-21 peptide conjugates and uses thereof |
Families Citing this family (13)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US6127153A (en) * | 1995-06-07 | 2000-10-03 | Neose Technologies, Inc. | Method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and gene encoding a polyglycosyltransferase |
| DE19735994A1 (en) * | 1997-08-19 | 1999-02-25 | Piepersberg Wolfgang Prof Dr | Synthesizing guanosine |
| US7244601B2 (en) | 1997-12-15 | 2007-07-17 | National Research Council Of Canada | Fusion proteins for use in enzymatic synthesis of oligosaccharides |
| CN1297907A (en) * | 1999-11-24 | 2001-06-06 | 上海博容基因开发有限公司 | Human acetylglactoside transferase 45 as one new kind of polypeptide and polynucleotides encoding this polypeptide |
| BRPI0519562A2 (en) | 2004-12-27 | 2009-01-27 | Baxter Int | proteinaceous construct, complex, method for prolonging the in vivo half-life of factor viii (fviii) or a biologically active derivative of the same pharmaceutical composition, and method for forming a proteinaceous construct |
| US20090196850A1 (en) | 2005-01-06 | 2009-08-06 | Novo Nordisk A/S | Anti-Kir Combination Treatments and Methods |
| JP2006223198A (en) * | 2005-02-17 | 2006-08-31 | National Institute Of Advanced Industrial & Technology | Compound library using transferase and method for producing the same |
| NZ572050A (en) | 2006-03-31 | 2011-09-30 | Baxter Int | Factor VIII conjugated to polyethylene glycol |
| US7645860B2 (en) | 2006-03-31 | 2010-01-12 | Baxter Healthcare S.A. | Factor VIII polymer conjugates |
| ES2515365T3 (en) | 2007-12-27 | 2014-10-29 | Baxter International Inc. | Method and compositions for specifically detecting physiologically acceptable polymer molecules |
| EP3423587B1 (en) | 2016-04-05 | 2020-06-24 | Council of Scientific & Industrial Research | A multifunctional recombinant nucleotide dependent glycosyltransferase protein and its method of glycosylation thereof |
| EP4263851A4 (en) * | 2020-12-17 | 2025-03-26 | Amyris, Inc. | MODIFIED ß-1,3-N-ACETYLGLUCOSAMINYLTRANSFERASE POLYPEPTIDES |
| CN114958791B (en) * | 2021-01-08 | 2023-07-28 | 中国科学院华南植物园 | Spermidine derivative glycosyltransferase LbUGT62 and encoding gene and application thereof |
Family Cites Families (6)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| SE451849B (en) * | 1985-12-11 | 1987-11-02 | Svenska Sockerfabriks Ab | VIEW TO SYNTHETIZE GYCLOSIDIC BINDINGS AND USE OF THIS RECEIVED PRODUCTS |
| US4925796A (en) * | 1986-03-07 | 1990-05-15 | Massachusetts Institute Of Technology | Method for enhancing glycoprotein stability |
| EP0481038B1 (en) * | 1990-04-16 | 2002-10-02 | The Trustees Of The University Of Pennsylvania | Saccharide compositions, methods and apparatus for their synthesis |
| US5308460A (en) * | 1992-10-30 | 1994-05-03 | Glyko, Incorporated | Rapid synthesis and analysis of carbohydrates |
| US5545553A (en) * | 1994-09-26 | 1996-08-13 | The Rockefeller University | Glycosyltransferases for biosynthesis of oligosaccharides, and genes encoding them |
| US6127153A (en) * | 1995-06-07 | 2000-10-03 | Neose Technologies, Inc. | Method of transferring at least two saccharide units with a polyglycosyltransferase, a polyglycosyltransferase and gene encoding a polyglycosyltransferase |
-
1995
- 1995-06-07 US US08/478,140 patent/US6127153A/en not_active Expired - Fee Related
-
1996
- 1996-06-03 AU AU59655/96A patent/AU709942B2/en not_active Ceased
- 1996-06-03 EP EP96916946A patent/EP0832277A4/en not_active Withdrawn
- 1996-06-03 WO PCT/US1996/008323 patent/WO1996040971A1/en not_active Ceased
- 1996-06-03 CA CA002224144A patent/CA2224144A1/en not_active Abandoned
- 1996-06-03 JP JP9500960A patent/JPH11506927A/en active Pending
-
1999
- 1999-06-24 US US09/338,943 patent/US6379933B1/en not_active Expired - Fee Related
-
2002
- 2002-03-07 US US10/096,129 patent/US20030207406A1/en not_active Abandoned
Cited By (38)
| Publication number | Priority date | Publication date | Assignee | Title |
|---|---|---|---|---|
| US8716240B2 (en) | 2001-10-10 | 2014-05-06 | Novo Nordisk A/S | Erythropoietin: remodeling and glycoconjugation of erythropoietin |
| US8008252B2 (en) | 2001-10-10 | 2011-08-30 | Novo Nordisk A/S | Factor VII: remodeling and glycoconjugation of Factor VII |
| US8076292B2 (en) | 2001-10-10 | 2011-12-13 | Novo Nordisk A/S | Factor VIII: remodeling and glycoconjugation of factor VIII |
| US8716239B2 (en) | 2001-10-10 | 2014-05-06 | Novo Nordisk A/S | Granulocyte colony stimulating factor: remodeling and glycoconjugation G-CSF |
| US8053410B2 (en) | 2002-06-21 | 2011-11-08 | Novo Nordisk Health Care A/G | Pegylated factor VII glycoforms |
| US8247381B2 (en) | 2003-03-14 | 2012-08-21 | Biogenerix Ag | Branched water-soluble polymers and their conjugates |
| US8853161B2 (en) | 2003-04-09 | 2014-10-07 | Novo Nordisk A/S | Glycopegylation methods and proteins/peptides produced by the methods |
| US8063015B2 (en) | 2003-04-09 | 2011-11-22 | Novo Nordisk A/S | Glycopegylation methods and proteins/peptides produced by the methods |
| US8791070B2 (en) | 2003-04-09 | 2014-07-29 | Novo Nordisk A/S | Glycopegylated factor IX |
| US7932364B2 (en) | 2003-05-09 | 2011-04-26 | Novo Nordisk A/S | Compositions and methods for the preparation of human growth hormone glycosylation mutants |
| US9005625B2 (en) | 2003-07-25 | 2015-04-14 | Novo Nordisk A/S | Antibody toxin conjugates |
| US8916360B2 (en) | 2003-11-24 | 2014-12-23 | Novo Nordisk A/S | Glycopegylated erythropoietin |
| US7842661B2 (en) | 2003-11-24 | 2010-11-30 | Novo Nordisk A/S | Glycopegylated erythropoietin formulations |
| US8633157B2 (en) | 2003-11-24 | 2014-01-21 | Novo Nordisk A/S | Glycopegylated erythropoietin |
| US8632770B2 (en) | 2003-12-03 | 2014-01-21 | Novo Nordisk A/S | Glycopegylated factor IX |
| US7956032B2 (en) | 2003-12-03 | 2011-06-07 | Novo Nordisk A/S | Glycopegylated granulocyte colony stimulating factor |
| US8361961B2 (en) | 2004-01-08 | 2013-01-29 | Biogenerix Ag | O-linked glycosylation of peptides |
| US8791066B2 (en) | 2004-07-13 | 2014-07-29 | Novo Nordisk A/S | Branched PEG remodeling and glycosylation of glucagon-like peptide-1 [GLP-1] |
| US8268967B2 (en) | 2004-09-10 | 2012-09-18 | Novo Nordisk A/S | Glycopegylated interferon α |
| US9200049B2 (en) | 2004-10-29 | 2015-12-01 | Novo Nordisk A/S | Remodeling and glycopegylation of fibroblast growth factor (FGF) |
| WO2006050247A2 (en) | 2004-10-29 | 2006-05-11 | Neose Technologies, Inc. | Remodeling and glycopegylation of fibroblast growth factor (fgf) |
| EP2586456A1 (en) | 2004-10-29 | 2013-05-01 | BioGeneriX AG | Remodeling and glycopegylation of fibroblast growth factor (FGF) |
| EP3061461A1 (en) | 2004-10-29 | 2016-08-31 | ratiopharm GmbH | Remodeling and glycopegylation of fibroblast growth factor (fgf) |
| US10874714B2 (en) | 2004-10-29 | 2020-12-29 | 89Bio Ltd. | Method of treating fibroblast growth factor 21 (FGF-21) deficiency |
| US9029331B2 (en) | 2005-01-10 | 2015-05-12 | Novo Nordisk A/S | Glycopegylated granulocyte colony stimulating factor |
| US9187546B2 (en) | 2005-04-08 | 2015-11-17 | Novo Nordisk A/S | Compositions and methods for the preparation of protease resistant human growth hormone glycosylation mutants |
| US20070154992A1 (en) * | 2005-04-08 | 2007-07-05 | Neose Technologies, Inc. | Compositions and methods for the preparation of protease resistant human growth hormone glycosylation mutants |
| EP2386571A2 (en) | 2005-04-08 | 2011-11-16 | BioGeneriX AG | Compositions and methods for the preparation of protease resistant human growth hormone glycosylation mutants |
| US8404809B2 (en) | 2005-05-25 | 2013-03-26 | Novo Nordisk A/S | Glycopegylated factor IX |
| US8911967B2 (en) | 2005-08-19 | 2014-12-16 | Novo Nordisk A/S | One pot desialylation and glycopegylation of therapeutic peptides |
| US8841439B2 (en) | 2005-11-03 | 2014-09-23 | Novo Nordisk A/S | Nucleotide sugar purification using membranes |
| US9187532B2 (en) | 2006-07-21 | 2015-11-17 | Novo Nordisk A/S | Glycosylation of peptides via O-linked glycosylation sequences |
| US8969532B2 (en) | 2006-10-03 | 2015-03-03 | Novo Nordisk A/S | Methods for the purification of polypeptide conjugates comprising polyalkylene oxide using hydrophobic interaction chromatography |
| US9050304B2 (en) | 2007-04-03 | 2015-06-09 | Ratiopharm Gmbh | Methods of treatment using glycopegylated G-CSF |
| US9493499B2 (en) | 2007-06-12 | 2016-11-15 | Novo Nordisk A/S | Process for the production of purified cytidinemonophosphate-sialic acid-polyalkylene oxide (CMP-SA-PEG) as modified nucleotide sugars via anion exchange chromatography |
| US8207112B2 (en) | 2007-08-29 | 2012-06-26 | Biogenerix Ag | Liquid formulation of G-CSF conjugate |
| US9150848B2 (en) | 2008-02-27 | 2015-10-06 | Novo Nordisk A/S | Conjugated factor VIII molecules |
| WO2019043457A2 (en) | 2017-09-04 | 2019-03-07 | 89Bio Ltd. | Mutant fgf-21 peptide conjugates and uses thereof |
Also Published As
| Publication number | Publication date |
|---|---|
| CA2224144A1 (en) | 1996-12-19 |
| US6127153A (en) | 2000-10-03 |
| EP0832277A4 (en) | 2000-10-11 |
| EP0832277A1 (en) | 1998-04-01 |
| US6379933B1 (en) | 2002-04-30 |
| WO1996040971A1 (en) | 1996-12-19 |
| AU709942B2 (en) | 1999-09-09 |
| AU5965596A (en) | 1996-12-30 |
| JPH11506927A (en) | 1999-06-22 |
Similar Documents
| Publication | Publication Date | Title |
|---|---|---|
| AU709942B2 (en) | Method of transferring at least two saccharide units with a polyglycosyltransferase | |
| US6780624B2 (en) | Glycosyltransferases for biosynthesis of oligosaccharides, and genes encoding them | |
| CA2493258C (en) | Synthesis of oligosaccharides, glycolipids, and glycoproteins using bacterial glycosyltransferases | |
| JP4460615B2 (en) | Campylobacter glycosyltransferase for the biosynthesis of gangliosides and ganglioside mimetics | |
| JP2003047467A (en) | Chondroitin synthetase | |
| MXPA97009505A (en) | Method to transfer at least two units of success with a poliglicoslitransfer | |
| AU714684C (en) | Glycosyltransferases for biosynthesis of oligosaccharides, and genes encoding them | |
| WO1998054331A2 (en) | High level expression of glycosyltransferases | |
| JP2002306184A (en) | Method for producing n-acetylglucosaminyl transferase |
Legal Events
| Date | Code | Title | Description |
|---|---|---|---|
| STCB | Information on status: application discontinuation |
Free format text: ABANDONED -- FAILURE TO RESPOND TO AN OFFICE ACTION |