International Journal of Molecular Biotechnological Research Original Research

Gene Annotation of Cancer Vaccine for Homo sapiens

  1. Nimilita Chakraborty Department of Biotechnology, KIIT University,Department of Biotechnology, KIIT University, Bhubaneswar

Abstract

Objectives: Gene annotation helps us to deduce the structural and functional aspects of a gene that encodes for a functional protein in our body. Thus, by determining the coding sequence and gene location we can derive meaningful insights as to what these genes do in our body. In this study, an unknown gene, cancer vaccine for Homo sapiens has been studied and annotated.

Methods: This study was based on a computational approach using various web interface tools to annotate an unknown gene taken from the NCBI Database. The chosen gene was structurally annotated using a GC% content calculator, visually represented using Microsoft Excel, Augustus for gene prediction, and RNAfold to determine the mRNA structure of the same gene. Functional annotation was done using BlastP, gene ontology was confirmed using the UniProt database, a phylogenetic tree was analyzed using HOGENOM database and TMHMM to visualize the transmembrane domain of the protein encoded by the gene, expression of the gene by Bgee, antibody analysis, subcellular localization, and functional analysis was accomplished using Human Protein Atlas, wolF PSORT and InterProScan respectively.

Results: After completing the gene annotation, the cancer vaccine for Homo sapiens query was found to be 99.9% similar to four-jointed box protein 1 precursor [Homo sapiens] which exhibits low cancer tissue specificity and is mostly related to renal and urothelial cancer.

Conclusion: The Cancer vaccine for Homo sapiens entry present in the NCBI Database, which had no annotation previously, was annotated structurally and functionally in this study. Now we can say this entry belongs to the gene coding for a four-box jointed protein-1 precursor protein which is useful for cancer diagnosis in the early stages and is related to poor prognosis of the disease. Often specific peptides are designed for FJX-1 protein which are beneficial in the treatment of cancers showing elevated expression of FJX1 proteins and are often used in the form of vaccines.

Keywords

References (22)

  1. Moore JH, Williams SM. New strategies for identifying gene-gene interactions in hypertension. Annals of Medicine. 2002;34(2):88-95. doi:10.1080/07853890252953473
  2. Mercer TR, Mattick JS. Understanding the regulatory and transcriptional complexity of the genome through structure. Genome Research. 2013;23(7):1081-1088. doi:10.1101/gr.156612.113
  3. Puente XS, Sánchez LM, Gutiérrez-Fernández A, Velasco G, López-Otín C. A genomic view of the complexity of mammalian proteolytic systems. Biochemical Society Transactions. 2005;33(2):331-334. doi:10.1042/bst0330331
  4. Pareek CS, Smoczynski R, Tretyn A. Sequencing technologies and genome sequencing. Journal of Applied Genetics. 2011;52(4):413-435. doi:10.1007/s13353-011-0057-x
  5. Ramsey J, Rasche H, Maughmer C, Criscione A, Mijalis E, Liu M, et al. Galaxy and Apollo as a biologist-friendly interface for high-quality cooperative phage genome annotation. PLOS Computational Biology. 2020;16(11):e1008214. doi:10.1371/journal.pcbi.1008214
  6. Afgan E, Baker D, Batut B, van den Beek M, Bouvier D, Čech M, et al. The Galaxy platform for accessible, reproducible and collaborative biomedical analyses: 2018 update. Nucleic Acids Research. 2018;46(W1):W537-W544. doi:10.1093/nar/gky379
  7. Wang Z, Chen Y, Li Y. A Brief Review of Computational Gene Prediction Methods. Genomics, Proteomics & Bioinformatics. 2004;2(4):216-221. doi:10.1016/s1672-0229(04)02028-5
  8. Hoff KJ, Stanke M. Predicting Genes in Single Genomes with AUGUSTUS. Current Protocols in Bioinformatics. 2018;65(1). doi:10.1002/cpbi.57
  9. Chan PP, Lowe TM. tRNAscan-SE: Searching for tRNA Genes in Genomic Sequences. Methods in Molecular Biology. 2019:1-14. doi:10.1007/978-1-4939-9173-0_1
  10. Gruber AR, Lorenz R, Bernhart SH, Neubock R, Hofacker IL. The Vienna RNA Websuite. Nucleic Acids Research. 2008;36(WebServer):W70-W74. doi:10.1093/nar/gkn188
  11. Oehmen C, Nieplocha J. ScalaBLAST: A Scalable Implementation of BLAST for High-Performance Data-Intensive Bioinformatics Analysis. IEEE Transactions on Parallel and Distributed Systems. 2006;17(8):740-749. doi:10.1109/tpds.2006.112
  12. Neumann RS, Kumar S, Shalchian-Tabrizi K. BLAST output visualization in the new sequencing era. Briefings in Bioinformatics. 2013;15(4):484-503. doi:10.1093/bib/bbt009
  13. Sussman JL, Lin D, Jiang J, Manning NO, Prilusky J, Ritter O, et al. Protein Data Bank (PDB): Database of Three-Dimensional Structural Information of Biological Macromolecules. Acta Crystallographica Section D Biological Crystallography. 1998;54(6):1078-1084. doi:10.1107/s0907444998009378
  14. The UniProt Consortium. UniProt: a hub for protein information. Nucleic Acids Research. 2014;43(D1):D204-D212. doi:10.1093/nar/gku989
  15. Jones P, Binns D, Chang HY, Fraser M, Li W, McAnulla C, et al. InterProScan 5: genome-scale protein function classification. Bioinformatics. 2014;30(9):1236-1240. doi:10.1093/bioinformatics/btu031
  16. Magwanga RO, Lu P, Kirungu JN, Cai X, Zhou Z, Wang X, et al. Whole Genome Analysis of Cyclin Dependent Kinase (CDK) Gene Family in Cotton and Functional Evaluation of the Role of CDKF4 Gene in Drought and Salt Stress Tolerance in Plants. International Journal of Molecular Sciences. 2018;19(9):2625. doi:10.3390/ijms19092625
  17. Krogh A, Larsson B, von Heijne G, Sonnhammer ELL. Predicting transmembrane protein topology with a hidden markov model: application to complete genomes11Edited by F. Cohen. Journal of Molecular Biology. 2001;305(3):567-580. doi:10.1006/jmbi.2000.4315
  18. Bastian FB, Roux J, Niknejad A, Comte A, Fonseca Costa SS, de Farias TM, et al. The Bgee suite: integrated curated expression atlas and comparative transcriptomics in animals. Nucleic Acids Research. 2020;49(D1):D831-D847. doi:10.1093/nar/gkaa793
  19. Breitwieser FP, Lu J, Salzberg SL. A review of methods and databases for metagenomic classification and assembly. Briefings in Bioinformatics. 2017;20(4):1125-1136. doi:10.1093/bib/bbx120
  20. Pontén F, Jirström K, Uhlen M. The Human Protein Atlas—a tool for pathology. The Journal of Pathology. 2008;216(4):387-393. doi:10.1002/path.2440
  21. Chai SJ, Ahmad Zabidi MM, Gan SP, Rajadurai P, Lim PVH, Ng CC, et al. An Oncogenic Role for Four-Jointed Box 1 (FJX1) in Nasopharyngeal Carcinoma. Disease Markers. 2019;2019:1-10. doi:10.1155/2019/3857853
  22. Salzberg SL. Next-generation genome annotation: we still struggle to get it right. Genome Biology. 2019;20(1). doi:10.1186/s13059-019-1715-2