Matheus Pedron Cassol, Alexandre Rafael Lenz, Rudinei Zacaria, Scheila de Avila e Silva
{"title":"基因组测序和注释的计算环境:生命研究人员项目应用的工作流程","authors":"Matheus Pedron Cassol, Alexandre Rafael Lenz, Rudinei Zacaria, Scheila de Avila e Silva","doi":"10.36995/j.recyt.2022.38.008","DOIUrl":null,"url":null,"abstract":"The current paper seeks to approach, using a workflow, basics subjects of the bioinformatic field and also useful informations to consider during the development of in silico researches. Installation and general usage of multiple softwares related to different sections of the genome annotation process were also presented. At last, an model organism, Staphylococcus aureus, was sequenced in two different softwares, SPAdes and IDBA-UD, seeking further comparison and evaluation of the process as a whole. The quality evaluation of the assemble was established by tests on QUAST, BUSCO and Augustus, supported by BLASTP. Results: QUAST evaluation returned genome coverage values above 98% in both test cases, pointing towards a trustworthy assemble for this organism. Via SPAdes were needed less computational resources, but, using IDBA-UD the sequences found were more contiguous. Results deriving from BUSCO showed only one expected gene difference. Some proteins and genes predicted by Augustus led to hits, sequences already studied in that organism, using the BLASTP program.","PeriodicalId":21243,"journal":{"name":"Revista de Ciencia y Tecnología","volume":"40 1","pages":""},"PeriodicalIF":0.0000,"publicationDate":"2022-10-31","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":"{\"title\":\"Computational Environment for Genomic Sequencing and Annotation: a workflow for application in projects by life researchers\",\"authors\":\"Matheus Pedron Cassol, Alexandre Rafael Lenz, Rudinei Zacaria, Scheila de Avila e Silva\",\"doi\":\"10.36995/j.recyt.2022.38.008\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"The current paper seeks to approach, using a workflow, basics subjects of the bioinformatic field and also useful informations to consider during the development of in silico researches. Installation and general usage of multiple softwares related to different sections of the genome annotation process were also presented. At last, an model organism, Staphylococcus aureus, was sequenced in two different softwares, SPAdes and IDBA-UD, seeking further comparison and evaluation of the process as a whole. The quality evaluation of the assemble was established by tests on QUAST, BUSCO and Augustus, supported by BLASTP. Results: QUAST evaluation returned genome coverage values above 98% in both test cases, pointing towards a trustworthy assemble for this organism. Via SPAdes were needed less computational resources, but, using IDBA-UD the sequences found were more contiguous. Results deriving from BUSCO showed only one expected gene difference. Some proteins and genes predicted by Augustus led to hits, sequences already studied in that organism, using the BLASTP program.\",\"PeriodicalId\":21243,\"journal\":{\"name\":\"Revista de Ciencia y Tecnología\",\"volume\":\"40 1\",\"pages\":\"\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2022-10-31\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"0\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"Revista de Ciencia y Tecnología\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.36995/j.recyt.2022.38.008\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"Revista de Ciencia y Tecnología","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.36995/j.recyt.2022.38.008","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}
Computational Environment for Genomic Sequencing and Annotation: a workflow for application in projects by life researchers
The current paper seeks to approach, using a workflow, basics subjects of the bioinformatic field and also useful informations to consider during the development of in silico researches. Installation and general usage of multiple softwares related to different sections of the genome annotation process were also presented. At last, an model organism, Staphylococcus aureus, was sequenced in two different softwares, SPAdes and IDBA-UD, seeking further comparison and evaluation of the process as a whole. The quality evaluation of the assemble was established by tests on QUAST, BUSCO and Augustus, supported by BLASTP. Results: QUAST evaluation returned genome coverage values above 98% in both test cases, pointing towards a trustworthy assemble for this organism. Via SPAdes were needed less computational resources, but, using IDBA-UD the sequences found were more contiguous. Results deriving from BUSCO showed only one expected gene difference. Some proteins and genes predicted by Augustus led to hits, sequences already studied in that organism, using the BLASTP program.