Identification of conserved regions from 230,163 SARS?CoV?2 genomes and their use in diagnostic PCR primer design

Cited 2 time in scopus
Metadata Downloads

Full metadata record

DC FieldValueLanguage
dc.contributor.authorHaeyoung Jeong-
dc.contributor.authorS Lee-
dc.contributor.authorJ Ko-
dc.contributor.authorM Ko-
dc.contributor.authorHwi Won Seo-
dc.date.accessioned2022-07-12T05:46:17Z-
dc.date.available2022-07-12T05:46:17Z-
dc.date.issued2022-
dc.identifier.issn1976-9571-
dc.identifier.urihttps://oak.kribb.re.kr/handle/201005/26972-
dc.description.abstractBackground: As the rapidly evolving characteristic of SARS-CoV-2 could result in false negative diagnosis, the use of as much sequence data as possible is key to the identification of conserved viral sequences. However, multiple alignment of massive genome sequences is computationally intensive. Objective: To extract conserved sequences from SARS-CoV-2 genomes for the design of diagnostic PCR primers using a bioinformatics approach that can handle massive genomic sequences efficiently. Methods: A total of 230,163 full-length viral genomes were retrieved from the NCBI SARS-CoV-2 Resources and GISAID EpiCoV database. This number was reduced to 14.11% following removal of 5'-/3'-untranslated regions and sequence dereplication. Fast, reference-based, multiple sequence alignments identified conserved sequences and specific primer sets were designed against these regions using a conventional tool. Primer sets chosen among the candidates were evaluated by in silico PCR and RT-qPCR. Results: Out of 17 conserved sequences (totaling 4.3 kb), two primer sets targeting the nsp2 and ORF3a genes were picked that exhibited > 99.9% in silico amplification coverage against the original dataset (230,163 genomes) when a 5% mismatch between the primers and target was allowed. In addition, the primer sets successfully detected nine SARS-CoV-2 variant RNA samples (Alpha, Beta, Gamma, Delta, Epsilon, Zeta, Eta, Iota, and Kappa) in experimental RT-qPCR validations. Conclusion: In addition to the RdRp, E, N, and S genes that are targeted commonly, our approach can be used to identify novel primer targets in SARS-CoV-2 and should be a priority strategy in the event of novel SARS-CoV-2 variants or other pandemic outbreaks.-
dc.publisherSpringer-
dc.titleIdentification of conserved regions from 230,163 SARS?CoV?2 genomes and their use in diagnostic PCR primer design-
dc.title.alternativeIdentification of conserved regions from 230,163 SARS?CoV?2 genomes and their use in diagnostic PCR primer design-
dc.typeArticle-
dc.citation.titleGenes & Genomics-
dc.citation.number8-
dc.citation.endPage912-
dc.citation.startPage899-
dc.citation.volume44-
dc.contributor.affiliatedAuthorHaeyoung Jeong-
dc.contributor.affiliatedAuthorHwi Won Seo-
dc.contributor.alternativeName정해영-
dc.contributor.alternativeName이시석-
dc.contributor.alternativeName고준상-
dc.contributor.alternativeName고민수-
dc.contributor.alternativeName서휘원-
dc.identifier.bibliographicCitationGenes & Genomics, vol. 44, no. 8, pp. 899-912-
dc.identifier.doi10.1007/s13258-022-01264-7-
dc.subject.keywordSARS-CoV-2-
dc.subject.keywordMultiple sequence alignment-
dc.subject.keywordRT-qPCR-
dc.subject.localSARS-CoV-2-
dc.subject.localSARS-Cov-2-
dc.subject.localmultiple sequence alignment-
dc.subject.localMultiple sequence alignment-
dc.subject.localRT-qPCR-
dc.description.journalClassY-
Appears in Collections:
Division of Research on National Challenges > Infectious Disease Research Center > 1. Journal Articles
Files in This Item:
  • There are no files associated with this item.


Items in OpenAccess@KRIBB are protected by copyright, with all rights reserved, unless otherwise indicated.