Статья

VGsim: scalable viral genealogy simulator for global pandemic

S. Vladimir, S. Vadim, P. Victor, B. Evgeni, De Maio Nicola, C. Russell,
2021

As an effort to help contain the COVID-19 pandemic, large numbers of SARS-CoV-2 genomes have been sequenced from all continents. More than one million viral sequences are publicly available as of April 2021. Many studies estimate viral genealogies from these sequences, as these can provide valuable information about the spread of the pandemic across time and space. Additionally such data are a rich source of information about molecular evolutionary processes including natural selection, for example allowing the identification and investigating the spread of new variants conferring transmissibility and immunity evasion advantages to the virus. To validate new methods and to verify results resulting from these vast datasets, one needs an efficient simulator able to simulate the pandemic to approximate world-scale scenarios and generate viral genealogies of millions of samples. Here, we introduce a new fast simulator VGsim which addresses this problem. The simulation process is split into two phases. During the forward run the algorithm generates a chain of events reflecting the dynamics of the pandemic using an hierarchical version of the Gillespie algorithm. During the backward run a coalescent-like approach generates a tree genealogy of samples conditioning on the events chain generated during the forward run. Our software can model complex population structure, epistasis and immunity escape. The code is freely available at https://github.com/Genomics-HSE/VGsim.

Цитирование

Похожие публикации

Источник

Версии

  • 1. Version of Record от 2021-04-27

Метаданные

Об авторах
  • S. Vladimir
    National Research University – Higher School of Economics
  • S. Vadim
    National Research University – Higher School of Economics
  • P. Victor
    National Research University – Higher School of Economics
  • B. Evgeni
    National Research University – Higher School of Economics
  • De Maio Nicola
    European Bioinformatics Institute
  • C. Russell
    National Research University – Higher School of Economics
Предметная рубрика
  • COVID-19
Название журнала
  • medRxiv : the preprint server for health sciences
Тип документа
  • preprint
Источник
  • lens