Pesquisa | Portal Regional da BVS (teste)

Towards accurate indel calling for oncopanel sequencing through an international pipeline competition at precisionFDA.

Gong, Binsheng; Lababidi, Samir; Kusko, Rebecca; Bouri, Khaled; Prezek, Sarah; Thovarai, Vishal; Prasanna, Anish; Maier, Ezekiel J; Golkaram, Mahdi; Sun, Xingqiang; Kyriakidis, Konstantinos; Kitajima, João Paulo; Ebrahim Sahraeian, Sayed Mohammad; Guo, Yunfei; Johanson, Elaine; Jones, Wendell; Tong, Weida; Xu, Joshua.

Sci Rep ; 14(1): 8165, 2024 04 08.

Artigo em Inglês | MEDLINE | ID: mdl-38589653

RESUMO

Accurately calling indels with next-generation sequencing (NGS) data is critical for clinical application. The precisionFDA team collaborated with the U.S. Food and Drug Administration's (FDA's) National Center for Toxicological Research (NCTR) and successfully completed the NCTR Indel Calling from Oncopanel Sequencing Data Challenge, to evaluate the performance of indel calling pipelines. Top performers were selected based on precision, recall, and F1-score. The performance of many other pipelines was close to the top performers, which produced a top cluster of performers. The performance was significantly higher in high confidence regions and coding regions, and significantly lower in low complexity regions. Oncopanel capture and other issues may have occurred that affected the recall rate. Indels with higher variant allele frequency (VAF) may generally be called with higher confidence. Many of the indel calling pipelines had good performance. Some of them performed generally well across all three oncopanels, while others were better for a specific oncopanel. The performance of indel calling can further be improved by restricting the calls within high confidence intervals (HCIs) and coding regions, and by excluding low complexity regions (LCR) regions. Certain VAF cut-offs could be applied according to the applications.

Assuntos

Sequenciamento de Nucleotídeos em Larga Escala , Mutação INDEL , Polimorfismo de Nucleotídeo Único

Synthetic Health Data Can Augment Community Research Efforts to Better Inform the Public During Emerging Pandemics.

Prasanna, Anish; Jing, Bocheng; Plopper, George; Miller, Kristina Krasnov; Sanjak, Jaleal; Feng, Alice; Prezek, Sarah; Vidyaprakash, Eshaw; Thovarai, Vishal; Maier, Ezekiel J; Bhattacharya, Avik; Naaman, Lama; Stephens, Holly; Watford, Sean; Boscardin, W John; Johanson, Elaine; Lienau, Amanda.

medRxiv ; 2023 Dec 13.

Artigo em Inglês | MEDLINE | ID: mdl-38168217

RESUMO

The COVID-19 pandemic had disproportionate effects on the Veteran population due to the increased prevalence of medical and environmental risk factors. Synthetic electronic health record (EHR) data can help meet the acute need for Veteran population-specific predictive modeling efforts by avoiding the strict barriers to access, currently present within Veteran Health Administration (VHA) datasets. The U.S. Food and Drug Administration (FDA) and the VHA launched the precisionFDA COVID-19 Risk Factor Modeling Challenge to develop COVID-19 diagnostic and prognostic models; identify Veteran population-specific risk factors; and test the usefulness of synthetic data as a substitute for real data. The use of synthetic data boosted challenge participation by providing a dataset that was accessible to all competitors. Models trained on synthetic data showed similar but systematically inflated model performance metrics to those trained on real data. The important risk factors identified in the synthetic data largely overlapped with those identified from the real data, and both sets of risk factors were validated in the literature. Tradeoffs exist between synthetic data generation approaches based on whether a real EHR dataset is required as input. Synthetic data generated directly from real EHR input will more closely align with the characteristics of the relevant cohort. This work shows that synthetic EHR data will have practical value to the Veterans' health research community for the foreseeable future.

RESUMO

Assuntos

RESUMO

ENVIAR RESULTADO:

SELEÇÃO DE REFERÊNCIAS

DETALHE DA PESQUISA