Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 Detection of atypical data in multicenter clinical trials using unsupervised statistical monitoring.

Literature DB >> 31331195

Detection of atypical data in multicenter clinical trials using unsupervised statistical monitoring.

Laura Trotta¹, Yuusuke Kabeya^2,3, Marc Buyse^4,5, Erik Doffagne¹, David Venet⁶, Lieven Desmet⁷, Tomasz Burzykowski^8,9, Akira Tsuburaya¹⁰, Kazuhiro Yoshida¹¹, Yumi Miyashita¹², Satoshi Morita¹³, Junichi Sakamoto^12,14, Paurush Praveen¹, Koji Oba^2,15.

Abstract

BACKGROUND/AIMS: A risk-based approach to clinical research may include a central statistical assessment of data quality. We investigated the operating characteristics of unsupervised statistical monitoring aimed at detecting atypical data in multicenter experiments. The approach is premised on the assumption that, save for random fluctuations and natural variations, data coming from all centers should be comparable and statistically consistent. Unsupervised statistical monitoring consists of performing as many statistical tests as possible on all trial data, in order to detect centers whose data are inconsistent with data from other centers.
METHODS: We conducted simulations using data from a large multicenter trial conducted in Japan for patients with advanced gastric cancer. The actual trial data were contaminated in computer simulations for varying percentages of centers, percentages of patients modified within each center and numbers and types of modified variables. The unsupervised statistical monitoring software was run by a blinded team on the contaminated data sets, with the purpose of detecting the centers with contaminated data. The operating characteristics (sensitivity, specificity and Youden's J-index) were calculated for three detection methods: one using the p-values of individual statistical tests after adjustment for multiplicity, one using a summary of all p-values for a given center, called the Data Inconsistency Score, and one using both of these methods.
RESULTS: The operating characteristics of the three methods were satisfactory in situations of data contamination likely to occur in practice, specifically when a single or a few centers were contaminated. As expected, the sensitivity increased for increasing proportions of patients and increasing numbers of variables contaminated. The three methods showed a specificity better than 93% in all scenarios of contamination. The method based on the Data Inconsistency Score and individual p-values adjusted for multiplicity generally had slightly higher sensitivity at the expense of a slightly lower specificity.
CONCLUSIONS: The use of brute force (a computer-intensive approach that generates large numbers of statistical tests) is an effective way to check data quality in multicenter clinical trials. It can provide a cost-effective complement to other data-management and monitoring techniques.

Entities: Disease Species

Keywords: Data quality; central statistical monitoring; fraud detection; operating characteristics; risk-based monitoring; simulations

Year: 2019 PMID： 31331195 DOI： 10.1177/1740774519862564

Source DB: PubMed Journal: Clin Trials ISSN： 1740-7745 Impact factor: 2.486

Keyword Cloud
Cited

1 in total

Review 1. Central statistical monitoring of investigator-led clinical trials in oncology.

Authors: Marc Buyse; Laura Trotta; Everardo D Saad; Junichi Sakamoto
Journal: Int J Clin Oncol Date: 2020-06-23 Impact factor: 3.402

1 in total