Background <p>The standard regulatory approach to assess replication success is the two-trials rule, requiring both the original and the replication study to be significant with effect estimates in the same direction. The sceptical <i>p</i>-value was recently presented as an alternative method for the statistical assessment of the replicability of study results.</p> Methods <p>We review the statistical properties of the sceptical <i>p</i>-value and compare those to the two-trials rule. We extend the methodology to non-inferiority trials and describe how to invert the sceptical <i>p</i>-value to obtain confidence intervals. We illustrate the performance of the different methods using real-world evidence emulations of randomized controlled trials (RCTs) conducted within the RCT DUPLICATE initiative.</p> Results <p>The sceptical <i>p</i>-value depends not only on the two <i>p</i>-values, but also on sample size and effect size of the two studies. It can be calibrated to have the same Type-I error rate as the two-trials rule, but has larger power to detect an existing effect. In the application to the results from the RCT DUPLICATE initiative, the sceptical <i>p</i>-value leads to qualitatively similar results than the two-trials rule, but tends to show more evidence for treatment effects compared to the two-trials rule.</p> Conclusion <p>The sceptical <i>p</i>-value represents a valid statistical measure to assess the replicability of study results and is useful in the context of real-world evidence emulations.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Assessing the replicability of RCTs in RWE emulations

  • Jeanette Köppe,
  • Charlotte Micheloud,
  • Stella Erdmann,
  • Rachel Heyard,
  • Leonhard Held

摘要

Background

The standard regulatory approach to assess replication success is the two-trials rule, requiring both the original and the replication study to be significant with effect estimates in the same direction. The sceptical p-value was recently presented as an alternative method for the statistical assessment of the replicability of study results.

Methods

We review the statistical properties of the sceptical p-value and compare those to the two-trials rule. We extend the methodology to non-inferiority trials and describe how to invert the sceptical p-value to obtain confidence intervals. We illustrate the performance of the different methods using real-world evidence emulations of randomized controlled trials (RCTs) conducted within the RCT DUPLICATE initiative.

Results

The sceptical p-value depends not only on the two p-values, but also on sample size and effect size of the two studies. It can be calibrated to have the same Type-I error rate as the two-trials rule, but has larger power to detect an existing effect. In the application to the results from the RCT DUPLICATE initiative, the sceptical p-value leads to qualitatively similar results than the two-trials rule, but tends to show more evidence for treatment effects compared to the two-trials rule.

Conclusion

The sceptical p-value represents a valid statistical measure to assess the replicability of study results and is useful in the context of real-world evidence emulations.