<link rel="stylesheet" href="styles.f3b1fba60ec7970c.css">

Improved A/B testing acceleration methods for parametric hypothesis testing: T-test comparison with CUPED, CUPED++ and Bayesian Estimator

Вантажиться...
Ескіз

Дата

Автори

Назва журналу

Номер ISSN

Назва тому

Анотація

The study aimed to compare statistical analysis methods to improve the testing of alternatives. The study evaluated four main methods: the classic T-test, the conventional and advanced method of Controlled Experiments Using Pre-Experimental Data (CUPED), and the Bayesian Estimator. The main results included a demonstration of the A/B testing process, and the described statistical analysis methods included detailed characteristics and examples of use. The simulations and practical application revealed that the T-test provides high accuracy with small samples, but its effectiveness decreases with increasing sample size due to high resource requirements. The calculator for this method demonstrated effectiveness in simple tasks but had limitations with large data. The conventional CUPED method has shown increased accuracy due to variation correction, but its effectiveness decreases when working with large and complex data sets. The written program for this method has shown to be effective in cases the previous data is well represented, but its capabilities are limited when processing large data sets. The improved version provided a significant improvement in both accuracy and processing speed, especially for large datasets, thanks to advanced modelling and optimisation. The code results confirmed that this method is highly efficient for complex experiments, particularly when processing large amounts of data. Moreover, the Bayesian Estimator demonstrated high accuracy due to the integration of prior knowledge but required more computational resources and time. The platform used for this method demonstrated the ability to account for uncertainty yet required complex model settings. The results highlighted the importance of ing the appropriate statistical analysis method depending on the scale and complexity of the data to ensure optimal accuracy and efficiency of testing.

Опис

Мова

Бібліографічний опис

Markov A. Improved A/B testing acceleration methods for parametric hypothesis testing: T-test comparison with CUPED, CUPED++ and Bayesian Estimator // Information Technologies and Computer Engineering. 2024. № 3 (21). С. 119-131. URI: https://itce.vn.ua/uk/journals/t-21-3-2024/pokrashcheni-metodi-prishvidshennya-a-b-testuvannya-dlya-otsinki-parametrichnikh-gipotez-porivnyannya-t-test-z-cuped-cuped-ta-bayesian-estimator.

Схвалення

Рецензія

Доповнено

Цитується в

Список використаної літератури (30)

  1. Allard, C., & Marchand, É. (2024). Bayesian and Minimax estimators of loss. Japanese Journal of Statistics and Data Science. doi: 10.1007/s42081-024-00261-2.
  2. Baik, S.M., Byon, E., & Ko, Y.M. (2023). Distributionally robust stratified sampling for stochastic simulations with multiple uncertain input models. ArXiv. doi: 10.48550/arXiv.2306.09020.
  3. Bertolino, F., Manca, M., Musio, M., Racugno, W., & Ventura, L. (2024). A new Bayesian discrepancy measure. Journal of the Italian Statistical Society, 33, 381-405. doi: 10.1007/s10260-024-00745-1.
  4. Chauvet, L.A., & Cruz, D.M. (2024). Computational modeling of decision-making in substance abusers: Testing Bechara’s hypotheses. Frontiers in Psychology, 15, article number 1281082. doi: 10.3389/fpsyg.2024.1281082.
  5. Cho, Y.W., Chow, S.-M., Marini, C.M., & Martire, L.M. (2024). Multilevel latent differential structural equation model with short time series and time-varying covariates: A comparison of frequentist and Bayesian estimators. Multivariate Behavioral Research, 59(5), 934-956. doi: 10.1080/00273171.2024.2347959.
  6. Cissé, A., Evangelopoulos, X., Carruthers, S., Gusev, V.V., & Cooper, A.I. (2024). HypBO: Accelerating black-box scientific experiments using experts’ hypotheses. In K. Larson (Ed.), Proceedings of the thirty-third international joint conference on artificial intelligence (pp. 3881-3889). Vienna: IJCAI. doi: 10.24963/ijcai.2024/429.
  7. Deiri, E. (2021). Expected Bayesian estimator and hierarchical Bayesian estimator for the parameter of a Rayleigh distribution reliability system under the progressive type-II data sample. Mathematical Researches, 7(3), 527-544. doi: 10.52547/mmr.7.3.527.
  8. Deng, A., Yuan, L.-H., & Salama-Manteau, A. (2021). Variance reduction for experiments with one-sided triggering using CUPED. ArXiv. doi: 10.48550/arXiv.2112.13299.
  9. Dogan, O., Taspinar, S., & Bera, A.K. (2020). A Bayesian robust chi-squared test for testing simple hypotheses. Journal of Econometrics, 222(2), 933-958. doi: 10.1016/j.jeconom.2020.07.046.
  10. Duthie, B. (2024). Fundamental statistical concepts and techniques in the biological and environmental sciences. New York: Chapman and Hall. doi: 10.1201/9781032692388.