端到端机器学习
Joshua Stapleton
Machine Learning Engineer


ks_2samp() 返回两个值:检验统计量、p 值。from scipy.stats import ks_2samp
# load the 1D data distribution samples for comparison
sample_1, sample_2 = training_dataset_sample, current_inference_sample
# perform the KS-test - ensure input samples are numpy arrays
test_statistic, p_value = ks_2samp(sample_1, sample_2)
if p_value < 0.05:
print("Reject null hypothesis - data drift might be occuring")
else:
print("Samples are likely to be from the same dataset")
根据新数据更新模型
新/推理数据不足?


端到端机器学习