与数据集移位探测和模式选择的批量正常化统计 (Unsupervised Model Drift Estimation with Batch Normalization Statistics for Dataset Shift Detection and Model Selection)

While many real-world data streams imply that they change frequently in a nonstationary way, most of deep learning methods optimize neural networks on training data, and this leads to severe performance degradation when dataset shift happens. However, it is less possible to annotate or inspect newly streamed data by humans, and thus it is desired to measure model drift at inference time in an unsupervised manner. In this paper, we propose a novel method of model drift estimation by exploiting statistics of batch normalization layer on unlabeled test data. To remedy possible sampling error of streamed input data, we adopt low-rank approximation to each representational layer. We show the effectiveness of our method not only on dataset shift detection but also on model selection when there are multiple candidate models among model zoo or training trajectories in an unsupervised way. We further demonstrate the consistency of our method by comparing model drift scores between different network architectures.

翻译：虽然许多真实世界的数据流意味着它们经常以非静止的方式变化,但大多数深层次的学习方法优化了培训数据神经网络,这导致发生数据集转移时出现严重的性能退化。然而,人类对新流数据进行笔记或检查的可能性较小,因此,人们希望以不受监督的方式测量在推论时间的模型漂移。在本文中,我们提出一种新的模型漂移估计方法,利用未贴标签的测试数据上的批量正常化层统计数据。为了补救流出输入数据可能的抽样错误,我们采用了对每个代表层的低位近似值。我们不仅在数据集移位探测上展示了我们的方法的有效性,而且在模型动物园或培训轨迹中存在多种候选模型时也展示了模型选择的有效性。我们进一步通过比较不同网络结构之间的模型漂移分数来展示我们方法的一致性。

相关内容

MoDELS

关注 43

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/