完成水质数据的多目标工具比特模型 (Multi-Target Tobit Models for Completing Water Quality Data)

Monitoring microbiological behaviors in water is crucial to manage public health risk from waterborne pathogens, although quantifying the concentrations of microbiological organisms in water is still challenging because concentrations of many pathogens in water samples may often be below the quantification limit, producing censoring data. To enable statistical analysis based on quantitative values, the true values of non-detected measurements are required to be estimated with high precision. Tobit model is a well-known linear regression model for analyzing censored data. One drawback of the Tobit model is that only the target variable is allowed to be censored. In this study, we devised a novel extension of the classical Tobit model, called the \emph{multi-target Tobit model}, to handle multiple censored variables simultaneously by introducing multiple target variables. For fitting the new model, a numerical stable optimization algorithm was developed based on elaborate theories. Experiments conducted using several real-world water quality datasets provided an evidence that estimating multiple columns jointly gains a great advantage over estimating them separately.

翻译：监测水中的微生物行为对于管理水媒病原体的公共卫生风险至关重要,尽管量化水中微生物生物浓度仍然具有挑战性,因为水样中许多病原体的浓度往往低于量化限度,从而产生审查数据。为了能够根据定量值进行统计分析,需要以高精确度对非检测测量的真实值进行估算。托比特模型是分析受审查数据的一个众所周知的线性回归模型。托比特模型的一个缺点是,只允许对目标变量进行检查。在本研究中,我们设计了经典托比特模型的新扩展,称为\emph{多目标托比特模型},以便同时通过引入多个目标变量来处理多个经过审查的变量。为适应新模型,根据精心制定的理论,开发了数字稳定优化算法。使用几个真实世界水质数据集进行的实验提供了证据,估计多个柱子在分别估算这些变量方面共同获得极大优势。

相关内容

MoDELS

关注 44

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日