StepMix: 一个用于广义混合模型的伪似然估计的Python软件包，支持外部变量 (StepMix: A Python Package for Pseudo-Likelihood Estimation of Generalized Mixture Models with External Variables) - 专知论文

会员服务 ·

0

伪似然 · 似然 · 潜在 · 类别 · 混合模型 ·

2023 年 4 月 11 日

StepMix: A Python Package for Pseudo-Likelihood Estimation of Generalized Mixture Models with External Variables

翻译：StepMix: 一个用于广义混合模型的伪似然估计的Python软件包，支持外部变量

Sacha Morin,Robin Legault,Zsuzsa Bakk,Charles-Édouard Giguère,Roxane de la Sablonnière,Éric Lacourse

from arxiv, Sacha Morin and Robin Legault contributed equally

StepMix is an open-source software package for the pseudo-likelihood estimation (one-, two- and three-step approaches) of generalized finite mixture models (latent profile and latent class analysis) with external variables (covariates and distal outcomes). In many applications in social sciences, the main objective is not only to cluster individuals into latent classes, but also to use these classes to develop more complex statistical models. These models generally divide into a measurement model that relates the latent classes to observed indicators, and a structural model that relates covariates and outcome variables to the latent classes. The measurement and structural models can be estimated jointly using the so-called one-step approach or sequentially using stepwise methods, which present significant advantages for practitioners regarding the interpretability of the estimated latent classes. In addition to the one-step approach, StepMix implements the most important stepwise estimation methods from the literature, including the bias-adjusted three-step methods with BCH and ML corrections and the more recent two-step approach. These pseudo-likelihood estimators are presented in this paper under a unified framework as specific expectation-maximization subroutines. To facilitate and promote their adoption among the data science community, StepMix follows the object-oriented design of the scikit-learn library and provides interfaces in both Python and R.

翻译：StepMix是一个开源软件包，用于估计带有外部变量（协变量和结果）的广义混合模型（潜在类别分析和潜在剖面分析）的伪似然估计方法（单步、两步和三步方法）。在社会科学的许多应用中，主要目标不仅是将个体聚类成为潜在的类别，而且使用这些类别来发展更复杂的统计模型。这些模型通常分为测量模型和结构模型两部分，前者将潜在类别与观察指标相关，后者将协变量和结果变量与潜在类别相关。测量模型和结构模型可以同时估计，也可以使用分步方法逐步估计。分步方法对于从业人员来说具有重要的优势，可以轻松解释估计的潜在类别。除了一步方法，StepMix还实现了文献中最重要的逐步估计方法，包括带有BCH和ML校正的偏差调整的三步方法和较新的两步方法。这些伪似然估计器在本文中以统一框架在特定的期望最大化子程序下进行介绍。为了方便和推广在数据科学社区的采用，StepMix遵循scikit-learn库的面向对象设计，并提供Python和R两种接口。

0

相关内容

伪似然

【2023新书】随机模型基础，815页pdf

【2023新书】随机模型基础，815页pdf

专知会员服务

105+阅读 · 2023年5月10日

【2023新书】使用Python进行统计和数据可视化，554页pdf

【2023新书】使用Python进行统计和数据可视化，554页pdf

专知会员服务

130+阅读 · 2023年1月29日

【干货书】工程和科学中的概率和统计，

【干货书】工程和科学中的概率和统计，

专知会员服务

58+阅读 · 2022年12月24日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

专知会员服务

37+阅读 · 2020年2月27日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

专知会员服务

246+阅读 · 2019年10月21日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【干货书】基于统计和机器学习的实用时间序列分析预测，Time Series Analysis Prediction

【干货书】基于统计和机器学习的实用时间序列分析预测，Time Series Analysis Prediction

专知

18+阅读 · 2022年4月9日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

利用动态深度学习预测金融时间序列基于Python

利用动态深度学习预测金融时间序列基于Python

量化投资与机器学习

18+阅读 · 2018年10月30日

【论文推荐】最新5篇图像分割（Image Segmentation）相关论文—多重假设、超像素分割、自监督、图、生成对抗网络

【论文推荐】最新5篇图像分割（Image Segmentation）相关论文—多重假设、超像素分割、自监督、图、生成对抗网络

专知

27+阅读 · 2018年2月7日

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

专知

23+阅读 · 2018年1月18日

【推荐】用Python/OpenCV实现增强现实

【推荐】用Python/OpenCV实现增强现实

机器学习研究会

15+阅读 · 2017年11月16日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

分别基于SVM和ARIMA模型的股票预测 Python实现附Github源码

分别基于SVM和ARIMA模型的股票预测 Python实现附Github源码

数据挖掘入门与实战

15+阅读 · 2017年9月9日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

Underlay频谱共享方式下信号参数估计和调制识别的方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

纵向数据的动态半参数建模及其统计推断

国家自然科学基金

0+阅读 · 2014年12月31日

复杂数据下含指标项半参数模型结构的统计推断及应用

国家自然科学基金

0+阅读 · 2014年12月31日

基于似然函数的统计推断

国家自然科学基金

5+阅读 · 2014年12月31日

非参数动态混合Copula模型：估计、推断及应用

国家自然科学基金

0+阅读 · 2013年12月31日

生物医学研究中不完全分类数据的统计推断

国家自然科学基金

0+阅读 · 2012年12月31日

区间删失数据的半参数回归模型的有效估计方法

国家自然科学基金

0+阅读 · 2012年12月31日

基于纵向数据的秩回归和分位数回归的有效参数估计

国家自然科学基金

0+阅读 · 2012年12月31日

流行病学中若干统计分析模型的推断

国家自然科学基金

2+阅读 · 2012年12月31日

缺失数据下部分线性单指标模型的经验似然推断

国家自然科学基金

0+阅读 · 2009年12月31日

Leveraging Evolutionary Changes for Software Process Quality

Arxiv

0+阅读 · 2023年5月29日

Disentangling Light Fields for Super-Resolution and Disparity Estimation

Arxiv

0+阅读 · 2023年5月29日

A Bayesian Approach for Clustering Constant-wise Change-point Data

Arxiv

0+阅读 · 2023年5月28日

Mixed-integer linear programming for computing optimal experimental designs

Arxiv

0+阅读 · 2023年5月27日

bqror: An R package for Bayesian Quantile Regression in Ordinal Models

Arxiv

0+阅读 · 2023年5月27日

On Calibrating Diffusion Probabilistic Models

Arxiv

0+阅读 · 2023年5月26日

Detecting and diagnosing prior and likelihood sensitivity with power-scaling

Arxiv

0+阅读 · 2023年5月26日

Negative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models

Arxiv

0+阅读 · 2023年5月26日

Bayesian Inversion for Nonlinear Imaging Models using Deep Generative Priors

Arxiv

0+阅读 · 2023年5月25日

Generative Adversarial Networks and Probabilistic Graph Models for Hyperspectral Image Classification

Arxiv

11+阅读 · 2018年2月10日

VIP会员

文章信息

相关主题

相关VIP内容

【2023新书】随机模型基础，815页pdf

【2023新书】随机模型基础，815页pdf

专知会员服务

105+阅读 · 2023年5月10日

【2023新书】使用Python进行统计和数据可视化，554页pdf

【2023新书】使用Python进行统计和数据可视化，554页pdf

专知会员服务

130+阅读 · 2023年1月29日

【干货书】工程和科学中的概率和统计，

【干货书】工程和科学中的概率和统计，

专知会员服务

58+阅读 · 2022年12月24日

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

INRIA最新「机器学习理论」新书，229页pdf原理性阐述机器学习

专知会员服务

69+阅读 · 2021年3月27日

INRIA 最新《机器学习理论》课程笔记，176页pdf

专知会员服务

51+阅读 · 2020年12月14日

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

【DeepMind】PolyGen: 一种三维网格的自回归生成模型，PolyGen: An Autoregressive Generative Model of 3D Meshes

专知会员服务

37+阅读 · 2020年2月27日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

【机器学习基础最新版】（Mathematics for Machine Learning），417页pdf

专知会员服务

246+阅读 · 2019年10月21日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

热门VIP内容

开通专知VIP会员享更多权益服务

《革新战术战场空间能力：反无人机系统》报告

《在单一作战合成环境（SSE）中运用人工智能与大型语言模型以提供灵活人文地形及可信角色组》报告

螺旋式开发作为战略资产：美军启示

《提示战争：大语言模型如何决定军事干预》报告

相关资讯

【干货书】基于统计和机器学习的实用时间序列分析预测，Time Series Analysis Prediction

【干货书】基于统计和机器学习的实用时间序列分析预测，Time Series Analysis Prediction

专知

18+阅读 · 2022年4月9日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

利用动态深度学习预测金融时间序列基于Python

利用动态深度学习预测金融时间序列基于Python

量化投资与机器学习

18+阅读 · 2018年10月30日

【论文推荐】最新5篇图像分割（Image Segmentation）相关论文—多重假设、超像素分割、自监督、图、生成对抗网络

【论文推荐】最新5篇图像分割（Image Segmentation）相关论文—多重假设、超像素分割、自监督、图、生成对抗网络

专知

27+阅读 · 2018年2月7日

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

最新5篇生成对抗网络相关论文推荐—FusedGAN、DeblurGAN、AdvGAN、CipherGAN、MMD GANS

专知

23+阅读 · 2018年1月18日

【推荐】用Python/OpenCV实现增强现实

【推荐】用Python/OpenCV实现增强现实

机器学习研究会

15+阅读 · 2017年11月16日

【论文】变分推断（Variational inference)的总结

【论文】变分推断（Variational inference)的总结

机器学习研究会

39+阅读 · 2017年11月16日

分别基于SVM和ARIMA模型的股票预测 Python实现附Github源码

分别基于SVM和ARIMA模型的股票预测 Python实现附Github源码

数据挖掘入门与实战

15+阅读 · 2017年9月9日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

相关论文

Leveraging Evolutionary Changes for Software Process Quality

Arxiv

0+阅读 · 2023年5月29日

Disentangling Light Fields for Super-Resolution and Disparity Estimation

Arxiv

0+阅读 · 2023年5月29日

A Bayesian Approach for Clustering Constant-wise Change-point Data

Arxiv

0+阅读 · 2023年5月28日

Mixed-integer linear programming for computing optimal experimental designs

Arxiv

0+阅读 · 2023年5月27日

bqror: An R package for Bayesian Quantile Regression in Ordinal Models

Arxiv

0+阅读 · 2023年5月27日

On Calibrating Diffusion Probabilistic Models

Arxiv

0+阅读 · 2023年5月26日

Detecting and diagnosing prior and likelihood sensitivity with power-scaling

Arxiv

0+阅读 · 2023年5月26日

Negative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models

Arxiv

0+阅读 · 2023年5月26日

Bayesian Inversion for Nonlinear Imaging Models using Deep Generative Priors

Arxiv

0+阅读 · 2023年5月25日

Generative Adversarial Networks and Probabilistic Graph Models for Hyperspectral Image Classification

Arxiv

11+阅读 · 2018年2月10日

相关基金

Underlay频谱共享方式下信号参数估计和调制识别的方法研究

国家自然科学基金

0+阅读 · 2015年12月31日

纵向数据的动态半参数建模及其统计推断

国家自然科学基金

0+阅读 · 2014年12月31日

复杂数据下含指标项半参数模型结构的统计推断及应用

国家自然科学基金

0+阅读 · 2014年12月31日

基于似然函数的统计推断

国家自然科学基金

5+阅读 · 2014年12月31日

非参数动态混合Copula模型：估计、推断及应用

国家自然科学基金

0+阅读 · 2013年12月31日

生物医学研究中不完全分类数据的统计推断

国家自然科学基金

0+阅读 · 2012年12月31日

区间删失数据的半参数回归模型的有效估计方法

国家自然科学基金

0+阅读 · 2012年12月31日

基于纵向数据的秩回归和分位数回归的有效参数估计

国家自然科学基金

0+阅读 · 2012年12月31日

流行病学中若干统计分析模型的推断

国家自然科学基金

2+阅读 · 2012年12月31日

缺失数据下部分线性单指标模型的经验似然推断

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员