不同数据处理图书馆的能源消耗 -- -- 探索性研究 (On the Energy Consumption of Different Dataframe Processing Libraries -- An Exploratory Study) - 专知论文

会员服务 ·

0

Learning · Analysis · Machine Learning · Processing（编程语言） · Attention ·

2022 年 9 月 12 日

On the Energy Consumption of Different Dataframe Processing Libraries -- An Exploratory Study

翻译：不同数据处理图书馆的能源消耗 -- -- 探索性研究

Shriram Shanbhag,Sridhar Chimalakonda

Background: The energy consumption of machine learning and its impact on the environment has made energy efficient ML an emerging area of research. However, most of the attention stays focused on the model creation and the training and inferencing phase. Data oriented stages like preprocessing, cleaning and exploratory analysis form a critical part of the machine learning workflow. However, the energy efficiency of these stages have gained little attention from the researchers. Aim: Our study aims to explore the energy consumption of different dataframe processing libraries as a first step towards studying the energy efficiency of the data oriented stages of the machine learning pipeline. Method: We measure the energy consumption of 3 popular libraries used to work with dataframes, namely Pandas, Vaex and Dask for 21 different operations grouped under 4 categories on 2 datasets. Results: The results of our analysis show that for a given dataframe processing operation, the choice of library can indeed influence the energy consumption with some libraries consuming 202 times lesser energy over others. Conclusion: The results of our study indicates that there is a potential for optimizing the energy consumption of the data oriented stages of the machine learning pipeline and further research is needed in the direction.

翻译：目标:我们的研究旨在探索不同数据框架处理图书馆的能源消耗情况,作为研究机器学习管道数据导向阶段的能源效率的第一步。方法:我们衡量用于数据框架的3个流行图书馆的能源消耗情况,即Pandas、Vaex和Dask,用于按2个数据集分为4类的21个不同操作,结果:我们的分析结果显示,对于某一数据框架处理作业,图书馆的选择确实能够影响能源消耗,而某些图书馆消耗的能源比其他图书馆少202倍。结论:我们的研究结果表明,有可能优化机器学习管道数据导向阶段的能源消耗,并需要在这方面进行进一步研究。

0

相关内容

Learning

自然语言处理顶会NAACL2022最佳论文出炉！

自然语言处理顶会NAACL2022最佳论文出炉！

专知会员服务

43+阅读 · 2022年6月30日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

专知会员服务

108+阅读 · 2020年5月1日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

【CIKM2019 Tutorial】Recent Developments of Deep Heterogeneous Information Network Analysis（深度异构信息网络分析的最新进展），附157页PDF免费下载

【CIKM2019 Tutorial】Recent Developments of Deep Heterogeneous Information Network Analysis（深度异构信息网络分析的最新进展），附157页PDF免费下载

专知会员服务

29+阅读 · 2019年11月3日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

罗巴代数的表示和罗巴代数在operad中的应用

国家自然科学基金

0+阅读 · 2015年12月31日

GAPDS在糖尿病引起的不育症中的作用及分子机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

淡水鱼贮藏过程鱼肉体系中肌苷酸变化规律及调控机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

Ghrelin对牛卵母细胞体外成熟的调控机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

超低介电PMO薄膜的可控制备及构效关系研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于蛋白质组学和代谢组学整合分析的Paraconiothyrium variable GHJ-4降解木质素的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

氧化物半导体薄膜晶体管的模型及参数提取方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

函数域上的超曲面中有理空间的密度

国家自然科学基金

0+阅读 · 2011年12月31日

苦荞麦抗肿瘤蛋白的结构及构效关系研究

国家自然科学基金

0+阅读 · 2009年12月31日

Pd/TiAl界面结合强度和热稳定性的基础研究

国家自然科学基金

0+阅读 · 2008年12月31日

A Survey of Machine Unlearning

Arxiv

0+阅读 · 2022年10月21日

A Stability Analysis of Modified Patankar-Runge-Kutta methods for a nonlinear Production-Destruction System

Arxiv

0+阅读 · 2022年10月21日

Free energy model of emotional valence in dual-process perceptions

Arxiv

0+阅读 · 2022年10月21日

Comparison of REML methods for the study of phenome-wide genetic variation

Arxiv

0+阅读 · 2022年10月21日

Video Summarization Overview

Arxiv

7+阅读 · 2022年10月21日

On Feature Learning in the Presence of Spurious Correlations

Arxiv

1+阅读 · 2022年10月20日

Reproducibility of the Methods in Medical Imaging with Deep Learning

Arxiv

0+阅读 · 2022年10月20日

Analyzing the Robustness of Decentralized Horizontal and Vertical Federated Learning Architectures in a Non-IID Scenario

Arxiv

0+阅读 · 2022年10月20日

Label Noise in Adversarial Training: A Novel Perspective to Study Robust Overfitting

Arxiv

0+阅读 · 2022年10月19日

Curriculum Learning: A Survey

Arxiv

24+阅读 · 2021年1月25日

VIP会员

文章信息

相关主题

Machine Learning

Processing（编程语言）

相关VIP内容

自然语言处理顶会NAACL2022最佳论文出炉！

自然语言处理顶会NAACL2022最佳论文出炉！

专知会员服务

43+阅读 · 2022年6月30日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

专知会员服务

108+阅读 · 2020年5月1日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

【CIKM2019 Tutorial】Recent Developments of Deep Heterogeneous Information Network Analysis（深度异构信息网络分析的最新进展），附157页PDF免费下载

【CIKM2019 Tutorial】Recent Developments of Deep Heterogeneous Information Network Analysis（深度异构信息网络分析的最新进展），附157页PDF免费下载

专知会员服务

29+阅读 · 2019年11月3日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

《在单一作战合成环境（SSE）中运用人工智能与大型语言模型以提供灵活人文地形及可信角色组》报告

《俄罗斯的未来战争方式第二部分：核威慑》报告

《提示战争：大语言模型如何决定军事干预》报告

《俄罗斯的未来战争方式第三部分：军事改革》报告

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

A Survey of Machine Unlearning

Arxiv

0+阅读 · 2022年10月21日

A Stability Analysis of Modified Patankar-Runge-Kutta methods for a nonlinear Production-Destruction System

Arxiv

0+阅读 · 2022年10月21日

Free energy model of emotional valence in dual-process perceptions

Arxiv

0+阅读 · 2022年10月21日

Comparison of REML methods for the study of phenome-wide genetic variation

Arxiv

0+阅读 · 2022年10月21日

Video Summarization Overview

Arxiv

7+阅读 · 2022年10月21日

On Feature Learning in the Presence of Spurious Correlations

Arxiv

1+阅读 · 2022年10月20日

Reproducibility of the Methods in Medical Imaging with Deep Learning

Arxiv

0+阅读 · 2022年10月20日

Analyzing the Robustness of Decentralized Horizontal and Vertical Federated Learning Architectures in a Non-IID Scenario

Arxiv

0+阅读 · 2022年10月20日

Label Noise in Adversarial Training: A Novel Perspective to Study Robust Overfitting

Arxiv

0+阅读 · 2022年10月19日

Curriculum Learning: A Survey

Arxiv

24+阅读 · 2021年1月25日

相关基金

罗巴代数的表示和罗巴代数在operad中的应用

国家自然科学基金

0+阅读 · 2015年12月31日

GAPDS在糖尿病引起的不育症中的作用及分子机制研究

国家自然科学基金

0+阅读 · 2015年12月31日

淡水鱼贮藏过程鱼肉体系中肌苷酸变化规律及调控机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

Ghrelin对牛卵母细胞体外成熟的调控机制研究

国家自然科学基金

0+阅读 · 2013年12月31日

超低介电PMO薄膜的可控制备及构效关系研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于蛋白质组学和代谢组学整合分析的Paraconiothyrium variable GHJ-4降解木质素的分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

氧化物半导体薄膜晶体管的模型及参数提取方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

函数域上的超曲面中有理空间的密度

国家自然科学基金

0+阅读 · 2011年12月31日

苦荞麦抗肿瘤蛋白的结构及构效关系研究

国家自然科学基金

0+阅读 · 2009年12月31日

Pd/TiAl界面结合强度和热稳定性的基础研究

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员