内嵌系统从汇编到部署的 ML执行环境 (TinyIREE: An ML Execution Environment for Embedded Systems from Compilation to Deployment) - 专知论文

会员服务 ·

0

编译器 · Extensibility · 回合 · Machine Learning · 缩放 ·

2022 年 5 月 28 日

TinyIREE: An ML Execution Environment for Embedded Systems from Compilation to Deployment

翻译：内嵌系统从汇编到部署的 ML执行环境

Hsin-I Cindy Liu,Marius Brehler,Mahesh Ravishankar,Nicolas Vasilache,Ben Vanik,Stella Laurenzo

from arxiv, 9 pages, 3 figures, to be published in IEEE Micro

Machine learning model deployment for training and execution has been an important topic for industry and academic research in the last decade. Much of the attention has been focused on developing specific toolchains to support acceleration hardware. In this paper, we present IREE, a unified compiler and runtime stack with the explicit goal to scale down machine learning programs to the smallest footprints for mobile and edge devices, while maintaining the ability to scale up to larger deployment targets. IREE adopts a compiler-based approach and optimizes for heterogeneous hardware accelerators through the use of the MLIR compiler infrastructure which provides the means to quickly design and implement multi-level compiler intermediate representations (IR). More specifically, this paper is focused on TinyIREE, which is a set of deployment options in IREE that accommodate the limited memory and computation resources in embedded systems and bare-metal platforms, while also demonstrating IREE's intuitive workflow that generates workloads for different ISA extensions and ABIs through LLVM.

翻译：过去十年来,为培训和执行而部署机器学习模型一直是工业和学术研究的一个重要议题,许多注意力都集中在开发支持加速硬件的具体工具链上。本文介绍IREE,这是一个统一的编译器和运行时间堆,其明确目标是将机器学习程序缩小到移动和边缘装置的最小脚印,同时保持将规模扩大到更大部署目标的能力。IREE采用基于编译器的方法,并通过使用MLIR编译器基础设施优化不同硬件加速器。MLIR编译器基础设施提供了快速设计和实施多级别编译器中间演示(IR)的手段。更具体地说,本文侧重于TinyIRREE,这是IREE的一套部署选项,它满足嵌入系统和光金属平台有限的记忆和计算资源,同时展示IREE的直觉工作流程,通过LLVM为不同的ISA扩展和ABI产生工作量。

0

相关内容

编译器

编译器（Compiler），是一种计算机程序，它会将用某种编程语言写成的源代码（原始语言），转换成另一种编程语言（目标语言）。

神经常微分方程教程，50页ppt，A brief tutorial on Neural ODEs

神经常微分方程教程，50页ppt，A brief tutorial on Neural ODEs

专知会员服务

74+阅读 · 2020年8月2日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Call for Nominations: 2022 Multimedia Prize Paper Award

Call for Nominations: 2022 Multimedia Prize Paper Award

CCF多媒体专委会

0+阅读 · 2022年2月12日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

补肾活血方对大鼠骨质疏松症模型Hedgehog信号通路调控及骨髓间充质干细胞成骨分化过程的实验研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于深度视觉注意机制的可逆水印及评价方法

国家自然科学基金

0+阅读 · 2013年12月31日

mirPS细胞与内皮祖细胞移植共同促进脑缺血损伤修复

国家自然科学基金

0+阅读 · 2013年12月31日

潘多拉菌中氯苯代谢的两个基因簇的转录调控研究

国家自然科学基金

0+阅读 · 2013年12月31日

金属衬底上氧化物薄膜外延生长机理研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于定理证明的多核并行程序验证

国家自然科学基金

0+阅读 · 2012年12月31日

lncRNA-UCA1通过PKM2参与膀胱癌细胞Warburg效应的机制

国家自然科学基金

0+阅读 · 2012年12月31日

原始态(naive)和始发态(primed)水牛诱导多能干细胞的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于ORP的锌净化过程控制参数化优化方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

DWI监测RNAi沉默AQP4治疗脑缺血半暗带的实验研究

国家自然科学基金

0+阅读 · 2009年12月31日

Detection of Poisoning Attacks with Anomaly Detection in Federated Learning for Healthcare Applications: A Machine Learning Approach

Arxiv

0+阅读 · 2022年7月18日

Fine-grained Data Access Control for Collaborative Process Execution on Blockchain

Arxiv

0+阅读 · 2022年7月18日

LambdaLite: Application-Level Optimization for Cold Start Latency in Serverless Computing

Arxiv

0+阅读 · 2022年7月17日

Towards Observability for Production Machine Learning Pipelines

Arxiv

0+阅读 · 2022年7月15日

POET: Training Neural Networks on Tiny Devices with Integrated Rematerialization and Paging

Arxiv

0+阅读 · 2022年7月15日

Computing Execution Times with eXecution Decision Diagrams in the Presence of Out-Of-Order Resources

Arxiv

0+阅读 · 2022年7月15日

Modeling and Executing Production Processes with Capabilities and Skills using Ontologies and BPMN

Arxiv

0+阅读 · 2022年7月15日

Multi: a Formal Playground for Multi-Smart Contract Interaction

Arxiv

0+阅读 · 2022年7月14日

Enable Deep Learning on Mobile Devices: Methods, Systems, and Applications

Arxiv

35+阅读 · 2022年4月25日

Distributed Machine Learning on Mobile Devices: A Survey

Distributed Machine Learning on Mobile Devices: A Survey

Arxiv

37+阅读 · 2019年9月18日

VIP会员

文章信息

相关主题

Machine Learning

相关VIP内容

神经常微分方程教程，50页ppt，A brief tutorial on Neural ODEs

神经常微分方程教程，50页ppt，A brief tutorial on Neural ODEs

专知会员服务

74+阅读 · 2020年8月2日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

【CMU卡内基梅隆大学】深度学习在计算机视觉的应用：方法，解释，因果与公平性

专知会员服务

83+阅读 · 2019年10月9日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《美陆军徒步机动作战条令手册》最新168页

【博士论文】基于不确定性的可靠性：现代机器学习中的选择性预测与可信部署

军事后勤数字化未来展望

《美海军后勤体系整合与创新挑战》最新报告

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Call for Nominations: 2022 Multimedia Prize Paper Award

Call for Nominations: 2022 Multimedia Prize Paper Award

CCF多媒体专委会

0+阅读 · 2022年2月12日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【推荐】SVM实例教程

【推荐】SVM实例教程

机器学习研究会

17+阅读 · 2017年8月26日

相关论文

Detection of Poisoning Attacks with Anomaly Detection in Federated Learning for Healthcare Applications: A Machine Learning Approach

Arxiv

0+阅读 · 2022年7月18日

Fine-grained Data Access Control for Collaborative Process Execution on Blockchain

Arxiv

0+阅读 · 2022年7月18日

LambdaLite: Application-Level Optimization for Cold Start Latency in Serverless Computing

Arxiv

0+阅读 · 2022年7月17日

Towards Observability for Production Machine Learning Pipelines

Arxiv

0+阅读 · 2022年7月15日

POET: Training Neural Networks on Tiny Devices with Integrated Rematerialization and Paging

Arxiv

0+阅读 · 2022年7月15日

Computing Execution Times with eXecution Decision Diagrams in the Presence of Out-Of-Order Resources

Arxiv

0+阅读 · 2022年7月15日

Modeling and Executing Production Processes with Capabilities and Skills using Ontologies and BPMN

Arxiv

0+阅读 · 2022年7月15日

Multi: a Formal Playground for Multi-Smart Contract Interaction

Arxiv

0+阅读 · 2022年7月14日

Enable Deep Learning on Mobile Devices: Methods, Systems, and Applications

Arxiv

35+阅读 · 2022年4月25日

Distributed Machine Learning on Mobile Devices: A Survey

Distributed Machine Learning on Mobile Devices: A Survey

Arxiv

37+阅读 · 2019年9月18日

相关基金

补肾活血方对大鼠骨质疏松症模型Hedgehog信号通路调控及骨髓间充质干细胞成骨分化过程的实验研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于深度视觉注意机制的可逆水印及评价方法

国家自然科学基金

0+阅读 · 2013年12月31日

mirPS细胞与内皮祖细胞移植共同促进脑缺血损伤修复

国家自然科学基金

0+阅读 · 2013年12月31日

潘多拉菌中氯苯代谢的两个基因簇的转录调控研究

国家自然科学基金

0+阅读 · 2013年12月31日

金属衬底上氧化物薄膜外延生长机理研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于定理证明的多核并行程序验证

国家自然科学基金

0+阅读 · 2012年12月31日

lncRNA-UCA1通过PKM2参与膀胱癌细胞Warburg效应的机制

国家自然科学基金

0+阅读 · 2012年12月31日

原始态(naive)和始发态(primed)水牛诱导多能干细胞的研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于ORP的锌净化过程控制参数化优化方法研究

国家自然科学基金

0+阅读 · 2011年12月31日

DWI监测RNAi沉默AQP4治疗脑缺血半暗带的实验研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员