前向或反向模式自动区分:区别是什么? (Forward- or Reverse-Mode Automatic Differentiation: What's the Difference?) - 专知论文

会员服务 ·

0

Haskell · 可理解性 · 原点 · 泛函 · Neural Networks ·

2022 年 12 月 21 日

Forward- or Reverse-Mode Automatic Differentiation: What's the Difference?

翻译：前向或反向模式自动区分:区别是什么?

Birthe van den Berg,Tom Schrijvers,James McKinna,Alexander Vandenbroucke

Automatic differentiation (AD) has been a topic of interest for researchers in many disciplines, with increased popularity since its application to machine learning and neural networks. Although many researchers appreciate and know how to apply AD, it remains a challenge to truly understand the underlying processes. From an algebraic point of view, however, AD appears surprisingly natural: it originates from the differentiation laws. In this work we use Algebra of Programming techniques to reason about different AD variants, leveraging Haskell to illustrate our observations. Our findings stem from three fundamental algebraic abstractions: (1) the notion of module over a semiring, (2) Nagata's construction of the 'idealization of a module', and (3) Kronecker's delta function, that together allow us to write a single-line abstract definition of AD. From this single-line definition, and by instantiating our algebraic structures in various ways, we derive different AD variants, that have the same extensional behaviour, but different intensional properties, mainly in terms of (asymptotic) computational complexity. We show the different variants equivalent by means of Kronecker isomorphisms, a further elaboration of our Haskell infrastructure which guarantees correctness by construction. With this framework in place, this paper seeks to make AD variants more comprehensible, taking an algebraic perspective on the matter.

翻译：自动差异( AD) 是许多学科的研究人员感兴趣的话题, 自机器学习和神经网络应用以来, 自动差异( AD) 已经越来越受欢迎。虽然许多研究人员欣赏并知道如何应用自动, 但对于真正理解基础过程来说仍然是一个挑战。但是, 从代数角度看, AD 似乎令人惊讶地自然: 它起源于差异法。在这项工作中, 我们使用编程技术代数来解释不同的AD变量, 利用Haskell 来说明我们的观察。我们的发现来自三个基本的代数抽象:(1) 模块在半数模型上的概念, (2) Nagata 构建“ 模块化模块” 和(3) Kronecker 的三角体功能, 这使得我们一起能够写出一个单线的AD抽象定义。从这一单线定义中, 并且通过以不同方式即刻现我们的代数结构结构, 我们从不同的AD变体变体变体模型中得出了相同的扩展行为, 但不同的恒变体特性, 主要是( 模拟) 计算复杂性。我们展示不同的变体的变体,, 也就是, 我们的变体, 以一种可变体结构的的方式, 以变体的结构的选择的的的, 的的的的。

0

相关内容

Haskell

Haskell 是一种纯函数式编程语言，于 1990 年在编程语言 Miranda 的基础上标准化，并且以 λ 演算为基础发展而来。

ICLR 2021杰出论文奖出炉，8篇论文上榜！

专知会员服务

26+阅读 · 2021年4月2日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

中国图象图形学学会CSIG

1+阅读 · 2021年11月11日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

瘦素调节2型糖尿病大鼠交感神经活性及压力反射敏感性的机制

国家自然科学基金

0+阅读 · 2015年12月31日

应用膜蛋白纳米组装研究EGFR/HER2过表达致癌的分子机理与结构

国家自然科学基金

0+阅读 · 2014年12月31日

混凝土Weibull统计尺寸效应理论模型改进研究

国家自然科学基金

0+阅读 · 2013年12月31日

稀土有机-无机杂化近红外量子剪裁纳米材料的制备与性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

多环（杂）芳烃桥联双金属化合物的合成及其性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

稀土掺杂石英基异构集成式微结构光纤特性及超连续谱产生机理的研究

国家自然科学基金

0+阅读 · 2012年12月31日

高迁移率族蛋白1通过上调肝癌病人kupffer细胞Toll样受体和IL-33表达来促进Th17细胞的功能

国家自然科学基金

0+阅读 · 2012年12月31日

基于室温固体氧化物燃料电池的超晶格电解质界面效应研究

国家自然科学基金

0+阅读 · 2012年12月31日

LIF受体乙酰化介导的代谢异常在乳腺癌中的功能及机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

TfR抗体和CTX修饰纳米载体介导hTERTC27治疗神经胶质瘤

国家自然科学基金

0+阅读 · 2009年12月31日

On some limitations of probabilistic models for dimension-reduction: Illustration in the case of probabilistic formulations of partial least squares

Arxiv

0+阅读 · 2023年2月22日

Variational inference in neural functional prior using normalizing flows: Application to differential equation and operator learning problems

Arxiv

0+阅读 · 2023年2月21日

Sharp analysis of EM for learning mixtures of pairwise differences

Arxiv

0+阅读 · 2023年2月20日

Reverse Differentiation via Predictive Coding

Arxiv

0+阅读 · 2023年2月20日

Learning to Increase the Power of Conditional Randomization Tests

Arxiv

0+阅读 · 2023年2月19日

Collocation methods for second and higher order systems

Arxiv

0+阅读 · 2023年2月17日

Efficiently Forgetting What You Have Learned in Graph Representation Learning via Projection

Arxiv

0+阅读 · 2023年2月17日

Doubly transitive equiangular tight frames that contain regular simplices

Arxiv

0+阅读 · 2023年2月17日

Learning with Differentiable Algorithms

Arxiv

11+阅读 · 2022年9月1日

Differentiable Dynamic Programming for Structured Prediction and Attention

Arxiv

56+阅读 · 2018年2月20日

VIP会员

文章信息

相关主题

Neural Networks

相关VIP内容

ICLR 2021杰出论文奖出炉，8篇论文上榜！

专知会员服务

26+阅读 · 2021年4月2日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

【新书】数字图像(影像)处理手第二版，2176pdf，Mathematical Methods in Imaging

专知会员服务

93+阅读 · 2020年2月12日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

《物联网（IoT）中的无人机通信高效控制》135页

《在GNSS信号降级环境中利用共识实现无人机集群稳健协调》

中程单向攻击无人机的战略意义：俄乌战争启示

《面向无人机集群的避障动态传感器覆盖算法》最新38页

相关资讯

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium9

中国图象图形学学会CSIG

0+阅读 · 2021年12月17日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium5

中国图象图形学学会CSIG

1+阅读 · 2021年11月11日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium4

中国图象图形学学会CSIG

0+阅读 · 2021年11月10日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

On some limitations of probabilistic models for dimension-reduction: Illustration in the case of probabilistic formulations of partial least squares

Arxiv

0+阅读 · 2023年2月22日

Variational inference in neural functional prior using normalizing flows: Application to differential equation and operator learning problems

Arxiv

0+阅读 · 2023年2月21日

Sharp analysis of EM for learning mixtures of pairwise differences

Arxiv

0+阅读 · 2023年2月20日

Reverse Differentiation via Predictive Coding

Arxiv

0+阅读 · 2023年2月20日

Learning to Increase the Power of Conditional Randomization Tests

Arxiv

0+阅读 · 2023年2月19日

Collocation methods for second and higher order systems

Arxiv

0+阅读 · 2023年2月17日

Efficiently Forgetting What You Have Learned in Graph Representation Learning via Projection

Arxiv

0+阅读 · 2023年2月17日

Doubly transitive equiangular tight frames that contain regular simplices

Arxiv

0+阅读 · 2023年2月17日

Learning with Differentiable Algorithms

Arxiv

11+阅读 · 2022年9月1日

Differentiable Dynamic Programming for Structured Prediction and Attention

Arxiv

56+阅读 · 2018年2月20日

相关基金

瘦素调节2型糖尿病大鼠交感神经活性及压力反射敏感性的机制

国家自然科学基金

0+阅读 · 2015年12月31日

应用膜蛋白纳米组装研究EGFR/HER2过表达致癌的分子机理与结构

国家自然科学基金

0+阅读 · 2014年12月31日

混凝土Weibull统计尺寸效应理论模型改进研究

国家自然科学基金

0+阅读 · 2013年12月31日

稀土有机-无机杂化近红外量子剪裁纳米材料的制备与性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

多环（杂）芳烃桥联双金属化合物的合成及其性能研究

国家自然科学基金

0+阅读 · 2012年12月31日

稀土掺杂石英基异构集成式微结构光纤特性及超连续谱产生机理的研究

国家自然科学基金

0+阅读 · 2012年12月31日

高迁移率族蛋白1通过上调肝癌病人kupffer细胞Toll样受体和IL-33表达来促进Th17细胞的功能

国家自然科学基金

0+阅读 · 2012年12月31日

基于室温固体氧化物燃料电池的超晶格电解质界面效应研究

国家自然科学基金

0+阅读 · 2012年12月31日

LIF受体乙酰化介导的代谢异常在乳腺癌中的功能及机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

TfR抗体和CTX修饰纳米载体介导hTERTC27治疗神经胶质瘤

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员