文本控制的愿景模型概念代数 (Concept Algebra for Text-Controlled Vision Models)

This paper concerns the control of text-guided generative models, where a user provides a natural language prompt and the model generates samples based on this input. Prompting is intuitive, general, and flexible. However, there are significant limitations: prompting can fail in surprising ways, and it is often unclear how to find a prompt that will elicit some desired target behavior. A core difficulty for developing methods to overcome these issues is that failures are know-it-when-you-see-it -- it's hard to fix bugs if you can't state precisely what the model should have done! In this paper, we introduce a formalization of "what the user intended" in terms of latent concepts implicit to the data generating process that the model was trained on. This formalization allows us to identify some fundamental limitations of prompting. We then use the formalism to develop concept algebra to overcome these limitations. Concept algebra is a way of directly manipulating the concepts expressed in the output through algebraic operations on a suitably defined representation of input prompts. We give examples using concept algebra to overcome limitations of prompting, including concept transfer through arithmetic, and concept nullification through projection. Code available at https://github.com/zihao12/concept-algebra.

翻译：本文涉及对文本指导的基因模型的控制, 即用户提供自然语言提示, 且该模型根据此输入生成样本。提示是直观的、一般性的和灵活的。但是, 有许多限制: 提示会以令人惊讶的方式失败, 并且往往不清楚如何找到能够引起某些预期目标行为的快速。制定方法克服这些问题的一个核心困难是, 失败是知道的 - 何时即时看- 失败 -- 如果您不能准确地说明该模型应该做什么的话, 则很难修复错误! 在本文中, 我们引入了“ 用户想要做什么” 的正规化, 即该模型所培训的生成过程隐含的潜在概念。这种正式化让我们能够找出提示的一些基本限制。然后我们用形式主义来开发概念代数来克服这些限制。概念代数是直接通过测算操作来调节输出中表达的概念, 即通过正确定义的投入表达速度来直接调节。我们用概念代数来举例说明“ 代数” 来克服概念的局限性, 包括通过算法/ 解释性概念/ 。

相关内容

MoDELS

关注 43

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

高效可扩展图神经网络的研究进展，Recent Advances in Efficient and Scalable Graph Neural Networks

专知会员服务

78+阅读 · 2022年3月15日

计算机科学课程与视频课件合集，Computer Science courses with video lectures

专知会员服务

37+阅读 · 2022年1月24日

威斯康辛大学《机器学习导论》2020秋季课程完结，课件、视频资源已开放

专知会员服务

16+阅读 · 2020年12月25日

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日