ILLUME: 通过与Jabber互动实现愿景语言模型合理化 (ILLUME: Rationalizing Vision-Language Models by Interacting with their Jabber) - 专知论文

会员服务 ·

0

INTERACT · MoDELS · 环 · Conformer · 训练数据 ·

2022 年 8 月 18 日

ILLUME: Rationalizing Vision-Language Models by Interacting with their Jabber

翻译：ILLUME: 通过与Jabber互动实现愿景语言模型合理化

Manuel Brack,Patrick Schramowski,Björn Deiseroth,Kristian Kersting

Bootstrapping from pre-trained language models has been proven to be an efficient approach for building foundation vision-language models (VLM) for tasks such as image captioning or visual question answering. However, it is difficult-if not impossible-to utilize it to make the model conform with user's rationales for specific answers. To elicit and reinforce commonsense reasons, we propose an iterative sampling and tuning paradigm, called ILLUME, that executes the following loop: Given an image-question-answer prompt, the VLM samples multiple candidate rationales, and a human critic provides minimal feedback via preference selection, used for fine-tuning. This loop increases the training data and gradually carves out the VLM's rationalization capabilities. Our exhaustive experiments demonstrate that ILLUME is competitive with standard supervised fine-tuning while using significantly fewer training data and only requiring minimal feedback.

翻译：实践证明,从经过培训的语文模型中引入引导是建立基本视觉语言模型(VLM)的高效方法,用于图像字幕或视觉问答等任务。然而,要使用模型使模型符合用户对具体答案的理由,即使并非不可能,也是困难的。为了获取和加强常识性理由,我们提议了一个迭代抽样和调试模式,称为ILLUME,该模式可实施以下循环:由于图像问答迅速,VLM抽样多个候选人理由,以及一位人类评论家通过选择优惠提供最低限度的反馈,用于微调。这一循环增加了培训数据,并逐渐将VLM的合理化能力绘制出来。我们详尽的实验表明,ILLUME与标准监管的微调具有竞争力,同时使用的培训数据要少得多,只需要最低限度的反馈。

0

相关内容

INTERACT

IFIP TC13 Conference on Human-Computer Interaction是人机交互领域的研究者和实践者展示其工作的重要平台。多年来，这些会议吸引了来自几个国家和文化的研究人员。官网链接：http://interact2019.org/

最新《自监督表示学习》报告，70页ppt

最新《自监督表示学习》报告，70页ppt

专知会员服务

86+阅读 · 2020年12月22日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

专知会员服务

108+阅读 · 2020年5月1日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

【CCL 2019】ATT-第19期：文本生成 |Text Generation: From the Perspective of Interactive Inference （张家俊）

【CCL 2019】ATT-第19期：文本生成 |Text Generation: From the Perspective of Interactive Inference （张家俊）

专知会员服务

43+阅读 · 2019年11月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

与心房颤动相关的HAND1等转录因子基因突变鉴定及功能研究

国家自然科学基金

0+阅读 · 2015年12月31日

水溶液中多种痕量重金属元素的高灵敏度激光诱导击穿光谱

国家自然科学基金

0+阅读 · 2014年12月31日

随机泛函微分方程的适定性与渐近性分析

国家自然科学基金

0+阅读 · 2012年12月31日

活化的PLC-γ及与Akt关联调控OA软骨基质代谢的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

TGF-β3对角膜基质体外三维培养模型中纤维化细胞外基质合成的影响

国家自然科学基金

0+阅读 · 2012年12月31日

衰老相关长链非编码RNA的鉴定及其功能研究

国家自然科学基金

0+阅读 · 2012年12月31日

玉米种子老化的表观遗传机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

HC-SCR反应中乙醇催化制氢与还原剂活化耦合研究

国家自然科学基金

0+阅读 · 2011年12月31日

泥沙在吸附和絮凝过程中表面电荷的特性研究

国家自然科学基金

0+阅读 · 2009年12月31日

食管癌中靶向调控fascin基因的 miRNA的鉴定及其表达调控机制

国家自然科学基金

0+阅读 · 2009年12月31日

Binding Language Models in Symbolic Languages

Arxiv

0+阅读 · 2022年10月6日

Learning to Prompt for Vision-Language Models

Arxiv

0+阅读 · 2022年10月6日

Distilling Task-specific Logical Rules from Large Pre-trained Models

Arxiv

0+阅读 · 2022年10月6日

Star-Graph Multimodal Matching Component Analysis for Data Fusion and Transfer Learning

Arxiv

0+阅读 · 2022年10月5日

Prompt Learning with Optimal Transport for Vision-Language Models

Arxiv

1+阅读 · 2022年10月3日

Language-Aware Soft Prompting for Vision & Language Foundation Models

Arxiv

0+阅读 · 2022年10月3日

A Hybrid Compositional Reasoning Approach for Interactive Robot Manipulation

Arxiv

0+阅读 · 2022年10月3日

Medical Image Understanding with Pretrained Vision Language Models: A Comprehensive Study

Arxiv

0+阅读 · 2022年9月30日

Language Models Can Teach Themselves to Program Better

Arxiv

0+阅读 · 2022年9月30日

Compositional Semantic Parsing with Large Language Models

Arxiv

0+阅读 · 2022年9月30日

VIP会员

文章信息

相关主题

相关VIP内容

最新《自监督表示学习》报告，70页ppt

最新《自监督表示学习》报告，70页ppt

专知会员服务

86+阅读 · 2020年12月22日

2020数据工程师成长路线图

专知会员服务

41+阅读 · 2020年9月6日

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

【2020新书】自然语言处理Python与spaCy实践，216页pdf，NLP with Python

专知会员服务

108+阅读 · 2020年5月1日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

166+阅读 · 2020年3月18日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

【CCL 2019】ATT-第19期：文本生成 |Text Generation: From the Perspective of Interactive Inference （张家俊）

【CCL 2019】ATT-第19期：文本生成 |Text Generation: From the Perspective of Interactive Inference （张家俊）

专知会员服务

43+阅读 · 2019年11月12日

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

49+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

【CMU博士论文】数据驱动决策中的激励、信息与不确定性

DGP双粒度提示框架：图增强大模型助力欺诈检测

【ICCV2025】ESSENTIAL：用于视频类增量学习的情景记忆与语义记忆整合

唯快不破：大型语言模型高效架构综述

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium1

中国图象图形学学会CSIG

0+阅读 · 2021年11月3日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

Binding Language Models in Symbolic Languages

Arxiv

0+阅读 · 2022年10月6日

Learning to Prompt for Vision-Language Models

Arxiv

0+阅读 · 2022年10月6日

Distilling Task-specific Logical Rules from Large Pre-trained Models

Arxiv

0+阅读 · 2022年10月6日

Star-Graph Multimodal Matching Component Analysis for Data Fusion and Transfer Learning

Arxiv

0+阅读 · 2022年10月5日

Prompt Learning with Optimal Transport for Vision-Language Models

Arxiv

1+阅读 · 2022年10月3日

Language-Aware Soft Prompting for Vision & Language Foundation Models

Arxiv

0+阅读 · 2022年10月3日

A Hybrid Compositional Reasoning Approach for Interactive Robot Manipulation

Arxiv

0+阅读 · 2022年10月3日

Medical Image Understanding with Pretrained Vision Language Models: A Comprehensive Study

Arxiv

0+阅读 · 2022年9月30日

Language Models Can Teach Themselves to Program Better

Arxiv

0+阅读 · 2022年9月30日

Compositional Semantic Parsing with Large Language Models

Arxiv

0+阅读 · 2022年9月30日

相关基金

与心房颤动相关的HAND1等转录因子基因突变鉴定及功能研究

国家自然科学基金

0+阅读 · 2015年12月31日

水溶液中多种痕量重金属元素的高灵敏度激光诱导击穿光谱

国家自然科学基金

0+阅读 · 2014年12月31日

随机泛函微分方程的适定性与渐近性分析

国家自然科学基金

0+阅读 · 2012年12月31日

活化的PLC-γ及与Akt关联调控OA软骨基质代谢的机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

TGF-β3对角膜基质体外三维培养模型中纤维化细胞外基质合成的影响

国家自然科学基金

0+阅读 · 2012年12月31日

衰老相关长链非编码RNA的鉴定及其功能研究

国家自然科学基金

0+阅读 · 2012年12月31日

玉米种子老化的表观遗传机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

HC-SCR反应中乙醇催化制氢与还原剂活化耦合研究

国家自然科学基金

0+阅读 · 2011年12月31日

泥沙在吸附和絮凝过程中表面电荷的特性研究

国家自然科学基金

0+阅读 · 2009年12月31日

食管癌中靶向调控fascin基因的 miRNA的鉴定及其表达调控机制

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员