以多重不确定性固定化方式通过文本反馈合成图像检索合成的图像检索器 (Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization) - 专知论文

会员服务 ·

0

查全率/召回率 · 正则化项 · 图像检索 · MoDELS · 振荡 ·

2022 年 11 月 29 日

Composed Image Retrieval with Text Feedback via Multi-grained Uncertainty Regularization

翻译：以多重不确定性固定化方式通过文本反馈合成图像检索合成的图像检索器

Yiyang Chen,Zhedong Zheng,Wei Ji,Leigang Qu,Tat-Seng Chua

We investigate composed image retrieval with text feedback. Users gradually look for the target of interest by moving from coarse to fine-grained feedback. However, existing methods merely focus on the latter, i.e, fine-grained search, by harnessing positive and negative pairs during training. This pair-based paradigm only considers the one-to-one distance between a pair of specific points, which is not aligned with the one-to-many coarse-grained retrieval process and compromises the recall rate. In an attempt to fill this gap, we introduce a unified learning approach to simultaneously modeling the coarse- and fine-grained retrieval by considering the multi-grained uncertainty. The key idea underpinning the proposed method is to integrate fine- and coarse-grained retrieval as matching data points with small and large fluctuations, respectively. Specifically, our method contains two modules: uncertainty modeling and uncertainty regularization. (1) The uncertainty modeling simulates the multi-grained queries by introducing identically distributed fluctuations in the feature space. (2) Based on the uncertainty modeling, we further introduce uncertainty regularization to adapt the matching objective according to the fluctuation range. Compared with existing methods, the proposed strategy explicitly prevents the model from pushing away potential candidates in the early stage, and thus improves the recall rate. On the three public datasets, i.e., FashionIQ, Fashion200k, and Shoes, the proposed method has achieved +4.03%, + 3.38%, and + 2.40% Recall@50 accuracy over a strong baseline, respectively.

翻译：我们用文本反馈来调查图像检索。用户通过从粗糙到细细的反馈逐渐寻找感兴趣的目标。但是, 现有方法仅仅侧重于后者, 即精细的搜索, 在培训期间使用正对和负对对。这种双向模式只考虑一对特定点之间的一到一距离, 这与一到多粗粗的检索进程不相符, 并会降低回调率。为了填补这一空白, 我们引入了一种统一的学习方法, 以同时模拟粗略和细微的回补, 同时考虑到多重的不确定性。但是, 现有方法的主要理念是将精细和粗粗粗的检索分别结合成小和大波动的数据点。具体地说, 我们的方法包含两个模块: 不确定性模型和不确定性规范。 (1) 不确定性模拟多粗重的查询, 在特性空间中引入相同的分布波动。 (2) 基于不确定性模型, 我们进一步引入不确定性+精准的Q+精细的检索模型, 从多重的精确度模型到匹配目标的F- 3。将精细的检索作为基础, 分别用于推进F- k 的F- 和精确的推移算, 因此, 将现有方法, 防止了现有方法的回调。

0

相关内容

查全率/召回率

查全率/召回率

【CVPR 2022】基于层次化视觉语言知识蒸馏的开放词汇单阶段检测，Improving Visual Grounding with Visual-Linguistic Verification and Iterative Reasoning

【CVPR 2022】基于层次化视觉语言知识蒸馏的开放词汇单阶段检测，Improving Visual Grounding with Visual-Linguistic Verification and Iterative Reasoning

专知会员服务

7+阅读 · 2022年3月19日

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

专知会员服务

10+阅读 · 2022年3月19日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

骨髓微环境调控骨髓瘤细胞RANKL表达的作用机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

ING3：原发性肝癌的诊断与治疗新靶点

国家自然科学基金

0+阅读 · 2012年12月31日

针刺对慢性应激抑郁大鼠脑细胞微环境的影响

国家自然科学基金

0+阅读 · 2012年12月31日

IKIP在直肠癌术前放疗敏感性预测中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

维吾尔族睡眠呼吸暂停患者呼吸调节功能及其遗传学基础研究

国家自然科学基金

0+阅读 · 2012年12月31日

MRP1/ABCC1基因3＇UTR单核苷酸多态性介导miRNA对原发性肝癌多药耐药性的影响

国家自然科学基金

0+阅读 · 2012年12月31日

原发性肝细胞癌微灌注及弹性模量状态与复发机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

REMg2TMx型多相合金的吸/放氢行为和衰减机理研究

国家自然科学基金

0+阅读 · 2009年12月31日

肾上腺源性及原发性高血压线粒体tRNAIle、tRNALeu(UUR)和tRNAlys基因突变的差异对比研究

国家自然科学基金

0+阅读 · 2009年12月31日

Uncertainty-Driven Dense Two-View Structure from Motion

Uncertainty-Driven Dense Two-View Structure from Motion

Arxiv

0+阅读 · 2023年2月1日

Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization

Arxiv

0+阅读 · 2023年2月1日

Learning Against Distributional Uncertainty: On the Trade-off Between Robustness and Specificity

Arxiv

0+阅读 · 2023年1月31日

NoiseTransfer: Image Noise Generation with Contrastive Embeddings

Arxiv

0+阅读 · 2023年1月31日

Better Uncertainty Calibration via Proper Scores for Classification and Beyond

Arxiv

0+阅读 · 2023年1月30日

Are Random Decompositions all we need in High Dimensional Bayesian Optimisation?

Arxiv

0+阅读 · 2023年1月30日

Domain Adaptive Video Semantic Segmentation via Cross-Domain Moving Object Mixing

Arxiv

0+阅读 · 2023年1月27日

Deep Image Retrieval: A Survey

Arxiv

16+阅读 · 2021年1月27日

Embedding Uncertain Knowledge Graphs

Arxiv

12+阅读 · 2019年2月26日

Consensus Based Medical Image Segmentation Using Semi-Supervised Learning And Graph Cuts

Arxiv

11+阅读 · 2018年5月21日

VIP会员

文章信息

相关主题

查全率/召回率

相关VIP内容

【CVPR 2022】基于层次化视觉语言知识蒸馏的开放词汇单阶段检测，Improving Visual Grounding with Visual-Linguistic Verification and Iterative Reasoning

【CVPR 2022】基于层次化视觉语言知识蒸馏的开放词汇单阶段检测，Improving Visual Grounding with Visual-Linguistic Verification and Iterative Reasoning

专知会员服务

7+阅读 · 2022年3月19日

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

【CVPR 2022】基于粗粒度和细粒度特征匹配的视频描述评估，EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching

专知会员服务

10+阅读 · 2022年3月19日

50+篇《神经架构搜索NAS》2020论文合集

专知会员服务

61+阅读 · 2020年3月19日

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

100+篇《自监督学习(Self-Supervised Learning)》论文最新合集

专知会员服务

165+阅读 · 2020年3月18日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

95+阅读 · 2020年3月12日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

【加州大学伯克利分校博士论文】通过自我监督预测学习泛化

专知会员服务

65+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

从社会学实验到行为仿真：理解基于Agent的观点动力学建模思维

中英文版《GPT-5 System Card速览》报告

ACL 2025 | 大模型结构化知识提示的泛化能力研究

【普林斯顿博士论文】大型模型的高效推理

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

IEEE ICKG 2022: Call for Papers

IEEE ICKG 2022: Call for Papers

机器学习与推荐算法

3+阅读 · 2022年3月30日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Latest News & Announcements of the Plenary Talk1

【ICIG2021】Latest News & Announcements of the Plenary Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年11月1日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

【论文推荐】最新八篇情感分析相关论文—Pair-wise判别器、多模态情感分析、上下文语境、Gated 卷积网络

专知

20+阅读 · 2018年6月29日

相关论文

Uncertainty-Driven Dense Two-View Structure from Motion

Uncertainty-Driven Dense Two-View Structure from Motion

Arxiv

0+阅读 · 2023年2月1日

Learning Generalized Zero-Shot Learners for Open-Domain Image Geolocalization

Arxiv

0+阅读 · 2023年2月1日

Learning Against Distributional Uncertainty: On the Trade-off Between Robustness and Specificity

Arxiv

0+阅读 · 2023年1月31日

NoiseTransfer: Image Noise Generation with Contrastive Embeddings

Arxiv

0+阅读 · 2023年1月31日

Better Uncertainty Calibration via Proper Scores for Classification and Beyond

Arxiv

0+阅读 · 2023年1月30日

Are Random Decompositions all we need in High Dimensional Bayesian Optimisation?

Arxiv

0+阅读 · 2023年1月30日

Domain Adaptive Video Semantic Segmentation via Cross-Domain Moving Object Mixing

Arxiv

0+阅读 · 2023年1月27日

Deep Image Retrieval: A Survey

Arxiv

16+阅读 · 2021年1月27日

Embedding Uncertain Knowledge Graphs

Arxiv

12+阅读 · 2019年2月26日

Consensus Based Medical Image Segmentation Using Semi-Supervised Learning And Graph Cuts

Arxiv

11+阅读 · 2018年5月21日

相关基金

骨髓微环境调控骨髓瘤细胞RANKL表达的作用机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

基于SURE/PURE准则的图像盲反卷积算法研究

国家自然科学基金

3+阅读 · 2013年12月31日

ING3：原发性肝癌的诊断与治疗新靶点

国家自然科学基金

0+阅读 · 2012年12月31日

针刺对慢性应激抑郁大鼠脑细胞微环境的影响

国家自然科学基金

0+阅读 · 2012年12月31日

IKIP在直肠癌术前放疗敏感性预测中的作用及分子机制

国家自然科学基金

0+阅读 · 2012年12月31日

维吾尔族睡眠呼吸暂停患者呼吸调节功能及其遗传学基础研究

国家自然科学基金

0+阅读 · 2012年12月31日

MRP1/ABCC1基因3＇UTR单核苷酸多态性介导miRNA对原发性肝癌多药耐药性的影响

国家自然科学基金

0+阅读 · 2012年12月31日

原发性肝细胞癌微灌注及弹性模量状态与复发机制研究

国家自然科学基金

0+阅读 · 2011年12月31日

REMg2TMx型多相合金的吸/放氢行为和衰减机理研究

国家自然科学基金

0+阅读 · 2009年12月31日

肾上腺源性及原发性高血压线粒体tRNAIle、tRNALeu(UUR)和tRNAlys基因突变的差异对比研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员