Scalar 还不够: 以矢量为基础的无偏见学习到排名 (Scalar is Not Enough: Vectorization-based Unbiased Learning to Rank) - 专知论文

会员服务 ·

0

Learning · 秩 · 无偏 · 标量 · 分解的 ·

2022 年 6 月 3 日

Scalar is Not Enough: Vectorization-based Unbiased Learning to Rank

翻译：Scalar 还不够: 以矢量为基础的无偏见学习到排名

Mouxiang Chen,Chenghao Liu,Zemin Liu,Jianling Sun

from arxiv, Accepted by KDD 2022

Unbiased learning to rank (ULTR) aims to train an unbiased ranking model from biased user click logs. Most of the current ULTR methods are based on the examination hypothesis (EH), which assumes that the click probability can be factorized into two scalar functions, one related to ranking features and the other related to bias factors. Unfortunately, the interactions among features, bias factors and clicks are complicated in practice, and usually cannot be factorized in this independent way. Fitting click data with EH could lead to model misspecification and bring the approximation error. In this paper, we propose a vector-based EH and formulate the click probability as a dot product of two vector functions. This solution is complete due to its universality in fitting arbitrary click functions. Based on it, we propose a novel model named Vectorization to adaptively learn the relevance embeddings and sort documents by projecting embeddings onto a base vector. Extensive experiments show that our method significantly outperforms the state-of-the-art ULTR methods on complex real clicks as well as simple simulated clicks.

翻译：无偏见的排名学习(LUCTR)旨在从有偏向的用户点击日志中培训一个公正的排名模型(LUCTR) 。当前的LUCTR方法大多基于测试假设( EH) 。该假设假设假设假定点击概率可以被分解成两个星标函数, 一个与排名特征有关, 另一个与偏差因素有关。不幸的是, 特性、偏差因素和点击之间的相互作用在实践中很复杂, 通常无法以这种独立的方式进行分解。将点击数据与 EH 匹配可能导致模型错误的区分并带来近似错误。在本文中, 我们提议以矢量为基础的 EH 方法, 并将点击概率作为两个矢量函数的点产值。这个解决方案是完整的, 因为它在安装任意点击功能时具有普遍性。基于这个假设, 我们提议了一个名为矢量的新模型, 通过投射嵌入基矢量, 来适应性地学习嵌入和整理文件的关联性。广泛的实验显示, 我们的方法在复杂的实际点击上, 以及简单的模拟点击上, 大大超越了最先进的土控方法。

0

相关内容

Learning

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

BMAL1调控间充质干细胞在糖尿病性牙周炎骨缺损修复的机制

国家自然科学基金

0+阅读 · 2014年12月31日

马尾松高抗旱家系应答干旱胁迫的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

山羊朊蛋白基因（PRNP）转录调控因子的筛选与鉴定

国家自然科学基金

0+阅读 · 2012年12月31日

SARI转录抑制机制及在急性髓细胞白血病发病中的作用

国家自然科学基金

0+阅读 · 2012年12月31日

针灸预处理防治运动性Th1/Th2细胞失衡的效应及作用机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

加工番茄可溶性固形物含量的全基因组关联分析与连锁作图

国家自然科学基金

0+阅读 · 2011年12月31日

我国小麦纹枯病菌Rhizoctonia cerealis的分子生态学研究

国家自然科学基金

0+阅读 · 2009年12月31日

氮素调控小麦籽粒淀粉粒粒级分布特征形成的生理生化机制

国家自然科学基金

0+阅读 · 2009年12月31日

肾上腺源性及原发性高血压线粒体tRNAIle、tRNALeu(UUR)和tRNAlys基因突变的差异对比研究

国家自然科学基金

0+阅读 · 2009年12月31日

慢性间断低氧对家兔颏舌肌运动皮质区调控上气道扩张肌的影响及作用机制

国家自然科学基金

0+阅读 · 2008年12月31日

Maximum Likelihood Imputation

Arxiv

0+阅读 · 2022年7月20日

FedDM: Iterative Distribution Matching for Communication-Efficient Federated Learning

Arxiv

0+阅读 · 2022年7月20日

Accommodating false positives within acoustic spatial capture-recapture, with variable source levels, noisy bearings and an inhomogeneous spatial density

Accommodating false positives within acoustic spatial capture-recapture, with variable source levels, noisy bearings and an inhomogeneous spatial density

Arxiv

0+阅读 · 2022年7月19日

Is Integer Arithmetic Enough for Deep Learning Training?

Arxiv

0+阅读 · 2022年7月18日

High-dimensional robust approximated M-estimators for mean regression with asymmetric data

Arxiv

0+阅读 · 2022年7月18日

Streaming Algorithms for Support-Aware Histograms

Arxiv

0+阅读 · 2022年7月18日

A General Framework for Pairwise Unbiased Learning to Rank

Arxiv

0+阅读 · 2022年7月18日

Adaptive Behavioral Model Learning for Software Product Lines

Arxiv

0+阅读 · 2022年7月16日

An Approach for Link Prediction in Directed Complex Networks based on Asymmetric Similarity-Popularity

Arxiv

0+阅读 · 2022年7月15日

Bug Fix Time Optimization Using Matrix Factorization and Iterative Gale-Shaply Algorithms

Arxiv

0+阅读 · 2022年7月14日

VIP会员

文章信息

相关主题

相关VIP内容

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

不可错过！《机器学习100讲》课程，UBC Mark Schmidt讲授

专知会员服务

76+阅读 · 2022年6月28日

Linux导论，Introduction to Linux，96页ppt

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

【跨语言BERT模型大集合】Transfer learning is increasingly going multilingual with language-specific BERT models

专知会员服务

54+阅读 · 2020年1月30日

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

Aspect-Oriented Syntax Network for Aspect-Based Sentiment Analysis，中山大学数据科学与计算机学院权小军教授，第八届全国社会媒体处理大会SMP2019

专知会员服务

19+阅读 · 2019年10月22日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

Deep Learning Based Detection and Correction of Cardiac MR Motion Artefacts During Reconstruction for High-Quality Segmentation

专知会员服务

59+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

【哈佛大学商学院课程Fall 2019】机器学习可解释性

【哈佛大学商学院课程Fall 2019】机器学习可解释性

专知会员服务

105+阅读 · 2019年10月9日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

人工智能驾驶：旧理念与新技术

美军手册：战术心理战分遣队与小组指南 | 68页

军事机器学习设计：关于开发自动化任务摘要系统的梯次化设计科学研究 | 2025最新93页

美国防部自主系统研制试验与鉴定指南 | 2025年最新200页

相关资讯

【ICIG2021】Latest News & Announcements of the Workshop

【ICIG2021】Latest News & Announcements of the Workshop

中国图象图形学学会CSIG

0+阅读 · 2021年12月20日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium8

中国图象图形学学会CSIG

0+阅读 · 2021年11月16日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

【ICIG2021】Latest News & Announcements of the Industry Talk1

【ICIG2021】Latest News & Announcements of the Industry Talk1

中国图象图形学学会CSIG

0+阅读 · 2021年7月28日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

相关论文

Maximum Likelihood Imputation

Arxiv

0+阅读 · 2022年7月20日

FedDM: Iterative Distribution Matching for Communication-Efficient Federated Learning

Arxiv

0+阅读 · 2022年7月20日

Accommodating false positives within acoustic spatial capture-recapture, with variable source levels, noisy bearings and an inhomogeneous spatial density

Accommodating false positives within acoustic spatial capture-recapture, with variable source levels, noisy bearings and an inhomogeneous spatial density

Arxiv

0+阅读 · 2022年7月19日

Is Integer Arithmetic Enough for Deep Learning Training?

Arxiv

0+阅读 · 2022年7月18日

High-dimensional robust approximated M-estimators for mean regression with asymmetric data

Arxiv

0+阅读 · 2022年7月18日

Streaming Algorithms for Support-Aware Histograms

Arxiv

0+阅读 · 2022年7月18日

A General Framework for Pairwise Unbiased Learning to Rank

Arxiv

0+阅读 · 2022年7月18日

Adaptive Behavioral Model Learning for Software Product Lines

Arxiv

0+阅读 · 2022年7月16日

An Approach for Link Prediction in Directed Complex Networks based on Asymmetric Similarity-Popularity

Arxiv

0+阅读 · 2022年7月15日

Bug Fix Time Optimization Using Matrix Factorization and Iterative Gale-Shaply Algorithms

Arxiv

0+阅读 · 2022年7月14日

相关基金

BMAL1调控间充质干细胞在糖尿病性牙周炎骨缺损修复的机制

国家自然科学基金

0+阅读 · 2014年12月31日

马尾松高抗旱家系应答干旱胁迫的分子机理

国家自然科学基金

0+阅读 · 2012年12月31日

山羊朊蛋白基因（PRNP）转录调控因子的筛选与鉴定

国家自然科学基金

0+阅读 · 2012年12月31日

SARI转录抑制机制及在急性髓细胞白血病发病中的作用

国家自然科学基金

0+阅读 · 2012年12月31日

针灸预处理防治运动性Th1/Th2细胞失衡的效应及作用机制研究

国家自然科学基金

0+阅读 · 2012年12月31日

加工番茄可溶性固形物含量的全基因组关联分析与连锁作图

国家自然科学基金

0+阅读 · 2011年12月31日

我国小麦纹枯病菌Rhizoctonia cerealis的分子生态学研究

国家自然科学基金

0+阅读 · 2009年12月31日

氮素调控小麦籽粒淀粉粒粒级分布特征形成的生理生化机制

国家自然科学基金

0+阅读 · 2009年12月31日

肾上腺源性及原发性高血压线粒体tRNAIle、tRNALeu(UUR)和tRNAlys基因突变的差异对比研究

国家自然科学基金

0+阅读 · 2009年12月31日

慢性间断低氧对家兔颏舌肌运动皮质区调控上气道扩张肌的影响及作用机制

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员