ApacheJIT: 及时失灵预测的大型数据集 (ApacheJIT: A Large Dataset for Just-In-Time Defect Prediction) - 专知论文

会员服务 ·

0

讲稿 · 数据集 · 学成 · MoDELS · Machine Learning ·

2022 年 4 月 30 日

ApacheJIT: A Large Dataset for Just-In-Time Defect Prediction

翻译：ApacheJIT: 及时失灵预测的大型数据集

Hossein Keshavarz,Meiyappan Nagappan

In this paper, we present ApacheJIT, a large dataset for Just-In-Time defect prediction. ApacheJIT consists of clean and bug-inducing software changes in popular Apache projects. ApacheJIT has a total of 106,674 commits (28,239 bug-inducing and 78,435 clean commits). Having a large number of commits makes ApacheJIT a suitable dataset for machine learning models, especially deep learning models that require large training sets to effectively generalize the patterns present in the historical data to future data.

翻译：在本文中,我们介绍ApacheJIT,这是一个用于 " 时对时的错误预测 " 的大型数据集。ApacheJIT由流行的Apache项目中清洁和诱虫软件变化构成。ApacheJIT共有106,674项承诺(28,239项诱虫和78,435项清洁承诺 ) 。大量承诺使ApacheJIT成为机器学习模型的合适数据集,特别是需要大型培训的深层学习模型,以便有效地将历史数据中存在的模式归纳为未来数据。

0

相关内容

【超赞的#C++#速查&信息图】“hacking c++ - Cheat Sheets & Infographics”

【超赞的#C++#速查&信息图】“hacking c++ - Cheat Sheets & Infographics”

专知会员服务

30+阅读 · 2022年3月8日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium7

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium7

中国图象图形学学会CSIG

0+阅读 · 2021年11月15日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

金属碳化物基低铂介孔催化材料的合成、界面设计与电催化性能研究

国家自然科学基金

0+阅读 · 2015年12月31日

高效Ag基阳极析氧催化剂的原位调控构筑及其电解水制氢研究

国家自然科学基金

0+阅读 · 2014年12月31日

微纳米结构Ti/Mg双连续相复合材料的可控制备及力学性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

高效光纤型SERS探针的界面设计、可控制备及性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

硅酸二钙的结构参数与水化活性的关系

国家自然科学基金

0+阅读 · 2012年12月31日

基于Linked Open Data的Web服务语义互操作关键技术

国家自然科学基金

0+阅读 · 2012年12月31日

芳代稠环类结晶诱导荧光增强材料

国家自然科学基金

0+阅读 · 2011年12月31日

基于模糊理论的冰情预报及冰凌灾害风险分析

国家自然科学基金

0+阅读 · 2011年12月31日

结晶高分子材料晶态结构与力学性能关系的同步辐射原位研究

国家自然科学基金

0+阅读 · 2009年12月31日

混凝土桥梁构件耐久性数值模拟

国家自然科学基金

0+阅读 · 2008年12月31日

Evolution through Large Models

Evolution through Large Models

Arxiv

0+阅读 · 2022年6月17日

Local Attention Graph-based Transformer for Multi-target Genetic Alteration Prediction

Arxiv

0+阅读 · 2022年6月17日

Evaluation of Contrastive Learning with Various Code Representations for Code Clone Detection

Arxiv

0+阅读 · 2022年6月17日

Online Score Statistics for Detecting Clustered Change in Network Point Processes

Arxiv

0+阅读 · 2022年6月16日

An Empirical Study on the Effectiveness of Data Resampling Approaches for Cross-Project Software Defect Prediction

Arxiv

0+阅读 · 2022年6月16日

Off-Policy Evaluation for Large Action Spaces via Embeddings

Arxiv

0+阅读 · 2022年6月16日

Conformal prediction set for time-series

Arxiv

0+阅读 · 2022年6月15日

CLEF. A Linked Open Data native system for Crowdsourcing

Arxiv

0+阅读 · 2022年6月1日

Adaptive Synthetic Characters for Military Training

Adaptive Synthetic Characters for Military Training

Arxiv

50+阅读 · 2021年1月6日

Time-Series Event Prediction with Evolutionary State Graph

Arxiv

14+阅读 · 2020年11月25日

VIP会员

文章信息

相关主题

Machine Learning

相关VIP内容

【超赞的#C++#速查&信息图】“hacking c++ - Cheat Sheets & Infographics”

【超赞的#C++#速查&信息图】“hacking c++ - Cheat Sheets & Infographics”

专知会员服务

30+阅读 · 2022年3月8日

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

图像分类技巧集，17页ppt《Bag of Tricks for Image Classification》

专知会员服务

96+阅读 · 2020年3月12日

【深度学习表格检测、信息提取和结构化】《Table Detection, Information Extraction and Structuring using Deep Learning》by Vihar Kurama

专知会员服务

38+阅读 · 2020年1月23日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

160+阅读 · 2019年10月12日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

机器学习入门的经验与建议

机器学习入门的经验与建议

专知会员服务

94+阅读 · 2019年10月10日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用

专知会员服务

41+阅读 · 2019年10月9日

热门VIP内容

开通专知VIP会员享更多权益服务

不确定环境下无人机三维路径规划研究 | 221页

远征作战军事后勤规划

大语言模型将如何改变军事指挥结构

美陆军能力集成与开发系统（ACIDS）流程指南 | 2025最新122页

相关资讯

ACM MM 2022 Call for Papers

ACM MM 2022 Call for Papers

CCF多媒体专委会

5+阅读 · 2022年3月29日

ACM TOMM Call for Papers

ACM TOMM Call for Papers

CCF多媒体专委会

2+阅读 · 2022年3月23日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium7

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium7

中国图象图形学学会CSIG

0+阅读 · 2021年11月15日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium6

中国图象图形学学会CSIG

2+阅读 · 2021年11月12日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium3

中国图象图形学学会CSIG

0+阅读 · 2021年11月9日

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

【ICIG2021】Check out the hot new trailer of ICIG2021 Symposium2

中国图象图形学学会CSIG

0+阅读 · 2021年11月8日

Hierarchically Structured Meta-learning

Hierarchically Structured Meta-learning

CreateAMind

27+阅读 · 2019年5月22日

A Technical Overview of AI & ML in 2018 & Trends for 2019

A Technical Overview of AI & ML in 2018 & Trends for 2019

待字闺中

18+阅读 · 2018年12月24日

disentangled-representation-papers

disentangled-representation-papers

CreateAMind

26+阅读 · 2018年9月12日

相关论文

Evolution through Large Models

Evolution through Large Models

Arxiv

0+阅读 · 2022年6月17日

Local Attention Graph-based Transformer for Multi-target Genetic Alteration Prediction

Arxiv

0+阅读 · 2022年6月17日

Evaluation of Contrastive Learning with Various Code Representations for Code Clone Detection

Arxiv

0+阅读 · 2022年6月17日

Online Score Statistics for Detecting Clustered Change in Network Point Processes

Arxiv

0+阅读 · 2022年6月16日

An Empirical Study on the Effectiveness of Data Resampling Approaches for Cross-Project Software Defect Prediction

Arxiv

0+阅读 · 2022年6月16日

Off-Policy Evaluation for Large Action Spaces via Embeddings

Arxiv

0+阅读 · 2022年6月16日

Conformal prediction set for time-series

Arxiv

0+阅读 · 2022年6月15日

CLEF. A Linked Open Data native system for Crowdsourcing

Arxiv

0+阅读 · 2022年6月1日

Adaptive Synthetic Characters for Military Training

Adaptive Synthetic Characters for Military Training

Arxiv

50+阅读 · 2021年1月6日

Time-Series Event Prediction with Evolutionary State Graph

Arxiv

14+阅读 · 2020年11月25日

相关基金

金属碳化物基低铂介孔催化材料的合成、界面设计与电催化性能研究

国家自然科学基金

0+阅读 · 2015年12月31日

高效Ag基阳极析氧催化剂的原位调控构筑及其电解水制氢研究

国家自然科学基金

0+阅读 · 2014年12月31日

微纳米结构Ti/Mg双连续相复合材料的可控制备及力学性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

高效光纤型SERS探针的界面设计、可控制备及性能研究

国家自然科学基金

0+阅读 · 2013年12月31日

硅酸二钙的结构参数与水化活性的关系

国家自然科学基金

0+阅读 · 2012年12月31日

基于Linked Open Data的Web服务语义互操作关键技术

国家自然科学基金

0+阅读 · 2012年12月31日

芳代稠环类结晶诱导荧光增强材料

国家自然科学基金

0+阅读 · 2011年12月31日

基于模糊理论的冰情预报及冰凌灾害风险分析

国家自然科学基金

0+阅读 · 2011年12月31日

结晶高分子材料晶态结构与力学性能关系的同步辐射原位研究

国家自然科学基金

0+阅读 · 2009年12月31日

混凝土桥梁构件耐久性数值模拟

国家自然科学基金

0+阅读 · 2008年12月31日

微信扫码咨询专知VIP会员