再现性需要统一的工件

2023 IEEE/ACM 2nd International Conference on AI Engineering – Software Engineering for AI (CAIN) Pub Date : 2023-05-01 DOI:10.1109/CAIN58948.2023.00025

Iordanis Fostiropoulos, Bowman Brown, L. Itti

{"title":"再现性需要统一的工件","authors":"Iordanis Fostiropoulos, Bowman Brown, L. Itti","doi":"10.1109/CAIN58948.2023.00025","DOIUrl":null,"url":null,"abstract":"Machine learning is facing a ‘reproducibility crisis’ where a significant number of works report failures when attempting to reproduce previously published results. We evaluate the sources of reproducibility failures using a meta-analysis of 142 replication studies from ReScience C and 204 code repositories. We find that missing experiment details such as hyperparameters are potential causes of unreproducibility. We experimentally show the bias of different hyperparameter selection strategies and conclude that consolidated artifacts with a unified framework can help support reproducibility.","PeriodicalId":175580,"journal":{"name":"2023 IEEE/ACM 2nd International Conference on AI Engineering – Software Engineering for AI (CAIN)","volume":"320 1","pages":"0"},"PeriodicalIF":0.0000,"publicationDate":"2023-05-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"0","resultStr":"{\"title\":\"Reproducibility Requires Consolidated Artifacts\",\"authors\":\"Iordanis Fostiropoulos, Bowman Brown, L. Itti\",\"doi\":\"10.1109/CAIN58948.2023.00025\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Machine learning is facing a ‘reproducibility crisis’ where a significant number of works report failures when attempting to reproduce previously published results. We evaluate the sources of reproducibility failures using a meta-analysis of 142 replication studies from ReScience C and 204 code repositories. We find that missing experiment details such as hyperparameters are potential causes of unreproducibility. We experimentally show the bias of different hyperparameter selection strategies and conclude that consolidated artifacts with a unified framework can help support reproducibility.\",\"PeriodicalId\":175580,\"journal\":{\"name\":\"2023 IEEE/ACM 2nd International Conference on AI Engineering – Software Engineering for AI (CAIN)\",\"volume\":\"320 1\",\"pages\":\"0\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2023-05-01\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"0\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"2023 IEEE/ACM 2nd International Conference on AI Engineering – Software Engineering for AI (CAIN)\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.1109/CAIN58948.2023.00025\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"2023 IEEE/ACM 2nd International Conference on AI Engineering – Software Engineering for AI (CAIN)","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.1109/CAIN58948.2023.00025","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 0

摘要

机器学习正面临着“可重复性危机”，在试图重现先前发表的结果时，大量工作报告失败。我们通过对来自ReScience C和204个代码库的142个复制研究的荟萃分析来评估可重复性失败的来源。我们发现缺少实验细节，如超参数是不可重复性的潜在原因。我们通过实验证明了不同超参数选择策略的偏差，并得出结论，统一框架下的整合工件有助于支持再现性。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

Reproducibility Requires Consolidated Artifacts

Machine learning is facing a ‘reproducibility crisis’ where a significant number of works report failures when attempting to reproduce previously published results. We evaluate the sources of reproducibility failures using a meta-analysis of 142 replication studies from ReScience C and 204 code repositories. We find that missing experiment details such as hyperparameters are potential causes of unreproducibility. We experimentally show the bias of different hyperparameter selection strategies and conclude that consolidated artifacts with a unified framework can help support reproducibility.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

2023 IEEE/ACM 2nd International Conference on AI Engineering – Software Engineering for AI (CAIN)

自引率

0.00%

发文量

期刊最新文献

safe.trAIn – Engineering and Assurance of a Driverless Regional Train Extensible Modeling Framework for Reliable Machine Learning System Analysis Maintaining and Monitoring AIOps Models Against Concept Drift Conceptualising Software Development Lifecycle for Engineering AI Planning Systems Reproducibility Requires Consolidated Artifacts