数据书:一个标准化的框架，用于数据科学项目中算法设计的动态文档

IASSIST quarterly Pub Date : 2021-09-26 DOI:10.29173/iq989

Anna Nesvijevskaia

{"title":"数据书:一个标准化的框架，用于数据科学项目中算法设计的动态文档","authors":"Anna Nesvijevskaia","doi":"10.29173/iq989","DOIUrl":null,"url":null,"abstract":"This paper proposes a standard documentation framework for Data Science projects, called Databook. It is a result of five years of action-research on multiple projects in several sectors of activity in France, and of a confrontation of standard theoretical Data Science processes, such as CRISP_DM, with the reality of the field. As a vector for knowledge sharing and capitalisation, the Databook has been identified as one of the main facilitators of Human Data Mediation. Transformed into an operational prototype of simple and minimalist documentation, it has since been tested then on about a hundred Data Science projects, has proven its benefits for the internal and external efficiency of Data Science projects, and can be turned into a more ambitious standard framework for data patrimony valorisation and data quality governance.","PeriodicalId":84870,"journal":{"name":"IASSIST quarterly","volume":" ","pages":""},"PeriodicalIF":0.0000,"publicationDate":"2021-09-26","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"1","resultStr":"{\"title\":\"DATABOOK : a standardised framework for dynamic documentation of algorithm design during Data Science projects\",\"authors\":\"Anna Nesvijevskaia\",\"doi\":\"10.29173/iq989\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"This paper proposes a standard documentation framework for Data Science projects, called Databook. It is a result of five years of action-research on multiple projects in several sectors of activity in France, and of a confrontation of standard theoretical Data Science processes, such as CRISP_DM, with the reality of the field. As a vector for knowledge sharing and capitalisation, the Databook has been identified as one of the main facilitators of Human Data Mediation. Transformed into an operational prototype of simple and minimalist documentation, it has since been tested then on about a hundred Data Science projects, has proven its benefits for the internal and external efficiency of Data Science projects, and can be turned into a more ambitious standard framework for data patrimony valorisation and data quality governance.\",\"PeriodicalId\":84870,\"journal\":{\"name\":\"IASSIST quarterly\",\"volume\":\" \",\"pages\":\"\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2021-09-26\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"1\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"IASSIST quarterly\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.29173/iq989\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"IASSIST quarterly","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.29173/iq989","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 1

摘要

本文为数据科学项目提出了一个标准的文档框架，称为Databook。这是对法国几个活动部门的多个项目进行了五年行动研究的结果，也是CRISP_DM等标准理论数据科学过程与该领域现实相对抗的结果。作为知识共享和资本化的载体，数据手册已被确定为人类数据中介的主要推动者之一。它被转变为一个简单、最低限度的文档的操作原型，此后在大约100个数据科学项目中进行了测试，证明了它对数据科学项目的内部和外部效率的好处，并可以转化为一个更雄心勃勃的数据遗产定价和数据质量治理标准框架。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

DATABOOK : a standardised framework for dynamic documentation of algorithm design during Data Science projects

This paper proposes a standard documentation framework for Data Science projects, called Databook. It is a result of five years of action-research on multiple projects in several sectors of activity in France, and of a confrontation of standard theoretical Data Science processes, such as CRISP_DM, with the reality of the field. As a vector for knowledge sharing and capitalisation, the Databook has been identified as one of the main facilitators of Human Data Mediation. Transformed into an operational prototype of simple and minimalist documentation, it has since been tested then on about a hundred Data Science projects, has proven its benefits for the internal and external efficiency of Data Science projects, and can be turned into a more ambitious standard framework for data patrimony valorisation and data quality governance.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

IASSIST quarterly

自引率

0.00%

发文量

期刊最新文献

Security and preservation of election data in Nigeria in the fourth industrial revolution Knowledge and perception of librarians towards cloud-based technology in academic libraries in southwest Nigeria Much new research, and advances for the IQ Data protection and right to privacy legislation in Kenya Guest editors’ notes