猪与蜂箱用例分析

International journal of database theory and application Pub Date : 2016-12-31 DOI:10.14257/IJDTA.2016.9.12.24

D. Kendal, Oded Koren, N. Perel

{"title":"猪与蜂箱用例分析","authors":"D. Kendal, Oded Koren, N. Perel","doi":"10.14257/IJDTA.2016.9.12.24","DOIUrl":null,"url":null,"abstract":"Corporations are changing their practices to data-driven big data initiatives, as big data analytics has provided companies with the ability to grow their businesses and increase competition. As the importance of data analytics grew, so accordingly did the size of the data to analyze, thus demanding a more powerful data platform. This paper shows a case study of two High Level Query Languages that are constructed on top of Hadoop MapReduce; Pig and Hive. By creating a query in each query language, both resulting in an identical output, and by running each query 30 times on 2 different sized files (120 runs total), this comparison provides a statistically significant conclusion.","PeriodicalId":13926,"journal":{"name":"International journal of database theory and application","volume":"12 1 1","pages":"267-276"},"PeriodicalIF":0.0000,"publicationDate":"2016-12-31","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"4","resultStr":"{\"title\":\"Pig Vs. Hive Use Case Analysis\",\"authors\":\"D. Kendal, Oded Koren, N. Perel\",\"doi\":\"10.14257/IJDTA.2016.9.12.24\",\"DOIUrl\":null,\"url\":null,\"abstract\":\"Corporations are changing their practices to data-driven big data initiatives, as big data analytics has provided companies with the ability to grow their businesses and increase competition. As the importance of data analytics grew, so accordingly did the size of the data to analyze, thus demanding a more powerful data platform. This paper shows a case study of two High Level Query Languages that are constructed on top of Hadoop MapReduce; Pig and Hive. By creating a query in each query language, both resulting in an identical output, and by running each query 30 times on 2 different sized files (120 runs total), this comparison provides a statistically significant conclusion.\",\"PeriodicalId\":13926,\"journal\":{\"name\":\"International journal of database theory and application\",\"volume\":\"12 1 1\",\"pages\":\"267-276\"},\"PeriodicalIF\":0.0000,\"publicationDate\":\"2016-12-31\",\"publicationTypes\":\"Journal Article\",\"fieldsOfStudy\":null,\"isOpenAccess\":false,\"openAccessPdf\":\"\",\"citationCount\":\"4\",\"resultStr\":null,\"platform\":\"Semanticscholar\",\"paperid\":null,\"PeriodicalName\":\"International journal of database theory and application\",\"FirstCategoryId\":\"1085\",\"ListUrlMain\":\"https://doi.org/10.14257/IJDTA.2016.9.12.24\",\"RegionNum\":0,\"RegionCategory\":null,\"ArticlePicture\":[],\"TitleCN\":null,\"AbstractTextCN\":null,\"PMCID\":null,\"EPubDate\":\"\",\"PubModel\":\"\",\"JCR\":\"\",\"JCRName\":\"\",\"Score\":null,\"Total\":0}","platform":"Semanticscholar","paperid":null,"PeriodicalName":"International journal of database theory and application","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.14257/IJDTA.2016.9.12.24","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"","JCRName":"","Score":null,"Total":0}

引用次数: 4

摘要

随着大数据分析为企业提供了发展业务和增加竞争的能力，企业正在将其实践转变为数据驱动的大数据计划。随着数据分析重要性的提高，需要分析的数据量也随之增加，因此需要一个更强大的数据平台。本文展示了基于Hadoop MapReduce构建的两种高级查询语言的案例研究;小猪和蜂巢。通过在每种查询语言中创建查询，两种查询语言都会产生相同的输出，并且在两个不同大小的文件上运行每个查询30次(总共运行120次)，这种比较提供了统计上显著的结论。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

Pig Vs. Hive Use Case Analysis

Corporations are changing their practices to data-driven big data initiatives, as big data analytics has provided companies with the ability to grow their businesses and increase competition. As the importance of data analytics grew, so accordingly did the size of the data to analyze, thus demanding a more powerful data platform. This paper shows a case study of two High Level Query Languages that are constructed on top of Hadoop MapReduce; Pig and Hive. By creating a query in each query language, both resulting in an identical output, and by running each query 30 times on 2 different sized files (120 runs total), this comparison provides a statistically significant conclusion.

求助全文

通过发布文献求助，成功后即可免费获取论文全文。去求助

来源期刊

International journal of database theory and application

自引率

0.00%

发文量