Optimization Factor Analysis of Large-Scale Join Queries on Different Platforms

Chao Yang,Qian Wang,Qing Yang,Huibing Zhang,Jingwei Zhang,Ya Zhou

Optimization Factor Analysis of Large-Scale Join Queries on Different Platforms

2017

Popular big data computing platforms, such as Spark, provide new computing paradigm for traditional database operations, such as queries. Except for the management ability of large-scale data, big data platforms earn the reputation for their simple programming interface and good performance of scaling out. But traditional databases have intrinsic optimization mechanisms for fundamental operators, which supports efficient and flexible data processing. It is very valuable to give a comprehensive view of these two kinds of platforms on data processing performance. In this paper, we focus on join operation, a primary and frequently used operator for both databases and big data analysis, design and conduct extensive experiments to test the performance of the two classic platforms under unified datasets and hardware, which will disclose the performance influence on computing schema, storage media, etc. Based on the experimental analysis, we also put forwards our advice on computing platform onsideration for different application scenarios.

Keywords:

Correction
Source
Cite
Save
Machine Reading By IdeaReader

References

Citations