跳到主要导航 跳到搜索 跳到主要内容

Correlation based file prefetching approach for Hadoop

  • Bo Dong
  • , Xiao Zhong
  • , Qinghua Zheng
  • , Lirong Jian
  • , Jian Liu
  • , Jie Qiu
  • , Ying Li

科研成果: 书/报告/会议事项章节会议稿件同行评审

21 引用 (Scopus)

摘要

Hadoop Distributed File System (HDFS) has been widely adopted to support Internet applications because of its reliable, scalable and low-cost storage capability. BlueSky, one of the most popular e-Learning resource sharing systems in China, is utilizing HDFS to store massive courseware. However, due to the inefficient access mechanism of HDFS, access latency of reading files from HDFS significantly impacts the performance of processing user requests. This paper introduces a two-level correlation based file prefetching approach, taking the characteristics of HDFS into consideration, to improve performance by reducing access latency. Four placement patterns to store prefetched data are presented, with policies to achieve trade-off between performance and efficiency of HDFS prefetching. Moreover, a dynamic replica selection algorithm is investigated to improve the efficiency of HDFS prefetching. The proposed prefetching approach has been implemented in BlueSky, and experimental results prove that correlation based file prefetching can significantly reduce access latency therefore improve performance of Hadoop-based Internet applications.

源语言英语
主期刊名Proceedings - 2nd IEEE International Conference on Cloud Computing Technology and Science, CloudCom 2010
出版商IEEE Computer Society
41-48
页数8
ISBN(印刷版)9780769543024
DOI
出版状态已出版 - 2010
活动2nd IEEE International Conference on Cloud Computing Technology and Science, CloudCom 2010 - Indianapolis, IN, 美国
期限: 30 11月 20103 12月 2010

出版系列

姓名Proceedings - 2nd IEEE International Conference on Cloud Computing Technology and Science, CloudCom 2010

会议

会议2nd IEEE International Conference on Cloud Computing Technology and Science, CloudCom 2010
国家/地区美国
Indianapolis, IN
时期30/11/103/12/10

学术指纹

探究 'Correlation based file prefetching approach for Hadoop' 的科研主题。它们共同构成独一无二的指纹。

引用此