摘要
Data pre-deployment in the HDFS (Hadoop distributed file systems) is more complicated than that in traditional file systems. There are many key issues need to be addressed, such as determining the target location of the data prefetching, the amount of data to be prefetched, the balance between data prefetching services and normal data accesses. Aiming to solve these problems, we employ the characteristics of digital ocean information service flows and propose a deployment scheme which combines input data prefetching with output data oriented storage strategies. The method achieves the parallelism of data preparation and data processing, thereby massively reducing I/O time cost of digital ocean cloud computing platforms when processing multi-source information synergistic tasks. The experimental results show that the scheme has a higher degree of parallelism than traditional Hadoop mechanisms, shortens the waiting time of a running service node, and significantly reduces data access conflicts.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 82-92 |
| 页数 | 11 |
| 期刊 | Acta Oceanologica Sinica |
| 卷 | 33 |
| 期 | 9 |
| DOI | |
| 出版状态 | 已出版 - 9月 2014 |
| 已对外发布 | 是 |
学术指纹
探究 'Research on data pre-deployment in information service flow of digital ocean cloud computing' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver