Hadoop中处理海量小文件的方法

微信公众号

网站二维码

首页 > 过刊浏览>2015年第24卷第11期 >157-161

Hadoop中处理海量小文件的方法增强出版
DOI:
                        
                    
作者:
                        
                        
                    
作者单位:
作者简介:
通讯作者:
中图分类号:
基金项目:2013年度科技部科技支撑计划(2013BAJ10B14-5)

Methods of Dealing With Massive Small Files in Hadoop

Author:

Affiliation:

Fund Project:

摘要

图/表

访问统计

参考文献

相似文献

引证文献

增强出版

文章评论

摘要:

针对Hadoop中提供底层存储的HDFS对处理海量小文件效率低下、严重影响性能的问题.设计了一种小文件合并、索引和提取方案,并与原始的HDFS以及HAR文件归档方案进行对比,通过一系列实验表明,本文的方案能有效减少Namenode内存占用,提高HDFS的I/O性能.

Abstract:

HDFS provides the underlying storage for Hadoop, however, the HDFS deals with massive small files inefficiently and decreases system performance seriously. To solve this problem, we designed a file merging, indexing and retrieval solution. Then through a series of experiments compared to the original HDFS and HAR solution, it can be shown that our scheme can effectively reduce the memory usage of Namenode and improve the I/O performance of HDFS.

参考文献

相似文献

引证文献

引用本文

李旭,李长云,张清清,胡淑新,周玲芳. Hadoop中处理海量小文件的方法.计算机系统应用,2015,24(11):157-161

复制

文章指标

点击次数:
下载次数:
HTML阅读次数:
引用次数:

历史

收稿日期:2015-03-08
最后修改日期:2015-05-18
录用日期:
在线发布日期: 2015-12-03
出版日期:

微信公众号

网站二维码

引用本文

分享

文章指标

历史