Application and research of massive big data storage system based on HBase
Pan Zhengjun, Lianfen Zhao · 2018
Because HBase only supports queries based on primary keys. When users don't know the primary keys and query data without primary keys, they can only get data through full table scanning. This way result in high cost and low efficiency, which can't meet real-time query requirements. In this paper, for the non primary key table to establish the secondary indexes, and then through the secondary indexes query the RowKey, by the RowKey required value and enhance the query efficiency. Through the analysis of the experimental results, the improved Hbase storage system has greatly improved the performance of big data insertion and query.