首页 | 本学科首页   官方微博 | 高级检索  
     

基于压缩-字对齐位图的天文海量数据实时索引
引用本文:刘应波,王 锋,季凯帆,邓 辉,戴 伟,梁 波. 基于压缩-字对齐位图的天文海量数据实时索引[J]. 计算机工程与应用, 2016, 52(1): 37-41
作者姓名:刘应波  王 锋  季凯帆  邓 辉  戴 伟  梁 波
作者单位:1.中国科学院 云南天文台,昆明 6500112.昆明理工大学 云南省计算机技术应用重点实验室,昆明 6505003.中国科学院大学,北京 100049
摘    要:澄江一米新真空大型天文望远镜(NVST)当前每天最大能产生2 TB,约十多万条的观测数据。由于这些数据量巨大并具有非结构化特性,使用离线构建索引会带来巨大时间开销,传统的关系型数据库难以满足快速索引和检索需求。针对这些问题,结合数据采集流程,提出了使用基于压缩的字对齐位图索引算法来在线实时构建索引。这种方式不仅克服了离线构建索引方式时,文件访问、FITS头读取和解析FITS头等操作带来的大量额外时间消耗问题,而且有助于解决海量太阳观测数据的高效检索难题。通过实验证明了在线实时构建索引方式能够极大地降低时间开销,也表明了该方式在天文海量数据索引和检索应用中的有效性和可行性。

关 键 词:字对齐位图索引  FastBit  海量数据  大型望远镜  

Massive astronomical real-time data indexing based on compressed word-aligned bitmap
LIU Yingbo,WANG Feng,JI Kaifan,DENG Hui,DAI Wei,LIANG Bo. Massive astronomical real-time data indexing based on compressed word-aligned bitmap[J]. Computer Engineering and Applications, 2016, 52(1): 37-41
Authors:LIU Yingbo  WANG Feng  JI Kaifan  DENG Hui  DAI Wei  LIANG Bo
Affiliation:1.Yunnan Observatories, Chinese Academy of Sciences, Kunming 650011, China2.Computer Technology Application Key Laboratory of Yunnan Province, Kunming University of Science and Technology, Kunming 650500, China3.University of Chinese Academy of Sciences, Beijing 100049, China
Abstract:At present, New Vacuum Solar Telescope(NVST) is generating data more than 2 TB per day, and due to the characteristic of massive non-structure data, it is beyond traditional database systems to search efficiently for a subset of these extremely large with low latency and response time and time overtime is increasing by the way of off-line index building. Aiming at these problems, combined with the strategy of on-line real data indexing in observation data capture system, using an approach of the compressed word-aligned bitmap index for observation data indexing and querying is proposed. This approach cannot only save the time cost of accessing FITS file, reading FITS file header and parsing header keyword in off-line indexing mode, but also can solve the problem of massive solar data retrieval. Experiments show that time overhead can be reduced by using real-time data indexing, and the effectiveness and feasibility of the method are proved in the experiments.
Keywords:word-aligned bitmap  FastBit  massive data  large telescope  
本文献已被 万方数据 等数据库收录!
点击此处可从《计算机工程与应用》浏览原始摘要信息
点击此处可从《计算机工程与应用》下载全文
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号