###
Journal of Software:2012.23(6):1561-1577

XML 数据流上Top-K 关键字查询处理
黎玲利,王宏志,高宏,李建中
(哈尔滨工业大学 计算机科学与技术学院,黑龙江 哈尔滨 150001)
Efficient Top-K Keyword Search on XML Streams
LI Ling-Li,WANG Hong-Zhi,GAO Hong,LI Jian-Zhong
(School of Computer Science and Technology, Harbin Institute of Technology, Harbin 150001, China)
Abstract
Chart / table
Reference
Similar Articles
Article :Browse 3169   Download 3360
Received:April 28, 2010    Revised:September 02, 2011
> 中文摘要: 利用关键字可以在模式未知的情况下对XML 数据进行查询.在当前的XML数据流上的关键字查询处理中,打分函数往往不能都满足各种用户不同的需求.提出了一种基于skyline 的XML 数据流上的Top-K 关键字查询.对于这种查询,不需要考虑影响结果与查询相关性的复杂因素,只需利用skyline 挑选与查询最相关的结果.提出了两种XML 数据流上的有效的基于skyline 的Top-K 关键查询处理算法,包括对单查询和多查询的处理算法.通过扩展实验对两种算法的有效性和可扩展性进行了验证.经过实验验证,所提出的查询处理算法的效率几乎不受关键字个数、查询结果数量、查询数量等参数的影响,运行时间和文档大小大致呈线性关系.
中文关键词: XML  数据流  关键字查询  Top-K  skyline
Abstract:Keywords are suitable for query XML streams without schema information. In current forms of keywords search on XML streams and rank functions do not always represent users' intensions. This paper addresses this problem in another aspect. In this paper, the skyline Top-K keyword queries, a novel kind of keyword queries on XML streams, are presented. For such queries, skyline is used to choose results on XML streams without considering the complicated factors influencing the relevance to queries. With skyline query processing techniques, two techniques, are presented to process skyline Top-K keyword single queries and multi-queries on XML streams efficiently. Extensive experiments are performed to verify the effectiveness and efficiency of these techniques presented in this paper. According to the experimental results, the algorithms are not sensitive to the parameters such as the number of keywords, the number of results, the number of queries, and the runtime is approximately linear to the size of document.
keywords: XML  streams  keyword search  Top-K  skyline
文章编号:     中图分类号:    文献标志码:
基金项目:国家自然科学基金(61003046, 61111130189); 国家重点基础研究发展计划(973)(2012CB316200); 高等学校博士学科点专项科研基金(20102302120054) 国家自然科学基金(61003046, 61111130189); 国家重点基础研究发展计划(973)(2012CB316200); 高等学校博士学科点专项科研基金(20102302120054)
Foundation items:
Reference text:

黎玲利,王宏志,高宏,李建中.XML 数据流上Top-K 关键字查询处理.软件学报,2012,23(6):1561-1577

LI Ling-Li,WANG Hong-Zhi,GAO Hong,LI Jian-Zhong.Efficient Top-K Keyword Search on XML Streams.Journal of Software,2012,23(6):1561-1577