Multi-query processing of XML data streams on multicore

In order to solve the problem of storage and query for massive XML data, a method of efficient storage and parallel query for a massive volume of XML data with Hadoop is proposed. This method can store massive XML data in Hadoop and the massive XML data is divided into many XML data blocks and loaded on HDFS. The parallel query method of massive XML data is proposed, which uses parallel XPath queries based on multiple predicate selection, and the results of parallel query can satisfy the requirement of query given by the user. In this chapter, the map logic algorithm and the reduce logic algorithm based on parallel XPath queries based using MapReduce programming model are proposed, and the parallel query processing of massive XML data is realized. In addition, the method of MapReduce query optimization based on multiple predicate selection is proposed to reduce the data transfer volume of the system and improve the performance of the system. Finally, the effectiveness of the proposed method is verified by experiment.

Download Full-text

Using a Pipeline Approach to Build Data Cube for Large XML Data Streams

Database Systems for Advanced Applications - Lecture Notes in Computer Science ◽

10.1007/978-3-642-40270-8_5 ◽

2013 ◽

pp. 59-73 ◽

Cited By ~ 1

Author(s):

Hao Gui ◽

Mark Roantree

Keyword(s):

Data Streams ◽

Data Cube ◽

Xml Data

Download Full-text

Approximate OLAP Query Processing over Uncertain and Imprecise Multidimensional Data Streams

Lecture Notes in Computer Science - Database and Expert Systems Applications ◽

10.1007/978-3-642-40173-2_15 ◽

2013 ◽

pp. 156-173 ◽

Cited By ~ 4

Author(s):

Alfredo Cuzzocrea

Keyword(s):

Query Processing ◽

Data Streams ◽

Multidimensional Data

Download Full-text

XML Data Integration

Advanced Applications and Structures in XML Processing ◽

10.4018/978-1-61520-727-5.ch015 ◽

2010 ◽

pp. 333-360 ◽

Cited By ~ 1

Author(s):

Yan Qi ◽

Huiping Cao ◽

K. Selçuk Candan ◽

Maria Luisa Sapino

Keyword(s):

Data Integration ◽

Query Processing ◽

Multiple Sources ◽

Data Types ◽

Xml Data ◽

Data Schema ◽

Integration Data ◽

Feedback Techniques ◽

Different Sources ◽

Xml Data Integration

In XML Data Integration, data/metadata merging and query processing are indispensable. Specifically, merging integrates multiple disparate (heterogeneous and autonomous) input data sources together for further usage, while query processing is one main reason why the data need to be integrated in the first place. Besides, when supported with appropriate user feedback techniques, queries can also provide contexts in which conflicts among the input sources can be interpreted and resolved. The flexibility of XML structure provides opportunities for alleviating some of the difficulties that other less flexible data types face in the presence of uncertainty; yet, this flexibility also introduces new challenges in merging multiple sources and query processing over integrated data. In this chapter, the authors discuss two alternative ways XML data/schema can be integrated: conflict-eliminating (where the result is cleaned from any conflicts that the different sources might have with each other) and conflict-preserving (where the resulting XML data or XML schema captures the alternative interpretations of the data). They also present techniques for query processing over integrated, possibly imprecise, XML data, and cover strategies that can be used for resolving underlying conflicts.

Download Full-text

Continuous and Progressive XML Query Processing and its Applications

Open and Novel Issues in XML Database Applications ◽

10.4018/978-1-60566-308-1.ch009 ◽

2010 ◽

pp. 181-197

Author(s):

Stéphane Bressan ◽

Wee Hyong Tok ◽

Xue Zhao

Keyword(s):

Query Processing ◽

Data Streams ◽

Ad Hoc ◽

Data Representation ◽

Future Trends ◽

Xml Query Processing ◽

Processing Techniques ◽

Xml Technologies ◽

State Of Art ◽

Open Issues

Since XML technologies have become a standard for data representation, a great amount of discussion has been generated by the persisting open issues and their possible solutions. In this chapter, the authors consider the design space for XML query processing techniques that can handle ad hoc and continuous XPath or XQuery queries over XML data streams. This chapter presents the state-of-art techniques in continuous and progressive XML query processing. They also discuss several open issues and future trends.

Download Full-text