Français Anglais
Accueil Annuaire Plan du site
Accueil > Evenements > Séminaires
Séminaire d'équipe(s) BD
Processing XML Queries and Updates on Map/Reduce Clusters
Dario Colazzo

19 April 2013, 14h30 - 19 April 2013, 16h00
Salle/Bat : 435/PCRI-N
Contact : dario.colazzo@lri.fr

Activités de recherche :

Résumé :
Very large XML documents are generated and processed in several contexts, in particular in those involving scientific data and logs. In order to process such large documents we have designed and implemented techniques based on data partitioning for the evaluation of XQuery queries and updates on Map/Reduce clusters.

The proposed technique applies when queries and updates are iterative, i.e., they iterate the same query/update operations on a sequence of subtrees of the input document. We have developed schema-less, static analysis techniques to i) recognize iterative queries/updates, and ii) extract path information to be used for data partitioning purposes. Our system exploits both dynamic and static data partitioning to distribute the processing load among the machines of a Map/Reduce cluster. To boost the I/O performance across the distributed file system, our system uses EXI compression at each stage of the computation, from data partitioning to query/update execution.

After an introduction to the main techniques behind our system, a demonstration will show its abilities in dealing with complex workloads and large documents.

Pour en savoir plus :
Séminaires
Heterogeneous Treatment Effects Estimation: When M
Raisonnement automatique
Thursday 02 June 2022 - 10h30
Salle : 2011 - DIG-Moulon
Naoufal Acharki .............................................

Witness Generation for JSON Schema
Langages et systèmes centrés données
Monday 30 May 2022 - 00h00
Salle : 455 - PCRI-N
Mohamed-Amine BAAZIZI .............................................

TUTORIAL CODALAB - Apprenez à organiser un challen
Wednesday 13 April 2022 - 00h00
Salle : 1 - DIG-Moulon
Adrien Pavao .............................................

Generative Neural Networks for Observational Causa
Raisonnement automatique
Thursday 07 April 2022 - 10h30
Salle : 2011 - DIG-Moulon
Diviyan Kalainathan .............................................

Datamining in Epi- and Phylogenetics
Tuesday 15 March 2022 - 11h00
Salle : 455 - PCRI-N
Thomas Haschka .............................................