Skip to Main content Skip to Navigation
Journal articles

Database System Support of Simulation Data

Hermano Lustosa 1 Fabio Porto 1 Pablo Blanco 1 Patrick Valduriez 2, 3 
3 ZENITH - Scientific Data Management
LIRMM - Laboratoire d'Informatique de Robotique et de Microélectronique de Montpellier, CRISAM - Inria Sophia Antipolis - Méditerranée
Abstract : Supported by increasingly efficient HPC infrastructure , numerical simulations are rapidly expanding to fields such as oil and gas, medicine and meteorology. As simulations become more precise and cover longer periods of time, they may produce files with terabytes of data that need to be efficiently analyzed. In this paper, we investigate techniques for managing such data using an array DBMS. We take advantage of multidimensional arrays that nicely models the dimensions and variables used in numerical simulations. However , a naive approach to map simulation data files may lead to sparse arrays, impacting query response time, in particular, when the simulation uses irregular meshes to model its physical domain. We propose efficient techniques to map coordinate values in numerical simulations to evenly distributed cells in array chunks with the use of equi-depth his-tograms and space-filling curves. We implemented our techniques in SciDB and, through experiments over real-world data, compared them with two other approaches: row-store and column-store DBMS. The results indicate that multidi-mensional arrays and column-stores are much faster than a traditional row-store system for queries over a larger amount of simulation data. They also help identifying the scenarios where array DBMSs are most efficient, and those where they are outperformed by column-stores.
Document type :
Journal articles
Complete list of metadata

Cited literature [19 references]  Display  Hide  Download
Contributor : Patrick Valduriez Connect in order to contact the contributor
Submitted on : Sunday, September 11, 2016 - 4:21:12 PM
Last modification on : Friday, August 5, 2022 - 3:03:28 PM
Long-term archiving on: : Monday, December 12, 2016 - 12:16:36 PM


Files produced by the author(s)


  • HAL Id : lirmm-01363738, version 1


Hermano Lustosa, Fabio Porto, Pablo Blanco, Patrick Valduriez. Database System Support of Simulation Data. Proceedings of the VLDB Endowment (PVLDB), VLDB Endowment, 2016, 9 (13), pp.1329-1340. ⟨lirmm-01363738⟩



Record views


Files downloads