RailwayDB: adaptive storage of interaction graphs

Soulé R.; Gedik, B.

RailwayDB: adaptive storage of interaction graphs

Files

RailwayDB adaptive storage of interaction graphs.pdf (1.45 MB)

Date

2016

Authors

Soulé R.

Gedik, B.

BUIR Usage Stats

4
views

20
downloads

Citation Stats

Abstract

We are living in an ever more connected world, where data recording the interactions between people, software systems, and the physical world is becoming increasingly prevalent. These data often take the form of a temporally evolving graph, where entities are the vertices and the interactions between them are the edges. We call such graphs interaction graphs. Various domains, including telecommunications, transportation, and social media, depend on analytics performed on interaction graphs. The ability to efficiently support historical analysis over interaction graphs requires effective solutions for the problem of data layout on disk. This paper presents an adaptive disk layout called the railway layout for optimizing disk block storage for interaction graphs. The key idea is to divide blocks into one or more sub-blocks. Each sub-block contains the entire graph structure, but only a subset of the attributes. This improves query I/O, at the cost of increased storage overhead. We introduce optimal integer linear program (ILP) formulations for partitioning disk blocks into sub-blocks with overlapping and nonoverlapping attributes. Additionally, we present greedy heuristics that can scale better compared to the ILP alternatives, yet achieve close to optimal query I/O. We provide an implementation of the railway layout as part of RailwayDB—an open-source graph database we have developed. To demonstrate the benefits of the railway layout, we provide an extensive experimental evaluation, including model-based as well as empirical results comparing our approach to baseline alternatives.

Source Title

The VLDB Journal

Publisher

Association for Computing Machinery

Keywords

Adaptive storage, I/O optimization, Interaction graphs, Graphic methods, Integer programming, Open source software, Optimization, Query processing, Railroads, Transportation, Effective solution, Experimental evaluation, Greedy heuristics, Historical analysis, Integer linear programs, Storage overhead, Digital storage

Permalink

http://hdl.handle.net/11693/36933

Published Version (Please cite this version)

http://dx.doi.org/10.1007/s00778-015-0407-0

Collections

Scholarly Publications - Computer Engineering

Language

English

Type

Article

Full item page

RailwayDB: adaptive storage of interaction graphs

Files

Date

Authors

Editor(s)

Advisor

Supervisor

Co-Advisor

Co-Supervisor

Instructor

BUIR Usage Stats

Citation Stats

Series

Abstract

Source Title

Publisher

Course

Other identifiers

Book Title

Keywords

Degree Discipline

Degree Level

Degree Name

Citation

Permalink

Published Version (Please cite this version)

Collections

Language

Type

RailwayDB: adaptive storage of interaction graphs

Files

Date

Authors

Editor(s)

Advisor

Supervisor

Co-Advisor

Co-Supervisor

Instructor

BUIR Usage Stats

Citation Stats

Share

Series

Abstract

Source Title

Publisher

Course

Other identifiers

Book Title

Keywords

Degree Discipline

Degree Level

Degree Name

Citation

Permalink

Published Version (Please cite this version)

Collections

Language

Type