Autopipelining for data stream processing

Tang, Y.; Gedik, B.

Autopipelining for data stream processing

dc.citation.epage	2354	en_US
dc.citation.issueNumber	12	en_US
dc.citation.spage	2344	en_US
dc.citation.volumeNumber	24	en_US
dc.contributor.author	Tang, Y.	en_US
dc.contributor.author	Gedik, B.	en_US
dc.date.accessioned	2016-02-08T09:33:43Z
dc.date.available	2016-02-08T09:33:43Z
dc.date.issued	2013	en_US
dc.department	Department of Computer Engineering	en_US
dc.description.abstract	Stream processing applications use online analytics to ingest high-rate data sources, process them on-the-fly, and generate live results in a timely manner. The data flow graph representation of these applications facilitates the specification of stream computing tasks with ease, and also lends itself to possible runtime exploitation of parallelization on multicore processors. While the data flow graphs naturally contain a rich set of parallelization opportunities, exploiting them is challenging due to the combinatorial number of possible configurations. Furthermore, the best configuration is dynamic in nature; it can differ across multiple runs of the application, and even during different phases of the same run. In this paper, we propose an autopipelining solution that can take advantage of multicore processors to improve throughput of streaming applications, in an effective and transparent way. The solution is effective in the sense that it provides good utilization of resources by dynamically finding and exploiting sources of pipeline parallelism in streaming applications. It is transparent in the sense that it does not require any hints from the application developers. As a part of our solution, we describe a light-weight runtime profiling scheme to learn resource usage of operators comprising the application, an optimization algorithm to locate best places in the data flow graph to explore additional parallelism, and an adaptive control scheme to find the right level of parallelism. We have implemented our solution in an industrial-strength stream processing system. Our experimental evaluation based on microbenchmarks, synthetic workloads, as well as real-world applications confirms that our design is effective in optimizing the throughput of stream processing applications without requiring any changes to the application code. © 1990-2012 IEEE.	en_US
dc.identifier.doi	10.1109/TPDS.2012.333	en_US
dc.identifier.issn	1045-9219	en_US
dc.identifier.uri	http://hdl.handle.net/11693/20712	en_US
dc.language.iso	English	en_US
dc.publisher	Institute of Electrical and Electronics Engineers	en_US
dc.relation.isversionof	http://dx.doi.org/10.1109/TPDS.2012.333	en_US
dc.source.title	IEEE Transactions on Parallel and Distributed Systems	en_US
dc.subject	Autopipelining	en_US
dc.subject	Parallelization	en_US
dc.subject	Stream processing	en_US
dc.subject	Adaptive control schemes	en_US
dc.subject	Experimental evaluation	en_US
dc.subject	Optimization algorithms	en_US
dc.title	Autopipelining for data stream processing	en_US
dc.type	Article	en_US

Files

Original bundle

Now showing 1 - 1 of 1

Name:: Autopipelining for data stream processing.pdf
Size:: 878.29 KB
Format:: Adobe Portable Document Format
Description:: Full printable version

Download

Collections

Scholarly Publications - Computer Engineering