Title page for ETD etd-04082008-164938

Type of Document Dissertation
Author Curtis-Maury, Matthew
Author's Email Address mfcurt@vt.edu
URN etd-04082008-164938
Title Improving the Efficiency of Parallel Applications on Multithreaded and Multicore Systems
Degree PhD
Department Computer Science
Advisory Committee
Advisor Name Title
Nikolopoulos, Dimitrios S. Committee Chair
Cameron, Kirk W. Committee Member
de Supinski, Bronis R. Committee Member
Feng, Wu-Chun Committee Member
Ribbens, Calvin J. Committee Member
  • power-aware computing
  • high-performance computing
  • performance prediction
  • multicore processors
  • runtime adaptation
  • concurrency throttling
Date of Defense 2008-03-19
Availability unrestricted
The scalability of parallel applications executing on multithreaded and multicore multiprocessors is often quite limited due to large degrees of contention over shared resources on these systems. In fact, negative scalability frequently occurs such that a non-negligable performance loss is observed through the use of more processors and cores. In this dissertation, we present a prediction model for identifying efficient operating points of concurrency in multithreaded scientific applications in terms of both performance as a primary objective and power secondarily. We also present a runtime system that uses live analysis of hardware event rates through the prediction model to optimize applications dynamically. We discuss a dynamic, phase-aware performance prediction model (DPAPP), which combines statistical learning techniques, including multivariate linear regression and artificial neural networks, with runtime analysis of data collected from hardware event counters to locate optimal operating points of concurrency. We find that the scalability model achieves accuracy approaching 95%, sufficiently accurate to identify improved concurrency levels and thread placements from within real parallel scientific applications.

Using DPAPP, we develop a prediction-driven runtime optimization scheme, called ACTOR, which throttles concurrency so that power consumption can be reduced and performance can be set at the knee of the scalability curve of each parallel execution phase in an application. ACTOR successfully identifies and exploits program phases where limited scalability results in a performance loss through the use of more processing elements, providing simultaneous reductions in execution time by 5%-18% and power consumption by 0%-11% across a variety of parallel applications and architectures. Further, we extend DPAPP and ACTOR to include support for runtime adaptation of DVFS, allowing for the synergistic exploitation of concurrency throttling and DVFS from within a single, autonomically-acting library, providing improved energy-efficiency compared to either approach in isolation.

  Filename       Size       Approximate Download Time (Hours:Minutes:Seconds) 
 28.8 Modem   56K Modem   ISDN (64 Kb)   ISDN (128 Kb)   Higher-speed Access 
  dissertation.pdf 1.67 Mb 00:07:42 00:03:57 00:03:28 00:01:44 00:00:08

Browse All Available ETDs by ( Author | Department )

dla home
etds imagebase journals news ereserve special collections
virgnia tech home contact dla university libraries

If you have questions or technical problems, please Contact DLA.