wPerf: Generic Off-CPU Analysis to Identify Bottleneck Waiting Events

Fang Zhou; Yifan Gan; Sixiang Ma; Yang Wang

2018 OSDI OSDI 2018

wPerf: Generic Off-CPU Analysis to Identify Bottleneck Waiting Events

Abstract

This paper tries to identify waiting events that limit the maximal throughput of a multi-threaded application. To achieve this goal, we not only need to understand an event's impact on threads waiting for this event (i.e., local impact), but also need to understand whether its impact can reach other threads that are involved in request processing (i.e., global impact). To address these challenges, wPerf computes the local impact of a waiting event with a technique called cascaded re-distribution; more importantly, wPerf builds a wait-for graph to compute whether such impact can indirectly reach other threads. By combining these two techniques, wPerf essentially tries to identify events with large impacts on all threads. We apply wPerf to a number of open-source multi-threaded applications. By following the guide of wPerf, we are able to improve their throughput by up to 4.83$\times$. The overhead of recording waiting events at runtime is about 5.1% on average.

🧭 Keyword Pioneer — off-cpu analysis

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics

Authors

Fang Zhou , Yifan Gan , Sixiang Ma , Yang Wang

Topics

Machine Learning > Application Areas > Efficient Computing

Keywords

throughput optimization performance profiling performance analysis off-cpu analysis thread contention wait-for graph bottleneck identification multi-threaded application bottleneck detection waiting event thread synchronization

Download PDF

Related papers

Arachne: Core-Aware Thread Management 2018

Adaptive Dynamic Checkpointing for Safe Efficient Intermittent Computing 2018

The FuzzyLog: A Partially Ordered Shared Log 2018

Sledgehammer: Cluster-Fueled Debugging 2018

Obladi: Oblivious Serializable Transactions in the Cloud 2018