Xu Liu

chapter

DR-BW: Identifying Bandwidth Contention in NUMA Architectures with Supervised Learning

Hao Xu, Shasha Wen, Alfredo Gimenez, Todd Gamblin, more

2017 IEEE International Parallel and Distributed Processing Symposium (IPDPS) > 367 - 376

2017 IEEE International Parallel and Distributed Processing Symposium (IPDPS)

Non-Uniform Memory Access (NUMA) architectures are widely used in mainstream multi-socket computer systems to scale memory bandwidth. Without a NUMA-aware design, programs can suffer from significant performance degradation due to inter-socket bandwidth contention. However, identifying bandwidth contention is challenging. Existing methods measure bandwidth consumption. However, consumption alone is...

chapter

Understanding Data Analytics Workloads on Intel(R) Xeon Phi(R)

Biwei Xie, Xu Liu, Sally A. McKee, Jianfeng Zhan, more

2016 IEEE 18th International Conference on High Performance Computing and Communications; IEEE 14th International Conference on Smart City; IEEE 2nd International Conference on Data Science and Systems (HPCC/SmartCity/DSS) > 206 - 215

2016 IEEE 18th International Conference on High Performance Computing and Communications; IEEE 14th International Conference on Smart City; IEEE 2nd International Conference on Data Science and Systems (HPCC/SmartCity/DSS)

The Intel® Xeon Phi™ is gaining popularity for high-performance computing (HPC) applications, but the performance of this many-core coprocessor with wide floating point SIMD units has yet to be explored on data analytics workloads. We construct a benchmark suite to explore the Xeon Phi™'s potential for use in data center servers. Our resulting PhiBench consists of eight representative data analytics...

chapter

StructSlim: A lightweight profiler to guide structure splitting

Probir Roy, Xu Liu

2016 IEEE/ACM International Symposium on Code Generation and Optimization (CGO) > 36 - 46

2016 IEEE/ACM International Symposium on Code Generation and Optimization (CGO)

Memory access latency continues to be a dominant bottleneck in a large class of applications on modern architectures. To optimize memory performance, it is important to utilize the locality in the memory hierarchy. Structure splitting can significantly improve memory locality. However, pinpointing inefficient code and providing insightful guidance for structure splitting is challenging. Existing tools...

chapter

Runtime Value Numbering: A Profiling Technique to Pinpoint Redundant Computations

Shasha Wen, Xu Liu, Milind Chabbi

2015 International Conference on Parallel Architecture and Compilation (PACT) > 254 - 265

2015 International Conference on Parallel Architecture and Compilation (PACT)

Redundant computations can severely degrade performance in HPC applications. Redundant computations arise due to various causes such as developers' inattention to performance, inappropriate choice of algorithms, and inefficient code generation, among others. Aliasing, limited optimization scopes, and insensitivity to input and execution contexts act as severe deterrents to static program analysis...

chapter

ScaAnalyzer: a tool to identify memory scalability bottlenecks in parallel programs

Xu Liu, Bo Wu

SC15: International Conference for High Performance Computing, Networking, Storage and Analysis > 1 - 12

SC15: International Conference for High Performance Computing, Networking, Storage and Analysis

It is difficult to scale parallel programs in a system that employs a large number of cores. To identify scalability bottlenecks, existing tools principally pinpoint poor thread synchronization strategies or unnecessary data communication. Memory subsystem is one of the key contributors to poor parallel scaling in multicore machines. State-of-the-art tools, however, either lack sophisticated capabilities...

INFONA - science communication portal

Search results for: Xu Liu

DR-BW: Identifying Bandwidth Contention in NUMA Architectures with Supervised Learning

Understanding Data Analytics Workloads on Intel(R) Xeon Phi(R)

StructSlim: A lightweight profiler to guide structure splitting

Runtime Value Numbering: A Profiling Technique to Pinpoint Redundant Computations

ScaAnalyzer: a tool to identify memory scalability bottlenecks in parallel programs

Filter options

Publication date

Keywords

INFONA - science communication portal

Search results for: Xu Liu

DR-BW: Identifying Bandwidth Contention in NUMA Architectures with Supervised Learning

Understanding Data Analytics Workloads on Intel(R) Xeon Phi(R)

StructSlim: A lightweight profiler to guide structure splitting

Runtime Value Numbering: A Profiling Technique to Pinpoint Redundant Computations

ScaAnalyzer: a tool to identify memory scalability bottlenecks in parallel programs

Add recipient

Sending message cancelled

Are you sure you want to cancel sending this message?

Send message

Filter options

Publication date

Date range setting

Set the date range to filter the displayed results. You can set a starting date, ending date or both. You can enter the dates manually or choose them from the calendar.

Keywords

Reporting an error / abuse

Sending the report failed

Accessibility options