Featured article Archives - Page 15 of 21

Gary Grider Sees End Of Parallel File Systems Before He Reaches Retirement Age

November 28, 2014 by Rob Farber Leave a Comment

HPC storage luminary Gary Grider, noted in his talk at the SC14 Seagate HPC User Group meeting, "The Future of Supercomputing" that he funded the development of the Lustre file-system (which can stream a data at a TB/s), and he believes it is possible we will see the end of such file-systems before he reaches retirement age. Instead, Gary sees object based storage similar to … [Read more...]

Seismic Changes in the Animation Industry

November 25, 2014 by Rob Farber Leave a Comment

TechEnablement caught up with DreamWorks CTO Lincoln Wallen after his plenary invited talk at SC14. We had the opportunity to ask Lincoln about our observation of seismic changes happening within the animation industry as technology enables small and mid-sized businesses to create studio quality animated characters for television, augmented reality, and eventually for movies … [Read more...]

Inside The IBM NVIDIA Volta plus NVlink 2017 Delivery for $325M DOE Procurements

November 24, 2014 by Rob Farber Leave a Comment

The U.S. Department of Energy unveiled plans to build two GPU-accelerated leadership class supercomputers (Summit at ORNL and Sierra at LLNL) in a combined $325M USD procurement to be installed in 2017 that will be based on next-generation IBM POWER servers incorporating NVIDIA® Volta GPU accelerators plus NVLink™ high-speed GPU interconnect technology. The announcement by U.S. … [Read more...]

Game Changers at SC14 – Obsidian and Landsvirkjun

November 21, 2014 by Rob Farber Leave a Comment

TechEnablement has identified two game-changing "must watch" SC14 attendees: (1) Obsidian Strategics and (2) Iceland's Landsvirkjun National Power Company. In combination, these two organizations have the potential to convert the worlds hunger for HPC and the Internet from climate destroying coal to renewable energy. As Jeff Goodell observed, "coal was supposed to be the engine … [Read more...]

CreativeC GPU And Intel Xeon Phi Cluster For SC14 Class Runs Mobile In Van

November 14, 2014 by Rob Farber Leave a Comment

Our all-day class at SC14 on Sunday November 16, “From ‘Hello World’ to Exascale Using x86, GPUs and Intel Xeon Phi Coprocessors” (tut106s1) received more than double our expected enrollment! Students will be able to run on both Intel Xeon Phi and GPU supercomputers at TACC via an Xsede allocation (thank you very much) and on a CreativeC supercomputer and visualization cluster … [Read more...]

Under $200 Intel Xeon Phi

November 13, 2014 by Rob Farber Leave a Comment

For a limited time Intel is selling Intel® Xeon Phi™ Coprocessor 31S1P for under $200. This offer is designed for Software developers to cost-effectively purchase systems or clusters from OEMs to modernize their code for greater levels of performance. See one of the OEMs at this link, or Intel your rep for eligibility requirements. Additionally, as part of this developer … [Read more...]

Morton Order Improves Performance

November 11, 2014 by Rob Farber Leave a Comment

Author Kerry Evans writes in his High Performance Parallelism Pearls chapter, "There are many facets to performance optimization but three issues to deal with right from the beginning are memory access, vectorization, and parallelization. Unless we can optimize these, we cannot achieve peak performance.” Specifically, this chapter examines a method of mapping multidimensional … [Read more...]

Sparse matrix-vector multiplication: parallelization and vectorization

November 10, 2014 by Rob Farber Leave a Comment

The chapter authors (Albert-Jan N. Yzelman, Dirk Roose, and Karl Meerbergen) note that, "Current hardware trends lead to an increasing width of vector units as well as to decreasing effective bandwidth-per-core. For sparse computations these two trends conflict.” For this reason they designed a usable and efficient data structure for vectorized sparse computations on … [Read more...]

Scalable Out-Of-Core Solvers On A Cluster

November 7, 2014 by Rob Farber Leave a Comment

This chapters documents the implementation of a parallel distributed memory out-of-core (OOC) solver for performing LU and Cholesky factorizations of a large dense matrix on clusters equipped with Intel Xeon Phi coprocessors. The code was ported from CUDA with high-level library routines in CUBLAS This matches well with the offload model for the coprocessor using the … [Read more...]

Heterogeneous MPI Optimization With ITAC

November 6, 2014 by Rob Farber Leave a Comment

This chapter focuses on the workload balance of MPI applications running in heterogeneous cluster environment consisting of Intel Xeon processors and Intel Xeon Phi coprocessors in a financial industry application that calculates Asian option payoffs. Three cases are considered: unbalanced symmetric MPI code, manual balancing with pre-calculated performances of the cluster … [Read more...]

« Previous Page