TOP   Events & Outreach  R-CCS Cafe  The 199th R-CCS Cafe - part 2

Title

Recent topics of High Precision and Low Precision Computing in HPC

Details
Date Mon, Oct 5, 2020
Time 4:40 pm - 5 pm (5:20 pm - 5:40 pm Discussion (speakers are requested to participate; we will take a 1 - 2 min. break in the beginning) )
City Online
Place

Online seminar on BlueJeans

  • If you are not affiliated with R-CCS and would like to attend R-CCS Cafe, please email us at r-ccs-cafe[at]ml.riken.jp.
Language Presentation Language: English
Presentation Material: English
Speakers

Toshiyuki Imamura

Team Leader, Large-scale Parallel Numerical Computing Technology Research Team

photo: Toshiyuki Imamura

Abstract

For the HPC community, common issues on computational speed and computational accuracy are generally considered to be conflicting. However, the diversity and enhancement of hardware and the high productivity of software have allowed users to choose the precision within the requirement of appropriate computational accuracy. These may provide us with enormous changes in scientific and technical computing, whereas it has been dominated by double-precision calculation for a long time. In the seminar, I will introduce the recent topics such as the relationship between high performance and precision and the relationship between modern hardware and computation accuracy, mainly focusing on the numerical libraries developed by my team in the above topics; i) establishment of higher precision software by massively-and-high-performance low-precision computing units, ii) algorithmic advancement of lower-precision units in scientific computing like HPL-AI benchmark, iii) idea of minimal-precision computing. The first is the realization of a DGEMM-equivalent matrix product using TensorCore(TC) by Mukunoki et al. This is an important fact. It is one of the academic case studies of the utilization of TC's. On the other hand, it suggests the possibility of controlling the number of double precision units by installing a sufficient number of low precision arithmetic units. The second refers to our HPL-AI result, of course, one of the world's four crowning benchmarks and its computation is based on a mixed precision of FP16, FP32, and FP64 formats. The essential point of HPL-AI is to bring out the high performance of low-precision arithmetic while preventing numerical instability and inaccuracy in low-precision arithmetic. It is not simply a matter of rewriting double to half. This is accomplished by a preliminary analysis of the computation target and patterns. The third is to promote the minimum system of computation, which is anticipated to change storage capacity, energy consumption, and minimum hardware requirements of the current floating-point unit. Users won't feel a big impact in terms of input/output, but the internal design of computers will be significantly enhanced.

Important Notes

  • Please turn off your video and microphone when you join the meeting.
  • The broadcasting may be interrupted or terminated depending on the network condition or any other unexpected event.
  • The program schedule and contents may be modified without prior notice.
  • Depending on the utilized device and network environment, it may not be able to watch the session.
  • All rights concerning the broadcasted material will belong to the organizer and the presenters, and it is prohibited to copy, modify, or redistribute the total or a part of the broadcasted material without the previous permission of RIKEN.

(Oct 1, 2020)