Friday, March 16, 2012

1203.3339 (Shinji Motok et al.)

Development of Lattice QCD Tool Kit on Cell Broadband Engine Processor    [PDF]

Shinji Motok, i Yoshiyuki Nakagawa, Keitaro Nagata, Koichi Hashimoto, Kiyoshi Mizumaru, Atsushi Nakamura
We report an implementation of a code for SU(3) matrix multiplication on Cell/B.E., which is a part of our project, Lattice Tool Kit on Cell/B.E.. On QS20, the speed of the matrix multiplication on SPE in single precision is 227GFLOPS and it becomes 20GFLOPS {this vaule was remeasured and corrcted.} together with data transfer from main memory by DNA transfer, which is 4.6% of the hardware peak speed (460GFLOPS), and is 7.4% of the theoretical peak speed of this calculation (268.77GFLOPS). We briefly describe our tuning procedure.
View original: http://arxiv.org/abs/1203.3339

No comments:

Post a Comment