Practical experience in the numerical dangers of heterogeneous computing
- 1 June 1997
- journal article
- research article
- Published by Association for Computing Machinery (ACM) in ACM Transactions on Mathematical Software
- Vol. 23 (2) , 133-147
- https://doi.org/10.1145/264029.264030
Abstract
Special challenges exist in writing reliable numerical library software for heterogeneous computing environments. Although a lot of software for distributed-memory parallel computers has been written, porting this software to a network of workstations requires careful consideration. The symptoms of heterogeneous computing failures can range from erroneous results without warning to deadlock. Some of the problems are straightforward to solve, but for others the solutions are not so obvious, or incur an unacceptable overhead. Making software robust on heterogeneous systems often requires additional communication. We describe and illustrate the problems encountered during the development of ScaLAPACK and the NAG Numerical PVM Library. Where possible, we suggest ways to avoid potential pitfalls, or if that is not possible, we recommend that the software not be used on heterogeneous networks.Keywords
This publication has 7 references indexed in Scilit:
- Applied Parallel Computing Computations in Physics, Chemistry and Engineering SciencePublished by Springer Nature ,1996
- PVMPublished by MIT Press ,1994
- A set of level 3 basic linear algebra subprogramsACM Transactions on Mathematical Software, 1990
- Algorithm 679: A set of level 3 basic linear algebra subprograms: model implementation and test programsACM Transactions on Mathematical Software, 1990
- An extended set of FORTRAN basic linear algebra subprogramsACM Transactions on Mathematical Software, 1988
- Algorithm 656: an extended set of basic linear algebra subprograms: model implementation and test programsACM Transactions on Mathematical Software, 1988
- Basic Linear Algebra Subprograms for Fortran UsageACM Transactions on Mathematical Software, 1979