Nonlinear Mapping Networks

12 August 2000

journal article
Published by American Chemical Society (ACS) in Journal of Chemical Information and Computer Sciences

Vol. 40 (6) , 1356-1362
https://doi.org/10.1021/ci000033y

Abstract

Among the many dimensionality reduction techniques that have appeared in the statistical literature, multidimensional scaling and nonlinear mapping are unique for their conceptual simplicity and ability to reproduce the topology and structure of the data space in a faithful and unbiased manner. However, a major shortcoming of these methods is their quadratic dependence on the number of objects scaled, which imposes severe limitations on the size of data sets that can be effectively manipulated. Here we describe a novel approach that combines conventional nonlinear mapping techniques with feed-forward neural networks, and allows the processing of data sets orders of magnitude larger than those accessible with conventional methodologies. Rooted on the principle of probability sampling, the method employs a classical algorithm to project a small random sample, and then “learns” the underlying nonlinear transform using a multilayer neural network trained with the back-propagation algorithm. Once trained, the neural network can be used in a feed-forward manner to project the remaining members of the population as well as new, unseen samples with minimal distortion. Using examples from the fields of image processing and combinatorial chemistry, we demonstrate that this method can generate projections that are virtually indistinguishable from those derived by conventional approaches. The ability to encode the nonlinear transform in the form of a neural network makes nonlinear mapping applicable to a wide variety of data mining applications involving very large data sets that are otherwise computationally intractable.

Keywords

This publication has 11 references indexed in Scilit:

Virtual Compound Libraries: A New Approach to Decision Making in Molecular Discovery Research
Journal of Chemical Information and Computer Sciences, 1998
A new method for analyzing protein sequence relationships based on Sammon maps
Protein Science, 1997
Modern Multidimensional Scaling
Published by Springer Nature ,1997
Synthesis and Applications of Small Molecule Libraries
Chemical Reviews, 1996
Self-Organizing Maps
Published by Springer Nature ,1995
Improving the efficiency of Sammon's nonlinear mapping by using clustering archetypes
Electronics Letters, 1978
A Triangulation Method for the Sequential Mapping of Points from N-Space to Two-Space
IEEE Transactions on Computers, 1977
A Heuristic Relaxation Method for Nonlinear Mapping in Cluster Analysis
IEEE Transactions on Systems, Man, and Cybernetics, 1973
A Nonlinear Mapping for Data Structure Analysis
IEEE Transactions on Computers, 1969
Adaptive Control Processes
Published by Walter de Gruyter GmbH ,1961