DOE OSTI · 1869509
CrossSim Inference Manual v2.0
Abstract
Neural networks are largely based on matrix computations. During forward inference, the most heavily used compute kernel is the matrix-vector multiplication (MVM): $W \vec{x} $. Inference is a first frontier for the deployment of next-generation hardware for neural network applications, as it is more readily deployed in edge devices, such as mobile devices or embedded processors with size, weight, and power constraints. Inference is also easier to implement in analog systems than training, which has more stringent device requirements. The main processing kernel used during inference is the MVM.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Xiao, Tianyao Patrick, Bennett, Christopher H., Feinberg, Benjamin, Marinella, Matthew J., Agarwal, Sapan. 2022-05-01. CrossSim Inference Manual v2.0. https://doi.org/10.2172/1869509
Cite the original work for its findings. Save a collection to share your selection of sources.