Abstract
The present disclosure relates to a method for quantization-aware training of a neural network including a neural network processing unit, NPU, the method comprising the following steps: obtaining a plurality of tensors of the neural network from an NPU inference; replacing the corresponding tensors during the training of the neural network with the plurality of tensors of the neural network from the NPU inference.
Full Text
What is claimed is:
The present disclosure relates to a method for quantization-aware training of a neural network including a neural network processing unit, NPU, the method comprising the following steps: obtaining a plurality of tensors of the neural network from an NPU inference; replacing the corresponding tensors during the training of the neural network with the plurality of tensors of the neural network from the NPU inference.
Timeline
Filed
05/20/2026Published
09/17/2026Granted
Not AvailableIPC Codes(1)
G06N 3/08:Learning methods