The Model Efficiency series has moved to the company blog
The story of making neural networks smaller and faster (number formats, quantization, pruning, and so on) is now serialized on the company blog. Here is where to go.
The story of making neural networks smaller and faster (number formats, quantization, pruning, and so on) is now serialized on the company blog. Here is where to go.