The Model Efficiency series has moved to the company blog

The story of making neural networks smaller and faster (number formats, quantization, pruning, and so on) is now serialized on the company blog. Here is where to go.

January 4, 2026 · 1 min · rick