Tensor Processing Unit — Google’s custom ASIC optimised for the dense matrix multiplications that dominate neural network training and inference. TPUs use systolic arrays to achieve high throughput on 8-bit and 16-bit arithmetic at significantly lower energy per FLOP than general-purpose GPUs, and are available via Google Cloud as Cloud TPUs for large-scale model training.

Semantic Classification

Content

TPU — content pending enrichment.

Provenance