AI product
Tensor Processing Unit (TPU) is a family of application-specific integrated circuits (ASICs) and managed services developed and operated by Google to accelerate machine learning workloads. TPUs are available as on-premises hardware and as a managed Cloud TPU service for training and inference of neural networks and deep learning models. First announced in 2016, TPUs are used inside Google's datacenters and offered to external customers via Google Cloud.
3 uses taken from transcripts — each links to the moment in the video.
Performs low-precision dense linear algebra for machine-learning inference with lower energy use and latency than the CPUs and GPUs described in the transcript.
Discussed as specialized AI hardware used by Google and consumed in large numbers by Anthropic as an alternative compute platform.
Provided the compute used for the multilingual pre-training experiments.
3 in the library.