AWS Uses Intel's Habana Gaudi for Large Language Models

Intel Habana Labs Gaudi Accelerator
(Image credit: Intel)

While Intel's Habana Gaudi offers somewhat competitive performance and comes with the Habana SynapseAI software package, it still falls short compared to Nvidia's CUDA-enabled compute GPUs. This, paired with limited availability, is why Gaudi hasn't been as popular for large language models (LLMs) like ChatGPT. 

Now that the AI rush is on, Intel's Habana is seeing broader deployments. Amazon Web Services decided to try Intel's 1st Generation Gaudi with PyTorch and DeepSpeed to train LLMs, and the results were promising enough to offer DL1 EC2 instances commercially.

Anton Shilov
Contributing Writer

Anton Shilov is a contributing writer at Tom’s Hardware. Over the past couple of decades, he has covered everything from CPUs and GPUs to supercomputers and from modern process technologies and latest fab tools to high-tech industry trends.

  • bit_user
    Why does the article URL say "aws-uses-intel-habana-gaudi-for-llvm" ? LLVM (Low Level Virtual Machine) is something completely different.
    https://en.wikipedia.org/wiki/LLVM
    As for the article, it'd be ideal if it had some stats on how Nvidia A100's or H100's perform on the same workload. However, it's definitely cool to hear about Gaudi in the wild. Virtually all I've heard about it, since Intel bought them, is about their kernel patches and driver work.

    Makes me wonder how AMD's MI250X is getting on...
    Reply