RTX A6000 vs RTX 3090 Deep Learning Benchmarks | Lambda

RTX A6000 vs RTX 3090 Deep Learning Benchmarks

August 9, 2021• 3 min read

Lambda is currently shipping servers and workstations with RTX 3090 and RTX A6000 GPUs. In this post, we benchmark the PyTorch training speed of these top-of-the-line GPUs. For more info, including multi-GPU training performance, see our GPU benchmarks for PyTorch & TensorFlow.

For training image models (convnets) with PyTorch, a single RTX A6000 is...

For training language models (transformers) with PyTorch, a single RTX A6000 is...

For training image models (convnets) with PyTorch, 8x RTX A6000 are...

For training language models (transformers) with PyTorch, 8x RTX A6000 are...

* In this post, 32-bit refers to TF32; Mixed precision refers to Automatic Mixed Precision (AMP).
GPUDirect peer-to-peer (via PCIe) is enabled for RTX A6000s, but does not work for RTX 3090s.

3090 vs A6000 convnet training speed with PyTorch

3090 vs A6000 language model training speed with PyTorch

Benchmark software stack

Lambda's benchmark code is available at the GitHub repo here.