
Benchmarking 30B LLMs Across SageMaker GPUs
The benchmark compares Qwen3-Coder-30B and NVIDIA Nemotron-3-Nano-30B across G5, G6, G6e, and G7 instances on Amazon SageMaker AI. It measures throughput, latency, and cost per token, highlighting price-performance gains from G7 NVIDIA Blackwell GPUs.
AWS Machine Learning Blog · 9d ago




















