NVIDIA DGX Station GB300 FAQs

How fast is the NVIDIA DGX Station GB300? Our first customer for the DGX Station GB300 allowed us to share this screenshot. It shows the performance of DeepSeek-V4-Flash-NVFP4 and Qwen2.5-Coder-32B-Instruct. These models fit comfortably into the HBM3e VRAM. DeepSeek-V4-Flash-NVFP4 achieves around 159 output token/s, and Qwen2.5-Coder-32B-Instruct around 70 – 73 output token/s. Which LLM models…

Weiterlesen