Smoke test of the Qwen 3.8 Max model (2.4T total parameters, 95B active) using community NVFP4 weights on 8x NVIDIA B300 SXM6 275GB GPUs, served via the midstream v0.27.1 container image.
| Field | Value |
|---|---|
| Model | Inferact/Qwen3.8-2.4T-A95B-NVFP4 (community NVFP4 quant of Qwen/Qwen3.8-2.4T-A95B) |
| Container Image | quay.io/vllm/automation-vllm:cuda-31490750632 |