llamaperf

Qwen3.8 27B

on NVIDIA RTX Pro 4000 Blackwell · Ollama

Tone: positive
Sep 7, 2026
Throughput
33.9 t/s gen

Summary

User reports Qwen3.8:27b at 33.91 t/s on 2x RTX PRO 4000s and an RTX PRO 2000, up from 12.2 t/s. Splitting the GPUs across VMs produced the gain. User also reports muse-glimmer:30B at 22.61 t/s on the same setup, up from 14.3 t/s.