llamaperf

RTX 3090 Ti

NVIDIA · 24GB · 1 report

See what fits on this GPU →
This page is thin (1 of 3 reports needed for indexing). Help fill it in.

Qwen3.6 27B

RTX 3090 Ti · llama.cpp · 196,608 ctx

Tone: positive
throughput:
100.0 t/s gen
quant:
Q8_0 (gguf)

Tensor split-mode improved t/s from 70+ to 100+. Peak 130 t/s reported. Power draw 750W+.