Model

Qwen3.8-27B

ggml-org/Qwen3.8-27B-GGUF

Prefill ladder and generation throughput of Qwen3.8-27B under llama.cpp, measured per GPU and per quant on the GREZA stand.