Model
Qwen3.8-27B
ggml-org/Qwen3.8-27B-GGUF
Prefill ladder and generation throughput of Qwen3.8-27B under llama.cpp, measured per GPU and per quant on the GREZA stand.
Model
ggml-org/Qwen3.8-27B-GGUF
Prefill ladder and generation throughput of Qwen3.8-27B under llama.cpp, measured per GPU and per quant on the GREZA stand.