Open-weight vision-language model from Qwen (Alibaba), released 2026-08-24, 180.0B total, a bespoke open-weights licence. The Flash tier of the Qwen3.8 generation, roughly 180B total parameters, aimed at low-cost high-throughput serving.
0
Quantized builds
1
Benchmark records
Quantized builds
Community and official quantized releases; "linked records" counts only benchmarks tied to that exact release.
No quantized builds cataloged yet.
Benchmark data
1 configurations
| Hardware | Framework | Quant | Decode | VRAM used | Samples | Evidence |
|---|---|---|---|---|---|---|
| 4x NVIDIA Tesla P40 | llama.cpp | UD-Q3_K_XL | 21.6 tok/s | — | 1 | L1 Reproduced |
1 of these records are not linked to a specific quantized build; the full set is on the search page.
Community discussion
No discussions yet — start the first one.
Start a discussion