Qwen3.8-Flash-Next

180BChatby Qwen (Alibaba) Start a discussion

Open-weight vision-language model from Qwen (Alibaba), released 2026-08-24, 180.0B total, a bespoke open-weights licence. The Flash tier of the Qwen3.8 generation, roughly 180B total parameters, aimed at low-cost high-throughput serving.

0

Quantized builds

1

Benchmark records

Quantized builds

Community and official quantized releases; "linked records" counts only benchmarks tied to that exact release.

No quantized builds cataloged yet.

Benchmark data

1 configurations

HardwareFrameworkQuantDecodeVRAM usedSamplesEvidence
4x NVIDIA Tesla P40llama.cppUD-Q3_K_XL21.6 tok/s—1L1 Reproduced

1 of these records are not linked to a specific quantized build; the full set is on the search page.

Community discussion

No discussions yet — start the first one.

Start a discussion