NVIDIA RTX 2060
GPU1
Run records
1
Models covered
1
Frameworks covered
L1 Reproduced × 1
Evidence mix
Typical measured performance
Published benchmark averages per model × framework × quantization on this hardware.
- llama.cpp · Q4_K_M Decode 56.1 · Prefill 131 ·
Key specs
- Vendor
- NVIDIA
- VRAM
- 6 GB
Run records
1 configurations
| Model | Framework | Quantization | Decode | Gen time | Samples |
|---|---|---|---|---|---|
| Qwen3.5-4B | llama.cpp | Q4_K_M | 56.1 tok/s | — | 1 |