NVIDIA RTX 5090 Laptop
GPU1
Run records
1
Models covered
1
Frameworks covered
L1 Reproduced × 1
Evidence mix
Typical measured performance
Published benchmark averages per model × framework × quantization on this hardware.
- llama.cpp · NVFP4 Decode 77.1 · Prefill 32.7 ·
Key specs
- Vendor
- NVIDIA
- VRAM
- 24 GB
Run records
1 configurations
| Model | Framework | Quantization | Decode | Gen time | Samples |
|---|---|---|---|---|---|
| Qwen3.8-27B | llama.cpp | NVFP4 | 77.1 tok/s | — | 1 |