Qwen3.8-27B · Q6_K
L0 自报AMD Radeon RX 7900 GRE + NVIDIA GTX 1080Ti · 27GB · llama.cpp
今天
12.53 tok/s
Decode
116.21 tok/s
Prefill
— s
TTFT
23 GB
VRAM
配置(最近一次实测):
Ubuntu Q6_K llama.cpp 10453
来源:Reddit · 1 次实测 · 1 个独立来源
AMD Radeon RX 7900 GRE + NVIDIA GTX 1080Ti · 27GB · llama.cpp
今天
12.53 tok/s
Decode
116.21 tok/s
Prefill
— s
TTFT
23 GB
VRAM
配置(最近一次实测):
来源:Reddit · 1 次实测 · 1 个独立来源
AMD Radeon RX 7900 GRE + NVIDIA GTX 1080Ti · 27GB · llama.cpp
今天
11.13 tok/s
Decode
71.46 tok/s
Prefill
— s
TTFT
17.2 GB
VRAM
配置(最近一次实测):
来源:Reddit · 1 次实测 · 1 个独立来源
2x NVIDIA RTX 3090 · 48GB · llama.cpp
GitHub
今天
43.65 tok/s
Decode
1,600.18 tok/s
Prefill
0.42 s
TTFT
44.12 GB
VRAM
配置(最近一次实测):
来源:GitHub · 1 次实测 · 1 个独立来源
共 3 条结果