Speed sweep

qwen3-8-27b-q4-k-m-rx-7900-xtx-24gb-llama-cpp-tp1-sweep

qwen3-8-27b-q4-k-m-rx-7900-xtx-24gb-llama-cpp-tp1-sweep

Record

Recipe
qwen3-8-27b-q4-k-m-rx-7900-xtx-24gb-llama-cpp-tp1
Measured
2026-08-15T14:22:43.715Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
12,0481,169.642.9418.9observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
qwen3-8-27b-q4-k-m-rx-7900-xtx-24gb-llama-cpp-tp1-sweep
measured at
2026-08-15T14:22:43.715Z
recipe id
qwen3-8-27b-q4-k-m-rx-7900-xtx-24gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
llama.cpp b10335 (ROCm)
latest point at
2026-08-15T14:22:43.715Z
max context tokens
2,048
peak generation tps
42.9
peak prompt tps
1,169.6
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
12,04842.942.91,024Unknown1,169.61
Provenance & metadata (1)