Speed sweep

qwen3-coder-30b-a3b-instruct-q5-k-m-ryzen-ai-max-plus-395-128gb-llama-cpp-tp1-sweep

qwen3-coder-30b-a3b-instruct-q5-k-m-ryzen-ai-max-plus-395-128gb-llama-cpp-tp1-sweep

Record

Recipe
qwen3-coder-30b-a3b-instruct-q5-k-m-ryzen-ai-max-plus-395-128gb-llama-cpp-tp1
Measured
2026-04-26T14:39:29.820Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
116,384576.980.5observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
qwen3-coder-30b-a3b-instruct-q5-k-m-ryzen-ai-max-plus-395-128gb-llama-cpp-tp1-sweep
measured at
2026-04-26T14:39:29.820Z
recipe id
qwen3-coder-30b-a3b-instruct-q5-k-m-ryzen-ai-max-plus-395-128gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
b8672
latest point at
2026-04-26T14:39:29.820Z
max context tokens
16,384
peak generation tps
80.46
peak prompt tps
576.85
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
116,38480.4680.4612822.43576.851
Provenance & metadata (1)