Speed sweep

qwen3-8-flash-next-iq4-xs-radeon-ai-pro-r9700-32gb-llama-cpp-tp1-sweep

qwen3-8-flash-next-iq4-xs-radeon-ai-pro-r9700-32gb-llama-cpp-tp1-sweep

Record

Recipe
qwen3-8-flash-next-iq4-xs-radeon-ai-pro-r9700-32gb-llama-cpp-tp1
Measured
2026-08-27T16:27:03.367Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
132,768537.425.64,747.2observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
qwen3-8-flash-next-iq4-xs-radeon-ai-pro-r9700-32gb-llama-cpp-tp1-sweep
measured at
2026-08-27T16:27:03.367Z
recipe id
qwen3-8-flash-next-iq4-xs-radeon-ai-pro-r9700-32gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
035e227 + Qwen4Exp graph-reuse/QSA threshold patch 0249e47
latest point at
2026-08-27T16:27:03.367Z
max context tokens
32,768
peak generation tps
25.639
peak prompt tps
537.394
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
132,76825.63925.6398732.907537.3941
Provenance & metadata (1)