Speed sweep

qwen3-6-27b-q4-k-m-dgx-spark-gb10-128gb-llama-cpp-tp1-sweep

qwen3-6-27b-q4-k-m-dgx-spark-gb10-128gb-llama-cpp-tp1-sweep

Record

Recipe
qwen3-6-27b-q4-k-m-dgx-spark-gb10-128gb-llama-cpp-tp1
Measured
2026-05-16T05:06:44.979Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
11,024224.455.2570.4observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
qwen3-6-27b-q4-k-m-dgx-spark-gb10-128gb-llama-cpp-tp1-sweep
measured at
2026-05-16T05:06:44.979Z
recipe id
qwen3-6-27b-q4-k-m-dgx-spark-gb10-128gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
llama.cpp b8214 f7db3f378; Luce DFlash e534780
latest point at
2026-05-16T05:06:44.979Z
max context tokens
1,024
peak generation tps
55.17
peak prompt tps
224.41
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
11,02455.1755.17128Unknown224.411
Provenance & metadata (1)