Speed sweep

qwen-3-14b-instruct-gguf-q4-k-m-intel-arc-pro-b70-32gb-llama-cpp-tp1-sweep

qwen-3-14b-instruct-gguf-q4-k-m-intel-arc-pro-b70-32gb-llama-cpp-tp1-sweep

Record

Recipe
qwen-3-14b-instruct-gguf-q4-k-m-intel-arc-pro-b70-32gb-llama-cpp-tp1
Measured
2026-07-05T01:53:40.580Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
12,04838.2240observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
qwen-3-14b-instruct-gguf-q4-k-m-intel-arc-pro-b70-32gb-llama-cpp-tp1-sweep
measured at
2026-07-05T01:53:40.580Z
recipe id
qwen-3-14b-instruct-gguf-q4-k-m-intel-arc-pro-b70-32gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
fdb1db877c526ec90f668eca1b858da5dba85560
latest point at
2026-07-05T01:53:40.580Z
max context tokens
2,048
peak generation tps
38.249
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
12,04838.24938.249128UnknownUnknown1
Provenance & metadata (1)