Speed sweep

qwen3-6-27b-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-foe4kr0t-sweep

qwen3-6-27b-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-foe4kr0t-sweep

Record

Recipe
qwen3-6-27b-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-foe4kr0t
Measured
2026-07-22T18:39:35.041Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
12,048501.122observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
qwen3-6-27b-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-foe4kr0t-sweep
measured at
2026-07-22T18:39:35.041Z
recipe id
qwen3-6-27b-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-foe4kr0t
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
fork YanissAmz/llama.cpp branch puzzle-port (base b566325)
latest point at
2026-07-22T18:39:35.041Z
max context tokens
2,048
peak generation tps
21.95
peak prompt tps
501.1
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
12,04821.9521.95128Unknown501.11
Provenance & metadata (1)