Speed sweep

llama-3-2-1b-instruct-q8-0-intel-arc-pro-b60-24gb-llama-cpp-tp1-sweep

llama-3-2-1b-instruct-q8-0-intel-arc-pro-b60-24gb-llama-cpp-tp1-sweep

Record

Recipe
llama-3-2-1b-instruct-q8-0-intel-arc-pro-b60-24gb-llama-cpp-tp1
Measured
2026-06-29T13:03:06.147Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
14,0968,333.5194.761.4observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
llama-3-2-1b-instruct-q8-0-intel-arc-pro-b60-24gb-llama-cpp-tp1-sweep
measured at
2026-06-29T13:03:06.147Z
recipe id
llama-3-2-1b-instruct-q8-0-intel-arc-pro-b60-24gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
latest point at
2026-06-29T13:03:06.147Z
max context tokens
4,096
peak generation tps
194.71
peak prompt tps
8,333.45
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
14,096194.71Unknown128Unknown8,333.451
Provenance & metadata (1)