Speed sweep

gpt-oss-120b-mxfp4-intel-arc-pro-b70-32gb-llama-cpp-tp2-sweep

gpt-oss-120b-mxfp4-intel-arc-pro-b70-32gb-llama-cpp-tp2-sweep

Record

Recipe
gpt-oss-120b-mxfp4-intel-arc-pro-b70-32gb-llama-cpp-tp2
Measured
2026-07-23T00:31:44.693Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
12,0481,124.759.4observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
gpt-oss-120b-mxfp4-intel-arc-pro-b70-32gb-llama-cpp-tp2-sweep
measured at
2026-07-23T00:31:44.693Z
recipe id
gpt-oss-120b-mxfp4-intel-arc-pro-b70-32gb-llama-cpp-tp2
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
mainline 788e07d (2026-07-17)
latest point at
2026-07-23T00:31:44.693Z
max context tokens
2,048
peak generation tps
59.44
peak prompt tps
1,124.7
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
12,04859.4459.44128601,124.71
Provenance & metadata (1)