Speed sweep

mistral-small-3-2-24b-instruct-2506-gguf-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-sweep

mistral-small-3-2-24b-instruct-2506-gguf-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-sweep

Record

Recipe
mistral-small-3-2-24b-instruct-2506-gguf-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1
Measured
2026-07-04T21:06:31.647Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
14,09627.31,501.8observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
mistral-small-3-2-24b-instruct-2506-gguf-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1-sweep
measured at
2026-07-04T21:06:31.647Z
recipe id
mistral-small-3-2-24b-instruct-2506-gguf-ud-q4-k-xl-intel-arc-pro-b70-32gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
llama.cpp fdb1db877 / llama-server 9763 dec5ca557
latest point at
2026-07-04T21:06:31.647Z
max context tokens
4,096
peak generation tps
27.297
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
14,09627.29727.297128UnknownUnknown1
Provenance & metadata (1)