Speed sweep

pg-01a02811-1362-78bd-b108-82c0815d9424-sweep

pg-01a02811-1362-78bd-b108-82c0815d9424-sweep

Record

Recipe
pg-liquidai-lfm2-5-8b-a1b-gguf-lfm2-5-8b-a1b-q4-k-m-gguf-q4-k-m-apple-m4-pro-48gb-llama-cpp-b4a8abf506
Measured
2026-08-22T06:03:22.839Z
Points
4

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
18,256144.5observed
132,832121.3observed
1127,93675.3observed
132,76829,317.2observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
pg-01a02811-1362-78bd-b108-82c0815d9424-sweep
measured at
2026-08-22T06:03:22.839Z
recipe id
pg-liquidai-lfm2-5-8b-a1b-gguf-lfm2-5-8b-a1b-q4-k-m-gguf-q4-k-m-apple-m4-pro-48gb-llama-cpp-b4a8abf506
schema version
local-ai-registry/v1

metrics

base memory bytes
6871744512
base memory context tokens
192
concurrency
1
decode32k context tokens
32,832
decode32k tps
121.263
decode8k context tokens
8,256
decode8k tps
144.54
decode max context tokens
127,936
decode max context tps
75.306
decode mode
non-mtp
inference engine version
unknown
latest point at
2026-08-22T06:03:22.839Z
max context tokens
127,936
max prompt tokens
127,872
memory8k bytes
6921732096
memory8k context tokens
8,256
memory max context bytes
7440809984
memory max context tokens
127,936
peak generation tps
121.263
peak memory bytes
7440809984
peak prompt tps
1,117.706
point count
112
ttft32k cached prompt tokens
28,668
ttft32k context tokens
32,768
ttft32k seconds
29.317

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
18,256144.54144.54Unknown6.446Unknown1
132,832121.263121.263Unknown6.93Unknown1
1127,93675.30675.306Unknown6.93Unknown1
132,768UnknownUnknownUnknownUnknownUnknown1
Provenance & metadata (1)

source

paths
publication:pg-20260827T060320709Z, run:01a02811-1362-78bd-b108-82c0815d9424
repository
local.ai