Speed sweep

pg-01a0139f-fde6-7039-95d0-b030d7b92181-sweep

pg-01a0139f-fde6-7039-95d0-b030d7b92181-sweep

Record

Recipe
pg-liquidai-lfm2-5-8b-a1b-gguf-lfm2-5-8b-a1b-q4-k-m-gguf-q4-k-m-apple-m4-pro-48gb-llama-cpp-b4a8abf506
Measured
2026-08-18T06:47:27.302Z
Points
4

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
18,256143.3observed
132,832121.7observed
1127,93675.2observed
132,76827,796.5observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
pg-01a0139f-fde6-7039-95d0-b030d7b92181-sweep
measured at
2026-08-18T06:47:27.302Z
recipe id
pg-liquidai-lfm2-5-8b-a1b-gguf-lfm2-5-8b-a1b-q4-k-m-gguf-q4-k-m-apple-m4-pro-48gb-llama-cpp-b4a8abf506
schema version
local-ai-registry/v1

metrics

base memory bytes
6778585088
base memory context tokens
192
concurrency
1
decode32k context tokens
32,832
decode32k tps
121.739
decode8k context tokens
8,256
decode8k tps
143.328
decode max context tokens
127,936
decode max context tps
75.18
decode mode
non-mtp
inference engine version
unknown
latest point at
2026-08-18T06:47:27.302Z
max context tokens
127,936
max prompt tokens
127,872
memory8k bytes
6830718976
memory8k context tokens
8,256
memory max context bytes
7364493312
memory max context tokens
127,936
peak generation tps
121.739
peak memory bytes
7364493312
peak prompt tps
1,178.854
point count
112
ttft32k cached prompt tokens
28,668
ttft32k context tokens
32,768
ttft32k seconds
27.796

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
18,256143.328143.328Unknown6.362Unknown1
132,832121.739121.739Unknown6.859Unknown1
1127,93675.1875.18Unknown6.859Unknown1
132,768UnknownUnknownUnknownUnknownUnknown1
Provenance & metadata (1)

source

paths
publication:pg-20260827T060320709Z, run:01a0139f-fde6-7039-95d0-b030d7b92181
repository
local.ai