Speed sweep

pg-019fd7b6-1a82-761c-8988-7c077f4c611c-sweep

pg-019fd7b6-1a82-761c-8988-7c077f4c611c-sweep

Record

Recipe
pg-liquidai-lfm2-5-8b-a1b-gguf-lfm2-5-8b-a1b-q4-k-m-gguf-q4-k-m-apple-m4-pro-48gb-llama-cpp-b4a8abf506
Measured
2026-08-06T15:34:23.616Z
Points
4

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
18,256143.4observed
132,832121.3observed
1127,93675.2observed
132,76827,834.8observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
pg-019fd7b6-1a82-761c-8988-7c077f4c611c-sweep
measured at
2026-08-06T15:34:23.616Z
recipe id
pg-liquidai-lfm2-5-8b-a1b-gguf-lfm2-5-8b-a1b-q4-k-m-gguf-q4-k-m-apple-m4-pro-48gb-llama-cpp-b4a8abf506
schema version
local-ai-registry/v1

metrics

base memory bytes
6782271488
base memory context tokens
192
concurrency
1
decode32k context tokens
32,832
decode32k tps
121.254
decode8k context tokens
8,256
decode8k tps
143.365
decode max context tokens
127,936
decode max context tps
75.167
decode mode
non-mtp
inference engine version
unknown
latest point at
2026-08-06T15:34:23.616Z
max context tokens
127,936
max prompt tokens
127,872
memory8k bytes
6834225152
memory8k context tokens
8,256
memory max context bytes
7355809792
memory max context tokens
127,936
peak generation tps
121.254
peak memory bytes
7355809792
peak prompt tps
1,177.231
point count
112
ttft32k cached prompt tokens
28,668
ttft32k context tokens
32,768
ttft32k seconds
27.835

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
18,256143.365143.365Unknown6.365Unknown1
132,832121.254121.254Unknown6.851Unknown1
1127,93675.16775.167Unknown6.851Unknown1
132,768UnknownUnknownUnknownUnknownUnknown1
Provenance & metadata (1)

source

paths
publication:pg-20260827T060320709Z, run:019fd7b6-1a82-761c-8988-7c077f4c611c
repository
local.ai