Speed sweep

qwen3-6-35b-a3b-q4-k-m-rtx-3090-24gb-llama-cpp-tp1-sweep

qwen3-6-35b-a3b-q4-k-m-rtx-3090-24gb-llama-cpp-tp1-sweep

Record

Recipe
qwen3-6-35b-a3b-q4-k-m-rtx-3090-24gb-llama-cpp-tp1
Measured
2026-06-14T00:41:56.889Z
Points
1

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatus
1131,0721,541167.930observed

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

id
qwen3-6-35b-a3b-q4-k-m-rtx-3090-24gb-llama-cpp-tp1-sweep
measured at
2026-06-14T00:41:56.889Z
recipe id
qwen3-6-35b-a3b-q4-k-m-rtx-3090-24gb-llama-cpp-tp1
schema version
local-ai-registry/v1

metrics

concurrency
1
inference engine version
BeeLlama v0.3.2 (build 10251, commit 100fc138f)
latest point at
2026-06-14T00:41:56.889Z
max context tokens
131,072
peak generation tps
167.92
peak prompt tps
1,541
point count
1

rows

concurrencycontext tokensdecode tok sdecode tok s per streamoutput tokenspeak vram gbprefill tok ssamples
1131,072167.92167.9240023.41,5411
Provenance & metadata (1)