Speed sweep
laguna-s-2-1-compressed-tensors-int4-group-32-w4a16-intel-arc-pro-b70-32gb-vllm-tp4-sweep
laguna-s-2-1-compressed-tensors-int4-group-32-w4a16-intel-arc-pro-b70-32gb-vllm-tp4-sweepRecord
- Measured
- 2026-07-25T02:36:51.747Z
- Points
- 1
Measured speed
| Concurrency | Context | Prefill | Decode | TTFT ms | Status |
|---|---|---|---|---|---|
| — | 8,192 | — | 94.9 | 5,960.3 | observed |
Remaining fields
Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.
- id
- laguna-s-2-1-compressed-tensors-int4-group-32-w4a16-intel-arc-pro-b70-32gb-vllm-tp4-sweep
- measured at
- 2026-07-25T02:36:51.747Z
- recipe id
- laguna-s-2-1-compressed-tensors-int4-group-32-w4a16-intel-arc-pro-b70-32gb-vllm-tp4
- schema version
- local-ai-registry/v1
metrics
- inference engine version
- 0.1.dev1172+g4a6fd8747.xpu; local source ef334233deabeaeedb607056a2db1c90edb3887c; XPU kernels 4772f727590c51b72add79350b913d098cf67872
- latest point at
- 2026-07-25T02:36:51.747Z
- max context tokens
- 8,192
- peak generation tps
- 94.92
- point count
- 1
rows
| concurrency | context tokens | decode tok s | decode tok s per stream | output tokens | peak vram gb | prefill tok s | samples |
|---|---|---|---|---|---|---|---|
| Unknown | 8,192 | 94.92 | Unknown | 512 | Unknown | Unknown | 1 |
Provenance & metadata (1)
source
- kind
- leaderboard
- repository
- www.localmaxxing.com ↗