Speed sweep
qwen3-fp8-intel-arc-pro-b70-32gb-vllm-tp4-sweep
qwen3-fp8-intel-arc-pro-b70-32gb-vllm-tp4-sweepRecord
- Measured
- 2026-06-19
- Points
- 1
Measured speed
| Concurrency | Context | Prefill | Decode | TTFT ms | Status |
|---|---|---|---|---|---|
| 1 | 262,144 | — | 71.7 | 95 | historical |
Remaining fields
Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.
- id
- qwen3-fp8-intel-arc-pro-b70-32gb-vllm-tp4-sweep
- measured at
- 2026-06-19
- recipe id
- qwen3-fp8-intel-arc-pro-b70-32gb-vllm-tp4
- schema version
- local-ai-registry/v1
metrics
- concurrency
- 1
- inference engine version
- g6607a80da (custom)
- latest point at
- 2026-06-19
- max context tokens
- 262,144
- peak generation tps
- 71.7
- point count
- 1
rows
| concurrency | context tokens | decode tok s | decode tok s per stream | output tokens | peak vram gb | prefill tok s | samples |
|---|---|---|---|---|---|---|---|
| 1 | 262,144 | 71.7 | 71.7 | 200 | 127.2 | Unknown | 1 |
Provenance & metadata (1)
source
- repository
- www.localmaxxing.com ↗