Speed sweep
glm-4-9b-chat-q4-0-rocmfp4-fast-radeon-ai-pro-r9700-32gb-llama-cpp-tp3-sweep
glm-4-9b-chat-q4-0-rocmfp4-fast-radeon-ai-pro-r9700-32gb-llama-cpp-tp3-sweepRecord
- Measured
- 2026-08-04T22:40:31.249Z
- Points
- 1
Measured speed
| Concurrency | Context | Prefill | Decode | TTFT ms | Status |
|---|---|---|---|---|---|
| 1 | 783 | 2,952.3 | 65.7 | 178.5 | observed |
Remaining fields
Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.
- id
- glm-4-9b-chat-q4-0-rocmfp4-fast-radeon-ai-pro-r9700-32gb-llama-cpp-tp3-sweep
- measured at
- 2026-08-04T22:40:31.249Z
- recipe id
- glm-4-9b-chat-q4-0-rocmfp4-fast-radeon-ai-pro-r9700-32gb-llama-cpp-tp3
- schema version
- local-ai-registry/v1
metrics
- concurrency
- 1
- inference engine version
- ROCmFPX-117
- latest point at
- 2026-08-04T22:40:31.249Z
- max context tokens
- 783
- peak generation tps
- 65.748
- peak prompt tps
- 2,952.29
- point count
- 1
rows
| concurrency | context tokens | decode tok s | decode tok s per stream | output tokens | peak vram gb | prefill tok s | samples |
|---|---|---|---|---|---|---|---|
| 1 | 783 | 65.748 | 65.748 | 256 | 9.14 | 2,952.29 | 1 |
Provenance & metadata (1)
source
- kind
- leaderboard
- repository
- www.localmaxxing.com ↗