Speed sweep
glm52-nvfp4-reap-469b-rtxpro6000-vllm-tp4-sweep
glm52-nvfp4-reap-469b-rtxpro6000-vllm-tp4-sweepRecord
- Measured
- 2026-08-24
- Accepted
- 2026-08-24
- Points
- 2
Measured speed
| Concurrency | Context | Prefill | Decode | TTFT ms | Status |
|---|---|---|---|---|---|
| 1 | — | — | 60 | — | historical |
| 1 | 64,000 | 5,100 | 45 | 12,000 | historical |
Remaining fields
Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.
- accepted at
- 2026-08-24
- id
- glm52-nvfp4-reap-469b-rtxpro6000-vllm-tp4-sweep
- measured at
- 2026-08-24
- recipe id
- glm52-nvfp4-reap-469b-rtxpro6000-vllm-tp4
- schema version
- local-ai-registry/v1
metrics
- concurrency
- 1
- inference engine version
- voipmonitor b12x build 20260608; exact commit unreported
- latest point at
- 2026-08-24
- max context tokens
- 64,000
- peak generation tps
- 60
- peak prompt tps
- 5,100
- point count
- 2
rows
| concurrency | context tokens | decode tok s | decode tok s per stream | output tokens | peak vram gb | prefill tok s | samples |
|---|---|---|---|---|---|---|---|
| 1 | Unknown | 60 | 60 | Unknown | Unknown | Unknown | 1 |
| 1 | 64,000 | 45 | 45 | Unknown | Unknown | 5,100 | 1 |
Provenance & metadata (1)
source
- commit
- 579d78b14112198f57aa141332328a79ab02c5a5
- paths
- README.md, docker-compose.yml, .env.example, launch.sh
- repository
- github.com/0xSero/glm-5.2-sm120 ↗