Recipe

qwen3-6-35b-a3b-mlx-4bit-apple-m4-pro-48gb-lmstudio-tp1

qwen3-6-35b-a3b-mlx-4bit-apple-m4-pro-48gb-lmstudio-tp1

Observed LocalMaxxing leaderboard run. Evidence for compatibility, not an executable launch contract.

Record

Status
candidate
Source
localmaxxing
Engine
lmstudio
Engine version
0.4.20+1
Accelerators
1
Tensor parallel
1
Context tokens
154,112
Max concurrency
4
chat
unknown
reasoning
unknown
tools
unknown
vision
unknown

Hugging Face model card

Identity

https://huggingface.co/mlx-community/Qwen3.6-35B-A3B-OptiQ-4bit
Repository
mlx-community/Qwen3.6-35B-A3B-OptiQ-4bit
Status
known
Link type
Exact Hub repository

Public Hugging Face repository confirmed by the Hub API.

Candidate: useful compatibility or speed evidence. The registry does not offer Run until promotion requirements are met.

Observed configuration

lmstudio

Evidence only · candidate · reference

Candidate evidence — not a Run contract

Source
https://www.localmaxxing.com/en/runs/cmsqri8th00qimr016ohrvv20

Observed source tokens

Mechanical split of the source command. Unverified against the engine CLI. The registry does not offer Run for this recipe.

    1. lms
    2. server
    3. start
    4. --port
    5. 1234
    6. --bind
    7. 127.0.0.1
    1. lms
    2. load
    3. qwen3.6-35b-a3b-optiq
    4. --context-length
    5. 154112
    6. --parallel
    7. 4
    8. --identifier
    9. benchmark-qwen36-35b-optiq
    10. --yes
FlagValue
--context-length154112
--parallel4
--identifierbenchmark-qwen36-35b-optiq

Measured speed

ConcurrencyContextPrefillDecodeTTFT msStatusSweep
154,112717.861.4762observedqwen3-6-35b-a3b-mlx-4bit-apple-m4-pro-48gb-lmstudio-tp1-sweep

Remaining fields

Identity, launch, related records, and measured speed are shown above. This is the rest of the normalized record.

hardware count
1
hardware id
apple-m4-pro-48gb
id
qwen3-6-35b-a3b-mlx-4bit-apple-m4-pro-48gb-lmstudio-tp1
model instance id
mlx-community-qwen3-6-35b-a3b-optiq-4bit--mlx-4bit
recipe source
localmaxxing
schema version
local-ai-registry/v1
speed sweep ids
qwen3-6-35b-a3b-mlx-4bit-apple-m4-pro-48gb-lmstudio-tp1-sweep
status
candidate

capabilities

engine

name
lmstudio
version
0.4.20+1

serving

max concurrency
4
max context tokens
154,112
tensor parallel
1
Provenance & metadata (3)

facts

capabilities.chat · provenance · captured at
2026-08-30T09:10:02Z
capabilities.chat · reason capability-not-verified
capabilities.chat · state unknown
capabilities.reasoning · provenance · captured at
2026-08-30T09:10:02Z
capabilities.reasoning · reason capability-not-verified
capabilities.reasoning · state unknown
capabilities.tools · provenance · captured at
2026-08-30T09:10:02Z
capabilities.tools · reason capability-not-verified
capabilities.tools · state unknown
capabilities.vision · provenance · captured at
2026-08-30T09:10:02Z
capabilities.vision · reason capability-not-verified
capabilities.vision · state unknown
engine.graph mode · provenance · captured at
2026-08-30T09:10:02Z
engine.graph mode · reason runtime-detail-not-published
engine.graph mode · state unknown
serving.kv cache tokens · provenance · captured at
2026-08-30T09:10:02Z
serving.kv cache tokens · reason kv-cache-capacity-not-published
serving.kv cache tokens · state unknown
serving.max concurrency · provenance · captured at
2026-08-31T23:03:15Z
serving.max concurrency · reason server-capacity-derived-from-source-evidence
serving.max concurrency · state known

metadata

localmaxxing · backend
metal
localmaxxing · batch size
1
localmaxxing · hardware label
M4 Pro
localmaxxing · notes
OptiQ mixed-precision MLX 4-bit quantization. Median of 3 timed runs after 1 warmup using the same deterministic 512-word source-prompt protocol at temperature 0. The model-specific chat template produced 547 prompt tokens; each run generated 254 output tokens. Prefill estimated as prompt tokens divided by LM Studio native time-to-first-token.
localmaxxing · observed command
lms server start --port 1234 --bind 127.0.0.1 && lms load qwen3.6-35b-a3b-optiq --context-length 154112 --parallel 4 --identifier benchmark-qwen36-35b-optiq --yes
localmaxxing · run id
cmsqri8th00qimr016ohrvv20
localmaxxing · tokenized · arguments
lms, load, qwen3.6-35b-a3b-optiq, --context-length, 154112, --parallel, 4, --identifier, benchmark-qwen36-35b-optiq, --yes
localmaxxing · tokenized · fidelity
faithful

localmaxxing · tokenized · steps

01234567
lmsserverstart--port1234--bind127.0.0.1Unknown
lmsloadqwen3.6-35b-a3b-optiq--context-length154112--parallel4--identifier

provenance

captured at
2026-08-30T09:10:02Z

sources

captured atkindurl
2026-08-30T09:10:02Znormalized-recipewww.localmaxxing.com/en/runs/cmsqri8th00qimr016ohrvv20