← Back to community benchmarks

Qwen3.5-122B-A10B-oQ3.5

M3 Ultra (60c) · 96 GB · 3bit · 2026-04-05
Performance
16k
tokens
1,949
PP tok/s
20.1
TG tok/s
8408
TTFT (ms)
62.4
Peak mem (GB)
Hardware
Chip M3 Ultra (60c)
Memory 96 GB
GPU Cores 60
Software
oMLX v0.3.2
macOS macOS 26.4
Context 16,384