● FitLLM — fit receipt · 8-bit · 8K tokens ctx · KV F16 · engine 2.8.1
WON'T FIT ✗
| weights | KV cache | linear state | overhead | reserve | total | short by |
|---|---|---|---|---|---|---|
| 25.9 GB | 0.5 GB | 0.1 GB | 3.4 GB | 8.5 GB | 38.4 / 32 GB | 6.4 GB |
→ Quantize to 4-bit and it fits (small quality cost).
Replay: npx fitllm-engine "Qwen 3.8 27B" --mac 32 · JSON · interactive
— embed:

Computed by the open fitllm-engine (MIT) from official config.json values — estimates; runtime varies. Ran it for real? Challenge this prediction — measured reports calibrate the engine for everyone.