Configured Workload Fits Comfortably
68.4 GB used out of 128 GB Unified Memory (59.6 GB Headroom). Zero disk swap required.
OPTIMAL FIT
Unified Memory Footprint
68.4 / 128.0 GB (53.4%)
IDEs & Containers (
16.0 GB)
Model Weights
42.0 GB
Llama 3.3 70B @ Q4_K_M
KV Context Buffer
4.4 GB
16K fp16 KV cache
Estimated Token Speed
13.2 tok/s
Based on 614 GB/s mem bus
Available Headroom
59.6 GB
Safe for background compiling
Lineup Viability Matrix for Selected Model & Apps
Click any configuration to switch