The buying problem
Same test.
Different box.
Run the same local model in LM Studio, keep 48 GPU offload constant, send the same small “hi,” and write down what the M3 Ultra and M4 Max actually do.
Test context from the post: Mac Studio M3 Ultra vs Mac Studio M4 Max · same local chat · 48 GPU offload · prompt: “hi”
Machine A
M3 Ultra
Machine B
M4 Max
This is an observation log, not a benchmark. No hardware is connected and no result is invented.