People are warning me about the prompt-processing speed of a MacBook Pro M5 Max with 128 GB RAM. My main concern is prompt ingestion / prefill latency and large-context handling — not raw token generation speed (which I think is OK). I only plan to use Qwen 3.5 / 3.6 / 3.7 models or similar mostly c