What's your workflow and what's the best way you have found to code with local LLM when your token generation is < 10 tk/sec?