I’ve found that there’s usually a set standard in the actual work tasks I do when using local LLM’s Around 10k usually goes to model instruction, then itself will spend around 30k looking for context and trying to understand the issue, then around another 10 usually for the actual work with usually