So, here's an update to my GRPO training on length constrained reddit posts summarization on 3x Mac minis - a new direction! Gist- been trying to test how good of a summarization model can be trained for summarization using exactly 64 tokens! So, once all the t-test and evals were done for LFM2.5.-3