Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Trying to train tiny LLMs on length constrained reddit posts summarization task using GRPO on 3xMac Minis - updates!

Via r/LocalLlama
Tuesday, May 5, 2026 · 8:15AM
Summary

So, here's an update to my GRPO training on length constrained reddit posts summarization on 3x Mac minis - a new direction! Gist- been trying to test how good of a summarization model can be trained for summarization using exactly 64 tokens! So, once all the t-test and evals were done for LFM2.5.-3

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories