Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Developing open source LLM from ground up from pretrain - rlhf(PPO/GRPO)

Via r/LocalLlama
Thursday, May 14, 2026 · 7:38PM
Summary

Hello I have been working on creating a LLM from ground up. It is based on deepseek architecture with heavily VRAM footprint reduced optimized(GUM+muon) Currently this is the json schema I am using which should suffice as to what currently is being pretrained. I have 2 6000 pro 600W Testing a 7B par

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories