Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Extracted MTP tensor GGUFs - smaller donor models for grafting.

Via r/LocalLlama
Thursday, May 7, 2026 · 11:28PM
Summary

The script to graft MTP tensors requires a full GGUF model file. I felt that was a bit hefty, so I asked local Gemma to write something to just extract what's required. The results are two faux GGUFs weighing in at just 900MB (35A3B) and 450MB (27B), containing only the tensors and fully compatible

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories