Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Got a 9B Abliterated Claude-Distilled model running for my local hermes

Via r/LocalLlama
Monday, Mar 30, 2026 · 5:50PM
Summary

My laptop only has 6GB of VRAM, which wasn't enough to run abliterated model for my local AI. I managed to completely offload the inference to a free Google Colab T4 GPU and route the API straight back to my local CLI terminal using a Cloudflare tunnel. spent 0$ so far... for a test.

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories