Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Local LLM evaluation advice after DPO on a psychotherapy dataset

Via r/LocalLlama
Saturday, Mar 28, 2026 · 12:28PM
Summary

I fine-tuned Gemma 3 4B on a psychotherapy dataset using DPO as part of an experiment to make a local chatbot that can act as a companion (yes, this is absolutely not intendended to give medical advice or be a therapist). I must thank whoever invented QLoRa and PeFT - I was able to run the finetunin

Continue reading the full article
Read at r/LocalLlama
www.reddit.com
Back to all stories