Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Making LLMs tell you how confident they really are through probe-targeted fine tuning.[R]

Via r/MachineLearning
Friday, May 29, 2026 · 5:15AM
Summary

Just wanted to share my research regarding probe-targeted fine-tuning (LoRa) for verbal confidence calibration., If you probe the hidden states of an instruct-tuned LLM, it can tell correct from incorrect answers at 0.76–0.88 AUROC. But when you ask it directly it tends to respond with confidence at

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories