Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

The loss curve said tie. The judges said otherwise. Seeking replication for an early LLM training result [R]

Via r/MachineLearning
Tuesday, Apr 28, 2026 · 2:43PM
Summary

TL;DR - I've written two novel functions that shape the training signal for LLMs. Early tests show people prefer responses from models trained with my functions by ~59.9%, but I'm just one guy with one GPU. Hoping someone with more resources can prove me right or wrong. The functions: Per-token gain

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories