Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

[D] Data curation and targeted replacement as a pre-training alignment and controllability method

Via r/MachineLearning
Sunday, Mar 29, 2026 · 6:53PM
Summary

Hi, r/MachineLearning: has much research been done in large-scale training scenarios where undesirable data has been replaced before training, such as any instances of violence, lying, or deception in the dataset? Most controllability work, like RLHF or constitutional AI, seems to be done post-train

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories