Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

What if a model could only learn what trusted LoRA adapters can express? [R]

Via r/MachineLearning
Tuesday, Jul 7, 2026 · 8:00PM
Summary

Hello I published a paper. Most defenses against fine-tuning poisoning try to detect malicious data or reduce its impact. I explored a different question: What if the model simply could not learn certain malicious updates? The idea is to constrain fine-tuning to a subspace learned from trusted LoRA

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories