Best AI News โ€” Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

Open-weight 4B models approach o3-level medical question answering in Swedish [P]

Via r/MachineLearning
Sunday, Jul 26, 2026 ยท 11:58AM
Summary

I have been running some experiments with smaller open-weight LLMs on multiple-choice questions of Swedish medical licensing exams. On a dataset called MedQA-SWE, GPT-4 scored 84% accuracy in 2024 and o3 scored 88% in 2025 on a smaller, overlapping dataset. With post-training (SFT) on data from earl

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories