HuggingFace benchmark datasets now let you filter by model size
Via r/LocalLlama
Wednesday, May 20, 2026 · 1:33PM
Summary
Quite useful to see which model under 32B performs best on swebenchverified for example. https://huggingface.co/datasets?benchmark=benchmark:official&sort=trending