Best AI News — Updated Every 3 Hours
Story Page
← All Stories
Home Community Story
Community

LLM rankings are not a ladder: experimental results from a transitive benchmark graph [D]

Via r/MachineLearning
Saturday, May 9, 2026 · 7:16PM
Summary

I built a small website called LLM Win: https://llm-win.com It turns LLM benchmark results into a directed graph: text If model A beats model B on benchmark X, add an edge A -> B. Then it searches for the shortest transitive chain between two models. The meme version is: text Can LLaMA 2 7B beat Cla

Continue reading the full article
Read at r/MachineLearning
www.reddit.com
Back to all stories