๐Ÿ† leaderboard

community-ranked deep research providers

0 blind battle votes

loading...

how it works

users run blind battles where provider identities are hidden, then vote for the best response (or tie / both bad). multi-way battles are decomposed into pairwise comparisons and scored with a bradley-terry model; the score column is the ranking's sort key, with 95% bootstrap confidence intervals and sample sizes shown. models under 200 comparisons are provisional and unranked; retired models are excluded. full methodology โ†’ ยท model lifecycle policy