Modelpicker icon

Web app · Productivity

Modelpicker

Keep score on the models you actually use — a private benchmark journal

Run the same prompt across GPT-5, Claude, Gemini, Llama or anything else, log which one won and why, and watch a leaderboard build from your own work: win rates, head-to-head records, and which model to reach for per kind of task. Free, local, no API key.

Sign in to open it → Free account · takes 10 seconds with Google

Stored in your browser only. Nothing you do in the web version is sent to our servers. Clear your browser data and it's gone — so export a backup from the in-app menu if you want to keep it.

Modelpicker web app — screen 1
Modelpicker web app — screen 2
Modelpicker web app — screen 3

In your browser

What you can do in Modelpicker on the web

1

Log a comparison in seconds

Task, the models you ran it across, the winner, how clear the win was, who answered fastest, and the note that explains it.

2

Win rate that means something

A model is only judged on the runs it actually entered, so adding a new contender never dilutes anyone's record.

3

Head-to-head records

Every pairing you have tested at least twice, scored only over the runs where both models were in the room.

4

Who to reach for

Twelve kinds of work — coding, summarizing, translation, reasoning, extraction, agentic and more — each with the model that wins them for you.

5

Your own roster

Start from the six models the iPhone app ships with, then add whatever you really test and drop what you never do.

6

Yours to take away

Export every run as CSV or a JSON backup. Nothing is ever sent to a model or a server.

Screens

Every screen, rebuilt for the browser

  1. 1

    Compare

    Your current leader, clear-win share, task filters and every logged run with its contenders and verdict.

  2. 2

    Leaderboard

    Wins by model with rate bars, share-of-wins donut, per-task leaders, head-to-head records and logging pace.

  3. 3

    Benchmark

    One comparison: winner, margin, contenders, your note and the head-to-head record it feeds.

  4. 4

    New comparison

    Pick the kind of work, tap the contenders, crown the winner, say how clear it was and why.

  5. 5

    Edit benchmark

    Change the winner, the contenders or the note — the leaderboard rebuilds itself.

  6. 6

    Settings

    Manage your model roster, see journal totals, export CSV or JSON, load sample data or erase.

Try it with sample data: Loads 22 comparisons from the last seven weeks across five models and eleven kinds of work, with the notes that explain each verdict. You can start fresh at any time.

Ready? It's free.

Sign in once and every All Things AI web app opens instantly. No install, no subscription, no data leaving your browser.

Sign in to open →

Your say

Is Modelpicker useful? Tell us what to build next.

0
0 up · 0 down

Sign in to vote.

Tell us what to build next.

Sign in to leave a comment or vote — it takes a few seconds.

Sign in to comment

No comments yet — be the first to say what this app should do next.

Questions

Is the web version of Modelpicker free?+

Yes. Sign in with a free All Things AI account and use every screen — no subscription, no in-app purchases.

Where is my data stored?+

Only in this browser's local storage. Nothing you enter is sent to our servers, so we can't see it, back it up, or restore it. Clearing site data, a private window, or another device starts empty. Export a backup from the in-app menu to keep it.

Do I need an account?+

You can read about the app without one, but you must sign in to open it. Sign-in is free and takes a few seconds with Google.

How is the web version different from the Modelpicker iPhone app?+

Modelpicker never calls a model: it has no API key and no network access, so you run the prompts wherever you normally do and record the result here. Latency and cost are recorded as your own judgment — which felt fastest, what you noted — not measured from an API. Runs live in this browser only, with no iCloud sync, so export the JSON backup if you want them on another device. There are no public benchmark scores to compare against; the whole point is that the only leaderboard here is yours.

More apps you can try in your browser

All web apps →