← Front page

How I AI · Tuesday, June 30, 2026

Weighted Leaderboard Recommends Models for Specific Tasks

Following a weighted analysis balancing human opinion and backend performance, a new leaderboard was generated, with Sonnet 46 and Gemini 3 Pro leading, followed by GPT 5.5. Recommendations were made per task: GPT 5.5 for PRDs, Sonnet 46 for prototyping and conversational tasks, and Opus 48/Sonnet 5 for codebases.

companyAnthropiccompanyGoogle

The tape

2 quotes
At the top of the list, Sonnet 46. Who would have thunk? And Gemini 3 Pro. Followed by what I think is my favorite 55. And at the bottom, poor brand new Sonnet 5 and really expensive 48.
Model by task. If you're writing a PRD, use GPT 5.5 because it will give you something comprehensive and clear. If you are prototyping, guess what? Sonnet 46. Pretty good. And if you want to chit chat with a model, again, Sonnet 46 has good vibes.
Heard on How I AI — “Sonnet 5 review: I ran 64 generations to find out if it's worth it, published Tuesday, June 30, 2026. Heardvine summarizes and quotes with attribution and timestamps, and links to the original everywhere.
Transcribed via Gemini audio transcription · $0.02
Weighted Leaderboard Recommends Models for Specific Tasks — Heardvine