I wanted a sign that would tell me which AI to open. So here it is. Pick your plan and what you’re trying to do; the needle is the recommendation.
I check which models you can actually get, then look through independent tests, papers, articles, and whatever useful first-hand posts turn up. Writing, research, coding, and files & app tasks count equally in the overall pick. That last category is about producing files and taking actions across apps, rather than just answering a question.
There’s no secret score. If the comparisons are thin, I make a call and mark it as a lean. A missing test doesn’t make a model bad. And a lab announcing something doesn’t mean it’s in your plan yet.
The colored wedges are just the four tools. They aren’t danger levels. This is an independent project, with no connection to any parks department.