Model comparison
ChatGPTClaudeGemini

Sol, Sonnet, Argon: comparing the new AI models

The AI week of 28 September–2 October 2026: model announcements, an external test and videos.

Illustration for Findbest.si
What you will learn
  1. Sol and Sonnet are candidates for your own tasks; access to Argon is initially limited.

  2. Compare model, tools and settings together: the name alone does not explain a result.

  3. Use the tasks below as an evaluation plan. Provider claims and creator videos are not Findbest measurements.

Findbest perspective

The short answer

An announced model is not necessarily a usable alternative. Our overview combines provider information with a plan for your own comparison. It does not claim a winner based on our own measurements.

The differences at a glance

Model comparison
CriterionGPT‑6.1 SolClaude Sonnet 5.5Gemini 4 Argon
Announced29 September28 September30 September
PositioningCoding and agentsWell-defined everyday tasksInitially cyber-defender access
Access checkCheck plan and administrator enablementCheck model selection in your productSelected Fairwind partners initially
Our suggested taskChange an existing codebaseProduce a document to specificationPrepare tasks and wait for access
Findbest hands-on testNot performed yetNot performed yetNot performed yet

First: ChatGPT and GPT are different things

ChatGPT is OpenAI’s application. GPT refers to model families used inside it and other products. Claude is Anthropic’s assistant; names such as Sonnet refer to models. Enabled tools, plan and working environment also affect the result. A new model release alone does not tell you which features your account can access.

Open practical example

Before comparing, record the application, selected model, date, plan and enabled tools. If the app chooses automatically, record that too. Without these details, a result is hard to interpret later.

Your practical comparison

What you need

A real task, an expected outcome and access to the comparison models. Use anonymized data and record model name, date, reasoning setting and tools.

  1. Choose three tasks: a known code defect, a structured analysis and a writing assignment with clear requirements.
  2. Give each model the same files and requirements. Set a time limit and correction budget in advance.
  3. Check outcomes: does it work, are the facts correct and were all requirements followed?
  4. Record waiting time, your own rework and actual costs separately.
  5. Repeat difficult tasks before deciding which model should handle them in future.

Example to get you started

Test prompt: “Group these anonymized comments into five themes. Include a supporting passage for each and mark missing information.” Criteria: correct grouping, no invented quotations and clearly identified gaps.

Things to consider

The table describes access and useful evaluation tasks. Our suggestions are editorial analysis. Different provider benchmarks do not form a common leaderboard. The creator videos below present outside experiences; titles and conclusions belong to their authors.

Videos to explore

Original provider and creator videos, in English. Titles belong to their authors; their assessments are not Findbest test results. Preview artwork consists of Findbest illustrations.

External comparison

I Tested Sonnet 5.5 vs. GPT-6.1 Sol on 6 Real Use Cases

Duncan Rogoff | Learn Claude Code · English

A connection to YouTube is made only when you load the video. Your IP address and technical device information will be transmitted.

Watch on YouTube

External comparison

GPT-6.1 Sol vs Sonnet 5.5 – My Real App Tests

AIex The AI Workbench · English

A connection to YouTube is made only when you load the video. Your IP address and technical device information will be transmitted.

Watch on YouTube

Explore the tools

ChatGPTView tool →Visit official website
ClaudeView tool →Visit official website
GeminiView tool →Visit official website

Read next

Ideas to put into practice →