Radio
Now Playing
Quickyla Radio โ€” Click to play
Open โ†’
3 min left

Claude Vs ChatGPT: How These AI Assistants Differ

For many, there's a clear winner in this battle of the artificial minds. ChatGPT was probably most people's first encounter with an LLM, considering how many users it picked up shortly after its lauโ€ฆ

Claude Vs ChatGPT: How These AI Assistants Differ
Engadget โ€” 8 August 2026
Text:
20 0 0

For many, there's a clear winner in this battle of the artificial minds.

ChatGPT was probably most people's first encounter with an LLM, considering how many users it picked up shortly after its launch in 2022. It also enjoyed the privileged position of having no real competition until Anthropic released Claude a few months later in March 2023. Since then, both companies have constantly improved their models and added new features. But because there's an overlap in features โ€” both platforms have a coding mode and a dedicated workspace mode โ€” it can be difficult to decide which tool is best for you.

Also, there's little difference between the two during everyday usage: You can use either of them to code, answer a quick query, summarize documents, organize your files, and, if you're feeling particularly blasphemous, write thought leadership pieces in the style of Herman Melville.

However, anything beyond that, and you'll need to start thinking about which LLM fits your use case better.

Measuring accuracy in LLMs can be tricky, as there's no straight answer. The specific model you're using and the prompt you feed into it play an important role in the quality of the output.

When it comes to flagship models โ€” Claude Fable 5 (Max) and GPT 5.6 Sol (Max) โ€” Claude is marginally more accurate according to the AA-Omniscience Accuracy benchmark . The scores stand at 61ย percent and 59ย percent, respectively. Because the difference is so marginal, you'll rarely notice it in day-to-day usage.

But, unless you're tokenmaxxing, you'll be using mid-tier models for most tasks. On Claude, this is Sonnet 5, and on ChatGPT, 5.6 Terra. Here, the scales are tipped in ChatGPT's favor: ChatGPT 5.6 Terra (Max) scores 46 percent, whereas Claude Sonnet 5 (Max) is significantly lower at 38 percent.

Another important factor when measuring accuracy is the tendency of the model to hallucinate. Ideally, if an LLM doesn't know the answer to something, it should flat out refuse to answer it instead of making stuff up, i.e., hallucinating. The benchmark for this is AA-Omniscience Hallucination Rate ,ย in which Claude has a significant leg up against ChatGPT. A lower score is better in this benchmark, and Claude's Fable 5 model scores 55ย percent compared to ChatGPT 5.6 Sol's score of 89 percent. The difference is even more stark in the mid-tier models: Claude Sonnet 5 scores 37ย percent, whereas ChatGPT 5.6 Terra scores 85 percent.

Read Full Story at Engadget โ†’
Advertisement
React:
Sources
Sponsored

More to Read

Hanwha Group and LG CNS tokenize trade receivables to enhanโ€ฆ
๐Ÿ’ป Technology
Hanwha Group and LG CNS tokenize trade receivables to enhance supply chain finance
CoinDesk ยท 13 days ago
7 Statesโ€™ Water Systems Hit by Cyberattacks Likely Tied to โ€ฆ
๐Ÿ’ป Technology
7 Statesโ€™ Water Systems Hit by Cyberattacks Likely Tied to Iran
Wired ยท 8 days ago
Trump administration exempts SpaceX's Starlink from foreignโ€ฆ
๐Ÿ’ป Technology
Trump administration exempts SpaceX's Starlink from foreign router ban
Ars Technica ยท 13 days ago
Hereโ€™s the biggest news you missed this weekend
๐ŸŒ World News
Hereโ€™s the biggest news you missed this weekend
NBC News ยท 14 days ago
Ghana's community service bill: A fix for the prison crisis?
๐ŸŒ World News
Ghana's community service bill: A fix for the prison crisis?
DW World ยท 13 days ago
The Vatican still hasnโ€™t learned how to deal with abuse allโ€ฆ
๐Ÿ•Œ Religion & Faith
The Vatican still hasnโ€™t learned how to deal with abuse allegations
Crux Now ยท 14 days ago
Full view