Radio
Now Playing
Quickyla Radio โ€” Click to play
Open โ†’
3 min left

Claude Vs ChatGPT: How These AI Assistants Differ

For many, there's a clear winner in this battle of the artificial minds. ChatGPT was probably most people's first encounter with an LLM, considering how many users it picked up shortly after its lauโ€ฆ

Claude Vs ChatGPT: How These AI Assistants Differ
Engadget โ€” 8 August 2026
Text:
12 0 0

For many, there's a clear winner in this battle of the artificial minds.

ChatGPT was probably most people's first encounter with an LLM, considering how many users it picked up shortly after its launch in 2022. It also enjoyed the privileged position of having no real competition until Anthropic released Claude a few months later in March 2023. Since then, both companies have constantly improved their models and added new features. But because there's an overlap in features โ€” both platforms have a coding mode and a dedicated workspace mode โ€” it can be difficult to decide which tool is best for you.

Also, there's little difference between the two during everyday usage: You can use either of them to code, answer a quick query, summarize documents, organize your files, and, if you're feeling particularly blasphemous, write thought leadership pieces in the style of Herman Melville.

However, anything beyond that, and you'll need to start thinking about which LLM fits your use case better.

Measuring accuracy in LLMs can be tricky, as there's no straight answer. The specific model you're using and the prompt you feed into it play an important role in the quality of the output.

When it comes to flagship models โ€” Claude Fable 5 (Max) and GPT 5.6 Sol (Max) โ€” Claude is marginally more accurate according to the AA-Omniscience Accuracy benchmark . The scores stand at 61ย percent and 59ย percent, respectively. Because the difference is so marginal, you'll rarely notice it in day-to-day usage.

But, unless you're tokenmaxxing, you'll be using mid-tier models for most tasks. On Claude, this is Sonnet 5, and on ChatGPT, 5.6 Terra. Here, the scales are tipped in ChatGPT's favor: ChatGPT 5.6 Terra (Max) scores 46 percent, whereas Claude Sonnet 5 (Max) is significantly lower at 38 percent.

Another important factor when measuring accuracy is the tendency of the model to hallucinate. Ideally, if an LLM doesn't know the answer to something, it should flat out refuse to answer it instead of making stuff up, i.e., hallucinating. The benchmark for this is AA-Omniscience Hallucination Rate ,ย in which Claude has a significant leg up against ChatGPT. A lower score is better in this benchmark, and Claude's Fable 5 model scores 55ย percent compared to ChatGPT 5.6 Sol's score of 89 percent. The difference is even more stark in the mid-tier models: Claude Sonnet 5 scores 37ย percent, whereas ChatGPT 5.6 Terra scores 85 percent.

Read Full Story at Engadget โ†’
Advertisement
React:
Sources
Sponsored

More to Read

Alonso pleased with Aston Martin upgrade as Newey targets 'โ€ฆ
๐Ÿ’ป Technology
Alonso pleased with Aston Martin upgrade as Newey targets 'respectability'
Sky Sports ยท 14 days ago
Apple announces Siloโ€™s season 4 return date
๐Ÿ’ป Technology
Apple announces Siloโ€™s season 4 return date
9to5Mac ยท 12 days ago
Anthropic upgrades Claude with new Opus 5 model, details heโ€ฆ
๐Ÿ’ป Technology
Anthropic upgrades Claude with new Opus 5 model, details here
9to5Mac ยท 14 days ago
Why Tesla Stock Crashed Today
๐Ÿ“ˆ Markets & Finance
Why Tesla Stock Crashed Today
Nasdaq News ยท 15 days ago
Hereโ€™s the biggest news you missed this weekend
๐ŸŒ World News
Hereโ€™s the biggest news you missed this weekend
NBC News ยท 12 days ago
ACC Portal Tracker
โšฝ Sports
ACC Portal Tracker
Yahoo Sports ยท 15 days ago
Full view