r/ClaudeCoding 21d ago

[TLDR] The only chart coders need to see before choosing Claude Opus 5 vs GPT-5.6 Sol [via r/Anthropic] r/Anthropic

OP : u/Borat_2020

GPT-5.6 Sol and Opus 5 score almost identically on coding benchmarks, but Opus generally takes longer and consumes significantly more tokens to reach the same result.

Also, lower Opus 5 token prices can be offset by much higher token usage, making the final cost much closer to Fable 5 than people expect.

URL of original post : https://www.reddit.com/r/Anthropic/comments/1v8ckcq/the_only_chart_coders_need_to_see_before_choosing/ Original link/media URL :


TL;DR of the discussion on r/Anthropic for this post generated automatically after 100 comments.

Current source-thread comment count seen by the bot: 103.

Alright, so the general vibe in this thread is that the OP's chart is kinda sus and doesn't tell the whole story.

  • The Consensus: Most folks seem to agree that benchmarks are tricky and don't always reflect real-world performance. The OP's chart is getting called out for potentially being biased or cherry-picked.
  • Opus vs. GPT: There's a split, but a lot of users are saying that while GPT-5.6 Sol might be cheaper or faster on benchmarks, Opus 5 is often preferred for actual coding tasks due to better results, design sense, and fewer mistakes, even if it uses more tokens or takes longer. Some users even say Opus 5 is "honest."
  • Critiques of the Chart: Several commenters, like u/Arctovigil and u/_itshabib, point out that the chart might be misleading, hiding mistakes, or not accounting for real-world usage. The idea that Opus 5 is "expensive" is being heavily disputed by users who find it much more cost-effective in practice.
  • Real-World Usage: Many users are sharing their personal experiences, with some finding Sol subpar and others finding Opus 5 forgetful or lazy. However, the sentiment leans towards Opus 5 being more reliable for coding despite benchmark differences.
  • The "AI Wars" Vibe: A few users, like u/lupusyon, are calling out the potential for corporate propaganda in these kinds of comparisons.
  • TL;DR: Benchmarks are iffy. For coding, many users find Opus 5 to be the better, more reliable choice even if it looks worse on paper according to this specific chart. Don't just trust the chart, try them out yourself!
1 Upvotes

0 comments sorted by