Pyyan / Compare / Claude Sonnet 5.5 vs GPT-6 Astra

Claude Sonnet 5.5 vs GPT-6 Astra

2 of 5

Language Models · verified 29 Sept 2026

Claude Sonnet 5.5Anthropiccurrent
GPT-6 AstraOpenAIcurrent
3 slots left
SpecificationClaude Sonnet 5.5GPT-6 Astra
SummaryAbove Opus 5.5 on Terminal-Bench, at half the input price.Saturates three of the benchmarks it reports, and ships disabled for enterprise.
Context1M1.05M
Input ($/Mtok)$2$10
Output ($/Mtok)$10$50
LicenceProprietaryProprietary
Max output128K128K
ParametersNot publishedNot published
Released28 Sep 20263 Sep 2026
Model IDclaude-sonnet-5-5gpt-6-astra
CategoryLanguage ModelsLanguage Models
OfficialAnthropic ↗—

Highlighted rows are where these differ.

Claude Sonnet 5.5

  • 70.6% on Terminal-Bench 4.0, against 66.4% for Opus 5.5 six days earlier
  • 1844 on GDPval-AA, level with Opus 5.5's 1846
  • More than 30% faster than Sonnet 5, and Anthropic says up to 30% less per task on tokens

Best for agentic coding at volume.

Full spec sheet →

GPT-6 Astra

  • 97.6% on FrontierMath Tier 4, 99.9% on ARC-AGI-3 and 100% on ExploitBench
  • Above 272K input tokens the whole request reprices: input and cache at 2x, output at 1.5x
  • Rolls out to OpenAI's Daybreak cyber programme first, and arrives switched off for enterprise

Best for hard reasoning and computer use.

Full spec sheet →