
ArticleLocal AIAtlassian Forge
Claude Sonnet 5.5 vs Opus 5.5 and GPT-6.1 Sol: the generalist and the specialist
Claude Sonnet 5.5 edged Claude Opus 5.5 on our general-programming benchmark, 0.7984 to 0.7926, a margin I read as level. On an Atlassian Forge app, Opus won 0.9767 to 0.5508. Every check behind both results, GPT-6.1 Sol and Sol Pro on the same two boards, how the benchmarks grade a running app, and the full bill for fine-tuning our own 27B model.