medium.com2 months agoGPT-5.5 vs Claude Opus 4.7 vs Gemini 3.1 Pro vs DeepSeek V4Apr 28, 2026 · Does it trail on benchmarks? Yes. On SWE-Bench Pro, DeepSeek scores 55.4% versus Opus 4.7's 64.3% and GPT-5.5's 58.6%. On GPQA Diamond, it hits ...Visit medium.com5BookmarkAdd to collection