SpaceXAI GPT 4.7 is Out and Benchmarks Look Good Close to Opus 5 Max for Agentic Coding
The new model builds on a larger base with extended training on complex tasks, boosting scores like 46.3% on CursorBench 4.0 for software engineering and 71.0% on DeepSWE v1.1.