Nextbigfuture iconNextbigfutureSep 21, 2026

SpaceXAI GPT 4.7 is Out and Benchmarks Look Good Close to Opus 5 Max for Agentic Coding

The new model builds on a larger base with extended training on complex tasks, boosting scores like 46.3% on CursorBench 4.0 for software engineering and 71.0% on DeepSWE v1.1.

Share this story

Send the public story page.

Useful takeaways from this story.

The new model builds on a larger base with extended training on complex tasks, boosting scores like 46.3% on CursorBench 4.0 for software engineering and 71.0% on DeepSWE v1.1.

It is way better than Grok 4.6 at the same price and speed.

Building the complete brief

The page is ready to read now. The fuller skim-friendly version will appear here automatically.

The useful part

The new model builds on a larger base with extended training on complex tasks, boosting scores like 46.3% on CursorBench 4.0 for software engineering and 71.0% on DeepSWE v1.1. It is way better than Grok 4.6 at the same price and speed.

Details worth keeping

It is way better than Grok 4.6 at the same price and speed.

Keep reading in the app

Open the app view to save this story, compare related coverage, and continue from the same source.

Open in app