The biggest AI thread today is SpaceX putting its muscle behind an unusual choice of venue: shipping Grok 4.7 as a commercial model with explicit, measurable performance claims and a clearer stance on long-horizon behavior. SpaceX says Grok 4.7 adds long-horizon processing—basically the ability to keep working across extended, multi-step tasks—while also layering in safety upgrades and evaluating the model across multiple benchmarks. The company’s most concrete datapoint is cost: it reports Grok 4.7 completed CursorBench 4.0 tasks at an average $4.69 per task, a reminder that capability is only half the story when customers are pricing every prompt.
The rollout also comes with a straightforward pricing ladder. Grok 4.7 starts at $2 per million input tokens and $6 per million output tokens, and SpaceX offers a twice-faster variant that costs twice as much—useful for teams that need latency over thrift, such as interactive coding or customer-facing assistants. The subtext is that SpaceX is positioning Grok not just as a chatbot, but as an execution engine: long-horizon processing, measurable task completion, and unit economics that don’t force enterprises to do mental gymnastics before buying in.