AInews: xAI Unveils Grok 4.7 With Expanded Reasoning for Developers

AInews: xAI introduced Grok 4.7 on September 21, 2026, positioning the model as its strongest system yet for software development, agentic tasks and professional knowledge work. The release is available through Cursor, Grok Build, the Grok API, third-party coding tools, model routers and cloud platforms, according to xAI’s launch announcement.
What did xAI release on September 21?
xAI released Grok 4.7 as a new flagship model for work that requires extended reasoning, code generation and research. The company says the system is built on a larger base model than Grok 4.6 and received further reinforcement learning on difficult, long-running tasks.
- Launch date: September 21, 2026, according to xAI.
- Primary uses: coding, agentic workflows and knowledge work, according to xAI.
- Model identifier: grok-4.7 on the xAI API, according to xAI’s release information.
- Inputs and outputs: text and images in, with text output, according to xAI’s published model details.
The launch followed several days of speculation about a new Grok version. Reports published before the announcement pointed to possible appearances in cloud-service quotas, but those reports did not establish a public release. xAI’s September 21 announcement supplied the first direct confirmation.
What capabilities does Grok 4.7 claim?
Grok 4.7 is designed to stay engaged with complex tasks for longer, check its own work and handle large amounts of context. Those claims come from xAI and describe the company’s intended use case rather than an independent performance verdict.
- Context window: 500,000 tokens, according to xAI’s model information.
- Reasoning controls: low, medium, high and xhigh settings, according to xAI’s release documentation.
- Default reasoning level: high, according to release details summarized by independent AI industry publications.
- Tools: function calling, web search, X search and code execution, according to xAI’s API documentation.
The model’s focus is practical rather than limited to chat. A coding agent can use tool calls, inspect files, revise code and run tests. A research workflow can retain more material in one session. The 500,000-token window also gives developers room to pass large repositories, technical documents or multi-step task histories without splitting every request.
How did the new model perform in published tests?
Early benchmark figures come mainly from xAI’s own launch materials and should be read as company-reported results. Independent coverage has repeated the figures, but the available reports do not establish that every test used identical prompts, tools or evaluation rules across competing systems.
- CursorBench 4.0: 46.3%, according to xAI’s reported benchmark results.
- DeepSWE v1.1: 71.0%, according to xAI’s reported benchmark results.
- EEBench: 64.0%, according to xAI’s reported benchmark results.
- HealthBench Professional: 56.7%, according to xAI’s reported benchmark results.
- Harvey Legal Agent Benchmark: 19.6%, according to xAI’s reported benchmark results.
These tests cover different abilities. Coding benchmarks measure software-engineering performance, while health and legal evaluations examine professional question answering or agent behavior. A single percentage cannot describe the model’s performance across every use case.
Users will also need to distinguish benchmark scores from production reliability. A system can produce a strong result under a controlled test and still require human review when it changes a live codebase, handles sensitive records or makes decisions in regulated fields.
Where can customers use Grok 4.7?
Grok 4.7 launched across developer products rather than as a single consumer-only update. xAI says customers can access it through Cursor and Grok Build, while API users can connect it to applications and coding agents.
- Cursor: available at launch, according to xAI and independent release coverage.
- Grok Build: available at launch, according to xAI.
- xAI API: available under the grok-4.7 name, according to xAI.
- Other routes: third-party coding harnesses, model routers and cloud platforms, according to xAI.
- GitHub Copilot: independent release reports said the model was added for Pro, Pro+, Max, Business and Enterprise plans on launch day.
Availability can vary by product, plan and region. API access also depends on account approval, endpoint support and a developer’s chosen integration. A model appearing in a routing service does not necessarily mean that every customer receives the same limits, latency or tool access.
How much does Grok 4.7 cost?
xAI priced the standard API model at $2 per million input tokens and $6 per million output tokens for prompts below 200,000 tokens, according to the company’s published pricing. Longer prompts use a higher price tier.
- Input: $2 per million tokens below the 200,000-token prompt threshold, according to xAI.
- Cached input: $0.50 per million tokens below that threshold, according to release documentation.
- Output: $6 per million tokens below the threshold, according to xAI.
- Longer prompts: $4 input, $1 cached input and $12 output per million tokens above 200,000 tokens, according to xAI’s listed rates.
Token pricing is only one part of the bill. Applications that run repeated tool calls, web searches or code execution can consume more tokens and generate extra service costs. Developers also need to account for storage, monitoring and the cost of human review.
What changed from Grok 4.6?
Grok 4.7 keeps the 500,000-token context window and the standard $2 input and $6 output rates associated with Grok 4.6 for shorter prompts, according to release reports. xAI instead presents the upgrade as a capability improvement aimed at harder tasks, stronger verification and longer-running work.
- Model scale: xAI describes Grok 4.7 as using a larger base model than Grok 4.6.
- Training: xAI says the newer system received extended reinforcement learning on harder and longer tasks.
- Self-checking: xAI says Grok 4.7 verifies its outputs more reliably.
- Pricing: the standard short-prompt API rates remain $2 per million input tokens and $6 per million output tokens, according to xAI.
The company also lists a faster version called Grok 4.7 Fast. Independent release coverage reported that the faster variant is available through Cursor and Grok Build at twice the standard token rates, while it was not listed as a public xAI API model at launch.
What should developers watch next?
The next test will come from real deployments. Developers will assess whether the model’s longer context and self-checking reduce debugging time, whether higher reasoning settings justify their latency and whether the system behaves consistently across different coding tools and cloud services.
- Independent benchmark replication could clarify how Grok 4.7 compares with rival systems.
- API documentation and regional rollout details will determine who can access each feature.
- Enterprise users will examine privacy, logging, retention and permission controls before wider deployment.
- Developers will compare the standard and Fast versions on cost, speed and task accuracy.
Grok 4.7 arrives as AI companies compete for developer workloads that involve more than text generation. The product’s success will depend on measurable software outcomes, predictable operating costs and the ability to keep human oversight in the loop when an automated agent changes code or handles sensitive professional work.


