{"id":3118,"date":"2026-09-23T07:26:13","date_gmt":"2026-09-23T07:26:13","guid":{"rendered":"https:\/\/dmarketertayeeb.com\/blog\/grok-4-7\/"},"modified":"2026-09-23T07:34:06","modified_gmt":"2026-09-23T07:34:06","slug":"grok-4-7","status":"publish","type":"post","link":"https:\/\/dmarketertayeeb.com\/blog\/grok-4-7\/","title":{"rendered":"Grok 4.7: API Pricing, Fast Access and Benchmarks"},"content":{"rendered":"\n<p><a href=\"https:\/\/x.ai\/news\/grok-4-7\/\">xAI announced Grok 4.7 on September 21, 2026<\/a> for coding and knowledge work. For developers, the key is to separate where the model runs from what it costs: the direct xAI API lists <code>grok-4.7<\/code> at $2 per million uncached input tokens, $0.50 per million cached input tokens and $6 per million output tokens below the long-context tier. A prompt reaching the roughly 200,000-token long-context boundary moves the whole request to higher rates; xAI&#8217;s pages differ at exactly 200,000. Grok 4.7 Fast is available through Cursor and Grok Build, but not the public xAI API.<\/p>\n\n\n<p>Use the API table below for direct xAI calls, Cursor&#8217;s own plan and context rules inside Cursor, and the product&#8217;s current controls in Grok Build. xAI&#8217;s launch benchmarks are vendor-reported snapshots, not an independent or up-to-date ranking against every model available today.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">What Grok 4.7 is<\/h2>\n\n\n<p>The xAI API model ID is <code>grok-4.7<\/code>. The current model page lists a 500,000-token context window, text and image input, text output, and reasoning effort levels of low, medium, high and xhigh, with high as the default. xAI describes 4.7 as its most capable model for coding, agentic tasks and knowledge work; that is the vendor&#8217;s characterization, not a result from an independent test in this guide. <a href=\"https:\/\/docs.x.ai\/developers\/models\/grok-4.7\">Check xAI&#8217;s current model page<\/a> before building against a specific capability or limit.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">Choose the access surface before comparing prices<\/h2>\n\n\n<p>Grok 4.7 is listed across the xAI API, Cursor and Grok Build. These are different products with different entitlement and billing rules. A model router may also add its own rate or show a third-party route price, so do not treat every listing as the xAI API rate.<\/p>\n\n\n<figure class=\"wp-block-table\"><table>\n<thead><tr><th>Surface<\/th><th>What the current docs establish<\/th><th>Access and billing boundary<\/th><\/tr><\/thead>\n<tbody>\n<tr><td><strong>xAI API<\/strong><\/td><td>Model ID <code>grok-4.7<\/code>; 500K context; standard model only.<\/td><td>Use xAI&#8217;s global API rates below. Fast is not offered on the public API. The US regional endpoint costs 10% more.<\/td><\/tr>\n<tr><td><strong>Cursor<\/strong><\/td><td>Every paid plan lists Grok 4.7. Cursor bills it through the Cursor Models usage pool.<\/td><td>Start (India) is fixed at medium effort and non-Fast. Pro or higher allows effort selection and Fast. Cursor has its own long-context threshold above 256K.<\/td><\/tr>\n<tr><td><strong>Grok Build<\/strong><\/td><td>xAI&#8217;s launch page and Build changelog list Grok 4.7.<\/td><td>Fast is available, but the Grok Build free tier does not include it. Check Build&#8217;s current plan and usage controls for the session you intend to run.<\/td><\/tr>\n<tr><td><strong>Consumer Grok<\/strong><\/td><td>The sources checked here establish API, Cursor and Build access.<\/td><td>They do not establish that a consumer subscription includes Grok 4.7. Do not infer app-plan access from API availability.<\/td><\/tr>\n<\/tbody>\n<\/table><\/figure>\n\n\n<p>Cursor&#8217;s listed standard rates are $2 input, $0.50 cached input and $6 output per million tokens; Fast is $4, $1 and $12. Cursor says prompts above 256K use twice standard rates, or three times standard rates in Fast, up to its 500K window. Those are Cursor-specific rates and limits, not the xAI API&#8217;s approximately 200K long-context boundary. <a href=\"https:\/\/prod.cursor.com\/help\/models-and-usage\/grok-4-7\">Cursor&#8217;s Grok 4.7 plan guide<\/a> also says enterprise administrators manage model access.<\/p>\n\n\n<p>For Build access, controls and product behavior, see our <a href=\"https:\/\/dmarketertayeeb.com\/blog\/grok-build-access-devices-memory-2026\">Grok Build access guide<\/a>. The <a href=\"https:\/\/dmarketertayeeb.com\/blog\/grok-bot-cloud-computer-agent\">Grok Bot&#8217;s persistent cloud-computer workflow<\/a> is a separate product built around a model; it is not another name for the Grok 4.7 model itself.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">Direct xAI API pricing, including the long-context tier<\/h2>\n\n\n<p>xAI quotes direct API text prices per million tokens. Its pricing table labels long context at <strong>200,000 tokens or more<\/strong>, while model and release-note prose describes the higher tier as applying when a prompt exceeds 200,000. The docs therefore differ at exactly the boundary. The examples here use 150,000 and 220,000 prompt tokens, not exactly 200,000; check the live billing page before estimating a request at the boundary.<\/p>\n\n\n<figure class=\"wp-block-table\"><table>\n<thead><tr><th>Prompt size<\/th><th>Uncached input \/ cached input \/ output per 1M<\/th><th>Billing rule<\/th><\/tr><\/thead>\n<tbody>\n<tr><td>Below 200K<\/td><td>$2.00 \/ $0.50 \/ $6.00<\/td><td>Short-context rates apply.<\/td><\/tr>\n<tr><td>Long context (pricing table: \u2265200K)<\/td><td>$4.00 \/ $1.00 \/ $12.00<\/td><td>Once the prompt reaches the long-context threshold, xAI says all tokens in that request use the long-context rates.<\/td><\/tr>\n<\/tbody>\n<\/table><\/figure>\n\n\n<p>For example, a 150,000-token prompt with 130,000 uncached tokens and 20,000 cached tokens, plus 3,000 output tokens, costs about <strong>$0.288<\/strong> at the below-200K rates: (130,000 \u00d7 $2 + 20,000 \u00d7 $0.50 + 3,000 \u00d7 $6) \u00f7 1,000,000. A 220,000-token prompt with 200,000 uncached and 20,000 cached tokens, plus the same 3,000 output tokens, costs about <strong>$0.856<\/strong> at long-context rates: (200,000 \u00d7 $4 + 20,000 \u00d7 $1 + 3,000 \u00d7 $12) \u00f7 1,000,000. These illustrations exclude server-side tool invocation charges and any regional premium.<\/p>\n\n\n<p>The US regional endpoint, <code>https:\/\/us.api.x.ai\/v1<\/code>, applies a 1.1\u00d7 premium to token rates. xAI also lists separate charges for server-side tools such as web search, X search and code execution. If a workload uses those tools, add their invocation charges to token usage; the model table alone is not the request&#8217;s full cost. See <a href=\"https:\/\/docs.x.ai\/developers\/pricing\">xAI&#8217;s current API pricing<\/a>.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">Grok 4.7 Fast: where it is available and what the rate table says<\/h2>\n\n\n<p>xAI describes Fast as the same Grok 4.7 model on faster infrastructure. Its pricing page says Fast is available only through Cursor and Grok Build, not the public xAI API; the Build free tier excludes it. In Cursor, Fast is available on Pro or higher, not the fixed medium\/non-Fast Start plan.<\/p>\n\n\n<figure class=\"wp-block-table\"><table>\n<thead><tr><th>xAI Fast prompt band<\/th><th>Input \/ cached input \/ output per 1M<\/th><th>Scope<\/th><\/tr><\/thead>\n<tbody>\n<tr><td>Below 200K<\/td><td>$4.00 \/ $1.00 \/ $12.00<\/td><td>Fast rate table; Cursor and Grok Build only.<\/td><\/tr>\n<tr><td>Above 200K<\/td><td>$6.00 \/ $1.50 \/ $18.00<\/td><td>Long-context Fast rates; not a public API option.<\/td><\/tr>\n<\/tbody>\n<\/table><\/figure>\n\n\n<p>There is a mismatch within xAI&#8217;s pricing page: its prose calls Fast twice the standard token rates, but the published long-context row is 1.5 times the standard long-context row ($6\/$1.50\/$18 versus $4\/$1\/$12). Use the explicit surface-specific price table and recheck current terms; do not derive the long-context Fast rate by doubling the standard rate. Cursor separately publishes its own Fast price schedule and a different long-context cutoff.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">What the launch benchmarks show\u2014and what they do not<\/h2>\n\n\n<p>xAI&#8217;s September 21 announcement reported Grok 4.7 results across coding, engineering and knowledge-work benchmarks. The excerpt below keeps the published effort settings attached to the scores. These are <strong>xAI-reported results<\/strong>, not measurements from our own run.<\/p>\n\n\n<figure class=\"wp-block-table\"><table>\n<thead><tr><th>Benchmark<\/th><th>Grok 4.7<\/th><th>Other scores in xAI&#8217;s launch table<\/th><\/tr><\/thead>\n<tbody>\n<tr><td>CursorBench 4.0<\/td><td>46.3% (xHigh)<\/td><td>Grok 4.6 High: 40.4%; GPT-5.6 Sol Max: 41.7%; Fable 5.1 Max: 51.8%.<\/td><\/tr>\n<tr><td>DeepSWE v1.1<\/td><td>71.0% (High effort; marked with an asterisk)<\/td><td>Grok 4.6 High: 65.2%; GPT-5.6 Sol Max: 72.7%; Fable 5.1 Max: 70.0%.<\/td><\/tr>\n<\/tbody>\n<\/table><\/figure>\n\n\n<p>The announcement&#8217;s table labels the Grok 4.7 column xHigh, but the 71.0% DeepSWE result has an asterisk and a footnote saying that score used high effort. The comparison columns use different effort labels, and the table does not include every current model. OpenAI announced GPT-6 Sol the next day, September 22; that newer model is absent from xAI&#8217;s September 21 comparison. So the table is useful as a dated, vendor-selected snapshot, not a current neutral ranking or proof that 4.7 wins across coding work. For the older Grok 4.6 query and its own benchmark context, keep the separate <a href=\"https:\/\/dmarketertayeeb.com\/blog\/grok-4-6-pricing-benchmarks-availability\">Grok 4.6 pricing and benchmark guide<\/a>.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">A fair way to evaluate it on your own tasks<\/h2>\n\n\n<p>Before making Grok 4.7 a default, run a small comparison on the work your team actually sends to an agent. Keep the test set and execution setup fixed:<\/p>\n\n\n<ul class=\"wp-block-list\">\n<li>Choose a handful of representative tasks, such as a failing-test fix, a multi-file change and a bug investigation. Use the same repository state, task prompt, allowed tools and time limit for each model.<\/li>\n<li>Record the model surface and variant, reasoning effort, context length, tool calls and price schedule. A direct API call, a Cursor run and a Grok Build session are not interchangeable billing routes.<\/li>\n<li>Score accepted changes, test results, scope errors, human rework and elapsed time. Re-run tasks that are nondeterministic rather than treating one completion as a stable result.<\/li>\n<li>Compare cost per accepted task, including tool charges and rework, alongside latency and quality. A cheaper token rate or a vendor benchmark score alone cannot answer that decision.<\/li>\n<\/ul>\n\n\n<p>This is a proposed evaluation protocol, not a test we ran for this article. The evidence checked here confirms where Grok 4.7 is listed and what its published price tiers are; it does not establish which model will perform best on your repository or workflow.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\">Sources and method<\/h2>\n\n\n<p>Product and price details were checked against the <a href=\"https:\/\/x.ai\/news\/grok-4-7\/\">xAI launch announcement<\/a>, <a href=\"https:\/\/docs.x.ai\/developers\/models\/grok-4.7\">xAI model documentation<\/a>, <a href=\"https:\/\/docs.x.ai\/developers\/pricing\">xAI API pricing<\/a>, the <a href=\"https:\/\/docs.x.ai\/developers\/release-notes\">xAI release notes<\/a>, the <a href=\"https:\/\/prod.cursor.com\/help\/models-and-usage\/grok-4-7\">Cursor Grok 4.7 guide<\/a> and <a href=\"https:\/\/prod.cursor.com\/help\/models-and-usage\/available-models\">Cursor&#8217;s current model and pricing page<\/a>. The launch benchmark scores are xAI-reported. We did not run an independent model test or verify account-specific usage in Cursor or Grok Build.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Grok 4.7 is available through the xAI API, Cursor and Grok Build. Compare their access and pricing rules, read the launch benchmarks with their effort settings, and calculate the direct API cost around the long-context tier.<\/p>\n","protected":false},"author":1,"featured_media":3117,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[180,178],"tags":[294,393,300,315],"class_list":["post-3118","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai-news","category-artificial-intelligence","tag-ai-model-comparison","tag-ai-model-releases","tag-ai-models","tag-developer-tools","has-featured-image"],"_links":{"self":[{"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/posts\/3118","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/comments?post=3118"}],"version-history":[{"count":1,"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/posts\/3118\/revisions"}],"predecessor-version":[{"id":3119,"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/posts\/3118\/revisions\/3119"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/media\/3117"}],"wp:attachment":[{"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/media?parent=3118"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/categories?post=3118"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/dmarketertayeeb.com\/blog\/wp-json\/wp\/v2\/tags?post=3118"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}