Key Information

Pricing

View as Markdown

All prices are in USD. For per-model details, see the models page.

Text API Pricing

ModelContextInput / 1M tokensCached input / 1M tokensOutput / 1M tokens
grok-4.7 (< 200k prompt tokens)500k$2.00$0.50$6.00
grok-4.7 (≥ 200k prompt tokens)500k$4.00$1.00$12.00
grok-4.6 (< 200k prompt tokens)500k$2.00$0.50$6.00
grok-4.6 (≥ 200k prompt tokens)500k$4.00$1.00$12.00
grok-4.5 (< 200k prompt tokens)500k$2.00$0.30$6.00
grok-4.5 (≥ 200k prompt tokens)500k$4.00$0.60$12.00
grok-4.3 (< 200k prompt tokens)1M$1.25$0.20$2.50
grok-4.3 (≥ 200k prompt tokens)1M$2.50$0.40$5.00
grok-4.20-0309-reasoning (< 200k prompt tokens)1M$1.25$0.20$2.50
grok-4.20-0309-reasoning (≥ 200k prompt tokens)1M$2.50$0.40$5.00
grok-4.20-0309-non-reasoning (< 200k prompt tokens)1M$1.25$0.20$2.50
grok-4.20-0309-non-reasoning (≥ 200k prompt tokens)1M$2.50$0.40$5.00
grok-build-0.1 (< 200k prompt tokens)256k$1.00$0.20$2.00
grok-build-0.1 (≥ 200k prompt tokens)256k$2.00$0.40$4.00
grok-4.20-multi-agent-0309 (< 200k prompt tokens)1M$1.25$0.20$2.50
grok-4.20-multi-agent-0309 (≥ 200k prompt tokens)1M$2.50$0.40$5.00

Prices shown per million tokens. Models listed with two rows use long context pricing: requests whose prompt reaches the listed token threshold are billed at the higher rate for all tokens in the request.

Imagine Pricing

ModelCost
grok-imagine-image-2.0$0.04 / image
grok-imagine-image$0.02 / image
grok-imagine-image-quality$0.05 / image
grok-imagine-video-1.5$0.080 / sec
grok-imagine-video$0.050 / sec

Voice Pricing

ModeCost
Speech to Speech (grok-voice-think-fast-2.0)$0.08 / min ($4.80 / hr) audio
$0.004 / text input
Speech to Text$0.10 / hr (REST), $0.20 / hr (Streaming)
Text to Speech$15.00 / 1M chars

Tools Pricing

Requests which make use of xAI provided server-side tools are priced based on two components: token usage and server-side tool invocations. Since the agent autonomously decides how many tools to call, costs scale with query complexity.

Token Costs

All standard token types are billed for the model used in the request:

  • Input tokens: Your query and conversation history

  • Reasoning tokens: Agent's internal thinking and planning

  • Completion tokens: The final response

  • Image tokens: Visual content analysis (when applicable)

  • Cached prompt tokens: Prompt tokens that were served from cache rather than recomputed

Tool Invocation Costs

ToolTool NameDescriptionCost / 1k Calls
Web Searchweb_searchSearch the internet and browse web pages$5
X Searchx_searchSearch X posts, user profiles, and threads$5 / 1k posts, $10 / 1k profiles
Code Executioncode_execution, code_interpreterRun Python code in a sandboxed environment$5
Image Generationimage_generationGenerate and edit imagesImagine API rates
File Attachmentsattachment_searchSearch through files attached to messages$10
Collections Searchcollections_search, file_searchQuery your uploaded document collections (RAG)$2.50
Image Understandingview_imageAnalyze images found during Web Search and X Search*Token-based
X Video Understandingview_x_videoAnalyze videos found during X Search*Token-based
Remote MCP ToolsSet by MCP serverConnect and use custom MCP tool serversToken-based
† All tool names work in the Responses API. In the gRPC API (Python xAI SDK), code_interpreter and file_search are not supported.
* Only applies to images and videos found by search tools — not to images passed directly in messages.

X Search is billed per item fetched rather than per call: every post returned by a search or thread fetch, including parent and quoted posts, counts toward the post rate, and every profile returned by a user search counts toward the profile rate.

For the view image and view x video tools, you will not be charged for the tool invocation itself but will be charged for the image tokens used to process the image or video.

Image Search is part of Web Search and is billed at the standard Web Search rate.

For Remote MCP tools, you will not be charged for the tool invocation but will be charged for any tokens used.

For more information on using Tools, please visit our guide on Tools.


Batch API Pricing

The Batch API lets you process large volumes of requests asynchronously at a discount to standard pricing. The size of the discount varies by model. Batch requests are queued and processed in the background, with most completing within 24 hours.

Real-time APIBatch API
Token pricingStandard ratesDiscounted rates (varies by model)
Response timeImmediate (seconds)Typically within 24 hours
Rate limitsPer-minute limits applyRequests don't count towards rate limits

The batch discount applies to all token types — input tokens, output tokens, cached tokens, and reasoning tokens. Batch discounts by model:

20% off standard rates

  • grok-4.3

  • grok-4.20-0309-reasoning

  • grok-4.20-0309-non-reasoning

  • grok-4.20-multi-agent-0309

Models not listed above have no batch discount.

To see a model's resulting batch prices, toggle "Show batch API pricing" on its detail page. Models that accept Batch with no discount show N/A.


Priority Processing Pricing

Priority Processing gives text requests higher scheduling priority for lower latency. Priority requests are billed at a 2x premium over standard rates.

StandardPriority
Token pricingStandard rates2x standard rates
Response timeStandard scheduling priorityHigher scheduling priority

The 2x multiplier applies to all token types — input, output, cached, and reasoning. Prompt caching discounts are applied before the multiplier.

You are only billed at the priority rate when the response confirms "service_tier": "priority". If the request is served at the default tier instead, standard rates apply.


Grok 4.7 Fast pricing (Cursor and Grok Build only)

Grok 4.7 Fast is the same Grok 4.7 model served on faster infrastructure, at twice the standard token rates. It is available only through Cursor and Grok Build; it is not available on the public xAI API, and Grok Build's free tier does not include it.

Prompt tokensInputCached inputOutput
Below 200k$4.00 / 1M$1.00 / 1M$12.00 / 1M
Above 200k$6.00 / 1M$1.50 / 1M$18.00 / 1M

Long-context rates apply once a request's prompt exceeds 200k tokens. Cursor bills its own fast variant through your Cursor plan.


US Regional Endpoint Pricing

Requests sent to the US regional endpoint, https://us.api.x.ai/v1, run inference in the United States; their token usage is billed at 1.1x the global token rates, a 10% premium.

Global endpointUS regional endpoint
Base URLhttps://api.x.ai/v1https://us.api.x.ai/v1
Token pricingStandard rates1.1x standard rates
ModelsAll models available to your teamCurrently grok-4.7 and grok-4.6 only

For grok-4.7 this is $2.20 / $0.55 / $6.60 per 1M tokens (input / cached input / output) below 200k prompt tokens, and $4.40 / $1.10 / $13.20 above. The 1.1x multiplier applies to input, output, and cached input tokens, including long-context rates. Prompt caching discounts are applied before the multiplier. See the Regional Endpoints documentation for the scope of the US processing and storage guarantee.


Files and Collections Pricing

Files and collections stored on the xAI platform are billed based on the amount of storage used.

ResourceRate
File storage$0.025 / GiB / day
Collection storage$0.10  / GiB / day

Download Costs

Downloading data from files and collections is charged at a flat rate based on the amount of data transferred:

ResourceRate
File downloads$0.20 / GiB downloaded
Collection downloads$0.20 / GiB downloaded

You can view and manage your files and collections through the xAI console or the xAI API.


Usage Guidelines Violation Fee

When your request is deemed to be in violation of our usage guideline by our system, we will still charge for the generation of the request.

For violations that are caught before generation in the Responses API, we will charge a $0.05 usage guideline violation fee per request.


Billing and Availability

Your model access might vary depending on various factors such as geographical location, account limitations, etc.

For how the bills are charged, visit Manage Billing for more information.

For the most up-to-date information on your team's model availability, visit Models Page on xAI Console.


Last updated:September 21, 2026