Grok

grok-4-fast-reasoning

xAIToken-based
Create API Key

xAI 历史 Grok 模型名称。官方 API 已于 2026 年 5 月 15 日将该名称重定向到 Grok 4.3;本站实际模型版本以接入渠道为准。

text
Starting price
Input / Output · 1M
Context
—
Maximum input window
Modalities
→

Pricing by Supplier

xAI
-50%
xAI 官方
Input$0.2$0.1/ 1M
Output$0.5$0.25/ 1M
Cache Read$0.05$0.025/ 1M

Capabilities / Supported modalities

StreamingSystem promptFunction callingToolsJSON modeStructured outputPrompt cachingReasoning
Input
Output

Provider & data privacy

Provider
xAIDocs
Tokenizer
Grok tokenizer (BPE)
License
Proprietary (commercial)Proprietary
Data retention48 daysNot used for upstream training by default

Performance

About grok-4-fast-reasoning

xAI 历史 Grok 模型名称。官方 API 已于 2026 年 5 月 15 日将该名称重定向到 Grok 4.3;本站实际模型版本以接入渠道为准。

Use cases and prompting

Starting points for evaluation; supported inputs and options are listed in API access.

Use cases to explore

  • Draft and revise text with explicit audience, tone and format requirements.
  • Summarize supplied documents and compare answers against the original sources.

Practical tips

Separate instructions from source material, describe the desired output and provide an example. Test tool use or structured output only when listed as supported.

Example prompt

Summarize the following document in five bullet points. Separate confirmed facts from open questions, cite the relevant passages and do not invent missing information. Document: [paste your text]

API access

Code samples

RequestPOST/v1/chat/completions
Example request
Parameters
ParameterTypeDefault / rangeDescription
reasoning_effort
enum
=medium
Controls how much the model thinks before answering
max_completion_tokens
integer>= 1Maximum tokens including hidden reasoning tokens
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
user
string—End-user identifier for abuse monitoring

Replace <YOUR_API_KEY> with the API key from your token settings.

Authentication

All requests must include Authorization: Bearer <TOKEN> header. Anthropic-formatted endpoints accept the x-api-key header instead.

Generate tokens from the Tokens page; you can scope them to specific models, groups, IPs, and rate-limits.

Supported parameters

Generation parameters
ParameterTypeDefault / rangeDescription
reasoning_effort
enum
=medium
Controls how much the model thinks before answering
max_completion_tokens
integer>= 1Maximum tokens including hidden reasoning tokens
stop
array—Up to 4 strings that stop generation
seed
integer—Deterministic sampling seed (best-effort)
stream
boolean
=false
Stream tokens via Server-Sent Events
response_format
object—Force JSON object or schema-conforming output
tools
array—Tool / function declarations the model may call
tool_choice
string
autononerequired
Tool-choice policy or specific tool name
user
string—End-user identifier for abuse monitoring

Rate limits

SupplierRPMTPMRPD
xAI530214K11K

RPM = requests per minute, TPM = tokens per minute, RPD = requests per day. Limits apply per token group.

Frequently asked questions about grok-4-fast-reasoning

What is grok-4-fast-reasoning?

xAI 历史 Grok 模型名称。官方 API 已于 2026 年 5 月 15 日将该名称重定向到 Grok 4.3;本站实际模型版本以接入渠道为准。

How do I call grok-4-fast-reasoning?

Create an API key with access to grok-4-fast-reasoning, then use the exact model ID and a supported endpoint from the API access section. Request fields depend on the selected endpoint.

How is grok-4-fast-reasoning priced?

Pricing depends on the selected provider group and the model billing unit. The current input, output, request, or media prices are shown on this page before sign-up.

How should I evaluate grok-4-fast-reasoning for my project?

Separate instructions from source material, describe the desired output and provide an example. Test tool use or structured output only when listed as supported.