Back to models
NewChat Models

xAI

Grok 4.5

Best Overall

Coding and engineering model with 500K context, image input, tool calling, and configurable reasoning.

CodingAgentic

At a glance

Know the model before you prompt.

xAI API specifications
Context window
500,000 tokens
Maximum output
Not verified
Inputs → output
Text + Images → Text
Knowledge cutoff
February 1, 2026

API model ID: grok-4.5

Capabilities & boundaries

What it supports. Where the limits are.

Tool support requires the appropriate API integration; a supported tool is not automatically active in every chat.

Supported API features and tools

  • Text and image input with a 500,000-token context
  • Coding and engineering workflows with configurable reasoning
  • Function calling and structured outputs
  • Responses API and Chat Completions, with tool support dependent on the endpoint
  • Web search, X search and code execution through configured provider tools

Before you choose

  • Reasoning cannot be disabled; Batch API is not supported.
  • The model card lists xhigh, but the dedicated reasoning guide says 4.5 treats xhigh as high. Do not present it as a separate deeper mode.
  • The dated overview gives a February 1, 2026 knowledge cutoff. Later events require supplied evidence or configured search; an output ceiling was not verified.
  • The text output modality does not include native audio, video or image generation. Check generated code and tool actions before using them.

Choose the reasoning effort

lowmediumhigh · default

Reasoning cannot be disabled. Use low, medium or high as distinct effective levels. Although the model card lists xhigh, the dedicated reasoning guide says xhigh requests on Grok 4.5 are treated as high. This guide preserves that caveat instead of promising an extra level of reasoning.

Unsupported settings: none.

  • The Grok 4.5 overview documents Responses and Chat Completions. Provider-side tools need the appropriate endpoint and integration.
  • The overview recommends a prompt_cache_key on Responses, or x-grok-conv-id on Chat Completions, to improve cache routing. A cache hit is not guaranteed.
  • The published aliases include grok-4.5-latest and grok-build-latest. Use the explicit model ID when you need to communicate which model you evaluated.
  • Reasoning-model requests reject presencePenalty, frequencyPenalty and stop. Recheck provider documentation when migrating SDKs or endpoints.

Put it to work

Start with a more useful prompt.

Original examples from EZ Ai Assist. Adapt these to your task and the features available in your workspace.

Workflow 01

Repair a form validation bug

Constrain the fix and protect existing behavior.

Review this form component, its validation rules and the failing example. Identify the specific bug and propose the smallest patch that fixes it without changing the visual design or unrelated behavior. Explain empty, malformed and boundary-value cases. Add focused test cases and distinguish tests you actually ran from tests you recommend running.

Workflow 02

Check an algorithm implementation

Ask for concrete counterexamples and a bounded correction.

Check this algorithm against the specification below. Work through the supplied examples and construct small counterexamples for edge cases. Explain any correctness or complexity problem, then propose a minimal correction with a short proof sketch and a test matrix. If the specification is ambiguous, state the ambiguity before choosing an interpretation.

Workflow 03

Evaluate experiment evidence

Separate measured results from plausible explanations.

Using only the experiment plan, measurements and notes I provide, summarize the result and assess whether the evidence supports the original hypothesis. Identify missing controls, measurement uncertainty and alternative explanations. Recommend one next experiment that would most reduce uncertainty. Do not invent observations or treat correlation as proof of cause.

Developer reference

xAI API pricing

These are global xAI API reference prices, not EZ Ai Assist subscription prices.

View EZ Ai Assist plans
Standard processing · USD per 1,000,000 tokens
Token type< 200,000prompt tokens≥ 200,000prompt tokens
Input$2.00$4.00
Cached input$0.30$0.60
Output$6.00$12.00
Additional API fees · USD
UsageRate and unit
Web Search$5.00per 1,000 calls
X Search: posts$5.00per 1,000 posts fetched
X Search: profiles$10.00per 1,000 profiles fetched
Code execution$5.00per 1,000 calls
  • At 200,000 prompt tokens or more, doubled rates apply to the entire request. Grok 4.5's cached-input rate is different from 4.6 and 4.7.
  • Reasoning usage is billed, and supported tools incur separate charges when enabled. X Search bills per fetched post/profile, not per search call.
  • Batch API is not supported. Priority processing charges 2× standard rates when the response confirms the priority tier.
  • These are global standard API rates. Check the exact endpoint, cache usage and tool fees before estimating a workload; EZ Ai Assist subscription billing is separate.

Common questions

A few things worth knowing.

What work suits Grok 4.5?

Its documented focus is coding, agentic software and engineering tasks. Give it a specific objective, relevant files and a validation boundary. Maintaining an evaluated workflow is different from assuming it is the best choice for every new task.

What are its context and knowledge limits?

The card lists 500,000 tokens of context; the model overview lists a February 1, 2026 knowledge cutoff. A maximum output ceiling is not verified here. Current facts should come from supplied evidence or enabled search, not assumed training knowledge.

Does xhigh add a deeper reasoning mode?

The model card lists xhigh, but the dedicated reasoning guide states that Grok 4.5 maps it to high. Treat low, medium and high as distinct effective choices unless xAI resolves the discrepancy. High is the documented default, and reasoning cannot be disabled.

Can it work with screenshots and tools?

The API accepts images and text and lists function calling and structured outputs. Search and execution tools require configuration. Image understanding is not native image generation, and API capability does not establish that a tool is connected in the app.

Why is cached input priced differently?

The published rate is $0.30 per million cached-input tokens below 200,000 prompt tokens, and $0.60 at or above that threshold. Those are 4.5-specific rates, not the $0.50/$1 rates of 4.6 and 4.7. Only actual cache-served usage receives the discount.

Can I use Batch API or copy all OpenAI parameters?

The model card says Batch API is unsupported. Similar API shapes do not imply identical parameters: xAI reasoning models reject presencePenalty, frequencyPenalty and stop. Verify the SDK, endpoint and settings before integrating.

Are these features and costs guaranteed in EZ Ai Assist?

No. The guide documents xAI reference behavior and prices. Check EZ Ai Assist for model access, available controls and plan terms. The app may expose a different subset of tools or use different limits from the provider API.

Check the source

Official documentation

Specifications and API prices checked on . Example prompts and workflow advice are editorial guidance from EZ Ai Assist.

Same provider