Tool support requires the appropriate API integration; a supported tool is not automatically active in every chat.
Supported API features and tools
Streaming responses
Function calling and structured outputs
Text and image input
Web search, file search, code interpreter, and MCP
Before you choose
No native audio or video support; native output is text.
Fine-tuning is not supported.
Image generation is not listed among this model’s supported Responses tools; do not inherit GPT-5’s tool list.
The gpt-5-mini-2025-08-07 API snapshot is scheduled to retire December 11, 2026.
Choose the reasoning effort
minimallowmedium · defaulthigh
The GPT-5 family documentation establishes minimal, low, medium, and high, with medium as default. For a well-defined task, compare a lower effort against your accuracy target; do not assume the Mini name means reasoning is disabled.
Unsupported settings: none, xhigh, max.
Chat Completions, Responses, and Batch are supported. Responses tool availability does not mean every tool is exposed in EZ Ai Assist.
272,000 tokens is the documented maximum input within the 400,000-token context; maximum output is 128,000 tokens.
The documented snapshot is gpt-5-mini-2025-08-07. API rate limits depend on usage tier; the free API tier is not supported.
Put it to work
Start with a more useful prompt.
Original examples from EZ Ai Assist. Adapt these to your task and the features available in your workspace.
Workflow 01
Normalize a product description
Apply a fixed schema without filling gaps with invented details.
Normalize each supplied product description into the fields name, category, material, dimensions, and stated limitations. Use only explicit source text, preserve units, and put null for missing values. Include a short evidence quote for each populated field. Flag contradictory measurements for review instead of choosing one. Return one record per product in the order provided.
Workflow 02
Draft replies from a support policy
Keep repeated communication tied to an approved source.
Draft a concise support reply for each ticket using only the attached policy excerpts. Address the customer’s stated problem, cite the applicable policy internally in a separate note, and avoid promising refunds, timelines, or account changes that the policy does not authorize. If the evidence is insufficient, ask one focused clarification and mark the case for human review.
Workflow 03
Check a release-note draft
Review a small artifact against a concrete list of delivered changes.
Compare this release-note draft with the merged-change summaries below. Identify unsupported claims, missing user-visible changes, and technical jargon that could confuse the intended audience. Suggest a concise rewrite using only verified changes. Keep breaking changes and required user actions explicit, and list anything that needs confirmation before publication.
Developer reference
OpenAI API pricing
These are OpenAI API reference prices, not EZ Ai Assist subscription prices.
These are Standard GPT-5 Mini prices, not GPT-5.4 Mini or GPT-5.6 Terra rates.
Eligible cached input has a separate discount. No additional long-context tier is listed in the reviewed table.
Measure reasoning/output usage as well as input. Tool charges and other service tiers can change the total; recheck the source before budgeting.
Common questions
A few things worth knowing.
What kinds of tasks suit GPT-5 Mini?
OpenAI describes it as a cost-efficient GPT-5 option for well-defined tasks and precise prompts. The examples here focus on normalization, policy-grounded replies, and bounded editing. Validate accuracy on your own cases instead of treating low cost as evidence of suitability.
What replaces this Mini snapshot?
The June 11, 2026 deprecation notice recommends GPT-5.6 Terra for gpt-5-mini-2025-08-07, scheduled to shut down December 11, 2026. Compare the replacement’s behavior, tools, and prices before switching an existing workflow.
Does Mini have the same knowledge cutoff as GPT-5?
No. This model’s documented cutoff is May 31, 2024, while GPT-5’s is September 30, 2024. Provide current evidence for time-sensitive tasks; a search tool is only available when the integration actually enables it.
Can it generate images?
It can accept images, but image generation is not listed in its supported Responses tools. Do not infer the capability from another GPT-5 family member. Its native output is text and its available app tools must be checked separately.
Which reasoning settings are supported?
Minimal, low, medium, and high are documented for the original GPT-5 family, with medium as default. None, xhigh, and max are not listed for this model. Evaluate lower effort against a clear accuracy threshold for repeated tasks.
How much input can I send?
The documented maximum input is 272,000 tokens, within a 400,000-token total context and a 128,000-token maximum output. A large context is not a guarantee of perfect extraction; test missing values, contradictions, and boundary cases.
Are these rates included in EZ Ai Assist?
No. The listed per-token rates are direct OpenAI API reference prices. EZ Ai Assist plans, allowances, model availability, and controls are separate. Verify the app’s supported workflow before using any example at scale.
Check the source
Official documentation
Specifications and API prices checked on . Example prompts and workflow advice are editorial guidance from EZ Ai Assist.