Tool support requires the appropriate API integration; a supported tool is not automatically active in every chat.
Supported API features and tools
Streaming responses
Function calling and structured outputs
Text and image input with prompt caching
Web search, file search, image generation tools, code interpreter, and MCP
Before you choose
No native audio or video support; native output is text.
Fine-tuning is not supported.
The gpt-5-2025-08-07 API snapshot is scheduled to retire December 11, 2026.
Minimal reasoning reduces deliberation; it is not the same as a newer model’s none setting.
Choose the reasoning effort
minimallowmedium · defaulthigh
The original GPT-5 family supports minimal, low, medium, and high, with medium as the documented default. Test minimal for narrowly defined tasks and compare higher effort on difficult work. Do not carry over GPT-5.1 or GPT-6 effort settings.
Unsupported settings: none, xhigh, max.
Chat Completions, Responses, and Batch are supported. Built-in tools require a compatible Responses integration.
272,000 tokens is the documented maximum input, within the 400,000-token total context. The 128,000-token output allowance is not extra space beyond the total context.
The documented snapshot is gpt-5-2025-08-07. API rate limits depend on usage tier; the free API tier is not supported.
Put it to work
Start with a more useful prompt.
Original examples from EZ Ai Assist. Adapt these to your task and the features available in your workspace.
Workflow 01
Plan a focused refactor
Keep a code change bounded by observable behavior and tests.
Plan a refactor of this function using the supplied callers and tests. Identify responsibilities that can be separated without changing behavior, list the invariants that must remain true, and show a small sequence of edits. For each step, name the test that would catch a regression. Do not introduce new dependencies or assume unprovided callers behave a certain way.
Workflow 02
Extract requirements from meeting notes
Turn a messy discussion into traceable decisions and open questions.
Read these meeting notes and extract agreed requirements, tentative ideas, rejected options, and unresolved questions into separate sections. Preserve who said what when names are provided and cite the note or timestamp for every requirement. Do not promote a suggestion into a decision. Finish with a short list of clarifications needed before implementation can begin.
Workflow 03
Assess a dashboard discrepancy
Compare the visible interface with a supplied definition of each metric.
Compare this dashboard screenshot with the metric definitions and sample records below. Identify inconsistent labels, calculations, units, or time windows that the evidence supports. Cite the visual region and data fields involved. Separate a likely display issue from a confirmed calculation error, and propose three reproducible checks before changing the dashboard.
Developer reference
OpenAI API pricing
These are OpenAI API reference prices, not EZ Ai Assist subscription prices.
These are Standard GPT-5 rates, not GPT-5 Pro or a later GPT generation.
Cached-input pricing requires eligible cached tokens. No separate long-context tier is listed in the reviewed table.
Reasoning tokens contribute to output billing; tools and other processing tiers may have separate charges. Recheck current rates before estimating spend.
Common questions
A few things worth knowing.
Is this the original GPT-5 model?
Yes. The API identifier is gpt-5 and the documented snapshot is gpt-5-2025-08-07. It is distinct from GPT-5.1, GPT-5.2, the GPT-5.6 family, and GPT-5 Pro; their limits, settings, and prices should not be substituted here.
When does the GPT-5 snapshot retire?
OpenAI’s June 11, 2026 notice schedules gpt-5-2025-08-07 for shutdown on December 11, 2026 and recommends GPT-5.6 Sol. Test a migration with your real tasks before the deadline. App availability must be checked separately.
Is minimal the same as turning reasoning off?
No. Minimal is the lowest documented effort for this original GPT-5 family; none is not listed. Medium is the documented default. Choose an effort explicitly when benchmarking a workflow and check actual token usage.
Can I send 400,000 input tokens?
No. The model reference separately lists a 272,000-token maximum input inside a 400,000-token total context. Output and reasoning also consume the available context; the 128,000-token output ceiling is not an extra allowance beyond it.
Does JSON output mean the answer is correct?
No. Structured outputs can constrain the response shape, but facts and calculations still require validation. Provide a clear schema in an integration that supports it and independently check any result used to trigger actions.
Can it generate images, audio, or video?
Its native output is text, with text and image input. Image generation is a listed Responses tool and requires an integration. Native audio and video are not supported; EZ Ai Assist may expose a different set of tools.
How should I compare its cost with my plan?
The table is a direct OpenAI API reference, not an EZ Ai Assist subscription price. For API comparisons, measure input, eligible cached input, reasoning/output, and tool charges. Use the site’s pricing page for subscription plans.
Check the source
Official documentation
Specifications and API prices checked on . Example prompts and workflow advice are editorial guidance from EZ Ai Assist.