Back to models
Chat Models

OpenAI

GPT-4.1 Mini

OpenAI non-reasoning model for instruction following and tool-assisted tasks, with a 1,047,576-token context window.

1M contextCompact

At a glance

Know the model before you prompt.

OpenAI API specifications
Context window
1,047,576 tokens
Maximum output
32,768 tokens
Inputs → output
Text + Images → Text
Knowledge cutoff
June 1, 2024

API model ID: gpt-4.1-mini

Capabilities & boundaries

What it supports. Where the limits are.

Tool support requires the appropriate API integration; a supported tool is not automatically active in every chat.

Supported API features and tools

  • Text and image input; text output
  • Streaming, function calling, and structured outputs
  • Predicted outputs and prompt caching
  • Web search, file search, code interpreter, and MCP through supported integrations
  • Fine-tuning capability, subject to current platform access restrictions

Before you choose

  • No native audio or video support.
  • Image generation is not a supported tool in this model’s reference; do not copy that capability from GPT-4.1 or GPT-4o Mini.
  • Fine-tuning capability is subject to platform restrictions: organizations that never fine-tuned cannot start new jobs, and inactive organizations lost access July 2, 2026. Remaining active customers lose new-job creation January 6, 2027; existing inference continues until the base model is deprecated.
  • Maximum output is 32,768 tokens, not the size of the input context. Long-document coverage still needs evaluation.

Non-reasoning model

GPT-4.1 Mini responds without a separate reasoning step. It has no configurable reasoning-effort levels. Give it a small set of explicit rules, a concrete output format, and representative examples; use a reasoning model comparison when the task needs deeper analysis.

  • Chat Completions, Responses, and Batch are supported. Built-in tools require the appropriate Responses integration.
  • The documented snapshot is gpt-4.1-mini-2025-04-14. Pin a snapshot when reproducibility matters.
  • The Assistants API retired August 26, 2026; its historical model-page listing is not a recommendation for a new integration.
  • Long-context rate limits apply above 128,000 input tokens. A rate-limit threshold is not a separate token price; usage limits also depend on your API tier.

Put it to work

Start with a more useful prompt.

Original examples from EZ Ai Assist. Adapt these to your task and the features available in your workspace.

Workflow 01

Extract a policy exception register

Preserve source references and separate explicit exceptions from missing guidance.

Read the supplied policy sections and build an exception register. For each explicit exception, return the rule, exception conditions, required approver, and exact section reference. Use not stated for missing fields. Keep conflicting clauses separate and flag them for review. Do not infer permissions from silence or treat a suggested practice as an approved exception.

Workflow 02

Make a narrowly scoped code edit

Define the permitted change before asking for a patch.

In the supplied code, add handling for the one edge case described below. Preserve the existing interface, dependencies, and behavior for all other inputs. Return a minimal proposed diff, explain how it satisfies the requirement, and supply test cases for the edge case and unchanged behavior. Do not claim to have run tests or modify unrelated code.

Workflow 03

Standardize product records

Use a fixed schema without filling gaps from general knowledge.

Normalize these product descriptions into a JSON array with name, stated_dimensions, stated_materials, and source_quote fields. Preserve units and exact product names. Use null when a required value is absent, and include a short source quote for every populated attribute. Do not infer dimensions from images or marketing language. Report ambiguous records separately for human review.

Developer reference

OpenAI API pricing

These are OpenAI API reference prices, not EZ Ai Assist subscription prices.

View EZ Ai Assist plans
Standard processing · USD per 1,000,000 tokens
Token typePrice
Input$0.40
Cached input$0.10
Output$1.60
  • These are Standard base-model prices, not training or fine-tuned inference rates.
  • Eligible cached input receives the listed discount. No separate long-context price is listed here; long-context throughput limits are a different constraint.
  • Batch and tool usage have their own pricing rules. Check the current pricing source before estimating total API spend.

Common questions

A few things worth knowing.

When is GPT-4.1 Mini worth evaluating?

Try it on bounded instruction-following tasks with measurable acceptance criteria: extracting fields, revising small code sections, or organizing supplied documents. Compare it with a reasoning model on difficult cases rather than assuming a smaller price means adequate accuracy.

Is the 1M context also its output limit?

No. The model reference specifies 1,047,576 tokens of context and a maximum 32,768-token output. Large inputs still require careful organization and spot checks; they do not guarantee every relevant passage will be retrieved.

Can I adjust reasoning effort?

No configurable reasoning step is documented for GPT-4.1 Mini. Use clearer instructions, examples, and output constraints rather than transferring effort settings from a GPT-5 or GPT-6 model.

Can it generate an image?

Its native output is text, and its own tool list does not include image generation. It can analyze image inputs. Do not confuse image understanding with the separate image-generation tools listed for other models.

Can a new customer fine-tune it?

The model has fine-tuning capability, but current platform access is restricted. New organizations and inactive fine-tuning customers cannot start jobs. Remaining active customers lose new-job creation January 6, 2027; this is not a scheduled base-model shutdown.

Which API should a developer use?

The reference supports Chat Completions, Responses, and Batch. The Assistants API has already retired. Use current endpoint documentation and confirm the tool and rate-limit requirements for your integration.

Are these capabilities included in my EZ Ai Assist plan?

The guide describes provider API capabilities and prices, not a subscription entitlement. Check the app for model access, supported inputs, and tools. Adapt the prompts to the features you have and verify their outputs.

Check the source

Official documentation

Specifications and API prices checked on . Example prompts and workflow advice are editorial guidance from EZ Ai Assist.

Same provider