Tool support requires the appropriate API integration; a supported tool is not automatically active in every chat.
Supported API features and tools
Streaming responses and multimodal understanding
Function calling and structured outputs
Context caching and Batch API
Code execution, Google Search grounding, and URL context
File search and Google Maps grounding
Computer use (Preview), plus Flex and Priority processing
Before you choose
No native image or audio generation; these models return text.
Live API is not supported; audio input is not a real-time voice session.
Tool access, upload limits, and permissions depend on the integration. Validate outputs against the original evidence.
A model-specific knowledge cutoff was not verified in the reviewed sources. Do not infer it from a release date or another Flash model.
This is Gemini 3.5 Flash, not the separate Flash-Lite or Flash Cyber models. Their pricing and capabilities are not interchangeable.
Gemini thinking levels
minimallowmedium · defaulthigh
Medium is the documented default, with minimal, low, medium, and high supported. Minimal does not guarantee that thinking is disabled. Match the level to the task, then measure answer quality, response time, and output tokens; the levels are not fixed token budgets.
Use gemini-3.5-flash. No shutdown date is announced in the reviewed Gemini Developer API lifecycle table.
For stateful or multi-turn tools, follow Google’s documented conversation-state and thought-signature requirements. Supported capabilities still need to be enabled by the integration.
Put it to work
Start with a more useful prompt.
Original examples from EZ Ai Assist. Adapt these to your task and the features available in your workspace.
Workflow 01
Normalize a reference packet
Extract comparable records from uneven source material.
Convert the reference documents below into a consistent table of item name, stated purpose, requirements, and limitations. Cite a source section for each nonempty field. Preserve meaningful differences in terminology and mark missing values as unknown. Add a separate list of contradictions for human review rather than resolving them by guesswork.
Workflow 02
Draft a grounded handoff
Turn supplied notes into a practical next-step summary.
Create a handoff note from this project history and the current issue list. Separate completed work, pending decisions, blockers, and proposed next steps. Include an evidence reference and owner only where the notes provide one. Keep the summary concise and flag ambiguous deadlines. Do not present proposed actions as completed.
Workflow 03
Review multimodal consistency
Look for discrepancies between visible material and written claims.
Compare the supplied product images with the written specification. Identify visible matches, contradictions, and requirements that cannot be checked from images alone. Cite the image label and specification section for each finding. Prioritize issues by user impact and propose a verification step without inventing measurements or hidden behavior.
Developer reference
Google API pricing
These are Google Gemini Developer API reference prices, not EZ Ai Assist subscription prices.
These are Standard rates for Gemini 3.5 Flash, not the promotional rates for newer Flash models.
Context-cache storage is an additional $1.00 per million token-hours. Thinking tokens are included in billed output.
Batch, Flex, and Priority have separate rates and availability. Grounding and other tools may add charges; consult the current model-specific pricing section.
Common questions
A few things worth knowing.
Is Gemini 3.5 Flash the same as Flash-Lite?
No. Flash, Flash-Lite, and Flash Cyber are different model variants. Select the exact model ID and consult its own limits and prices rather than combining claims from the family.
Why is the knowledge cutoff marked Not verified?
The reviewed sources did not establish a model-specific cutoff for this entry. A release date or a sibling model’s cutoff would not be a reliable substitute. Supply current evidence for time-sensitive questions.
Are the newer Flash promotional prices used here?
No. The current Standard rates listed for 3.5 Flash are $1.50 input, $0.15 cached input, and $9.00 output per million tokens. The guide keeps these separate from other Flash promotions.
Does Legacy mean I must migrate immediately?
No shutdown date was announced for this model in the reviewed lifecycle table. Legacy identifies an earlier generation here. Evaluate newer models and watch provider notices, but do not treat the label alone as an API deadline.
Can I use these prompts in EZ Ai Assist?
Yes—adapt these original editorial examples to the inputs and controls available in your workspace. Share only material you are authorized to provide. They are starting points, not benchmarks or a promise of a particular result.
Does the input limit guarantee accurate long-document answers?
No. The 1,048,576-token input limit is capacity, not a recall guarantee. Organize sources with names and sections, request citations, and check them. The separate maximum output is 65,536 tokens; actual requests must also fit integration limits.
Are these the prices of my EZ Ai Assist plan?
No. The table describes direct Google API token usage. EZ Ai Assist subscriptions and app capabilities are separate. Use our pricing page for current plans and check the app for enabled models and controls.
Check the source
Official documentation
Specifications and API prices checked on . Example prompts and workflow advice are editorial guidance from EZ Ai Assist.