EZ AI ASSIST MODEL DIRECTORY

234 Models.One Interface.

Access every major AI model for chat, image generation, video generation, and music generation from one unified EZ Ai Assist workspace.

143

Chat Models

Frontier chat, reasoning, coding, research, and agentic models.

46

Image Tools

Prompt-to-image engines for creative and production visual work.

32

Video Models

Text-to-video and cinematic generation models across leading providers.

13

Music Models

Prompt-to-song, sound, voice, and audio generation models.

234 results

Browse the EZ Ai Assist model catalog.

Each card links to a static marketing detail page. The directory does not trigger real model selection or app workflows.

chat

OpenAI

GPT-6.1 Sol

GPT-6.1 Sol is OpenAI’s reasoning model for complex coding and professional work, with image input and a 1.05M-token context window.

ReasoningCoding
NewGPT-6
chat

OpenAI

GPT-6 Sol

GPT-6 Sol is an OpenAI model for complex coding and tool-assisted workflows, with image input and configurable reasoning.

ReasoningCoding
NewGPT-6
chat

OpenAI

GPT-6 Luna

GPT-6 Luna is an efficient OpenAI model for focused, high-volume tasks, with image input, structured outputs, and configurable reasoning.

FastEfficient
chat

OpenAI

GPT-6 Astra

GPT-6 Astra is an OpenAI model for complex reasoning, coding, and research, with image input and a 1.05M-token context window.

ReasoningCoding
New
chat

OpenAI

GPT-5.6 Sol

Best for Coding

GPT-5.6 Sol is an OpenAI model for professional reasoning and coding, with image input, configurable reasoning effort, and a 1.05M-token context window.

1M+ contextCoding
New
chat

OpenAI

GPT-5.6 Terra

Best Overall

OpenAI model balancing reasoning and cost for everyday work, with image input, structured outputs, and configurable reasoning effort.

1M+ contextCoding
New
chat

OpenAI

GPT-5.6 Luna

Fastest

OpenAI model for cost-sensitive, high-volume tasks, with image input, structured outputs, and configurable reasoning effort.

1M+ contextCoding
New
chat

OpenAI

GPT-5.5

OpenAI reasoning model for complex professional work and coding, with image input, tool support, and a large context window.

1M+ contextCoding
New
chat

OpenAI

GPT-5.5 Pro

OpenAI reasoning model for demanding analysis, with a 1,050,000-token context window and text/image input.

1M+ contextReasoning
Deprecated
chat

OpenAI

GPT-5.3 Codex

Best for Coding

OpenAI coding specialist for repository work and tools. Deprecated; API shutdown is scheduled for April 1, 2027.

CodingTool use
New
chat

OpenAI

GPT-5.4

OpenAI model for professional work, coding, and tool-assisted tasks, with a 1,050,000-token context window.

1M contextFlagship
New
chat

OpenAI

GPT-5.4 Mini

Best Overall

Efficient OpenAI model for coding, image understanding, and tool-assisted workflows with a 400,000-token context window.

BalancedExtended context
Deprecated
chat

OpenAI

GPT-5.4 Nano

Fastest

OpenAI model for classification, extraction, and ranking. Deprecated; API shutdown is scheduled for April 1, 2027.

FastSimple tasks
New
chat

OpenAI

GPT-5.4 Pro

Best for Research

OpenAI reasoning model for difficult analysis via the Responses API, with a 1,050,000-token context window.

Deep reasoningCritical tasks
Deprecated
chat

OpenAI

GPT-5.1

OpenAI model for coding and configurable reasoning. Deprecated; API shutdown is scheduled for April 1, 2027.

CodingAgentic
chat

OpenAI

GPT-5.2 Thinking

OpenAI model for document analysis, coding, and professional reasoning, available through the API as gpt-5.2.

ReasoningKnowledge work
chat

OpenAI

GPT-5.2 Pro

OpenAI reasoning model for complex professional analysis via the Responses API, with a 400,000-token context window.

ReasoningKnowledge work
Deprecated
chat

OpenAI

o3

OpenAI reasoning model for technical and visual analysis. Its deprecated API snapshot retires December 11, 2026.

Reasoning
Deprecated
chat

OpenAI

o3-pro

OpenAI model for extended reasoning through Responses, without streaming. Its deprecated API snapshot retires December 11, 2026.

ReasoningExtended thinking
Deprecated
chat

OpenAI

GPT-5

OpenAI model for coding and configurable reasoning. Its deprecated API snapshot retires December 11, 2026.

FlagshipMultimodal
Deprecated
chat

OpenAI

GPT-5 Mini

OpenAI GPT-5 variant for well-defined, cost-sensitive tasks. Its deprecated API snapshot retires December 11, 2026.

Lower costGeneral purpose
Deprecated
chat

OpenAI

GPT-5 Nano

OpenAI GPT-5 variant for focused summarization and classification. Its deprecated API snapshot retires December 11, 2026.

FastHigh volume
Deprecated
chat

OpenAI

GPT-5 Pro

OpenAI model with high-only reasoning for complex analysis. Its deprecated API snapshot retires December 11, 2026.

Deep analysisReasoning
chat

OpenAI

GPT-4.1

OpenAI non-reasoning model for instruction following, coding, and long documents, with a 1,047,576-token context window.

Long contextCoding
chat

OpenAI

GPT-4.1 Mini

OpenAI non-reasoning model for instruction following and tool-assisted tasks, with a 1,047,576-token context window.

1M contextCompact
chat

OpenAI

GPT-4o Mini

OpenAI small model for focused text and image-input tasks, with structured outputs and a 128,000-token context window.

MultimodalFast
chat

OpenAI

ChatGPT Latest

OpenAI’s rolling chat-latest alias for the latest Instant model used in ChatGPT, with text and image input.

Rolling aliasChatGPT
New
chat

Anthropic

Claude Sonnet 5.5

Anthropic model for coding and everyday professional work, with adaptive thinking, image input, and a 1M-token context window.

Adaptive thinking1M context
New
chat

Anthropic

Claude Opus 5.5

Claude Opus 5.5 is Anthropic’s model for demanding coding and analysis, with always-on adaptive thinking and optional API Fast Mode.

Adaptive thinkingFast Mode
Legacy
chat

Anthropic

Claude Opus 5

Best for Coding

Anthropic’s active legacy Opus model for complex coding and analysis, with adaptive thinking and a 1M-token context window.

CodingComplex reasoning
Legacy
chat

Anthropic

Claude Sonnet 5

Best Overall

Anthropic’s active legacy Sonnet model for coding and general work, with adaptive thinking and a 1M-token context window.

BalancedGeneral purpose
chat

Anthropic

Claude Fable 5.1

Claude Fable 5.1 is Anthropic’s model for demanding reasoning and long-running agentic work, with always-on adaptive thinking.

Adaptive thinkingLong-horizon work
Legacy
chat

Anthropic

Claude Fable 5

Best for Research

Anthropic’s active legacy Fable model for difficult reasoning and long-running work, with always-on adaptive thinking.

ResearchComplex reasoning
Retired
chat

Anthropic

Claude Opus 4.1

Historical Anthropic model for coding and reasoning. Retired on Anthropic-operated platforms August 5, 2026; partner availability can differ.

ReasoningFlagship
Legacy
chat

Anthropic

Claude Opus 4.8

Best for Research

Anthropic’s active legacy Opus model with optional adaptive thinking, image input, and a 1M-token context.

1M contextAgentic
Legacy
chat

Anthropic

Claude Opus 4.7

Best for Research

Anthropic’s legacy Opus model with adaptive thinking, five effort levels, and a 1M-token context for supplied material.

1M contextAgentic
Legacy
chat

Anthropic

Claude Sonnet 4.6

Best for Coding

Anthropic’s active legacy Sonnet model for coding and analysis, with optional adaptive thinking and a 1M-token context.

1M contextCoding
Legacy
chat

Anthropic

Claude Opus 4.6

Best for Research

Anthropic’s active legacy Opus model with a 1M-token context and adaptive thinking for established coding and analysis workflows.

1M contextDeep reasoning
Deprecated
chat

Anthropic

Claude Sonnet 4.5

Anthropic model with manual extended thinking. Deprecated; retirement on Anthropic-operated platforms is scheduled for November 30, 2026.

Reasoning
Legacy
chat

Anthropic

Claude Opus 4.5

Best for Writing

Anthropic’s active legacy Opus model with manual extended thinking, three effort levels, and a 200K-token context.

WritingAgentic
chat

Anthropic

Claude Haiku 4.5

Anthropic’s efficiency-oriented Haiku model for bounded tasks, with optional extended thinking, image input, and a 200K-token context.

FastNear-frontier
New
chat

Google

Gemini 3.8 Flash

Google’s Gemini 3.8 Flash supports multimodal understanding, three thinking levels, and tool-assisted workflows with promotional API pricing.

FastMultimodal
chat

Google

Gemini 3.7 Flash

Google’s Gemini 3.7 Flash combines multimodal inputs, configurable thinking, and structured tool workflows for repeatable everyday tasks.

Chat
Preview
chat

Google

Gemini 3.1 Pro

Best for Coding

Google’s Gemini 3.1 Pro Preview supports complex synthesis, coding, and multimodal analysis with a 1M-token input limit.

ReasoningCoding
Retired
chat

Google

Gemini 3 Pro

Historical guide to Google’s retired Gemini 3 Pro Preview API, with specifications, migration checks, and the Gemini 3.1 Pro replacement.

Chat
Preview
chat

Google

Gemini 3.1 Pro (Custom Tools)

Best for Coding

Google’s Gemini 3.1 Pro Preview variant favors custom tools over bash; evaluate tool selection and answer quality for your workflow.

Tool useAgentic
Legacy
chat

Google

Gemini 3.6 Flash

Best Overall

Google’s earlier Gemini 3.6 Flash supports multimodal workflows and four thinking levels; the recommended replacement for Gemini 3 Flash Preview.

1M contextAgentic
Legacy
chat

Google

Gemini 3.5 Flash

Best Overall

Google’s earlier Gemini 3.5 Flash supports multimodal analysis, thinking controls, and tools, with its own model-specific API rates.

FastAgentic
Preview
chat

Google

Gemini 3 Flash

Best Overall

Google’s Gemini 3 Flash Preview supports multimodal inputs and tool workflows, with separate audio rates and a Gemini 3.6 Flash migration path.

FlashValue
New
chat

Google

Gemini 3.5 Flash-Lite

Google’s Gemini 3.5 Flash-Lite supports high-volume multimodal tasks, configurable thinking, and document parsing with a 1M-token input limit.

FastEfficient
New
chat

Google

Gemini 3.1 Flash-Lite

Best Value

Google’s Gemini 3.1 Flash-Lite supports translation, extraction, and lightweight multimodal workflows, with minimal thinking by default.

ValueHigh volume
Legacy
chat

Google

Gemini 2.5 Flash-Lite

Best Value

Google’s stable Gemini 2.5 Flash-Lite supports budget-sensitive multimodal tasks with optional thinking. Direct API access is restricted to prior users.

BudgetMultimodal
Legacy
chat

Google

Gemini 2.5 Pro

Best for Coding

Google’s stable Gemini 2.5 Pro supports complex reasoning and coding with a 1M-token input limit. Direct API access is restricted to prior users.

ReasoningCoding
Legacy
chat

Google

Gemini 2.5 Flash

Best Overall

Google’s stable Gemini 2.5 Flash combines multimodal input and adjustable thinking budgets. Direct API access is restricted to prior users.

FastReasoning
chat

Perplexity

Sonar

Perplexity’s lightweight web-grounded Sonar model for focused questions, with citations and a documented transition to the Agent API.

SearchLightweight
chat

Perplexity

Sonar Pro

Perplexity’s Sonar Pro supports deeper search and a 200K-token context. Explore its reference pricing and Agent API transition.

CitationsReasoning
chat

Perplexity

Sonar Reasoning Pro

Best for Research

Perplexity’s Sonar Reasoning Pro combines multi-step analysis with web retrieval. Distinct from the retired sonar-reasoning API identifier.

ResearchReasoning
chat

Perplexity

Sonar Deep Research

Perplexity’s Sonar Deep Research synthesizes broad web research into reports, with separate search, citation, and reasoning charges.

Deep researchSearch
New
chat

Meta

Muse Spark 1.3

Meta’s latest Muse Spark for coding and multi-step tool workflows, with a 1M-token context and max reasoning on the Standard tier.

ReasoningCoding
chat

Meta

Muse Spark 1.2

Meta’s previous Muse Spark version supports text, image, video, audio, and PDF inputs for long-context reasoning and tool-assisted work.

MultimodalReasoning
chat

Meta

Muse Spark 1.1

Meta’s original Muse Spark version supports multimodal understanding and reasoning with a 1M-token context and Standard-tier API pricing.

MultimodalReasoning
New
chat

xAI

Grok 4.7

Coding and knowledge-work model with 500K context, image input, and configurable reasoning effort.

Reasoning controlsReasoning
chat

xAI

Grok 4.6

Coding and agentic model with 500K context, structured outputs, and four reasoning effort levels.

Chat
New
chat

xAI

Grok 4.5

Best Overall

Coding and engineering model with 500K context, image input, tool calling, and configurable reasoning.

CodingAgentic
chat

xAI

Grok 4.3

Best Overall

1M-context model for instruction-driven work and tool calling, with configurable reasoning that can be disabled.

1M contextInstruction following
chat

xAI

Grok Build 0.1

Best for Coding

Coding-focused model with 256K context, image input, function calling, and structured outputs.

CodingAgentic
chat

xAI

Grok 4.20 Reasoning

Best for Research

Reasoning-enabled Grok 4.20 with 1M context, image input, tool calling, and structured outputs.

Reasoning1M context
chat

xAI

Grok 4.20 Non-Reasoning

Fastest

Non-reasoning Grok 4.20 for direct-response workflows, with 1M context, image input, and tool calling.

FastCost-efficient
chat

xAI

Grok 4.20 Multi-Agent

Best for Research

Beta research model coordinating 4 or 16 agents, with 1M context and a leader that synthesizes the answer.

Multi-agentDeep research
New
chat

DeepSeek

DeepSeek V4.1 Flash

DeepSeek V4.1 Flash combines text and vision with thinking controls, long context, and peak/off-peak direct API pricing.

FastTime-aware pricing
New
chat

DeepSeek

DeepSeek V4 Flash

Best Overall

V4 Flash 0731 on OpenRouter, with long-context text workflows and hosting limits distinct from newer direct API aliases.

1M contextVersioned model
New
chat

DeepSeek

DeepSeek V4 Pro

Best for Research

V4 Pro 0423 on OpenRouter for long-context reasoning, with provider-specific pricing and direct API version caveats.

Deep reasoning1M context
New
chat

Groq

GPT-OSS 120B

Best Overall

OpenAI's larger open-weight reasoning model on Groq, with tool use, structured outputs, and a 131K context window.

500 tok/sOpen model
New
chat

Groq

GPT-OSS 20B

Fastest

Compact open-weight reasoning on Groq for focused text tasks, with structured outputs, tool use, and automatic caching.

1000 tok/sLightweight
Retired
chat

Groq

Compound Mini

Historical single-tool Groq system retired September 21, 2026. Explore its former workflow and migration considerations.

Historical systemMigration guide
Retired
chat

Groq

Compound

Historical multi-tool Groq system retired September 21, 2026, with no direct replacement listed in its shutdown notice.

Historical systemMigration guide
Enterprise access
chat

Groq

LLaMA 3.3 70B Versatile

Meta's 70B text model on Groq with 131K context. Free/developer access ended; committed-spend enterprise contracts are exempt.

Text workflows131K context
Enterprise access
chat

Groq

LLaMA 3.1 8B Instant

Compact LLaMA text model on Groq. Free/developer access ended; eligible enterprise contracts retain access under Groq's exception.

CompactText workflows
Enterprise access
chat

Groq

Qwen3.6 27B

Groq-hosted text and vision model with thinking modes and 131K context. Self-service access ended; an enterprise exception remains.

MultimodalThinking modes
New
chat

Mistral AI

Mistral Medium 3.5

Best Overall

Text-and-vision model for coding and agent workflows, with 256K context and configurable reasoning.

AgenticCoding
New
chat

Mistral AI

Mistral Small 4

Fastest

Hybrid text-and-vision model combining instruction following, reasoning and coding with 256K context.

HybridCoding
New
chat

Mistral AI

Ministral 3 8B

Fastest

Compact 8B text-and-vision model for extraction and constrained workflows, with 256K context.

TextVision
New
chat

Mistral AI

Ministral 3 3B

Best Value

Small 3B text-and-vision model for narrow classification and extraction tasks, with 256K context.

EfficientVision
Flagship
chat

Mistral AI

Mistral Large 3

Best for Research

Open-weight mixture-of-experts model for text, vision and tool-assisted work, with 256K context.

Open-weightMultimodal
Vision
chat

Mistral AI

Ministral 3 14B

14B text-and-vision model for document and image workflows, with 256K context and open weights.

TextVision
New
chat

MiniMax

MiniMax M3

Best Overall

Multimodal coding model with text, image and video input and up to 1M context, subject to endpoint capacity.

1M contextMultimodal
chat

MiniMax

MiniMax M2.7

Best Overall

Text-focused coding and office-work model with 204,800 context, tool calling and always-on thinking.

CodingReasoning
Fast
chat

MiniMax

MiniMax M2.7 Highspeed

Fastest

Faster-serving M2.7 endpoint with 204,800 context, always-on thinking and separate API rates.

FastCoding
Legacy
chat

MiniMax

MiniMax M2-her

Best Overall

Legacy dialogue and role-play model with published 64K context and dedicated conversation settings.

DialogueRole-play
Legacy
chat

MiniMax

MiniMax M2.5

Legacy coding and tool-workflow model with 204,800 context and always-on thinking.

CodingReasoning
Legacy
chat

MiniMax

MiniMax M2.5 Highspeed

Legacy faster-serving M2.5 variant with 204,800 context and distinct input/output API rates.

FastCoding
Legacy
chat

MiniMax

MiniMax M2.1

Best for Coding

Legacy multilingual programming and refactoring model with 204,800 context and always-on thinking.

CodingMultilingual
Legacy
chat

MiniMax

MiniMax M2.1 Highspeed

Legacy faster-serving M2.1 endpoint for multilingual code workflows, with 204,800 context.

FastMultilingual
Legacy
chat

MiniMax

MiniMax M2

Original legacy M2 for text, reasoning and tool workflows, with 204,800 context and qualified output limits.

AgenticReasoning
New
chat

Kimi

Kimi K3

Best for Coding

Multimodal Kimi with 1M context, always-on thinking and adjustable reasoning effort for coding and knowledge work.

1M contextCoding
New
chat

Kimi

Kimi K2.7 Code

Best for Coding

Coding-focused Kimi with 256K context, visual inputs and always-on thinking for multi-step software work.

256K contextCoding
chat

Kimi

Kimi K2.7 Code HighSpeed

Best for Coding

Faster-serving K2.7 Code endpoint with the same model behavior, 256K context and separately priced API usage.

High speedCoding
chat

Kimi

Kimi K2.6

Best Overall

Multimodal Kimi with 256K context and optional thinking for coding, document analysis and visual review.

256K contextOptional thinking
Retired
chat

Kimi

Kimi K2.5

Historical multimodal Kimi with 256K context, retired on the Kimi platform August 31, 2026. Migration reference for K3.

Historical256K context
New
chat

Cohere

Command A+

Best for Research

Multimodal Cohere MoE with hybrid reasoning, 48-language support and a 128K context window.

Vision48 languages
chat

Cohere

Command A

Best for Research

Text-focused Cohere model for grounded answers and tool workflows, with 256K context and 8K output.

RAG256K context
chat

Cohere

Command A Reasoning

Best for Research

Hybrid text reasoning with optional thinking budgets, 256K context and 32K output for evidence-heavy workflows.

ReasoningResearch
chat

Cohere

Command A Vision

Best for Research

Text and image understanding for documents, tables and charts, with up to 20 images per request and no tool use.

VisionDocuments
chat

Cohere

Command A Translate

Best Overall

Specialized text translation across 23 languages, with separate 8K input and 8K output limits.

Translation23 languages
chat

Cohere

Command R7B

Best Overall

Compact 7B text model for grounded answers and tool workflows, with 128K context and low API token rates.

RAGTool use
chat

Cohere

Command R+

Best for Research

Dated August 2024 model for grounded answers and multi-step tools; distinct from deprecated Command R+ aliases.

RAGCitations
chat

Cohere

Command R

Dated August 2024 text model for cost-conscious grounded answers, citations and tool workflows with 128K context.

RAGTool use
New
chat

Z.ai

GLM-5.3

Best for Coding

Text-only flagship for complex coding and agent workflows with 1M context and always-on reasoning.

1M contextReasoning
New
chat

Z.ai

GLM-5.3-Flash

Best Value

Native multimodal GLM-5 model for images, video, files and text with 1M context and low API token rates.

Multimodal1M context
New
chat

Z.ai

GLM-5.3-FlashX

Fastest

Faster-serving GLM-5.3-Flash tier with native multimodal inputs, 1M context and separately priced API access.

Fast servingMultimodal
New
chat

Z.ai

GLM-5.2

Best Overall

Text model for long-horizon engineering with 1M context, 128K output and configurable thinking.

1M contextLong horizon
New
chat

Z.ai

GLM-5.1

Best Overall

Text model for sustained engineering and iterative agent workflows with 200K context and optional thinking.

200K contextAgentic
chat

Z.ai

GLM-5-Turbo

Fastest

OpenClaw-oriented text model with 200K context; confirm current Turbo endpoint access and API pricing.

FastCoding
chat

Z.ai

GLM-5V-Turbo

Best Overall

Multimodal coding model with 200K context for screenshots and visual workflows; current API rates remain unverified.

VisionFast
chat

Z.ai

GLM-5

Best Overall

Text reasoning and coding model with 200K context, 128K output and optional thinking.

200K contextReasoning
chat

Z.ai

GLM-4.6V

Best Value

Vision model for images, video and documents with 128K context, 32K output and native function calling.

VisionDocuments
chat

Z.ai

GLM-4.7

Best Overall

Previous-generation text model with 200K context, turn-level thinking and coding-oriented tool workflows.

ReasoningFlagship
chat

Z.ai

GLM-4.6

Best for Coding

Text model with 200K context, hybrid thinking and explicitly enabled streaming tool-call arguments.

200K contextCoding
chat

Z.ai

GLM-4.5

Best Overall

General-purpose text model with 128K context, 96K output and hybrid thinking for coding and analysis.

BalancedGeneral purpose
chat

Z.ai

GLM-4.5-Air

Fastest

Lightweight GLM-4.5 text tier with 128K context, hybrid thinking and lower API token rates.

128K contextLightweight
chat

Z.ai

GLM-4.7-FlashX

Best Value

Paid lightweight GLM-4.7 tier for fast text and coding workflows with 200K context and 128K output.

High throughputLow cost
chat

Z.ai

GLM-4.7-Flash

Best Value

Lightweight text model with 200K context and currently free direct API token rates, subject to provider limits.

Free tierLightweight
chat

Z.ai

GLM-4.5-X

Premium high-speed GLM-4.5 text tier with 128K context, hybrid thinking and separate API rates.

Fast serving128K context
chat

Z.ai

GLM-4.5-AirX

Faster-serving lightweight GLM-4.5 tier with 128K context and higher API rates than standard Air.

LightweightFast serving
chat

Z.ai

GLM-4.5V

Visual reasoning model for images, video and documents with 64K context, 16K output and switchable thinking.

Vision64K context
chat

Z.ai

GLM-4.6V-FlashX

Best Value

Paid lightweight vision tier with 128K context, 32K output and native function calling at low API token rates.

VisionLow cost
chat

Z.ai

GLM-4-32B-0414-128K

Best Value

Older 32B text model with 128K context, 16K output and low direct API rates for instruction-following tasks.

Text128K context
chat

Z.ai

GLM-5 Plus

Best for Coding

Enhanced GLM-5 with deeper reasoning and extended output.

ReasoningExtended output
chat

Z.ai

GLM-5 Air

Fastest

Lightweight GLM-5 for ultra-fast and cost-effective work.

LightweightCost-effective
New
chat

Xiaomi MiMo

MiMo-V2.5-Pro

Best Overall

Flagship model with 1M context, 128K output, deep thinking, and web search.

1M contextWeb search
New
chat

Xiaomi MiMo

MiMo-V2.5

Omni full-modal understanding with 1M context and 128K output.

1M contextOmni
New
chat

Alibaba Qwen

Qwen3.7-Max

Best Overall

Closed-weights flagship with 1M context, native extended thinking, and agentic coding.

1M contextAgentic coding
New
chat

Alibaba Qwen

Qwen3.7-Plus

Best Overall

Recommended default with native multimodal support, 1M context, and agentic coding.

1M contextMultimodal
chat

Alibaba Qwen

Qwen3.6-Plus

Best Overall

Balanced production workhorse with 1M context and built-in tools.

1M contextBuilt-in tools
chat

Alibaba Qwen

Qwen3.6-Flash

Fastest

Low-cost, high-speed tier with 1M context and full tool support.

1M contextTool support
New
chat

Alibaba Qwen

Qwen3 Max

Best Overall

Trillion-parameter flagship for top reasoning, coding, and multilingual performance.

Trillion paramsMultilingual
chat

Alibaba Qwen

Qwen3 Coder Plus

Best for Coding

Agentic coding specialist and SWE-Bench leader with long-context repository reasoning.

CodingRepo reasoning
chat

Alibaba Qwen

Qwen3 VL Plus

Best for Writing

Native vision-language model with spatial reasoning and 1M-context video analysis.

VisionVideo analysis
New
chat

Alibaba Qwen

Qwen3 Next 80B

Fastest

Efficient 80B MoE with 3B active parameters for fast, capable, budget-friendly work.

80B MoEFast
New
chat

Alibaba Qwen

Qwen 3.6

Best Overall

Latest flagship with hybrid thinking and 128K context.

Hybrid thinking128K context
chat

Alibaba Qwen

Qwen 2.5 Coder

Best for Coding

Coding specialist 32B parameter model optimized for code.

Coding32B
chat

Alibaba Qwen

Qwen 2.5 72B

Large general-purpose model with strong all-around performance.

General purpose72B
chat

Alibaba Qwen

QwQ 32B

Best for Research

Reasoning specialist with chain-of-thought capabilities.

Reasoning32B
image

Google

Nano Banana

Fast AI image generation for rapid visual ideation.

Fast imageCreative
Latest
image

Google

Nano Banana 2

Gemini 3.1 Flash image generation for fast, capable visual work.

GeminiFast
New
image

Google

Nano Banana 2 Lite

Nano Banana 2 Lite is a lighter, faster Google image model for creative work in the Nano Banana family.

FastImage generation
image

Google

Nano Banana Pro

Advanced Gemini 3 Pro image generation for higher fidelity visual output.

ProHigh fidelity
Latest
image

Google

Imagen 4 Fast

Google's latest Imagen 4 model optimized for fast image creation.

FastImagen
image

Google

Imagen 4

Google's flagship Imagen 4 model for high-quality prompt-to-image generation.

FlagshipImage
image

Google

Imagen 4 Ultra

Google's highest quality image model for premium visual generation.

UltraHigh quality
New
image

OpenAI

GPT Image 2.5 Sunburst

GPT Image 2.5 Sunburst is an OpenAI image generation model for turning creative ideas into visuals.

Image generationCreative
New
image

OpenAI

GPT Image 2.5 Flare

GPT Image 2.5 Flare offers another OpenAI image generation option for exploring visual concepts and creative directions.

Image generationCreative
image

OpenAI

ChatGPT Image Latest

ChatGPT Image Latest is an OpenAI image model in the EZ Ai Assist catalog.

Image
New
image

OpenAI

GPT Image 2

OpenAI's latest flagship image generation model.

FlagshipImage
image

OpenAI

GPT Image 2 (2026-04-21)

GPT Image 2 (2026-04-21) is an OpenAI image model in the EZ Ai Assist catalog.

Image
Latest
image

OpenAI

GPT Image 1.5

OpenAI flagship image model for versatile creative workflows.

FlagshipCreative
image

OpenAI

GPT Image 1

OpenAI advanced image generation for prompt-driven visual creation.

ImageCreative
image

OpenAI

GPT Image 1 Mini

Cost-efficient multimodal image model for lighter visual workloads.

Cost-efficientMultimodal
image

OpenAI

DALL-E 3

OpenAI reliable image generation model for creative prompt-to-image work.

ReliableCreative
New
image

xAI

Grok Imagine Pro

Higher quality xAI image generation through Grok Imagine.

ProCreative
image

xAI

Grok Imagine

xAI's latest creative image generation model.

CreativeImage
Latest
image

FLUX

FLUX.2 Pro

32B parameter flagship image model with best-in-class detail.

32BDetail
image

FLUX

FLUX.2 Flex

Typography specialist for flexible, prompt-aware image generation.

TypographyFlexible
Latest
image

FLUX

FLUX.2 Turbo

Ultra-fast FLUX.2 generation for rapid creative output.

FastFLUX.2
image

FLUX

FLUX.1 Schnell

Ultra-fast generation model for speed-focused image workflows.

FastSchnell
image

Stability AI

Stable Image Ultra

Flagship model for highest-quality Stable image generation.

FlagshipHigh quality
image

Stability AI

Stable Image Core

Balanced quality and speed for everyday image generation.

BalancedSpeed
image

Stability AI

SD 3.5 Large

8B parameter flagship Stable Diffusion model for image generation.

8BFlagship
image

Stability AI

SD 3.5 Large Turbo

4-step distilled model for high-speed Stable Diffusion generation.

TurboFast
image

Stability AI

SD 3.5 Medium

2.5B parameter model balancing quality and efficiency.

2.5BBalanced
New
image

ByteDance

Seedream 4.5

Next-generation text-to-image model optimized for high-quality output.

Text-to-imageQuality
New
image

ByteDance

Seedream 5.0 Lite

State-of-the-art text-to-image model in a lighter configuration.

LiteText-to-image
New
image

ByteDance

Seedream 4.5 Sequential

Multi-image set generation with unified sequential visual consistency.

SequentialMulti-image
New
image

ByteDance

Seedream 5.0 Lite Sequential

Cost-efficient multi-image set generation for sequential creative workflows.

SequentialLite
image

ByteDance

Seedream 4.0

ByteDance flagship image generation model for creative production.

FlagshipImage
image

ByteDance

Seedream 4.0 Sequential

Multi-image set generation with consistent sequential output.

SequentialMulti-image
New
image

Ideogram

Ideogram V3 Quality

Highest-quality V3 model for realistic, detailed image generation.

QualityRealistic
New
image

Ideogram

Ideogram V3 Balanced

Balanced V3 model for realistic output with practical speed.

BalancedRealistic
New
image

Ideogram

Ideogram V3 Turbo

Fastest V3 model for photorealistic image generation.

TurboPhotorealistic
New
image

Ideogram

Ideogram V3 Transparent

Transparent-background PNG generation for production design assets.

TransparentPNG
image

Ideogram

Ideogram V2

State-of-the-art inpainting and strong prompt adherence.

InpaintingPrompt adherence
image

Ideogram

Ideogram V2 Turbo

Faster V2 model for inpainting and prompt-driven design output.

TurboInpainting
image

Ideogram

Ideogram V2A

Precise prompt-following model for high-fidelity image generation.

Prompt fidelityImage
image

Ideogram

Ideogram V2A Turbo

Fastest V2A model for high-fidelity text and image output.

TurboText
New
image

Alibaba Qwen

Qwen Image Max

Flagship Qwen image model for highest-quality visual generation.

FlagshipQuality
New
image

Alibaba Qwen

Qwen Image 2512

Latest Qwen image model with enhanced prompt adherence.

LatestPrompt adherence
image

Alibaba Qwen

Qwen Image

20B MMDiT next-generation text-to-image model.

20BText-to-image
image

Alibaba Qwen

Jib Mix Qwen

Qwen variant tuned for natural, flexible image generation.

QwenCreative
New
image
HI

Higgsfield

Marketing Studio

Marketing Studio from Higgsfield focuses on marketing-oriented visual creation and creative production.

MarketingCreative production
New
video

Google

Gemini Omni Flash

Gemini Omni Flash brings fast, multimodal video generation to Google's creative model lineup.

Video generationMultimodal
Latest
video

Google

VEO 3.1

Google's flagship video model with native 1080p generation.

1080pFlagship
Fast
video

Google

VEO 3.1 Fast

Faster VEO 3.1 model for native 1080p video workflows.

Fast1080p
Budget
video

Google

VEO 3.1 Lite

Affordable VEO 3.1 model for 720p and 1080p video generation.

BudgetVideo
video

Google

VEO 3

Google's previous flagship video generation model.

VideoFlagship
video

Google

VEO 3 Fast

Faster VEO 3 model for synchronized video generation.

FastVideo
video

Google

VEO 2

High-quality image-to-video and text-to-video generation.

Image-to-videoText-to-video
video

Runway

Gen-4 Aleph

Gen-4 Aleph is a Runway video model in the EZ Ai Assist catalog.

Video
video

Runway

Gen-4 Turbo

Gen-4 Turbo is a Runway video model in the EZ Ai Assist catalog.

Video
New
video

Runway

Gen-4.5 Turbo

Faster, cheaper variant of Runway Gen-4.5.

TurboFast
Latest
video

Runway

Gen-4.5

Runway flagship model for professional video generation.

FlagshipVideo
video

Runway

Gen-3a Turbo

Budget-friendly fast video generation model.

BudgetFast
video

Runway

Runway Act Two

Runway Act Two is a Runway video model in the EZ Ai Assist catalog.

Video
Latest
video

Pika

Pika 2.2

Enhanced creative video generation model.

CreativeVideo
video

Pika

Pika 1.5

Quick stylized video generation.

StylizedQuick
Latest
video

Luma AI

Ray 2

Luma flagship video model for best-in-class cinematic generation.

CinematicFlagship
video

Luma AI

Ray 1.5

Fast cinematic video generation.

FastCinematic
New
video

Kling

Kling 2.5 Turbo Pro

Kling 2.5 Turbo Pro for fast, high-quality video generation.

TurboHigh quality
video

Kling

Kling V3 Pro

Kling 3.0 Pro for fast, high-quality video generation.

ProHigh quality
video

Kling

Kling V3 Standard

Kling 3.0 Standard for reliable video generation.

StandardVideo
Latest
video

MiniMax

Hailuo 2.3

MiniMax flagship cinematic video generation.

CinematicFlagship
Fast
video

MiniMax

Hailuo 2.3 Fast

Same quality as Hailuo 2.3 with faster turnaround.

FastCinematic
video

MiniMax

Hailuo 0.2

Hailuo 0.2 is a MiniMax video model in the EZ Ai Assist catalog.

Video
New
video

xAI

Grok Imagine Video 2

xAI's latest text-to-video model through Grok Imagine.

Text-to-videoGrok
video

xAI

Grok Imagine Video

xAI's original text-to-video generation model.

Text-to-videoCreative
New
video

ByteDance

Seedance 2.0

Seedance 2.0 is a ByteDance video generation model for motion-led creative work.

Video generationMotion
New
video

ByteDance

Seedance 2.5

Seedance 2.5 is the newer Seedance generation in the ByteDance catalog for video creation and motion workflows.

Video generationMotion
New
video

ByteDance

Seedance 1 Pro

ByteDance flagship video model for cinematic generation.

FlagshipCinematic
New
video

ByteDance

Seedance 1.5 Pro

Seedance 1.5 Pro for cinematic, lively video generation.

CinematicPro
New
video

ByteDance

Seedance 1.5 Pro Fast

Faster Seedance 1.5 Pro for quicker video workflows.

FastPro
New
video

ByteDance

Seedance V1 Pro Fast

Coherent multi-shot text-to-video generation with faster output.

Multi-shotFast
New
video

ByteDance

Seedance V1 Pro 480p

Budget 480p text-to-video generation for lightweight workflows.

480pBudget
New
music

Suno

Suno v5.5

Latest Suno model for richer vocals, longer songs, custom lyrics, and styles.

Up to 4 minLyrics
music

Suno

Suno v5

Fast Suno music generation with superior expression and full songs with vocals.

Up to 4 minLyrics
New
music

MiniMax

MiniMax Music 2.6

Latest MiniMax via WaveSpeed with vocals, lyrics, and instrumentation.

Up to 4 minLyrics
music

MiniMax

MiniMax Music 2.5

Flagship full songs with vocals, lyrics, and instrumentation.

Up to 4 minLyrics
music

MiniMax

MiniMax Music 1.5

Vocals and instrumentation for complete prompt-to-song workflows.

Up to 4 minLyrics
New
music
GL

Google Lyria

Lyria 3.5

Lyria 3.5 is a Google Lyria model for music generation and creative audio work.

Music generationGoogle
New
music
GL

Google Lyria

Lyria 002

Lyria 002 is a Google Lyria music generation option for exploring musical ideas through audio.

Music generationAudio
New
music
GL

Google Lyria

Lyria 3 Pro

Google DeepMind flagship music model for rich, full-length tracks.

Up to ~2 minMusic
New
music
GL

Google Lyria

Lyria 3 Clip

Fast short-form Lyria 3 clips ideal for stings and loops.

Up to 30sLoops
New
music

Stability AI

Stable Audio 3

Latest Stable Audio via WaveSpeed for high-fidelity music and sound.

Up to ~3 minAudio
music

Stability AI

Stable Audio 2.5

High-quality music and sound from prompts.

Up to ~3 minAudio
New
music
SO

Sonilo

Sonilo Text-to-Music

Quick instrumental generation for backing tracks and loops.

Up to 3 minInstrumental
music

ElevenLabs

ElevenLabs Music

ElevenLabs Music turns text prompts into music for creative audio projects.

Up to 4 minMusic