Nano Banana Pro
Strong multilingual text, structured visual layouts, reference fusion, and precise natural-language edits.
nano-banana-proAPI price preview
19% below official
MODEL CATALOG
Search the model routes available through PixMind, compare capabilities and billing units, then copy the stable model ID into your integration.
76 results
Strong multilingual text, structured visual layouts, reference fusion, and precise natural-language edits.
nano-banana-proAPI price preview
19% below official
High-control image generation and editing for posters, products, infographics, and iterative design work.
gpt-image-2API price preview
17% below official
Commercial-grade visuals with strong composition, detail, reference consistency, and high-resolution output.
seedream-5.0-proAPI price preview
25% below official
Context-aware image editing that preserves subjects and style across targeted changes.
flux-kontext-proPricing to be announced
Fast text-to-image generation with readable typography, flexible ratios, and detailed output.
qwen-image-2.0Pricing to be announced
A balanced route for generation, local editing, multiple references, and production image workflows.
wan2.7-imagePricing to be announced
Distinctive art direction, polished aesthetics, and reference-driven visual exploration.
mj-v7API price preview
63% below official
PixMind's free image generation route for quick everyday text-to-image creation.
z-imageComing next
Pricing to be announced
A cost-efficient Nano Banana Pro route covering text-to-image and image editing with reference support.
nano-banana-pro-liteComing next
Pricing to be announced
Google's Nano Banana 2 high-speed image generation across 1K, 2K, and 4K resolutions.
nano-banana-2API price preview
31% below official
The original Nano Banana image generation route for versatile text-to-image and editing work.
nano-bananaComing next
Pricing to be announced
An economical Nano Banana 2 variant balancing speed, quality, and cost for everyday generation.
nano-banana-2-ecoAPI price preview
38% below official
Midjourney's anime and illustration-tuned model for stylized characters and vibrant artwork.
mj-niji7Coming next
Pricing to be announced
Midjourney V8.2 with refined aesthetics, stronger prompt adherence, and richer art direction.
mj-v8.2Coming next
Pricing to be announced
Midjourney V8.1 delivering elevated visual quality and consistency for polished creative output.
mj-v8.1Coming next
Pricing to be announced
ByteDance Seedream 5.0 commercial-grade image generation with strong composition and detail.
seedream-5.0Coming next
Pricing to be announced
Seedream 4.5 balanced image generation with reliable quality for everyday creative work.
seedream-4.5Coming next
Pricing to be announced
Seedream 4.0 image generation covering text-to-image and editing for production visuals.
seedream-4.0Coming next
Pricing to be announced
GPT Image 1.5 high-control generation and editing for design-heavy visuals.
gpt-image-1.5Coming next
Pricing to be announced
GPT Image 4o versatile image generation with precise prompt following and editing.
gpt-image-4oComing next
Pricing to be announced
An economical GPT Image 2 variant for cost-aware generation and editing at scale.
gpt-image-2-ecoComing next
Pricing to be announced
Wan 2.6 image generation and editing from Alibaba for multilingual creative workflows.
wan2.6-imageComing next
Pricing to be announced
Wan 2.7 Image Pro, Alibaba's higher-tier route for sharper, production-ready visuals.
wan2.7-image-proComing next
Pricing to be announced
Qwen Image 3.0 Pro with improved text rendering and high-resolution output.
qwen-image-3.0-proAPI price preview
50% below official
Qwen Image 2.0 Pro text-to-image generation with readable typography and detail.
qwen-image-2.0-proComing next
Pricing to be announced
xAI image generation for expressive, high-impact visuals across seven common aspect ratios and larger batches.
grok-imagine-2-ecoAPI price preview
15% below official
Fast Krea text-to-image generation with flexible aspect ratios, deterministic seeds, and 1K or 2K output.
krea-2-turboAPI price preview
56% below official
Grok Imagine 1.5 image generation with bold, expressive visual styling.
grok-imagine-1.5Coming next
Pricing to be announced
Recraft V4 vector generation for clean, scalable, print-ready artwork.
recraftv4_vectorComing next
Pricing to be announced
Recraft V4 Pro vector generation with premium quality for professional design.
recraftv4_pro_vectorComing next
Pricing to be announced
Recraft image-to-vector conversion that turns raster images into editable vectors.
recraft-vectorizeComing next
Pricing to be announced
Cinematic video generation with image guidance, shot control, first/last frames, and native audio.
veo-3.1API price preview
20% below official
Native 2K video with synchronized audio from text, images, first/last frames, and multimodal references.
minimax-h3API price preview
Multi-reference video creation with coordinated characters, scenes, motion, and sound.
seedance-2.0-proAPI price preview
20% below official
Native 30-second single-shot video with up to 50 multimodal references, 4K output, and region-level editing.
seedance-2.5API price preview
20% below official
High-fidelity image animation and camera or subject motion control for production shots.
kling-v3-motion-controlAPI price preview
20% below official
Versatile text, image, reference, and first/last-frame workflows through one video route.
wan2.7-videoAPI price preview
20% below official
Wan 3.0 video generation with multimodal references, synchronized audio, 2-30 second clips, and up to 1080p output.
wan3.0-videoAPI price preview
14% below official
Detailed video generation for longer creative shots with image guidance and audio-capable output.
sora-2-proAPI price preview
20% below official
Fast video ideation across text-to-video, image animation, transitions, effects, and sound.
pixverse-v5API price preview
20% below official
Reference-led character performance and expressive motion for short-form creative video.
happyhorse-1.0API price preview
20% below official
Kuaishou Kling 3.0 Turbo, a fast text-to-video and image-to-video route for quick drafts.
kling-3.0-turboAPI price preview
19% below official
ByteDance Seedance 1.5 Pro video generation with audio-synced output.
seedance-1.5-proAPI price preview
16% below official
Seedance 2.0 Fast, a quicker Seedance route for faster video turnaround.
seedance-2.0-fastComing next
Pricing to be announced
Seedance 2.0 Mini, a lightweight variant for efficient short video generation.
seedance-2-0-miniComing next
Pricing to be announced
OpenAI Sora 2 video generation for text-to-video and image-to-video creative shots.
sora-2Coming next
Pricing to be announced
sora-2-ecoComing next
Pricing to be announced
veo-3Coming next
Pricing to be announced
veo-3.0-fastComing next
Pricing to be announced
veo-3.1-ecoComing next
Pricing to be announced
Veo 3.1 Fast Eco, the most cost-optimized Veo route for fast, budget-friendly video.
veo-3.1-fast-ecoComing next
Pricing to be announced
PixVerse V6 video generation with text-to-video, image animation, and effects.
pixverse-v6Coming next
Pricing to be announced
wan2.6-i2vComing next
Pricing to be announced
wan2.6-t2vComing next
Pricing to be announced
Wan 2.6 I2V Flash, a rapid image-to-video variant for quick animation.
wan2.6-i2v-flashComing next
Pricing to be announced
Happy Horse 1.1, an upgraded route for reference-led character performance video.
happyhorse-1.1API price preview
Anthropic's Mythos-class route for autonomous knowledge work, long-form reasoning, and coding. Choose Fable 5 when an agent must retain context across extended tasks, inspect visual inputs, and coordinate tools before producing a concise result.
claude-fable-5Anthropic's flagship route for demanding reasoning, end-to-end software work, visual analysis, and autonomous tools. It fits high-stakes workflows where solution quality and persistence matter more than choosing the lowest-cost model.
claude-opus-5A balanced Claude route for production coding, agents, professional writing, and visual reasoning. Sonnet 5 is the practical default when you need strong multi-step work with lower cost than the Opus tier.
claude-sonnet-5A high-capability Opus route for complex reasoning, coding, vision, and tool-driven execution. Use it for long tasks that benefit from careful planning, repository context, and consistent decisions across many steps.
claude-opus-4-8A cost-optimized Claude Opus 4.8 variant focused on coding and tool-driven workloads. It is useful when you want Opus-class task handling while reducing the entry token cost for repeated engineering runs.
claude-opus-4-8-ccPlanned multimodal route for text, vision, reasoning, streaming, and tool use.
google/gemini-proComing next
Pricing to be announced
Planned reasoning-focused route for technical analysis, coding, and streamed responses.
deepseek/deepseek-reasonerComing next
Pricing to be announced
Higher-capability DeepSeek V4 text route for complex reasoning, software architecture, careful code review, and long-document analysis with a one-million-token context.
deepseek-v4-proCost-efficient DeepSeek V4 text route for high-volume extraction, classification, rewriting, summaries, and lightweight coding with a one-million-token context.
deepseek-v4-flashZhipu AI's large reasoning route for complex software engineering and long-horizon agents. Its long-context design suits repository-scale analysis, multi-step tool use, and streamed technical work where earlier decisions must remain consistent.
glm-5.3A cost-efficient multimodal GLM route for focused coding, visual reasoning, and agent subtasks. Choose the Flash tier for high-volume tool workflows that need long-context behavior without the full GLM-5.3 entry cost.
glm-5.3-flashThe flagship GPT-5.6 route for complex reasoning, coding, vision, and agentic execution. Sol is the quality-first choice for command-line work, multi-step engineering, and tool workflows that can justify its higher token price.
gpt-5.6-solThe balanced GPT-5.6 tier for everyday coding, reasoning, visual input, and tools. Terra sits between Sol and Luna when a production workflow needs capable agent behavior without paying for the flagship route on every request.
gpt-5.6-terraThe lightweight GPT-5.6 tier for high-volume, latency-sensitive text work. Use Luna for chat, extraction, classification, and simple streamed workflows where cost efficiency matters more than deep multimodal reasoning.
gpt-5.6-lunaGoogle's frontier preview route for complex reasoning, software engineering, and multimodal analysis. It accepts text, images, and video through PixMind and is best reserved for difficult workflows that need Pro-tier depth and tool use.
gemini-3.1-pro-previewA high-efficiency Gemini model that brings strong coding, reasoning, and multimodal understanding to Flash-tier workloads. It fits responsive agents and parallel processing where throughput and capable visual analysis must stay balanced.
gemini-3.5-flashA responsive multimodal route for coding, agent workflows, and polished web or application output. Choose it for production tasks that need stronger reasoning and visual context while retaining Flash-tier efficiency.
gemini-3.6-flashGoogle's most capable Flash-tier route for long-horizon software engineering, autonomous agents, and complex multi-step workflows. Through PixMind it combines text, image, audio, video, and PDF input with reasoning, tools, coding, and streamed text output.
gemini-3.8-flashGoogle's multimodal Flash route for fast agentic workflows, coding, and reliable multi-step reasoning. It is a practical option for responsive products that combine text, image, or video context with streamed output.
gemini-3.7-flashGoogle's most cost-efficient Flash tier for focused, repeatable subtasks. Use it for high-volume chat, extraction, classification, and lightweight agents where a smaller price footprint matters more than maximum reasoning depth.
gemini-3.5-flash-liteSEPARATE API PRICING
Studio credits include the creation interface, workflow tools, storage, and product services. API pricing is a separate pay-as-you-go ledger for programmatic volume. Launch previews are set about 20% below the equivalent Studio entry configuration; the final request quote still changes with resolution, duration, audio, quality, and output count.