Nur für kurze ZeitJahresmitgliedschaft:30 % Rabattplus unbegrenzter Zugriff auf GPT Image, MiniMax H3 und mehr
Neuer AI Chat · Täglich kostenlose Testnutzungen
Jetzt upgraden
Pixmind
API Platform

MODEL CATALOG

Choose the model. Keep the same API.

Search the model routes available through PixMind, compare capabilities and billing units, then copy the stable model ID into your integration.

Image, video, and LLM routes are live. Individual planned models remain clearly labeled on their cards.
76model routes
19providers
43live routes

76 results

GoogleAvailable

Nano Banana Pro

Strong multilingual text, structured visual layouts, reference fusion, and precise natural-language edits.

Text to imageImage editingMulti-reference+1
nano-banana-pro

API price preview

$0.13/image

19% below official

OpenAIAvailable

GPT Image 2

High-control image generation and editing for posters, products, infographics, and iterative design work.

Text to imageImage editingMulti-reference+1
gpt-image-2

API price preview

$0.050/image

17% below official

ByteDanceAvailable

Seedream 5.0 Pro

Commercial-grade visuals with strong composition, detail, reference consistency, and high-resolution output.

Text to imageImage editingMulti-reference+1
seedream-5.0-pro

API price preview

$0.060/image

25% below official

Black Forest LabsAvailable

FLUX.1 Kontext Pro

Context-aware image editing that preserves subjects and style across targeted changes.

Text to imageImage editingMulti-reference+1
flux-kontext-pro

Pricing to be announced

AlibabaAvailable

Wan 2.7 Image

A balanced route for generation, local editing, multiple references, and production image workflows.

Text to imageImage editingMulti-reference+1
wan2.7-image

Pricing to be announced

MidjourneyAvailable

Midjourney V7

Distinctive art direction, polished aesthetics, and reference-driven visual exploration.

Text to imageMulti-referenceCharacter consistency+1
mj-v7

API price preview

$0.030/image

63% below official

PixmindPlanned

Pixmind

PixMind's free image generation route for quick everyday text-to-image creation.

Text to image
z-image

Coming next

Pricing to be announced

GooglePlanned

Nano Banana Pro Eco

A cost-efficient Nano Banana Pro route covering text-to-image and image editing with reference support.

Text to imageImage editingMulti-reference+1
nano-banana-pro-lite

Coming next

Pricing to be announced

GoogleAvailable

Nano Banana 2

Google's Nano Banana 2 high-speed image generation across 1K, 2K, and 4K resolutions.

Text to imageImage editingMulti-reference+1
nano-banana-2

API price preview

$0.090/image

31% below official

GooglePlanned

Nano Banana

The original Nano Banana image generation route for versatile text-to-image and editing work.

Text to imageImage editingMulti-reference
nano-banana

Coming next

Pricing to be announced

GoogleAvailable

Nano Banana 2 Eco

An economical Nano Banana 2 variant balancing speed, quality, and cost for everyday generation.

Text to imageImage editingMulti-reference+1
nano-banana-2-eco

API price preview

$0.050/image

38% below official

MidjourneyPlanned

Midjourney Niji 7

Midjourney's anime and illustration-tuned model for stylized characters and vibrant artwork.

Text to imageImage editingMulti-reference
mj-niji7

Coming next

Pricing to be announced

MidjourneyPlanned

Midjourney V8.2

Midjourney V8.2 with refined aesthetics, stronger prompt adherence, and richer art direction.

Text to imageImage editingMulti-reference
mj-v8.2

Coming next

Pricing to be announced

MidjourneyPlanned

Midjourney V8.1

Midjourney V8.1 delivering elevated visual quality and consistency for polished creative output.

Text to imageImage editingMulti-reference
mj-v8.1

Coming next

Pricing to be announced

ByteDancePlanned

Seedream 5.0

ByteDance Seedream 5.0 commercial-grade image generation with strong composition and detail.

Text to imageImage editingMulti-reference+1
seedream-5.0

Coming next

Pricing to be announced

ByteDancePlanned

Seedream 4.5

Seedream 4.5 balanced image generation with reliable quality for everyday creative work.

Text to imageImage editingMulti-reference+1
seedream-4.5

Coming next

Pricing to be announced

ByteDancePlanned

Seedream 4.0

Seedream 4.0 image generation covering text-to-image and editing for production visuals.

Text to imageImage editingMulti-reference+1
seedream-4.0

Coming next

Pricing to be announced

OpenAIPlanned

GPT Image 1.5

GPT Image 1.5 high-control generation and editing for design-heavy visuals.

Text to imageImage editingMulti-reference
gpt-image-1.5

Coming next

Pricing to be announced

OpenAIPlanned

GPT Image 4o

GPT Image 4o versatile image generation with precise prompt following and editing.

Text to imageImage editingMulti-reference
gpt-image-4o

Coming next

Pricing to be announced

OpenAIPlanned

GPT Image 2 Eco

An economical GPT Image 2 variant for cost-aware generation and editing at scale.

Text to imageImage editingMulti-reference+1
gpt-image-2-eco

Coming next

Pricing to be announced

AlibabaPlanned

Wan 2.6 Image

Wan 2.6 image generation and editing from Alibaba for multilingual creative workflows.

Text to imageImage editingMulti-reference
wan2.6-image

Coming next

Pricing to be announced

AlibabaPlanned

Wan 2.7 Image Pro

Wan 2.7 Image Pro, Alibaba's higher-tier route for sharper, production-ready visuals.

Text to imageImage editingMulti-reference+1
wan2.7-image-pro

Coming next

Pricing to be announced

AlibabaAvailable

Qwen Image 3.0 Pro

Qwen Image 3.0 Pro with improved text rendering and high-resolution output.

Text to imageImage editingMulti-reference+2
qwen-image-3.0-pro

API price preview

$0.040/image

50% below official

AlibabaPlanned

Qwen Image 2.0 Pro

Qwen Image 2.0 Pro text-to-image generation with readable typography and detail.

Text to image
qwen-image-2.0-pro

Coming next

Pricing to be announced

xAIAvailable

Grok Imagine 2.0 Eco

xAI image generation for expressive, high-impact visuals across seven common aspect ratios and larger batches.

Text to imageHigh resolution
grok-imagine-2-eco

API price preview

$0.090/image

15% below official

KreaAvailable

Krea 2 Turbo

Fast Krea text-to-image generation with flexible aspect ratios, deterministic seeds, and 1K or 2K output.

Text to imageHigh resolution
krea-2-turbo

API price preview

$0.010/image

56% below official

Grok AIPlanned

Grok Imagine 1.5

Grok Imagine 1.5 image generation with bold, expressive visual styling.

Text to imageImage editingMulti-reference
grok-imagine-1.5

Coming next

Pricing to be announced

RecraftPlanned

Recraft V4 Vector

Recraft V4 vector generation for clean, scalable, print-ready artwork.

Text to imageSVG vector output
recraftv4_vector

Coming next

Pricing to be announced

RecraftPlanned

Recraft V4 Pro Vector

Recraft V4 Pro vector generation with premium quality for professional design.

Text to imageSVG vector output
recraftv4_pro_vector

Coming next

Pricing to be announced

RecraftPlanned

Recraft Image to Vector

Recraft image-to-vector conversion that turns raster images into editable vectors.

Image editingSVG vector output
recraft-vectorize

Coming next

Pricing to be announced

GoogleAvailable

Veo 3.1

Cinematic video generation with image guidance, shot control, first/last frames, and native audio.

Text to videoImage to videoNative audio+1
veo-3.1

API price preview

$0.36/sec

20% below official

MiniMaxAvailable

MiniMax H3

Native 2K video with synchronized audio from text, images, first/last frames, and multimodal references.

Text to videoImage to videoNative audio+2
minimax-h3

API price preview

$0.090/sec
ByteDanceAvailable

Seedance 2.0 Pro

Multi-reference video creation with coordinated characters, scenes, motion, and sound.

Text to videoImage to videoMulti-reference+1
seedance-2.0-pro

API price preview

$0.072/sec

20% below official

ByteDanceAvailable

Seedance 2.5

Native 30-second single-shot video with up to 50 multimodal references, 4K output, and region-level editing.

Text to videoImage to videoMulti-reference+2
seedance-2.5

API price preview

$0.072/sec

20% below official

KuaishouAvailable

Kling V3

High-fidelity image animation and camera or subject motion control for production shots.

Text to videoImage to videoMulti-reference+1
kling-v3-motion-control

API price preview

$0.16/sec

20% below official

AlibabaAvailable

Wan 2.7 Video

Versatile text, image, reference, and first/last-frame workflows through one video route.

Text to videoImage to videoMulti-reference+1
wan2.7-video

API price preview

$0.12/sec

20% below official

AlibabaAvailable

Wan 3.0 Video

Wan 3.0 video generation with multimodal references, synchronized audio, 2-30 second clips, and up to 1080p output.

Text to videoImage to videoMulti-reference+2
wan3.0-video

API price preview

$0.060/sec

14% below official

OpenAIAvailable

Sora 2 Pro

Detailed video generation for longer creative shots with image guidance and audio-capable output.

Text to videoImage to videoNative audio+1
sora-2-pro

API price preview

$0.048/sec

20% below official

PixVerseAvailable

PixVerse V5

Fast video ideation across text-to-video, image animation, transitions, effects, and sound.

Text to videoImage to videoFirst / last frame+1
pixverse-v5

API price preview

$0.067/sec

20% below official

MiniMaxAvailable

Happy Horse 1.0

Reference-led character performance and expressive motion for short-form creative video.

Text to videoImage to videoMulti-reference+1
happyhorse-1.0

API price preview

$0.176/sec

20% below official

Kling AIAvailable

Kling 3.0 Turbo

Kuaishou Kling 3.0 Turbo, a fast text-to-video and image-to-video route for quick drafts.

Text to videoImage to video
kling-3.0-turbo

API price preview

$0.13/sec

19% below official

ByteDancePlanned

Seedance 2.0 Fast

Seedance 2.0 Fast, a quicker Seedance route for faster video turnaround.

Text to videoImage to video
seedance-2.0-fast

Coming next

Pricing to be announced

ByteDancePlanned

Seedance 2.0 Mini

Seedance 2.0 Mini, a lightweight variant for efficient short video generation.

Text to videoImage to video
seedance-2-0-mini

Coming next

Pricing to be announced

OpenAIPlanned

Sora 2

OpenAI Sora 2 video generation for text-to-video and image-to-video creative shots.

Text to videoImage to video
sora-2

Coming next

Pricing to be announced

OpenAIPlanned

Sora 2 Eco

An economical Sora 2 variant for cost-efficient video generation.

Text to videoImage to video
sora-2-eco

Coming next

Pricing to be announced

GooglePlanned

Veo 3

Google Veo 3 cinematic video generation with image guidance and audio.

Text to videoImage to video
veo-3

Coming next

Pricing to be announced

GooglePlanned

Veo 3.0 Fast

Veo 3.0 Fast, a quicker Veo route for rapid video iteration.

Text to videoImage to video
veo-3.0-fast

Coming next

Pricing to be announced

GooglePlanned

Veo 3.1 Eco

An economical Veo 3.1 variant balancing quality and cost.

Text to videoImage to video
veo-3.1-eco

Coming next

Pricing to be announced

GooglePlanned

Veo 3.1 Fast Eco

Veo 3.1 Fast Eco, the most cost-optimized Veo route for fast, budget-friendly video.

Text to videoImage to video
veo-3.1-fast-eco

Coming next

Pricing to be announced

PixVersePlanned

PixVerse V6

PixVerse V6 video generation with text-to-video, image animation, and effects.

Text to videoImage to video
pixverse-v6

Coming next

Pricing to be announced

AlibabaPlanned

Wan 2.6 I2V

Wan 2.6 image-to-video from Alibaba for animating still images.

Image to video
wan2.6-i2v

Coming next

Pricing to be announced

AlibabaPlanned

Wan 2.6 T2V

Wan 2.6 text-to-video from Alibaba for turning prompts into motion.

Text to video
wan2.6-t2v

Coming next

Pricing to be announced

AlibabaPlanned

Wan 2.6 I2V Flash

Wan 2.6 I2V Flash, a rapid image-to-video variant for quick animation.

Image to video
wan2.6-i2v-flash

Coming next

Pricing to be announced

AnthropicAvailable

Claude Fable 5

Anthropic's Mythos-class route for autonomous knowledge work, long-form reasoning, and coding. Choose Fable 5 when an agent must retain context across extended tasks, inspect visual inputs, and coordinate tools before producing a concise result.

TextVisionReasoning+3
claude-fable-5
Input$8.07/1M tokens
Output$40.33/1M tokens
AnthropicAvailable

Claude Opus 5

Anthropic's flagship route for demanding reasoning, end-to-end software work, visual analysis, and autonomous tools. It fits high-stakes workflows where solution quality and persistence matter more than choosing the lowest-cost model.

TextVisionReasoning+3
claude-opus-5
Input$4.04/1M tokens
Output$20.17/1M tokens
AnthropicAvailable

Claude Sonnet 5

A balanced Claude route for production coding, agents, professional writing, and visual reasoning. Sonnet 5 is the practical default when you need strong multi-step work with lower cost than the Opus tier.

TextVisionReasoning+3
claude-sonnet-5
Input$1.62/1M tokens
Output$8.07/1M tokens
AnthropicAvailable

Claude Opus 4.8

A high-capability Opus route for complex reasoning, coding, vision, and tool-driven execution. Use it for long tasks that benefit from careful planning, repository context, and consistent decisions across many steps.

TextVisionReasoning+3
claude-opus-4-8
Input$4.04/1M tokens
Output$20.17/1M tokens
AnthropicAvailable

Claude Opus 4.8 CC

A cost-optimized Claude Opus 4.8 variant focused on coding and tool-driven workloads. It is useful when you want Opus-class task handling while reducing the entry token cost for repeated engineering runs.

TextReasoningTool calling+2
claude-opus-4-8-cc
Input$2.56/1M tokens
Output$12.77/1M tokens
GooglePlanned

Gemini Pro

Planned multimodal route for text, vision, reasoning, streaming, and tool use.

TextVisionImage input+4
google/gemini-pro

Coming next

Pricing to be announced

DeepSeekPlanned

DeepSeek Reasoner

Planned reasoning-focused route for technical analysis, coding, and streamed responses.

TextReasoningCoding+1
deepseek/deepseek-reasoner

Coming next

Pricing to be announced

DeepSeekAvailable

DeepSeek V4 Pro

Higher-capability DeepSeek V4 text route for complex reasoning, software architecture, careful code review, and long-document analysis with a one-million-token context.

TextReasoningCoding+1
deepseek-v4-pro
Input$1.12/1M tokens
Output$3.36/1M tokens
DeepSeekAvailable

DeepSeek V4 Flash

Cost-efficient DeepSeek V4 text route for high-volume extraction, classification, rewriting, summaries, and lightweight coding with a one-million-token context.

TextReasoningCoding+1
deepseek-v4-flash
Input$0.38/1M tokens
Output$1.12/1M tokens
智谱 AIAvailable

GLM-5.3

Zhipu AI's large reasoning route for complex software engineering and long-horizon agents. Its long-context design suits repository-scale analysis, multi-step tool use, and streamed technical work where earlier decisions must remain consistent.

TextReasoningTool calling+2
glm-5.3
Input$1.48/1M tokens
Output$5.17/1M tokens
智谱 AIAvailable

GLM-5.3-Flash

A cost-efficient multimodal GLM route for focused coding, visual reasoning, and agent subtasks. Choose the Flash tier for high-volume tool workflows that need long-context behavior without the full GLM-5.3 entry cost.

TextVisionImage input+2
glm-5.3-flash
Input$0.15/1M tokens
Output$0.52/1M tokens
OpenAIAvailable

GPT-5.6 Sol

The flagship GPT-5.6 route for complex reasoning, coding, vision, and agentic execution. Sol is the quality-first choice for command-line work, multi-step engineering, and tool workflows that can justify its higher token price.

TextVisionReasoning+2
gpt-5.6-sol
Input$3.90/1M tokens
Output$23.39/1M tokens
OpenAIAvailable

GPT-5.6 Terra

The balanced GPT-5.6 tier for everyday coding, reasoning, visual input, and tools. Terra sits between Sol and Luna when a production workflow needs capable agent behavior without paying for the flagship route on every request.

TextVisionReasoning+2
gpt-5.6-terra
Input$1.56/1M tokens
Output$9.36/1M tokens
OpenAIAvailable

GPT-5.6 Luna

The lightweight GPT-5.6 tier for high-volume, latency-sensitive text work. Use Luna for chat, extraction, classification, and simple streamed workflows where cost efficiency matters more than deep multimodal reasoning.

TextStreaming
gpt-5.6-luna
Input$0.16/1M tokens
Output$0.94/1M tokens
GoogleAvailable

Gemini 3.1 Pro Preview

Google's frontier preview route for complex reasoning, software engineering, and multimodal analysis. It accepts text, images, and video through PixMind and is best reserved for difficult workflows that need Pro-tier depth and tool use.

TextVisionImage input+4
gemini-3.1-pro-preview
Input$2.02/1M tokens
Output$12.10/1M tokens
GoogleAvailable

Gemini 3.5 Flash

A high-efficiency Gemini model that brings strong coding, reasoning, and multimodal understanding to Flash-tier workloads. It fits responsive agents and parallel processing where throughput and capable visual analysis must stay balanced.

TextVisionImage input+4
gemini-3.5-flash
Input$1.43/1M tokens
Output$8.65/1M tokens
GoogleAvailable

Gemini 3.6 Flash

A responsive multimodal route for coding, agent workflows, and polished web or application output. Choose it for production tasks that need stronger reasoning and visual context while retaining Flash-tier efficiency.

TextVisionImage input+4
gemini-3.6-flash
Input$1.01/1M tokens
Output$5.19/1M tokens
GoogleAvailable

Gemini 3.8 Flash

Google's most capable Flash-tier route for long-horizon software engineering, autonomous agents, and complex multi-step workflows. Through PixMind it combines text, image, audio, video, and PDF input with reasoning, tools, coding, and streamed text output.

TextVisionImage input+7
gemini-3.8-flash
Input$1.01/1M tokens
Output$5.19/1M tokens
GoogleAvailable

Gemini 3.7 Flash

Google's multimodal Flash route for fast agentic workflows, coding, and reliable multi-step reasoning. It is a practical option for responsive products that combine text, image, or video context with streamed output.

TextVisionImage input+2
gemini-3.7-flash
Input$1.01/1M tokens
Output$5.19/1M tokens
GoogleAvailable

Gemini 3.5 Flash-Lite

Google's most cost-efficient Flash tier for focused, repeatable subtasks. Use it for high-volume chat, extraction, classification, and lightweight agents where a smaller price footprint matters more than maximum reasoning depth.

TextVisionImage input+2
gemini-3.5-flash-lite
Input$0.41/1M tokens
Output$3.37/1M tokens

SEPARATE API PRICING

API usage is priced independently from PixMind Studio

Studio credits include the creation interface, workflow tools, storage, and product services. API pricing is a separate pay-as-you-go ledger for programmatic volume. Launch previews are set about 20% below the equivalent Studio entry configuration; the final request quote still changes with resolution, duration, audio, quality, and output count.