Skip to main content
MeiGen supports multiple AI models for image and video generation. Use the comparison table below to pick one, then scroll to its section for detailed specs.

Quick Comparison

GPT image 1.5 was retired on 2026-04-22 (replaced by GPT Image 2.0). Midjourney V7 was replaced by V8.1 on 2026-05-09.

Image Models

GPT Image 2.0

Credit matrix: All resolutions consume welcome or purchased credits (premium model — daily free credits don’t apply). 4K is paid-only. Aspect ratio does not affect pricing — credits depend only on resolution × quality. The current default model and MeiGen’s best overall image model. Excels at instruction following, in-image text rendering, and complex multi-element scenes.
Stricter moderation: prompts involving real people, celebrities, brand logos, or violent/adult content may be rejected. Rephrase to drop specific names or try a different model.

Nanobanana 2

The most affordable Gemini model with the widest range of aspect ratios, including ultra-wide (8:1) and ultra-tall (1:8) options. Good for quick iterations and exploration. 4K is paid-only.
Content moderation: Nanobanana 2 has stricter content filtering than other models. Prompts that mention celebrities, brand names, real people, or face-swap scenarios are more likely to be rejected. If your prompt is flagged, try rephrasing to avoid specific names or use a different model.

Nanobanana Pro

The premium Gemini model with higher image quality. Better at complex scenes, fine details, and text rendering. Choose this when quality matters more than cost. 4K is paid-only.

Seedream 5.0 Pro

ByteDance’s flagship Seedream tier. Accepts up to 10 reference images — the highest of any model on MeiGen — making it the best pick for multi-subject composition and for keeping a character consistent across images. 2K is paid-only.

Seedream 5.0 Lite

ByteDance’s latest lightweight model. Offers good quality at low cost with optional 3K resolution output. Strong at photorealistic content and Chinese-style aesthetics. 3K is paid-only.

Seedream 4.5

The previous generation Seedream model. Still produces solid results, particularly good for product photography and marketing visuals. 4K is paid-only.

Grok Imagine Quality

xAI’s high-quality image generation model. Supports both text-to-image and image-to-image (up to 3 reference images for editing or guided generation). 2K is paid-only.

Midjourney V8.1

Midjourney’s photorealistic flagship. Returns 4 candidate images per generation (charged once). Choose 2K for native high-resolution output, High quality for finer detail at no extra cost. 2K is paid-only. Features:
  • Content and style reference modes
  • Advanced parameters (stylize, chaos, raw, etc.) for fine-tuning the output
  • Auto-translates non-English prompts to English
  • Native 2K output toggle
  • Quality (Standard / High) toggle

Flux 2 Klein

A base model from Black Forest Labs, suited to everyday general-purpose creation.

Z Image Turbo

The fastest and cheapest model — ideal for quick drafts when you need results in seconds.
No reference image support. Z Image Turbo generates images from text only — you cannot upload reference images with this model. If you need to use reference images, choose a different model.

Agnes Image 2.1 Flash

The cheapest model on MeiGen at 1 credit per image, and the only basic-tier image model that accepts a reference image. Good for high-volume drafting and low-cost image-to-image iteration.

Video Models

Seedance 2.0

ByteDance’s latest video generation model. Supports text-to-video, image-to-video (first/last frame), and video continuation. All outputs include AI-generated audio. The adaptive aspect ratio automatically matches the reference image/video dimensions. Seedance 2.0 is also the only video model with character assets: save a character once, then reference it with @ in your prompt to keep the same subject across shots. Three tiers, picked from the tier switch (Mini / Fast / Pro) in the video sidebar. Same Seedance 2.0 model, same params, just different rendering quality and price:
  • Mini (default): cheapest and most lightweight; 480p / 720p, supports reference-video continuation
  • Fast: quick turnaround at 480p / 720p
  • Pro: higher fidelity, with native 1080p and 4K rendering

Pricing

Per-second. Rate depends on tier (Mini / Fast / Pro) × whether a reference video is uploaded: Mini and Fast only support 480p / 720p. Pro adds native 1080p and 4K (4K is direct-generation only). All three tiers support reference-video continuation; Pro has no 480p With-reference-video rate. With-reference-video billing rule (all tiers): billable seconds = max(reference video duration + your selected duration, min billable) Output length is always your selected duration (4–15), independent of the reference video’s length.

Examples

If generation fails, the full credit amount is automatically refunded — including all per-second charges.

Grok Video 1.5

xAI’s image-to-video model. A starting (first-frame) image is required — pure text-to-video is not supported. The output aspect ratio always follows your input image, and AI-generated audio is included.

Pricing

Examples:
If generation fails, the full credit amount is automatically refunded.

Happyhorse 1.1

Alibaba’s video model — text-to-video and single-image first-frame image-to-video. AI-generated audio included. No reference-video continuation. Auto aspect ratio infers the output ratio from your prompt or reference image. Typical 720p runs take about 2–4 minutes; the time scales with the clip length you choose, and first-frame image-to-video generally finishes faster than pure text-to-video. 1080p takes noticeably longer.

Pricing

Examples:
If generation fails, the full credit amount is automatically refunded.

Veo 3.1

Google’s video generation model. Creates short clips with AI-generated audio synced to the prompt. Auto aspect ratio infers the output ratio from your prompt or reference image. Pick this for premium visual quality with native audio; for longer durations (>8s) or wider ratio support, use Seedance 2.0. Two tiers — pick one from the tier switch (Fast / Pro) in the video sidebar. Same Veo 3.1 model, same params, just different rendering quality and price:
  • Fast (default) — quick turnaround, balanced quality
  • Pro — higher fidelity, longer rendering time

Pricing

Per generation. Price depends on tier (Fast / Pro) × duration (4 / 6 / 8s). All resolutions (720p / 1080p / 4k) share the same price — picking a higher resolution costs no extra credits. Two caveats: 4k requires a paid account, and it takes noticeably longer to render.

Examples

If generation fails, the full credit amount is automatically refunded.

Agnes Video 2.0

The only basic-tier video model — and the only one your daily free credits cover. Supports text-to-video and single-image first-frame image-to-video at a fixed 5 seconds / 480p. For higher resolution or longer clips, use Seedance 2.0, Happyhorse 1.1, or Veo 3.1.

Pricing

Duration is fixed at 5 seconds, so every generation costs 10 credits.
If generation fails, the full credit amount is automatically refunded.

Known Limitations

This model is more sensitive to content involving real people, celebrities, brand logos, and face-swap requests. Prompts that work on other models may be rejected here. If you get a content violation error, rephrase your prompt to remove specific names or try a different model.
Both base models generate images from text only. The reference image upload section will be hidden when either is selected. If you need to use reference images, switch to any other model.
Midjourney V8.1 accepts only 1 reference image, while most other models accept up to 5 (Seedream 5.0 Pro accepts up to 10). In exchange it offers more control over how the reference is used (content vs. style reference).
Unlike other models that return a single image, Midjourney V8.1 returns 4 candidate images per generation. All candidates are saved and can be browsed in the image detail dialog. You are only charged once (15 credits at 1K / 20 credits at 2K) for all 4 images.
Unlike Seedance 2.0, Happyhorse does not support video continuation. Reference image input is single first-frame only. For continuation, use Seedance 2.0.
GPT Image 2.0 applies stricter moderation than other models to real people, celebrities, brand logos, and adult/violent content. Rephrase your prompt to drop specific names, or try a different model.
Seedance bills per second, not per image — and uploading a reference video uses a separate rate table with a minimum-billable floor. See the Seedance 2.0 section above for the full rates and formula.
The Mini tier (the default) is the lightweight option, capped at 720p (it does support reference-video continuation). For 1080p or 4K output, switch to Pro.
Grok Video 1.5 is image-to-video only — you must upload a starting (first-frame) image. Pure text-to-video is not supported. The output aspect ratio follows your input image. If you want text-to-video, use Seedance 2.0, Happyhorse 1.1, or Veo 3.1.
Grok Video 1.5 outputs at 480p or 720p only. For higher resolutions (1080p / 4k), use Seedance 2.0 or Veo 3.1.
Veo 3.1 only supports 16:9 / 9:16 (plus Auto). Square (1:1), 4:3 / 3:4 portrait, and other ratios are not available. For wider ratio support, use Seedance 2.0 or Happyhorse 1.1.
720p / 1080p / 4k all cost the same on Veo 3.1, so 4k costs no extra credits. Two things to know: 4k is locked for free accounts (selecting it opens the upgrade dialog), and rendering takes longer — 4k typically takes 6–8 minutes versus 1–4 minutes for 720p / 1080p. The same 4k lock applies to Seedance 2.0 Pro.