Compare AI image & video models on MeiGen — GPT Image 2.0, Seedream 5.0 Pro, Midjourney V8.1, Nano Banana 2, Seedream 5.0 Lite, Flux 2 Klein, Agnes Image 2.1 Flash, Seedance 2.0, Happyhorse 1.1, Veo 3.1, Agnes Video 2.0. Features, pricing, aspect ratios side by side.
MeiGen supports multiple AI models for image and video generation. Use the comparison table below to pick one, then scroll to its section for detailed specs.
All resolutions consume welcome or purchased credits (premium model — daily free credits don’t apply). 4K is paid-only. Aspect ratio does not affect pricing — credits depend only on resolution × quality.The current default model and MeiGen’s best overall image model. Excels at instruction following, in-image text rendering, and complex multi-element scenes.
Stricter moderation: prompts involving real people, celebrities, brand logos, or violent/adult content may be rejected. Rephrase to drop specific names or try a different model.
The most affordable Gemini model with the widest range of aspect ratios, including ultra-wide (8:1) and ultra-tall (1:8) options. Good for quick iterations and exploration. 4K is paid-only.
Content moderation: Nanobanana 2 has stricter content filtering than other models. Prompts that mention celebrities, brand names, real people, or face-swap scenarios are more likely to be rejected. If your prompt is flagged, try rephrasing to avoid specific names or use a different model.
The premium Gemini model with higher image quality. Better at complex scenes, fine details, and text rendering. Choose this when quality matters more than cost. 4K is paid-only.
ByteDance’s flagship Seedream tier. Accepts up to 10 reference images — the highest of any model on MeiGen — making it the best pick for multi-subject composition and for keeping a character consistent across images. 2K is paid-only.
ByteDance’s latest lightweight model. Offers good quality at low cost with optional 3K resolution output. Strong at photorealistic content and Chinese-style aesthetics. 3K is paid-only.
The previous generation Seedream model. Still produces solid results, particularly good for product photography and marketing visuals. 4K is paid-only.
xAI’s high-quality image generation model. Supports both text-to-image and image-to-image (up to 3 reference images for editing or guided generation). 2K is paid-only.
Midjourney’s photorealistic flagship. Returns 4 candidate images per generation (charged once). Choose 2K for native high-resolution output, High quality for finer detail at no extra cost. 2K is paid-only.Features:
Content and style reference modes
Advanced parameters (stylize, chaos, raw, etc.) for fine-tuning the output
The fastest and cheapest model — ideal for quick drafts when you need results in seconds.
No reference image support. Z Image Turbo generates images from text only — you cannot upload reference images with this model. If you need to use reference images, choose a different model.
The cheapest model on MeiGen at 1 credit per image, and the only basic-tier image model that accepts a reference image. Good for high-volume drafting and low-cost image-to-image iteration.
All tiers (Pro 480p has no With-reference-video rate)
Resolutions
Mini and Fast: 480p / 720p; Pro: 480p / 720p / 1080p / 4K
Aspect ratios
adaptive, 16:9, 4:3, 1:1, 3:4, 9:16, 21:9
Audio
Auto-generated
ByteDance’s latest video generation model. Supports text-to-video, image-to-video (first/last frame), and video continuation. All outputs include AI-generated audio. The adaptive aspect ratio automatically matches the reference image/video dimensions. Seedance 2.0 is also the only video model with character assets: save a character once, then reference it with @ in your prompt to keep the same subject across shots.Three tiers, picked from the tier switch (Mini / Fast / Pro) in the video sidebar. Same Seedance 2.0 model, same params, just different rendering quality and price:
Mini (default): cheapest and most lightweight; 480p / 720p, supports reference-video continuation
Fast: quick turnaround at 480p / 720p
Pro: higher fidelity, with native 1080p and 4K rendering
Per-second. Rate depends on tier (Mini / Fast / Pro) × whether a reference video is uploaded:
Resolution
Mini Direct
Mini With reference video
Fast Direct
Fast With reference video
Pro Direct
Pro With reference video
480p
9/sec
7/sec
13/sec
8/sec
14/sec
—
720p
16/sec
13/sec
20/sec
14/sec
22/sec
16/sec
1080p
—
—
—
—
42/sec
28/sec
4K
—
—
—
—
100/sec
—
Mini and Fast only support 480p / 720p. Pro adds native 1080p and 4K (4K is direct-generation only). All three tiers support reference-video continuation; Pro has no 480p With-reference-video rate.With-reference-video billing rule (all tiers):billable seconds = max(reference video duration + your selected duration, min billable)Output length is always your selected duration (4–15), independent of the reference video’s length.
Follows the first-frame image automatically (no manual ratio selection)
Audio
Auto-generated
xAI’s image-to-video model. A starting (first-frame) image is required — pure text-to-video is not supported. The output aspect ratio always follows your input image, and AI-generated audio is included.
Alibaba’s video model — text-to-video and single-image first-frame image-to-video. AI-generated audio included. No reference-video continuation. Auto aspect ratio infers the output ratio from your prompt or reference image.Typical 720p runs take about 2–4 minutes; the time scales with the clip length you choose, and first-frame image-to-video generally finishes faster than pure text-to-video. 1080p takes noticeably longer.
Google’s video generation model. Creates short clips with AI-generated audio synced to the prompt. Auto aspect ratio infers the output ratio from your prompt or reference image. Pick this for premium visual quality with native audio; for longer durations (>8s) or wider ratio support, use Seedance 2.0.Two tiers — pick one from the tier switch (Fast / Pro) in the video sidebar. Same Veo 3.1 model, same params, just different rendering quality and price:
Fast (default) — quick turnaround, balanced quality
Per generation. Price depends on tier (Fast / Pro) × duration (4 / 6 / 8s). All resolutions (720p / 1080p / 4k) share the same price — picking a higher resolution costs no extra credits. Two caveats: 4k requires a paid account, and it takes noticeably longer to render.
The only basic-tier video model — and the only one your daily free credits cover. Supports text-to-video and single-image first-frame image-to-video at a fixed 5 seconds / 480p. For higher resolution or longer clips, use Seedance 2.0, Happyhorse 1.1, or Veo 3.1.
This model is more sensitive to content involving real people, celebrities, brand logos, and face-swap requests. Prompts that work on other models may be rejected here. If you get a content violation error, rephrase your prompt to remove specific names or try a different model.
Flux 2 Klein / Z Image Turbo — No reference images
Both base models generate images from text only. The reference image upload section will be hidden when either is selected. If you need to use reference images, switch to any other model.
Midjourney V8.1 — Single reference image only
Midjourney V8.1 accepts only 1 reference image, while most other models accept up to 5 (Seedream 5.0 Pro accepts up to 10). In exchange it offers more control over how the reference is used (content vs. style reference).
Midjourney V8.1 — Returns 4 candidate images
Unlike other models that return a single image, Midjourney V8.1 returns 4 candidate images per generation. All candidates are saved and can be browsed in the image detail dialog. You are only charged once (15 credits at 1K / 20 credits at 2K) for all 4 images.
Happyhorse 1.1 — No reference video continuation
Unlike Seedance 2.0, Happyhorse does not support video continuation. Reference image input is single first-frame only. For continuation, use Seedance 2.0.
GPT Image 2.0 — Stricter moderation
GPT Image 2.0 applies stricter moderation than other models to real people, celebrities, brand logos, and adult/violent content. Rephrase your prompt to drop specific names, or try a different model.
Seedance bills per second, not per image — and uploading a reference video uses a separate rate table with a minimum-billable floor. See the Seedance 2.0 section above for the full rates and formula.
Seedance 2.0 Mini — 480p / 720p only
The Mini tier (the default) is the lightweight option, capped at 720p (it does support reference-video continuation). For 1080p or 4K output, switch to Pro.
Grok Video 1.5 — First frame required
Grok Video 1.5 is image-to-video only — you must upload a starting (first-frame) image. Pure text-to-video is not supported. The output aspect ratio follows your input image. If you want text-to-video, use Seedance 2.0, Happyhorse 1.1, or Veo 3.1.
Grok Video 1.5 — 480p / 720p only
Grok Video 1.5 outputs at 480p or 720p only. For higher resolutions (1080p / 4k), use Seedance 2.0 or Veo 3.1.
Veo 3.1 — Limited aspect ratios
Veo 3.1 only supports 16:9 / 9:16 (plus Auto). Square (1:1), 4:3 / 3:4 portrait, and other ratios are not available. For wider ratio support, use Seedance 2.0 or Happyhorse 1.1.
Veo 3.1 — 4k costs no extra credits, but needs a paid account
720p / 1080p / 4k all cost the same on Veo 3.1, so 4k costs no extra credits. Two things to know: 4k is locked for free accounts (selecting it opens the upgrade dialog), and rendering takes longer — 4k typically takes 6–8 minutes versus 1–4 minutes for 720p / 1080p. The same 4k lock applies to Seedance 2.0 Pro.