Compare current MeiGen image and video models, including GPT Image 2, Nano Banana, Seedance 2.5, Seedance 2.0, and Veo 3.1. See capabilities, pricing, and limits.
MeiGen supports multiple AI models for image and video generation. Use the comparison table below to pick one, then scroll to its section for detailed specs.
Per-second (Mini 9–16/sec, Fast 13–20/sec, Pro 14–100/sec)
Video, ~2-6min
First/last frame or content reference
✓
Three tiers (Mini / Fast / Pro), up to 4K on Pro, 4–15s, audio
Grok Video 1.5
Per-second (480p 16/sec, 720p 28/sec)
Video, ~35s
First frame (required)
✓
First-frame image-to-video only, 4–15s, ratio follows input, audio
Veo 3.1
Per generation (Fast 80–165, Pro 220–435)
Video, ~1-8min
Up to 2 (first/last frame)
Web only
Two tiers (Fast / Pro), 4/6/8s, 720p/1080p/4k same price, audio
Agnes Video 2.0
10 per video (480p, 2/sec × 5s)
Video, ~2min
First frame (single image)
✓
Fixed 5s / 480p, the only video model covered by daily free credits
✓ = available on the iOS/Android app and web. Web only = browser only for now.GPT image 1.5 was retired on 2026-04-22 (replaced by GPT Image 2.0). Midjourney V7 was replaced by V8.1 on 2026-05-09. Happyhorse 1.1 was retired on 2026-07-31.
All resolutions consume welcome or purchased credits (premium model — daily free credits don’t apply). 4K is paid-only. Aspect ratio does not affect pricing — credits depend only on resolution × quality.The current default model and MeiGen’s best overall image model. Excels at instruction following, in-image text rendering, and complex multi-element scenes.
Stricter moderation: prompts involving real people, celebrities, brand logos, or violent/adult content may be rejected. Rephrase to drop specific names or try a different model.
The most affordable Gemini model with the widest range of aspect ratios, including ultra-wide (8:1) and ultra-tall (1:8) options. Good for quick iterations and exploration. 4K is paid-only.
Content moderation: Nanobanana 2 has stricter content filtering than other models. Prompts that mention celebrities, brand names, real people, or face-swap scenarios are more likely to be rejected. If your prompt is flagged, try rephrasing to avoid specific names or use a different model.
The premium Gemini model with higher image quality. Better at complex scenes, fine details, and text rendering. Choose this when quality matters more than cost. 4K is paid-only.
ByteDance’s flagship Seedream tier. Accepts up to 10 reference images — the highest of any model on MeiGen — making it the best pick for multi-subject composition and for keeping a character consistent across images. 2K is paid-only.
ByteDance’s latest lightweight model. Offers good quality at low cost with optional 3K resolution output. Strong at photorealistic content and Chinese-style aesthetics. 3K is paid-only.
The previous generation Seedream model. Still produces solid results, particularly good for product photography and marketing visuals. 4K is paid-only.
xAI’s Grok Imagine 2.0 image model. Supports both text-to-image and image-to-image (up to 3 reference images for editing or guided generation), with a focus on precise editing, infographics, ads, and UI/UX mockups. A premium model — daily free credits don’t apply.
Midjourney’s photorealistic flagship. Returns 4 candidate images per generation (charged once). Choose 2K for native high-resolution output, High quality for finer detail at no extra cost. 2K is paid-only.Features:
Content and style reference modes
Advanced parameters (stylize, chaos, raw, etc.) for fine-tuning the output
The fastest and cheapest model — ideal for quick drafts when you need results in seconds.
No reference image support. Z Image Turbo generates images from text only — you cannot upload reference images with this model. If you need to use reference images, choose a different model.
The cheapest model on MeiGen at 1 credit per image, and the only basic-tier image model that accepts a reference image. Good for high-volume drafting and low-cost image-to-image iteration.
Web/App: up to 2 as first/last frame or content reference. REST API: up to 30 content references (referenceMode: content)
Reference video
2–30 seconds, up to 200MB; actual duration is detected automatically
Resolutions
480p, 720p (default 480p)
Aspect ratios
adaptive, 16:9, 4:3, 1:1, 3:4, 9:16, 21:9
Audio
Auto-generated
Tiers
None — do not send tier
Seedance 2.5 is the long-duration Seedance model on MeiGen. It supports text-to-video, first/last-frame image-to-video, content-reference images, and reference-video continuation. The Web and App pickers currently accept two content references; the REST API accepts up to 30. Frame references and reference-video continuation use adaptive ratio so the output follows the input media; text-to-video and content-reference requests can use an explicit supported ratio.
Direct generation bills only the selected output duration. With a reference video:billable seconds = max(reference video duration + selected output duration, ceil(selected output duration × 5 / 3))The minimum-billable floor is 7 seconds for a 4-second output, 9 for 5 seconds, and 50 for 30 seconds. Output length is always the selected 4–30 seconds; the billing floor does not lengthen the generated clip.
Scenario
Calculation
Total
480p × 5s direct
18 × 5
90
720p × 5s direct
39 × 5
195
480p, 2s reference + 5s output
11 × max(2+5, 9)
99
720p, 2s reference + 5s output
24 × max(2+5, 9)
216
480p × 30s direct
18 × 30
540
If generation fails, the full credit amount is automatically refunded.
Web/App: up to 2 as first/last frame or content reference. REST API: up to 9 content references (referenceMode: content)
Reference video
All three tiers. Upload must be 2–15s and ≤50MB — the platform detects the actual duration automatically
Resolutions
Direct: Mini/Fast 480p/720p, Pro 480p/720p/1080p/4K. With reference video: Mini/Fast 480p/720p, Pro 720p/1080p (no 480p, no 4K)
Aspect ratios
adaptive, 16:9, 4:3, 1:1, 3:4, 9:16, 21:9
Audio
Auto-generated
Seedance 2.0 is the three-tier Seedance option. It supports text-to-video, image-to-video (first/last frame or content reference), and video continuation. All outputs include AI-generated audio. The adaptive aspect ratio automatically matches the reference image/video dimensions. Seedance 2.0 is also the only video model with character assets: save a character once, then reference it with @ in your prompt to keep the same subject across shots.Two reference-image modes: first/last frame (default, up to 2 images) sets the opening and/or closing frame; content reference guides style or subject instead of fixing a frame position. The Web and App content-reference pickers accept up to 2 images, while the REST API accepts up to 9 with referenceMode: content.Three tiers, picked from the tier switch (Mini / Fast / Pro) in the video sidebar. Same Seedance 2.0 model, same params, just different rendering quality and price:
Mini (default): cheapest and most lightweight; 480p / 720p
Fast: quick turnaround at 480p / 720p
Pro: higher fidelity, with native 1080p and 4K rendering (4K is direct-generation only)
All three tiers support reference-video continuation — continuation output tops out at 1080p even on Pro.
Per-second. Rate depends on tier (Mini / Fast / Pro) × whether a reference video is uploaded:
Resolution
Mini Direct
Mini With reference video
Fast Direct
Fast With reference video
Pro Direct
Pro With reference video
480p
9/sec
7/sec
13/sec
8/sec
14/sec
—
720p
16/sec
13/sec
20/sec
14/sec
22/sec
16/sec
1080p
—
—
—
—
42/sec
28/sec
4K
—
—
—
—
100/sec
—
Mini and Fast only support 480p / 720p. Pro adds native 1080p and 4K for direct generation; with a reference video, Pro is limited to 720p / 1080p (no 480p rate).With-reference-video billing rule (all tiers):billable seconds = max(reference video duration + your selected duration, min billable)Output length is always your selected duration (4–15), independent of the reference video’s length.
Follows the first-frame image automatically (no manual ratio selection)
Audio
Auto-generated
xAI’s image-to-video model. A starting (first-frame) image is required — pure text-to-video is not supported. The output aspect ratio always follows your input image, and AI-generated audio is included.
Google’s video generation model. Creates short clips with AI-generated audio synced to the prompt. Auto aspect ratio infers the output ratio from your prompt or reference image. Pick this for premium visual quality with native audio; for longer durations (>8s), use Seedance 2.5.Two tiers — pick one from the tier switch (Fast / Pro) in the video sidebar. Same Veo 3.1 model, same params, just different rendering quality and price:
Fast (default) — quick turnaround, balanced quality
Per generation. Price depends on tier (Fast / Pro) × duration (4 / 6 / 8s). All resolutions (720p / 1080p / 4k) share the same price — picking a higher resolution costs no extra credits. Two caveats: 4k requires a paid account, and it takes noticeably longer to render.
The only basic-tier video model — and the only one your daily free credits cover. Supports text-to-video and single-image first-frame image-to-video at a fixed 5 seconds / 480p. For longer clips, use Seedance 2.5; for higher resolutions, use Seedance 2.0 Pro or Veo 3.1.
This model is more sensitive to content involving real people, celebrities, brand logos, and face-swap requests. Prompts that work on other models may be rejected here. If you get a content violation error, rephrase your prompt to remove specific names or try a different model.
Flux 2 Klein / Z Image Turbo — No reference images
Both base models generate images from text only. The reference image upload section will be hidden when either is selected. If you need to use reference images, switch to any other model.
Midjourney V8.1 — Single reference image only
Midjourney V8.1 accepts only 1 reference image, while most other models accept up to 5 (Seedream 5.0 Pro accepts up to 10). In exchange it offers more control over how the reference is used (content vs. style reference).
Midjourney V8.1 — Returns 4 candidate images
Unlike other models that return a single image, Midjourney V8.1 returns 4 candidate images per generation. All candidates are saved and can be browsed in the image detail dialog. You are only charged once (15 credits at 1K / 20 credits at 2K) for all 4 images.
GPT Image 2.0 — Stricter moderation
GPT Image 2.0 applies stricter moderation than other models to real people, celebrities, brand logos, and adult/violent content. Rephrase your prompt to drop specific names, or try a different model.
Seedance bills per second, not per image — and uploading a reference video uses a separate rate table with a minimum-billable floor. See the Seedance 2.0 section above for the full rates and formula.
Seedance 2.5 — No Mini / Fast / Pro tier
Seedance 2.5 has one quality level and supports 480p / 720p. Do not send tier; use Seedance 2.0 Pro when you specifically need 1080p or 4K.
Seedance 2.0 Mini — 480p / 720p only
The Mini tier (the default) is the lightweight option, capped at 720p. For 1080p or 4K output, switch to Pro — note that reference-video continuation on Pro tops out at 1080p (4K needs direct generation with no reference video).
Grok Video 1.5 — First frame required
Grok Video 1.5 is image-to-video only — you must upload a starting (first-frame) image. Pure text-to-video is not supported. The output aspect ratio follows your input image. If you want text-to-video, use Seedance 2.5, Seedance 2.0, or Veo 3.1.
Grok Video 1.5 — 480p / 720p only
Grok Video 1.5 outputs at 480p or 720p only. For higher resolutions (1080p / 4k), use Seedance 2.0 or Veo 3.1.
Veo 3.1 — Limited aspect ratios
Veo 3.1 only supports 16:9 / 9:16 (plus Auto). Square (1:1), 4:3 / 3:4 portrait, and other ratios are not available. For wider ratio support, use Seedance 2.5 or Seedance 2.0.
Veo 3.1 — 4k costs no extra credits, but needs a paid account
720p / 1080p / 4k all cost the same on Veo 3.1, so 4k costs no extra credits. Two things to know: 4k is locked for free accounts (selecting it opens the upgrade dialog), and rendering takes longer — 4k typically takes 6–8 minutes versus 1–4 minutes for 720p / 1080p. The same 4k lock applies to Seedance 2.0 Pro.