Which Sume video models accept a reference video?
Nine Sume video ids take a video input: the four Seedance 2.x ids, Wan 3.0, both MiniMax H3 models, Omni Flash and Genjutsu. Kling and Grok do not.

Nine Sume video ids accept a video input: seedance-2.5, seedance-2, seedance-2-fast, seedance-2-mini, wan-3.0, minimax-h3, minimax-h3-max and gemini-omni-flash-1.1, plus a Higgsfield motion-transfer model, which needs one. kling-3 and grok-imagine-video-1.5 do not take one.
A video input does three different jobs depending on the model, so read the mode before you choose. Facts are from Sume's catalog and video docs, read 2026-10-01.
What does the video do, and what are the limits?
| Model | Role of the video | Limits |
|---|---|---|
wan-3.0 | reference | up to 5 videos, 15 seconds combined, at least 16 fps |
minimax-h3, minimax-h3-max | reference | up to 3 videos of 2 to 15 seconds, 15 combined; 12 inputs in all |
gemini-omni-flash-1.1 | reference, or edit source | reference: up to 3 videos of 3 seconds each; edit: one video_url |
| Higgsfield motion-transfer model | motion transfer source | one source video plus 1 to 8 reference images, 480p or 720p |
| Seedance 2.x ids | reference | supported; counts not listed in the catalog |
Which model cannot run without a video?
The Higgsfield motion-transfer model needs one source video, plus 1 to 8 reference images, at 480p or 720p. The other video-input ids treat the video as optional.
Which one edits a video from a prompt?
Only gemini-omni-flash-1.1 has an edit mode: send video_url and a prompt to POST /v1/video-router/generate. It cannot be combined with frames or reference lists, and aspect_ratio is rejected.
curl -X POST https://api.sume.com/v1/video-router/generate \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: edit-001" \
-d '{
"model": "gemini-omni-flash-1.1",
"prompt": "Make it night, keep the motion identical",
"video_url": "https://example.com/clip.mp4",
"resolution": "720p"
}'Sources
Related posts
More in Models
- YouTube fhd, qhd, uhd thumbnails in the API: a 3840 AI image
YouTube's API notes fhd, qhd and uhd thumbnail keys for some videos. Sume's ChatGPT Image 2.5 image_size accepts a 3840 edge, enough for a 4K image.
- An OpenRouter-compatible video API: sume/auto or a pinned model
Sume's POST /v1/videos follows OpenRouter's video generation API field for field. Let sume/auto pick the model, or pin a catalog id like seedance-2.5.
- Image generation API with reference images: POST /v1/images
Send a prompt plus public HTTPS reference images to Sume's POST /v1/images. Pin a catalog model or send sume/auto; the catalog lists each model's limits.
- Video 1.0 and Image 1.0 are retiring soon: move to sume/auto
Sume Video 1.0 and Image 1.0 are retiring soon and already run as aliases for the Auto path. New integrations call /v1/videos or /v1/images with sume/auto.
Written by Sume