OmniHuman 1.5 API: native 1080p vs Sume avatar 720p
BytePlus says OmniHuman 1.5 outputs native 1080p from one image plus audio. Sume's Avatar Video currently outputs 720p; its video models list 1080p separately.
BytePlus describes OmniHuman 1.5 as video from a single image and multimodal prompts, with native 1080p output. Sume's Avatar Video is not that model: its resolution is currently 720p. If 1080p matters more than the avatar workflow, Sume's general video models list 1080p in their catalog entries.
What does the BytePlus page claim?
The page lists four capabilities: video from a single image with audio, image and text prompts; native 1080p-resolution output; rhythmic, emotional and multi-person performances; and gesture control with camera movement. These are vendor claims as read 2026-09-30; this post does not test them.
What limits does Sume's Avatar Video have?
From the Avatar videos docs, read 2026-09-30:
| Field | Documented value |
|---|---|
resolution | Currently 720p |
aspect_ratio | 1:1, 3:4, 9:16, 4:3, 16:9; default 9:16 |
| Avatars per video | One resolved avatar per final video |
| Quality | standard, plus (default), max |
So is multi-person output possible on Sume?
Not in one Avatar Video. The docs say current execution supports one resolved avatar per final video and expects scene backgrounds to resolve to one shared scene. For two speakers, see AI avatar conversation video with two speakers for how Sume handles that case.
Where does Sume list 1080p?
The video generation catalog returns supported_resolutions per model, with example values such as 480p, 720p and 1080p. Read that field for the model you plan to use; it is the source of truth for resolution, as described in Video generation. The quality tier max on Avatar Video is slower turnaround, not a higher resolution.
Sources
Related posts
More in Models
- PixVerse API and R2 world model: Sume has no live session
PixVerse R2 is a persistent real-time world. Sume has no live session or action controls: video jobs return a clip you poll for or get by webhook.
- Recraft V4.1 Flash API: Sume lists V4, webp, text-only
Recraft V4.1 Flash is live in Recraft Studio and its API. Sume lists recraft/recraft-v4, not a Flash id: webp output, text-to-image only, no references.
- Medical transcription API: what Sume STT does and doesn't
Sume has one speech-to-text route, sume/stt-1.0, with no medical model option. What it takes and returns, so you can decide for clinical audio.
- Seedance 1.5 Pro retires Nov 11: which Seedance ids Sume lists
ByteDance retires Seedance 1.5 Pro on November 11, 2026. Sume's video docs list seedance-2.5 and seedance-2; use bare catalog ids and check the models endpoint.
Written by Sume