AI presenter for a SaaS demo video, with or without a product image
For a SaaS demo, omit `product_image` for a productless avatar video, or add one. Each Sume avatar job runs 4 to 60 seconds, so longer demos are split.

A SaaS demo presenter on Sume is one POST /v1/avatar-1.0/talking-video call with a ready avatar and a script. A SaaS product has no physical item to hold, so omit product_image and you get a productless avatar video. Each job must estimate at 4 to 60 seconds, so a longer demo is several jobs.
HeyGen's updates feed lists BrandKit and Video Podcast among its recent announcements; this post covers only the Sume route. Facts are from the avatar video docs, read 2026-10-01.
When do I add a product image?
product_image is optional. Leave it out for a presenter talking over a scene, and add it when a physical or packaged item should appear in the shot. A screenshot of your UI is a different job: use a photo scene (scene: { "type": "photo", "image_url": "https://..." }) as a scene reference, or a prompt scene for direction. Media URLs must be fetchable public HTTPS.
How long can the demo be?
Scripts and multi-scene plans are accepted when the estimated duration is 4 to 60 seconds inclusive. Shorten the script or split it into several jobs. With ordered video_inputs you can compose hook, demo and silence beats in one video, but the total must still land in the same window. If you turn on inline captions, an estimate above 60 seconds is rejected.
What does it cost?
Avatar video is billed per second and the rate depends on quality (standard, plus default, max) and whether a product is present.
| Quality | No product | With product |
|---|---|---|
standard | $0.184 per second | $0.194 per second |
plus (default) | $0.245 per second | $0.258 per second |
max | $0.55 per second | $0.58 per second |
What should I do for a five-minute demo?
Split the script into scenes of under 60 seconds, make one job per scene, and join them with Timeline 1.0; the steps are in avatar video longer than 60 seconds.
Sources
Related posts
More in Use cases
- AI video background swap for a fashion lookbook with Seedance 2.5
Seedance 2.5 describes green screen editing that swaps a background and keeps the subject. On Sume, send the garment clip as a video reference, 4 to 30 seconds.
- Seedance 2.5 green screen editing: what Sume can do
ByteDance says Seedance 2.5 improves green screen editing. Sume's docs document an edit mode only for Gemini Omni Flash 1.1; Seedance takes video references.
- AI workout video generator: Seedance 2.5, 4 to 30 seconds
For a workout demo, one Seedance 2.5 request runs 4 to 30 seconds on Sume, with video and audio references; most other catalog models stop at 15 seconds.
- Synthesia Avatar Builder credits: 14 per option, and Sume jobs
Synthesia Avatar Builder charges 14 credits per generated option. Sume creates an avatar with one job per request via POST /v1/avatar-1.0/generate.
Written by Sume