Resolve Cloud Presentations subtitles: burn cues into a copy
Resolve 21.1 added subtitle support to Cloud Presentations. For a review copy outside Resolve, Sume burns cues onto a video; SRT uploads are unsupported.

Blackmagic says Cloud Presentations in DaVinci Resolve 21.1 gained subtitle support. Sume does not touch that feature; what it offers for a review copy is burned-in captions: POST /v1/video-captions draws cues onto a video, and SRT uploads are unsupported, so pass the text as cues.
Vendor facts are from the 21.1 announcement; Sume facts from Video captions, read 2026-10-01.
What did Resolve 21.1 add to Cloud Presentations?
The announcement says Blackmagic Cloud Presentations has been updated with subtitle support, annotation sharing and comment replies, to make collaboration with clients on review and approval easier. It gives no more detail than that sentence, so this post does not describe how it works.
How do I make a reviewable copy with subtitles in Sume?
Standalone captions burn text onto an existing public HTTPS video URL. If you already have the wording, send cues (or segments) with text, start and end in seconds. The docs say that overlay path skips speech-to-text.
The words are then part of the pixels: the viewer cannot switch them off, so keep the clean version too.
curl -X POST https://api.sume.com/v1/video-captions \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-H "Idempotency-Key: review-cues-001" \
-d '{
"video_url": "https://media.sume.com/artifacts/example/clean.mp4",
"style": "punch",
"cues": [
{ "text": "Review copy, v2", "start": 0, "end": 2.5 },
{ "text": "Logo moves left", "start": 2.5, "end": 5 }
]
}'Can I upload an SRT file?
No. The docs state that SRT uploads and provider task ids are unsupported, and tell you to pass phrase-level text as cues or segments instead. Convert the SRT lines to text, start and end yourself; the three fields are the whole shape.
What does it cost and how long can the video be?
The docs say each accepted standalone caption job reserves and captures $0.20 USD of Sume usage for videos up to 60 seconds under the current fixed estimate, and tell you to confirm live pricing in GET /v1/catalog and OpenAPI.
| Item | Docs say |
|---|---|
| Input | video_url, public HTTPS |
| Authored text | cues / segments skip speech-to-text |
| SRT | Unsupported |
| Price | $0.20 USD up to 60 seconds, fixed estimate |
What if the clip has no speech?
Without cues, a silent clip fails as caption_no_speech. Passing cues avoids that; see silent clips and overlay cues. For the difference between burned-in text and a sidecar file, read captions vs subtitles.
Sources
Related posts
More in Use cases
- Multimaster output sizes: repeat video trim with new width and height
For several deliverable sizes from one master, run Sume video trim once per size with output width and height 256 to 2160, fps 24/25/30/60, precision exact.
- Resolve 21.1 source viewer aspect ratio: check clips first
Resolve 21.1's source viewer keeps each clip's original aspect ratio. To check clips before conforming, Sume video inspect returns stills with width and height.
- AI presenter for a SaaS demo video, with or without a product image
For a SaaS demo, omit `product_image` for a productless avatar video, or add one. Each Sume avatar job runs 4 to 60 seconds, so longer demos are split.
- AI video background swap for a fashion lookbook with Seedance 2.5
Seedance 2.5 describes green screen editing that swaps a background and keeps the subject. On Sume, send the garment clip as a video reference, 4 to 30 seconds.
Written by Sume