Lyria 3.5 has no Batch API: queue many tracks on Sume instead
Google's Lyria 3.5 page says batch, Flex and Priority inference are not supported. Sume accepts many music jobs and queues them past your concurrency.

Google's Lyria 3.5 model page says batch, Flex and Priority inference are not supported (read 2026-10-01). On Sume you can still submit many music jobs: valid jobs are accepted as queued when your workspace is at its processing limit, and run as slots free up.
How does queueing work?
Sume's generation admission docs say concurrency is a dispatch limit, not a submit limit. Processing concurrency is plan-based: Free 1, Pro 4, Startup 8, Scale 20. Queue capacity defaults to max(3, concurrency x 5). A full queue returns 429 queue_full, and too many requests returns 429 rate_limited.
What does each side document?
| Topic | What is documented |
|---|---|
| Google batch, Flex, Priority | Not supported on the Lyria 3.5 page |
| Sume submit when busy | Accepted as queued, runs as slots free up |
| Sume concurrency by plan | Free 1, Pro 4, Startup 8, Scale 20 |
| Sume queue full | 429 queue_full |
How do I submit a list?
One POST per track, each with its own Idempotency-Key, then store the job ids. Do not treat queued as a failure.
import os
import requests
H = {"Authorization": "Bearer " + os.environ["SUME_API_KEY"]}
names = ["Maya", "Leo", "Sam"]
jobs = []
for i, n in enumerate(names):
r = requests.post(
"https://api.sume.com/v1/music-router/generate",
headers={**H, "Idempotency-Key": f"bday-batch-{i}"},
json={"model": "lyria-3.5", "prompt": f"Birthday pop song for {n}"},
)
r.raise_for_status()
jobs.append(r.json()["id"])
print(jobs)Does Sume price differ for batch-style use?
Sume's price is a fixed $0.125 per generation with no batch discount in the docs. Google's page does not give a Sume-comparable number here, so do not assume a saving either way.
Sources
Related posts
More in Developers
- Same Lyria 3.5 prompt, different song: why, and no seed on Sume
Google says Lyria results vary between calls. Sume's music API has no seed. How to keep the take you like and generate a few to choose from.
- MAI-Image-2.6-Flash API: not on Sume, models to use instead
Microsoft's MAI-Image-2.6 and 2.6-Flash launched 2026-09-04 in Foundry preview. Sume's image catalog does not list them; check GET /v1/images/models.
- MAI-Voice-2.1: 23 languages and voice matching; Sume takes a voice id
Microsoft says MAI-Voice-2.1 covers 23 languages and matches a voice from a short clip. Sume TTS takes a ready voice id or avatar, not reference audio.
- MAI-Voice-2.1-Flash lists 45 ms; Sume TTS is an async job
Microsoft lists MAI-Voice-2.1-Flash at about 45 ms and MAI-Voice-2.1 at about 550 ms. Sume TTS 1.0 is non-streaming: a job with a poll URL or webhook.
Written by Sume