Gemini TTS voice design voice_ id vs Sume voice ids

Gemini voice design returns a persistent voice_ id from a text prompt. Sume TTS accepts only a voice UUID or voi_ library id, and rejects other shapes with 400.

4 min readSume
All posts

A Gemini voice_... id will not work as a Sume voice.id. Sume accepts a TTS voice UUID (8-4-4-4-12 hex) or a Voices library id (voi_ plus 32 hex), and any other shape is rejected synchronously with 400 public_reason=invalid_voice_id before a job is queued or credits are reserved.

Gemini facts are from its speech-generation guide; Sume facts from the API reference. Read 2026-10-01.

What does Gemini voice design create?

The guide says you can generate a custom vocal persona from a natural-language description in Google AI Studio or with POST /v1beta/voices and type="prompted", which returns a persistent voice_... id and a sample_audio WAV preview. Stateful voices are limited to 200 per project, shared across prompted and replicated voices, with a one-year retention.

Voice ids from the Gemini guide and the Sume API reference, read 2026-10-01.
ItemGemini 3.8 TTSSume TTS 1.0
Id shapevoice_... (stateful), voicekey_... (stateless replicated)UUID or voi_ + 32 hex
Wrong shapeNot covered here400 invalid_voice_id, no credits reserved

How do I choose a voice on Sume?

Either copy a voice id verbatim, or send avatar_id or avatar_handle and let Sume resolve the voice. The reference calls the avatar reference the discoverable selector: list avatars with GET /v1/avatar-1.0/avatars and use one whose voice.status is ready. If both an avatar and voice.id are set they must match, or the request fails with 400.

Can I design a voice from a prompt on Sume?

Not through the TTS request: it takes an existing id, not a description. Making a voice from a text description is covered separately in AI voice from a text description.

Sources

Related posts

More in Developers

All Developers posts

Written by Sume