Gemini tiers unlock by spend and days. Sume limits follow the plan
Gemini upgrades a project after spend and elapsed days. Sume sets request and concurrency limits by plan, and a prepaid top-up does not raise concurrency.

Gemini raises a project's rate limits as it meets spend and time thresholds, while Sume ties limits to the subscription plan. Google's page lists Tier 2 as $100+ spending plus 3 days elapsed and Tier 3 as $1,000+ plus 30 days. Sume's docs say processing concurrency is plan-only and that prepaid top-ups do not raise it.
Google facts are from its rate limits page (listed under Sources) and Sume facts from Generation admission and Authentication, read 2026-09-30.
How does each side decide your limits?
Google says tiers upgrade automatically once a project meets the qualification criteria. Sume reads the plan of the workspace the key belongs to.
| Plan | Processing concurrency | Queue capacity | Writes per minute |
|---|---|---|---|
| Free | 1 | 5 | 120 |
| Pro | 4 | 20 | 300 |
| Startup | 8 | 40 | 600 |
| Scale | 20 | 100 | 1200 |
Will topping up credits lift my concurrency?
No. The docs say generation concurrency is plan-only and prepaid top-ups do not raise the processing limit. Admin overrides can raise the effective concurrency_limit (limit_source: admin_override), and organization workspaces have a floor of 10. The dashboard Concurrency tab is the source of truth, exposed as generation_limits.concurrency_limit.
What does the queue do when concurrency is full?
It accepts valid jobs as queued while queue capacity remains and moves them to processing as slots open. Only when accepted capacity (concurrency plus queue) is full does a submit fail with 429 queue_full. Enterprise request limits are contract-specific; until provisioned, an Enterprise key resolves to the Scale row.
What are the limits of this comparison?
The two systems measure different things: Gemini limits are requests and tokens per project, Sume's headline limit here is concurrent generation jobs and requests per key. The static table is the shipped default; read your own generation_limits before planning a batch.
Sources
Related posts
More in Pricing
- Genjutsu 1080p: Sume serves it at 480p or 720p only
Higgsfield lists Genjutsu output up to 1080p. On Sume, Genjutsu Motion Transfer accepts 480p (default) or 720p, and any other value is rejected.
- GPT Image 2.5 reference image cost: Runway credit vs Sume tokens
Runway charges 1 credit per GPT Image 2.5 reference image, once per request. Sume prices reference images as input image tokens inside the endpoint price.
- GPT-Live 1 pricing per second vs Sume's per-job audio billing
GPT-Live 1 voice sessions cost $0.05 per minute billed per second. Sume bills TTS per character and STT per audio minute, each reserved at submit.
- Hedra API pricing: estimate endpoint vs Sume dry_run
Hedra says you can price the exact request with an estimate endpoint. Sume's hosted MCP has dry_run and max_spend_usd gates that preview cost before a job.
Written by Sume