FLUX.2 prompt guide: structure, length and no negatives
Black Forest Labs' FLUX.2 prompt rules: subject, action, style, context; 30-80 words; no negative prompts. Which apply to flux.2-pro on Sume.

Write FLUX.2 prompts as subject, then action, then style, then context, put the most important element first, and aim for 30 to 80 words. Describe what you want instead of what you do not: FLUX.2 has no negative prompts. Those are Black Forest Labs' rules for its [pro] and [max] models. On Sume you send the finished prompt to black-forest-labs/flux.2-pro.
The prompt rules come from BFL's prompting guide; the Sume side from the Image API docs. All read 2026-09-29.
How should I structure a FLUX.2 prompt?
BFL's framework is Subject + Action + Style + Context: the main focus, what it is doing, the artistic approach or medium, and the setting, lighting or mood. It adds that word order matters, because FLUX.2 pays more attention to what comes first, so lead with the subject and the key action and leave secondary details for the end.
| Length | BFL says it fits |
|---|---|
| Short, 10-30 words | Quick concepts and style exploration |
| Medium, 30-80 words | Usually ideal for most projects |
| Long, 80+ words | Complex scenes requiring detailed specifications |
Does FLUX.2 take a negative prompt?
No. BFL's guide says FLUX.2 does not support negative prompts and gives the workaround: instead of "no blur", write "sharp focus throughout"; instead of "no people", describe an "empty scene". Sume's Image API request table has no negative-prompt field either, so the advice carries over unchanged.
Which FLUX.2 controls does Sume pass through?
BFL's quick reference lists a seed, and guidance and steps for [flex]. Sume's docs list seed as part of the schema that no model advertises, so sending one returns 400 unsupported_parameter. They do not list guidance or steps at all, so leave them out and put the control in the prompt.
What does a structured call look like on Sume?
One black-forest-labs/flux.2-pro image is estimated at $0.04 on Sume. The prompt below follows BFL's order: subject, action, style, context.
curl -X POST https://api.sume.com/v1/images \
-H "Authorization: Bearer $SUME_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "black-forest-labs/flux.2-pro",
"prompt": "A ceramicist shaping a bowl on a wheel, hands wet with slip, shot on a 35mm film camera, warm morning light in a small studio, sharp focus throughout",
"aspect_ratio": "3:2"
}'What else is in BFL's prompting guide?
The guide also covers hex color codes, JSON structured prompts for complex scenes, camera and lens references for photorealism, and multi-language prompting. Its best-practice line for color is to tie a code to an object: "The car is #FF0000" instead of "use red #FF0000 in the image". Try each technique on one image before you build a template around it.
Do these rules apply to the other Sume models?
BFL wrote the guide for FLUX.2 [pro] and [max]. This post makes no claim that the structure or the length bands transfer to other models on Sume. Test a few prompts per model, and read GET /v1/images/models for what each one accepts.
Sources
- Image API
- [Black Forest Labs docs: FLUX.2 [pro] and [max] prompting guide (read 2026-09-29)](https://docs.bfl.ml/guides/prompting_guide_flux2)
Related posts
More in Models
- FLUX 3 vs FLUX 2: which one you can call on Sume today
FLUX 3 vs FLUX 2 for image work: BFL's notes start FLUX 3 with video, and Sume's image catalog serves FLUX.2 Pro and Flex. What that means for stills.
- FLUX 3 video 20-second clips: what Sume's video API offers
FLUX 3 Video makes up to 20 s clips at 1920x1088. Sume does not list it, but seedance-2.5 and wan-3.0 accept 20 s requests. Durations by model.
- Gap or click when joining audio files: use wav, not mp3
Sume's Timeline audio joins up to 20 Sume-hosted parts with no silence at the seams. Its docs say to keep wav, because mp3 adds padding at every edge.
- Gemini Live Avatar vs a video avatar API: live or rendered?
Gemini 3.8 Live Avatar streams a lip-synced face in a conversation. Sume's avatar video renders a finished clip from a script. Which job needs which.
Written by Sume