Firecrawl scrape formats: the seven Sume supports
Firecrawl scrape lists summary, branding, question and more. Sume's scrape accepts seven formats: markdown, links, html, json, screenshot, rawHtml, images.

Sume's scrape formats array accepts seven values: markdown (the default), links, html, json, screenshot, rawHtml and images. Firecrawl's scrape reference lists more, including Summary, Branding, Product, Menu, Audio, Video, Question and Highlights. Those have no Sume equivalent.
Vendor list from Firecrawl's scrape reference, Sume's from its OpenAPI enum and MCP tool description, read 2026-09-30.
Which Firecrawl formats exist on Sume?
The names that match are Markdown, HTML, Raw HTML, Links, Images, Screenshot and JSON. Sume spells them markdown, html, rawHtml, links, images, screenshot and json.
| Firecrawl format | On Sume |
|---|---|
| Markdown | markdown (default) |
| HTML | html |
| Raw HTML | rawHtml, returned as raw_html |
| Links | links |
| Images | images, returned as image_urls |
| Screenshot | screenshot |
| JSON | json with json_schema or json_prompt |
| Summary, Branding, Product, Menu | Not available |
| Audio, Video, Question, Highlights | Not available |
| Raw Base64, Change Tracking | Not available |
What if I relied on summary or question?
Use the json format, which is Sume's structured-answer path. A json_prompt of up to 2,000 characters can ask for a summary or one answer; see the question-format post. For brand colors or product fields there is nothing to switch on, so describe the fields in json_schema and read them from the page.
What else shapes the output?
only_main_content defaults to true, so navigation and footers are trimmed. timeout_ms goes up to 30000, wait_for_ms from 0 to 10000, and actions takes up to 8 steps. Output is bounded and a truncated marker says when it was cut; page content is untrusted.
How many formats can one request ask for?
formats takes 1 to 7 unique values, so one call can return markdown and links together. Ask only for what you will read: rawHtml and html are large.
const res = await fetch("https://api.sume.com/v1/firecrawl/scrape", {
method: "POST",
headers: {
Authorization: "Bearer " + process.env.SUME_API_KEY,
"Content-Type": "application/json",
},
body: JSON.stringify({ url: "https://example.com", formats: ["markdown", "links"] }),
});
console.log(res.status, await res.json());Sources
Related posts
More in Developers
- Firecrawl scrape timeout: 60000 ms default vs Sume's 30000 cap
Firecrawl's scrape timeout defaults to 60000 ms and goes to 300000. Sume's crawl scrape takes timeout_ms up to 30000, so send 30000 or less.
- Firecrawl waitFor and wait actions: Sume's 10 s combined budget
Firecrawl allows 60 s of waitFor plus wait actions combined. Sume's crawl_scrape allows 10000 ms combined, and the total must stay below timeout_ms.
- FLUX image URL expires after 10 minutes: what Sume returns
Black Forest Labs says generated FLUX image URLs expire after 10 minutes. Sume returns Sume-hosted signed URLs; here is how to fetch and keep the file.
- frame_images plus input_references in one request: which wins?
When a Sume video request carries both frame_images and input_references, frame_images wins and the job runs as image-to-video. What that means for references.
Written by Sume