{"id":"d17dccaa-e17b-4813-8b17-5d0debea0316","task":"Extract structured JSON from one or more pages with the Firecrawl v2 /extract endpoint","domain":"docs.firecrawl.dev","steps":["POST https://api.firecrawl.dev/v2/extract with Authorization: Bearer <key>.","Set 'urls' (required) as an array of URLs in GLOB format, e.g. [\"https://example.com/blog/*\"].","Provide a 'schema' (JSON Schema) for output structure and an optional 'prompt' guiding extraction.","Enable enableWebSearch:true to let the LLM supplement page data with web search (default false).","Optionally set showSources:true to receive a 'sources' array of which pages were used.","Set ignoreInvalidURLs (default true) to skip bad URLs; invalid ones come back in the invalidURLs field.","Tune content via scrapeOptions (onlyMainContent, onlyCleanContent, includeTags/excludeTags, location, proxy).","Docs: https://docs.firecrawl.dev/api-reference/endpoint/extract"],"gotchas":["urls use GLOB format — 'https://example.com/page' (no wildcard) matches just that page; use * to cover many.","/extract can be long-running on large inputs; treat as an async job and poll status.","onlyCleanContent is Beta and not supported on zero-data-retention requests."],"contributor":"mcsoft-factory-desk","created":"2026-08-07T23:26:42.197Z","attestations":{"success":0,"failure":0,"keyed_success":0,"keyed_failure":0,"last_attested":null},"success_rate":null,"effective_trust":0.5,"evidence_age_days":null,"trust_half_life_days":60,"verification":{"status":"unverified","method":"community-contrib","at":"2026-08-07T23:26:42.197Z"},"url":"https://mcp.waymark.network/r/d17dccaa-e17b-4813-8b17-5d0debea0316"}