Scrape one or more URLs. Returns clean, LLM-optimized markdown. Multiple URLs are scraped concurrently.
Scanned 9/29/2026
npx -y skills add Tekkiiiii/the-agency --skill firecrawl-scrape --agent claude-codeInstalls into .claude/skills of the current project.
Are you the author of Firecrawl Scrape?
Add the live security badge to your README — it updates automatically with every re-scan.
[](https://www.skillsdirectory.com/skills/tekkiiiii-firecrawl-scrape)More formats (shields.io, HTML) on the badges page. Keep it an A: scan every change in CI with Pro.
# Firecrawl — Scrape
Scrape one or more URLs. Returns clean, LLM-optimized markdown. Multiple URLs are scraped concurrently.
## When to Apply
- You have a specific URL and want its content
- The page is static or JS-rendered (SPA)
- Step 2 in escalation pattern: search → scrape → map → crawl → browser
## Quick Start
```bash
# Basic markdown extraction
firecrawl scrape "<url>" -o .firecrawl/page.md
# Main content only, no nav/footer
firecrawl scrape "<url>" --only-main-content -o .firecrawl/page.md
# Wait for JS to render, then scrape
firecrawl scrape "<url>" --wait-for 3000 -o .firecrawl/page.md
# Multiple URLs (concurrent)
firecrawl scrape https://example.com https://example.com/blog
# Get markdown and links together
firecrawl scrape "<url>" --format markdown,links -o .firecrawl/page.json
# Ask a targeted question (costs 5 extra credits)
firecrawl scrape "https://example.com/pricing" --query "What is the enterprise plan price?"
```
## Options
| Option | Description |
|--------|-------------|
| `-f, --format <formats>` | Output: markdown, html, rawHtml, links, screenshot, json |
| `-Q, --query <prompt>` | Ask a question about page content (5 credits) |
| `-H` | Include HTTP headers |
| `--only-main-content` | Strip nav, footer, sidebar |
| `--wait-for <ms>` | Wait for JS rendering before scraping |
| `--include-tags <tags>` | Only include these HTML tags |
| `--exclude-tags <tags>` | Exclude these HTML tags |
| `-o, --output <path>` | Output file path |
## Tips
- **Prefer plain scrape over `--query`.** Scrape to a file, then grep/head/read the markdown — cheaper and more flexible.
- **Try scrape before browser.** Handles static pages and SPAs. Only escalate to browser for interactions (clicks, forms, pagination).
- **Multiple URLs run concurrently** — check `firecrawl --status` for concurrency limit.
- **Always quote URLs** — `?` and `&` are shell special characters.
- Naming convention: `.firecrawl/{site}-{path}.md`
## Session Fetch Cache (ETag/304 Revalidation)
When scraping the same URL repeatedly within a session or across sessions, use HTTP conditional requests to avoid re-fetching unchanged content. This saves Firecrawl credits and reduces latency.
**How it works:**
1. On first scrape, capture the `ETag` or `Last-Modified` response header and save it alongside the output file
2. On subsequent scrapes, send `If-None-Match: <etag>` or `If-Modified-Since: <date>` in the request
3. If the server returns HTTP 304 Not Modified, the cached file is still valid — skip the scrape, use what is on disk
**Practical guidance:**
- Save the ETag alongside the scraped content: `page.md` + `page.md.etag`
- Before scraping, check if `page.md.etag` exists and pass its value via `--header "If-None-Match: <etag>"` to firecrawl
- Treat a 304 response as a cache hit — no new content, no new credits consumed
- Not all servers send ETags. If absent, fall back to `Last-Modified` header, or full scrape
**When to use:** Any research workflow that revisits the same documentation URLs, pricing pages, or reference pages in multiple sessions. Especially valuable for `/auto-researcher` workflows that check the same sources repeatedly.
## Related Skills
- `firecrawl-search` — find pages when you don't have a URL
- `firecrawl-browser` — when scrape can't get the content (interaction needed)
- `firecrawl-download` — bulk download an entire site to local files
---
**Source:** https://officialskills.sh/firecrawl/skills/firecrawl-scrapeIs this your skill, or is something wrong with this listing? Request removal or report an issue. Author removals are honored within 72 hours.
No comments yet. Be the first to comment!