Firecrawl Company Directories

by firecrawl94cc91229d6cISC185 starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated 6 weeks ago

Extract structured company lists from directories with Firecrawl. Use for scraping YC, Crunchbase, Product Hunt, G2, startup directories, category directories, or custom company databases into JSON, CSV, CRM-ready lists, or research tables.

AI-generated overview

Extracts structured company lists from startup and company directories using Firecrawl, producing JSON, CSV, or Markdown exports.

What it does
This skill guides an agent through turning startup or company directories into structured lists. It plans Firecrawl collection using browser mode for filtered, paginated, or infinite-scroll directories and scrape/map for public static listings, then captures visible fields such as name, description, industry, stage, location, team size, funding, tags, and profile or website URLs. It produces a Markdown export with a summary, company table or JSON/CSV link, sources, and rerun inputs, plus a defined JSON shape. It also sets quality expectations around deduplication, pagination tracking, and noting rate limits, login walls, or CAPTCHA blocks.
When to use it
Use it when you need to scrape directories such as YC, Crunchbase, Product Hunt, G2, or custom company databases into JSON, CSV, CRM-ready lists, or research tables. It fits cases where directory listings require filters, pagination, infinite scroll, or profile clicks, or where public static listings can be scraped or mapped.
Requirements
Requires a Firecrawl API key (FIRECRAWL_API_KEY) for hosted Firecrawl requests, plus network access to the target directories. It ships no scripts; it is instructions only.

Firecrawl Company Directories

Use this to turn startup or company directories into structured lists.

Onboarding Interview

Infer the directory, filters, result count, and output format from context. If the source is clear, proceed immediately.

Ask at most 1-3 concise questions only if blocked, such as the directory URL/name, required filters, or target result count.

Firecrawl Collection Plan

Use Firecrawl browser when the directory needs filters, pagination, infinite scroll, or profile clicks. Use scrape/map when listings are public and static.

Suggested sources include YC companies, Crunchbase, Product Hunt, G2 categories, or any custom directory URL.

Extraction Fields

Capture fields that are visible:

  • name
  • description
  • industry/category
  • stage/founded/location/team size/funding when visible
  • tags
  • directory profile URL
  • company website URL

Leave unavailable fields blank. Do not infer.

Final Deliverable

markdown
# Company Directory Export: [Source]
## Summary[Filters, count extracted, limitations]
## Companies[Table or link to JSON/CSV]
## Sources[Directory pages and profiles used]
## Rerun Inputsworkflow: firecrawl-company-directoriesdirectory: [source]filters: [criteria]max_results: [number]output: [json/csv/markdown]

JSON Shape

Use source, filters, extractedAt, totalResults, and companies[] with name, url, description, industry, stage, founded, location, teamSize, funding, tags, profileUrl, and websiteUrl.

Quality Bar

  • Deduplicate companies.
  • Track pagination progress.
  • Note rate limits, login walls, or CAPTCHA blocks.

Source and attribution

Source:firecrawl/firecrawl-workflowsinskills/firecrawl-company-directoriesat commit94cc912

License: ISC

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal

Firecrawl Company Directories Agent Skill | SourceWeft