Crunchbase Scraper

Company profiles,
as a spreadsheet.

Give it a list of crunchbase.com organization URLs and get their profiles back as rows: the company name and description, the website, when it was founded, headcount, industries, where it is, its Crunchbase rank and operating status, and its LinkedIn, X and Facebook links. Sixteen columns, one row per company. One thing to know before you start: Crunchbase sits behind Cloudflare, and this scraper needs a residential proxy to return anything at all.

one-time 500 free rows$0.002 per row after16 columnsCSV · XLSX · JSON
How it works

A company list in,
their profiles out.

The input is the organization page - the URL you would have opened yourself. Paste the ones you already have and the job reads each profile.

  1. STEP 1Sign in to the platform.
  2. STEP 2Open the Crunchbase Scraper.
  3. STEP 3Paste crunchbase.com/organization/ URLs, one per line - or upload a CSV, XLSX, TXT or Parquet file.
  4. STEP 4Configure a residential proxy - without one the run returns no rows.
  5. STEP 5Set a limit per query, or leave it at zero to take everything.
  6. STEP 6Click Get Data.

One row per company, each tagged with the URL it came from - so a run across a portfolio's worth of organizations still reconciles back to your input list.

Why teams use it

The firmographics,
in one file.

Profile fields, not a scrape of the page

Sixteen named columns come back per organization - the description, the website, the founding date, the headcount, the industries, the location, the rank, the operating status and three social links. You get fields to filter on, not HTML to pick apart afterwards.

The proxy is a real requirement, not a footnote

Crunchbase sits behind Cloudflare, and it does not yield to a datacentre address. This scraper needs a residential proxy configured against the job, and without one a run completes and returns nothing. We would rather you read that here than discover it on an empty export.

One control, set before the run

A cap on the rows per query, chosen before the job starts, or left at zero to take everything. There is no sort option on this scraper - you are naming the companies, so the order is the order you gave.

What you get back

Sixteen columns,
one row per company.

Each row carries the organization profile as Crunchbase presents it: what the company is, where it is, how big it is, how Crunchbase ranks it, and where to find it elsewhere on the web.

The column list below is the header row of real run exports, corroborated by the column array the run UI renders - the two agree exactly. What it is not is a sample: read the note under the table before you write code against any of these fields, because it says plainly what we have and have not seen this scraper return.

Data dictionary

Sixteen columns,
named and nothing more.

The names are exact - they are the workbook header row and the run UI's column array, which match. The descriptions say what each column is for. They do not describe its format, and the note under the table explains why not.

query
The crunchbase.com organization URL you submitted, repeated on every row that came from it.
company
The organization’s name as Crunchbase records it.
description
The profile’s description of what the company does.
website
The company’s own website.
founded
When the company was founded.
employees
The headcount Crunchbase shows for the company.
industries
The industries the profile is tagged with.
city
The city in the company’s headquarters location.
region
The region or state in that location.
country
The country in that location.
rank
The rank Crunchbase assigns the organization.
operating_status
Whether the company is still trading, as Crunchbase records it.
linkedin
The company’s LinkedIn link from the profile.
twitter
The company’s X (Twitter) link from the profile.
facebook
The company’s Facebook link from the profile.
logo
The company’s logo image from the profile.

Read this before you plan around the table above. Every run export we hold for this scraper came back empty - five runs across 2026-06-30 to 2026-07-15, each a workbook with this header row and no data row beneath it. The cause is not a bug: Crunchbase is behind Cloudflare, and a run without a residential proxy completes and returns nothing. So the sixteen names are solid - two independent sources carry them, in this order - and nothing else here is. We publish no example values and no format for any field: not whether founded is a year or a full date, not whether employees is a band or a number, not whether rank is global or per category, not how industries separates its tags, not what vocabulary operating_status uses. Stating any of those would mean guessing. Point a free-tier run with your own proxy configured at one organization and read the first file you get: that is the only sample that will tell you the truth about your data.

Run controls

Set on the job,
not in the spreadsheet.

This scraper has one control besides the query box - a cap on the rows per query. The rest of what you decide up front is how you hand over the organization URLs, and the proxy the job runs through.

Organization URL input Limit per query Paste one per line CSV upload XLSX upload TXT upload Parquet upload Residential proxy
Common workflows

Three jobs people
most often run here.

A few examples of how teams use company profile data to answer a question they actually have.

Enrichment

Fill in a list you already have

You have the companies; what you are missing is the shape of them. One run turns a column of organization URLs into headcount, industries, location and website, so the list can be segmented instead of just read.

RevOps
Market mapping

Map a category on evidence

Take every organization in a space and put their descriptions, industries and locations side by side. What a market looks like is much easier to see in one file than across fifty browser tabs.

Strategy
Sourcing

Screen a pipeline before the first call

Founding date, headcount and operating status are the fields that decide whether a company is worth a conversation. Having them next to the social links means the desk research is done before anyone picks up the phone.

Investment
Pricing

Pay only for the companies
you actually pull.

No subscription, no minimum, no recurring bill. Your first 500 rows are on us - after that, pay-as-you-go at the same flat rate as every other scraper here.

Free tier

500 free rows - $0

Every new account, one-time. No credit card required. Per-query limits, file upload and every export format included. The proxy this scraper needs is yours to supply, on the free tier and after it.

$0 forever
Pay-as-you-go

$0.002 per row, after the free tier

The same flat rate as every other scraper on the platform. The pre-flight estimator shows the row count and credit cost before a run starts - no surprise bills, no compute units to translate.

Most popular
Volume

Custom · high volume

Volume pricing, dedicated workers and an SLA for continuous monitoring or very large historical pulls. Tell us your numbers and we will quote.

Talk to us
10% off your first paid run.Use code LIVESCRAPER10 at checkout.
Sign up
Pairs well with

The profile,
and everything around it.

The legal bit

Is it legal to scrape
Crunchbase profiles?

Short answer: organization pages are public business data - and the constraint here is technical and contractual rather than privacy.

A Crunchbase organization page is shown to anyone who opens it, signed in or not. The company name, the description, the website, the location and the social links are published to be read. Collecting publicly visible company information for research is long-established practice, and nothing here touches a login or a paywall.

This export is business data rather than personal data - a company, its size, its industries and its links, not a named individual. That is a lighter obligation than a reviews or profiles export, though it does not make it none: a one-person company is still a person, and a list of small firms can shade into personal data faster than people expect.

The constraint worth planning around is Crunchbase's own. The site is actively protected by Cloudflare, and its terms restrict automated access - so this is a terms question as well as a technical one, and if you have a subscription or an API agreement with them you should check it before scaling. We run no third-party trackers on the data layer, and your exports auto-delete after 30 days.

livescraper.app · principles
Public organization pages only
No logins, no accounts touched
Needs your own residential proxy!
Business data, not personal data
Exports auto-delete (30 days)
Check Crunchbase's own terms before scaling.
Common questions

Things people
ask before signing up.

The questions we hear most. Anything else? Talk to us - humans, not bots, write the answers.

How do I scrape Crunchbase company profiles?+
Using the Crunchbase Scraper:
  1. Sign in to the platform.
  2. Open the Crunchbase Scraper.
  3. Paste crunchbase.com/organization/ URLs, one per line - or upload a CSV, XLSX, TXT or Parquet file.
  4. Configure a residential proxy - without one the run returns no rows.
  5. Set a limit per query, or leave it at zero to take everything.
  6. Click Get Data.
Do I need a proxy for this one?+
Yes, and it is the single most important thing on this page. Crunchbase is protected by Cloudflare, which does not serve data to a datacentre address. This scraper needs a residential proxy configured against the job. Without one the run completes normally and the export comes back with a header row and nothing under it.
Can I search by company name or domain instead of a URL?+
Not with this scraper. This one takes crunchbase.com/organization/ URLs, one per line. Resolving a name or a domain to the right organization is a separate service - the Crunchbase Search Scraper - so if your list is company names rather than links, that is the one you want first.
What comes back for each company?+
Sixteen columns: the query you submitted, the company name and description, the website, the founding date, the headcount, the industries, the city, region and country, the Crunchbase rank, the operating status, and the LinkedIn, X and Facebook links plus the logo.
Why does this page show no example values?+
Because we will not print values we have not seen. Every run export we hold for this scraper returned zero rows - five runs between 2026-06-30 and 2026-07-15, each a workbook with the header row and no data beneath it, all made without a residential proxy. That makes the sixteen column names trustworthy, because two independent sources list them in this order, and makes any claim about their contents a guess. Our other pages describe field formats because we measured them; this one stays silent because we could not.
My run came back empty. What went wrong?+
Almost certainly the proxy. A Crunchbase job without a residential exit address is served a Cloudflare challenge rather than a profile, and the job finishes cleanly with nothing in it - the same empty file our own test runs produced. Check the proxy first, and check the URL is an organization page, before you look anywhere else.
Can I sort or limit the companies?+
You can limit them. Set a limit per query, or leave it at zero to take everything. There is no sort control on this scraper - you are naming the organizations yourself, so the order is the one you supplied.
How much does it cost?+
The first 500 rows are free and one-time, with no credit card. After that it is $0.002 per row - the same flat rate as every other scraper on the platform. The estimator shows the cost of a run before it starts. The proxy is a separate cost, and it is yours.

Your first 500 companies,
on the house.

500 one-time free rows on every new account - no expiry. After that it is $0.002 per row, pay-as-you-go - no card on file until you say so. Bring your own residential proxy.

Activates instantly · no card required

Scrape Crunchbase company profiles at scale

Livescraper's Crunchbase Scraper turns a list of organization pages into company data. You submit crunchbase.com/organization/ URLs - typed one per line, or uploaded as a CSV, XLSX, TXT or Parquet file - cap the rows per query if you want to, and download the profiles as a clean CSV, Excel or JSON file. One row per company, sixteen columns.

Each row carries the profile as Crunchbase presents it: the company name and description, the website, the founding date, the headcount, the industries the profile is tagged with, the city, region and country of its headquarters, the rank Crunchbase assigns it, its operating status, and its LinkedIn, X and Facebook links alongside the logo. RevOps teams use it to enrich a list of organization URLs into something segmentable. Strategy teams map a category by putting descriptions, industries and locations side by side. Investment teams screen a pipeline on founding date, headcount and operating status before the first call.

One requirement is not optional and is stated here rather than buried. Crunchbase is protected by Cloudflare, and this scraper needs a residential proxy to return anything: without one, a job finishes cleanly and the export contains a header row and no data. That is the behaviour our own test runs produced, and it is the first thing to check if a run comes back empty. Note too that this scraper takes organization URLs only - resolving a company name or a domain to the right profile is the separate Crunchbase Search Scraper.

This page also does something the sibling pages do not, on purpose. It publishes no example values and no format for any field - not whether the founding date is a year or a full date, not whether the headcount is a band or a number, not what vocabulary the operating status uses - because every run export we hold for this scraper returned zero rows, and a value we have not seen is not a value worth printing. The sixteen column names are exact and cross-checked; everything else is for your first free-tier run to tell you. Start free: your first 500 rows cost nothing and need no credit card, and after that it is $0.002 per row, flat.