Crunchbase Search Scraper

A name in,
the right company out.

Give it a company name, a domain or a URL and get back what Crunchbase matches to it - the name it holds, the link to the profile, its permalink, a short description and the entity type. Six columns, one row per matching company. One thing to know before you start: Crunchbase sits behind Cloudflare, and this scraper needs a residential proxy to return anything at all.

one-time 500 free rows$0.002 per row after6 columnsCSV · XLSX · JSON
How it works

Messy input in,
Crunchbase identities out.

The input is whatever your list already holds. A trading name, a website, a full URL - all three go in the same box, one per line.

  1. STEP 1Sign in to the platform.
  2. STEP 2Open the Crunchbase Search Scraper.
  3. STEP 3Paste company names, domains or URLs, one per line - or upload a CSV, XLSX, TXT or Parquet file.
  4. STEP 4Configure a residential proxy - without one the run returns no rows.
  5. STEP 5Set a limit per query, or leave it at zero to take everything.
  6. STEP 6Click Get Data.

Every row is tagged with the query it came from - which matters here more than usual, because one query can match more than one company.

Why teams use it

The step before
every other lookup.

A name, a domain or a URL

All three are accepted in the same field, one per line, or uploaded as a CSV, XLSX, TXT or Parquet file. That is the point of this scraper: your list rarely arrives as tidy Crunchbase links, and this is what turns it into them.

One row per match, not per query

A search is not a resolution. A query that matches several companies produces several rows, each tagged with the query it came from - so you can see the ambiguity and decide, rather than having one answer picked for you silently.

The proxy is a real requirement

Crunchbase sits behind Cloudflare, and it does not yield to a datacentre address. This scraper needs a residential proxy configured against the job, and without one a run completes and returns nothing. We would rather you read that here than discover it on an empty export.

What you get back

Six columns,
one row per match.

Each row is a company Crunchbase matched to your query: the name it holds, a link to the profile, its permalink, a short description and the entity type.

The column list below is the header row of real run exports, corroborated by the column array the run UI renders - the two agree exactly. What it is not is a sample: read the note under the table before you write code against any of these fields, because it says plainly what we have and have not seen this scraper return.

Data dictionary

Six columns,
named and nothing more.

The names are exact - they are the workbook header row and the run UI's column array, which match. The descriptions say what each column is for. They do not describe its format, and the note under the table explains why not.

query
The name, domain or URL you submitted, repeated on every row that came from it. With one query able to match more than one company, this is the column that tells you which rows belong together.
name
The company name Crunchbase holds for the match - which is not necessarily the string you searched for.
crunchbase_url
The link to the matched company on Crunchbase.
permalink
Crunchbase’s permalink for the match. Whether this is a bare slug or a full URL is not something we have observed, so the page does not say - see the note below before you build a join on it.
short_description
The short description Crunchbase carries for the match.
entity_type
The type Crunchbase classifies the match as.

Read this before you plan around the table above. Every run export we hold for this scraper came back empty - four runs across 2026-06-30 to 2026-07-15, each a workbook with this header row and no data row beneath it. The cause is not a bug: Crunchbase is behind Cloudflare, and a run without a residential proxy completes and returns nothing. So the six names are solid - two independent sources carry them, in this order - and nothing else here is. In particular we do not tell you the format of permalink, the vocabulary of entity_type, how many matches a query returns, or in what order they come back. The obvious next step is to feed crunchbase_url into the Crunchbase Scraper for the full profile - that is what the two are for - but confirm the shape of your own first file before you wire the two together. Point a free-tier run with your own proxy at one company and read what you get.

Run controls

Set on the job,
not in the spreadsheet.

This scraper has one control besides the query box - a cap on the rows per query, which here caps the matches you take per search. The rest is how you hand over the list, and the proxy the job runs through.

Company name input Domain input URL input Limit per query Paste one per line CSV upload XLSX upload TXT upload Parquet upload Residential proxy
Common workflows

Three jobs people
most often run here.

A few examples of what teams do when their list of companies is not yet a list of links.

Resolution

Turn a messy list into Crunchbase links

A spreadsheet of trading names or websites is not something you can look anything up with. One run gives each of them a Crunchbase identity and a profile link, which is the thing every later step needs.

RevOps
Deduplication

See where a name is ambiguous

Because a query returns a row per match rather than one best guess, the rows themselves show you which names are ambiguous. That is the difference between catching a mismatch now and finding it in a report later.

Data
Pipeline

Feed the profile scraper

Search first, profile second. Resolve your list here, then hand the Crunchbase links to the Crunchbase Scraper for the full sixteen-column profile on each one.

Research
Pricing

Pay only for the matches
you actually pull.

No subscription, no minimum, no recurring bill. Your first 500 rows are on us - after that, pay-as-you-go at the same flat rate as every other scraper here.

Free tier

500 free rows - $0

Every new account, one-time. No credit card required. Per-query limits, file upload and every export format included. The proxy this scraper needs is yours to supply, on the free tier and after it.

$0 forever
Pay-as-you-go

$0.002 per row, after the free tier

The same flat rate as every other scraper on the platform. Remember a row here is a match, not a query - an ambiguous name costs more than a precise one. The estimator shows the cost of a run before it starts.

Most popular
Volume

Custom · high volume

Volume pricing, dedicated workers and an SLA for continuous monitoring or very large historical pulls. Tell us your numbers and we will quote.

Talk to us
10% off your first paid run.Use code LIVESCRAPER10 at checkout.
Sign up
Pairs well with

Find the company,
then read the profile.

The legal bit

Is it legal to search
Crunchbase this way?

Short answer: you are reading a public search result - and the constraint here is technical and contractual rather than privacy.

Crunchbase search results are shown to anyone who runs the search, signed in or not. A company name, a link, a one-line description and a type are published to be read. Collecting publicly visible company information for research is long-established practice, and nothing here touches a login or a paywall.

This export is business data rather than personal data - a company and a pointer to its profile, not a named individual. That is a lighter obligation than a reviews or profiles export, though it does not make it none: a one-person company is still a person, and the entity type column exists precisely because not everything on Crunchbase is a company. We have not seen its values, so treat an unexpected one with care rather than assuming.

The constraint worth planning around is Crunchbase's own. The site is actively protected by Cloudflare, and its terms restrict automated access - so this is a terms question as well as a technical one, and if you have a subscription or an API agreement with them you should check it before scaling. We run no third-party trackers on the data layer, and your exports auto-delete after 30 days.

livescraper.app · principles
Public search results only
No logins, no accounts touched
Needs your own residential proxy!
Business data, not personal data
Exports auto-delete (30 days)
Check Crunchbase's own terms before scaling.
Common questions

Things people
ask before signing up.

The questions we hear most. Anything else? Talk to us - humans, not bots, write the answers.

How do I search Crunchbase for a list of companies?+
Using the Crunchbase Search Scraper:
  1. Sign in to the platform.
  2. Open the Crunchbase Search Scraper.
  3. Paste company names, domains or URLs, one per line - or upload a CSV, XLSX, TXT or Parquet file.
  4. Configure a residential proxy - without one the run returns no rows.
  5. Set a limit per query, or leave it at zero to take everything.
  6. Click Get Data.
How is this different from the Crunchbase Scraper?+
This one finds companies; the other one reads them. Give this a name, a domain or a URL and it returns the Crunchbase matches - six columns, one row per match. Give the Crunchbase Scraper an organization URL and it returns that company's full profile - sixteen columns, one row per company. Most people use this first and that second.
Do I need a proxy for this one?+
Yes, and it is the single most important thing on this page. Crunchbase is protected by Cloudflare, which does not serve data to a datacentre address. This scraper needs a residential proxy configured against the job. Without one the run completes normally and the export comes back with a header row and nothing under it.
Does one query always return one company?+
No - the output is one row per matching company, so an ambiguous name can produce several rows and a precise one fewer. Every row carries the query it came from, so you can group by it and decide which match you meant. That is also why a row here costs the same as any other row: an ambiguous search is a more expensive search.
Can I join the results straight into the Crunchbase Scraper?+
That is what the pair is for, and crunchbase_url is the field to use. We will not promise you the exact shape of permalink, because we have never seen a value in it - whether it is a bare slug or a full URL is something your first run will show you. Check one file before you wire an automated join between the two.
Why does this page show no example values?+
Because we will not print values we have not seen. Every run export we hold for this scraper returned zero rows - four runs between 2026-06-30 and 2026-07-15, each a workbook with the header row and no data beneath it, all made without a residential proxy. That makes the six column names trustworthy, because two independent sources list them in this order, and makes any claim about their contents a guess. Our other pages describe field formats because we measured them; this one stays silent because we could not.
My run came back empty. What went wrong?+
Almost certainly the proxy. A Crunchbase job without a residential exit address is served a Cloudflare challenge rather than results, and the job finishes cleanly with nothing in it - the same empty file our own test runs produced. Check the proxy first, before you start rewriting your queries.
How much does it cost?+
The first 500 rows are free and one-time, with no credit card. After that it is $0.002 per row - the same flat rate as every other scraper on the platform. Remember that a row here is a match rather than a query. The estimator shows the cost of a run before it starts, and the proxy is a separate cost that is yours.

Your first 500 matches,
on the house.

500 one-time free rows on every new account - no expiry. After that it is $0.002 per row, pay-as-you-go - no card on file until you say so. Bring your own residential proxy.

Activates instantly · no card required

Search Crunchbase by company name, domain or URL

Livescraper's Crunchbase Search Scraper turns the list you actually have into Crunchbase identities. You submit company names, domains or URLs - typed one per line, or uploaded as a CSV, XLSX, TXT or Parquet file - cap the matches per query if you want to, and download the results as a clean CSV, Excel or JSON file. All three input forms go in the same box, which is the point: a list of trading names or websites is not something you can look anything up with until it has been resolved.

Each row is one matching company: the query you submitted, the name Crunchbase holds, the link to the profile, its permalink, a short description and the entity type. The output is one row per match rather than one per query, so an ambiguous name produces several rows and you can see the ambiguity instead of having a single answer chosen for you. Every row carries its query, so grouping back to your input list is straightforward.

This is the lookup half of a pair. Once a search has resolved your list, the Crunchbase Scraper reads each organization page and returns the full sixteen-column profile - founding date, headcount, industries, location, rank, operating status and social links. Search first, profile second, is how most teams use the two.

Two things are worth knowing before you start. Crunchbase is protected by Cloudflare, and this scraper needs a residential proxy to return anything: without one, a job finishes cleanly and the export contains a header row and no data. And this page publishes no example values and no format for any field - not the shape of the permalink, not the vocabulary of the entity type, not how many matches a query returns - because every run export we hold returned zero rows, and a value we have not seen is not a value worth printing. The six column names are exact and cross-checked; everything else is for your first free-tier run to tell you. Start free: your first 500 rows cost nothing and need no credit card.