Crunchbase Scraper for Company Profiles and Funding Rounds

Crunchbase shows part of every company profile to anonymous visitors and masks the rest. Our Crunchbase scraper collects the open part: industries, status, rounds, investors.

Plans from €169/month · Free project assessment · Reply within 1 business day

Crunchbase Scraper
Solutions

Crunchbase Scraping as a Managed Service

You name the scope and we deliver the table. A scope can be a list of company names or domains to match against Crunchbase permalinks, a set of hubs - an industry, a region, a founding year - or the funding rounds of companies you already follow. We build the crawler, run it once or on your schedule, and repair it when the profile layout changes.

Crunchbase puts a session check in front of its profiles, so pacing, proxy rotation, rendering and CAPTCHA handling are part of the service. We never log in and never go behind the paywall: a value the site masks for anonymous visitors arrives as an empty, flagged cell, and the field documentation says which columns that can affect.

What Our Crunchbase Data Scraper Collects

A company lives at /organization/ plus a permalink - a slug of its name, with a short suffix where two companies share one. The permalink is the key of every row. From the open part of the profile we return:

  • Identity - organization name, permalink, profile URL, Legal Name, Also Known As, website and the company description.
  • Classification - Industries, Hub Tags such as Unicorn, Company Type such as For Profit, and Operating Status, which reads Active or Closed.
  • Location and size - headquarters city, region and country, the Headquarters Regions line (San Francisco Bay Area, West Coast, Western US) and the employee range.
  • Funding summary - Last Funding Type, Number of Funding Rounds, Number of Investors and Number of Lead Investors.
  • Funding rounds - one row per round with Transaction Name, Funding Type, Funding Stage, the lead investors and the address of the round's own page.
  • Deals - Number of Investments for companies that invest themselves, acquisitions made, and Acquired by where the company was bought.
  • Public market - IPO Status and Stock Symbol for listed companies.
  • Scores - CB Rank, Growth Score (built on company activity, operational metrics and investments) and Heat Score (market interest and media activity), the two scores with their change in points over the past quarter.
  • Company contacts - the Phone Number and Contact Email a company lists for itself, not those of its people.

Not every cell is readable. For visitors who are not signed in Crunchbase masks part of the financial detail - typically the Money Raised of a round, the Total Funding Amount and a round's Announced Date, and on some profiles the Founded Date as well. We deliver what the page shows: a masked value becomes an empty cell with a flag, so an analyst can tell a figure that was never disclosed from one that was not visible. Founders, team members and investors who are private individuals are people, not company fields, and are left out by default.

What Our Crunchbase Data Scraper Collects
Hubs, Funding Round Pages and a Crunchbase Dataset With History

Hubs, Funding Round Pages and a Crunchbase Dataset With History

Company profiles are one of three public page types worth collecting.

  • Hubs. A hub is a ready-made list at /hub/ plus a slug: FinTech startups, United States startups, startups founded in a given year. Its overview carries Number of Organizations, Industries, Industry Groups, Location, CB Rank (Hub), Number of Funding Rounds and Total Funding Amount, and the organizations it shows are where the crawl of a sector begins. The biggest hubs carry Top 10K in their name; narrower ones sit beside them - the same industry in one country, one city or one funding stage - and their lists merge on the permalink.
  • Funding round pages. Every round has its own address under /funding_round/, built from the company permalink, the round type and a short hash. It names the organization, the Funding Type, the Funding Stage - Seed, Early Stage Venture and onwards - and the lead investors, which is enough to build an investor-to-company table.
  • Profile tabs. Financial Details, Tech Details and Growth Outlook sit on their own addresses under the profile, as does the investments tab of a company that invests. The profile also carries a line of alternatives and possible competitors that links one company to its neighbours.

A Crunchbase dataset earns its keep on the second run. Repeated passes record when Operating Status flips to Closed, when Last Funding Type moves from Seed to Series A, when a new round joins the list and how the scores drift from quarter to quarter - a history the site shows only as today's state.

About Crunchbase and Its Public Layer

Crunchbase is a company database: organizations, the investors behind them and the deals that connect the two. It began as the startup database of the technology news site TechCrunch, later became an independent company based in San Francisco, and now presents itself as a platform for business intelligence on private and public companies, used by sales teams, investors and analysts.

Three things about its structure matter to a data buyer. First, everything is an entity with its own address - organizations, funding rounds, acquisitions, hubs - and an investor is an organization too, so a venture firm has the same kind of profile as the startups it backs. Second, the vocabulary is closed. Funding Type runs from Pre-Seed and Seed through the lettered series to Venture - Series Unknown, Private Equity, Debt Financing and Secondary Market, and Funding Stage groups those into a handful of stages. Closed lists mean the columns filter without cleaning.

Third, the site has two layers. The public layer shows the skeleton of each profile and most of its descriptive fields; the rest is kept for signed-in accounts. This page, like our service, deals with the first layer only.

Get a Quote
dev_w
25

Developers

customers
500+

Customers worldwide

pages
1 500 000 000+

Pages extracted

stime
15000+

Hours saved for our clients

Plans

Crunchbase Scraping Plans and Pricing

Airplane

€199 / one-time

setup fee - included

Data limits100,000
Frequencyone-time
Run timeup to 5 days
Data storing7 days

Helicopter

€169 / mo

setup fee €499

Data limits250,000
Frequencymonthly
Run timeup to 5 days
Data storing14 days

Glasses

€229 / mo

setup fee €499

Data limits1,000,000
Frequencyweekly
Run timeup to 5 days
Data storing30 days

DNA

€549 / mo

setup fee €799

Data limits3,000,000
Frequency3 times daily
Run timesame day
Data storing90 days

Delivery: CSV, JSON or XLSX files, our API or a JSON feed. MCP server on request.

Full plan details →

Why Scrape Crunchbase Company Profiles

Crunchbase is where a company's category, stage and backers are written down in one place and in one vocabulary. Teams collect it for work that needs many companies at once:

  • Market maps. A hub plus Industries, headquarters and Operating Status gives the list of active companies in a sector and a region - the denominator behind a market-size slide.
  • Firmographics for a CRM. Accounts matched by website pick up Legal Name, Company Type, employee range, Last Funding Type and IPO Status, in the same wording for every record.
  • Investor mapping. Lead investors per round show who backs whom, which firms co-invest and which funds are active at Seed as opposed to later stages.
  • Stage triggers. A change of Last Funding Type or a new round on the list is the moment a vendor, a recruiter or a bank wants to know about, and a weekly pass catches it.
  • Momentum and hygiene. Growth Score and Heat Score rank a watchlist by movement; Operating Status set to Closed removes dead companies from a startup database before anyone calls them.

Our Blog

Reads Our Latest News & Blog

Learn how to use web scraping to solve data problems for your organization

Web Scraping for Sentiment Analysis: Sources, Fields and Data Traps

Web Scraping for Sentiment Analysis: Sources, Fields and Data Traps

Where opinion text lives on the public web, what a review record contains on each kind of site, which fields a model needs besides the text, and what quietly breaks a sentiment dataset - with examples from our own projects.

Empower Your Business with Google Maps Data

Empower Your Business with Google Maps Data

Using web scraping to extract Google Maps data will help you quickly and efficiently find businesses in any industry, city, state, or region. And the extracted contact information, such as phone or email address, social networks, or links to web pages, will help to contact them.

How to Generate Business Leads Using Web Scraping

How to Generate Business Leads Using Web Scraping

Lead scraping is a reliable way to get the right customer contacts to market your product, saves organizations time, and helps you better understand the audience you want to attract. Contacts include emails, phone numbers, or social media profiles.

scrapeit logo

About ScrapeIt

ScrapeIt is a web scraping agency that works as a managed service: you describe the company data you need, and we build, run and repair the crawlers that deliver it. Company directories like Crunchbase are their own kind of job - entity pages that link to each other, closed vocabularies, and a public layer that ends where the sign-in begins. We work inside that public layer, document every field, and handle personal data under GDPR. We assess a project for free, and a sample of real profiles can be ordered ahead of the full run. Plans start from EUR 169 a month, or EUR 199 for a one-time extract.

FAQ

Do you offer a Crunchbase API for company and funding data?

Yes. The ScrapeIt Crunchbase API returns the permalink, Legal Name, Industries, Operating Status, Last Funding Type and the list of funding rounds as JSON from an endpoint we host, refreshed on the schedule you set. The same data also comes as CSV or XLSX files, a JSON feed or straight into your database.

Is it legal to scrape Crunchbase?

Yes. Company names, descriptions, industries, headquarters, operating status and funding round types are shown on Crunchbase to anyone who opens a profile. We collect only publicly available data - everything a visitor can see on Crunchbase - and we collect it legally. Profiles of founders and other people are personal data and stay out of the dataset by default.

Does Crunchbase data scraping reach fields hidden behind a sign-in?

No. We never log in and never go behind the paywall, so values that Crunchbase masks for anonymous visitors - money raised, total funding, some dates - are delivered as empty cells with a flag. What is shown openly on company profiles, funding round pages and hubs is what the dataset contains.

Can you export Crunchbase data to Excel or CSV?

Yes. A Crunchbase data export arrives as CSV, Excel (XLSX), JSON or JSONLines, with one table for companies and one for funding rounds, joined on the permalink. Files go to e-mail, SFTP, Amazon S3, Google Drive or your own database, as a one-off or on a schedule.

How often should Crunchbase company data be refreshed?

It depends on the field. Growth Score and Heat Score are published with their change over the past quarter, so a monthly pass is enough for them, while new funding rounds and a switch of Operating Status to Closed deserve a weekly check on the companies you follow. Every run is stamped with its date, and that stamp is what turns snapshots into history.

How does it Work?

Step 1 - Make a Request

You share your needs, expectations, and desired timeframe. We’ll suggest the best solution based on your request and budget.

Step 2 - Configuring Custom Web Crawlers

Our specialists configure the crawlers and extract a sample dataset for your review before proceeding with the full-scale extraction.

Step 3 - Collect and Deliver

Once you approve the sample, we launch the project and start full data collection. We gather, filter, and structure the data for easy use, delivering it on time in your preferred format.

Step 4 - Maintain and Support

Our team manages ongoing processes, monitors website changes, and supports all data extraction cycles. We can also help integrate data into your systems or create dashboards to simplify analysis.

Request a Quote

Tell us more about you and your project information.
Which sites, which fields, how often. A couple of lines is enough.

We reply within 1 business day. No obligation.

Crunchbase data from €169/month. Free project assessment, reply within 1 business day.

scrapiet

Scrapeit Sp. z o.o.
10/208 Legionowa str., 15-099, Bialystok, Poland
NIP: 5423457175
REGON: 523384582