SimilarWeb Scraper - rank and engagement estimates for any domain
Competitive research runs into the same wall every time: you can measure your own website precisely and you can measure nobody else's at all. Third-party rank and engagement estimates exist to fill that gap, and the SimilarWeb Scraper collects them in bulk. Submit up to 1,000 domains - a bare domain or a full URL both work - and each returns an overview block carrying global, category and country rank along with the movement on each, the two-letter code of the country the country rank refers to, bounce rate, pages per visit, average visit duration, and four fields describing the company behind the site. Fifteen documented fields per domain, one row each, exportable as CSV, JSON or Excel.
The most important thing on this page is not a feature. Figures like these are modelled rather than measured: nobody outside a company can observe its traffic, so estimates are inferred from panels and sampling and published as approximations. That methodology is legitimate and widely relied on, and it has a specific consequence for how you should use the output. These numbers are strong for comparison and weak as absolutes. Ranking three competitors against one another, or watching whether a rival's position is climbing across successive runs, is exactly what the data supports. Quoting a bounce rate as measured fact in a board pack or a valuation model is where the same data becomes a liability. For your own properties your own analytics remain authoritative; the value here is coverage of domains you have no access to.
Four details will save whoever writes the importer an afternoon. Bounce rate arrives as a decimal fraction rather than a percentage - the documented example is 0.524, meaning roughly 52% - and displaying it unscaled produces a rate of half a percent, which is the most common error with this field. Average visit duration is a pre-formatted string such as 00:02:27 rather than a number of seconds, so it needs parsing before it can be sorted or averaged. The two company size fields both end in _min and represent the lower bound of a band rather than a figure, so they should be read as "at least this many" and never as a headcount. And the documented block contains no total-visits field at all, despite monthly visits being the metric most people expect, which is worth confirming against real output on the free tier before any report depends on it.
Two of the fifteen fields overlap conceptually with Company Insights, which is the dedicated firmographics service: twenty flat columns with an exact founding year, a normalised industry token, full address, phone and social accounts. The company fields here are context for the traffic data rather than a segmentation basis, so the split is straightforward - Company Insights answers who a company is, this answers how its website performs relative to others. On the legal side, this service differs from most of the catalogue in one respect worth flagging. The pages read are publicly viewable, but the figures are a vendor's compiled commercial product rather than incidental public information, which makes automated access a terms question and republication a separate risk from collection, since compiled databases attract protection in some jurisdictions. Internal research and competitive comparison is the ordinary use; publishing the figures as your own dataset warrants your own legal view. No personal data is involved, nothing behind a login is touched, and exports auto-delete after 30 days. Start free: your first 500 rows cost nothing and need no credit card.