
Apify
import.io
Octoparse
Bright Data
ParseHub
Zyte
Scrapy
Data Miner
BrowserCloud.io
Firecrawl
puppeteer
Browserbase
Playwright
Selenium
Multilogin
GoLogin
Apify is a JavaScript & Node.js based data extraction tool for websites that crawls lists of URLs and automates workflows on the web. With Apify you can manage and automatically scale a pool of headless Chrome / Puppeteer instances, maintain queues of URLs to crawl, store crawling results locally or in the cloud, rotate proxies and much more.
Apify
BrowserCloud.ioNo BrowserCloud.io videos yet. You could help us improve this page by suggesting one.
Our company has been working with BrowserCloud.io for ~1 year, we use their API to run 150-200 parallel Puppeteer sessions for our web crawling solution. The skillful support team that can offer customizations for our specific needs
Based on our record, Apify should be more popular than BrowserCloud.io. It has been mentiond 44 times since March 2021. We are tracking product recommendations and mentions on various public social media platforms and blogs. They can help you identify which product is more popular and what people think of it.
Data collection: Apify actors, one per source, that scrape the open-data endpoints and normalize them. Quebec RBQ ships a daily bulk CSV (inside a 10.8 MB zip, ~924k rows that dedupe to ~54k active licences). Ontario HCRA has no bulk file โ it's an internal JSON API behind the public registry. - Source: dev.to / 16 days ago
Create a free Apify account and grab your API token from Settings โ API & Integrations. - Source: dev.to / 24 days ago
BYOK. It runs on your own Apify token. No shared keys, no lock-in, no licensing chokepoint โ a lesson the whole "Proxycurl shut down and stranded everyone" saga taught the space. - Source: dev.to / about 1 month ago
You need apify-client installed (pip install apify-client pandas scikit-learn). Get a free Apify API token at apify.com โ no card required, every account starts with $5 of credit. - Source: dev.to / about 2 months ago
A free Apify account (for the API token). - Source: dev.to / about 2 months ago
Try to run puppeteer or playwright for these purposes. It depends on how many pages you need to get If you need to scale your web-scraping - you can try https://browsercloud.io. Source: over 3 years ago
Hello How much will you pay AWS for running 10-20 browsers in parallel for a long time? I'm just asking because we're building the https://browsercloud.io service as an alternative. It should be much cheaper than running chromium on AWS/lambda. Really interesting to compare. Source: almost 4 years ago
Https://browsercloud.io ;) More info is needed, what site? LinkedIn? How many requests to the site? It might be you need proxy / user-agent rotation and so on. Source: almost 4 years ago
Hi Do you have a permanent workload for scraping? We're developing our service for these cases like yours :) It's like "puppeteer cloud", optimized AMD servers that run chromium browsers, you can run 10-20-50 parallel sessions. We have a "usage-based" plan billed per second and it might be cheaper than running your own machine (50-100$/mo?) if you have high workloads from time to time. Take a look:... Source: over 4 years ago
They have rest JSON API for sending and a common SMTP gateway, we use it https://browsercloud.io , works fine. Source: over 4 years ago
import.io - Import. io helps its users find the internet data they need, organize and store it, and transform it into a format that provides them with the context they need.
Firecrawl - Turn any website into LLM-ready data.
Octoparse - Octoparse provides easy web scraping for anyone. Our advanced web crawler, allows users to turn web pages into structured spreadsheets within clicks.
puppeteer - Puppeteer is a Node library which provides a high-level API to control headless Chrome or Chromium...
Bright Data - World's largest proxy service with a residential proxy network of 72M IPs worldwide and proxy management interface for zero coding.
Browserbase - A web browser for your AI