Check out https://crawlee.dev/ to get started.
Crawlee helps you build reliable scrapers. Fast. It is an intuitive, customizable open-source Node.js library for web scraping and browser automation.
Quickly scrape data, store it, and avoid getting blocked with headless browsers, smart proxy rotation, and auto-generated human-like headers and fingerprints.
#crawlee #webscraping #coding
🤿 Dive into Crawlee
Support the project by leaving a ⭐ on GitHub - https://github.com/apify/crawlee
Check it out on NPM - https://www.npmjs.com/
Join Crawlee's Discord community: / discord
📖Contents of this video:
0:00 Introduction
5:37 Agenda
6:40 What is Crawlee?
8:21 Crawl the way you want!
10:21 Ready to use templates
11:16 Proxies and Sessions
12:35 Crawling Context
13:33 Using Router
15:38 Context helpers: enqueueLinks
17:28 Context helpers: parseWithCheerio
18:11 Saving data
19:26 Running in the cloud
20:00 Apify platform
20:55 Q&A
41:12 Get in touch!