Introducing open-source web scraping library Crawlee • Build reliable scrapers. Fast.

Опубликовано: 02 Ноябрь 2024
на канале: Apify
2,477
43

Check out https://crawlee.dev/ to get started.

Crawlee helps you build reliable scrapers. Fast. It is an intuitive, customizable open-source Node.js library for web scraping and browser automation.

Quickly scrape data, store it, and avoid getting blocked with headless browsers, smart proxy rotation, and auto-generated human-like headers and fingerprints.

#crawlee #webscraping #coding

🤿 Dive into Crawlee

Support the project by leaving a ⭐ on GitHub - https://github.com/apify/crawlee
Check it out on NPM - https://www.npmjs.com/
Join Crawlee's Discord community:   / discord  

📖Contents of this video:

0:00 Introduction
5:37 Agenda
6:40 What is Crawlee?
8:21 Crawl the way you want!
10:21 Ready to use templates
11:16 Proxies and Sessions
12:35 Crawling Context
13:33 Using Router
15:38 Context helpers: enqueueLinks
17:28 Context helpers: parseWithCheerio
18:11 Saving data
19:26 Running in the cloud
20:00 Apify platform
20:55 Q&A
41:12 Get in touch!