PINGDOM_CHECK

Learn web scraping

Guides, tutorials and courses on web scraping and data extraction — from your first request to production pipelines.

Zyte Blog — field notes from the world of data extraction
Use case

Best Web Scraping Companies (Software + Services)

Looking for the best web scraping company? Compare top providers by success rates, data quality, compliance standards, and enterprise support.

Arnold Alexander10 min read
Zyte Blog — field notes from the world of data extraction
Use case

The best web scraping tools in 2026

Compare the best headless browsers for web scraping in 2026. Learn when to use Playwright, Puppeteer, Selenium, or Zyte API’s managed CDP browser for scalable, anti-ban scraping.

Arnold Alexander10 min read
Zyte Blog — field notes from the world of data extraction
How To

Web scraping for pricing intelligence: how to track competitor prices at scale

Compare the best headless browsers for web scraping in 2026. Learn when to use Playwright, Puppeteer, Selenium, or Zyte API’s managed CDP browser for scalable, anti-ban scraping.

Mitch Holt10 min read
Zyte Blog — field notes from the world of data extraction
How To

Best headless browsers for web scraping in 2026

Compare the best headless browsers for web scraping in 2026. Learn when to use Playwright, Puppeteer, Selenium, or Zyte API’s managed CDP browser for scalable, anti-ban scraping.

Arnold Alexander10 min read
Zyte Blog — field notes from the world of data extraction
How To

Best proxy providers for web scraping in 2026 | Zyte

Compare the best proxy providers for web scraping in 2026. Learn which residential, ISP, and mobile proxies work best—and when teams move beyond proxies to automation.

Arnold Alexander10 min read
Zyte Blog — field notes from the world of data extraction
How To

Hybrid Scraping: The Architecture for the Modern Web

John Rooney4 min read
Zyte Blog — field notes from the world of data extraction

The Modern Scrapy Developer's Guide (Part 3): Auto-Generating Page Objects with the Web Scraping Copilot

In this guide, we'll show you how to use Web Scraping Copilot (our VS Code extension) to automatically write 100% of your Items, Page Objects, and even your unit tests.

5 min read
Zyte Blog — field notes from the world of data extraction
Scraping practice

The Modern Scrapy Developer's Guide (Part 2): Page Objects with scrapy-poet

In this guide, we'll fix this by refactoring our spider to a professional, modern standard using Scrapy Items and Page Objects (via crapy-poet). We will completely separate our crawling logic from our parsing logic.

John Rooney5 min read
Zyte Blog — field notes from the world of data extraction
Scraping practice

The Modern Scrapy Developer's Guide (Part 1): Building Your First Spider

In this definitive guide, we will walk you through, step-by-step, how to build a real, multi-page crawling spider. You will go from an empty folder to a clean JSON file of structured data in about 15 minutes

John Rooney4 min read
Zyte Blog — field notes from the world of data extraction
Scraping practice

How to Scrape Search Engine Results

From SEO audits to market intelligence, scraping search engine results data can give you the insights you need to make smarter, faster business decisions.

Karlo Jeđud5 min read
Zyte Blog — field notes from the world of data extraction
Scraping practice

Scrape Web Pages and Files Using Python, wget, and Zyte

The command-line utility wget (pronounced "web-get") can download online files. This free network downloader may run in the background without user intervention.

Karlo Jeđud7 min read
Zyte Blog — field notes from the world of data extraction
Traffic

Using curl with a Proxy for Web Scraping

When it comes to command-line tools for HTTP requests, few are as versatile and powerful as curl. Loved by developers and system administrators alike, curl makes fetching web resources straightforward.

Karlo Jeđud8 min read