PINGDOM_CHECK

Field notes from the world of data extraction.

Articles, interviews and analysis on how data is gathered, used and fought over — written by the people closest to it.

Browser bother: Three painkillers for headless scraping headaches
Scraping practice

Browser bother: Three painkillers for headless scraping headaches

This article shares three strategies to operationalize large-scale browser automation yourself,and what alternatives exist.

Theresia Tanzil10 min read
The Right AI For the Right Problem: How Zyte Solved Web Data's Trilemma of Cost, Quality, and Flexibility
AI-assisted data extraction

The Right AI For the Right Problem: How Zyte Solved Web Data's Trilemma of Cost, Quality, and Flexibility

This article reveals the approach that makes once-unviable web data projects possible.

Iain Lennon15 min read
Zyte Blog — field notes from the world of data extraction
Web data application

Why Might a Business Use Web Scraping to Collect Data?

Businesses use data to better understand consumer behavior, track market trends, study competitive landscapes, and enhance internal processes. Choosing the correct web scraping vendor is critical for companies who want to capitalize on web scraping's potential.

Karlo Jeđud10 min read
Responsibly Using Big Data to Train LLMs: A Practical Demonstration
Developer interest

Responsibly Using Big Data to Train LLMs: A Practical Demonstration

Join Joachim Asare, AI/ML Engineer & Master’s in Design Engineering @Harvard University, as he explores responsible methods for extracting and leveraging big data to train LLMs. This session covers key ethical considerations, including privacy, transparency, and fairness throughout the AI development lifecycle.

Joachim Asare1 min read
Scraping Smarter, Not Harder: Unlocking Public Data with Scrapy
Developer interest

Scraping Smarter, Not Harder: Unlocking Public Data with Scrapy

Explore how to overcome the challenges of collecting publicly available data from websites protected by advanced security systems like Cloudflare Turnstile.

Tamas Deak1 min read
Unblockers vs Zyte API: What’s the Real Cost of Bans?
Web data collection

Why AI is changing the game for data buyers in 2025

Discover how AI, data marketplaces, and economies of scale are making web data more accessible than ever.

Cleber Alexandre10 min read
Buy or Build? The Four Roads to Acquiring Web Data
Web data collection

Buy or Build? The Four Roads to Acquiring Web Data

Weighing your options from full control to full service.

Theresia Tanzil10 min read
Zyte Blog — field notes from the world of data extraction
Scraping practice

Golang Web Scraping in 2025: Tools, Techniques and Best Practices

Go (Golang)—a language built for speed, efficiency, and concurrency. Whether you’re scraping large datasets, handling high-throughput requests, or managing complex site interactions, Golang will deliver.

Karlo Jeđud10 min read
Zyte Blog — field notes from the world of data extraction
Scraping practice

Advanced Use Cases for Session Management in Web Scraping

In this article, we’ll explore the sophisticated techniques that help manage modern bot defenses, why they matter, and how Zyte API gives you an edge in maintaining seamless, efficient, and cost-effective data extraction.

Karlo Jeđud10 min read
Web Scraping in 2025: Four Essentials For Developers and Data Buyers to Stay Ahead
Scraping strategy

Web Scraping in 2025: Four Essentials For Developers and Data Buyers to Stay Ahead

Gain exclusive insights from The 2025 Web Scraping Industry Report and understand how AI is shaping web scraping and data practices in 2025.

Cleber Alexandre2 min read
Play Before You Scrape: Explore Zyte API Settings with Playground
Product Update

Play Before You Scrape: Explore Zyte API Settings with Playground

Discover the best way to configure your scrapers using Zyte API Playground

Cleber Alexandre10 min read
Zyte Blog — field notes from the world of data extraction
Anti-ban

From Basic to Advanced Ways of Managing Bans in Web Scraping

Discover a comprehensive guide on managing bans in web scraping, from basic strategies to advanced techniques, ensuring efficient and ethical data extraction.

Karlo Jeđud10 min read