Scraping practice
Articles from the Zyte blog about Scraping practice.

Podcast Ep08 - Scrapy, Python and mushroom soup
Scrapy's core handles crawling well and deliberately leaves almost everything else out: no bundled browser, no opinion about how you shape your data, no built-in answer for every anti-bot wrinkle. What it gives you instead is a clean way to add those things at the edges, exactly when you need them and never before.

The missing middle ground in scrapy-playwright just got filled
You can now choose which Python browser library you want to use with scrapy-playwright. I go through why this is a huge deal for the right scraping demographic.

AI won’t fix your data quality (until you answer these three questions)
In our interview, a QA expert warns - before you delegate web scraping quality assurance to AI, make sure you can describe what ‘good’ looks like for yourself.

The recipe for a request: Scaling data extraction through investigation
Learn how an investigative mindset helps scale data extraction from single requests to millions daily by building resilient, efficient scraping systems.

Teaching AI to scrape like a pro: how we measure LLMs’ data quality
AI-enabled code editors can now conjure scraping code on command. But is it any good? Here’s how Zyte re-engineered LLMs with Web Scraping Copilot to drive best-in-class output.

Hybrid scraping: The architecture for the modern web
Learn how hybrid scraping combines headless browsers and lightweight HTTP clients to bypass JavaScript challenges efficiently. Reduce RAM usage, improve speed, and scale your web scraping pipelines with session reuse and TLS fingerprinting.

How I trade gold using e-ink, live data and an old Raspberry Pi
Track real-world gold and silver retail prices automatically using Zyte API, Python, and a Raspberry Pi with an e-ink display. Learn how to scrape rendered HTML, parse prices, and build an always-on trading dashboard.

Scrapy in 2026: New release brings modern async crawling standards
Scrapy 2.14.0 modernizes the framework with native async/await, smarter scheduling, and cleaner spider configuration. Here’s what the new release means for production crawlers in 2026.

How to build a daily industry news digest
Learn how data analyst Anshika Khandelwal automated a daily AI funding news digest using n8n and Zyte API. Discover how to pull articles, classify funding stories, and deliver a curated newsletter that saves 10+ hours per week.

Browser bother: Three painkillers for headless scraping headaches
This article shares three strategies to operationalize large-scale browser automation yourself,and what alternatives exist.

Overcoming web scraping challenges of Puppeteer and Playwright
Explore issues like browser farm management, IP rotation, and anti-scraping measures that can complicate large-scale operations.
![JSON Parsing with Python [Practical Guide]](/_next/image/?url=https%3A%2F%2Fextract.zyte.com%2Fapi%2Fmedia%2Ffile%2Fjson-parsing-with-python-hero-1-1200x630-qxGk7q0hXsCKlTSSTGo73YeO5qEHaa.png&w=3840&q=75)
JSON Parsing with Python [Practical Guide]
Learn how to parse JSON data with Python. This guide covers libraries, methods, and advanced tools like JMESPath and ChompJS.




_HFpro5d6k3.png&w=256&q=75)
_E4PyVpfAxa.png&w=256&q=75)


-(1).png&w=1920&q=75)
-(1)_VZGHqxCgXV.png&w=1920&q=75)