Field notes from the world of data extraction.
Articles, interviews and analysis on how data is gathered, used and fought over — written by the people closest to it.

Kill your product - why sacrificing your cash cow can be the path to growth
Letting go of the past is the best way to embrace the future. Retiring a flagship product isn’t a sign of failure; it’s a commitment to innovation.

From script to system: 10 building blocks to scale web scraping
Scaling your business’ web data gathering – acquiring, monitoring and storing a growing amount of data from a growing number of sources over time – requires holistic planning.

Scrape Web Pages and Files Using Python, wget, and Zyte
The command-line utility wget (pronounced "web-get") can download online files. This free network downloader may run in the background without user intervention.

New in Zyte: Scroll Control, Lower Costs, and More
As the web continues to evolve, Zyte API is evolving right alongside it—adding powerful new features and refinements designed to make data extraction smarter, faster, and more adaptable than ever.

The future of Scrapy: Smarter, faster and ready for AI-powered scraping
What does the future hold for the tool some describe as “the gift that revolutionised web scraping”?

Rise of the Data Vendor: How Outsourcing is Transforming Supply and Fuelling Businesses
With the emergence of managed data extraction vendors, businesses no longer need to gather web data themselves.

Quality, focus and scale: Three ways data outsourcing benefits businesses
The Strategic Case for Buying Web Data: Quality, Focus, and Scale

What AI Builders Need to Know About the Training Data Copyright Debate
The generative AI gold rush is upon us, with astounding new products and capabilities that are fuelled by web data and raising questions about AI copyright.

Ten years since Scrapy 1.0: The stats and stories behind your favorite framework
See what 10 years of Scrapy 1.0 has built — in milestones and metrics.

Web Scraping Best Practices
In this article learn everything about web scraping best practices, from the pros. At Zyte, we extract web data everyday for our customers, Read are our tips!

Using curl with a Proxy for Web Scraping
When it comes to command-line tools for HTTP requests, few are as versatile and powerful as curl. Loved by developers and system administrators alike, curl makes fetching web resources straightforward.

What’s your data type? Solving the procurement problem
Engagements with data suppliers break down when buyers don’t have a clear project concept. Understanding and articulating your needs is paramount. Meet the three types of data buyers. Which one are you?




_HFpro5d6k3.png&w=256&q=75)
_E4PyVpfAxa.png&w=256&q=75)


-(1).png&w=1920&q=75)
-(1)_VZGHqxCgXV.png&w=1920&q=75)