Field notes from the world of data extraction.
Articles, interviews and analysis on how data is gathered, used and fought over — written by the people closest to it.

I'm not the same developer I was before LLMs
I've been running a series of conversations with developers at Zyte to understand what's actually changed in the way they work since LLMs showed up. Not the headlines. The day-to-day. What they delegate, what they don't, what they notice, what surprises them. This one was different on two counts.

Flatcar Linux for web scrapers: deploy immutable containers with just one config file
The next time you spin up a VPS to give it a persistent home, you spend the better part of an afternoon rebuilding from memory. Here's a tool to help using Flatcar Linux

My agentic coding setup: Claude Code, multi-agent orchestration, and how I actually work
Ayan's 4 agent team, using Claude's /goal, and the models and coding agents he uses to code effectively.

llms.txt isn’t dead: How we put dev docs in AI’s spotlight
Marketers are giving up on the idea of plain-text pages - but llms.txt and Markdown are how we’ll get our docs in the hands of LLMs and developers.

The great wall of data: The complexities of web scraping in the Asian market
While the technological arms race of web data access is universal, the battleground in Asia has its own unique rules of engagement.

Actually, web scraping APIs are cheaper
Many data teams still think running a proxy-based scraping stack is most cost-effective. Industry pressures and our research disprove that idea.

The science of compliance: Tech tips for a legal data pipeline
New legal and regulatory compulsions for web data have significant business consequences. So, how can technologists engineer their company’s risk profile lower?

AI won’t fix your data quality (until you answer these three questions)
In our interview, a QA expert warns - before you delegate web scraping quality assurance to AI, make sure you can describe what ‘good’ looks like for yourself.

Why 10 million tokens won’t save your AI agent (and what will)
New models can process larger inputs, and confuse themselves in the process. Context management techniques can solve the problem.

Web scraping APIs vs proxies: A head-to-head comparison
Proxies are essential to scraping at scale. So, how do full-stack web scraping APIs compare?

OpenClaw and Claude helped me buy the perfect sneakers using Zyte API
Quickly compare e-commerce products across any site with an agent, a skill and an AI-powered web scraping API.

Legal clarity comes with compliance demands
Explore how new regulations like the EU AI Act and California AB 2013 are reshaping AI data compliance in 2026. Learn why provenance, transparency, and lawful sourcing are now critical.




_HFpro5d6k3.png&w=256&q=75)
_E4PyVpfAxa.png&w=256&q=75)


-(1).png&w=1920&q=75)
-(1)_VZGHqxCgXV.png&w=1920&q=75)