PINGDOM_CHECK

#ExtractSummit2026 The world's largest web scraping conference returns. Austin Oct 7–8 · Dublin Nov 10–11.

Register now
Data Services
Pricing
Login
Try Zyte APIContact Sales
  • Unblocking and Extraction

    Zyte API

    The ultimate API for web scraping. Avoid website bans and access a headless browser or AI Parsing

    Ban Handling

    Headless Browser

    AI Extraction

    SERP

    Enterprise

    DocumentationSupport

    Hosting and Deployment

    Scrapy Cloud

    Run, monitor, and control your Scrapy spiders however you want to.

    Coding Agent Add-Ons

    Agentic Web Data

    Plugins that give coding agents the context to build production Scrapy projects. Starts with Claude Code.

  • Data Services
  • Pricing
  • Browse

    • BlogArticles, podcasts, videos
    • Case studiesCustomer outcomes
    • White papersIn-depth reports
    • DocumentationGuides & API reference
    • EventsConferences, webinars, recordings

    Subscribe

    • NewsletterSwiftly delivered
    • Discord communityExtract Data community
  • Product and E-commerce

    From e-commerce and online marketplaces

    Data for AI

    Collect and structure web data to feed AI

    Job Posting

    From job boards and recruitment websites

    Real Estate

    From Listings portals and specialist websites

    News and Article

    From online publishers and news websites

    Search

    Search engine results page data (SERP)

    Social Media

    From social media platforms online

  • Meet Zyte

    Our story, people and values

    Contact us

    Get in touch

    Support

    Knowledge base and raise support tickets

    Terms and Policies

    Accept our terms and policies

    Open Source

    Our open source projects and contributions

    Web Data Compliance

    Guidelines and resources for compliant web data collection

    Join the team building the future of web data
    We're Hiring
    Trust Center
    Security, compliance & certifications
Login
Try Zyte APIContact Sales
P

Author

Pablo Hoffman

Pablo has worked in open source for 13+ years. He founded Insophia in 2007 — the first Uruguayan company built exclusively on Python — where his team created and open-sourced Scrapy, now the standard Python web crawling framework. In 2010 he founded Scrapinghub (now Zyte) to build dedicated crawling products around Scrapy, growing it from 3 people to a staff of 70 while continuing to lead Scrapy's open-source development.

Pablo's writing documents much of Zyte's (then Scrapinghub's) early product and open-source history — Smart Proxy Manager's launch, ScrapyRT, Portia, and Frontera — alongside community-facing pieces on conferences (EuroPython, PyCon Philippines), engineering culture (the move to Slack, remote distributed teams), and JavaScript rendering in Scrapy via Splash. Together they form a first-hand record of the company's product evolution from roughly 2014 onward.

Backconnect proxies explained: How to use them in a scraping project?
Proxies

Introducing Zyte Smart Proxy Manager — Free Trials & New Plans

We are introducing a new Zyte Smart Proxy Manager plans that better suit your web scraping proxy needs. Try Zyte Smart Proxy Manager for completely free.

Pablo Hoffman·2 min read·February 18, 2020
Backconnect proxies explained: How to use them in a scraping project?
Proxies

A sneak peek inside Zyte Smart Proxy Manager

Take a look behind the scenes inside Smart Proxy Manager , the world's smartest web scraping proxy network.

Pablo Hoffman·6 min read·February 15, 2019
Backconnect proxies explained: How to use them in a scraping project?
Product Update

The Zyte Smart Proxy Manager Story: Enhancing Scraping Efficiency

Discover Smart Proxy Manager (Crawlera), the world's smartest proxy network tailored for web scraping, eliminating proxy management hassles.

Pablo Hoffman·6 min read·February 7, 2019
GDPR: Public and Personal Data Update
Open-source

Improving Access to Peruvian Congress Bills with Scrapy

Improving Access to Peruvian Congress Bills with Scrapy - Learn how Scrapy is improving access to Peruvian Congress bills for greater transparency.

Pablo Hoffman·4 min read·July 13, 2016
The Road to Loading JavaScript in Portia: A Technical Journey
Open-source

The Road to Loading JavaScript in Portia: A Technical Journey

The Road to Loading JavaScript in Portia - Learn about the journey of adding JavaScript support to Portia. Extract data from dynamic websites more efficiently.

Pablo Hoffman·4 min read·August 3, 2015
EuroPython 2015: Uniting Pythonistas in Europe
Developer interest

EuroPython 2015: Uniting Pythonistas in Europe

Europython 2015 - Calling all Python enthusiasts! Join us at Europython 2015 for insightful talks and networking opportunities.

Pablo Hoffman·3 min read·July 21, 2015
StartupChats: Embracing Remote Working for Success
Scraping strategy

StartupChats: Embracing Remote Working for Success

StartupChats: Remote Working - Tune in to StartupChats as they discuss the advantages and challenges of remote working.

Pablo Hoffman·1 min read·July 17, 2015
A Practical Guide To Web Data QA Part IV
Developer interest

PyCon Philippines 2015: Celebrating Python and Community

PyCon Philippines 2015 - Join us at PyCon Philippines 2015. Discover the latest trends and innovations in the Python community.

Pablo Hoffman·3 min read·July 15, 2015
Google Summer of Code 2015: Empowering Open Source Projects
Open-source

Google Summer of Code 2015: Empowering Open Source Projects

Google Summer of Code 2015 - Get involved in the Google Summer of Code with Zyte. Explore exciting projects and opportunities for students.

Pablo Hoffman·1 min read·June 25, 2015
Embracing The Future Of Work: How To Communicate Remotely
Developer interest

Github repository: Manage Vacations Distributed Team

Leverage a github repository to help manage employee personal vacations, filter each country and their public holidays. Efficiently manage vacations for distributed teams.

Pablo Hoffman·3 min read·June 8, 2015
Gender Inequality Across Programming Languages
Developer interest

Gender Inequality Across Programming Languages

Gender inequality across programming languages is a hot topic. The study is based on UK profiles to determine the gender of a profile covering 80% of the users.

Pablo Hoffman·2 min read·May 27, 2015
Want To Predict Fitbit’s Quarterly Revenue? Eagle Alpha Did It Using Web Scraped Product Data
Open-source

Frontera: The Brain Behind The Crawls

Frontera, formerly Crawl Frontier, is an open-source framework to manage our crawling logic and sharing it between spiders in our Scrapy projects.

Pablo Hoffman·5 min read·April 22, 2015
A Practical Guide to Web Data QA (Part V): Navigating Broad Crawls
Open-source

Scrape Data Visually With Portia And Scrapy Cloud

Note: Portia is no longer available for new users. It has been disabled for all the new organizations from August 20, 2018, onward.

Pablo Hoffman·4 min read·April 7, 2015
Chats with Rinar Solutions: Insights into Remote Working
Developer interest

Why We Moved To Slack

We are veterans in the chat group arena. We have been using one form of another since we started Zyte in 2010 and I've been personally using corporate

Pablo Hoffman·3 min read·March 16, 2015
Chats with Rinar Solutions: Insights into Remote Working
Developer interest

History of Zyte : A Journey of Innovation

History of Zyte - Learn about the journey of Zyte and how we evolved from Zyte to a leading web scraping and data extraction platform.

Pablo Hoffman·2 min read·March 16, 2015
A Practical Guide To Web Data QA Part IV
Developer interest

Handling JavaScript In Scrapy With Splash

Handling modern websites that entirely run on Javascript? In this article, learn how to use Splash to render JavaScript-based pages in your Scrapy spiders.

Pablo Hoffman·5 min read·March 2, 2015
A Practical Guide to Web Data QA (Part V): Navigating Broad Crawls
Product Update

New Changes to Our Scrapy Cloud Platform: Enhanced Performance and Features

New Changes to Our Scrapy Cloud Platform - Stay up-to-date with the latest changes to Scrapy Cloud. Enhance your web scraping workflow with new features.

Pablo Hoffman·3 min read·January 23, 2015
A Practical Guide to Web Data QA (Part V): Navigating Broad Crawls
Product Update

Introducing ScrapyRT: An API for Scrapy Spiders

Introducing ScrapyRT: An API for Scrapy Spiders - Make the most of your Scrapy spiders with ScrapyRT. Explore its functionalities as an API for seamless integration.

Pablo Hoffman·1 min read·January 22, 2015
Chats with Rinar Solutions: Insights into Remote Working
Developer interest

Looking Back at 2014: Highlights and Milestones

Looking Back at 2014 - Take a trip down memory lane and see the milestones and breakthroughs at Zyte in 2014.

Pablo Hoffman·3 min read·December 31, 2014
Open source at Zyte
Open-source

Open source at Zyte

Open Source at Zyte, Now Zyte - Embrace the open-source movement at Zyte. Learn how we contribute to the community and promote transparency.

Pablo Hoffman·2 min read·January 18, 2014
Marcos Campal Is A ScrapingHubber!
Developer interest

Marcos Campal Is A ScrapingHubber!

Marcos Campal is a Zyteber - Get to know one of our talented team members, Marcos Campal. Learn about his contributions to the world of web scraping.

Pablo Hoffman·1 min read·October 1, 2013
Proxy management: In-house or off-the-shelf proxy solutions?
Product Update

Introducing Smart Proxy Manager

Introducing Zyte Smart Proxy Manager - Enhance your web scraping with Smart Proxy Manager. Explore its powerful features and benefits as a smart proxy manager.

Pablo Hoffman·1 min read·May 11, 2013
4 simple Steps for effective Automated Data QA Process
How To

Git Workflow For Scrapy Projects

Git Workflow for Scrapy Projects - Streamline your Scrapy projects with an efficient Git workflow. Improve collaboration and project management.

Pablo Hoffman·2 min read·March 6, 2013
Zyte Blog — field notes from the world of data extraction
Product Update

How To Fill Login Forms Automatically

We often have to write spiders that need to fill login forms to sites. Our customers provide us with the site, username and password, and we do the rest.

Pablo Hoffman·3 min read·October 26, 2012
A Practical Guide to Web Data QA (Part V): Navigating Broad Crawls
Developer interest

Spiders Activity Graphs

Spiders Activity Graphs - Visualize your spiders' performance with activity graphs. Optimize your web scraping process with actionable insights.

Pablo Hoffman·2 min read·August 25, 2012
A Practical Guide To Web Data QA Part IV
Product Update

Scrapy 0.15: Dropping Support for Python 2.5

Scrapy 0.15 Dropping Support for Python 2.5 - Important update for Scrapy users! Discover the changes in the latest release and the end of Python 2.5 support.

Pablo Hoffman·1 min read·February 27, 2012

Services

Zyte Data

Coding tools & hacks straight to your inbox. Bi-weekly dosage of all things code.

Talk to us

Web Scraping API

Zyte API

Coding tools & hacks straight to your inbox. Bi-weekly dosage of all things code.

Sign Up

Developers

Zyte Developers

Coding tools & hacks straight to your inbox. Bi-weekly dosage of all things code.

Join Us
    • Zyte API
    • Ban Handling
    • AI Extraction
    • SERP
    • Enterprise
    • Scrapy Cloud
    • Agentic Web Data
    • Pricing
    • Product & E-commerce
    • Data for AI
    • Job Posting
    • Real Estate
    • News & Articles
    • Search
    • Social Media
    • Blog
    • Learn
    • Case Studies
    • Webinars
    • White Papers
    • Join our community
    • Documentation
    • Meet Zyte
    • Contact us
    • Jobs
    • Support
    • Terms and Policies
    • Trust Center
    • Do not sell
    • Cookie settings
    • Web Data Compliance
    • Open Source
    • What is Web Scraping
    • Web Scraping in Python: Ultimate Guide
    • Stop getting blocked, start scraping
  • EWDCI logoMost loved workplace certificateZyte rewardISO 27001 iconG2 rewardG2 rewardG2 reward
    XFacebookInstagramYouTubeLinkedInDiscord

    © Zyte Group Limited 2026