Open-source
Articles from the Zyte blog about Open-source.

Treat your AI skills like software, starting with evals
Most AI skills are never tested — and it shows. Here's how Zyte evaluates scraping skills like real software, catching failures demos miss.

AI won’t fix your data quality (until you answer these three questions)
In our interview, a QA expert warns - before you delegate web scraping quality assurance to AI, make sure you can describe what ‘good’ looks like for yourself.

Stop using Python requests for web scraping: Use these modern modules instead
While the 'Requests' library remains the default choice for many Python developers due to its reliability and extensive documentation, the Python HTTP landscape has evolved considerably. Modern alternatives now offer significant advantages, including built-in asynchronous support, HTTP/2 compatibility, enhanced performance, and up-to-date TLS handling.

The future of Scrapy: Smarter, faster and ready for AI-powered scraping
What does the future hold for the tool some describe as “the gift that revolutionised web scraping”?

Ten years since Scrapy 1.0: The stats and stories behind your favorite framework
See what 10 years of Scrapy 1.0 has built — in milestones and metrics.

The rise of Scrapy: How an open-source scraping framework conquered the web
The story of Scrapy reflects the broader evolution of the web itself and the ongoing quest to harness its ever-expanding ocean of information.

Sustainability in Open Source | Fireside Chat
Learn how successful open-source projects balance community value with sustainable growth. Industry leaders share insights on monetization, maintenance, and building thriving communities.

A Deep Dive into Zyte's Open-Source Libraries
Discover how Zyte’s open-source libraries like ClearHTML, Extruct, Chomp.js, and more simplify web data extraction and processing.

4 essential Scrapy plugins for building efficient and effective spiders
Here are four essential Scrapy plugins we use to build efficient web crawlers for our customers.

The Scraper’s System Part 2: Explorer’s Compass to analyze websites
Learn about the scrapers system: Explorer’s Compass to analyze websites.

Python Web Scraper Tools & Libraries
Everything you need to know about python web scraper tools and libraries including Requests, BeautifulSoup, Selenium and Scrapy.

How Scrapy makes web crawling easy and accurate
Get the best value for your web crawling project by using Scrapy. An awesome framework you should learn and incorporate for easy and accurate web crawling.
More articles on Open-source
- Scrapy in 2026: New release brings modern async crawling standards
- The new economics of web data: Smaller scraping just got cheaper
- Rise of the Data Vendor: How Outsourcing is Transforming Supply and Fuelling Businesses
- Quality, focus and scale: Three ways data outsourcing benefits businesses
- What AI Builders Need to Know About the Training Data Copyright Debate
- Selenium, Puppeteer, Playwright: Which tool is right for web scraping at scale?
- Dateparser: A Little But Powerful Date Parsing Library
- Scrapy Update: Better Broad Crawl Performance
- Building Spiders Made Easy | GUI For Scrapy Shell
- ScrapyRT: Turn Websites into Real-Time APIs
- Spidermon: Zyte's secret to data quality
- Meet Spidermon: Our battle tested spider monitoring library
- Scraping The Steam Game Store With Scrapy
- How to crawl the web with Scrapy
- This Month In Open Source At Zyte August 2016
- Meet Parsel: The Selector Library Behind Scrapy
- Scrapy Tips from the Pros (July 2016): Tips for Effective Scraping
- Improving Access to Peruvian Congress Bills with Scrapy
- Scrapely: The Brains Behind Portia Spider
- This Month in Open Source at Zyte (June 2016): Community Highlights
- Scraping Websites Based On ViewStates With Scrapy
- Scrapy Tips from the Pros (March 2016 Edition): Mastering the Craft
- This Month In Open Source At Zyte March 2016
- Scrapy Tips from the Pros (February 2016 Edition): Continuous Learning
- Portia: The Open-source Alternative To Kimono Labs
- Chats with Rinar Solutions: Insights into Remote Working
- Parse Natural Language Dates With Dateparser
- Aduana: Link Analysis to Crawl the Web at Scale
- Scrapy on the Road to Python 3 Support: Modernizing the Framework
- The Road to Loading JavaScript in Portia: A Technical Journey
- Aduana: Link Analysis With Frontera | Zyte
- Frontera: The Brain Behind The Crawls
- Scrape Data Visually With Portia And Scrapy Cloud
- Skinfer: Inferring JSON Schemas Made Easy
- Handling JavaScript In Scrapy With Splash
- Portia: The Open-Source Visual Web Scraper
- Open source at Zyte
- Autoscraping Casts A Wider Net









