Open-source

Articles from the Zyte blog about Open-source.

skills-are-software
Open-source

Treat your AI skills like software, starting with evals

Most AI skills are never tested — and it shows. Here's how Zyte evaluates scraping skills like real software, catching failures demos miss.

Neha Setia Nagpal14 min read
AI won’t fix your data quality (until you answer these three questions)
Scraping practice

AI won’t fix your data quality (until you answer these three questions)

In our interview, a QA expert warns - before you delegate web scraping quality assurance to AI, make sure you can describe what ‘good’ looks like for yourself.

Neha Setia Nagpal10 min read
Zyte Blog — field notes from the world of data extraction
Developer interest

Stop using Python requests for web scraping: Use these modern modules instead

While the 'Requests' library remains the default choice for many Python developers due to its reliability and extensive documentation, the Python HTTP landscape has evolved considerably. Modern alternatives now offer significant advantages, including built-in asynchronous support, HTTP/2 compatibility, enhanced performance, and up-to-date TLS handling.

Ayan Pahwa6 min read
The future of Scrapy: Smarter, faster and ready for AI-powered scraping
Open-source

The future of Scrapy: Smarter, faster and ready for AI-powered scraping

What does the future hold for the tool some describe as “the gift that revolutionised web scraping”?

Robert Andrews6 min read
Ten years since Scrapy 1.0: The stats and stories behind your favorite framework
Open-source

Ten years since Scrapy 1.0: The stats and stories behind your favorite framework

See what 10 years of Scrapy 1.0 has built — in milestones and metrics.

Cleber Alexandre5 min read
The rise of Scrapy: How an open-source scraping framework conquered the web
Developer interest

The rise of Scrapy: How an open-source scraping framework conquered the web

The story of Scrapy reflects the broader evolution of the web itself and the ongoing quest to harness its ever-expanding ocean of information.

Theresia Tanzil10 min read
Sustainability in Open Source | Fireside Chat
How To

Sustainability in Open Source | Fireside Chat

Learn how successful open-source projects balance community value with sustainable growth. Industry leaders share insights on monetization, maintenance, and building thriving communities.

Shane Evans2 min read
Data in 2026: Bets and forecasts from web experts
Open-source

A Deep Dive into Zyte's Open-Source Libraries

Discover how Zyte’s open-source libraries like ClearHTML, Extruct, Chomp.js, and more simplify web data extraction and processing.

Neha Setia Nagpal1 min read
4 essential Scrapy plugins for building efficient and effective spiders
Open-source

4 essential Scrapy plugins for building efficient and effective spiders

Here are four essential Scrapy plugins we use to build efficient web crawlers for our customers.

Neha Setia Nagpal1 min read
Zyte Blog — field notes from the world of data extraction
Open-source

The Scraper’s System Part 2: Explorer’s Compass to analyze websites

Learn about the scrapers system: Explorer’s Compass to analyze websites.

Neha Setia Nagpal8 min read
4 simple Steps for effective Automated Data QA Process
Open-source

How Scrapy makes web crawling easy and accurate

Get the best value for your web crawling project by using Scrapy. An awesome framework you should learn and incorporate for easy and accurate web crawling.

Attila Toth5 min read
Dateparser: A Little But Powerful Date Parsing Library
Open-source

Dateparser: A Little But Powerful Date Parsing Library

Meet Dateparser, a potent date parsing library simplifying date extraction from HTML pages. Ideal for various applications like command-line tools, chatbots, and more.

Marc Hernandez Cabot3 min read

More articles on Open-source