Field notes from the world of data extraction.

Articles, interviews and analysis on how data is gathered, used and fought over — written by the people closest to it.

How web data turns e-commerce listings into retail intelligence
Web data collection

How web data turns e-commerce listings into retail intelligence

Discover how web data enables digital shelf analytics vendors to track prices, availability, and product trends at scale—fueling real-time retail intelligence and competitive advantage.

Theresia Tanzil5 min read
The seven habits of highly effective data teams
Scraping strategy

The seven habits of highly effective data teams

Discover the seven habits that set high-performing data teams apart—from treating data as a product to ensuring data trust, quality, and decision impact. Learn how leading teams scale reliable data systems.

Robert Andrews5 min read
Zyte Blog — field notes from the world of data extraction
Data quality

How to ensure data quality in your Scrapy web scraping projects using Spidermon and Claude Code

Spidermon is an open-source monitoring framework for Scrapy. You attach it to your spider, define what "success" looks like, and it automatically checks your crawl results after the spider closes, flagging anything that doesn't meet your standards.

Ayan Pahwa5 min read
Zyte Blog — field notes from the world of data extraction
Web data collection

Why your API responses look like gibberish: the gzip decompression trap

The script was working. Requests were going out, responses were coming back with HTTP 200. But the response body was unreadable noise, a wall of binary characters that crashed the JSON parser and reported "no data found". No error code, no timeout, no network failure; just garbage where structured data should be.

Ayan Pahwa6 min read
Dawn of the autonomous data pipeline
Use case

Dawn of the autonomous data pipeline

Discover how autonomous, agent-driven data pipelines are transforming web scraping in 2026, enabling self-healing systems, API discovery, and end-to-end automation.

Theresia Tanzil5 min read
Are programming practices relevant anymore?
Web scraping APIs

Are programming practices relevant anymore?

From LLM-powered extraction to agentic pipelines, here's how AI is reshaping every stage of the web scraping workflow in 2026 -- and what it means for your stack.

Mikhail Korobov5 min read
AI is the new engine for web scraping
Web scraping APIs

AI is the new engine for web scraping

From LLM-powered extraction to agentic pipelines, here's how AI is reshaping every stage of the web scraping workflow in 2026 -- and what it means for your stack.

Theresia Tanzil, Robert Andrews5 min read
Super-powers, toll booths and the new era of data collection
Web data collection

Super-powers, toll booths and the new era of data collection

Explore how AI is transforming web scraping, the rise of advanced anti-bot systems, and what the future holds for data collection in an increasingly controlled internet.

Fabien Vauchelles5 min read
Is your AI coding assistant stuck in the past?
Developer interest

Is your AI coding assistant stuck in the past?

AI coding tools like Copilot can suggest outdated code and practices. Learn why this happens and how developers can ensure they use the latest standards.

Theresia Tanzil5 min read
Data outcomes are top of the scraping stack
Web scraping APIs

Data outcomes are top of the scraping stack

By 2026, scraping shifts from proxy management to API-first data outcomes. Learn why unified scraping APIs are replacing traditional stacks.

Theresia Tanzil, Robert Andrews7 min read
No-code web scraping workflows are here: Introducing the Zyte integration for Zapier
Integration

No-code web scraping workflows are here: Introducing the Zyte integration for Zapier

Discover Zyte’s Zapier integration to automate web data extraction and connect it to 10,000+ apps. Build powerful workflows without coding and turn web data into actionable insights instantly.

Robert Andrews10 min read
The Scrapy whisperer: Adrian Chaves on Web Scraping Copilot
AI-assisted data extraction

The Scrapy whisperer: Adrian Chaves on Web Scraping Copilot

An interview with Scrapy maintainer Adrian Chaves on Zyte’s Web Scraping Copilot, AI-generated parsing code, and building reliable scraping workflows.

Neha Setia Nagpal10 min read