#ExtractSummit2026 The world's largest web scraping conference returns. Austin Oct 7–8 · Dublin Nov 10–11

Register now
Data Services
Login
Try Zyte APIContact Sales
  • Unblocking and Extraction

    Zyte API

    The ultimate API for web scraping. Avoid website bans and access a headless browser or AI Parsing

    Ban Handling

    Headless Browser

    AI Extraction

    SERP

    Enterprise

    DocumentationSupport

    Hosting and Deployment

    Scrapy Cloud

    Run, monitor, and control your Scrapy spiders however you want to.

    Coding Agent Add-Ons

    Agentic Web Data

    Plugins that give coding agents the context to build production Scrapy projects. Starts with Claude Code.

  • Data Services
  • Zyte API

    Zyte Data

    Scrapy Cloud

  • Browse

    • BlogArticles, podcasts, videos
    • Case studiesCustomer outcomes
    • White papersIn-depth reports
    • DocumentationGuides & API reference
    • EventsConferences, webinars, recordings

    Subscribe

    • NewsletterSwiftly delivered
    • Join our community2,000+ web scraping engineers
  • Product and E-commerce

    From e-commerce and online marketplaces

    Data for AI

    Collect and structure web data to feed AI

    Job Posting

    From job boards and recruitment websites

    Real Estate

    From Listings portals and specialist websites

    News and Article

    From online publishers and news websites

    Search

    Search engine results page data (SERP)

    Social Media

    From social media platforms online

  • Meet Zyte

    Our story, people and values

    Contact us

    Get in touch

    Support

    Knowledge base and raise support tickets

    Terms and Policies

    Accept our terms and policies

    Open Source

    Our open source projects and contributions

    Web Data Compliance

    Guidelines and resources for compliant web data collection

    Affiliate Program

    Join Zyte’s affiliate program and start earning commissions today

    Join the team building the future of web data
    We're Hiring
    Trust Center
    Security, compliance & certifications
Login
Try Zyte APIContact Sales

Open source

We build web scraping tools in the open

Zyte started as the team behind Scrapy, and open source is still how we work. Zyte engineers maintain Scrapy and the libraries underneath it, fixes land upstream before they reach our own products, and what we publish stays free to use, commercially or not.

Browse librariesJoin the community

Paid maintainers

Core Scrapy maintainers are Zyte engineers, on the clock.

Upstream first

What we build for customers lands in the public repos.

Free, always

BSD-licensed, no paywalled features, no rug pulls.

Flagship projects

The tools we maintain

The projects we lead, use in production every day, and support publicly.

The leading web scraping framework

Scrapy

The open source framework for web crawling at scale, created by Zyte's co-founders and maintained by our engineers. An async engine, batteries-included middlewares, and a plugin ecosystem the whole industry builds on.

  • Async crawling engine, built for scale
  • Middlewares, pipelines and extensions
  • Created and maintained by Zyte engineers
64,296pablohoffmanpablohoffmanwRARwRARdangradangrakmikekmikeGallaecioGallaecio+369$ pip install scrapy
  • Parsing architecture

    web-poet & scrapy-poet

    Page objects for web scraping: keep extraction logic separate from the crawler so it can be reused, unit tested, and swapped per site without touching your spiders.

    • Extraction logic decoupled from crawlers
    • Unit-testable parsers
    • Swap providers per site
    106kmikekmikewRARwRARBurnzZBurnzZGallaecioGallaeciovictor-torresvictor-torres+8
  • Monitoring & QA

    Spidermon

    Battle-tested monitoring for Scrapy. Validate the items you extract, assert on job-level stats, and get alerted the moment a spider quietly starts returning junk.

    • Schema validation on extracted items
    • Job stat monitors and thresholds
    • Slack, email and custom alerts
    561gatufogatuforennerocharennerochajesuslosadajesuslosadafurther-readingfurther-readingVMRuizVMRuiz+54
  • Zyte API integration

    scrapy-zyte-api

    The official Scrapy plugin for Zyte API. Unblocking, headless browser requests and AI extraction from inside your existing spiders, with sessions and retries handled for you.

    • Unblocking and sessions handled
    • Headless browser requests
    • AI extraction from your spiders
    43GallaecioGallaeciowRARwRARBurnzZBurnzZkmikekmikesortafreelsortafreel+12
  • Browser automation

    scrapy-playwright

    Playwright, wired into Scrapy's download layer. Render JavaScript-heavy pages, click and scroll through them, and keep the rest of your project unchanged.

    • Render JavaScript-heavy pages
    • Scripted clicks, scrolls and waits
    • Drop-in download handler
    1,445elacuestaelacuestawRARwRARGallaecioGallaeciohmtbrhmtbraxiom-of-choiceaxiom-of-choice+9
  • Selectors · Python

    Parsel

    The selector library at the heart of Scrapy, usable on its own. Query HTML and XML with XPath or CSS, chain the two, and pull out text or attributes with a couple of lines.

    • XPath and CSS selectors, chainable
    • Works standalone, no Scrapy needed
    • The engine behind every Scrapy spider
    1,354GallaecioGallaeciowRARwRAReliasdorneleseliasdorneleskmikekmikepablohoffmanpablohoffman+49
  • Open standard · AI agents

    Agent Skills

    Portable, openly packaged skills that teach any coding agent how Zyte engineers build scrapers: modern Scrapy patterns, page objects and pytest fixtures instead of hallucinated selectors.

    • Works with any Agent Skills-compatible agent
    • Scrapy + web-poet scaffolding, not guesswork
    • Free to install and use
    9github-actions[bot]github-actions[bot]

Also ours

Smaller libraries, doing the unglamorous work

Pieces of Scrapy we release on their own, so you can use them anywhere.

  • Selectors

    cssselect

    CSS selectors for Python, translated to XPath

    310SimonSapinSimonSapinwRARwRARredappleredapple+26
  • Web utilities

    w3lib

    Web-related helpers for scraping projects

    418wRARwRARkmikekmikealbertedwardsonalbertedwardson+47
  • Date parsing

    dateparser

    Parse human-readable dates in 18 languages

    2,855waqasshabbirwaqasshabbirAllactagaAllactagagavishpoddargavishpoddar+142
  • Price parsing

    price-parser

    Pull amount and currency out of raw text

    346kmikekmikeWinterComesWinterComeslopuhinlopuhin+14
  • Number parsing

    number-parser

    Parse numbers written in natural language

    129arnavkapoorarnavkapoorserhii73serhii73noviluninoviluni+3
  • Item loading

    itemloaders

    Populate items with XPath and CSS

    49GallaecioGallaeciowRARwRARejulioejulio+48
  • Data containers

    itemadapter

    One interface for any data container class

    69elacuestaelacuestawRARwRARGallaecioGallaecio+5
  • Real-time API

    scrapyrt

    Run spider logic behind a single HTTP request

    884pawelmhmpawelmhmchekunkovchekunkovradostyleradostyle+14
  • Join the community

    Join 2,000+ engineers who share what's actually working.

    Join the community

Services

Zyte Data

Fully managed web data extraction, delivered to your spec.

Explore Zyte Data

Web Scraping API

Zyte API

Scrape any website at scale with automatic proxy rotation and ban handling.

Sign Up

Developers

Zyte Developers

Docs, tools, and a community to help you build and scale scrapers.

Join Us
    • Zyte API
    • Ban Handling
    • AI Extraction
    • SERP
    • Enterprise
    • Scrapy Cloud
    • Agentic Web Data
    • Pricing
    • Product & E-commerce
    • Data for AI
    • Job Posting
    • Real Estate
    • News & Articles
    • Search
    • Social Media
    • Blog
    • Learn
    • Case Studies
    • Webinars
    • White Papers
    • Join our community
    • Join our Affiliate Program
    • Documentation
    • Meet Zyte
    • Contact us
    • Jobs
    • Support
    • Terms and Policies
    • Trust Center
    • Do not sell
    • Cookie settings
    • Web Data Compliance
    • Open Source
    • What is Web Scraping
    • Web Scraping in Python: Ultimate Guide
    • Stop getting blocked, start scraping
  • Logo EWDCILogo Most Loved WorkplaceLogo Job TogetherISO 27001 SealMedal Leader Europe Winter 2025Fastest Implementation Winter 2025Logo Leader Winter 2025Grid Leader Spring 2025Grid Leader Summer 2025Leader Fall 2025Leader Winter 2026
    XFacebookInstagramYouTubeLinkedInDiscord

    © Zyte Group Limited 2026