This file is 207 lines long; read all of them.
From a plain-English prompt to a working Scrapy spider.
Not using exclusively Claude Code? See Zyte Coding Agent Add-Ons for alternatives.
claude plugin marketplace add zytedata/claude-skills
claude plugin install zyte-web-data@zyte-aiIf Claude Code is already running, reload plugins in the active session:
/reload-pluginsIf /reload-plugins isn't available (e.g. in the VS Code extension), restart Claude Code.
See also: Discovering and installing plugins
This is Zyte's official Claude Code plugin that generates production-ready Scrapy spiders with web-poet page objects from a plain-English prompt. Give it a URL and describe what you want to extract. It handles site exploration, schema discovery, code generation, and smoke testing: no boilerplate, no manual selector hunting.
The plugin explores the target site, discovers available fields, and presents a schema for your approval before generating a single line of code. After you confirm the schema, it creates a Scrapy project with all dependencies configured, generates web-poet page objects and test fixtures, wires up the spider, and runs a smoke test to verify that extraction is working before handing the project back to you.
Optionally, use /zyte to deploy directly to Scrapy Cloud for scheduled runs, job history, and monitoring. A free tier is available.
The /scrape skill works on any website with repeating structured content: detail pages linked from a listing or category page. Examples from the skill:
- Product catalogs
- Job listings
- Recipes
The /scrape skill orchestrates two stages automatically:
1. Plan and validate the scrape → /scrape-plan
2. Build the project and spider → /scrapy-extra
Each stage feeds directly into the next. When the pipeline completes, you have a runnable spider and a passing test suite:
uv run scrapy crawl <spider_name>
uv run pytest fixtures/| Skill | Description |
|---|---|
scrape |
End-to-end web scraping workflow — from URL to working spider with web-poet page objects |
| Skill | Description |
|---|---|
scrape-plan |
Plan the scrape and author a validated extraction spec: discover fields, download diverse pages, compare HTML variants, optional browser review |
scrape-analyze-page |
Extract all available fields with values from a detail page |
scrapy-extra |
Hands-on Scrapy coding: write/debug spiders, web-poet page objects, and projects; configure scrapy-poet and scrapy-zyte-api |
| Skill | Description |
|---|---|
zyte |
Interact with Zyte's APIs and cloud services: set up your Zyte account and credentials; deploy projects, schedule spiders, list/stop jobs, and view items or logs on Scrapy Cloud; query historical Zyte API usage stats; look up Zyte API pricing and per-website costs; and answer how-to and documentation questions about Zyte from the official docs |
- Claude Code (CLI or desktop app)
uv— used to create and manage the Scrapy project
Project dependencies (scrapy, scrapy-poet, scrapy-zyte-api, web-poet, extruct, price-parser, pytest) are installed automatically by the skills.
Any scraping prompt triggers the skill automatically. For example:
Scrape books.toscrape.com
The plugin walks you through schema approval interactively, then generates a complete, tested Scrapy project.
We recommend enabling automatic updates:
- Enter
/pluginin a Claude Code session - Select Marketplaces → zyte-ai → Enable auto-update
To update manually:
claude plugin marketplace update zytedata/claude-skillsThen, in a Claude Code session:
/reload-pluginsIf /reload-plugins isn't available (e.g. in the VS Code extension), restart Claude Code.
We automatically evaluate skills and track both wall time and cost. We measure and aim to improve these metrics over time.
If you find any issue — such as prompts that did not work as expected, or that caused excessive wall time or cost — please open a GitHub issue.
Provide as much detail as possible to help us reproduce the issue. You are welcome to anonymize target websites or other data.
No. The generated spider is a standard Scrapy project that runs locally with uv. A Zyte account is required only if you want to deploy to Scrapy Cloud or use Zyte API to access sites that block standard scrapers. If you want to use Zyte API, you'll need an account to generate an API key.
The generated project includes scrapy-zyte-api as a dependency. Enabling headless browser rendering requires a Zyte API key. The /zyte skill guides you through setting up your credentials.
The project template includes scrapy, scrapy-poet, scrapy-zyte-api, web-poet, extruct, price-parser, and pytest. All dependencies are installed automatically via uv sync.
Yes. The plugin generates a standard Scrapy project. Run it directly with:
uv run scrapy crawl <spider_name>You can extend, modify, and deploy it independently of Claude Code.
See LICENSE.md for the Zyte End User License Agreement.
