Agent skill

web-scraper

使用CSS选择器从网页提取数据,支持分页、限速和多种输出格式。

Stars 163
Forks 31

Install this agent skill to your Project

npx add-skill https://github.com/majiayu000/claude-skill-registry/tree/main/skills/other/other/web-scraper

Metadata

Additional technical details for this skill

short description
从网站爬取数据

SKILL.md

Web Scraper Tool

Description

Extract structured data from web pages using CSS selectors with rate limiting and pagination support.

Trigger

  • /scrape command
  • User requests web data extraction
  • User needs to parse HTML

Usage

bash
# Scrape single page
python scripts/web_scraper.py --url "https://example.com" --selector ".item" --output data.json

# Scrape with multiple selectors
python scripts/web_scraper.py --url "https://example.com" --selectors "title:.title,price:.price,link:a@href"

# Scrape multiple pages
python scripts/web_scraper.py --urls urls.txt --selector ".product" --output products.json --delay 2

Tags

scraping, web, html, data-extraction, automation

Compatibility

  • Codex: ✅
  • Claude Code: ✅

Expand your agent's capabilities with these related and highly-rated skills.

Didn't find tool you were looking for?

Be as detailed as possible for better results