Agent skill · Data & Analytics

ai-web-scraping-scrapegraph

AI-powered web scraping - extract data using natural language prompts

gooseworks-aigithub.com/gooseworks-aiGitHub ↗
claude-codecodexcursorMIT
Install
npx skills add gooseworks-ai/goose-skills --skill ai-web-scraping-scrapegraph --agent claude-code

Same command for any agent — swap --agent for codex, cursor, copilot.

Facts
Files in the skill folder: 2
SKILL.md size: 12 KB
Bundled scripts: none
Path: skills/research-tools/capabilities/ai-web-scraping-scrapegraph/SKILL.md
Open the folder on GitHub →
Where it comes from
Stars: 1,091
Language: Python

Weekly change comes from our own snapshots, not the repository page — it measures attention, not adoption.

From the SKILL.md

# ScrapeGraph AI - Intelligent Web Scraping ## Setup Read your credentials from ~/.gooseworks/credentials.json: ```bash export GOOSEWORKS_API_KEY=$(python3 -c "import json;print(json.load(open('$HOME/.gooseworks/credentials.json'))['api_key'])") export GOOSEWORKS_API_BASE=$(python3 -c "import json;print(json.load(open('$HOME/.gooseworks/credentials.json')).get('api_base','https://api.gooseworks.ai'))") ``` If ~/.gooseworks/credentials.json does not exist, tell the user to run: `npx gooseworks login` All endpoints use Bearer auth: `-H "Authorization: Bearer $GOOSEWORKS_API_KEY"` Extract web content using AI with natural language prompts. ## Capabilities - **Start SmartScraper**: Extract content from a webpage using AI by providing a natural language prompt and a URL - **Start SearchScraper**: Start a new AI-powered web search request - **Scrape**: Extract raw HTML content from web pages with JavaScript rendering support - **Start SmartCrawler**: Start a new web crawl request with AI extraction or markdown conversion - **Start Sitemap**: Extract all URLs from a website sitemap automatically - **Start Markdownify**: Convert any webpage into clean, readable Markdown format - **Get Sear

What's inside
Steps it walks through
  1. Setup
  2. Capabilities
  3. Usage
  4. Start SmartScraper
  5. Start SearchScraper
  6. Scrape
  7. Start SmartCrawler
  8. Start Sitemap
  9. Start Markdownify
  10. Get SearchScraper Status (free)
  11. Get Markdownify Status (free)
  12. Get Sitemap Status (free)
  13. Get SmartCrawler Status (free)
  14. Get SmartScraper Status (free)
Ships with 1 file
  • skill.meta.json
Commands it runs
export GOOSEWORKS_API_KEY=$(python3 -c "import json;print(json.load(open('$HOME/.gooseworks/credentials.json'))['api_key'])")
export GOOSEWORKS_API_BASE=$(python3 -c "import json;print(json.load(open('$HOME/.gooseworks/credentials.json')).get('api_base','https://api.gooseworks.ai'))")
curl -s -X POST $GOOSEWORKS_API_BASE/v1/proxy/orthogonal/run \
curl -s -X POST $GOOSEWORKS_API_BASE/v1/proxy/orthogonal/search \
curl -s -X POST $GOOSEWORKS_API_BASE/v1/proxy/orthogonal/details \
More from goose-skills
All skills →
About this skill
What does the ai-web-scraping-scrapegraph skill do?

AI-powered web scraping - extract data using natural language prompts

How do I install it?

Run `npx skills add gooseworks-ai/goose-skills --skill ai-web-scraping-scrapegraph --agent claude-code` — it drops the skill into your project so the agent can pick it up. Swap the --agent value for codex, cursor or copilot if you use one of those.

Where does this skill come from?

From gooseworks-ai/goose-skills, a repository with 1,091 stars. We read it straight from the repository tree rather than a submitted listing, so what you see here is what is actually published.

Is a popular skill a good skill?

Not necessarily. Stars measure attention, not adoption — a repository can trend for a week and be abandoned. That is why we show the weekly change from our own snapshots next to the total, instead of a single flattering number.

Keep going