Web Scraping

Browse Web Scraping agent skills in Data Processing and compare related workflows, tools, and use cases.

41 skills
A
data-scraper-agent

by affaan-m

data-scraper-agent helps build a repeatable public-data pipeline for web scraping, enrichment, and storage. It is designed for monitoring jobs, prices, news, repos, sports, and listings on a schedule using GitHub Actions, with outputs to Notion, Sheets, or Supabase. Best for ongoing tracking, not one-off extractions.

Web Scraping
Favorites 0GitHub 156.1k
B
remote-browser

by browser-use

remote-browser helps sandboxed agents control a headless browser for Browser Automation. Use it to open pages, inspect state, click indexed elements, type input, take screenshots, and connect to local apps or CDP-backed browser sessions.

Browser Automation
Favorites 0GitHub 84.9k
B
browser-use

by browser-use

browser-use is a browser automation skill for opening pages, inspecting state, clicking indexed elements, typing into fields, taking screenshots, and reusing a persistent browser session. Use it for reliable form filling, navigation, and logged-in workflows with the browser-use CLI.

Browser Automation
Favorites 0GitHub 84.9k
C
zyte-api-automation

by ComposioHQ

zyte-api-automation helps agents run Zyte API workflows through Composio Rube MCP by discovering current tool schemas, checking the zyte_api connection, and executing Web Scraping tasks with less guesswork.

Web Scraping
Favorites 0GitHub 67.5k
C
zenserp-automation

by ComposioHQ

zenserp-automation helps agents run Zenserp tasks through Composio Rube MCP. Use it to connect Zenserp, discover current tool schemas first, and collect SERP data for SEO research.

Seo Research
Favorites 0GitHub 67.5k
C
webscraping-ai-automation

by ComposioHQ

webscraping-ai-automation is a Claude skill for Web Scraping through Composio's Rube MCP, with setup checks, tool discovery, and schema-first usage guidance.

Web Scraping
Favorites 0GitHub 67.5k
C
serpdog-automation

by ComposioHQ

serpdog-automation helps agents run Serpdog tasks through Composio Rube MCP by discovering current tool schemas, checking the serpdog connection, and executing SEO/SERP research workflows with less guesswork.

Seo Research
Favorites 0GitHub 67.5k
C
scrapingbee-automation

by ComposioHQ

scrapingbee-automation helps agents run ScrapingBee web scraping tasks through Composio Rube MCP by discovering current tool schemas first, checking the ScrapingBee connection, and then executing safer workflows.

Web Scraping
Favorites 0GitHub 67.5k
C
scrapingant-automation

by ComposioHQ

scrapingant-automation helps agents run Scrapingant tasks through Composio Rube MCP by discovering live tool schemas first, checking the scrapingant connection, then executing the right tool. Use it for MCP-based Web Scraping workflows, not standalone scraper scripts.

Web Scraping
Favorites 0GitHub 67.5k
C
scrapfly-automation

by ComposioHQ

scrapfly-automation helps agents run Scrapfly web scraping tasks through Composio Rube MCP by discovering current tools, checking the Scrapfly connection, and avoiding stale schemas.

Web Scraping
Favorites 0GitHub 67.5k
C
scrapegraph-ai-automation

by ComposioHQ

scrapegraph-ai-automation skill guide for using Scrapegraph AI through Composio Rube MCP: set up the MCP connection, discover current schemas with RUBE_SEARCH_TOOLS, and run web scraping workflows.

Web Scraping
Favorites 0GitHub 67.5k
C
PhantomBuster Automation

by ComposioHQ

PhantomBuster Automation helps Claude operate PhantomBuster through Composio MCP to list agents and scripts, monitor quotas, inspect runs, and support controlled web scraping or lead generation workflows.

Web Scraping
Favorites 0GitHub 67.5k
C
parsehub-automation

by ComposioHQ

parsehub-automation is a Claude skill guide for running ParseHub workflows through Composio Rube MCP. It covers setup context, RUBE_SEARCH_TOOLS discovery, connection checks, and safe usage patterns for Web Scraping tasks.

Web Scraping
Favorites 0GitHub 67.5k
C
hyperbrowser-automation

by ComposioHQ

hyperbrowser-automation helps agents run Hyperbrowser workflows through Composio Rube MCP by discovering current tool schemas first, checking the Hyperbrowser connection, and then executing browser automation tasks.

Browser Automation
Favorites 0GitHub 67.5k
C
Firecrawl Automation

by ComposioHQ

Firecrawl Automation helps Claude Code run Firecrawl through Composio to scrape pages, crawl sites, extract structured data, batch process URLs, and map site structures with scoped, credit-aware workflows.

Web Scraping
Favorites 0GitHub 67.5k
C
diffbot-automation

by ComposioHQ

diffbot-automation helps Claude use Diffbot through Composio Rube MCP by discovering current tool schemas, checking the Diffbot connection, and running structured web data tasks safely.

Web Scraping
Favorites 0GitHub 67.5k
C
browseai-automation

by ComposioHQ

browseai-automation helps Claude run Browse AI workflows through Composio Rube MCP, with mandatory tool discovery, connection checks, and current schemas before execution.

Browser Automation
Favorites 0GitHub 67.4k
C
brightdata-automation

by ComposioHQ

brightdata-automation helps agents run Bright Data workflows through Composio Rube MCP by discovering current tool schemas, checking the Bright Data connection, and executing tasks with less guesswork.

Web Scraping
Favorites 0GitHub 67.4k
C
Apify Automation

by ComposioHQ

Apify Automation is a Claude skill for running Apify Actors through Composio: connect MCP, run sync or async scraping jobs, fetch datasets, create tasks, and inspect logs.

Web Scraping
Favorites 0GitHub 67.4k
C
agentql-automation

by ComposioHQ

agentql-automation helps Claude run AgentQL browser automation through Composio Rube MCP with schema-first tool discovery, connection checks, and safer execution steps.

Browser Automation
Favorites 0GitHub 67.4k
A
browser-automation

by alirezarezvani

browser-automation helps agents build Playwright workflows for scraping, form filling, screenshots, downloads, session handling, and structured data extraction. Includes recipes, API references, and helper scripts for production browser automation, not E2E testing.

Browser Automation
Favorites 0GitHub 22.2k
J
baoyu-url-to-markdown

by JimLiu

baoyu-url-to-markdown converts live URLs to Markdown with a vendored baoyu-fetch CLI using Chrome CDP, site adapters, and generic fallback. Review Bun runtime needs, first-time EXTEND.md setup, and usage for X, YouTube, Hacker News, and rendered pages.

Format Conversion
Favorites 0GitHub 13.2k
H
huggingface-datasets

by huggingface

Use the huggingface-datasets skill for Hugging Face Dataset Viewer API workflows to validate datasets, resolve splits, preview and paginate rows, search text, apply filters, and fetch parquet links or statistics. It is a practical huggingface-datasets guide for read-only dataset exploration.

Web Scraping
Favorites 0GitHub 10.4k
T
burpsuite-project-parser

by trailofbits

burpsuite-project-parser searches and extracts data from Burp Suite project files (.burp) using Burp Suite Professional and the burpsuite-project-file-parser extension. Use it for security audit findings, proxy history, site map entries, and regex searches across captured HTTP traffic.

Security Audit
Favorites 0GitHub 5k