Scraping

Web scraping and data extraction

Showing 289-312 of 708 skills
JosiahSiegel

ffmpeg-audio-processing

by JosiahSiegel

Complete audio encoding and normalization system. PROACTIVELY activate for: (1) Audio codec selection (AAC, MP3, Opus, FLAC), (2) Loudness normalization (EBU R128, loudnorm), (3) Audio extraction from video, (4) Format conversion, (5) Volume adjustment and dynamics, (6) Noise reduction and EQ, (7) Channel operations (stereo/mono/surround), (8) Sample rate and bit depth conversion, (9) Audio fade in/out and crossfades, (10) Podcast and broadcast processing chains. Provides: Codec comparison tables, loudness standards reference, two-pass normalization scripts, professional mastering chains. Ensures: Broadcast-compliant audio with proper loudness and quality.

CLI Tools 51 7mo ago
openclaw

Instructions

by openclaw

All versions of all skills that are on clawhub.com archived

Processing 4.5K 7mo ago
parcadei

firecrawl-scrape

by parcadei

Scrape web pages and extract content via Firecrawl MCP

Processing 3.9K 7mo ago
joaquimscosta

playwright-cli

by joaquimscosta

Browser automation via Playwright CLI for navigating pages, interacting with elements, capturing screenshots, and testing web applications through shell commands. Use when user mentions "playwright", "browser automation", "take a screenshot", "browser testing", "headless browser", "web testing", "fill out a form", "e2e test", or needs to automate browser workflows from the command line. This is a pre-installed CLI tool — do NOT install anything via npx or npm. Invoke this skill first, then use playwright-cli bash commands.

CLI Tools 21 6mo ago
Takazudo

dev-figma-capture

by Takazudo

Capture web pages and send them to Figma as editable design files. Use when: (1) User wants to capture a webpage to Figma, (2) User says 'figma capture', 'send to figma', 'capture to figma', (3) User provides URLs to convert to Figma designs

Automation 12 6mo ago
idanbeck

playwright-skill

by idanbeck

Browser automation for web tasks. Use when the user needs to automate browser interactions, take screenshots, extract data from websites, fill forms, or perform web scraping. Supports persistent sessions for logged-in state.

Auth 12 7mo ago
Takazudo

headless-browser

by Takazudo

Browser automation skill with two efficiency tiers. Tier 1: lightweight headless-check.js for quick checks, screenshots, error detection. Tier 2: playwright-cli for interactions (click, fill, navigate). Use when: (1) Quick webpage health checks, (2) Taking screenshots, (3) Checking console/network errors, (4) Simple interactions like clicking buttons or filling forms, (5) Multi-step browser automation. Use MCP Playwright only for complex scenarios requiring persistent context or rich introspection.

Auth 12 6mo ago
LdotJdot

webbrowser

by LdotJdot

"浏览网页。与 read skill 风格一致:直接 exec 调用 exe,传参执行,stdout 为结果。"

Processing 47 6mo ago
JochenYang

remotion-best-practices

by JochenYang

Best practices for Remotion - Video creation in React

Animation 20 6mo ago
ZhanlinCui

pdf

by ZhanlinCui

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI Tools 182 7mo ago
besoeasy

crawl-websites-at-scale

by besoeasy

"Scrape websites at scale using Scrapy, a Python web crawling and scraping framework. Use when: (1) Crawling multiple pages or entire sites, (2) Extracting structured data from HTML/XML, or (3) Building automated data pipelines from web sources."

Processing 132 6mo ago
besoeasy

phone-specs-scraper

by besoeasy

"Scrape phone specifications from GSM Arena, PhoneDB, and alternative sites. Use when: (1) Comparing smartphone specs, (2) Researching device features, or (3) Building phone comparison tools."

Design 132 28mo ago
besoeasy

using-web-scraping

by besoeasy

Search and scrape public web content with headless Chrome and DuckDuckGo using safe practices.

Scraping 132 28mo ago
lijigang

ljg-xray-book

by lijigang

Deep structure extraction from books using the Epiplexity principle - maximizing computational investment to extract maximum learnable structure from any book.

CLI Tools 458 7mo ago
Casper-Studios

firecrawl-scraping

by Casper-Studios

Web page and website scraping with Firecrawl API. Use this skill when scraping web articles, blog posts, documentation pages, paywalled content, or JavaScript-heavy sites. Triggers on requests to scrape websites, extract article content, convert pages to markdown, or handle anti-bot protection.

Docs Gen 12 7mo ago
Casper-Studios

apify-scrapers

by Casper-Studios

Social media and web scraping using Apify actors. Use this skill when scraping Twitter/X tweets, Reddit posts, LinkedIn posts, Instagram profiles/posts/reels, Facebook pages/posts/groups, TikTok videos, YouTube content, Google Maps businesses/reviews, contact enrichment (emails/phones from websites), or when auto-detecting URL type to scrape. Triggers on requests to scrape social media, get trending posts, extract business info, find contact details, or extract content from social URLs.

CLI Tools 12 7mo ago
dkyazzentwatwa

audio-trimmer

by dkyazzentwatwa

Cut, trim, and edit audio segments with fade effects, speed control, concatenation, and basic audio manipulations.

Linting 95 8mo ago
dkyazzentwatwa

image-metadata-tool

by dkyazzentwatwa

Extract EXIF metadata from images including GPS coordinates, camera settings, and timestamps. Map photo locations and strip metadata for privacy.

Processing 95 8mo ago
dkyazzentwatwa

color-palette-extractor

by dkyazzentwatwa

Extract dominant colors from images, generate color palettes, and export as CSS, JSON, or ASE with K-means clustering.

Processing 95 8mo ago
yangliu2060

video-creator

by yangliu2060

AI短视频创作与多平台发布,使用即梦MCP生成视频,使用Playwright MCP自动发布到YouTube/TikTok/Instagram/Facebook/LinkedIn/Twitter等平台。

Prompts 40 8mo ago
nonameplum

swift-async-stream-patterns

by nonameplum

Patterns and best practices for building robust AsyncStream and AsyncSequence types, learned from swift-async-algorithms.

Code Gen 17 7mo ago
alexei-led

testing-e2e

by alexei-led

E2E testing with Playwright MCP for browser automation, test generation, and UI testing. Use when discussing E2E tests, Playwright, browser testing, UI automation, visual testing, or accessibility testing. Supports TypeScript tests and Go/HTMX web applications.

Accessibility 35 7mo ago
jackspace

cloudflare-browser-rendering

by jackspace

Complete knowledge domain for Cloudflare Browser Rendering - Headless Chrome automation with Puppeteer and Playwright on Cloudflare Workers for screenshots, PDFs, web scraping, and browser automation workflows. Use when: taking screenshots, generating PDFs from HTML or URLs, web scraping content, crawling websites, browser automation tasks, testing web applications, managing browser sessions, performing batch browser operations, integrating with AI for content extraction, or encountering browser rendering errors, XPath selector errors, browser timeout issues, concurrency limits, memory exceeded errors, or "Cannot read properties of undefined (reading 'fetch')" errors. Keywords: browser rendering cloudflare, @cloudflare/puppeteer, @cloudflare/playwright, puppeteer workers, playwright workers, screenshot cloudflare, pdf generation workers, web scraping cloudflare, headless chrome workers, browser automation, puppeteer.launch, playwright.chromium.launch, browser binding, session management, puppeteer.sessions, puppeteer.connect, browser.close, browser.disconnect, XPath not supported, browser timeout, concurrency limit, keep_alive, page.screenshot, page.pdf, page.goto, page.evaluate, incognito context, session reuse, batch scraping, crawling websites

Cloud 16 9mo ago
Decodo

decodo-scraper

by Decodo

Search Google, scrape web pages, Amazon product pages, YouTube subtitles, or Reddit (post/subreddit) using the Decodo Scraper OpenClaw Skill.

CLI Tools 151 6mo ago