抓取

网页抓取与数据提取

显示 289-312 / 共 708 个技能
JosiahSiegel

ffmpeg-audio-processing

JosiahSiegel

Complete audio encoding and normalization system. PROACTIVELY activate for: (1) Audio codec selection (AAC, MP3, Opus, FLAC), (2) Loudness normalization (EBU R128, loudnorm), (3) Audio extraction from video, (4) Format conversion, (5) Volume adjustment and dynamics, (6) Noise reduction and EQ, (7) Channel operations (stereo/mono/surround), (8) Sample rate and bit depth conversion, (9) Audio fade in/out and crossfades, (10) Podcast and broadcast processing chains. Provides: Codec comparison tables, loudness standards reference, two-pass normalization scripts, professional mastering chains. Ensures: Broadcast-compliant audio with proper loudness and quality.

CLI 工具 51 7个月前
openclaw

Instructions

openclaw

All versions of all skills that are on clawhub.com archived

数据处理 4472 7个月前
parcadei

firecrawl-scrape

parcadei

Scrape web pages and extract content via Firecrawl MCP

数据处理 3930 7个月前
joaquimscosta

playwright-cli

joaquimscosta

Browser automation via Playwright CLI for navigating pages, interacting with elements, capturing screenshots, and testing web applications through shell commands. Use when user mentions "playwright", "browser automation", "take a screenshot", "browser testing", "headless browser", "web testing", "fill out a form", "e2e test", or needs to automate browser workflows from the command line. This is a pre-installed CLI tool — do NOT install anything via npx or npm. Invoke this skill first, then use playwright-cli bash commands.

CLI 工具 21 6个月前
Takazudo

dev-figma-capture

Takazudo

Capture web pages and send them to Figma as editable design files. Use when: (1) User wants to capture a webpage to Figma, (2) User says 'figma capture', 'send to figma', 'capture to figma', (3) User provides URLs to convert to Figma designs

自动化 12 6个月前
idanbeck

playwright-skill

idanbeck

Browser automation for web tasks. Use when the user needs to automate browser interactions, take screenshots, extract data from websites, fill forms, or perform web scraping. Supports persistent sessions for logged-in state.

认证鉴权 12 7个月前
Takazudo

headless-browser

Takazudo

Browser automation skill with two efficiency tiers. Tier 1: lightweight headless-check.js for quick checks, screenshots, error detection. Tier 2: playwright-cli for interactions (click, fill, navigate). Use when: (1) Quick webpage health checks, (2) Taking screenshots, (3) Checking console/network errors, (4) Simple interactions like clicking buttons or filling forms, (5) Multi-step browser automation. Use MCP Playwright only for complex scenarios requiring persistent context or rich introspection.

认证鉴权 12 6个月前
LdotJdot

webbrowser

LdotJdot

"浏览网页。与 read skill 风格一致:直接 exec 调用 exe,传参执行,stdout 为结果。"

数据处理 47 6个月前
JochenYang

remotion-best-practices

JochenYang

Best practices for Remotion - Video creation in React

动画 20 6个月前
ZhanlinCui

pdf

ZhanlinCui

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI 工具 182 7个月前
besoeasy

crawl-websites-at-scale

besoeasy

"Scrape websites at scale using Scrapy, a Python web crawling and scraping framework. Use when: (1) Crawling multiple pages or entire sites, (2) Extracting structured data from HTML/XML, or (3) Building automated data pipelines from web sources."

数据处理 132 6个月前
besoeasy

phone-specs-scraper

besoeasy

"Scrape phone specifications from GSM Arena, PhoneDB, and alternative sites. Use when: (1) Comparing smartphone specs, (2) Researching device features, or (3) Building phone comparison tools."

设计 132 28个月前
besoeasy

using-web-scraping

besoeasy

Search and scrape public web content with headless Chrome and DuckDuckGo using safe practices.

抓取 132 28个月前
lijigang

ljg-xray-book

lijigang

Deep structure extraction from books using the Epiplexity principle - maximizing computational investment to extract maximum learnable structure from any book.

CLI 工具 458 7个月前
Casper-Studios

firecrawl-scraping

Casper-Studios

Web page and website scraping with Firecrawl API. Use this skill when scraping web articles, blog posts, documentation pages, paywalled content, or JavaScript-heavy sites. Triggers on requests to scrape websites, extract article content, convert pages to markdown, or handle anti-bot protection.

文档生成 12 7个月前
Casper-Studios

apify-scrapers

Casper-Studios

Social media and web scraping using Apify actors. Use this skill when scraping Twitter/X tweets, Reddit posts, LinkedIn posts, Instagram profiles/posts/reels, Facebook pages/posts/groups, TikTok videos, YouTube content, Google Maps businesses/reviews, contact enrichment (emails/phones from websites), or when auto-detecting URL type to scrape. Triggers on requests to scrape social media, get trending posts, extract business info, find contact details, or extract content from social URLs.

CLI 工具 12 7个月前
dkyazzentwatwa

audio-trimmer

dkyazzentwatwa

Cut, trim, and edit audio segments with fade effects, speed control, concatenation, and basic audio manipulations.

Lint 与格式化 95 8个月前
dkyazzentwatwa

image-metadata-tool

dkyazzentwatwa

Extract EXIF metadata from images including GPS coordinates, camera settings, and timestamps. Map photo locations and strip metadata for privacy.

数据处理 95 8个月前
dkyazzentwatwa

color-palette-extractor

dkyazzentwatwa

Extract dominant colors from images, generate color palettes, and export as CSS, JSON, or ASE with K-means clustering.

数据处理 95 8个月前
yangliu2060

video-creator

yangliu2060

AI短视频创作与多平台发布,使用即梦MCP生成视频,使用Playwright MCP自动发布到YouTube/TikTok/Instagram/Facebook/LinkedIn/Twitter等平台。

提示词 40 8个月前
nonameplum

swift-async-stream-patterns

nonameplum

Patterns and best practices for building robust AsyncStream and AsyncSequence types, learned from swift-async-algorithms.

代码生成 17 7个月前
alexei-led

testing-e2e

alexei-led

E2E testing with Playwright MCP for browser automation, test generation, and UI testing. Use when discussing E2E tests, Playwright, browser testing, UI automation, visual testing, or accessibility testing. Supports TypeScript tests and Go/HTMX web applications.

无障碍 35 7个月前
jackspace

cloudflare-browser-rendering

jackspace

Complete knowledge domain for Cloudflare Browser Rendering - Headless Chrome automation with Puppeteer and Playwright on Cloudflare Workers for screenshots, PDFs, web scraping, and browser automation workflows. Use when: taking screenshots, generating PDFs from HTML or URLs, web scraping content, crawling websites, browser automation tasks, testing web applications, managing browser sessions, performing batch browser operations, integrating with AI for content extraction, or encountering browser rendering errors, XPath selector errors, browser timeout issues, concurrency limits, memory exceeded errors, or "Cannot read properties of undefined (reading 'fetch')" errors. Keywords: browser rendering cloudflare, @cloudflare/puppeteer, @cloudflare/playwright, puppeteer workers, playwright workers, screenshot cloudflare, pdf generation workers, web scraping cloudflare, headless chrome workers, browser automation, puppeteer.launch, playwright.chromium.launch, browser binding, session management, puppeteer.sessions, puppeteer.connect, browser.close, browser.disconnect, XPath not supported, browser timeout, concurrency limit, keep_alive, page.screenshot, page.pdf, page.goto, page.evaluate, incognito context, session reuse, batch scraping, crawling websites

云服务 16 9个月前
Decodo

decodo-scraper

Decodo

Search Google, scrape web pages, Amazon product pages, YouTube subtitles, or Reddit (post/subreddit) using the Decodo Scraper OpenClaw Skill.

CLI 工具 151 6个月前