抓取

网页抓取与数据提取

显示 145-168 / 共 708 个技能
obra

browsing

obra

Use when you need direct browser control - teaches Chrome DevTools Protocol for controlling existing browser sessions, multi-tab management, form automation, and content extraction via use_browser MCP tool

性能 345 6个月前
NeverSight

extract-transcripts

NeverSight

Extract readable transcripts from Claude Code and Codex CLI session JSONL files

认证鉴权 203 7个月前
actionbook

extract

actionbook

Extract structured data from websites and produce an executable Playwright script plus extracted data. Use when the user wants to scrape, extract, pull, collect, or harvest data from any website — product listings, tables, search results, feeds, profiles, or any repeating content.

CLI 工具 1584 6个月前
NeverSight

playwright-best-practices

NeverSight

Provides Playwright test patterns for resilient locators, Page Object Models, fixtures, web-first assertions, and network mocking. Must use when writing or modifying Playwright tests (.spec.ts, .test.ts files with @playwright/test imports).

数据处理 203 7个月前
jmagly

nl-router

jmagly

Translation table: docs/simple-language-translations.md

代码生成 204 7个月前
NeverSight

playwright-best-practices

NeverSight

Provides Playwright test patterns for resilient locators, Page Object Models, fixtures, web-first assertions, and network mocking. Must use when writing or modifying Playwright tests (.spec.ts, .test.ts files with @playwright/test imports).

数据处理 203 7个月前
appautomaton

pdf

appautomaton

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI 工具 153 7个月前
dkyazzentwatwa

business-card-scanner

dkyazzentwatwa

Extract contact information from business card images using OCR - name, company, email, phone, address.

CLI 工具 95 8个月前
Mindrally

cheerio-parsing

Mindrally

Expert guidance for HTML/XML parsing using Cheerio in Node.js with best practices for DOM traversal, data extraction, and efficient scraping pipelines.

数据处理 243 7个月前
AutoForgeAI

playwright-cli

AutoForgeAI

Automates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to navigate websites, interact with web pages, fill forms, take screenshots, test web applications, or extract information from web pages.

CLI 工具 1769 6个月前
clerk

clerk-testing

clerk

E2E testing for Clerk apps. Use with Playwright or Cypress for auth flow tests.

认证鉴权 69 7个月前
brightdata

scrape

brightdata

Scrape any webpage as clean markdown via Bright Data Web Unlocker API. Bypasses bot detection and CAPTCHA. Requires BRIGHTDATA_API_KEY and BRIGHTDATA_UNLOCKER_ZONE environment variables.

文档生成 255 7个月前
vm0-ai

firecrawl

vm0-ai

Firecrawl web scraping API via curl. Use this skill to scrape webpages, crawl websites, discover URLs, search the web, or extract structured data.

CLI 工具 77 8个月前
aAAaqwq

API Provider Status Skill

aAAaqwq

OpenRouter VIP - ✅ 可用

API 开发 89 6个月前
einverne

chrome-devtools

einverne

Browser automation, debugging, and performance analysis using Puppeteer CLI scripts. Use for automating browsers, taking screenshots, analyzing performance, monitoring network traffic, web scraping, form automation, and JavaScript debugging.

CLI 工具 121 9个月前
einverne

cloudflare-browser-rendering

einverne

Guide for implementing Cloudflare Browser Rendering - a headless browser automation API for screenshots, PDFs, web scraping, and testing. Use when automating browsers, taking screenshots, generating PDFs, scraping dynamic content, extracting structured data, or testing web applications. Supports REST API, Workers Bindings (Puppeteer/Playwright), MCP servers, and AI-powered automation. (project)

认证鉴权 121 9个月前
vuejs-ai

vue-testing-best-practices

vuejs-ai

Use for Vue.js testing. Covers Vitest, Vue Test Utils, component testing, mocking, testing patterns, and Playwright for E2E testing.

抓取 2801 7个月前
yonatangross

browser-tools

yonatangross

OrchestKit orchestration wrapper for browser automation. Adds security rules, rate limiting, and ethical scraping guardrails on top of the upstream agent-browser skill. Use when automating browser workflows, capturing web content, or extracting structured data from web pages.

智能体 224 6个月前
BrownFineSecurity

ffind

BrownFineSecurity

Advanced file finder with type detection and filesystem extraction for analyzing firmware and extracting embedded filesystems. Use when you need to analyze firmware files, identify file types, or extract ext2/3/4 or F2FS filesystems.

代码评审 814 8个月前
alinaqi

playwright-testing

alinaqi

E2E testing with Playwright - Page Objects, cross-browser, CI/CD

认证鉴权 705 8个月前
qdhenry

extract-video-frames

qdhenry

Extracts frames and timestamped audio segments from video files (GIF, MP4, MOV) at configurable intervals and stores them in a directory with a manifest file. Use when analyzing video content, preparing frames for visual review, extracting audio for transcription, or creating frame+audio sequences for another agent to process.

智能体 1333 6个月前
victorgarciaesgi

regle-typescript

victorgarciaesgi

TypeScript support for type-safe Regle form validation, rules, and component props.

抓取 494 7个月前
lofcz

pdf-processor

lofcz

Extracts text and tables from PDF files, fills forms, and merges documents. Use when working with PDF files or when the user mentions PDFs, forms, or document extraction.

API 开发 638 10个月前
wcygan

playwright-cli

wcygan

Automates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to navigate websites, interact with web pages, fill forms, take screenshots, test web applications, or extract information from web pages. Keywords: browser, automation, playwright, web testing, screenshot, form, click, navigate, scrape

CLI 工具 194 7个月前