抓取

网页抓取与数据提取

显示 265-288 / 共 708 个技能
saadshahd

shape

saadshahd

Bridge WHAT (intent) to HOW (implementation). Use when spec is clear but approach is not. Triggers on "shape this", "how should I build", "implementation approach".

认证鉴权 34 6个月前
saadshahd

consult

saadshahd

Use when asking "code like [expert]", "what would [expert] say", "idiomatic", "best practice", "panel", "debate", or needing domain guidance. Triggers on expert names, style requests, tradeoff questions, or "stuck on".

代码评审 34 6个月前
kazukinagata

reading-invoice

kazukinagata

請求書の画像を読み取り構造化データを返す。 他のスキルから呼び出されるほか、直接ユーザーが呼び出すことも可能。

Docker 357 6个月前
liyecom

playwright

liyecom

Web 自动化测试与浏览器控制

提示词 33 7个月前
tebjan

vvvv-patching

tebjan

Explains vvvv gamma visual programming patterns — dataflow, node connections, regions (ForEach/If/Switch/Repeat/Accumulator), channels for reactive data flow, event handling (Bang/Toggle/FrameDelay/Changed), patch organization, and common anti-patterns (circular dependencies, polling vs reacting, ignoring Nil). Use when the user asks about patching best practices, dataflow patterns, event handling, or how to structure visual programs.

代码生成 32 6个月前
zephyrwang6

markdown-to-image

zephyrwang6

将 Markdown 内容转换为精美的图片海报。特别适合将播客摘要、文章内容转为社交媒体分享图片。固定 3:4 比例,支持 YouTube 视频封面作为头图。触发词:「转图片」「Markdown 转图片」「生成海报」「生成分享图」「把这个转成图片」。

文档生成 336 7个月前
triggerdotdev

trigger-config

triggerdotdev

Configure Trigger.dev projects with trigger.config.ts. Use when setting up build extensions for Prisma, Playwright, FFmpeg, Python, or customizing deployment settings.

数据库 30 7个月前
acedergren

firecrawl

acedergren

Web scraping and search CLI returning clean Markdown from any URL (handles JS-rendered pages, SPAs). Use when user requests: (1) "search the web for X", (2) "scrape/fetch URL content", (3) "get content from website", (4) "find recent articles about X", (5) research tasks needing current web data, (6) extract structured data from pages. Outputs LLM-friendly Markdown, handles authentication via firecrawl login, supports parallel scraping for bulk operations. Automatically writes to .firecrawl/ directory. Triggers: web scraping, search web, fetch URL, extract content, Firecrawl, scrape website, get page content, web research, site map, crawl site.

数据处理 26 7个月前
zephyrwang6

x-blogger-analyzer

zephyrwang6

分析 X/Twitter 博主的内容风格、创作策略和增长原因。当用户输入 X/Twitter 博主链接(如 https://x.com/username 或 https://twitter.com/username)并要求分析时触发。支持:(1) 抓取博主推文内容,(2) 分析爆款原因和增长策略,(3) 提炼内容风格和创作频率,(4) 生成完整分析报告保存到笔记。

云服务 336 7个月前
armanzeroeight

refactoring-advisor

armanzeroeight

Provides refactoring recommendations and step-by-step improvement plans. Use when planning refactoring, improving code structure, or reducing technical debt.

代码生成 29 9个月前
bahayonghang

pdf

bahayonghang

Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multiple PDFs into one, splitting PDFs apart, rotating pages, adding watermarks, creating new PDFs, filling PDF forms, encrypting/decrypting PDFs, extracting images, and OCR on scanned PDFs to make them searchable. If the user mentions a .pdf file or asks to produce one, use this skill.

CLI 工具 16 6个月前
YPares

read-bin-docs

YPares

Straightforward text extraction from document files (text-based PDF only for now, no OCR or docx). Use when you just need to read/extract text from binary documents.

CLI 工具 29 9个月前
zephyrwang6

pdf

zephyrwang6

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI 工具 336 8个月前
EthanAlgoX

sherpa-onnx-tts

EthanAlgoX

Local text-to-speech via sherpa-onnx (offline, no cloud)

Git 与版本控制 65 7个月前
zephyrwang6

Working Nomads 远程工作爬取工具

zephyrwang6

爬虫使用 Playwright è¿›è¡Œé¡µé¢æ¸²æŸ“ï¼Œæ”¯æŒåŠ¨æ€åŠ è½½çš„ Angular 应用。

数据处理 336 7个月前
ZhihaoAIRobotic

summarize

ZhihaoAIRobotic

Summarize or extract text/transcripts from URLs, podcasts, and local files (great fallback for “transcribe this YouTube/video”).

数据处理 158 7个月前
agenticnotetaking

reduce

agenticnotetaking

Extract structured knowledge from source material. Comprehensive extraction is the default — every insight that serves the domain gets extracted. For domain-relevant sources, skip rate must be below 10%. Zero extraction from a domain-relevant source is a BUG. Triggers on "/reduce", "/reduce [file]", "extract insights", "mine this", "process this".

自动化 3482 6个月前
organvm-iv-taxis

pdf

organvm-iv-taxis

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI 工具 16 7个月前
manutej

apache-airflow-orchestration

manutej

Complete guide for Apache Airflow orchestration including DAGs, operators, sensors, XComs, task dependencies, dynamic workflows, and production deployment

自动化 61 10个月前
MaxiDonkey

pdf-extract

MaxiDonkey

Extrait le texte et les tableaux des fichiers PDF, remplit les formulaires, fusionne les documents. À utiliser lors du travail avec des fichiers PDF ou lorsque l'utilisateur mentionne les PDF, les formulaires ou l'extraction de documents.

抓取 60 6个月前
MiniMax-AI

webapp-testing

MiniMax-AI

Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.

自动化 2983 10个月前
pacphi

webapp-testing

pacphi

Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.

自动化 15 7个月前
pacphi

playwright

pacphi

Browser automation, web scraping, and visual testing with Playwright on Display :1. Use for navigating web pages, clicking elements, filling forms, taking screenshots, executing JavaScript, and visual verification of web applications. Visual access available via VNC on port 5901.

抓取 15 7个月前
openclaw

4chan-reader

openclaw

Browse 4chan boards and extract thread discussions into structured text files. Use when you need to fetch catalog information or specific thread content (including post text and file metadata) from 4chan boards like /a/, /vg/, /v/, etc.

抓取 4465 7个月前