Scraping

Web scraping and data extraction

Showing 265-288 of 708 skills
saadshahd

shape

by saadshahd

Bridge WHAT (intent) to HOW (implementation). Use when spec is clear but approach is not. Triggers on "shape this", "how should I build", "implementation approach".

Auth 34 6mo ago
saadshahd

consult

by saadshahd

Use when asking "code like [expert]", "what would [expert] say", "idiomatic", "best practice", "panel", "debate", or needing domain guidance. Triggers on expert names, style requests, tradeoff questions, or "stuck on".

Code Review 34 6mo ago
kazukinagata

reading-invoice

by kazukinagata

請求書の画像を読み取り構造化データを返す。 他のスキルから呼び出されるほか、直接ユーザーが呼び出すことも可能。

Docker 357 6mo ago
liyecom

playwright

by liyecom

Web 自动化测试与浏览器控制

Prompts 33 7mo ago
tebjan

vvvv-patching

by tebjan

Explains vvvv gamma visual programming patterns — dataflow, node connections, regions (ForEach/If/Switch/Repeat/Accumulator), channels for reactive data flow, event handling (Bang/Toggle/FrameDelay/Changed), patch organization, and common anti-patterns (circular dependencies, polling vs reacting, ignoring Nil). Use when the user asks about patching best practices, dataflow patterns, event handling, or how to structure visual programs.

Code Gen 32 6mo ago
zephyrwang6

markdown-to-image

by zephyrwang6

将 Markdown 内容转换为精美的图片海报。特别适合将播客摘要、文章内容转为社交媒体分享图片。固定 3:4 比例,支持 YouTube 视频封面作为头图。触发词:「转图片」「Markdown 转图片」「生成海报」「生成分享图」「把这个转成图片」。

Docs Gen 336 7mo ago
triggerdotdev

trigger-config

by triggerdotdev

Configure Trigger.dev projects with trigger.config.ts. Use when setting up build extensions for Prisma, Playwright, FFmpeg, Python, or customizing deployment settings.

Database 30 7mo ago
acedergren

firecrawl

by acedergren

Web scraping and search CLI returning clean Markdown from any URL (handles JS-rendered pages, SPAs). Use when user requests: (1) "search the web for X", (2) "scrape/fetch URL content", (3) "get content from website", (4) "find recent articles about X", (5) research tasks needing current web data, (6) extract structured data from pages. Outputs LLM-friendly Markdown, handles authentication via firecrawl login, supports parallel scraping for bulk operations. Automatically writes to .firecrawl/ directory. Triggers: web scraping, search web, fetch URL, extract content, Firecrawl, scrape website, get page content, web research, site map, crawl site.

Processing 26 7mo ago
zephyrwang6

x-blogger-analyzer

by zephyrwang6

分析 X/Twitter 博主的内容风格、创作策略和增长原因。当用户输入 X/Twitter 博主链接(如 https://x.com/username 或 https://twitter.com/username)并要求分析时触发。支持:(1) 抓取博主推文内容,(2) 分析爆款原因和增长策略,(3) 提炼内容风格和创作频率,(4) 生成完整分析报告保存到笔记。

Cloud 336 7mo ago
armanzeroeight

refactoring-advisor

by armanzeroeight

Provides refactoring recommendations and step-by-step improvement plans. Use when planning refactoring, improving code structure, or reducing technical debt.

Code Gen 29 9mo ago
bahayonghang

pdf

by bahayonghang

Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multiple PDFs into one, splitting PDFs apart, rotating pages, adding watermarks, creating new PDFs, filling PDF forms, encrypting/decrypting PDFs, extracting images, and OCR on scanned PDFs to make them searchable. If the user mentions a .pdf file or asks to produce one, use this skill.

CLI Tools 16 6mo ago
YPares

read-bin-docs

by YPares

Straightforward text extraction from document files (text-based PDF only for now, no OCR or docx). Use when you just need to read/extract text from binary documents.

CLI Tools 29 9mo ago
zephyrwang6

pdf

by zephyrwang6

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI Tools 336 8mo ago
EthanAlgoX

sherpa-onnx-tts

by EthanAlgoX

Local text-to-speech via sherpa-onnx (offline, no cloud)

Git & VCS 65 7mo ago
zephyrwang6

Working Nomads 远程工作爬取工具

by zephyrwang6

爬虫使用 Playwright è¿›è¡Œé¡µé¢æ¸²æŸ“ï¼Œæ”¯æŒåŠ¨æ€åŠ è½½çš„ Angular 应用。

Processing 336 7mo ago
ZhihaoAIRobotic

summarize

by ZhihaoAIRobotic

Summarize or extract text/transcripts from URLs, podcasts, and local files (great fallback for “transcribe this YouTube/video”).

Processing 158 7mo ago
agenticnotetaking

reduce

by agenticnotetaking

Extract structured knowledge from source material. Comprehensive extraction is the default — every insight that serves the domain gets extracted. For domain-relevant sources, skip rate must be below 10%. Zero extraction from a domain-relevant source is a BUG. Triggers on "/reduce", "/reduce [file]", "extract insights", "mine this", "process this".

Automation 3.5K 6mo ago
organvm-iv-taxis

pdf

by organvm-iv-taxis

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI Tools 16 7mo ago
manutej

apache-airflow-orchestration

by manutej

Complete guide for Apache Airflow orchestration including DAGs, operators, sensors, XComs, task dependencies, dynamic workflows, and production deployment

Automation 61 10mo ago
MaxiDonkey

pdf-extract

by MaxiDonkey

Extrait le texte et les tableaux des fichiers PDF, remplit les formulaires, fusionne les documents. À utiliser lors du travail avec des fichiers PDF ou lorsque l'utilisateur mentionne les PDF, les formulaires ou l'extraction de documents.

Scraping 60 6mo ago
MiniMax-AI

webapp-testing

by MiniMax-AI

Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.

Automation 3K 10mo ago
pacphi

webapp-testing

by pacphi

Toolkit for interacting with and testing local web applications using Playwright. Supports verifying frontend functionality, debugging UI behavior, capturing browser screenshots, and viewing browser logs.

Automation 15 7mo ago
pacphi

playwright

by pacphi

Browser automation, web scraping, and visual testing with Playwright on Display :1. Use for navigating web pages, clicking elements, filling forms, taking screenshots, executing JavaScript, and visual verification of web applications. Visual access available via VNC on port 5901.

Scraping 15 7mo ago
openclaw

4chan-reader

by openclaw

Browse 4chan boards and extract thread discussions into structured text files. Use when you need to fetch catalog information or specific thread content (including post text and file metadata) from 4chan boards like /a/, /vg/, /v/, etc.

Scraping 4.5K 7mo ago