- Home
- /
- Categories
- /
- Scraping
Scraping
Web scraping and data extraction
拼多多/1688 商品爬虫
by MimonWish
Pinduoduo 1688 ecommerce product scraper
firmware-extraction
by tangjunyi23
Firmware extraction and filesystem analysis for IoT devices. Use when analyzing firmware binaries, extracting filesystems with binwalk, identifying firmware format/structure, locating key files after extraction, or performing initial reconnaissance on router/camera/IoT firmware images. Triggers on tasks involving .bin/.img/.trx/.chk firmware files.
playwright
by brettatoms
Browser automation for web testing and interaction. Use for navigating pages, filling forms, clicking elements, taking screenshots, and inspecting page content. Maintains stateful browser session across commands.
site-crawler
by mindmorass
Crawl and extract content from websites
playwright-testing
by dtinth
Playwright testing. Use this skill to write and run automated tests for web applications using Playwright.
claude-command-firecrawl-scrape
by monkey1sai
Converted from Claude plugin command "scrape" (firecrawl). Use when the
claude-command-firecrawl-map
by monkey1sai
Converted from Claude plugin command "map" (firecrawl). Use when the
claude-command-firecrawl-crawl
by monkey1sai
Converted from Claude plugin command "crawl" (firecrawl). Use when the
firmware-decryption
by tangjunyi23
Firmware decryption, deobfuscation, and unpacking for encrypted IoT firmware images. Use when firmware entropy analysis reveals encrypted/obfuscated content, when binwalk extraction fails due to encryption, when decrypting vendor-specific firmware encryption (D-Link, Netgear, TP-Link, Hikvision, Dahua, ZTE), or when reversing custom XOR/AES/DES encryption applied to firmware update files.
brightdata
by multicam
Progressive four-tier URL content scraping with automatic fallback strategy. USE WHEN user says "scrape this URL", "fetch this page", "get content from", "can't access this site", "use Bright Data", "pull content from URL", or needs to retrieve web content that may have bot detection or access restrictions.
extract-diagrams
by clearsmog
Extract CeTZ diagrams to SVG from Typst files. For TikZ extraction, configure at project level.
extract-spark-meetings
by johnie
Extract meeting summaries and action items from Spark Mail shared links. Processes single URLs (pass as argument) or batch processes unchecked links from links.md. Use when working with Spark Mail shared meeting links.
xiaohongshu-skill
by 1uokun
小红书内容发布技能,提供检查登录状态和发布图文内容的功能。不依赖MCP,使用内置JavaScript脚本执行小红书相关操作。
research-intelligence
by spitoglou
Extract insights, analyze claims, and synthesize knowledge from research content. Use when processing academic papers, articles, podcasts, videos, meeting transcripts, or any content where the goal is to extract wisdom, analyze arguments, summarize findings, or compile references. Triggers include "analyze this paper", "extract key insights", "summarize the research", "what are the main claims", "extract wisdom from", "compile references", "critique this argument".
pdf-analysis
by bahayonghang
This skill should be used when the user asks to "解析PDF", "解读文档", "分析PDF文件", "PDF解读", "extract content from PDF", "analyze PDF document", "parse academic paper", or provides a PDF file path for content extraction and analysis. Provides comprehensive PDF document analysis and content extraction capabilities for WeChat content creation.
ask-refactoring-readability
by NavanithanS
Refactor code for readability using DRY, meaningful names, and modularization.
ogie
by DobroslavRadosavljevic
Extract OpenGraph, Twitter Cards, and metadata from URLs or HTML. Use when building link previews, SEO tools, or scraping webpage metadata.
article-extractor
by jrajasekera
Extract clean article content from URLs and save as markdown. Triggers when user provides a webpage URL and wants to download it, extract content, get a clean version without ads, capture an article for offline reading, save an article, grab content from a page, archive a webpage, clip an article, or read something later. Handles blog posts, news articles, tutorials, documentation pages, and similar web content. Supports Wayback Machine for dead links or paywalled content. This skill handles the entire workflow - do NOT use web_fetch or other tools first, just call the extraction script directly with the URL.
pdf-processor
by aig787
Process PDF files for text extraction, form filling, and document analysis. Use when you need to extract content from PDFs, fill forms, or analyze document structure.
ffmpeg-daemon
by fast-gateway-protocol
Fast video/audio processing via FGP daemon - 5-20x faster than spawning ffmpeg per operation. Use when user needs to convert videos, extract audio, trim clips, resize, add watermarks, or transcode. Triggers on "convert video", "extract audio", "trim video", "compress video", "ffmpeg", "video editing", "transcode".
by wollfoo
Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Factory needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale. Sử dụng khi: xử lý PDF, trích xuất, ghép file, chia nhỏ, điền form PDF.
PR Feedback Training-First Loop
by violetio
Extract, learn, and integrate PR feedback into the Violet brain
by Krosebrook
Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.
by ashleytower
Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.