Scraping

Web scraping and data extraction

Showing 457-480 of 700 skills
horuz-ai

pdf

by horuz-ai

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI Tools 5 7mo ago
mikefilsaime-groove

pdf

by mikefilsaime-groove

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI Tools 5 5mo ago
aashari

mail-meetings

by aashari

Find meeting invites, calendar events, and meeting-related emails (notes, agendas, reschedules). Use when user asks about meetings in their email, upcoming invites, or wants to see meeting notes. Arguments: optional time range or "upcoming", "today", "this week".

CLI Tools 5 4mo ago
aashari

mail-action-items

by aashari

Extract action items, tasks, and to-dos from recent emails. Scan email bodies for requests, deadlines, approvals needed, and follow-ups. Use when user wants to know what they need to do based on their email, or asks "what do I need to act on?" Arguments: optional time range (default: last 3 days) or account filter.

CLI Tools 5 4mo ago
krishagel

browser-control

by krishagel

Full browser control for authenticated web interactions using Playwright scripts

CLI Tools 5 7mo ago
SherifEldeeb

registry-forensics

by SherifEldeeb

Analyze Windows Registry hives for forensic investigation. Use when investigating malware persistence, user activity, system configuration changes, or evidence of program execution. Supports offline registry analysis from disk images or extracted hives.

Code Review 5 6mo ago
bdmorin

analyze-malware

by bdmorin

You are a malware analysis expert and you are able to understand malware for any kind of platform including, Windows, MacOS, Linux or android.

Docs Gen 5 4mo ago
aashari

mail-expenses

by aashari

Extract financial transactions, expenses, receipts, payments, and invoices from email. Summarize spending with amounts, merchants, and categories. Use when user asks about expenses, spending, receipts, payments, or financial transactions from email. Arguments: time range like "last 24 hours", "this month", "last week", or a specific date range.

CLI Tools 5 4mo ago
katyella

research-extract

by katyella

Ingest and analyze content from YouTube, podcasts, blogs, PDFs, and audio files. Extract structured insights using parallel agent teams. Generate Show Notes and Cheat Sheet HTML variants. Use /research-extract when you want to analyze any content source and extract key insights, quotes, themes, challenges, solutions, frameworks, and external resources.

Academic 5 5mo ago
bdmorin

analyze-answers

by bdmorin

You are a PHD expert on the subject defined in the input section provided below.

Code Gen 5 4mo ago
bdmorin

analyze-threat-report

by bdmorin

You are a super-intelligent cybersecurity expert.

Analytics 5 4mo ago
krishagel

pdf

by krishagel

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI Tools 5 8mo ago
adaptationio

ac-insight-extractor

by adaptationio

Extract insights from autonomous coding sessions. Use when learning from completions, extracting patterns, analyzing decisions, or improving future performance.

Agents 11 6mo ago
iota9star

querying-json

by iota9star

Extracts specific fields from JSON files efficiently using jq instead of reading entire files, saving 80-95% context. Use this skill when querying JSON files, filtering/transforming data, or getting specific field(s) from large JSON files

Processing 10 7mo ago
founderjourney

article-extractor

by founderjourney

Extract clean article content from URLs, removing ads, navigation, and clutter. Save as readable text files for research, archiving, or offline reading.

Scraping 10 6mo ago
samhvw8

chrome-devtools

by samhvw8

"Browser automation via Puppeteer CLI scripts (JSON output). Capabilities: screenshots, PDF generation, web scraping, form automation, network monitoring, performance profiling, JavaScript debugging, headless browsing. Actions: screenshot, scrape, automate, test, profile, monitor, debug browser. Keywords: Puppeteer, headless Chrome, screenshot, PDF, web scraping, form fill, click, navigate, network traffic, performance audit, Lighthouse, console logs, DOM manipulation, element selector, wait, scroll, automation script. Use when: taking screenshots, generating PDFs from web, scraping websites, automating form submissions, monitoring network requests, profiling page performance, debugging JavaScript, testing web UIs."

CLI Tools 10 7mo ago
iota9star

querying-yaml

by iota9star

Extracts specific fields from YAML files efficiently using yq instead of reading entire files, saving 80-95% context. Use this skill when querying YAML files, filtering/transforming configuration data, or getting specific field(s) from large YAML files like docker-compose.yml or GitHub Actions workflows

Processing 10 7mo ago
samhvw8

pdf

by samhvw8

"PDF document processing and manipulation. Tools: Python (PyPDF2, pdfplumber, reportlab), CLI tools. Capabilities: text extraction, table extraction, form filling, merge/split documents, create PDFs, add annotations, watermarks, page manipulation. Actions: extract, create, merge, split, fill, annotate PDFs. Keywords: PDF, text extraction, table extraction, form fill, PDF form, merge PDF, split PDF, create PDF, reportlab, PyPDF2, pdfplumber, annotation, watermark, page rotation, PDF metadata, bookmarks, OCR. Use when: extracting text/tables from PDFs, filling PDF forms, merging/splitting documents, creating PDFs programmatically, adding annotations/watermarks, processing PDFs at scale."

CLI Tools 10 7mo ago
founderjourney

firecrawl

by founderjourney

Web scraping, search, and data extraction using Firecrawl API. Use when users need to fetch web content, discover URLs on sites, search the web, or extract structured data from pages.

Embeddings 10 6mo ago
founderjourney

digitaliza-data-extractor

by founderjourney

Extract and prepare client data for digitalizaweb.vercel.app LinkTree-style digital cards. Use when: (1) Processing restaurant/business client folders containing screenshots, scraped HTML, or LinkTree data, (2) Extracting brand colors from logos/images, (3) Generating Digitaliza-ready JSON with slug, name, links, colors, and theme configuration, (4) Batch processing multiple client folders for 100+ restaurants project, (5) User mentions "digitaliza", "tarjeta digital", "linktree", "extraer datos de cliente", or "procesar carpeta de restaurante".

Processing 10 6mo ago
founderjourney

pdf

by founderjourney

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs to fill in a PDF form or programmatically process, generate, or analyze PDF documents at scale.

CLI Tools 10 6mo ago
partme-ai

stitch-mcp-get-project

by partme-ai

Retrieves the detailed metadata of a specific Stitch project.

Processing 4 5mo ago
arlenagreer

playwright-cli

by arlenagreer

Automates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to navigate websites, interact with web pages, fill forms, take screenshots, test web applications, extract information from web pages, debug web apps, record browser sessions as video, mock or intercept API requests, manage browser cookies/localStorage, generate Playwright test code, capture execution traces, or run multiple browser sessions concurrently.

CLI Tools 4 5mo ago
89jobrien

PDF Processing

by 89jobrien

Extract text and tables from PDF files, fill forms, merge documents.

Processing 4 7mo ago