ML Ops

Machine learning operations

Showing 241-264 of 1815 skills
davila7

tensorrt-llm

by davila7

Optimizes LLM inference with NVIDIA TensorRT for maximum throughput and lowest latency. Use for production deployment on NVIDIA GPUs (A100/H100), when you need 10-100x faster inference than PyTorch, or for serving models with quantization (FP8/INT4), in-flight batching, and multi-GPU scaling.

ML Ops 29.8K 6mo ago
davila7

ray-train

by davila7

Distributed training orchestration across clusters. Scales PyTorch/TensorFlow/HuggingFace from laptop to 1000s of nodes. Built-in hyperparameter tuning with Ray Tune, fault tolerance, elastic scaling. Use when training massive models across multiple machines or running distributed hyperparameter sweeps.

Analytics 29.8K 6mo ago
davila7

llama-cpp

by davila7

Runs LLM inference on CPU, Apple Silicon, and consumer GPUs without NVIDIA hardware. Use for edge deployment, M1/M2/M3 Macs, AMD/Intel GPUs, or when CUDA is unavailable. Supports GGUF quantization (1.5-8 bit) for reduced memory and 4-10× speedup vs PyTorch on CPU.

CLI Tools 29.8K 6mo ago
davila7

unsloth

by davila7

Expert guidance for fast fine-tuning with Unsloth - 2-5x faster training, 50-80% less memory, LoRA/QLoRA optimization

i18n 29.8K 6mo ago
tech-leads-club

subagent-creator

by tech-leads-club

Guide for creating AI subagents with isolated context for complex multi-step workflows. Use when users want to create a subagent, specialized agent, verifier, debugger, or orchestrator that requires isolated context and deep specialization. Works with any agent that supports subagent delegation. Triggers on "create subagent", "new agent", "specialized assistant", "create verifier".

Agents 4.9K 5mo ago
tech-leads-club

security-threat-model

by tech-leads-club

Repository-grounded threat modeling that enumerates trust boundaries, assets, attacker capabilities, abuse paths, and mitigations, and writes a concise Markdown threat model. Trigger only when the user explicitly asks to threat model a codebase or path, enumerate threats/abuse paths, or perform AppSec threat modeling. Do not trigger for general architecture summaries, code review, or non-security design work.

Code Gen 4.9K 5mo ago
davila7

deepspeed

by davila7

Expert guidance for distributed training with DeepSpeed - ZeRO optimization stages, pipeline parallelism, FP16/BF16/FP8, 1-bit Adam, sparse attention

Processing 29.8K 6mo ago
skills-directory

codex

by skills-directory

Use when the user asks to run Codex CLI (codex exec, codex resume) or references OpenAI Codex for code analysis, refactoring, or automated editing

Code Review 1.4K 5mo ago
bitwize-music-studio

about

by bitwize-music-studio

Provides information about the bitwize-music plugin and its creator. Use when the user asks about the plugin, its purpose, or capabilities.

Code Gen 366 5mo ago
elie222

wait

by elie222

Pause execution for a user-specified duration

CLI Tools 11.7K 4mo ago
Starlitnightly

bulk-rna-seq-deconvolution-with-bulk2single

by Starlitnightly

Turn bulk RNA-seq cohorts into synthetic single-cell datasets using omicverse's Bulk2Single workflow for cell fraction estimation, beta-VAE generation, and quality control comparisons against reference scRNA-seq.

Code Gen 1.2K 8mo ago
qwibitai

use-local-whisper

by qwibitai

Use when the user wants local voice transcription instead of OpenAI Whisper API. Switches to whisper.cpp running on Apple Silicon. WhatsApp only for now. Requires voice-transcription skill to be applied first.

CLI Tools 30.3K 4mo ago
openclaw

summarize

by openclaw

Summarize or extract text/transcripts from URLs, podcasts, and local files (great fallback for “transcribe this YouTube/video”).

Processing 383.8K 5mo ago
openclaw

model-usage

by openclaw

Use CodexBar CLI local cost usage to summarize per-model usage for Codex or Claude, including the current (most recent) model or a full model breakdown. Trigger when asked for model-level usage/cost data from codexbar, or when you need a scriptable per-model summary from codexbar cost JSON.

CLI Tools 383.8K 5mo ago
benjaminasterA

ai-engineer

by benjaminasterA

"Build production-ready LLM applications, advanced RAG systems, and"

Embeddings 64 5mo ago
benjaminasterA

3d-web-experience

by benjaminasterA

"Expert in building 3D experiences for the web - Three.js, React Three Fiber, Spline, WebGL, and interactive 3D scenes. Covers product configurators, 3D portfolios, immersive websites, and bringing ..."

Processing 64 5mo ago
NeoLabHQ

sadd:do-and-judge

by NeoLabHQ

Execute a task with sub-agent implementation and LLM-as-a-judge verification with automatic retry loop

Agents 1.3K 5mo ago
elizaOS

openai-whisper

by elizaOS

Local speech-to-text with the Whisper CLI (no API key).

API Dev 18.8K 5mo ago
openclaw

openai-whisper

by openclaw

Local speech-to-text with the Whisper CLI (no API key).

API Dev 383.8K 5mo ago
jaechang-hits

geniml

by jaechang-hits

"Geniml is a Python library for genomic interval machine learning. Train and apply region2vec embeddings to convert BED file regions into numeric vectors, load and index genomic interval datasets for ML pipelines, search embedding spaces with BEDSpace, and evaluate embedding quality. Use for chromatin accessibility clustering, regulatory element classification, cross-sample region comparison, and building ML models on genomic intervals."

Accessibility 279 5mo ago
getsentry

django-perf-review

by getsentry

Django performance code review. Use when asked to "review Django performance", "find N+1 queries", "optimize Django", "check queryset performance", "database performance", "Django ORM issues", or audit Django code for performance problems.

Database 875 5mo ago
elie222

step-by-step

by elie222

Execute tasks one step at a time with user confirmation

ML Ops 11.7K 4mo ago
huggingface

hf-mcp

by huggingface

Use Hugging Face Hub via MCP server tools. Search models, datasets, Spaces, papers. Get repo details, fetch documentation, run compute jobs, and use Gradio Spaces as AI tools. Available when connected to the HF MCP server.

Automation 10.9K 6mo ago
guanyang

advanced-evaluation

by guanyang

This skill should be used when the user asks to "implement LLM-as-judge", "compare model outputs", "create evaluation rubrics", "mitigate evaluation bias", or mentions direct scoring, pairwise comparison, position bias, evaluation pipelines, or automated quality assessment.

Processing 935 5mo ago