DevOps
Vmware Aria
Use this skill whenever the user needs VMware Aria Operations data — performance metrics, alerts, capacity planning, anomaly detection, and automated reports...
Productivity
builder-stats
Use this skill when the user says "builder stats", "stats check", "AI-native check-in", "scan my machine", "run my stats", "how AI-native am I", or when a user pastes an install-an…
Data
garmin-health-analysis
Talk to your Garmin data naturally - "what was my fastest speed snowboarding?", "how did I sleep last night?", "what was my heart rate at 3pm?". Access 20+ metrics (sleep stages, B…
AI / ML
empirical-prompt-tuning
Fetch and execute mizchi's empirical-prompt-tuning skill at runtime. Use when evaluating or iteratively refining an agent-facing prompt (skill / slash command / task prompt / CLAUD…
Data
enrich-context
Augments a Wren project with business context missing from schema: enum meanings, units, NULL semantics, sentinels, soft-delete rules, synonyms, time conventions, cross-system IDs,…
Productivity
pm-metrics-critic
Critiques a metrics dashboard, success-criteria section, or proposed North Star metric against the repo's metrics guide. Surfaces common failures: aggregate metrics hiding segment …
AI / ML
ce-optimize
Run metric-driven iterative optimization loops. Define a goal, add measurement scaffolding, execute parallel experiments across approaches, score results against gates or quality j…
Productivity
rz-quarterly-review
Triggers on first Sunday of January/April/July/October or explicit quarterly review commands. Runs an 80-100 minute process: pulls metrics, re-scores channels on a 5-criteria frame…
AI / ML
inference-perf-baseline-bridge
Connects inference benchmark outputs to the perf-baseline registry, capturing canonical metrics (TTFT, ITL, throughput, cache hit rate) under inference_perfbench_v1 with workload-t…
Automation
os-improvement-loop
Coordinates concurrent multi-agent improvement cycles using shared event bus and memory. Each cycle executes, evaluates (KEEP/DISCARD), emits friction events, persists metrics and …
Testing
loop-evals
Designs the evaluation harness for agent loops, establishing trustworthy verification through a 7-layer suite with false-completion-rate and repair-productivity as core metrics. En…
Business
pm-progress-auditor
Audit a status update, exec review, board email, all-hands talking point, or dashboard callout for credibility leaks before sending. Flags overstated claims, cherry-picked windows,…
Engineering
run
Sustained metric-improvement loop with atomic commits, auto-rollback, and experiment logging. Iterates with specialist agents, commits atomically, and auto-rolls back on regression…
Documentation
docs-design
Design or measure documentation systems: plan information architecture, mode taxonomy, README/quickstart/reference structures, tool contracts, error messaging, versioning, and acce…
AI / ML
compoundos
Design and operate a self-improving AI business operating system with nine integrated components: strategy, prioritization, knowledge, operations, department agents, projects, auto…
DevOps
observabilityaudit
Comprehensive observability audit scoring 18 dimensions: logging, tracing (OpenTelemetry), metrics (RED/USE), dashboards, alerts, SLOs, error tracking, retention, sampling, probes,…
Productivity
pm-north-star-selector
Selects a single North Star metric for a product. Weighs candidates across behavioral, value-delivered, and financial dimensions, emphasizing explainability, adoption + retention c…
Business
plg-skill
Guides SaaS teams on product-led growth strategy: evaluating PLG readiness, selecting freemium or trial models, defining activation and PQLs, crafting self-serve onboarding, buildi…
Research
research-platform
Aggregates an operator's owned analytics, public metrics, and prior evaluations into a sourced evidence base for X, LinkedIn, TikTok, YouTube, and Instagram. Tags every data point …
Testing
shadow-mode-runner
Coordinates SHADOW mode operation where the agent runs parallel to human input without delivering output or incurring billing. Measures agreement rates and generates promotion repo…
Data
x-tweet-search-by-query
Executes advanced X/Twitter searches via query string and returns normalized results including tweet text, author details, engagement metrics, media, and pagination cursor. Support…
Research
experimentation
Designs and runs controlled experiments (A/B tests, RCTs, offline hypothesis tests) including hypothesis framing, randomization, power analysis, metric selection, variance reductio…
Engineering
performance-budgets
Define, track, and enforce performance thresholds across time, size, and count dimensions using metrics, thresholds, percentiles, and consequences. Covers Core Web Vitals, RAIL, Li…
Security
harness-score
Generates a 5-dimension harness readiness scorecard (harnessFit, compileConfidence, taskCoverage, toolSafety, memoryUsefulness) plus estimated cost and scaffold readiness from a gi…
AI / ML
pm-experimentation-ab
Designs and reviews A/B tests with statistical rigor: pre-registered hypotheses, power analysis, primary metrics with guardrails, segment reads, and clear iterate/pivot/persevere d…
Data
datajunction-query
Query DataJunction nodes, generate SQL, fetch metric data, explore lineage, and visualize results through APIs or compatible tools. Pair with datajunction for concepts, datajunctio…
Productivity
analyst-modes
Defines five analyst operating modes (Query, Dashboard, Document Compile, Extract, Challenge) with procedures, output formats, and metrics. Includes write-gates for dashboard and s…
Business
pm-okr-metric-validity-audit
Runs structured validity audits on OKRs, KPIs, North Star Metrics, or success-metric sets. Flags vanity metrics, unfalsifiable results, output-vs-outcome confusion, gameable target…
Business
professional-indemnity-profit-per
Captures near-misses, claim letters, insurance notifications, and lessons learned. Equips managing partners, management committees, and COO/CFO roles in German mid-sized law firms …
Business
equity-partner-modell
Analyzes equity structures, fixed-share models, salary partner tracks, counsel roles, and partner entry/exit scenarios for mid-sized German law firms, delivering metrics, decision …
DevOps
salesforce-agentforce-stdm-observer-skill
Monitors live Salesforce Agentforce sessions via STDM and Data Cloud, surfacing faithfulness scores, action telemetry, and quality metrics under least-privilege access. Answers rea…
DevOps
forge-observability
Production observability with OpenTelemetry covering traces, metrics, and logs with correlation. Includes SDK initialization, span conventions, error handling, RED/USE metrics, sam…
Productivity
github-dashboard
GitHub repository analytics dashboard — stars, forks, contributors, issues, pull requests, recent activity, and top contributors. Use when the brief asks for a GitHub repo dashboar…
Research
hypothesis
Write testable product hypotheses with clear success metrics, baselines, targets, and timeframes. Produces a structured statement, supporting evidence, validation plan with method …
Productivity
token-defaults-score
Runs a repeatable audit that turns every safe token-saving default on, locks it against regression, and tracks the rest as roadmap items—without shipping unwitnessed claims. Re-mea…
Testing
armstat
Lightweight per-arm turn-log analyzer reporting billable turns, HTTP status counts, cache metrics with 5m/1h split, token-weighted hit rate, and thinking budget. Safe on live NDJSO…
Business
role-redesign-for-ai
Redesigns a role after AI has shifted its workload — mapping pre/post task inventories, redefining core responsibilities, and updating metrics plus growth paths. Use when AI change…
AI / ML
dos-self-improve
Runs an isolated, kernel-driven self-improvement loop: proposes changes, verifies them in clean worktrees, and keeps only those where an independent witness confirms metric gains a…
Design
voice-agent-design
Design voice AI agents for phone or in-app use: conversation flows, interruption handling, escalation logic, and metrics that flag poor caller experience. Outputs specs covering pe…
Testing
eval-result-interpreter
Analyzes AI agent evaluation pass/fail results from CSV exports or custom evaluators using triage frameworks to deliver SHIP/ITRATE/BLOCK verdicts, root cause analysis, mode classi…
Data
metric-analyst
Use when the task involves defining, calculating, or implementing business metrics or KPIs. Covers KPI definition, SQL metric logic, Excel formulas, churn, retention, revenue, conv…
AI / ML
simplicio-autoresearch
Evolutionary optimize-by-metric loop that mutates targets, evaluates against fixed criteria, commits on improvement or reverts via git, and breaks plateaus after stagnation. Includ…
Security
pr-control-log
Analyzes completed PRs to audit AI behavior by extracting autonomous decisions, assumptions, and spec deviations from git logs, diffs, and PR bodies. Generates structured control l…
DevOps
stream-aggregation-helper
Design and tune VictoriaMetrics stream aggregation rules to cut cardinality, sampling rate, or query load. Handles vmagent config, metric aggregation choices, interval tuning, pipe…
DevOps
alloy
Configures Grafana Alloy for OpenTelemetry collection and telemetry pipelines. Supports the Alloy language, components for metrics/logs/traces/profiles, data export to Grafana Clou…
Data
metricflow_ingest
Maps MetricFlow semantic models and metrics into ktx semantic layer sources. Handles primitive tables, inheritance flattening, metric types, model refs, and provides worked example…
Productivity
pm-rate-team
Assesses individual contributions across a configurable window (default 2 weeks). Generates contributor metrics including merged PRs, review cycles, issue throughput, participation…
Data
data-domain
Enforces data engineering standards: warehouse access is read-only, queries follow sample-then-scale, every metric links to logged query and data hash, evaluations stay isolated fr…
Business
pm-okrs-kpis
Define or review OKRs, KPIs, and North Star Metrics from product strategy. Enforces outcome-focused key results with baselines, targets, and one objective per team per cycle, balan…
Data
formatting-and-highlighting
Provides formatting rules for metrics: decimals, prefixes/suffixes, currency, percentages, scaling (K/M/bp), separators, sign handling, and display modes including rich text, URLs,…
Content
hit-lab
把内容创作变成可校准的预测循环——打分 → 盲预测 → T+3d 复盘 → 进化 rubric,适用任何能被量化(播放/阅读/收听/点击)的内容;内置一份观点视频 rubric,其他形态可借此起步。触发词:"初始化"/"打分这篇"/"启动预测"/"已发布"/"复盘"/"升级 rubric"/"推荐选题"/"抓热点"/"挖问题"/"状态"/"找对标"/"lea…
AI / ML
build-eval
Designs and implements rigorous evaluations for LLM agents, multi-agent systems, skills, and prompts—covering frameworks like DeepEval, Braintrust, and RAGAS plus metrics such as p…
Business
promotion-case-builder
Builds a structured promotion case document including accomplishments, metrics, scope evidence, stakeholder impact, and compensation context, with a clear request for advancement. …
Data
flow-metrics
Computes DORA and Flow Framework metrics—cycle time, lead time, throughput, WIP, flow efficiency—from Jira changelogs, optionally joined with Jira Align for program/portfolio rollu…
Data
modeling-pigment-applications
Provides the mental model and decision framework for designing Pigment applications, covering core concepts like dimensions, metrics, transaction lists, sparsity, and scope, plus c…
DevOps
create-kit
Creates a new easy-db-lab kit (kit.yaml, K8s manifests, metrics, dashboards, docs) for any database or workload. Detects internal vs external mode, researches the workload, tracks …
DevOps
gitlab-portfolio
Aggregates cross-repo health signals from GitLab and GitHub projects registered in a vault. Scans frontmatter for repo metadata, fetches open issues/MRs and stale indicators via pa…
AI / ML
self-optimize
Analyzes system performance and failure patterns to autonomously improve skill prompts. Reads logs, identifies weak skills, rewrites prompts, commits changes, and rolls back if met…
AI / ML
state-estimator-evaluate-bags
Evaluate Moleworks ROS2 mole_estimator performance on MCAP/rosbag2 datasets by replaying sensor bags, recording reprocessed outputs with state and graph topics, then running offlin…
Data
libra
Query and manage Libra/DataTester A/B experiments including details, traffic allocation, app lists, parameter-path search, reports, metric groups, realtime dashboards, and test use…
Showing the top 60 of 1,355. See the full list →