773 free skills
Observability skills
Skills for observability — structured logging, distributed tracing, metrics dashboards, alerting, and incident response runbooks.
Sourced from real, public repositories — synced daily, never invented.
773 free skills
Skills for observability — structured logging, distributed tracing, metrics dashboards, alerting, and incident response runbooks.
Sourced from real, public repositories — synced daily, never invented.
15 tools across six categories
13 of them never send your data anywhere
Free · No signup · No trial clock
SEE THE DIRECTORY

salt-logmanager
taupikpirdian/skills
Structured logging with trace-ID propagation, data masking, and segment timing for Go apps using the SALT-Indonesia/salt-pkg logmanager library — plus middleware for Gin, Echo, Gorilla Mux, gRPC, RabbitMQ, and Resty. Use when adding request logging, distributed tracing, or sensitive-data masking with this library.
structured-log
codealive-ai/ceo-ai-os
Write structured entries to daily memory log with facts, concepts, and files — use after any meaningful work, research, or decision
distributed-tracing-setup
prompt-or-die-labs/hyper-forge
Configure distributed tracing with Jaeger, Zipkin, or Datadog for microservices observability
distribute-outreach
yoyothesheep/claude-skills
Backlink and citation outreach for a newly published post. Finds Substack writers, journalists, and bloggers likely to reference the post's key claim, generates personalized pitches, and maintains a running outreach contacts file. Use after a post is published and socially distributed.
distribute-social
yoyothesheep/claude-skills
Social distribution for a newly published post. Drafts Reddit comments/posts and a LinkedIn post in one run. Invokes reddit-content and linkedin-content as sub-skills. Use when a post is published and ready to distribute to social channels.
security-incident-response
kentoshimizu/sw-agent-skills
Security incident workflow for triage, containment, eradication, and recovery evidence handling. Use when suspected or confirmed security incidents require coordinated response actions; do not use for proactive threat modeling or routine vulnerability backlog grooming.
create-runbook
jerelvelarde/chalk-skills
Create an operational runbook when the user asks to document a procedure, write a runbook, create an ops guide, or document how to handle a specific operational task
runbook-authoring
kentoshimizu/sw-agent-skills
Runbook authoring workflow for clear, executable operational procedures for on-call responders. Use when incident response, rollback, failover, or maintenance steps must be documented before production use; do not use for production feature logic design.
detect-anomalies
axiomhq/cli
Detect anomalies in Axiom datasets using statistical analysis. Use when looking for unusual patterns, volume spikes, outliers, or new error types in observability data.
deepgram-js-voice-agent
deepgram/deepgram-js-sdk
Use when writing or reviewing JavaScript/TypeScript in this repo that builds an interactive voice agent via `agent.deepgram.com/v1/agent/converse`. Covers `client.agent.v1.createConnection()` / `connect()`, `sendSettings`, `sendMedia`, runtime updates, event handling, and function-call responses. Use `deepgram-js-text-to-speech` for one-way synthesis, `deepgram-js-speech-to-text` or `deepgram-js-conversational-stt` for transcription only, and `deepgram-js-management-api` for project/model admin rather than live agent runtime. Triggers include "voice agent", "agent converse", "full duplex", "barge-in", "function calling", and "agent.v1".
mlflow-evaluation
databricks-solutions/ai-dev-kit
MLflow 3 GenAI agent evaluation. Use when writing mlflow.genai.evaluate() code, creating @scorer functions, using built-in scorers (Guidelines, Correctness, Safety, RetrievalGroundedness), building eval datasets from traces, or setting up instrumenting apps for tracing, trace ingestion and production monitoring.
producthunt-launches
browser-act/skills
Scrape Product Hunt daily/weekly/monthly/yearly leaderboard launches with full product details, maker profiles, and website contact info. Use when user mentions Product Hunt, producthunt, PH scraper, product hunt launches, product hunt leaderboard, scrape product hunt, product hunt data, PH daily launches, product hunt upvotes, product hunt maker info, extract product hunt, product hunt today, top products product hunt, product hunt archive, PH products, product hunt email extraction, product hunt contact info, producthunt.com scraping, get product hunt launches, product hunt API alternative. Also applies to: startup launch monitoring, new product discovery, maker/founder contact enrichment, product hunt lead generation, daily product hunt digest, competitive product tracking.
scholar-monitor
joshzyj/open-scholar-skill
Current-awareness literature monitoring for social sciences. Tracks new publications from top journals (ASR, AJS, Demography, Social Forces, NHB, NCS, Science Advances, APSR, PNAS), preprint servers (arXiv cs.CL for LLM research, cs.CY for computational social science, econ.GN), and custom author/keyword watchlists. Delta-based and /loop-safe: each invocation fetches only papers published since the last tick, groups them by user-defined category, generates 2–3 sentence LLM summaries of title+abstract, and auto-ingests into the scholar-knowledge graph. Delivers digests to the user's phone via Telegram (using the existing MCP channel) or ntfy.sh push, optionally email over SMTP — with a file-based audit trail always written to output/monitor/. Modes: default fetch (all due enabled sources), targeted fetch (source_id argument), all (force-fetch all enabled), preview (dry-run — show what would fetch without network calls), init, list, status, add, remove, configure delivery, digest. Designed for /loop (e.g., '/loop 24h /scholar-monitor arxiv-llm' or '/loop 7d /scholar-monitor'). Fully idempotent: cadence_days filter drops redundant ticks. State persisted in ~/.claude/scholar-monitor/ (configurable via SCHOLAR_MONITOR_DIR). Works alongside scholar-lit-review (retrospective reviews) and scholar-knowledge (knowledge graph accumulation).
x-system
x-cmd/skill
This skill provides comprehensive system administration and monitoring tools through x-cmd CLI, including process management, macOS system utilities, network configuration, disk health monitoring, and storage analysis. This skill should be used when users need to perform system administration tasks, monitor system performance, manage network configurations, or troubleshoot system issues from command line interfaces.
create-structured-logger
dykyi-roman/awesome-claude-code
Generates Structured Logger for PHP 8.4. Creates PSR-3 structured logging setup with Monolog processors, correlation ID propagation, and context middleware. Includes unit tests.
trace
arielbk/trace
Bind the current session to a Trace task, and re-enter prior task context. Use when the user names a specific piece of work they are starting, resuming, or continuing — including "scope out X", "define a new X", "plan the work for X", "build X", "I need to build X from scratch", "tackle X", "add X to Y", "work on X", or "get back to X". Covers features, bugs, and refactors. Trace fires FIRST whenever the user commits to a new, named piece of work — whether or not they say "feature" or "task", and even when they also want to brainstorm, scope, or plan it — so the session is bound before any planning or creative skill runs. Does NOT apply when actively debugging or investigating a running system ("this endpoint is throwing errors", "help me debug X"), nor when the user is explicitly deferring the build to explore options first ("before we build anything, let's explore the design space") — neither is a committed new piece of work. Also use when explicitly asked to bind a session to a task, or when session-start context reports no active task during real project work.
distribution-ops
accolver/skill-maker
Designs channel-specific go-to-market plans, asset checklists, sequencing, and kill criteria for a chosen offer or product. Use when the opportunity and offer are already chosen and the next question is how to reach buyers through the right channels instead of generic marketing advice.
runbook
pablof7z/skills
Learn durable procedures from repeated delegation. Use when an agent should recognize a recurring request, recover how this user expects it done, execute from a lightweight runbook, or turn a successful new task into reusable procedural memory without making the user design the process.
deployment-runbook
pfangueiro/claude-code-agents
Deployment procedures, health checks, and rollback strategies. Use this skill when deploying applications, performing health checks, managing releases, or handling deployment failures. Provides systematic deployment workflows, verification scripts, and troubleshooting guides. Complements the devops-automation agent.
miso
curtisnewbie/miso
Use the miso Go framework for backend microservices. Use when working with the miso framework (https://github.com/curtisnewbie/miso) for: (1) Creating new microservices with component-based architecture, (2) Implementing RESTful APIs with Gin integration, (3) Database operations with GORM, (4) Configuration management with Viper, (5) Error handling with structured MisoErr types, (6) Distributed tracing via Rail context, (7) Distributed tasks with cron scheduling, (8) Bootstrap lifecycle management, (9) Service discovery and middleware integration, (10) Health checks and monitoring, (11) Performance profiling with pprof/FlightRecorder, (12) Request validation, (13) Caching strategies, (14) Kafka messaging, (15) Utility middleware (crypto, JWT, expr, Lua, money, ZooKeeper), (16) Util packages (atom.Time, json, strutil, slutil, retry, randutil, async, osutil, testutil, flags, excel, copyutil), (17) HTTP reverse proxy (HttpProxy) with filter pipelines, config-driven dynamic access filters (whitelist, bearer/basic/remote auth), and WebSocket access control
sentry-setup-ai-monitoring
getsentry/sentry-for-cursor
Setup Sentry AI Agent Monitoring in any project. Use this when asked to add AI monitoring, track LLM calls, monitor AI agents, or instrument OpenAI/Anthropic/Vercel AI/LangChain/Google GenAI. Automatically detects installed AI SDKs and configures the appropriate Sentry integration.
applicationinsights-web-ts
microsoft/skills
Instrument browser/web apps with the Application Insights JavaScript SDK (@microsoft/applicationinsights-web). Use for Real User Monitoring (RUM) — page views, clicks, AJAX/fetch dependencies, exceptions, custom events, and browser-side GenAI agent traces correlated to backend OpenTelemetry traces. Covers SDK Loader Script and npm setup, framework extensions (React, React Native, Angular), Click Analytics, telemetry initializers, and OTel GenAI semantic conventions for agent/tool/model spans emitted from the browser.
sentry-setup-logging
getsentry/sentry-for-cursor
Setup Sentry Logging in any project. Use this when asked to add Sentry logs, enable structured logging, setup console log capture, or integrate logging with Sentry. Supports JavaScript, TypeScript, Python, Ruby, React, Next.js, and other frameworks.
create-docker-compose-production
dykyi-roman/awesome-claude-code
Generates Docker Compose production configurations for PHP projects. Creates hardened stacks with resource limits, restart policies, and monitoring.
data-quality-checks
mohitagw15856/pm-claude-skills
Design the data quality checks for a table or pipeline across the standard dimensions. Use when asked to add data quality tests, define DQ checks, catch bad data before it hits dashboards, or set up monitoring for a dataset. Produces a checks plan across completeness, validity, uniqueness, freshness, consistency, and accuracy — each with the rule, severity, and where it runs (dbt test / Great Expectations / SQL assertion).
incident-response-lifecycle
vahagn-madatyan/netsec-skills-suite
>-
monitoring-alerting
johnqtcg/awesome-skills
>
sentry-setup-tracing
getsentry/sentry-for-claude
Setup Sentry Tracing (Performance Monitoring) in any project. Use this when asked to add performance monitoring, enable tracing, track transactions/spans, or instrument application performance. Supports JavaScript, TypeScript, Python, Ruby, React, Next.js, and Node.js.
sentry-setup-logging
getsentry/sentry-for-claude
Setup Sentry Logging in any project. Use this when asked to add Sentry logs, enable structured logging, setup console log capture, or integrate logging with Sentry. Supports JavaScript, TypeScript, Python, Ruby, React, Next.js, and other frameworks.
tools-unity-sentry
tjboudreaux/cc-plugin-unity-gamedev
Sentry Unity SDK integration patterns for error tracking, performance monitoring, transactions, spans, and custom instrumentation.
observability
incidentfox/incidentfox
Log, metric, and trace analysis methodology. Use when analyzing logs, investigating errors, querying metrics, or correlating signals across observability backends (Coralogix, Datadog, CloudWatch).
rabbitmq-management
delorenj/skills
Expert guidance for RabbitMQ message broker management, including MQTT, streams, virtual hosts, monitoring, troubleshooting, and CLI administration
dx-optimizer
belokonm/claude-supercode-skills
Expert in optimizing the end-to-end developer journey. Specializes in Internal Developer Portals (IDP), DORA metrics, and on-call health. Use when improving developer experience, building internal platforms, measuring engineering productivity, or reducing developer friction.
vector
terminalskills/skills
Expert guidance for Vector, the high-performance observability data pipeline built in Rust by Datadog. Helps developers collect, transform, and route logs, metrics, and traces from any source to any destination with minimal resource usage. Vector replaces Logstash, Fluentd, and Filebeat with a single, faster tool.
my-incident-response
itsnavee/claude-forge
Use when there's a production incident or outage — structured triage, investigation, fix, verification, and postmortem. Also use for "site is down", "production issue", "incident", or "outage".
dd-monitors
datadog/pup
Monitor management - create, update, mute, and alerting best practices.
structured-logging
kmshihab7878/claude-code-setup
Structured logging patterns with loguru and structlog for Python services, JSON output, correlation IDs, and Loki/Grafana integration
structured-logging
brpaz/agent-skills
Design language-agnostic structured logs with consistent fields, correlation IDs, safe context, and production-ready observability guidance.
logging
1mangesh1/dev-skills-collection
Logging setup, structured logging, and log management. Use when user asks to "add logging", "set up structured logging", "configure log levels", "create a logger", "set up log rotation", "send logs to ELK", "configure Winston", "set up Pino", "add request logging", "implement audit logging", "log formatting", "log correlation", "debug logging", "log sampling", "log filtering", or mentions logging best practices, structured logging, log aggregation, log levels, observability, log rotation, or centralized logging.
aimx
blizhan/aimx
Use when autoresearch, log_experiment, experiment analysis, or automatic iteration workflows need to inspect local Aim repositories with aimx, collect run params, metric summaries, traces, image evidence, compare training runs, or summarize model results without mutating Aim data.
incident-response
ffsshhttiikk/opencode-agents-skills
Security incident handling procedures
runbook-writer
00prabalk00/claude-skills
Create operational runbooks from tribal knowledge, commands, logs, and recovery steps. Use when recurring incidents lack clean response docs.
observability
iii-hq/skills
>-
dependency-update
microsoft/aspire
Guides dependency version updates by checking nuget.org for latest versions, triggering the dotnet-migrate-package Azure DevOps pipeline, and monitoring runs. Use this when asked to update external NuGet dependencies.
AgentObservability
majiayu000/claude-skill-registry
|
observability-service-health
elastic/cursor-plugins
>
gcp-monitoring
tomz/agent-skills
GCP observability — Cloud Monitoring, Cloud Logging, Cloud Trace, Error Reporting, Profiler
portfolio-health-prioritization
fourteenwm/ppc-ai-skills
Account prioritization criteria for portfolio health monitoring and daily briefings. Auto-invoke when determining which accounts need investigation, prioritizing daily actions, running portfolio health checks, or deciding investigation order. Provides 5-tier priority system with portfolio-specific thresholds (Portfolio A ±5%, Portfolio B ±8%).
documentation-bc-architecture-generator
fernandoartalf/al-copilot-skills-collection
Create well-structured Business Central architecture documents in the `openspec/architecture/` folder as the long-lived companion to an existing approved technical spec, following the ARCH-001 template structure. Captures business and technical context, architectural drivers traced back to user-story ACs, numbered Architecture Decision Records (ADR-N) with Context/Decision/Consequence/Alternatives-rejected, a logical component view (ASCII or mermaid), key runtime flows, data architecture (entity-relationship + storage/ownership/key-derivation tables), cross-cutting concerns (security, reliability, performance, observability, localisation, RapidStart/portability, upgrade/migration), coexistence with legacy code, and known constraints/limitations. Writes the file as `ARCH-NNN-<kebab-title>.architecture.md` with frontmatter linking back to its spec, user story, and CCN. Use whenever the user asks to create, draft, generate, or write an architecture document, ARCH document, ADR, design rationale, or technical companion document for a spec, with trigger phrases such as "create architecture doc for SPEC-NNN", "draft ARCH document", "write ADRs for [feature]", "generate architecture companion to the spec", "document the why behind SPEC-NNN", or "create an ARCH file in openspec/architecture".
bc-telemetry-generator
fernandoartalf/al-copilot-skills-collection
Instruments Business Central AL codeunits with Application Insights telemetry using the System Application Telemetry codeunit. Analyzes an existing codeunit supplied by the user, generates a dedicated Feature Usage Telemetry codeunit (or extends an existing one) with helper procedures for LogStart/LogEnd with automatic duration, LogError with call stack, LogFeatureUsage, and LogPerformanceWarning. Adds Telemetry.LogMessage calls with consistent event IDs (PREFIX-NNN), meaningful custom dimensions (Dictionary of [Text, Text]), proper Verbosity levels (Normal, Warning, Error, Critical), correct DataClassification (SystemMetadata, CustomerContent, EndUserIdentifiableInformation), and TelemetryScope::ExtensionPublisher. Covers five telemetry categories — lifecycle events (started/completed), error tracking (validation/posting failures), performance monitoring (slow operation warnings with threshold comparison), feature usage analytics (discount types, payment methods, order size distribution), and contextual telemetry (environment, version, anonymized user). Generates KQL queries for Application Insights analysis. Use when adding telemetry to codeunits, instrumenting AL code with Application Insights, tracking feature usage, monitoring performance, logging errors with custom dimensions, creating telemetry helper codeunits, or implementing observability for BC extensions.
infrastructure-diagrams
kroegha/infrastructure-diagrams
Create professional Azure, hybrid, and on-premises infrastructure architecture diagrams using Python's Diagrams library. Use when asked to create architecture diagrams, infrastructure diagrams, cloud diagrams, network diagrams, system architecture visualizations, or data center layouts. Supports Azure (VMs, networking, storage, databases, containers, security), on-premises (servers, databases, networking equipment, monitoring), Kubernetes, and hybrid cloud scenarios. Outputs PNG, SVG, or PDF files.
structured-logging
fulgidus/garland
Apply structured logging conventions. Use when writing, reviewing, or configuring any log statement in application code. Do NOT trigger for test code, CLI tools printing user-facing output, or scripts where stdout is the product.
logging
bigpapicb/universal-claude-skills
Structured logging patterns, log levels, trace IDs, and what to log vs what never to log. Use when setting up logging, choosing log levels, implementing request tracing, or reviewing log output quality.
structured-log
codealive-ai/ai-cofounder
Write structured entries to daily memory log with facts, concepts, and files — use after any meaningful work, research, or decision
distributed-tracing
camoneart/claude-code
Implement distributed tracing with Jaeger and Tempo to track requests across microservices and identify performance bottlenecks. Use when debugging microservices, analyzing request flows, or implementing observability for distributed systems.
incident-response
njones17/ai-agent-master-cyber-skills-list
Handle security incidents with IR playbooks and procedures. Implement detection, containment, eradication, and recovery processes. Use when responding to security events or building incident response capabilities.
incident-response
kienbui1995/magic-powers
Use when handling production incidents - outage triage, root cause analysis, communication, postmortem writing
loom-prometheus
cosmix/claude-code-setup
Prometheus monitoring and alerting for cloud-native observability. Use for writing PromQL queries, configuring scrape targets, creating alerting and recording rules, instrumenting applications, and setting up service discovery. Not for dashboards (use loom-grafana) or log analysis (use loom-logging-observability).
alert-manager
modsetter/surfsense
Configure SEO alerts for ranking drops, traffic changes, technical issues, competitor movements. SEO预警/排名监控
azure-observability
shoppinh/my-skills-101
Azure Observability Services including Azure Monitor, Application Insights, Log Analytics, Alerts, and Workbooks. Provides metrics, APM, distributed tracing, KQL queries, and interactive reports. USE FOR: Azure Monitor, Application Insights, Log Analytics, Alerts, Workbooks, metrics, APM, distributed tracing, KQL queries, interactive reports, observability, monitoring dashboards. DO NOT USE FOR: instrumenting apps with App Insights SDK (use appinsights-instrumentation), querying Kusto/ADX clusters (use azure-kusto), cost analysis (use azure-cost-optimization).