PluginBench
Skill
Review
Audit score 70

prometheus-configuration

wshobson/agents

Set up Prometheus for comprehensive metric collection, storage, and monitoring of infrastructure and applications.

What is prometheus-configuration?

Prometheus Configuration provides guidance for setting up Prometheus to collect metrics, store time-series data, and implement alerting across infrastructure and applications. Use this when implementing metrics collection, setting up monitoring infrastructure, or configuring alerting systems.

  • Configure metric scraping from targets with appropriate intervals
  • Create recording rules for expensive queries and metric aggregation
  • Design alert rules for infrastructure and application monitoring
  • Implement service discovery for dynamic target management
  • Set up metric retention policies based on storage capacity
  • Use relabeling for metric cleanup and standardization

How to install prometheus-configuration

npx skills add https://github.com/wshobson/agents --skill prometheus-configuration
Prerequisites
  • Prometheus server installed and running
  • Understanding of metric naming conventions
  • Access to targets that expose metrics endpoints
  • Basic knowledge of PromQL (Prometheus Query Language)
Claude Code
Cursor
Windsurf
Cline

How to use prometheus-configuration

  1. 1.Define scrape targets in prometheus.yml with appropriate intervals (15-60s typical)
  2. 2.Configure relabeling rules to clean up and standardize metric names
  3. 3.Create recording rules for frequently-used or expensive queries
  4. 4.Design alert rules with clear thresholds and notification channels
  5. 5.Implement service discovery if managing dynamic infrastructure
  6. 6.Monitor Prometheus itself to ensure reliability
  7. 7.Test queries using the Prometheus API or web UI

Use cases

Good for
  • Setting up monitoring for microservices infrastructure
  • Configuring Prometheus to scrape metrics from multiple application instances
  • Creating recording rules to pre-compute expensive aggregations
  • Implementing alerting rules for SLO violations and system anomalies
  • Designing high-availability Prometheus deployments with federation
Who it's for
  • DevOps engineers
  • Site reliability engineers (SREs)
  • Infrastructure architects
  • Application developers implementing observability
  • Platform teams managing monitoring infrastructure

prometheus-configuration FAQ

What scrape interval should I use?

Typical scrape intervals range from 15-60 seconds depending on your needs; shorter intervals provide more granularity but increase storage and CPU usage.

How do I check if Prometheus is scraping targets correctly?

Use curl http://localhost:9090/api/v1/targets to view all targets and their scrape status.

When should I use recording rules?

Use recording rules to pre-compute expensive queries that run frequently, reducing query latency and Prometheus CPU usage.

How do I implement high availability for Prometheus?

Run multiple Prometheus instances scraping the same targets, then use federation or tools like Thanos/Cortex for deduplication and long-term storage.

What retention policy should I set?

Configure retention based on your storage capacity and query needs; typical values range from 15 days to several months.

Full instructions (SKILL.md)

Source of truth, from wshobson/agents.


name: prometheus-configuration description: Set up Prometheus for comprehensive metric collection, storage, and monitoring of infrastructure and applications. Use when implementing metrics collection, setting up monitoring infrastructure, or configuring alerting systems.

Prometheus Configuration

Complete guide to Prometheus setup, metric collection, scrape configuration, and recording rules.

Purpose

Configure Prometheus for comprehensive metric collection, alerting, and monitoring of infrastructure and applications.

When to Use

  • Set up Prometheus monitoring
  • Configure metric scraping
  • Create recording rules
  • Design alert rules
  • Implement service discovery

Detailed patterns and worked examples

Detailed pattern documentation lives in references/details.md. Read that file when the navigation tier above is insufficient.

Best Practices

  1. Use consistent naming for metrics (prefix_name_unit)
  2. Set appropriate scrape intervals (15-60s typical)
  3. Use recording rules for expensive queries
  4. Implement high availability (multiple Prometheus instances)
  5. Configure retention based on storage capacity
  6. Use relabeling for metric cleanup
  7. Monitor Prometheus itself
  8. Implement federation for large deployments
  9. Use Thanos/Cortex for long-term storage
  10. Document custom metrics

Troubleshooting

Check scrape targets:

curl http://localhost:9090/api/v1/targets

Check configuration:

curl http://localhost:9090/api/v1/status/config

Test query:

curl 'http://localhost:9090/api/v1/query?query=up'

Related Skills

  • grafana-dashboards - For visualization
  • slo-implementation - For SLO monitoring
  • distributed-tracing - For request tracing