PluginBench
Skill
Pass
Audit score 90

health-check

codewithmukesh/dotnet-claude-kit

Multi-dimensional health assessment for .NET projects with letter grades (A-F) using Roslyn MCP tools.

What is health-check?

Evaluates .NET projects across 8 dimensions (build health, code quality, architecture, test coverage, dead code, API surface, security, documentation) and produces a structured report card with GPA and actionable recommendations. Use this when onboarding to a project, before release, or during quarterly reviews to establish baselines and prioritize technical debt.

  • Grades 8 dimensions independently (Build Health, Code Quality, Architecture, Test Coverage, Dead Code, API Surface, Security Posture, Documentation) using Roslyn MCP tools and CLI commands
  • Converts grades to GPA (A=4.0 through F=0.0) and averages only assessed dimensions
  • Produces a structured report card with key findings, triage decisions, and priority recommendations
  • Includes effort estimates and named types/files for each recommendation
  • Applies exact rubric thresholds without curving grades by project size
  • Supports partial assessments (quick health checks) or full 8-dimension reviews

How to install health-check

npx skills add https://github.com/codewithmukesh/dotnet-claude-kit --skill health-check
Prerequisites
  • A .NET project with Roslyn MCP tools available
  • Access to `references/grading-rubric.md` for canonical rubrics and thresholds
  • Ability to run `dotnet build` and `dotnet list package` commands
Claude Code
Cursor
Windsurf
Cline

How to use health-check

  1. 1.Choose your assessment scope: full 8-dimension (onboarding/pre-release), quick health (1-4 dimensions), or targeted (post-refactor/post-update)
  2. 2.Read `references/grading-rubric.md` to understand grade thresholds for each dimension
  3. 3.Collect data per dimension using the specified MCP tools or CLI commands (build errors, antipatterns, test coverage, dead code, API surface, vulnerable packages, documentation)
  4. 4.Apply the triage gate: read summaries not violation lists, drop suppressed findings, set aside medium-priority items, grade only high-confidence findings, verify against project invariants
  5. 5.Grade each dimension using exact rubric thresholds and convert to GPA (A-F scale)
  6. 6.Generate the report card with grades table, key findings, overall GPA, and priority recommendations with effort estimates

Use cases

Good for
  • Onboarding to an unfamiliar project to establish a baseline health score
  • Pre-release quality gate or monthly/quarterly maintenance reviews
  • Measuring progress after a cleanup sprint (re-grade to show improvement)
  • Prioritizing technical debt by identifying lowest-grade dimensions
  • Quick mid-sprint health checkpoint before a demo or major merge
Who it's for
  • Engineering leads and architects reviewing project health
  • Teams onboarding to new codebases
  • Release managers conducting pre-release quality gates
  • Technical debt prioritization committees
  • Developers conducting quarterly codebase reviews

health-check FAQ

What if I cannot assess all 8 dimensions?

Mark unassessed dimensions as 'Not assessed' and exclude them from GPA calculation. Only average the dimensions you actually graded. A dimension marked 'Not assessed' is never scored as an F.

How do I handle suppressed findings or convention-filtered results?

Record the count and suppression configuration in the report so suppression stays visible, but do not count them toward the grade. Similarly, ignore `conventionFiltered` results in dead-code analysis.

Should I grade on a curve based on project size?

No. Apply rubric thresholds exactly regardless of project size. 15 warnings is a C whether the project is small or large; this prevents standards from eroding.

When should I use this instead of the code-review skill?

Use health-check for whole-project assessment and baseline grading. Use code-review (via the code-reviewer agent) for per-change review or deep dives into specific dimensions like Code Quality.

How do I show progress after a cleanup sprint?

Re-run the assessment on the same dimensions you graded before, compare the new grades to the previous report, and include a trend comparison table in the new report card.

Full instructions (SKILL.md)

Source of truth, from codewithmukesh/dotnet-claude-kit.


name: health-check description: > Multi-dimensional health assessment for .NET projects with letter grades (A-F) using Roslyn MCP tools. Evaluates 8 dimensions: build health, code quality, architecture, test coverage, dead code, API surface, security posture, and documentation. Produces a structured report card with actionable recommendations. Load this skill when: "health check", "how healthy is this", "project health", "code quality report", "grade this project", "assess codebase", "quality audit", "technical assessment", "codebase review", "report card".

/health-check — 8-Dimension Project Assessment

What

Runs a data-driven health assessment across 8 dimensions, each graded A-F with the specific data points that produced the grade, and rolls them into a GPA. Gut feeling is not a grade: every dimension uses MCP tools or CLI commands, and every grade below A comes with specific, prioritized, effort-estimated fixes — "add test classes for OrderService, PaymentProcessor, ShippingCalculator" is actionable; "improve test coverage" is not.

This skill owns the canonical grading system for the kit. The full rubrics, GPA scale, and report template live in references/grading-rubric.md — load that file when running an assessment.

Tone is diagnostic, not punitive: a C grade is an improvement path, not a failure.

When

  • Onboarding to an unfamiliar or new project — set the baseline
  • "How healthy is this?", "grade this project", "codebase review", "report card"
  • Pre-release quality gate, or monthly/quarterly maintenance review
  • After a cleanup sprint (/de-sloppify) — re-grade to show progress
  • Tech-debt prioritization — lowest grades get the next sprint's attention

How

Step 1: Choose Scope

ScenarioDimensions
Full assessment (onboarding, pre-release, monthly review)All 8
Quick health (mid-sprint checkpoint, before a demo, after a merge)1-4 only
After major refactor1 (Build), 3 (Architecture), 4 (Tests)
Post-dependency update1 (Build), 7 (Security)
After cleanup sprintRe-grade only the cleaned dimensions

Step 2: Run the Dimensions

Read references/grading-rubric.md for the grade thresholds, then collect data per dimension. For deep code-quality dimensions, delegate to the code-reviewer agent with the code-review skill.

#DimensionData source
1Build Healthdotnet build --no-restore — errors + warnings
2Code QualityMCP detect_antipatterns — read summary, grade high-confidence only
3ArchitectureMCP get_project_graph + detect_circular_dependencies (projects AND types)
4Test CoverageMCP get_test_coverage_map — check applicable first (structural, not line coverage)
5Dead CodeMCP find_dead_code(scope: "solution") — grade high-confidence; ignore conventionFiltered
6API SurfaceMCP get_public_api + find_references — overexposure, return-type consistency
7Security Posturedotnet list package --vulnerable --include-transitive + secrets/auth spot check (deep dive: /security-scan)
8DocumentationXML doc coverage on public APIs + README currency

Step 2.5: Triage Gate (before any grade is assigned)

Detector output is evidence, not a grade. Pass every finding through this gate first — it is what stops a noisy count becoming a wrong letter.

  1. Read summary, not the violation list. summary.byId is complete even when the list is truncated. Never sample a truncated list and extrapolate.
  2. Drop suppressed. Record the count and summary.suppressionConfig in the report so suppression stays visible.
  3. Set aside medium. These are review items, not grade inputs. Summarise them by category; do not fix or count them.
  4. Grade high only. These are wrong regardless of context.
  5. Check invariants. If a signal contradicts something the target repo's CLAUDE.md documents as deliberate, the invariant wins — verify before grading.
  6. Read before asserting. Never describe a finding you have not opened. If you cannot open all of them, report the ones you did and say so.

Fill in the triage table from references/grading-rubric.md as you go. If the table cannot be filled, the dimension is not ready to grade.

Step 3: Grade and Aggregate

Apply the rubric thresholds exactly — never grade on a curve ("pretty good for a project this size" is how standards erode; 15 warnings is a C regardless of project size). Convert to GPA (A=4.0 … F=0.0), averaging only the dimensions actually graded — a dimension marked "Not assessed" is excluded from the GPA, never scored as an F.

Step 4: Report

Produce the report card from the template in references/grading-rubric.md: grades table with key findings, overall GPA, and priority recommendations — each with named types/files, priority order, and effort estimates. If a previous report exists, append the trend comparison table.

Example

User: /health-check

Claude: Running full 8-dimension assessment...

| Dimension | Grade | Key Finding |
|-----------|-------|-------------|
| Build Health | A | 0 errors, 2 warnings |
| Code Quality | B | 3 high-confidence findings in 4.2K lines; 31 medium untriaged |
| Architecture | A | Clean direction, 0 cycles |
| Test Coverage | Not assessed | Integration-driven suite — structural metric invalid |
| Dead Code | B | 5 unused methods (79 convention-discovered, not counted) |
| API Surface | B | 2 overexposed service types |
| Security | A | 0 vulnerable packages |
| Documentation | D | 12/30 public APIs documented |

Overall GPA: 3.1 (B) — averaged over 7 graded dimensions.

Triage: 44 AP005 raw → all log-and-rethrow wrappers (medium); 2 AP004 real.

Priority: (1) `SystemSeeder` → `TimeProvider`, ~15 min; (2) XML docs on the 8
endpoint classes, ~1 day; (3) review the 44 catch blocks or suppress by path.

Related

  • references/grading-rubric.md — canonical rubrics, GPA scale, report template
  • /de-sloppify — cleanup pipeline for the issues a health check surfaces
  • /security-scan — deep 6-layer scan behind Dimension 7
  • /code-review — per-change review (this skill grades the whole project)
  • /verify — pass/fail pipeline for a change set, not a graded assessment