PluginBench
Skill
Pass
Audit score 90

build-loop-claude-code

buildgreatproducts/builder-os

Disciplined build→review→test→fix loop for Claude Code feature work.

What is build-loop-claude-code?

Automates quality-gated feature implementation by running code through a structured cycle: build from plan or prompt, review with `/review` and `/security-review`, test end-to-end, fix issues, and repeat until complete. Use when you need features built to a high standard with every increment verified before marking done.

  • Reads work from a plan file (with checkboxes) or a direct feature prompt and builds the first unchecked task
  • Runs Claude Code's `/review` command and `/security-review` for auth, payments, input, or data-access changes
  • Executes end-to-end testing including full test suite, user flow walkthrough, and error-state verification
  • Fixes all review findings and test failures in scope, re-running review until clean
  • Marks tasks complete only after passing review and testing; repeats until all plan tasks are checked off
  • Reports what was built, review findings fixed, verification method, and next steps

How to install build-loop-claude-code

npx skills add https://github.com/buildgreatproducts/builder-os --skill build-loop-claude-code
Prerequisites
  • Claude Code installed and available in your editor
  • A plan file in the repo (roadmap.md, plan.md, or similar with `- [ ]` checkboxes) or a clear feature prompt
  • Existing test suite that can be run to verify no regressions
Claude Code
Cursor
Windsurf
Cline

How to use build-loop-claude-code

  1. 1.Trigger the skill by saying 'run the build loop', 'build the next task', 'continue the plan', or 'build this feature properly'
  2. 2.If no plan exists, restate your feature request as a verifiable goal with 2–4 success criteria and confirm scope
  3. 3.The skill reads the first unchecked task from the plan (or your prompt) and implements it
  4. 4.Review runs automatically; you fix any findings, then re-run review until clean
  5. 5.Test end-to-end: run the test suite, walk the user flow including empty/loading/error states
  6. 6.Fix any test failures and loop back through review until passing
  7. 7.The skill marks the task complete, moves to the next unchecked task, and repeats until done
  8. 8.Review the final report showing what was built, findings fixed, and what needs your attention

Use cases

Good for
  • Building a new feature from a roadmap with multiple ordered tasks and checkboxes
  • Implementing a refactor plan where each step must be reviewed and tested before moving to the next
  • Adding a payment or auth feature where security review is mandatory before marking done
  • Fixing a bug where the fix must pass review, existing tests, and new test coverage before completion
  • Continuing work on a partially-built feature by resuming the next unchecked task in the plan
Who it's for
  • Teams using Claude Code who want automated quality gates on feature work
  • Developers building from a structured plan or roadmap with ordered tasks
  • Projects where every code change must pass review and testing before completion
  • Codebases with security-sensitive surfaces (auth, payments, user input, data access)

build-loop-claude-code FAQ

What if the plan has tasks that seem wrong or out of order?

Ask one specific question about the task rather than guessing. Don't relitigate plan decisions or skip ahead — tasks are ordered intentionally.

Do I need to write tests for new code?

Yes. The skill adds tests for new logic and runs the full test suite to ensure nothing that passed before is broken.

What counts as 'security-review' work?

Changes touching auth, payments, user input, or data access trigger `/security-review`. All findings in scope must be fixed before marking the task complete.

What if testing finds a bug in code I didn't touch?

Note pre-existing issues in the final report instead of fixing silently. Focus on fixing only the code you changed.

Can I use this without a plan file?

Yes. If no plan exists, state your feature as a verifiable goal with 2–4 success criteria, confirm scope, and the skill builds and verifies it the same way.

Full instructions (SKILL.md)

Source of truth, from buildgreatproducts/builder-os.


name: build-loop-claude-code description: Use when building features with Claude Code in any codebase and the work should go through a disciplined build → review → test → fix loop. Triggers on "run the build loop", "build the next task", "continue the plan", "build this feature properly", or any request to implement work from a plan file or a direct feature prompt. Builds from the plan (or the prompt if no plan exists), runs Claude Code's /review (plus /security-review for sensitive surfaces) and fixes every issue found, tests and verifies the feature end to end, fixes anything testing surfaces, and reports back once complete. Repeats until all plan tasks are checked off. license: MIT metadata: author: BuilderOS version: "1.0"

Claude Code Build Loop

Quality-gated feature work: nothing ships on "it compiles" — every increment is built, reviewed, tested end to end, and fixed before the user hears "done."

Source of work

  • A plan file exists (roadmap, refactor plan, or task list with - [ ] checkboxes — search the repo): work the first unchecked task. Tasks are ordered intentionally — never skip ahead. If the plan references spec docs, read only the sections relevant to the current task.
  • No plan (or the request is outside it): build from the user's prompt. Restate it as a verifiable goal with 2–4 success criteria and confirm scope in one message before building.

The loop

Run per task (or per prompted feature). Do not advance until every step passes.

  1. Build. Implement exactly what the task specifies. Simplest implementation that satisfies it, surgical changes, no speculative scope. Match existing project conventions.

  2. Review. Run /review on the changed code. If the change touches auth, payments, user input, or data access, also run /security-review. Fix all findings in scope — bugs, security issues, edge cases, performance, style in files you touched. If the project has a design system spec (design tokens file, DESIGN.md, theme config), check UI changes against it — no hardcoded colors, type, or spacing that bypass tokens. Note pre-existing issues in untouched code for the report instead of fixing silently. Re-run /review until clean. If a finding contradicts the task or spec, the spec wins — flag the disagreement.

  3. Test end to end. Run the task's verification step (or the success criteria). Run the full test suite — everything that passed before must still pass. Add tests for new logic. Then exercise the feature as a user would: run the app, walk the real flow including empty, loading, and error states.

  4. Fix. Anything testing finds goes back through the loop: fix → /review → re-test. Never mark a failing task complete; never start the next task with the app broken.

  5. Continue. Mark the task - [x], update any progress/status line in the plan, and loop to the next task until the requested scope is complete.

  6. Report. When done, tell the user: what was built and plan progress, review findings fixed and anything deferred, how it was verified (tests + flow walked), and what needs their attention next. Be honest about anything flaky or partially verified.

Rules

  • Skipped review or untested work = unfinished work.
  • Don't relitigate plan decisions; if a task seems wrong, ask one specific question rather than guessing.
  • Discovered work no task covers? Surface it and propose a task — never silently expand scope.