PluginBench
Skill
Pass
Audit score 90

ce-work

everyinc/compound-engineering-plugin

Execute a plan or concrete work prompt end-to-end with local verification.

What is ce-work?

Executes bounded implementation tasks from a plan document, specification, or concrete work request. Use this when you have a clear scope and need implementation plus local verification without the full shipping workflow; use ce-debug for open-ended bugs instead.

  • Executes implementation tasks from a plan, spec path, or work description
  • Performs local verification and records evidence for each completed task
  • Handles recovery and inspection of existing runs without re-executing completed work
  • Supports Return-to-Caller Mode for outer orchestrators to own final verification gates
  • Manages workspace setup, branch placement, and pre-work inventory before writing
  • Integrates with cross-model execution when needed via structured controller handoff

How to install ce-work

npx skills add https://github.com/everyinc/compound-engineering-plugin --skill ce-work
Prerequisites
  • A writable canonical checkout (established during Phase 1)
  • Either a plan path, work description, or run ID for recovery
  • References directory from the skill must be readable by the harness
Claude Code
Cursor
Windsurf
Cline

How to use ce-work

  1. 1.Provide a plan path, work description, or leave blank to use the latest plan
  2. 2.The skill reads input-triage.md to classify your request and handle recovery if needed
  3. 3.Phase 1 establishes workspace, resolves execution engine, and selects strategy
  4. 4.Phase 2 executes implementation tasks, records verification evidence, and commits changes
  5. 5.In standalone mode, Phase 3-4 completes code review and shipping; in Return-to-Caller Mode, the result is returned for the caller to own remaining gates

Use cases

Good for
  • Implementing a feature from a detailed plan document created by ce-plan
  • Executing a concrete build request with clear scope and acceptance criteria
  • Recovering or inspecting a previous implementation run by run ID
  • Handing off verified implementation to an outer orchestrator for review and shipping
  • Executing trivial changes (one or two files, no behavioral change) without full planning
Who it's for
  • Developers executing bounded implementation tasks from specifications
  • Orchestrator workflows that need implementation and verification separate from shipping
  • Teams using ce-plan to scope work before ce-work executes it
  • Engineers debugging with ce-debug for open-ended issues

ce-work FAQ

When should I use ce-work instead of ce-debug?

Use ce-work when you have a clear plan, specification, or bounded work request. Use ce-debug for open-ended bugs where the root cause and scope are not yet known.

What happens if a file is already dirty before ce-work starts?

In standalone mode, you are asked once whether to include or exclude that file. In Return-to-Caller Mode, the run returns blocked, naming the collision and recovery steps.

Does ce-work handle code review and shipping?

In standalone mode, yes—it reads shipping-workflow.md and requires an actual code-review receipt or authorized skip before shipping. In Return-to-Caller Mode, code review and shipping are owned by the invoking workflow.

Can ce-work resume a previous run?

Yes. Provide the run ID during Phase 0 input triage. Recovery never re-executes completed verification or reruns completed tasks.

What is Return-to-Caller Mode?

A mode where ce-work performs implementation and local verification only, then returns a structured result to the invoking orchestrator, which owns final verification gates and shipping.

Full instructions (SKILL.md)

Source of truth, from everyinc/compound-engineering-plugin.


name: ce-work description: Execute a plan or concrete work prompt end-to-end. Use when implementing from a plan document, a spec path, or a clear build request; use ce-debug for open-ended bugs. Use when an outer orchestrator needs implementation and local verification only, without the shipping tail. argument-hint: "[Plan path, work description, or recovery request with run id; blank uses latest] | [mode:return-to-caller [implementation_engine:<compact-json>] [implementation_run:<safe-id>] <plan path> for outer orchestrators]"

Work Execution Command

Outcome

  • Result: A fully implemented, locally verified change set from a plan, specification, or concrete work prompt.
  • Next consumer: In standalone use, the shipping workflow takes the verified change through review and delivery. In Return-to-Caller Mode, the invoking workflow receives the structured implementation and verification result and owns its remaining gates.
  • Done: Every in-scope task is complete, required verification evidence is recorded, relevant checks pass, and the run reaches either its owned shipping handoff (with a code-review receipt or explicit skip phrase — see Phase 3-4), a complete return result, or an explicit blocker.
  • Intent: Finish the requested feature without renegotiating the plan or transferring canonical integration authority. Workers receive bounded units; the host orchestrator inspects actual changes and owns authoritative verification and canonical commits.

Execution Workflow

Bundled references must be read, never approximated. Resolve each reference or script path named below from this skill's loaded SKILL.md directory, using the full skill path the harness supplied, and never glob the target repository to find a bundled file. Read each reference when you enter the phase it governs; a read made before that phase does not satisfy it, and a reference this file says to read again is read again at its step even when already in context. If the harness does not expose the skill directory, or a required file cannot be read, stop before the action it governs and report which file is missing. Do not reconstruct its rules from memory; report the missing reference instead of continuing natively.

Phase 0: Input Triage

Recovery activation comes first. Before classifying the input as a plan, a path, a blank, or a bare prompt, recognize requests to resume, inspect, reap, or clean up an existing run. Recovery never dispatches a new worker, selects a new route, discovers another plan, reruns completed verification, or enters either shipping path. If the run id is missing, ask for it; never guess one.

Before any other input decision, read references/input-triage.md. It decides source resolution, control tokens, recovery, read-only discovery, plan readiness, non-code routing, blank input, and bare-prompt sizing. Three rules from it hold here:

  • A bare prompt that is Trivial — one or two files, no behavioral change — skips the task list but still resolves its execution engine before writing. A purely mechanical diff also ships without a post-PR watch. When either is uncertain, take the fuller route.
  • A bare prompt that ce-plan already sized in this session is executed, not planned again. A decision the user would weigh is asked as a question, never as a route back to ce-plan or ce-brainstorm.
  • If that reference cannot be read, stop; never treat control tokens or a non-executable artifact as code work.

When triage selects Return-to-Caller Mode, read references/return-to-caller.md immediately and record that it governs how this run ends. If it cannot be read, stop before any mutation; do not fall back to standalone behavior.

Phase 1: Quick Start

  1. Establish the workspace. Before moving branches, editing, dispatching, or committing, read references/workspace-setup.md. It decides the writable checkout, plan clarification, branch placement, the pre-work inventory, already-dirty files, and task setup. Never write without a writable canonical checkout, and never write on the real default branch unless the user explicitly directed that in this session.

    Do not commit or publish anything the user did not offer. When a unit needs a file that was already dirty, standalone mode asks once whether to include or exclude that file. Return-to-Caller Mode neither asks nor edits it; it returns blocked, naming the collision and how to recover.

  2. Resolve the engine, then strategy. After bounded plan intake and task derivation, but before selecting a unit for execution, writing, dispatching, or committing, read references/execution-engines.md and complete its route selection. It applies with or without a typed binding; native execution is eligible only when that reference selects it or exhausts an allowed fallback. The engine choice never changes which reference governs how the run ends.

    If cross-model execution is selected, read references/cross-model-execution.md before any content or authority crosses to the other model. It defines controller initialization, the post-init engine lock, bounded egress, transactions, recovery, and receipts.

    Before choosing inline, serial, or parallel execution, and before dispatching any worker, read references/execution-strategy.md. It decides scheduling, isolation, the packet each worker receives, worker lifecycle, and integration. The host orchestrator keeps authoritative verification and makes the canonical commits.

Phase 2: Execute

Before the first implementation write, including on the Trivial route, read references/implementation-loop.md. It decides how evidence is chosen, verification, when to stop a unit, incremental commits, following existing patterns, continuous testing, where simplification stops, UI work, progress tracking, and settled decisions.

The commit rule from this file stays in force throughout: every implementation commit names only that unit's owned files. A bare git commit can absorb the user's pre-existing index, so it is forbidden.

Phase 3-4: Quality Check and Finishing Work

After the tasks and local verification are complete, standalone mode reads references/shipping-workflow.md before any quality check or delivery. It decides simplification, code-review receipts and fallbacks, leftover findings, final validation, and delivery.

Code-review completion gate (standalone only). Code review must actually happen before shipping. The run is not done, must not call a commit or shipping skill, and must not report that shipping is complete until the shipping reference has recorded either an actual completed ce-code-review receipt or one of its exact authorized skip states. Never substitute a mental self-review or findings already applied earlier. This rule does not apply in Return-to-Caller Mode.

Return-to-Caller Mode

Return-to-Caller Mode performs implementation and local verification only. It must not enter Phase 3-4 or run final simplification, code review, PR creation, CI watching, babysitting, or any other standalone shipping action; the caller owns those steps.

Immediately before emitting the result, read references/return-to-caller.md again. It alone defines the full return result, the check that evidence is complete, the route and model records, recovery semantics, and standalone_shipping_skipped: true. Do not build a complete result from this file.

If that required read fails after planning or implementation created state, preserve every changed file, commit, workspace, and controller record. Return the minimum blocked result from this file: status: blocked, plan_path, run_id when known, changed_state, blockers naming the missing reference, and recovery_path. Do not erase partial state, report success, or fall into the standalone shipping path.

Related skills

More from everyinc/compound-engineering-plugin and the wider catalog.

CEce-work-beta logo

ce-work-beta

everyinc/compound-engineering-plugin

Execute work plans with optional Codex delegation for systematic feature delivery.

1.9k installs
CEce-worktree logo

ce-worktree

everyinc/compound-engineering-plugin

Create isolated git worktrees for parallel development without disturbing your main checkout.

2.9k installsAudited
COcoding-tutor logo

coding-tutor

everyinc/compound-engineering-plugin

Personalized coding tutorials that evolve with your knowledge, using your actual codebase and spaced repetition.

2.5k installs
COcompound-docs logo

compound-docs

everyinc/compound-engineering-plugin

Capture solved problems as categorized documentation with YAML frontmatter for fast lookup

1.5k installsAudited
DHdhh-rails-style logo

dhh-rails-style

everyinc/compound-engineering-plugin

This skill should be used when writing Ruby and Rails code in DHH's distinctive 37signals style. It applies when writing Ruby code, Rails applications, creating models, controllers, or any Ruby file. Triggers on Ruby/Rails code generation, refactoring requests, code review, or when the user mentions DHH, 37signals, Basecamp, HEY, or Campfire style. Embodies REST purity, fat models, thin controllers, Current attributes, Hotwire patterns, and the "clarity over cleverness" philosophy.

709 installs
FRfrontend-design logo

frontend-design

everyinc/compound-engineering-plugin

Build web interfaces with genuine design quality, not AI slop. Use for any frontend work - landing pages, web apps, dashboards, admin panels, components, interactive experiences. Activates for both greenfield builds and modifications to existing applications. Detects existing design systems and respects them. Covers composition, typography, color, motion, and copy. Verifies results via screenshots before declaring done.

625 installsAudited