PluginBench
Skill
Review
Audit score 70

reflect

cursor/plugins

Mine conversations for durable learnings and route them into skill edits.

What is reflect?

Reflect spawns three parallel review subagents to analyze the active transcript, surface learnings, and synthesize them into concrete edits on existing skills or new skill proposals. Use it when you want to capture and codify patterns from a conversation into your skill library.

  • Locates and reads the active transcript from the current workspace
  • Spawns three parallel reviewers (judgment, tooling, divergent lens) to analyze the conversation
  • Synthesizes findings into Accepted/Rejected/Backlog categories
  • Enforces structural checks to move lessons into lint rules or metadata when appropriate
  • Routes approved edits to existing skills or the create-skill workflow
  • Files backlog items to your devex tracker automatically

How to install reflect

npx skills add https://github.com/cursor/plugins --skill reflect
Prerequisites
  • Active workspace with agent-transcripts directory configured
  • Access to the create-skill workflow for substantive edits
  • Devex/backlog tracker for filing backlog items
  • MCP access for context lookups (tickets, traces, chat threads)
Claude Code
Cursor
Windsurf
Cline

How to use reflect

  1. 1.Invoke with 'reflect' or '/reflect' command during or after a conversation
  2. 2.Review the synthesizer's Accepted/Rejected/Backlog output when presented
  3. 3.Approve which items to apply and confirm any routing redirects
  4. 4.For trivial edits, the parent skill applies directly; for substantive changes, hand to create-skill
  5. 5.Verify touched skills with SKILL.md validator if available
  6. 6.Review the summary of applied edits, new skills, and backlog items

Use cases

Good for
  • After a complex debugging session, capture the pattern into a new troubleshooting skill
  • Following a tool integration, codify the setup steps and gotchas into an existing integration skill
  • When a conversation reveals a gap in agent behavior, propose a description tune or new skill
  • After mentoring a junior agent through a workflow, extract and share the best practices
  • When multiple conversations repeat the same problem, synthesize a solution into a reusable skill
Who it's for
  • Agent developers building and maintaining skill libraries
  • Teams standardizing patterns across multiple agents
  • Organizations capturing institutional knowledge from agent conversations
  • Skill authors iterating on existing skills based on real usage

reflect FAQ

When should I skip invoking reflect?

Skip when the conversation is trivial, off-topic, or already covered by an existing skill that the parent followed correctly. One-off solutions are not durable learnings.

How does reflect find the right transcript?

It lists transcripts in the active workspace's agent-transcripts/ directory, checks the first JSONL line for the opening user prompt, and takes the matching path. If no path resolves, it digests the session and passes that instead.

Why do reviewers need MCP access?

Reviewers need MCPs to look up context like tickets, chat threads, and observability traces that are referenced in the transcript. Readonly mode strips MCPs, so agent mode is required.

What happens to backlog items?

Backlog items are filed automatically to your team's devex/backlog tracker. Only Accepted items wait for user approval before being applied.

Can I redirect how an edit is routed?

Yes. After reviewing the synthesizer's output, you can pick which items to apply and redirect their routing (e.g., from a new skill to an edit of an existing one).

Full instructions (SKILL.md)

Source of truth, from cursor/plugins.


name: reflect description: Spawn three parallel review subagents over the active transcript, surface learnings, and route each to a concrete edit on an existing skill. Use when the user says reflect. disable-model-invocation: true

Reflect

Mine the current conversation for durable learnings, then route them into skill edits.

When to invoke

Invoke when the user says "reflect" or "/reflect". Skip when the conversation is trivial, off-topic, or already covered by an existing skill the parent followed correctly. One-offs are not learnings.

Process

1. Locate the active transcript

The parent finds its own transcript file before fanning out. The system prompt names the active workspace's agent-transcripts/ directory. Use that path. Do not glob across ~/.cursor/projects/*/. That crosses workspace boundaries and reads private chats from unrelated projects.

ls -t <agent-transcripts>/*.jsonl <agent-transcripts>/*/*.jsonl <agent-transcripts>/*/subagents/*.jsonl 2>/dev/null | head -10

Three transcript layouts: legacy flat (<id>.jsonl), current nested (<id>/<id>.jsonl), and subagent (<parent>/subagents/<child>.jsonl).

For each candidate, read the first JSONL line and check that message.content[0].text contains the conversation's opening user prompt. Take the matching path. If no path resolves, write a tight digest of the session and pass that instead.

2. Spawn three reviewers in parallel

One message, three Task calls, subagent_type: generalPurpose, with model set as below, agent mode (readonly: false). Reviewers need MCP access for context lookups (tickets, chat threads, observability traces referenced in the transcript). Readonly strips MCPs.

Each reviewer and the synthesizer name a role line in the pstack-models.mdc rule and a default. Set model to that line's value, or to the default if the rule or the line is missing. Leave model unset when the value is auto or inherit-parent. If the Task tool rejects a slug, use the default and say so. If it rejects the default, use the closest valid slug of the same family from its error message.

LensRole lineDefault modelPrompt template
Judgmentreflect judgment, divergent, synthesizerclaude-opus-5-5-maxreferences/judgment-reviewer.md
Toolingreflect toolinggpt-5.6-sol-maxreferences/tooling-reviewer.md
Divergentreflect judgment, divergent, synthesizerclaude-opus-5-5-maxreferences/divergent-reviewer.md

Pass each template verbatim, substituting the transcript path or digest where marked. Reviewers return findings in the Task response body.

3. Synthesize

One Task call, subagent_type: generalPurpose, with model from the reflect judgment, divergent, synthesizer line (default claude-opus-5-5-max), agent mode (readonly: false). The synthesizer's quality check includes spot-verifying citations, which can require MCP access. Readonly strips MCPs. Use references/synthesizer.md verbatim, with each reviewer's full output inlined where marked. The synthesizer returns a structured Accepted / Rejected / Backlog list.

4. Structural enforcement check

Sanity-check the synthesizer's Accepted list. For any item that would be enforced more reliably by a lint rule, script, metadata flag, or runtime check, move it from Accepted to Backlog. See the encode-lessons-in-structure principle skill.

5. Apply

Before applying any Accepted edit, present the synthesizer's full Accepted/Rejected/Backlog output to the user and wait for explicit approval. The user picks which subset to apply and may redirect routings. Skill changes affect every future agent in the org. Do not auto-apply.

Backlog items file to whatever devex / backlog tracker your team uses automatically. Only the Accepted list waits for approval.

For each approved Accepted item, follow the Routing field exactly:

  • Trivial existing-skill edit (a one-line bullet, a tightened sentence, a stale fact corrected): parent does directly.
  • Substantive existing-skill edit (a new section, a new pattern table, more than ~10 lines): hand to Cursor's built-in create-skill skill and run its draft / test / iterate loop.
  • tune description: <skill path> (the skill exists but didn't trigger when it should have): hand to create-skill and run its description-optimization loop.
  • new skill via create-skill: <kebab-name>: hand creation to create-skill. Do not invent the shape ad hoc.

If your environment ships a SKILL.md validator, run it on every touched skill before declaring done. Skip this step if it doesn't.

6. Summarize for the user

Short list, no preamble:

  • Edits applied: <skill path>. What changed, one line each.
  • New skills created: <skill path>. One line each (rare).
  • Backlog filed to the devex tracker: <issue title> (<tags>). One line each.
  • Dropped: one line per rejected finding + reason from the synthesizer.