agentharnesses.dev

A field guide to agentic coding harnesses, for teams that have been told to adopt one.


Harnesses > Workflow methodologies

pstack

Draft. This entry has not been checked against its sources yet.

A Cursor plugin in which one command, /poteto-mode, routes each task to one of 23 playbooks that require reproducing problems first, verifying on the real surface, review by several models, and small stacked pull requests.

MaintainerLauren Tan (poteto), published in Cursor's plugin repository
CategoryWorkflow methodologies
HostsCursor
Components47 skills, 2 subagent definitions
LicenseMIT
Version0.15.5
Source codehttps://github.com/cursor/plugins/tree/main/pstack
Websitehttps://cursor.com/marketplace/cursor/pstack
Docshttps://github.com/cursor/plugins/tree/main/pstack/docs/guide
DeepWikihttps://deepwiki.com/cursor/plugins
Last researched2026-09-28, at commit ecc249f

Features

Definitions are on the compare page.

FeatureSupportedNotes
Stages of work it covers
BrainstormingpartialNo brainstorming skill, and the rules discourage asking what the agent can check itself. figure-it-out presents framing and tradeoffs before a long run.
Written planspartialOnly the multi-phase-plan playbook writes a plan, to Cursor's agent store by default rather than the repository. README: 'the best spec is code.'
Test firstpartialtdd runs only when asked or when a cheap test exists. Bug fixes must commit the failing repro before the fix.
Debugging processyesbug-fix playbook: reproduce, binary-search hypotheses against runtime evidence, revert what a refuted hypothesis motivated.
Code reviewyesinterrogate runs one read-only reviewer per configured model, three by default. Shipping needs a verdict from an agent that did not write the code.
Verificationyesprinciple-prove-it-works: verify against the real artifact. Replies must label every claim as measured, inferred, or a guess.
Learning capturepartial/reflect turns a transcript into proposed skill edits, applied only after you approve. Nothing is loaded at session start.
Branch to PRyesopening-a-pr, babysit, and shipping: worktree off main, small ordered commits, narrow PRs, independent verification before squash-merge.
How it runs
Single entry pointyes/poteto-mode matches the task against 23 playbook descriptions and copies the chosen playbook's steps into the todo list.
Automatic activationno46 of 47 skills are marked disable-model-invocation. You start with /poteto-mode, which then stays on across turns.
Session-start hooknoNo hooks. /setup-pstack writes an always-applied rule that holds model choices only.
Subagentsyes
Parallel agentsyesParallel explorers, reviewers, and design candidates; some playbooks run one Cursor cloud agent per pull request.
Git worktreesyesWork happens in a worktree off main; parallel design candidates each get their own.
Adopting it
Multiple hostsnoCursor only. Community ports exist for Claude Code, Codex, OpenCode, Gemini CLI, and Pi.
Per-project setupyesEnable in the committed .cursor/settings.json. The model rule is always per user.
Team extensionspartial/automate-me writes your own <handle>-mode skill to use alongside it. Adding playbooks to poteto-mode means forking.
No extra servicesyesSome steps use /deslop and control skills from the separate cursor-team-kit plugin, and some playbooks use Cursor cloud agents.

How it works

The workflow

You start with /poteto-mode <goal>. The skill reads its own index of 23 principles, matches the task against 23 one-line playbook descriptions, and opens a todo list whose first items are the chosen playbook's steps "copied in verbatim". A step the agent decides to skip stays in the list with a one-line reason. Large or cross-cutting work, or work you plan to leave running, goes to figure-it-out, which writes a one-off playbook for the task. Once entered, poteto-mode stays on for later turns until you opt out.

The playbooks cover investigation, bug fixes, performance problems, hill-climbing on a metric, runtime and trace forensics, features, refactoring, prototypes, visual parity, authoring skills, evals, babysitting and shipping pull requests, autonomous runs, orchestrating multi-day programs, pausing and resuming, multi-phase plans, worktree cleanup, and opening a PR. Every code playbook ends with Opening a PR.

Two examples:

The author does not plan by default: "i don't believe in planning. the best spec is code." A plan document appears only in the multi-phase-plan playbook, and a script (check-plan.mjs) checks its structure.

Components

How it steers the agent

All steering is text or tooling; there are no hooks.

Files it writes

There is no standing memory file. Lessons go into skill edits through /reflect, and only with your approval.

Install

In Cursor, run /add-plugin pstack, then /setup-pstack to pick models per role and a reasoning budget, then start a new chat. For the full set of steps, also install cursor-team-kit, which provides /deslop, control-ui, and control-cli; pstack refers to them but does not include them. To enable it for a whole team, commit {"plugins": {"pstack": {"enabled": true}}} in .cursor/settings.json.

pstack only runs on Cursor. It depends on Cursor-specific tools: subagent options such as cloud execution, /loop, /goal, .mdc rules, and Cursor's transcript files. Community ports, none maintained by the author, include pstack-claude (Claude Code, Codex, OpenCode, Gemini), pstack-pi (Pi), and pstack-generic.

Scripts need git, gh, bun, and node. worktree-audit.sh assumes macOS versions of stat and date.

When to choose it

When not to

Sources