mrkeyoor.com_
Sat 03 Oct 03:10 UTC
Dev Toolsevaluationupdated 03 Oct 2026

pstack-claude review

pstack is a port of Lauren Tan's Cursor workflow stack for Claude Code, Codex, Pi, and several other coding-agent harnesses. It routes substantial coding tasks through named skills for investigation, architecture, implementation, review, and verification, so the agent follows a repeatable process instead of improvising one each time.

Verdict

We did not execute pstack-claude because our harness had no supported JavaScript ecosystem path and the repo had no Dockerfile, so this verdict rests on inspected documentation and repository activity rather than a lab pass. Its 31 public skills make sense for Claude Code, Codex, or Pi users who want a strong process imposed on substantial tasks. Skip it if automatic routing, parallel model cost, or uneven cross-runtime verification is more process than you want.

We ran it

Screenshot of pstack-claude (github.com/michael-denyer/pstack-claude)

Answers from our run

Did you run pstack-claude yourself?

No. Its code is JavaScript, and it carries no manifest our lab installs from, and no Dockerfile, so there was nothing standard to install, build or test. This review is written from the repository's own documentation.

Who should not use pstack-claude?

Developers who want a quiet, minimal agent setup: the Claude Code and Codex plugins install a startup routing hook, and Codex requires you to trust it.

What are the alternatives to pstack-claude?

Cursor pstack, Superpowers, Spec Kit. We did not execute pstack-claude because our harness had no supported JavaScript ecosystem path and the repo had no Dockerfile, so this verdict rests on inspected documentation and repository activity rather than a lab pass.

Setup3/5Plugin install is short, but hooks and optional tools add decisions
Docs5/5Runtime limits, dependencies, sync rules, and checks are explicit
Community3/5775 stars, two open issues and PRs, and an October 3 push
Maturity3/5Active v0.9 port with dedicated paths for three main runtimes

Who it’s for

Claude Code, Codex, or Pi users who want an opinionated workflow around non-trivial coding tasks.
Teams that want bug fixes to begin with reproduction and end with passing evidence.
Developers willing to configure role-specific models and supervise parallel agent work.
Maintainers who use GitHub CLI and want skills for CI repair, PR review, and merge readiness.

Who it’s NOT for

Developers who want a quiet, minimal agent setup: the Claude Code and Codex plugins install a startup routing hook, and Codex requires you to trust it.
Teams expecting identical behavior on every advertised harness: the reference says delegation and multi-model workflows remain unverified on Prime Agent, opencode, and Gemini CLI.
Privacy-sensitive work that cannot be sent to a model provider: pstack has no server or telemetry, but its README says material read by the skills, including transcripts, goes to that provider.
Users trying to minimize model spend: arena, swarm, and review skills can dispatch several agents, and the Pi guide warns that each Claude subagent adds token charges.
Teams that want upstream pstack unchanged: this port tracks upstream while keeping declared policy forks of its own.

Setup reality

We did not run commit 92debb7 in our sandbox. The harness reported no supported ecosystem for its JavaScript classification, and the repository had no Dockerfile, so there are no install, build, or test results from us to report.

The documented Claude Code and Codex route is a marketplace plus plugin install. Codex asks you to trust the startup hook. Pi uses pi install, while shared-skill installs require a clone and links under ~/.agents/skills. Some workflows also need GitHub CLI, Bun, Graphite CLI, or the plugin-dev companion.

The project documents its own Bun-based generator and test commands, but those are maintainer instructions, not results from our lab. Runtime support varies: Claude Code, Codex, and Pi have dedicated integration work, while several other harnesses rely on shared skill discovery and adaptation.

Thirty-one public skills turn one prompt into a coding process

pstack begins with poteto-mode, a router for work that needs more than a direct edit. A bug playbook reproduces the failure, investigates how and why it happens, delegates the fix, and reruns the failing case. If the change crosses a function boundary, the workflow brings in an architecture pass first. The output is supposed to include the change plus evidence from the failing and passing states.

The repository documents 54 skill directories: 31 public skills and 23 principle references. The public set covers explanation, architecture, test-first fixes, parallel attempts, adversarial review, CI repair, merge conflicts, PR maintenance, and technical writing. That range is pstack's appeal. You do not have to remember a separate prompt for each stage. The cost is accepting somebody else's view of when each stage belongs.

Automatic routing activates on multi-file and uncertain work

Claude Code and Codex receive a startup hook that loads a short routing instruction. It calls poteto-mode when a task changes more than one file, alters a called signature, needs a design choice, involves an unknown-cause bug, or concerns performance. Small work proceeds directly. Codex asks you to trust the hook through /hooks, and setup-pstack can disable automatic routing.

Those triggers are reasonable, but they change the texture of daily work. A two-file maintenance fix can become a staged exercise with more model calls and intermediate artifacts. Teams should try the router on a week of real issues, then compare review quality, elapsed time, and token use with their existing prompts. The repository does not claim that every substantial task benefits from the same amount of ceremony.

What happened when we ran it

We did not run commit 92debb7 in our sandbox. The harness classified the repository as JavaScript but had no supported ecosystem path for it, and there was no Dockerfile to provide a fallback. That means we have no measured install duration, dependency footprint, build result, or test result. Any sentence implying that pstack passed on our machine would be false.

The repository does publish maintainer commands using Bun: install from the lockfile, run the generator, and execute the tests directory. Its contribution guide lists extra checks for scripts, workflows, Markdown links, generated files, and runtime integrations. Those instructions show what contributors are expected to run. They are documentation, not a substitute for our missing execution result at commit 92debb7.

Claude Code, Codex, and Pi get different integration layers

Claude Code uses a marketplace plugin and namespaced commands. Codex gets a native plugin manifest, its own hook mapping, and optional command shortcuts. Pi receives an extension that supplies subagent, messaging, question, and wake-up tools, plus routing at each agent start. This is adaptation work, not a folder of identical prompts copied across products.

Support becomes thinner outside those three. The reference says opencode skill discovery and reading were checked on version 1.18.25. Prime Agent and Gemini CLI rely on documented shared-directory discovery, and their live sessions were not tested. More importantly, delegation and multi-model workflows remain unverified on Prime Agent, opencode, and Gemini CLI. Use those routes as skills-only experiments, not equivalent replacements for the dedicated integrations.

Parallel skills trade more model calls for a broader review

arena runs several attempts and combines their best parts. interrogate asks three different models to attack a diff. swarm splits work among several workers. These are useful when competing designs or a risky change justify independent views, but parallelism multiplies calls. The setup skill exists partly to assign different models and reasoning effort to each role.

Pi's documentation is unusually candid about billing: Claude used through Pi is charged per token, and every subagent adds to that bill even if you hold a Claude subscription elsewhere. Codex also needs its multi-agent feature enabled for the parallel skills, with a sequential fallback documented. Set low-cost defaults for routine roles, reserve panels for consequential changes, and do not let automatic routing decide the budget by accident.

Local scripts avoid telemetry, while prompts still leave the machine

pstack says it runs no server and collects no telemetry. Scripts execute locally, and PR tools use the existing GitHub CLI login. The same paragraph states that anything the skills ask an agent to read, including session transcripts, goes to the selected model provider. Local execution therefore does not make the full workflow local.

The distinction matters for incident logs, customer code, and transcripts containing credentials. Review the skills that read history, set provider retention according to your policy, and keep secret-bearing material out of agent context. Some optional workflows need Bun, Graphite, or GitHub authentication, adding local authority even when no pstack service exists. The project's security policy explicitly treats unintended commands, secret exfiltration, and access outside the pointed repository as valid security findings.

An October 3 push shows an active port with a sync burden

GitHub listed 775 stars and two combined open issues and pull requests, and the repository was pushed on October 3, 2026. The latest GitHub release returned by the API was v0.9.57 from October 1. Main's changelog already contained later 0.9.x entries, so the source was moving faster than the latest release record when checked.

This is a maintained port, not a frozen copy. It tracks an upstream revision, applies mechanical Cursor-to-Claude substitutions, and records deliberate policy forks in tools/forks.json. That discipline makes drift visible, while every upstream change still creates merge and verification work for the maintainer. pstack is a good fit when you want that opinionated process across Claude Code, Codex, or Pi. If all you need is a few reusable prompts, 54 skill directories and an automatic router are needless weight.

Alternatives

ProjectWhat it isPick it when
Cursor pstack gh↗The upstream Cursor plugin that this port tracks and translates.pick this instead when Cursor is your agent harness and you want the upstream workflow without a portability layer.
Superpowers gh↗A coding-agent skill set built around planning, test-first work, review, and completion checks.pick this instead when you want a smaller process framework without pstack's broad runtime and role configuration.
Spec Kit gh↗A specification-driven workflow for turning requirements into implementation tasks.pick this instead when written specifications are the center of your process rather than automatic task routing.

What people are saying

  1. [github-trending] michael-denyer/pstack-claude

Sources

  1. pstack README
  2. pstack runtime and skill reference
  3. pstack repository metadata
  4. pstack changelog
  5. pstack security policy

More dev tools reviews

terraform-provider-aws · tailcat · touchHLE · effect · SwitchHosts · Duo-animation · the whole board →