Back to authors
Ship Shit Dev

Ship Shit Dev

201 Skills published on GitHub.

testing-cicd-init

>-

testingcivitest
UncategorizedView skill →

testing-expert

Framework-agnostic testing strategy — which level to test at, what coverage numbers mean, how to design a test that survives refactoring, how to choose test data, and how to kill flakes. Use when deciding what is worth testing, setting or defending a coverage target, reviewing the shape of an existing suite, or diagnosing a flaky or slow test. Framework-specific work routes to a specialist skill instead of being answered here.

testingstrategycoverageflakinesstest-design
UncategorizedView skill →

artifacts-builder

Suite of tools for creating elaborate, multi-component claude.ai HTML artifacts using modern frontend web technologies (React, Tailwind CSS, @agenticindiedev/ui). Use for complex artifacts requiring state management or shared UI components - not for simple single-file HTML/JSX artifacts.

artifactsfrontendhtml
UncategorizedView skill →

biome-validator

Validate Biome 2.3+ configuration and detect outdated patterns. Ensures proper schema version, domains, assists, and recommended rules. Use before any linting work or when auditing existing projects.

biomelinterformattervalidationcode-quality
UncategorizedView skill →

bun-validator

Validate Bun workspace configuration and detect common monorepo issues. Ensures proper workspace setup, dependency catalogs, isolated installs, and Bun 1.3+ best practices. Use when setting up a Bun monorepo, before adding workspace dependencies, auditing an existing Bun workspace, or validating package.json in CI.

bunmonorepoworkspacevalidationpackage-manager
UncategorizedView skill →

clerk-validator

Validate Clerk authentication configuration and detect deprecated patterns. Ensures proper proxy.ts usage (Next.js 16), ClerkProvider setup, and modern auth patterns. Use before any Clerk work or when auditing existing auth implementations.

clerkauthenticationvalidationnextjsnestjs
UncategorizedView skill →

content-script-developer

Expert in browser extension content scripts, DOM integration, and safe page augmentation across modern web apps. Use when building or updating a browser-extension content script, injecting UI into third-party pages, or handling SPA navigation and dynamic DOM changes.

browser-extensioncontent-scriptdom
UncategorizedView skill →

devcontainer-setup

Scaffolds a complete VS Code Dev Container configuration with Docker, docker-compose, and optional Claude Code CLI support. Activates when asked to "set up devcontainer", "add docker development environment", "configure dev container", or containerize a development workflow.

devcontainerdockersetup
UncategorizedView skill →

env-setup

>-

environmentdotenvsecretsconfigurationscaffoldingvalidation
UncategorizedView skill →

fullstack-workspace-init

Initialize Shipshit.dev full-stack product workspaces through npx @shipshitdev/v0, then customize and verify the generated repo. Use for new product scaffolds or post-v0 workspace setup.

fullstackworkspacev0
UncategorizedView skill →

linter-formatter-init

Set up Biome (default) or ESLint + Prettier, Vitest testing, and pre-commit hooks for any JavaScript/TypeScript project. Uses Bun as the package manager. Use this skill when initializing code quality tooling for a new project or adding linting to an existing one.

lintingformattingsetup
UncategorizedView skill →

package-architect

Design and maintain TypeScript packages in a monorepo, including exports and build configuration. Use when creating or restructuring monorepo packages, defining package.json exports, or setting up tsconfig references.

packagesmonorepotypescript
UncategorizedView skill →

project-init-orchestrator

Selects the correct project initialization route and orchestrates setup. Triggers on "initialize project", "set up new project", "bootstrap project", or when scaffolding a new Shipshit.dev product repo. Use v0 for new Shipshit.dev product repos; use lower-level setup skills only for existing repo repair, customization, or small additions.

project-initscaffoldingorchestrationsetupmonorepo
UncategorizedView skill →

invalid-allowed-tools

Exercises rejection of a YAML list in the allowed-tools field.

fixturevalidation
UncategorizedView skill →

invalid-codex-path-claim

Exercises rejection of an unsupported Codex project command path.

fixturevalidation
UncategorizedView skill →

invalid-execution-parameter

Exercises detection of harness-owned execution parameters.

fixturevalidation
UncategorizedView skill →

invalid-model

Exercises rejection of a concrete model identifier.

fixturevalidation
UncategorizedView skill →

invalid-platform-marker

Exercises rejection of inert platform comment markers.

fixturevalidation
UncategorizedView skill →

invalid-plugin-description

Reports portable validation evidence when checking a fixture skill.

fixturevalidation
UncategorizedView skill →

invalid-plugin-version

Reports portable validation evidence when checking a fixture skill.

fixturevalidation
UncategorizedView skill →

invalid-route

Exercises detection of a dangling skill route in ordinary prose.

fixturevalidation
UncategorizedView skill →

invalid-side-effect

Exercises detection of an implicitly invocable mutating skill.

fixturevalidation
UncategorizedView skill →

invalid-triggers

Exercises rejection of dead nested activation metadata.

fixturevalidation
UncategorizedView skill →

valid-composable-writer

Exercises a callable writer with an in-body authorization gate.

fixturevalidation
UncategorizedView skill →

valid-library-reference

Documents libraries without declaring them as skill routes.

fixturevalidation
UncategorizedView skill →

valid-portable

Reports portable validation evidence when checking a fixture skill.

fixturevalidation
UncategorizedView skill →

architect

Sketch types, signatures, and module structure before code, then stay in the loop while implementation fills in. Use for architect this, design this, or non-trivial work where jumping to code would lock in the wrong shape.

architecturedesigntypesmodules
UncategorizedView skill →

arena

Spawn N parallel candidates at the same task, pick a base, and graft the strongest parts of the losers into it. Use for arena this, throw it in the arena, or when one attempt at a non-trivial artifact would lock in the wrong shape.

fan-outbakeoffdesignsynthesis
UncategorizedView skill →

blast-radius

Find what a small-looking change could break somewhere else, and prove the one fact it is safe because of by running real code. Use for blast radius of X, what could this break, or reviewing a small diff you do not trust.

impactsafetyreviewverification
UncategorizedView skill →

create-verification-skill

Generate a project-local verification skill that drives the app the way a user does. Use for create-verification-skill, make a verify skill for this repo, or when a project has no scripted way to prove UI, CLI, or service behavior.

verificationharnessfeature-mapproject-local
UncategorizedView skill →

figure-it-out

Design an auditable playbook when no narrower one fits. Use for figure it out, a large migration, an ambitious multi-part change, or work a human reviews after stepping away. Scales rigor to the task, runs a hypothesis loop, and logs decisions via show-me-your-work.

playbookmigrationaudithypothesis
UncategorizedView skill →

how

Walk through how a subsystem works. Use for "how does X work", code walkthroughs before changing something, and placement or ownership questions. Explains architecture, runtime flow, and onboarding mental models. Can critique architecture. Use why for motivation.

architecturewalkthroughonboardingcritique
UncategorizedView skill →

interrogate

Adversarial multi-reviewer pass over a diff. Use for interrogate, adversarial review, multi-model review, challenge this, stress test this code, find blind spots, or tear this apart. Several independent reviewers challenge the change. The lead synthesizes a verdict and does not auto-apply fixes.

reviewadversarialmulti-reviewerquality
UncategorizedView skill →

maintain-verification-skill

Keep a project's verification skill and feature map honest. Parallel source readers per feature, one live session driving every feature, at most one PR of proven corrections. Use for maintain-verification-skill or audit the verify skill.

verificationmaintenancefeature-map
UncategorizedView skill →

no-comments

Review comments on a diff, delete narration and workaround sermons, fix accepted findings, and offer encodings for claimed constraints. Use before review or when asked to strip comments.

commentsreviewcleanup
UncategorizedView skill →

pstack

Playbook orchestrator for verified, unslopped engineering work. Matches a task to a named playbook, applies the principles index, and routes to how, why, architect, arena, swarm, interrogate, tdd, and related skills. Use for pstack, poteto-mode, /pstack, or requests to work in this style.

orchestratorplaybooksverificationarchitecturereview
UncategorizedView skill →

recall

Rebuild recent working context from chat history, live state, and the shared record, then hand back a tight current-state brief. Use for recall my work on X, catch me up, what have I been working on, or where did I leave off.

contextresumebriefing
UncategorizedView skill →

setup-pstack

Configures canonical Shipshit Pstack adapters for selected harnesses while preserving existing role choices. Use when setting up Pstack, migrating duplicate installations, or enabling a supported optional adapter.

pstackworkflow
UncategorizedView skill →

show-me-your-work

Keep a reviewable decision trail for long-running or unattended work. A TSV log with one row per decision (what, why, evidence, result). Local by default. Commit it when a reviewer needs the trail to trust the result. Use for show-me-your-work, autonomous or multi-phase runs, or work a human reviews after stepping away.

auditdecisionstrailverification
UncategorizedView skill →

swarm

Fan out N parallel workers, drain them, and return one report. Use for swarm this, or parallel coverage, races, gauntlets, and exploration partitions.

fan-outcoverageracereport
UncategorizedView skill →

teach

Explain a body of work so a person actually understands it. Runs how and why and weaves what they find into one plain explanation, built up diagram by diagram. Use for teach me this, help me really understand X, or explain this change or subsystem.

teachingexplanationonboarding
UncategorizedView skill →

technical-writing

Layered technical-writing standard for docs, RFCs, READMEs, PR descriptions, and commit messages. Diátaxis structure, Google developer style sentences, STE instruction rules, Global English syntax. Use for technical-writing or when writing or reviewing those surfaces.

docswritingdiataxisstyle
UncategorizedView skill →

why

Investigate why code is shaped the way it is. Use for design rationale, regressions, postmortems, or data-backed thresholds. Queries available evidence categories in parallel, then returns a cited read on decisions and tradeoffs. Use how for runtime behavior.

investigationrationalehistoryevidence
UncategorizedView skill →

landing-page-vercel

Scaffolds a production-ready static landing page with working email capture form, analytics, and responsive design. Activates on "create landing page", "build a landing page", "launch page for product", or similar requests. Optionally deploys to Vercel on explicit request.

landing-pagevercelfrontend
UncategorizedView skill →

monitoring-setup

Sets up production monitoring for NestJS and Next.js apps — Sentry error tracking, Google Analytics, and operational signals (BullMQ queue depth, Postgres slow queries, connection saturation) with alerts on each. Activates when users need error tracking, production monitoring, analytics, queue/database observability, or alerting on operational health.

monitoringsentryanalyticsobservabilityalertingbullmqpostgres
UncategorizedView skill →

context-optimization

>-

contextoptimizationagents
UncategorizedView skill →

advanced-evaluation

Design and operate LLM-as-a-Judge evaluation systems using direct scoring, pairwise comparison, rubric calibration, evaluator bias mitigation, confidence scoring, and automated quality assessment. Use when building LLM-as-judge systems, comparing model responses, calibrating rubrics, debugging inconsistent evaluations, or designing A/B tests for prompt or model changes.

evaluationllm-as-judgequalitybias-mitigation
UncategorizedView skill →

agent-browser

Automates browser interactions for web testing, form filling, screenshots, and data extraction. Use when the user needs to navigate websites, interact with web pages, fill forms, take screenshots, test web applications, or extract information from web pages.

browserautomationtesting
UncategorizedView skill →

comment-mode

Granular feedback on drafts without rewriting. Generates highlighted HTML with click-to-reveal inline comments. Use when user says "comment on this", "leave comments on", "give feedback on", or asks for feedback on a draft. Supports multiple lenses—editor feedback, POV simulation ("as brian would react"), or focused angles ("word choice only", "weak arguments"). A granular alternative to rewrites that lets users review feedback incrementally without losing their voice.

feedbackcommentsediting
UncategorizedView skill →

context-degradation

Recognize, diagnose, and mitigate patterns of context degradation in agent systems. Use when context grows large, agent performance degrades unexpectedly, or debugging agent failures.

contextagentsreliability
UncategorizedView skill →

Page 4 of 5 · 201 results