Build Loop Codex

buildgreatproducts/builder-os/skills/build-loop-codex

by buildgreatproductsfb74cac0fc9dMIT228 starsListed Oct 8, 2026Updated Oct 8, 2026Repository updated 3 months ago

Use when building features with **Codex** (OpenAI Codex CLI) in any codebase and the work should go through a disciplined build → review → test → fix loop. Triggers on "run the build loop", "build the next task", "continue the plan", "build this feature properly", or any request to implement work from a plan file or a direct feature prompt. Builds from the plan (or the prompt if no plan exists), runs Codex's `/review` on uncommitted changes and fixes every issue found, tests and verifies the feature end to end, fixes anything testing surfaces, and reports back once complete. Repeats until all plan tasks are checked off.

Instructions onlySoftware Development
AI-generated overview

Runs a disciplined build, review, test and fix loop for feature work using the OpenAI Codex CLI.

What it does
This skill guides an agent through a per-task quality loop: implement the task, run Codex's /review on uncommitted changes and fix findings, test the feature end to end, and fix anything testing surfaces. It works from the first unchecked task in a plan file, or from the user's prompt when no plan exists, and repeats until the requested scope is complete. It finishes with a report covering what was built, plan progress, review findings fixed or deferred, and how the work was verified.
When to use it
Use it when implementing features with the OpenAI Codex CLI and the work should go through a build, review, test and fix cycle. It fits requests such as running the build loop, building the next task, continuing a plan, or implementing a feature properly from a plan file or direct prompt.
Requirements
Requires the OpenAI Codex CLI with its /review command available, plus a codebase to work in. It is instructions only and ships no scripts; it may need to run the project's tests and application to verify features.

Codex Build Loop

Quality-gated feature work: nothing ships on "it compiles" — every increment is built, reviewed, tested end to end, and fixed before the user hears "done."

Source of work

  • A plan file exists (roadmap, refactor plan, or task list with - [ ] checkboxes — search the repo): work the first unchecked task. Tasks are ordered intentionally — never skip ahead. If the plan references spec docs, read only the sections relevant to the current task.
  • No plan (or the request is outside it): build from the user's prompt. Restate it as a verifiable goal with 2–4 success criteria and confirm scope in one message before building.

The loop

Run per task (or per prompted feature). Do not advance until every step passes.

  1. Build. Implement exactly what the task specifies. Simplest implementation that satisfies it, surgical changes, no speculative scope. Match existing project conventions.

  2. Review. Run /review and select "Review uncommitted changes". If the change touches auth, payments, user input, or data access, run a second pass via "Custom review instructions" (e.g. "Focus on security vulnerabilities and unvalidated input"). Fix all findings in scope — bugs, security issues, edge cases, performance, style in files you touched. If the project has a design system spec (design tokens file, DESIGN.md, theme config), check UI changes against it — no hardcoded colors, type, or spacing that bypass tokens. Note pre-existing issues in untouched code for the report instead of fixing silently. Re-run /review until clean. If a finding contradicts the task or spec, the spec wins — flag the disagreement.

  3. Test end to end. Run the task's verification step (or the success criteria). Run the full test suite — everything that passed before must still pass. Add tests for new logic. Then exercise the feature as a user would: run the app, walk the real flow including empty, loading, and error states.

  4. Fix. Anything testing finds goes back through the loop: fix → /review → re-test. Never mark a failing task complete; never start the next task with the app broken.

  5. Continue. Mark the task - [x], update any progress/status line in the plan, and loop to the next task until the requested scope is complete.

  6. Report. When done, tell the user: what was built and plan progress, review findings fixed and anything deferred, how it was verified (tests + flow walked), and what needs their attention next. Be honest about anything flaky or partially verified.

Rules

  • Skipped review or untested work = unfinished work.
  • Don't relitigate plan decisions; if a task seems wrong, ask one specific question rather than guessing.
  • Discovered work no task covers? Surface it and propose a task — never silently expand scope.

Source and attribution

Source:buildgreatproducts/builder-osinskills/build-loop-codexat commitfb74cac

License: MIT

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal

More from buildgreatproducts/builder-os

Product Planner

buildgreatproducts

Vision intake conversation followed by generation of three product documents — `docs/product-vision.md` (strategy and brand), `docs/prd.md` (technical spec for coding agents), and `docs/product-roadmap.md` (phased build plan with task checkboxes). Also captures the founder's answers as `docs/VISION.md`. Use when the founder says "plan my product", "plan a product", "define my vision", "generate a PRD", "create a roadmap", "spec out my idea", "help me build something", or wants to convert an idea into shippable spec documents.

Awaiting classification228updated 3 months ago

Idea Validator

buildgreatproducts

Pressure-tests a product idea, producing a validation report with flaws, competition, MVP test and a verdict.

Business & Finance228updated 3 months ago

Idea Generator

buildgreatproducts

Guides a founder through a structured conversation to discover a product idea and writes it to docs/product-idea.md.

Productivity & Workflow228updated 3 months ago

Design System

buildgreatproducts

Translates an image (or a set of image references — screenshots, mockups, Figma URLs, live websites) into two mirrored design-system artifacts: `docs/design.md` (YAML tokens + prose, following Google's open [design.md](https://github.com/google-labs-code/design.md) format, for the coding agent) and `docs/design.html` (a self-contained, token-driven style guide rendering every token and component live, for the human to read). Reads the imagery, asks targeted clarifying questions, derives the design tokens (colors, typography, spacing, rounded, components), and writes both files. Fully standalone — requires no other document or skill. Use when the founder says "create a design system", "design from image", "translate image to design", "create design.md", "image to design system", "extract design tokens", or shares an image with no other clear intent.

Awaiting classification228updated 3 months ago

Design Better

buildgreatproducts

Applies UX/UI craft heuristics to frontend code generation and review, deferring visual style to a design system file.

Design & Creative228updated 3 months ago

Build Loop Cursor

buildgreatproducts

Guides a disciplined build, review, test and fix loop for implementing planned features with Cursor.

Software Development228updated 3 months ago