Meter
Status: Running in an MSP billing stack
Evidence: In production for an MSP; one reconciliation audit covered 202 accounts and 931 work items (self-reported)
Start with the field work: systems deployed into billing, operations, and media workflows. Then inspect the public extracts and demos. Deeper backend checks live in the demo operations page.
The production repositories stay private. Each case study names the operator problem, deployment state, evidence basis, and public extract where one exists. Case studies
Status: Running in an MSP billing stack
Evidence: In production for an MSP; one reconciliation audit covered 202 accounts and 931 work items (self-reported)
Status: Shipped and actively maintained
Evidence: Six vendor integrations share one connector contract; architecture notes estimate 69% fewer daily API calls
Status: Live, pre-launch
Evidence: A 23-step media pipeline with 11 programmatic scorers, five optional LLM evals, and a human release gate
Every entry names the file worth reading first. Entries tagged cold-clone reviewed were verified by an independent engineer working from a fresh clone before the repository went public. 27 public extracts
What it is: An agent that takes an approved GitHub issue and returns a tested, reviewable pull request.
Read first: The agent runner's test suite, which exercises the plan, patch, and review loop end to end.
What it is: An append-only decision ledger with a human publish gate, enforced inside Postgres by triggers, a hash chain, and unique indexes.
Read first: Attack tests that try to rewrite history and race two sessions against the publish gate; the migrations under them close what the tests document.
What it is: Human review as an API: agents submit work, real reviewers return a consensus verdict.
Read first: A README walkthrough of the offline verification harnesses, with a 173-test suite behind it.
What it is: Disposable cloud servers with a lease. When it expires, teardown writes a receipt proving the machine is gone.
Read first: A hermetic reaper suite covering lease expiry, escalation, and the receipts a destroyed server leaves behind.
What it is: Postgres row-level-security multi-tenancy, tested by an adversarial suite that attempts cross-tenant reads.
Read first: SQL attack tests that probe the tenant boundary from inside a session, including the one documented way across it.
What it is: Finds billing discrepancies, proposes typed fixes, and requires a person to approve every invoice change.
Read first: The verification workflow CI runs on every push, plus a README that walks the local run.
What it is: Migrates a WordPress export to Astro and fails the build when any origin URL is missing from the built site.
Read first: Parity tests plus the synthetic Harborline fixture the verifier runs against. No client URLs are in the repository.
Grades a codebase's technical debt and prices the paydown.
Turns AI coding-session transcripts into durable, searchable lessons.
Scores and routes inbound leads; a 24-case model eval suite gates every accuracy change.
A terminal dashboard for running a fleet of AI coding agents in tmux.
Fans one prompt out to a fleet of AI coding agents in tmux panes, then collects their work.
One CLI that routes a prompt to Claude, Codex, Cursor, or Gemini, with JSON output and real exit codes.
PostgreSQL fuzzy matching for public filings, with hermetic SQL assertions and a bounded live demo on Texas open data.
An autonomous coding agent that works a Linear backlog end to end and opens the pull requests.
The sanitized public kit behind Throughline: one four-method connector contract, the sync engine, sample CRM data, and tests.
One Zod-validated operation registers as an HTTP route, a CLI command, and an opt-in MCP tool, with OpenAPI emitted from the same registry.
Reads, injects, and exports 1Password secrets from the command line so they never have to sit in a committed .env.
Scores Claude prompts against a stored eval suite, with string-match or rubric grading you can re-run.
A Claude Code stop hook that sends the agent back to work when the last turn still lists next steps, and stops when it is asking you a question.
Forty-nine MCP tools over Meta's Marketing API, with a smoke test that starts the server and lists every tool without credentials.
Compiles a JSON show definition into a vMix preset bundle a studio machine can open.
Suggests the next Claude Code skill from a local library using fuzzy search, TF-IDF, and optional embeddings.
One command definition writes Raycast, PopClip, Dropzone, and Shortcuts artifacts, and will not emit a surface that cannot represent that command's input.
Drives Behringer X Air and Midas MR mixers from the terminal, and a parity test fails the build if the docs drift from the CLI.
Claude Code plugin packs for building, testing, and deploying, where CI runs the plugin validator over every pack on every push.
Resolves llm model keys from 1Password at prompt time. refs.json stores the op:// reference; a failed read aborts the call.
The rest of the public estate lives on the GitHub profile.
Availability: Live check · starts in your browser
Status: Live check pending
What you’ll see: A real server using synthetic sample data
Availability: Live check · starts in your browser
Status: Live check pending
What you’ll see: A real server using synthetic sample data
Availability: Live check · starts in your browser
Status: Live check pending
What you’ll see: A real server using synthetic sample data
Availability: Live check · starts in your browser
Status: Live check pending
What you’ll see: A real server using synthetic sample data
Availability: Live check · starts in your browser
Status: Live check pending
What you’ll see: A real server using synthetic sample data
Availability: Live check · starts in your browser
Status: Live check pending
What you’ll see: A real server using synthetic sample data
Availability: Live check · starts in your browser
Status: Live check pending
What you’ll see: A real server using synthetic sample data
If a row here is the reason you want to talk, email me at [email protected].