No client logos yet. Something better: two production systems we built and run ourselves, in the open. Clone them, run the tests, read the threat model.
Not a questionnaire and not client logos: first-party systems we built and operate, public and re-runnable.
A from-scratch, single-VPS runner on gitolite, rootless Docker, and sops/age. It ships this very site, with a written threat model that names what it does not defend.
“It is not a multi-tenant CI platform.”
Single-tenant, single-VPS, on purpose. Knowing where a system stops is production discipline, not a gap.
Built on the Claude Agent SDK: idea → orchestrator → parallel research workers → critic loop → policy gate → deployed site. The exact stack we sell: orchestration, evals, guardrails, cost control.
“It does not guarantee revenue. It builds the machine and optimizes on evidence; the market decides.”
Straight from its own CAPABILITIES.md. We built the machine and drove one idea to a deployed site. We never imply revenue or a track record.
No trust required, and no vanity metrics. The repos are public — read the design docs and the guardrail code and judge the work directly.
Honest scope limits, written down, not marketing.
A security hole closed by construction, not by a check bolted on later.
Spend and named-contact actions are predicates gated to a human. No stage advances on vibes.
The system states its own limits before it sells you anything.
No client case studies yet. The strength: our own systems are public and re-runnable, so you judge the engineering directly instead of taking our word.
A 45-minute working session on your actual AI project, not a sales deck. We find where it breaks and name the exact thing standing between you and launch.