Bounded agents that do real work, with logs you can trust.
Agent demos are easy. Agents that survive contact with real data, real permissions, and real edge cases are not. I build bounded agents that use tools, respect permissions, log every action, and hand off to a human when they should. Connected to your workflows, measured by evals, and safe to put in front of real work.
- Investment
- From $25k
Fixed project fee, scoped to the agent. Complexity (tools, autonomy, integrations) sets the band.
All services
Bounded agents that actually shipBounded AI agents with tools, permissions, logs, and evals. Not fragile demos.
Manual work, automated with control
Support agents you own, not rent
Every lead answered in seconds
Bounded agents that actually ship
- From $25k
Prototypes taken to production
Embedded AI leadership, on demand
Most agent demosnever reach production.
The gap between a demo and a deployed agent is permissions, logging, evals, and failure handling. A demo answers a happy-path prompt. A production agent has to know what it is allowed to do, prove what it did, and fail safely when the input is strange. That gap is the work.
- Agents that work in a demo and fall apart on real data
- No record of what the agent did or why
- No permission boundaries, so the blast radius is unknown
- No evals, so quality is a vibe, not a number
The buildBounded agents with tool use, permissions, logs, evals, and approval.
Bounded agents with tool use, permissions, logs, evals, and approval.
Received escalation from the WhatsApp channel.
Order #8412 shipped twice — drafting a refund for the duplicate.
Scoped to a real job, given only the tools and permissions it needs, and instrumented so you can see and trust every run.
- A bounded agent scoped to a specific job
- Tool use wired to your real systems
- Permission boundaries and a known blast radius
- Full run logs and tool-call history
- Evals and human approval for the high-stakes steps
A small number of moves,each one verifiable.
Every stage ships its own deliverables, so you can see progress and correct course before the next one starts.
- 01Scope the job
We define exactly what the agent does, what it can touch, and where a human stays in the loop.Job spec · Permission map · Tool list
- 02Build the agent
I build the agent, its tools, and its guardrails against your real systems.Working agent · Tools · Guardrails
- 03Instrument and eval
Logging, tool-call history, and an eval suite go in so quality is measured, not assumed.Run logs · Eval suite · Approval flow
- 04Ship and own
You own the agent and its instrumentation. An optional retainer keeps the evals honest as it scales.Documentation · Handoff · Optional retainer
I do not sell an autonomous workforce.I sell bounded agents that work.
No agent army, no AI employees, no autonomous company. Bounded agents with logs, permissions, approvals, and a number for how well they do the job.
I do not advise on AI from the outside. I build these systems in my own products first, live with the failure modes, and rebuild the parts that break under real users. The patterns I bring to your build are the ones I have already paid for in mine.
See the things I have built
See the things I have builtOne project fee.Scoped to the job.
Fixed project fee, scoped to the agent. Complexity (tools, autonomy, integrations) sets the band.
A bounded production agent with tools, permissions, logs, and evals.
Multi-tool, autonomous, or deeply integrated agents sit higher in the band.
Before you reach out
Start with the AI Systems Intensive.
$5,000, one focused engagement to map the work and decide what to build. The fee credits into the build. If a build is wrong for you, I will say so and you keep the plan.
Start with a short call. Straight answer either way.
We confirm fit, scope the work, and decide whether to start with the Intensive or go straight to the build.
One letter, every Sunday.Working systems, not hot takes.
Weekly. No spam. Unsubscribe anytime.









