Pick the tier that matches your situation
Every engagement starts with a scoping discovery call. Final price depends on your agent’s complexity and authority model. The ranges below are directional anchors — they reflect where most engagements land, not a fixed quote.
Tier 1
Starter Audit
From $2,500 AUD
One agent · assessment and risk register · no implementation
- 30-min scoping discovery call
- Authority model and blast-radius map
- Top-5 launch blockers identified and ranked
- Risk register you own and maintain going forward
- Code Production Hardening or Transaction Controls assessment
Most common entry point
Tier 2
Fleet Pilot
From $7,500 AUD
One agent · full engagement · discovery through launch-readiness handoff
- Everything in Starter Audit
- Full production audit across all seven failure modes
- Targeted hardening and controls implementation
- Approval gate and dry-run architecture
- Audit trail and rollback path documentation
- Reusable governance worksheet your team applies across the fleet
- Both Code Production Hardening and Transaction Controls if your agent crosses both failure modes
Tier 3
Full Deployment
From $18,000 AUD
3+ agents · fleet-wide controls · governance architecture
- Everything in Fleet Pilot, applied across your agent fleet
- Fleet-wide authority model and permission scope design
- Shared audit trail and observability architecture
- Agent-to-agent trust boundary design
- Behaviour baseline and drift detection framework
- Governance policy documentation for compliance review
Final scope and price are confirmed on the discovery call. If your situation is smaller or more complex than these anchors, say so in the inquiry form and we will scope accordingly. No obligation until we both agree on fit.
Fleet Pilot vs build-in-house vs hire a consultancy
If you are weighing these options, the decision usually comes down to three questions: how fast do you need it, how reusable does the output need to be, and what does a production incident actually cost you.
| What you’re comparing |
RFE Online Fleet Pilot |
Build in-house |
Hire a generalist consultancy |
| Time to first controls in production |
2–3 weeks from discovery call |
4–12 weeks depending on team familiarity |
6–16 weeks (discovery, scoping, proposal cycles) |
| Governance model based on live fleet evidence |
Yes — RFE Online’s 10-agent production fleet |
No — builds from theory and first principles |
No — typically framework-based, not production-tested |
| Reusable worksheet your team owns |
Yes — included in Fleet Pilot and above |
Yes — but you build it yourself |
Sometimes — often locked behind ongoing retainer |
| Fixed-scope engagement (no scope creep) |
Yes — confirmed on discovery call |
Depends on internal resourcing and priorities |
Rarely — generalist consultancies expand scope to expand billing |
| Addresses both code and transaction failure modes |
Yes — both plays available in one engagement |
Depends on team expertise in both domains |
Usually one or the other; specialists cost more |
| Ongoing retainer required |
No — engagement ends at launch-readiness handoff |
No — but requires ongoing internal resource |
Often — recurring billing is the consultancy model |
The honest case for building in-house: if your engineering team already has production AI governance experience and bandwidth, they can implement this faster than any external engagement. The case for RFE Online: if they don’t, the gap between what they think they know and what a live adversarial production environment reveals is where incidents happen.
If you came here from the trust-readiness cluster
The trust-readiness scorecard, hardening vs behaviour testing analysis, and agent fleet hardening case studies all point to the same production question: your agent holds authority it should not hold without governance, and the gap between its current state and a defensible production posture is where incidents happen.
Trust Readiness Scorecard
Know your governance gap score before pricing
Five questions, sixty seconds. If your agent scores high on governance urgency, Fleet Pilot is almost certainly the right tier. If you are pre-production and early-stage, Starter Audit gives you the artefact you need without over-committing.
Take the scorecard →
Trust Audit Sample
See what a production audit actually delivers
The trust audit sample shows the format and detail level of the risk register and authority model map you receive. Before committing to a tier, reading what the output looks like will tell you whether the depth matches what your team needs.
Read the trust audit page →