Two agents, separately deployed and separately tested: one measures for match officials, the other prices in-play markets inside an operator's approved system, and a person decides in both.
Take an aggregate override rate apart before trusting it: split it by cohort, and by deployment, because a small number of hard cohorts carry most of the overrides in both.
Slice performance — reported separately, not only in aggregateIllustrative example
Slice
Failure rate
Lift
Lift vs. threshold
Status
Occluded and crowded incidents
9.1%
3.7×
Review
Single-action micro-markets
5.4%
2.2×
Review
Restarts after an interruption
3.7%
1.5×
Watch
Clear, settled cases
1.5%
0.6×
Normal
Bar: override-rate lift vs. clear-case baseline · scale 0–4.0× · tick marks the 2.0× review threshold2 of 4 slices over threshold
Evidence-linked improvement
A cycle ends in a test, not a meeting
Nothing closes on a review meeting. It closes on a case the next release has to pass, and that suite is what the next output either side is measured against.
Improvement cycle · five stagesSwitchback — the path turns at Improve and returns at Learn
01Detect
Override rate rises in one cohort.
02Diagnose
If the band held, the fault is in the definition, the version or the hand-off.
03Improve
Every change ships against a version with the incidents attached.
04Verify
The release is blocked while an affected case fails.
05Learn
The case becomes permanent, and the separation controls are re-tested.
Learn → DetectThe return edge. The next detection runs against a suite one case longer.
Typical build scope
Twelve workstreams across six weeks
The build scope read against the delivery timeline. Week structure follows the six-week plan — discovery, sources, output workflow, evaluation, integration, then production validation and handover.
WorkstreamWeek 1Week 2Week 3Week 4Week 5Week 6
01Deployment separation and boundary definition.
02Venue and market-feed source assessment.
03Definition, tolerance band and house-rule mapping.
04Input ingestion and normalisation.
05Output logic and tolerance binding.
06Confidence scoring and unresolved routing.
07Official or trader workflow.
08Per-deployment system integration.
09Separation and tolerance cases.
10Guardrails and approval controls.
11Decision-trail instrumentation.
12Deployment, documentation and Agent Care handover.
12 workstreams · 6 weeks · bar shows the weeks a workstream is active — several run in parallelFinal scope and sequence confirmed in discovery
Engagement tiers
What each tier includes
Rows are the capabilities named in each tier's scope. Higher tiers include everything below them.
Capability✓ in scope · — not at this tierPilotOne competition or bookProductionProduction deploymentAdvancedMultiple competitions / books
Introduced at Pilot
Output to your filed definitions✓✓✓
Separation, and a human decision✓✓✓
Measurement-consistency baseline✓✓✓
Introduced at Production
Reporting by competition and market—✓✓
Decision workflow in your systems—✓✓
Approved write-back—✓✓
Trading-platform integration—✓✓
Introduced at Advanced
Multi-competition rule sets——✓
Multi-desk trading approvals——✓
High event volume——✓
Multi-jurisdiction filing controls——✓
Build priceFrom $5,000From $8,000Custom quote
Final build priceConfirmed after discovery based on integrations, workflow complexity, transaction volume, approval controls and deployment requirements.
Separate from buildBuild pricing is separate from recurring Agent Care, which covers managed monitoring, evaluations, incidents and verified improvements after launch.
What we need from you
What you bring, and what we build with it
Each input maps to a piece of build scope and a week in the delivery timeline.
You bringWe build with it
01Your filed definitions or your house rules→Input ingestion and definition mappingWeek 1
02Representative events from one deployment→Output baseline, band binding and version stampingWeek 2
03Your tolerance bands and who signs them→Definition, tolerance band and house-rule mappingWeek 1
04Access to relevant APIs, feeds or exports→Venue and market-feed assessment, then integration setupWeek 2
05Calls and prices you would not want repeated→Separation cases and failure-mode testingWeek 4
06What a measurement is never allowed to decide→Confidence scoring, unresolved routing, guardrails and separation controlsWeek 3
07Named officials or traders to act on it→Decision workflow, then pilot and production validationWeeks 5–6
Nothing else is requiredDeployment, documentation and Agent Care handover are ours.
Delivery timeline
Four phases across six weeks
The bands sit on the weeks the work occupies, so the fifth carries evaluation and launch together.
PhaseW1W2W3W4W5W6
DiscoveryW1
BuildW2 – W3
EvaluateW4 – W5
Pilot & LaunchW5 – W6
Week focusW1Deployment separation, definition mapping and the automation boundaryW2Source integration and the output baselineW3Output workflow, tolerance logic and decision controlsW4Evaluation suite, separation tests and failure-mode testingW5Per-deployment integration, pilot events and targeted correctionsW6One competition round run under the officials or the trading desk, then handover
Reading the bandEach bar covers only the weeks its work is named in. The fifth week carries two kinds of work.
At the end of W6The round closes validation and Agent Care picks up monitoring.
DurationSix-week plan shown · typical delivery 4–6 weeks depending on scope confirmed in discovery.
Next step · Sports & Fitness agent
Build one deployment, on its own side of the line.
Tell us which side you are on — a competition's officials, or a licensed book's desk. Your governing body signs the definitions, or your licensee files the house rules; we build to one of them.