Answer 'where is my order' from the order record itself, disclose that it is AI before anything else, and leave promising a date and granting a remedy to a person.
Do not read the aggregate until it is split: two enquiry types carry the hand-offs, and averaging them into the routine ones makes both invisible. Nestack reports the hand-off rate per enquiry type.
Slice performance — reported separately, not only in aggregateIllustrative example
Slice
Failure rate
Lift
Lift vs. threshold
Status
Delayed or missed delivery dates
3.5%
3.3×
Review
Refund and remedy requests
2.7%
2.6×
Review
Lost or damaged parcels
1.9%
1.8×
Watch
Routine in-transit status
0.6%
0.6×
Normal
Bar: hand-off-rate lift vs. routine-status baseline · scale 0–4.0× · tick marks the 2.0× review threshold2 of 4 slices over threshold
Evidence-linked improvement
Unless it becomes a test, it is only a note
A cycle closes when the failure is a regression case the next release has to pass. That suite is what the next enquiry answered is measured against.
Improvement cycle · five stagesSwitchback — the path turns at Improve and returns at Learn
01Detect
Hand-off rate rises in an enquiry type.
02Diagnose
The transcripts and the records behind them are read until one cause holds.
03Improve
The change ships against a version, with the enquiries that exposed it attached.
04Verify
Release is blocked until the affected regression cases pass again.
05Learn
The case joins the permanent suite and the service playbook.
Learn → DetectThe return edge. The next detection runs against a suite one case longer.
Typical build scope
Twelve workstreams across six weeks
The build scope read against the delivery timeline. Week structure follows the six-week plan — discovery, sources, answering workflow, evaluation, integration, then production validation and handover.
WorkstreamWeek 1Week 2Week 3Week 4Week 5Week 6
01Enquiry workflow discovery and boundary definition.
02Order and carrier source assessment.
03Disclosure, policy-corpus and remedy-cap rule mapping.
04Order and shipment-record ingestion.
05Answer logic and source binding.
06Confidence scoring and hand-off routing.
07Person-on-request workflow.
08Order-system and service-desk integration.
09Promise and remedy cases.
10Guardrails and escalation controls.
11Enquiry-trail instrumentation.
12Deployment, documentation and Agent Care handover.
12 workstreams · 6 weeks · bar shows the weeks a workstream is active — several run in parallelFinal scope and sequence confirmed in discovery
Engagement tiers
What each tier includes
Rows are the capabilities named in each tier's scope. Higher tiers include everything below them.
Capability✓ in scope · — not at this tierPilotOne market, one channelProductionProduction order and deskAdvancedMultiple markets / brands
Introduced at Pilot
Answering from your own records✓✓✓
Person on request✓✓✓
Answer-accuracy baseline✓✓✓
Introduced at Production
Reporting by service level—✓✓
Hand-off workflow in your systems—✓✓
Approved write-back—✓✓
Order-system integration—✓✓
Introduced at Advanced
Multi-market policy corpora——✓
Multi-stage remedy approvals——✓
High enquiry volume——✓
Multi-market service controls——✓
Build priceFrom $5,000From $8,000Custom quote
Final build priceConfirmed after discovery based on integrations, workflow complexity, transaction volume, approval controls and deployment requirements.
Separate from buildBuild pricing is separate from recurring Agent Care, which covers managed monitoring, evaluations, incidents and verified improvements after launch.
What we need from you
What you bring, and what we build with it
Each input maps to a piece of build scope and a week in the delivery timeline.
You bringWe build with it
01Your order records and policy text→Order and shipment-record ingestion and source bindingWeek 1
02Representative past enquiries→Answer baseline, source binding and the policy corpusWeek 2
03Your disclosure text and remedy caps→Disclosure, policy-corpus and remedy-cap rule mappingWeek 1
04Access to relevant APIs, feeds or exports→Order and carrier source assessment, then integration setupWeek 2
05Answers you would not want relied on→Promise cases and the evaluation suiteWeek 4
06What an answer may never promise→Confidence scoring, hand-off routing, guardrails and remedy controlsWeek 3
07Named people to take handed-over enquiries→Person-on-request workflow, then pilot and production validationWeeks 5–6
Nothing else is requiredDeployment, documentation and Agent Care handover are ours.
Delivery timeline
Four phases across six weeks
Each phase sits on the weeks it actually occupies, and week 5 carries both evaluation and launch work.
PhaseW1W2W3W4W5W6
DiscoveryW1
BuildW2 – W3
EvaluateW4 – W5
Pilot & LaunchW5 – W6
Week focusW1Enquiry workflow discovery, policy mapping and the boundaryW2Order and carrier integration and the answer baselineW3Answering workflow, confidence logic and hand-off controlsW4Evaluation suite, promise cases and failure-mode testingW5Service-desk integration, a pilot market and correctionsW6One peak week answered under supervision, then Agent Care handover
Reading the bandEach bar covers only the weeks its work is named in. The week 5 overlap is real, not padding.
At the end of W6Validation closes on live enquiries, and Agent Care picks up monitoring.
DurationSix-week plan shown · typical delivery 4–6 weeks depending on scope confirmed in discovery.
Next step · Transportation AI agent
Build a WISMO agent that discloses first and promises nothing.
Show us your order records, your policy text and the dates you state to customers. Tell us who signs off the disclosure wording and who owns the remedy caps — those two names set the outer edge of what the bot may ever say.