Nestack Agent Care
Industries / Accounting / Tax-return prep agent

Accounting AI agent · Tax-return prep

Tax-Return Preparation & Review AI Agent

Ingest source documents, roll the prior-year file forward, populate the return in your prep software and raise review notes with the source behind each entry — the signing preparer decides, signs and stays accountable.

4–6 weeksTypical delivery
Your stackDeployment
Pre-signaturePreparer review
Agent CareAfter launch

What this agent does

Populates the return, never the position it takes

In
01

A W-2, 1099, K-1 or trial balance arrives from the client portal, and the agent reads it into the return.

02

The prior-year file rolls forward, and every carried figure keeps the source return it came from.

Reason
03

The trial balance and source documents populate the return, form by form, in the firm's prep software.

04

The firm's configured entity elections, apportionment rules and mapping conventions are applied, not chosen.

05

Every populated figure is bound to the document behind it, and what the documents leave unclear is marked.

Decide
06

A late K-1, an unreconciled 1099 or a refundable-credit claim without its documentation is flagged.

07

Each review note reaches the signing preparer with its source document attached.

Out
08

The source document, the populated entry, the review note and the preparer's disposition are retained against the return.

09

Execute write actions only inside the approval boundaries agreed during implementation.

Product statement

The agent populates the return and drafts review notes; the signing preparer decides every position, signs and stays the preparer of record.

Example workflow

One return, document to sign-off

AgentHuman
1Source documents receivedW-2, 1099, K-1, trial balance or prior-year file
2Prior year rolled forwardCarried balances, elections, entities and schedules, each tied to the return it came from
3Return populatedForm-by-form entries, review notes, flagged items and confidence
4Controls appliedSource-document checks, due-diligence checklist, apportionment rules and confidence threshold
No human action required

Stages 1 to 4 run unaided, and nothing is filed at any of them — the agent is populating, and the preparer's lane opens at the confidence gate.

5DecisionBranches at the confidence threshold
High confidence

Goes to the signing preparer to review.

Low confidence

Adds a due-diligence review first.

Preparer review

The return is held with its source documents, its review notes and the confidence.

Approve · Edit · Send to reviewer
Approved — ready to sign
6Prep software updatedOnly where write access and approval policy allow it
7Outcome evaluatedEntry accuracy, preparer edits, due-diligence outcomes and post-review corrections
Edits

Every preparer edit is counted in the evaluation.

What should not run autonomously

Human approval stays in control

Outside the boundary — human approval required8 items
Signing the return.
Transmitting before Form 8879 is signed.
Deciding whether a position has substantial authority.
Determining filing status without review.
Automation boundaryAgent acts unaided
Populate the return from source documents and.
Apply the firm's configured entity, election and apportionment rules.
Bind each entry to the document it came from, and mark what is unclear.
Flag due-diligence gaps and unusual items, and hold.
Any populated field runs inside the boundaries agreed at implementation, never ahead of a signed consent or a review.
Deciding whether to disclose a position on Form 8275.
Sending taxpayer data to a third party without a signed §7216 consent.
Attesting to Form 8867 due diligence.
Changes to mapping rules or automation thresholds.

Example output

One entry in the return, annotated

Everything the agent populates is attached to the document it came from.

Preparation output · single returnIllustrative example
Return
Populated entry
Amount
Source document
Confidence
Preparer of record
Partnership K-1 return
Guaranteed payment carried to Schedule E as ordinary income
$42,000
K-1, Box 4a
91%
Preparer name and PTIN
As receivedTaken from the K-1 and the prior-year file — nothing on this side is a position the agent decided.
Source documents used K-1, Box 4a Partnership agreement Prior-year Schedule E
Why this entryThe K-1 states it as a guaranteed payment; whether it changes.
ActionApproveEditSend to reviewer
What the score decidesBelow the configured threshold the return picks up a due-diligence read before it reaches.

Value

Where AI adds value

The same four claims, placed at the point in the workflow where each one applies.

Where the value landsValue 01 – 04
Every returnFrom the source documents
03Preparation

Populate from the documents

Draw on the documents on file and the firm's configured elections and apportionment rules.

01Approved path

The signature is the position

Routine entries arrive already populated, tied to their documents.

02Human review

Send review to the risky returns

Due-diligence gaps and low-confidence entries are marked, so the preparer's read starts where risk concentrates.

04Build an evidence trail

The entry, the document it came from and the preparer who signed stay on the return.

Integrations

Typical integrations

Five system groups connect to the same agent. Which of them are in scope is decided in discovery.

Tax prep softwareDrake · UltraTax · Lacerte
ProSeries · CCH Axcess
Source-document intakeSafeSend · Intuit Link
K-1 aggregation portals
Practice managementClient portals · e-organizers
Prior-year engagement files

Agent

Tax-return preparation & review

Reads the documents
Populates the return
Holds for review

Signature and transmissionForm 8879 e-signature
ERO / e-file systems
Observability & evaluationOpenTelemetry · Langfuse
Supported monitoring/evaluation sources

Integration availability depends on the client's existing systems and API access.

Agent controls

Six layers between the model and the filed return

Every control wraps the one within it. What the set does not catch is named in the map below.

L6 · Outermost — last line of defenceInward → L1 · closest to the model
L6Rollback / safe modeNarrow the agent to data capture when evaluation or production signals degrade.Roll back
L5Version monitoringTrack model, prompt, mapping-rule and configuration changes.Track
L4Return trailRecord the source document, the entry, the review note and the preparer's disposition.Record
L3Preparer reviewHold the return for the signing preparer; it governs release, not whether a position is right.Gate
L2Diligence guardrailsTest entries against the due-diligence checklist and consent status; a failure holds the return.Restrict
L1Confidence thresholdsRoute low-confidence entries to a due-diligence read before the preparer sees them.Require review
Model coreReturn populated — entries, review notes, flagged items and confidence
L1 – L2Test whether an entry may stand
L3Puts the signature in the preparer's hands
L4 – L5Keep the entry and the document behind it
L6Narrows to data capture when signals degrade

How Nestack evaluates it

Evaluate the full preparation workflow — not only the filed position.

Coverage runs the whole depth of the workflow, and every layer is cut by slice.

Surface — the return the client signs
Depth of coverage ▼
E1Final-output evaluationDid every populated entry match its source document?
E2Step-level evaluationDid the agent use the right prior-year file, elections and apportionment rules?
E3Tool evaluationDid it read and write the correct return and the correct field?
E4Confidence calibrationDo low-confidence entries actually attract more preparer edits?
E5Slice evaluationHow does performance change across specific return types?
E6Business outcomeHow many entries needed a preparer edit or a correction after review?
Floor — the outcome the firm signs for

Failure modes

Where each failure originates in the agent

Seven failure modes, placed at the stage each one originates.

Agent lifecycleDirection of processing →
01 · Retrieval1 mode
RP-03

Superseded prior-year source

A rolled-forward balance is read from a year the client later amended.

Stage gathersSource documents, prior-year file and firm rules
02 · Reasoning2 modes
RP-04

Unsupported position drafted

An entry is drafted with no citation the substantial-authority standard would accept.

RP-06

Credit claimed without diligence

A refundable credit is populated from rolled-forward data with no current-year interview on file.

Stage proposesPopulated entries, flags and confidence
03 · Tool / write2 modes
RP-02

Unconsented data transfer

Source documents reach a third-party process before a signed §7216 consent is on file.

RP-05

Stale election carried forward

A prior-year method or election rolls into this year's return without a reviewer prompt.

Stage writesOnly where write access and approval policy allow it
04 · Output1 mode
RP-01

Preparer identifier omitted

The signing preparer's PTIN is missing from the populated return before it reaches review.

Stage returnsThe return the signing preparer reviews and signs
05 · Change / Version1 mode
RP-07

Silent mapping regression

A software or rule-set update widens what the agent will populate without a reviewer prompt.

Stage tracksModel, prompt, mapping rules and apportionment config
Sev-1 · transmits outside the boundary Sev-2 · unsupported entry reaches the return Sev-3 · source degrades, entry routes to review

Affected slices

One review-note rate can hide where the risk sits

A review sample drawn evenly across return types can look acceptable in aggregate while a handful of cohorts absorb most of the review notes. Nestack reports the review-note rate by slice, not only in total.

Slice performance — reported separately, not only in aggregateIllustrative example
SliceFailure rateLift Lift vs. thresholdStatus
Partnership K-1s received late6.9%3.6× Review
Multi-state apportionment returns5.1%2.7× Review
Returns with refundable credits3.3%1.7× Watch
Single-state wage-income returns1.8%0.7× Normal
Bar: review-note-rate lift vs. the single-state wage-income baseline · scale 0–4.0× · tick marks the 2.0× review threshold 2 of 4 slices over threshold

Evidence-linked improvement

The loop closes on a regression case, not an explanation

The loop shuts when the wrong position is a regression case, not when it has been explained. That suite is what the next return prepared is measured against.

Improvement cycle · five stagesSwitchback — the path turns at Improve and returns at Learn
01Detect

Review-note rate rises in a return-type slice.

02Diagnose

The reviewer who accepted it retraces the entry to its document until the cause narrows to one.

03Improve

Whatever changes ships against a version, with the returns that prompted it attached.

04Verify

Nothing ships until the affected positions pass a second time.

05Learn

The suite grows by one case; so does the review checklist.

Learn → DetectThe return edge. The next return prepared is checked against a suite one case longer.

Typical build scope

Twelve workstreams across six weeks

The build scope read against the delivery timeline. Week structure follows the six-week plan — discovery, documents, preparation workflow, evaluation, integration, then production validation and handover.

Workstream Week 1Week 2Week 3Week 4Week 5Week 6
01Workflow discovery and boundary definition and it is logged..
02Document and software source assessment.
03Election and apportionment-rule mapping and rule mapping.
04Document ingestion and normalisation.
05Roll-forward logic and document binding.
06Confidence scoring and review-note routing.
07Preparer review workflow.
08Prep-software and e-file integration.
09Position and due-diligence cases.
10Guardrails and review controls.
11Return-trail instrumentation.
12Deployment, documentation and Agent Care handover.
12 workstreams · 6 weeks · bar shows the weeks a workstream is active — several run in parallel Final scope and sequence confirmed in discovery

Engagement tiers

What each tier includes

Rows are the capabilities named in each tier's scope. Higher tiers include everything below them.

Capability✓ in scope · — not at this tier PilotOne prep-software seat ProductionProduction prep-software integration AdvancedMultiple offices / entities
Introduced at Pilot
Populating to your documents and rules
Preparer review
Preparation-accuracy baseline
Introduced at Production
Reporting by return type
Review workflow in your systems
Approved write-back
Prep-software integration
Introduced at Advanced
Multi-jurisdiction rules
Multi-stage preparer approvals
High return volume
Multi-state return controls
Build price From $5,000 From $8,000 Custom quote
Final build priceConfirmed after discovery based on integrations, workflow complexity, return volume, review controls and deployment requirements.
Separate from buildBuild pricing is separate from recurring Agent Care, which covers managed monitoring, evaluations, incidents and verified improvements after launch.

What we need from you

What you bring, and what we build with it

Each input maps to a piece of build scope and a week in the delivery timeline.

You bringWe build with it
01Your prep-software chart and entity structure Document ingestion and entry mappingWeek 1
02Representative prior-year returns Roll-forward baseline and entry-to-document bindingWeek 2
03Your firm's elections and due-diligence policies Election, apportionment and boundary-definition mappingWeek 1
04Access to relevant APIs, feeds or exports Prep-software and document-source assessment, then integration setupWeek 2
05Positions you would not want signed Credit cases and failure-mode testingWeek 4
06What no return may assert unreviewed Confidence scoring, flag routing, guardrails and review controlsWeek 3
07Named signing preparers to review drafts Preparer review workflow, then pilot and production validationWeeks 5–6
Nothing else is required Deployment, documentation and Agent Care handover are ours.

Delivery timeline

Four phases across six weeks

Bands follow the real work rather than the plan, which is why evaluation and pilot share week 5.

Phase W1W2W3W4W5W6
Discovery W1
Build W2 – W3
Evaluate W4 – W5
Pilot & Launch W5 – W6
Week focus W1Return workflow discovery, election mapping and the automation boundary W2Source integration and the roll-forward baseline W3Preparation workflow, confidence logic and review controls W4Position and due-diligence cases, guardrails and failure-mode testing W5E-file integration, pilot returns and targeted corrections W6One filing cycle prepared under the signing preparer, then handover
Reading the bandA band covers the weeks its work is named in, and no others. The week 5 overlap is real, not padding.
At the end of W6Once the cycle validates, Agent Care owns the running agent.
DurationSix-week plan shown · typical delivery 4–6 weeks depending on scope confirmed in discovery.

Next step · Accounting AI agent

Build a return-preparation agent around your firm's review chain.

Show us your source documents, your prep software and who signs. The position, the signature and the transmission stay with your preparer of record.

Nestack Agents · Tax-return preparation & reviewAGT-ACC-16 · Agent Care available after launch