search

How Generative AI Is Changing Tax Research and Compliance Workflows

5/25/2026

Tax research has a structure that generative AI does not respect, and understanding the mismatch is the whole subject.

Tax research runs on an authority hierarchy. The statute governs. Regulations interpret it. Rulings, procedures, and other published guidance carry defined weight. Cases carry weight depending on the court and the jurisdiction. Secondary sources — treatises, journals, blog posts — carry none, though they are useful for finding the authority that does.

A language model produces fluent text at a uniform confidence level regardless of where in that hierarchy its content came from, whether the source was authoritative or a message board, whether the provision has since been amended, or whether the citation exists at all.

That is not a limitation to work around. It is the defining characteristic of the tool, and everything below follows from it.

Why Fabricated Citations Are Worse in Tax Than Anywhere Else

Every profession has been warned that these models invent sources. In tax practice the consequence is specific and severe, and it is worth stating plainly.

The research file is the penalty defense.

Whether a position is supported — and whether the preparer and the taxpayer are protected from penalties — depends on the existence and weight of actual authority. The standards governing return positions turn on what authority exists and what it says. A preparer who relied on a citation that does not exist has not merely failed to find support; they have no support, and the file evidences a failure of diligence rather than its exercise.

That is a worse position than having done no research at all. An unresearched position is an unresearched position. A position supported by a fabricated citation, in a file, with a date on it, is a documented failure to exercise due diligence — and the practitioner conduct rules impose diligence and competence obligations independently of the penalty provisions.

The operating rule that follows is absolute: no citation enters a memo, a workpaper, or a return position unless the practitioner has read it in the primary source. Not a summary of it. Not the model's characterization of it. The source.

Where It Genuinely Helps

The tool is useful, and dismissing it costs real time. Seven applications, in descending order of value.

  1. Issue spotting and framing. Give it a messy client fact pattern — de-identified — and ask what questions a tax adviser should research. This is its best use: it produces breadth quickly, it surfaces the issue you had not considered, and being wrong costs nothing because the output is a research agenda rather than an answer.
  2. Orientation in an unfamiliar area. When you need to research something outside your usual practice, the hardest part is not knowing the terms of art well enough to search effectively. A model will give you the vocabulary, the general shape of the area, and the names of the concepts — after which you can search authority properly.
  3. Adversarial testing of your conclusion. Give it your position and ask it to argue against it, identify the weakest link, and state what facts would change the answer. The model has no stake in your conclusion, and since you are asking for arguments rather than authority, its citation weakness does not bite.
  4. Summarizing a document you supply. Grounded summarization — pasting a long ruling, a lengthy agreement, or a plan document and asking for a summary — is far safer than asking the model to recall something, because the source is in front of it. Verify against the document, and this saves genuine time.
  5. Drafting the memo structure once you have the authority. You supply the analysis; it supplies organization and prose.
  6. Explaining a provision you have already read, in plain language, as a check on your own understanding. If its explanation diverges from yours, one of you has misread something and it is worth finding out which.
  7. Translating a conclusion for a client, with instructions not to add anything you did not say.

Notice what is not on that list: answering the tax question.

Where It Fails, Specifically in Tax

Citations. It invents statutory sections, regulation numbers, revenue procedure and ruling numbers, and case names — often in plausible formats, attached to propositions that sound right.

Currency. Provisions that were amended, expired, sunset, or superseded. Inflation-adjusted amounts. Effective dates. A model will state a rule that was true in some prior period with no indication that it has changed, and tax is a field where a substantial share of the content has a date attached.

State and local law. Thin, inconsistent, and unreliable across jurisdictions. Anything multistate requires primary research.

Provisions interacting. Tax outcomes frequently depend on several provisions operating together, and this is where models produce confident answers that are individually plausible and jointly wrong.

Arithmetic on client figures. Do not use it to compute.

Distinguishing authority from commentary. It trained on both and it does not reliably signal which it is reproducing.

Knowing when it does not know. The absence of expressed uncertainty is the core risk, and it is why the verification step cannot be discretionary.

The Verification Chain

A workflow that captures the benefit and closes the exposure. Six steps, in order.

  1. Frame with the model, never answer with it. Produce the issue list, not the conclusion.
  2. Search actual authority in a tax research platform or the primary sources directly.
  3. Read the primary source. The statute, the regulation, the ruling — not an editorial summary, and not the platform's headnote.
  4. Check currency. Amendments, effective dates, later guidance, and whether the provision has been superseded or its regulations reproposed. This is the step most often skipped, and the one where a model's output is most likely to have been stale.
  5. Document what you read, where, and when — with the verified citation.
  6. If the model suggested an authority, verify two things: that it exists, and that it says what was claimed. Both. A real citation attached to a proposition it does not support is the failure mode practitioners least expect, and it is common — the model retrieves a genuine section number and characterizes it wrongly.

Retrieval-Grounded Tools Are a Different Category

Worth distinguishing, because the market conflates them.

An open-ended chatbot generates text from what it learned during training. It has no source for any particular statement and no ability to show you one.

A retrieval-grounded research tool searches an actual database of tax authority, retrieves documents, and generates an answer citing what it retrieved. That is a materially better architecture for this work, because the cited documents exist and can be opened.

It does not remove the verification step. Retrieval can surface a real authority that is irrelevant to the question, or relevant and superseded, and the generated summary can still mischaracterize what was retrieved. What changes is that verification becomes fast — the document is one click away — rather than a search from scratch.

The practical guidance: prefer tools that show you the source, and treat the citation list as the useful output rather than the prose above it.

Confidentiality Comes First

Repeated from our post on AI use in accounting practice because the exposure is severe and the mistake is easy.

Do not enter client information into a consumer AI tool. Professional confidentiality obligations apply regardless of the technology, and for tax practitioners the separate statutory restriction on disclosure and use of tax return information requires client consent in a prescribed form, with penalties for violation. Pasting a client's return data into a public chatbot to research their issue is, on a plain reading, a disclosure without the required consent.

The workable controls: an enterprise or business tier with contractual terms addressing data use and processing location, de-identification before input — which works fine, because the model is helping with framing and language rather than with the client's identity — and a written firm policy that says which tools are permitted for what.

What the Research File Must Contain

The penalty-protection point, made concrete. A defensible research workpaper contains:

The question, stated precisely, with the relevant facts as understood at the time.

The authority relied upon, cited specifically, with the relevant text attached or excerpted — not described. Attaching the provision is the single most valuable habit in tax research documentation, because it proves what was actually read.

The currency check — the date of the research, the source consulted, and confirmation that the authority was current as of that date.

The analysis, showing how the authority applies to these facts, including the contrary authority considered and why it was distinguished.

The conclusion, and the level of confidence.

Who performed it and when, and who reviewed it.

"I researched it and it's fine" is not a research file. Neither is a memo citing authority that nobody attached. And a memo citing authority that does not exist is affirmatively harmful — which is the whole reason the verification step matters.

Structured coverage is available through the AI for Accountants Certificate Program, AI Applications for Accountants, the AI for Accountants Strategy and Research Specialist series, tax practitioner regulations, penalties, and security, and ethics training and professional conduct for accounting and tax professionals.

Firm Policy for Tax Research Specifically

Beyond a general AI policy, tax practice needs four rules:

No unverified citation, ever. Written down, trained on, and enforced in review. The reviewer's job includes checking that cited authority was attached.

Approved tools named, with retrieval-grounded research tools distinguished from general chatbots and with consumer tiers prohibited for client information.

Attach, don't describe. A firm standard that research memos include the relevant authority text. This makes the verification failure visible in review rather than discoverable on examination.

Currency is part of the research. A memo without a research date and a confirmation of currency is incomplete, because tax authority has a shelf life.

Where the Time Is Actually Saved

An honest accounting, because the productivity claims in this area are inflated.

Genuinely saved: issue framing, orientation in unfamiliar areas, drafting structure, client-facing translation, summarizing documents you supply, and generating counterarguments. For a practitioner who writes a lot, this is real — plausibly a meaningful share of the non-reading time.

Not saved: the reading. Reading the statute, the regulation, and the guidance, and thinking about how they apply to these facts, is the work — and it is the part that cannot be delegated to something that cannot tell you when it is guessing.

Which is a reasonable trade. The tool removes friction from everything around the analysis and none from the analysis itself, and a practitioner who understands that boundary gets the benefit without acquiring the risk.

Where Practitioners Get This Wrong

  • Citing authority the model produced without reading it in the source — the failure that matters most
  • Verifying that a citation exists but not that it says what was claimed
  • Skipping the currency check, in a field where much of the content is dated
  • Asking the model for the answer rather than for the questions
  • Pasting client return information into a consumer tool
  • Relying on it for state and local law
  • Accepting a confident answer where several provisions interact
  • A research memo describing authority rather than attaching it
  • Treating a retrieval-grounded tool's summary as verified because the citation was real
  • No research date in the file, so currency cannot be established later
  • Believing the productivity claims and reducing the time budgeted for reading

The sentence to carry into any tax research workflow: the model is excellent at telling you what to look up and unreliable at telling you what it says. Used that way it is a real improvement. Used the other way it produces a file that documents a failure of diligence with a date on it.

Frequently Asked Questions

Why are fabricated citations especially dangerous in tax practice?

Because the research file is the penalty defense. Whether a position is supported depends on the existence and weight of actual authority, so a preparer who relied on a citation that does not exist has no support at all — and a dated file containing a fabricated citation evidences a failure of diligence rather than its exercise. That is a worse position than having done no research.

What is the best use of generative AI in tax research?

Issue spotting and framing. Give it a de-identified fact pattern and ask what questions a tax adviser should research. It produces breadth quickly, surfaces the issue you had not considered, and being wrong costs nothing because the output is a research agenda rather than an answer.

What must be verified about an AI-suggested citation?

Two things: that the authority exists, and that it says what was claimed. Practitioners check the first and skip the second, but the common failure mode is a genuine section number attached to a proposition it does not support.

Are retrieval-grounded research tools safe to rely on?

Safer, not sufficient. A tool that searches an actual authority database and cites what it retrieved is a materially better architecture, because the documents exist and can be opened. But retrieval can surface a real authority that is irrelevant or superseded, and the generated summary can still mischaracterize it. What changes is that verification becomes fast rather than unnecessary.

Can client tax information be entered into an AI tool?

Not a consumer tool. Professional confidentiality applies regardless of technology, and the separate statutory restriction on disclosure and use of tax return information requires client consent in a prescribed form with penalties for violation. The workable approach is an enterprise tier with contractual data terms, de-identification before input, and a written firm policy.

What should a defensible tax research file contain?

The question and the facts as understood, the authority relied upon with the relevant text attached rather than described, a currency check with the research date and source, the analysis including contrary authority considered, the conclusion and confidence level, and who performed and reviewed it. Attaching the authority is the habit that both proves what was read and makes a verification failure visible in review.

CPATrainingCenter.com 9715 Rod Road Suite A Alpharetta, GA 30022 1-770-410-1219 support@CPATrainingCenter.com
Certifications CPA CFP Enrolled Agent Payroll
Licensing & Events Securities Insurance Webinars Seminars
Stay Up To Date
Need Training Or Resources In Other Areas? Try Our Other Training Center Sites:
HR Banking Financial Services Insurance Mortgage Payroll Real Estate Safety
Training By Delivery Format & Subjects Covered:
Special Promotions Online Training Resource Materials Seminars Webinars All CPA/Accounting Subjects
Facebook Copyright CPATrainingCenter.com 2026