Evaluation 10 min read

AI in regulated professional work, Mid-2026

One structure links family law, valuation, and building inspections: a signed document others rely on. How each field's regulator answered the AI question.

A family lawyer’s affidavit, a valuer’s report, and a building inspector’s pre-purchase report have the same structure. A professional attends to the facts, applies judgement, and signs a document that someone else will rely on: a court, a lender, a purchaser at the moment of maximum commitment. The name at the bottom is the unit of accountability. The professional indemnity attaches to that name and to no other.

That structure is why these three professions are worth reading together in mid-2026. Each of their regulators has now been asked the same question (what happens when AI touches a document a professional has signed?) and each has answered differently. The courts wrote rules and enforced them. RICS wrote principles and made them mandatory. Building inspection in Australia has, so far, no answer at all. The distance between those three answers, when the underlying technology is identical in all three cases, tells an operator most of what they need to know.

This note consolidates and updates the three separate industry notes we published in June 2026. The evidence base for every claim below is the State of AI hub and the full Mid-2026 literature review, which surveys ten capability categories across applied AI and went through three independent fact-check passes before publication.

What do family law, valuation, and building inspection have in common?

In each of them, a professional’s name attaches to a document that someone with no obligation of charity will rely on. A family law affidavit is read by opposing solicitors who are paid to find errors, and by judicial officers under time pressure. A valuation is read by a lender, an owner, or an opposing expert who will test the comparable evidence. An inspection report carries a purchasing decision and is read, when things go wrong, by a tribunal. In all three, the reader’s reliance runs to the professional’s name and credentials, not to any tool that contributed to the document’s production.

The liability profile has the same shape too, and it is uneven in the same way. Inspection liability concentrates around the defects that were present, within scope, and missed, not the concealed ones a competent inspector could not have seen. A legal citation that does not exist is a categorical failure, not a marginal one. A valuation figure on contested property gets tested line by line by an opposing expert. In each field, a tool that is excellent at the easy cases and weak at the hard ones inverts the risk profile the professional actually carries.

That is the frame in which every AI question in these professions has to sit. Not “is the tool impressive” but “what reaches the signed document, and who verified it on the way”.

How has each regulator answered the AI question?

Three different ways: the courts wrote rules and have enforced them, RICS wrote principles and made them mandatory, and building inspection in Australia has no answer yet.

The courts: rules, with enforcement. On 16 April 2026 the Federal Court issued practice note GPN-AI, requiring the practitioner responsible for a document to confirm that legal authorities cited in it exist and support the propositions stated, alongside the facts and evidence relied on. NSW had already prohibited generative AI for affidavits and expert reports without leave from February 2025. Victoria, Queensland, and South Australia have parallel notes and guidelines in force. The enforcement is not hypothetical. In late 2025, in Mertz & Mertz (No 3) [2025] FedCFamC1A 222, the Federal Circuit and Family Court ordered a solicitor to pay $10,000 in costs thrown away after AI-fabricated citations were filed in a list of authorities, and referred the solicitor, along with two counsel, to their state conduct regulators. The question “is AI use permitted in legal work” has moved from open to answered: permitted, with a verification obligation that attaches personally to the practitioner.

Valuation: principles, made mandatory. RICS made its responsible-AI standard mandatory for regulated surveyors on 9 March 2026. The standard does not single out automated valuation models or defect detection by name. It requires risk-based governance, client disclosure of material AI use, and clear human accountability for the outcome, for any AI application a regulated surveyor relies on. In Australia, the API and the state property regulators have not yet matched RICS with a binding professional standard, but the direction of travel is visible, and the sensible planning assumption is that the Australian floor rises to meet the RICS one. A valuation-specific reading of the evidence, including the strongest Australian study and the build rules that follow from production work, is in AI in property valuation: the evidence and the design rules.

Building inspection: silence, so far. No Australian regulator has issued a binding position on AI in inspection work. The only legislated pathway anywhere is in the United States, where Florida permitted software-based plan review via House Bill 683, effective 1 July 2025. No independent accuracy evaluation of the plan-review products in that market (CodeComply.AI, CivCheck, PlanCheckPro.AI) has been published, and the regulatory pathway in Australia has not opened. Silence is not permission. It means the standard an inspector will eventually be measured against is being set now, elsewhere, and the legal profession’s lesson is available to be learned privately rather than publicly.

What does the independent evidence say about the tools?

In all three fields, vendor accuracy claims run ahead of the independent evidence, and the failure modes concentrate exactly where the liability does.

Legal AI has the most independent measurement, because the products are the most mature. Harvey raised US$200 million at an US$11 billion valuation in March 2026 and Thomson Reuters announced one million professional users of CoCounsel in February 2026; these are serious commercial products in wide use at sophisticated firms. Against that, the preregistered Stanford RegLab study, published in the Journal of Empirical Legal Studies in 2025, found Lexis+ AI hallucinating in 17 per cent of responses and Westlaw’s AI research product in 33 per cent, both marketed at the time with hallucination-free language. Those figures come from a peer-reviewed study by an institution with no commercial interest in the result, which makes them the credible baseline for any purchasing decision.

Valuation has the most honest vendor figure. CoreLogic’s own account, given by a company spokesperson to trade media rather than in a formal published report, puts almost 90 per cent of its Australian AVM estimates within 15 per cent of sale price since early 2024. That is a vendor claim, not an independently audited figure, but it is the most credible accuracy anchor in the Australian market. A widely circulated figure of “94.2 per cent within 5 per cent” could not be traced to any credible source, vendor or independent, and should be treated as marketing. Read plainly, the honest figure supports automated valuation as a triage and sanity-check layer, and does not support it as the final word on any contested property.

Building inspection has the most dangerous distribution. Vendor case studies for defect detection from imagery converge on roughly 90 to 95 per cent accuracy for visually distinct defect classes such as water staining and roof membrane damage; no independent benchmark confirms that range at the precision vendors quote it. (What video-capable models can and cannot actually see is surveyed separately in AI and video, Mid-2026.) The same systems perform poorly on subtle, internal, or occluded defects, and those are precisely the defects that produce inspection liability. The technology is strongest at the defects an inspector would have found easily and weakest at the defects an inspector might have missed. The exposure profile is the opposite of what the marketing implies.

Two cross-cutting findings complete the picture. First, the best independent evidence on grounded summarisation, the capability underneath every “ask your firm’s documents a question” tool, still shows leading systems hallucinating in roughly 13 per cent of responses, with omission rather than fabrication as the dominant failure mode. An omitted fact in family violence material, or an omitted qualification in a post-inspection answer to a client, is not a recoverable error. Second, any vendor whose published evidence rests on a single secondary source or an unspecified internal benchmark is showing you a procurement red flag, not a capability.

Where does AI fit when a signature carries the liability?

Upstream of the signature: assembling inputs, drafting, and retrieval, with verification built into the routine and the professional’s judgement authoritative on everything that reaches the document.

The defensible pattern is the same in all three professions. Research and retrieval, with the professional reading the underlying source before relying on it: the solicitor reads the authority before citing it, the valuer reads the comparable evidence, the inspector’s eye stays authoritative on what the photographs mean. First-cut drafting from structured input, with the professional editing rather than approving wholesale: routine legal documents from instructions and precedent, valuation reports from a structured field record, inspection reports assembled from the inspector’s on-site narration. The closest validated analogue for that last capability is clinical AI scribes, where a multi-site study across five academic health systems found around 16 minutes saved per 8 hours of patient care among adopters, with an error profile dominated by omission; both findings translate directly. And client communication templated against the firm’s own prior reasoning, so routine questions are answered consistently rather than from memory or from scratch each time.

The most underrated shift is also common to all three: knowledge synthesis at the firm level. Current frontier models can hold an organisation’s accumulated material in working memory at a scale they could not eighteen months ago. A boutique family law firm with twenty years of file notes, submissions, and consent orders; a valuation practice with two decades of reports and expert witness statements; an inspection firm with years of completed reports and twelve months of post-inspection client questions on every job: each is sitting on a corpus no off-the-shelf product can replicate, because no off-the-shelf product has access to the firm’s own work. The value is not that a tool replaces senior judgement. It is that the firm’s accumulated judgement becomes legible to every professional in the firm, on demand, subject to the same verification discipline as every other intermediate input.

What should a principal do while the floors keep rising?

Build the verification routine now, before it is mandated, and make it something you could describe to a judge, a regulator, or a tribunal without embarrassment.

For legal practitioners this is no longer optional. GPN-AI has made verification a professional obligation, and the costs order and conduct referrals in Mertz showed the courts are prepared to act on it. For valuers, the RICS standard is the visible shape of what is coming, and a practice that adopts risk-based governance and client disclosure before the Australian floor rises will find the transition uneventful. For inspectors, the absence of a rule is the opportunity: a firm that can describe in plain language what AI does and does not do in its inspection process is writing the standard it will eventually be measured against, on its own terms.

The operational moves are the same everywhere. Be specific, internally, about which tools are in use, on what kinds of matters or jobs, with what verification routine attached, and be willing to make that visible to clients. Build the verification step into the workflow, not into individual discretion. Treat the firm’s own material as the asset most worth making legible before licensing anything. And on timing: the frontier is moving quickly enough that nobody needs to be first, and slowly enough that eighteen months without forming a view leaves a practice visibly behind a peer who spent those months running small structured pilots and learning what to trust.

The question worth working through is the same in all three professions. Not “which AI product should we license” but “what is the verification routine around any AI-assisted output, what does our accumulated experience look like once every professional in the firm can ask it a question, and how would we describe both to the court, the regulator, or the tribunal that eventually asks”. We work through exactly that with professional services practices in Perth. If those sound like the right questions for your practice, start with a conversation.

Questions practices ask

Can Australian lawyers use AI to draft court documents? Yes, with obligations attached. Federal Court practice note GPN-AI requires the responsible practitioner to confirm that cited authorities exist and support the propositions stated, and NSW prohibits generative AI for affidavits and expert reports without leave of the court. Verified AI-assisted drafting is permitted; unverified output filed with a court has already produced costs orders and conduct referrals.

Are automated property valuations accurate enough to replace a valuer? No. The most credible Australian figure is CoreLogic’s own, which puts almost 90 per cent of its AVM estimates within 15 per cent of sale price, and it is a vendor claim rather than an independent benchmark. That supports using AVMs for triage and sanity checks, not as the final word on contested property, where a valuer weighs the model output as one input among several.

Can AI reliably detect building defects from photographs? Only the easy ones. Vendor case studies converge on roughly 90 to 95 per cent accuracy for visually distinct defects such as water staining, with no independent benchmark confirming that precision, and performance drops on subtle, internal, or occluded defects. Those are the defects that produce inspection liability, so the systems work as a second pass on photographs already gathered, not as a substitute for a trained inspector.

Do professionals have to tell clients when AI was used? It depends on the regulator. RICS requires regulated surveyors to disclose material AI use to clients under its standard, mandatory since 9 March 2026. Australian courts require practitioners to verify AI-assisted material rather than disclose it in every instance. Where no rule exists yet, being able to describe what AI did and did not do in your process is the posture that ages well.

Published 29 August 2026

Perth AI Consulting delivers AI opportunity analysis for small and medium businesses. Start with a conversation.

Prepared by Claude, directed and approved by PAC.

More from Thinking

Evaluation 11 min read

AI in property valuation: the evidence, the design rules, and what it could become

The best Australian evidence on vision AI in valuation measures a different task than the one vendors demo. The findings, and the design rules that follow.

Evaluation 7 min read

Eleven cells moved. Here is what they mean for your business.

Reading the September 2026 State of AI verdict table: what improved, what declined, and what to do differently this quarter.

Evaluation 7 min read

Competitor intelligence for small business: what AI can and cannot see

What AI-assisted competitor intelligence really is for a small business: the public sources worth watching, what they cannot tell you, and the legal line.

Technical 9 min read

The business knowledge base: evidence, risks, and how to build one

What a business knowledge base actually is, what the evidence says it delivers, the security and privacy realities, and how we build one that holds up.

Evaluation 8 min read

What AI can see in your customer data (and what it cannot)

What AI can genuinely find in the customer records an SME already holds, what it cannot, and when a spreadsheet honestly beats a model.

Building 7 min read

What an AI quoting engine actually does

What an AI quoting engine takes in, what it drafts, what the evidence says about accuracy and speed, and why the final price stays with a human.

Adoption 6 min read

Australia's AI adoption gap is bigger than the 12% headline suggests

ABS says 12% of Australian businesses use AI. The real story is 35% of large businesses against 11% of small ones, and the barrier isn't the technology.

Building 7 min read

Why we let AI run the interviews (and why we never let it pretend to be human)

AI-conducted interviews compress weeks of stakeholder discovery into days, standardise what gets asked, and lower the guard that distorts honest answers.

Adoption 14 min read

How AI capability actually moves through a business

The decisive variable in SME AI adoption is the human absorption sequence, not the tooling. A working framework from observation across WA businesses.

Evaluation 7 min read

AHPRA advertising rules for psychologist websites

Recovery stories, 'specialist', 'clinical psychologist', and endorsement titles are where psychology sites breach the National Law. A practical read-through.

Adoption 4 min read

Customer service AI has finally grown up

Chatbots and AI receptionists earned their bad reputation. What changed, and how the mature version answers every call without replacing anyone.

Evaluation 6 min read

Who can use the titles 'Dr', 'Specialist', and 'Surgeon'?

AHPRA restricts 'specialist' and 'surgeon' to specific registrations, and 'Dr' has its own rule. What health practice websites can and cannot claim.

Adoption 5 min read

Your best people hate writing reports

The operators you promote are brilliant at the work and allergic to reporting. A scheduled AI call interviews them, drafts the briefing, they approve it.

Building 6 min read

Your website isn't just for humans anymore

How to build a chatbot that keeps itself up to date, can't leak client information, and won't answer beyond what you've published.

Evaluation 7 min read

Can you show Google reviews on your health practice website?

AHPRA bans clinical testimonials, even true ones, but service reviews are fine. What that means for the Google reviews widget on your practice site.

Evaluation 7 min read

What AHPRA's advertising rules mean for your website

Your practice website is advertising under the National Law. What AHPRA's rules prohibit, who is responsible, and how to check your own site.

Evaluation 8 min read

Is it safe to paste client data into ChatGPT?

Short answer: it depends on one setting, and most people have it wrong. What ChatGPT, Claude and Copilot do with your data, and what the Privacy Act expects.

Evaluation 4 min read

What a good AI audit actually delivers

The audit report named one recommendation specific enough to check, and what the Build that followed looked like: one real engagement, generalised.

Evaluation 7 min read

AI and video, Mid-2026: the models can watch now, not just listen

AI could always transcribe video. It can now read the frames as well, and every hour of footage a business owns becomes something it can question.

Building 7 min read

Case study: a 119-page AML/CTF program in three days

How we built a seven-document AML/CTF compliance pack for a small accounting practice in three days, working from 31 confirmed assumptions.

Building 11 min read

From evidence base to delivery: a production AI methodology

How we delivered 34 evidence-anchored AI briefings to a WA peer-advisory chapter: fact-checked literature review, multi-agent verification, one method.

Technical 9 min read

The six functions of a working AI system

A working AI system is six functions doing six jobs. When all six connect, hallucinations get caught, outputs hold steady, and models become swappable.

Technical 7 min read

Supervised autonomy: the middle path for AI architecture

Between drafts you approve and agents you hope about sits the middle path: an envelope of authorised routine work, supervised, audited, and yours to widen.

Evaluation 5 min read

The state of applied AI in Mid-2026

Our literature review of applied AI in mid-2026: ten capability categories, three fact-check passes, written for operational leaders.

Technical 9 min read

How to design a PHI redaction system for clinical AI

PHI redaction is part of a clinical AI tool's architecture, not a feature you add. What the literature says it should look like, and how we built it.

Building 9 min read

How we built on-device de-identification so AI never sees real names

Most AI privacy is a policy. Ours is architecture: an NER model runs in the browser and strips names before anything leaves the device.

Technical 7 min read

Your agency's clients are about to ask why this costs so much

A solo consultant built in three weeks what your agency quoted twelve for. The client doesn't know why yet. The agencies that survive change what they sell.

Adoption 6 min read

What do you love doing? What do you hate doing?

Ask people what they love doing and what they hate doing, then show them AI is coming for the second list. Why the reframe works, and how it fails.

Technical 7 min read

Why I don't use n8n (and what I do instead)

n8n demos well. But a compelling demo and a reliable production system are different things, and the distance between them is where businesses get hurt.

Technical 10 min read

Your codebase was not built for AI. That's the actual problem.

Amazon's mandatory meeting about AI breaking production is an architecture story: codebases built for human maintainers only, now maintained by AI.

Adoption 4 min read

Your team has AI licences. You don't have an AI system.

Fifteen people, fifteen separate AI accounts, no shared context. The problem isn't the tool; it's the architecture around it. Here's the fix.

Building 7 min read

Your $2,000 day starts the night before: our system keeps you on the tools, not on the phone

Optimised routes overnight, automatic customer notifications, and promises the system keeps or corrects. A scheduling system that protects your daily rate.

Evaluation 4 min read

The fastest way for an executive to get across AI

AI moves faster than any executive can track. One focused conversation, one written report, and a decision you can act on: your time stays on the business.

Building 6 min read

Your IT department will take 18 months. You need this working by next quarter.

Senior leaders know what they need built; the gap is time. A prototype gets the tool working now and hands IT a validated blueprint for later.

Building 8 min read

We built an AI invoice verifier. Here's where it hits a wall.

We built an AI invoice verifier and watched a fake beat a real invoice. Why document analysis alone cannot stop fraud, and the five layers that can.

Building 5 min read

How to build an AI chatbot that doesn't lie to your customers

Woolworths scripted its AI to talk about its mother. The business fix is honesty; the technical fix is architecture that prevents fabrication by design.

Technical 9 min read

Why AI safety features are load-bearing architecture, not political decoration

The 'woke AI' label came from real failures, but they were engineering failures, not safety failures. The difference matters wherever errors have consequences.

Adoption 3 min read

Woolworths' AI told a customer it had a mother. That's a problem.

Woolworths' AI assistant Olive was scripted to talk about its mother and uncle. When callers realised, trust broke instantly. The fix is honesty.

Evaluation 5 min read

Google is no longer the only way your customers find you

Customers now find businesses through ChatGPT, Perplexity, and Gemini. The sites AI cites are structured differently to the sites Google ranks.

Evaluation 6 min read

The personal workflow analysis: what watching a real workday reveals about automation

People describe the work they value, not the work that eats their time. Recording a real workday reveals the automation opportunities interviews miss.

Evaluation 11 min read

An AI audit that starts with your business

How an operations-first AI audit works: what it looks for, how the evidence is collected, what the report contains, and what it tells you to skip.

Building 6 min read

What production AI teaches you that demos never will

The gap between a demo and a working system is where the useful lessons live. Architecture, framing, privacy, adoption: the patterns repeat every time.

Adoption 6 min read

The psychology of why your team won't use AI

You buy the tool, run the demo, and three months later nobody is using it. Five predictable psychological barriers, each with a strategy that works.

Technical 4 min read

Stop telling AI what NOT to do (and what to say instead)

Instructions built on prohibitions make AI cautious and generic. Describing what you want instead transforms the output, and the reason comes from psychology.

Building 5 min read

How we turned generic AI into a specialist: and what that means for your business

Mediocre AI output is rarely the model's fault. Three structural changes that turn the same model from generic to specialist-grade.

Evaluation 6 min read

Your business has 9 customer touchpoints. AI can fix the 6 you're dropping.

You pay to get customers to your door, then lose them to missed follow-up. AI can handle the six touchpoints most businesses drop.

Technical 6 min read

What happens to your data when you press 'Send' on an AI tool

Businesses send customer data to AI tools without knowing what happens during processing. The spectrum of AI privacy is wider than you think.