Can ChatGPT make serious legal or accounting errors?

Quick answer

Yes, frequently. A 2024 Stanford study showed ChatGPT hallucinates on 58-82% of US law questions depending on complexity, and regularly invents jurisprudence. In accounting, it confuses French/Belgian charts of accounts, forgets local tax specifics, suggests wrong entries. Absolute rule: never sign an AI-generated legal or accounting document without qualified human validation.

Can ChatGPT make serious legal or accounting errors, on the ground?

Consumer LLMs are particularly weak on law and accounting for four reasons. (1) Corpus, training data contains little recent European jurisprudence, even less Belgian law. (2) Local specificity, a US LLM poorly knows the Belgian PCG, French intra-EU VAT, sectoral collective agreements. (3) Response pressure, asked about non-existent jurisprudence, ChatGPT invents plausible ones rather than admit absence. The Mata v. Avianca case (USA 2023): a lawyer submitted 6 ChatGPT-invented decisions, disciplinary sanctions followed. (4) No updates, Belgian law evolves quarterly, GPT-4 remains frozen in 2023. "Sector-specific tools like Harvey AI, Doctrine.fr or Wolters Kluwer Co-Pilot are built for legal and accounting work, with a verified domain corpus and sourced citations," adds Lorenzo Eeman, founder of PROEMA. SME rule of thumb: ChatGPT to brainstorm a question, human expert to settle it.

What the 2026 numbers say on Can ChatGPT make serious legal or accounting errors

Public benchmarks converge on three signals. ChatGPT hit 900 million weekly active users in early 2026 (OpenAI announcement reported by TechCrunch on February 27, 2026). Google AI Overviews reached 47 % of European queries in March 2026 (Semrush Sensor 2026). Perplexity reported +800 % year-over-year query growth. In practical terms: informational traffic leaving Google's blue links for answer engines is no longer marginal, for a B2C F&B site, it typically runs 15-25 % of measurable traffic via Cloudflare AI Crawl Control or GA4 « ai-referrer » segments.

Why Can ChatGPT make serious legal or accounting errors isn't optional for serious brands

The 5W Citation Source Audit Q1 2026 shows LLMs concentrate citations on a tiny set of sources: Wikipedia (13.15 % at ChatGPT) + Reddit (11.97 %) = 25 % of citations, followed by vertical databases (Yelp, TripAdvisor, IMDB depending on context). For F&B brands, the problem is binary: either you're in the sources LLMs read, or you never show up, there is no « page 2 » of LLM citation. PROEMA's documented discipline targets exactly this presence: structure content via Schema.org, publish on hubs crawlers actually read, and lock down Author/Person + sameAs Wikidata to clear the confidence filter.

PROEMA operational rule for Can ChatGPT make serious legal or accounting errors

Translation: stop watching from the bench. By June 2026, a B2C F&B brand with no Schema.org Author/Person, no sameAs Wikidata, and no FAQPage gets approximately zero LLM citations on long-tail informational queries, confirmed across PROEMA verticals (expertvin.be, expertcafe.be, zeroproof.one). The fix isn't theoretical: it's three concrete deliverables (Schema markup audit + Wikidata entry + FAQ playbook 5-blocs structure) executed in six to eight weeks.

At a glance
TaskGPT-4Claude 3.5Dedicated tools
US jurisprudence citation58-82% invention35-45%<5% (Harvey)
EU multi-country VAT20-30% error15-25%<2% (accounting SW)
Template contract draft40% obsolete clauses30%<10% (LegalTech)
Jurisprudence synthesis25-40% imprecise20-30%<5% (Doctrine.fr)