← Research
Research

How will AI change law?

The citation is the profession's unit of trust, and it is the thing generative models are worst at. Everything else in this follows from that.

Last reviewed: 28 August 2026

What the hallucination case database now records, what the first preregistered evaluation of Lexis and Westlaw found, what the English Divisional Court settled in Ayinde and Al-Haroun, and the trainee rung nobody is measuring.

Question this page answersAll 780 questions this research covers

By moving the cost of verification onto the person least equipped to carry it, and by removing the early-career work through which lawyers learned to carry it. Law runs on a currency that generative models counterfeit convincingly and cheaply: the citation. A fabricated authority has a name, a year, a court and a neutral citation number. The form is correct when the content does not exist. That single property explains most of what has happened to the profession since 2023.

This page uses the evidence base that law has and other professions do not: a public, dated, growing record of what goes wrong when the checking fails, maintained by an outsider and now cited by courts.

Nearly two thousand decisions, and the largest group is not lawyers#

Damien Charlotin's AI Hallucination Cases database tracks decisions in which a court or tribunal has explicitly found or implied that a party relied on hallucinated material. It excludes mere allegations. As at its update of 27 August 2026 it recorded 1,963 cases, the earliest from the second quarter of 2023.

By jurisdiction: United States 1,345, Canada 214, Australia 98, United Kingdom 62, Israel 57, with more than thirty other countries represented. By nature of the defect: fabricated material in 1,634 entries, misrepresented authority in 816, false quotations in 528.

Then the number that reframes the story. By the party responsible: self-represented litigants 1,127, lawyers 784, judges 29, expert witnesses 15.

The dominant coverage of this problem is professional embarrassment, and the majority of the cases are not lawyers at all. They are people without lawyers who found a tool that produced something that looked like a legal argument. That is an access-to-justice phenomenon wearing the costume of a professional scandal, and the two need different remedies. The database owner is careful that this counts only decisions where the court addressed the point, so the true universe is larger.

The 29 judges are the entry to sit with. Fabricated authority has reached the bench itself.

The tools sold as hallucination-free#

Magesh and colleagues at Stanford ran the first preregistered empirical evaluation of commercial AI legal research tools, published in the Journal of Empirical Legal Studies in 2025. Their stated reason for doing it was that providers had described retrieval-augmented generation as eliminating hallucinations, or had guaranteed hallucination-free legal citations.

They wrote over 200 legal queries across four categories, preregistered the dataset with the Open Science Foundation before running anything, and graded responses on whether they were both correct and grounded in the sources cited.

One mechanism in the paper matters more than the headline rates. Westlaw produced the longest answers, averaging 350 words against 219 for Lexis+ AI and 175 for Ask Practical Law AI. More words means more falsifiable propositions and more chances to be wrong, and it also means more to check. The paper puts it directly: lengthier answers require substantially more time to check, verify and validate, because every proposition and citation has to be independently evaluated.

That is the trap in one sentence. The tool that produces the most useful-looking output imposes the most verification, and the verification is invisible, unbilled and easy to skip. This research has called the general form of that problem the verifier's discount: the work of checking is systematically undervalued relative to the work of producing.

What the English courts settled in June 2025#

In Ayinde v London Borough of Haringey and Al-Haroun v Qatar National Bank [2025] EWHC 1383 (Admin), decided on 6 June 2025, the Divisional Court, Dame Victoria Sharp P and Johnson J, dealt with two referrals under the Hamid jurisdiction and used them to state the position for the profession.

On what the tools can do, at paragraph 6:

Freely available generative artificial intelligence tools, trained on a large language model such as ChatGPT are not capable of conducting reliable legal research. Such tools can produce apparently coherent and plausible responses to prompts, but those coherent and plausible responses may turn out to be entirely incorrect. The responses may make confident assertions that are simply untrue. They may cite sources that do not exist. They may purport to quote passages from a genuine source that do not appear in that source.

On the duty, at paragraph 7, lawyers using such tools must check the accuracy of the research by reference to authoritative sources before using it, and the court names them: the government legislation database, the National Archives judgments database, the official Law Reports, and the databases of reputable legal publishers. At paragraph 8 the duty extends to lawyers relying on the work of others who used AI, on the same footing as relying on a trainee or a pupil.

The facts underneath. In Ayinde, grounds of claim cited five authorities that do not exist; one of them carried a real neutral citation number belonging to an entirely different case. In Al-Haroun, a claim seeking damages of £89.4 million, a schedule of forty-five citations was put before the court of which eighteen did not exist, and of those that did, many contained neither the quotations nor the propositions attributed to them. One of the fabricated authorities was attributed to the very judge hearing the application.

The holding with the widest practical reach is at paragraph 81:

A lawyer is not entitled to rely on their lay client for the accuracy of citations of authority or quotations that are contained in documents put before the court by the lawyer. It is the lawyer's professional responsibility to ensure the accuracy of such material.

At paragraph 23 the court set out the full range of its powers: public admonition, a costs order, a wasted costs order, striking out, referral to a regulator, contempt proceedings, and referral to the police. At paragraph 31 it added that, save in exceptional circumstances, admonishment alone is unlikely to be a sufficient response. In Ayinde it found the threshold for contempt proceedings met and declined to initiate them, giving five reasons including that the barrister concerned was extremely junior and apparently operating outside her level of competence, and stating that the decision is not a precedent.

The paragraph the profession has under-read#

Paragraph 9 does something the coverage largely missed. It puts the obligation on people who did not touch the tool.

practical and effective measures must now be taken by those within the legal profession with individual leadership responsibilities (such as heads of chambers and managing partners) and by those with the responsibility for regulating the provision of legal services... For the future, in Hamid hearings such as these, the profession can expect the court to inquire whether those leadership responsibilities have been fulfilled.

Read alongside the reasons given for not pursuing contempt in Ayinde, where the court referred to the regulator the question of whether those supervising the barrister's pupillage had complied with requirements as to supervision, work allocation and competence, the direction is clear. The court treated a fabricated citation as a supervision failure with a junior at the end of it, and said future hearings will ask who was supervising. That is the same accountability question this research has put to leadership elsewhere, in who supervises work they cannot do.

The regulator's answer was to forbid the machine from proposing case law#

On 6 May 2025 the Solicitors Regulation Authority authorised Garfield.Law Ltd, described as the first purely AI-based firm permitted to provide regulated legal services in England and Wales. It handles small claims for unpaid debts up to £10,000.

The conditions are the interesting part. Among them: the user must approve each stage, named regulated solicitors remain accountable for all system outputs, and the AI is precluded from proposing case law.

A regulator willing in principle to authorise a firm with no human fee-earners drew its line at exactly the function the empirical evidence says is unreliable. That is a more sophisticated regulatory response than either the prohibitionist or the permissive reading suggests, and it points at where the profession is heading: not a ban on AI, but a boundary drawn task by task around the operations where fabrication is both likely and consequential. The general version of that boundary is set out in the Delegation Boundary Map.

The rung that is disappearing, and the number that does not exist#

The work most exposed in legal practice is document review, first drafts, research memoranda and citation checking. It is also, historically, how a junior lawyer learned what an authority is worth: by reading fifty of them badly, being corrected, and eventually developing the reflex that something is off.

Reading the Magesh findings next to the Charlotin database gives the uncomfortable version. The tools produce output whose defects are detectable only by someone with the judgement that used to be built by doing the work the tools now do. The verification requires the capability; the capability was built by the task; the task has been automated.

No published dataset tracks trainee solicitor or pupil barrister hours by task before and after 2023. This page does not have the number and will not manufacture one. The nearest available evidence is general rather than legal: Brynjolfsson, Chandar and Chen find employment among 22 to 25 year olds in highly AI-exposed occupations running about 19 per cent below where it would sit on the comparison trend, through reduced hiring rather than dismissal, and concentrated in occupations where AI substitutes rather than complements. Whether law sits in the substituting group is currently an argument rather than a measurement. The structural case is at missing rungs.

What remains unresolved#

What a firm or chambers can do this quarter#

Key research and primary sources

Judgment paragraphs are quoted from the published text on the judiciary website. Database figures were read from the source on the date stated and will drift, so the date is given.

On the underlying failure mode, what an AI hallucination is and why AI sounds so confident. On who carries the checking, who owns verification and the verifier's discount. On the boundary, the Delegation Boundary Map and meaningful human oversight. On the profession-level method, deskilling risk by profession, and on the neighbouring cases, medicine and consulting. On the junior end, missing rungs.

About this research#

Rahim Hirji is the author of SuperSkills (Kogan Page, 2026), keynote speaker on AI and human capability, and founder of The SuperSkills Intelligence Company. The Divisional Court judgment was read in full at the primary source and paragraph numbers are given for every proposition attributed to it. Database figures were taken directly from the source page on 28 August 2026 and reflect its update of 27 August 2026. No figure for trainee hours is given anywhere on this page because none could be found. This is commentary on the professional and evidential position, not legal advice.

How this research works  ·  Reviewed quarterly  ·  Found an error? Tell me and it is corrected on the page.

Cite this

Hirji, R. (2026). How will AI change law? The SuperSkills Intelligence Company. Last reviewed 28 August 2026. thesuperskills.com/research/how-will-ai-change-law

Questions answered on this page

How will AI change law?

By moving the cost of verification onto the person least equipped to carry it, and by removing the early-career tasks through which lawyers learned to carry it. The citation is the profession's unit of trust and it is the thing generative models are worst at producing reliably. Damien Charlotin's database recorded 1,963 court decisions worldwide involving hallucinated material as at 27 August 2026, from Q2 2023 onwards. The English Divisional Court has held that lawyers using AI have a professional duty to check the output against authoritative sources, and that responsibility extends to heads of chambers and managing partners.

How reliable are AI legal research tools?

Better than a general chatbot and far short of the claims made for them. Magesh and colleagues ran the first preregistered empirical evaluation of RAG-based legal research tools, published in the Journal of Empirical Legal Studies in 2025, using over 200 handwritten legal queries. LexisNexis's Lexis+ AI answered 65 per cent of queries accurately and hallucinated on more than one in six. Thomson Reuters's Westlaw AI-Assisted Research was accurate 42 per cent of the time and hallucinated on one-third of responses. Ask Practical Law AI gave incomplete answers to 62 per cent of queries. The paper was written because providers had described their systems as eliminating hallucinations or guaranteeing hallucination-free citations.

What did the court decide in Ayinde v Haringey?

In Ayinde v London Borough of Haringey and Al-Haroun v Qatar National Bank [2025] EWHC 1383 (Admin), 6 June 2025, the Divisional Court held at paragraph 6 that freely available generative AI tools trained on a large language model are not capable of conducting reliable legal research, and at paragraph 7 that those who use them have a professional duty to check the accuracy of that research against authoritative sources before using it. At paragraph 81 it held that a lawyer is not entitled to rely on their lay client for the accuracy of citations. It set out the full range of powers available, from public admonition and wasted costs to contempt proceedings and referral to the police, and said that admonishment alone is unlikely to be a sufficient response save in exceptional circumstances.

Will AI replace junior lawyers?

The evidence for that specific claim does not exist, and this page declines to assert it. What can be said is narrower and better founded. The tasks most exposed in legal work are document review, first drafts and citation checking, which are also the tasks through which junior lawyers historically learned to judge an authority. Brynjolfsson, Chandar and Chen find employment among 22 to 25 year olds in highly AI-exposed occupations about 19 per cent below the comparison trend, concentrated where AI substitutes rather than complements, through reduced hiring rather than dismissal. No published dataset tracks trainee solicitor or pupil barrister hours by task before and after 2023, so the specific legal claim rests on inference.

In this hub

Professions and sectors

Where the pressure lands first, profession by profession.

Bring this into your organisation

If this describes something happening in your teams, say so.

Keynotes, board sessions and advisory work, drawing on research across more than 200 organisations in 30 countries. Tell me the room, the date and the shift you need. A reply within 24 hours.

Start a conversation

Topics and audiences  ·  All research

The hallucination database now records more filings by people representing themselves than by lawyers. The English Divisional Court has ruled in Ayinde and Al-Haroun. For a firm, the work is a verification standard that survives a wasted costs application. AI advisory for CEOs and boards.

Box of Amazing

Rahim’s free weekly letter on AI and human capability

If this was useful, the weekly letter is where the thinking happens first. Most of what ends up on this site starts there. Weekly essays on AI, capability and the future of work. Read by 25,000 people, every week since 2017. Free, and one click to stop.

Opens Substack to confirm. No pitch in it, unsubscribe in one click, and nobody follows up because you read something.