The answer looked legal before anyone checked it

People are already using general-purpose chatbots for real legal problems. A new study asks a simple question that accuracy tests often miss: who checks the answer before someone acts on it?

The researchers examined 153 first-person Reddit narratives published between January 2023 and January 2026. Most described no independent check and no timely community review. Independent verification appeared in 17.3% of the posts. Thirty threads, or 19.6%, put AI-generated material in front of a community that could still influence the user's next step.

The distinction is important. A confident legal-looking letter can feel credible because of its structure and tone, even when the person reading it cannot judge the law underneath.

The paper does not show that every unverified answer was wrong. It shows that, in these public stories, the burden of deciding was often left with the person least equipped to carry it.

Drafting was the main use

The team collected material from three legal subreddits and three AI-tool communities. A broad first pass found 10,289 candidate posts and 85,533 comments. Human review narrowed that to 153 distinct narratives about actual, non-hypothetical legal use.

Seventy-seven per cent of the final posts involved drafting. Users described demand letters, complaints, settlement documents, motions and court filings. ChatGPT appeared in 90.2% of the narratives, with smaller numbers mentioning Claude, Gemini or Grok.

Some people did make deliberate checks. Fifteen posts, or 9.8%, described comparing answers from several models. Others consulted an authoritative source or a professional. Around 30 threads showed a different route: asking Reddit users to review what the model had produced.

These practices overlap, so their percentages should not be added together. And silence is not proof that no check happened. A user may have verified the answer offline and simply left that detail out of the post.

The crowd was not evenly useful

The 153 posts attracted 5,341 comments and replies, but scrutiny was heavily concentrated. The 20 most-discussed posts generated 82.5% of all reactions.

Technology communities averaged 38.4 reactions per post. Legal communities averaged 2.5. The legal forums produced much more substantive legal discussion as a share of their replies, while technology forums were more supportive and more diffuse.

That creates a difficult trade-off. A person may get far more attention in an AI enthusiast community, but attention is not the same thing as qualified checking. Posting in a legal forum may bring more relevant scrutiny and much less of it.

The researchers call the fuller pattern distributed counsel: a model produces material, a user directs and applies it, and a community reviews it. It is not a substitute for a lawyer. It is an improvised support system built where professional help is hard to reach.

The method checked its own classifier

The study combined close human coding with automated analysis. Gemini 2.5 Flash classified the stance of 4,993 reactions judged to be human, while a keyword method offered a more conservative comparison.

Two doctoral law researchers independently re-annotated a sample of 30 posts and 89 reactions. Against the adjudicated labels, the model-assisted stance classification reached a Cohen's kappa of 0.74. The authors still treat it as a descriptive aid, not ground truth.

That caution matters because Reddit is not a representative survey. People choose what to post, successful or dramatic stories may travel further and promotional activity cannot be ruled out. Reported outcomes were not checked against court records or case files.

The current UK Civil Justice Council is considering related questions around AI-prepared court documents. Its June update said litigants in person present distinct and evolving challenges. A final report was still expected later in 2026.

What is observed, interpreted and still unknown

Confirmed: the five-author preprint was submitted on 13 August 2026. It covers six subreddits from January 2023 to January 2026 and analyses 153 canonical narratives with 5,341 associated reactions.

Observed in the sample: 17.3% of posts reported independent verification, 30 threads invited community evaluation and most narratives reported neither. Drafting appeared in 77% of posts, and engagement was far higher in technology communities than legal ones.

The researchers' interpretation: professional-looking language and emotional reassurance can give advice practical force before its accuracy is established, shifting verification work from regulated professionals to users and online communities.

Important limits: these are self-reported Reddit accounts, not a measure of all chatbot use. The study cannot establish the advice's accuracy, legal effect or causal role, and missing verification details may simply have happened offline.

Still open: how often private users check answers, what a safe verification path should look like and whether accessible professional services can meet the need that chatbots are filling. A polished answer can be useful. It can also arrive without anyone responsible for checking it.

Sources

  1. Owens et al. — AI legal-advice study recordPrimary preprint record submitted 13 August 2026. Source for authorship, sample, research question and headline findings.
  2. Owens et al. — full AI legal-advice manuscriptFull primary manuscript. Source for collection, coding, verification rates, forum comparison, human validation, interpretation and limitations.
  3. Civil Justice Council — use of AI in court documentsOfficial June 2026 update. Source for the current UK review and its treatment of litigants in person as a distinct area needing further work.