> Mata v. Avianca: Court Cases Invented by ChatGPT, Filed in Federal Court: Case study: lawyers filed court cases ChatGPT invented, and ChatGPT said they were real. Why an AI vouching for itself is not a check.
>
> Evidentiality Framework for AI. Early findings, October 2026. Web version: https://evidentiality-framework.org/case-mata-v-avianca.html. Text CC BY 4.0.

Case study

# Mata v. Avianca: Court Cases Invented by ChatGPT, Filed in Federal Court

In 2023, two New York lawyers cited earlier court cases in a lawsuit. The cases didn’t exist. ChatGPT had made them up, and when asked, it said they were real.

**What are the labels?** The Evidentiality Framework asks an AI to mark each claim it writes: (g) generated, its own work; (u) given, passed to it by someone else; or (m) checked against a named source. The marks stay on a claim while people work with it. [How the labels work](../../labels.html).

---

## What Happened

- A man sued the airline Avianca. The case was moved to a New York federal court. On 1 March 2023 his lawyers filed a brief, a written argument to the court, citing earlier cases, including “Varghese v. China Southern Airlines Co., Ltd., 925 F.3d 1339 (11th Cir. 2019)”. That case doesn’t exist.
- On 15 March, Avianca’s lawyers told the court they couldn’t find several of the cases. The court ordered copies. In April the lawyers filed what they said were excerpts.
- The lawyer who did the research had asked ChatGPT “Is Varghese a real case”. It told him the cases were real and could be found in the main legal databases. The lawyer who signed the filing hadn’t read any of the cases.
- On 22 June 2023 the judge fined the two lawyers and their firm $5,000, and ordered them to send the court’s ruling to the real judges falsely named as authors of the fake rulings.
- The judge wrote that “there is nothing inherently improper about using a reliable artificial intelligence tool for assistance.” He found bad faith, based on “acts of conscious avoidance and false and misleading statements to the Court”.

---

## Follow the Claim

This is an **illustration** of how the labels would have worked, not a test. It only holds if the conditions under “What Would Have Had to Be True” held.

Follow one case: from ChatGPT to a federal court

**Key:** as it happened, the type gets bigger as the claim sounds more certain. With labels: red (g) generated: written by the AI; blue (m) checked against a named source.

**As it happened****With labels**

Early 2023 · ChatGPT → a lawyer

As it happenedVarghese v. China Southern Airlines Co., Ltd., 925 F.3d 1339 (11th Cir. 2019)A case reference in the exact legal format.

With labels(g)Varghese v. China Southern Airlines Co., Ltd., 925 F.3d 1339 (11th Cir. 2019)(/g)If the tool used the labels: marked as written by the AI.

Before 1 March 2023 · The check

As it happenedNobody looked the case up in a legal databaseThe case was never found, because it doesn’t exist.

With labels(m)No such case exists(/m: the court’s ruling of 22 June 2023, checked 3 Oct 2026)A legal database search could show this at the time. We name the source that confirmed it later.

1 March 2023 · Lawyers → the court

As it happenedCited in a signed court filingNow it’s legal argument, with a lawyer’s name on it.

With labelsLeft out. The check found nothing, so the claim never reaches the finished document.Finished documents carry no labels. They only carry claims that passed the check.

The case does not exist. The reference is as filed, quoted in the court’s ruling.

Asking the AI to check itself

**Key:** as it happened, the type gets bigger as the claim sounds more certain. With labels: red (g) generated: written by the AI.

**As it happened****With labels**

2023 (the exact date was disputed in court) · Lawyer → ChatGPT: “Is Varghese a real case”

As it happenedChatGPT says the cases are realThe AI vouches for itself.

With labels(g)Varghese is a real case(/g)Still red. An AI checking its own answer is still the AI’s own claim. A check needs a source outside the AI.

The court found the lawyer’s explanations of when he asked this were not consistent.

---

## How the Labels Could Have Helped

- **Self-checks stay red.** Asking the AI whether its answer is real produces another (g), not an (m). This is the clearest lesson of the case.
- **Filing waits for a check.** A case nobody has found in a legal database isn’t cited.
- **The signing lawyer can see what was checked.** He signed without reading the cases. Labels would have shown him that nobody had.

---

## What Would Have Had to Be True

The [four conditions every case shares](../../cases.html#assume): the AI tool used the labels; it labelled its own work correctly (the weakest link: in [our tests](../../check.html), AI sometimes mislabels its own work); the label stayed on when the text was copied; and someone owned a rule that unchecked claims don’t go further. In this case:

- The rule is owned by the lawyer who signs the filing, who already has that duty.

---

## What Already Existed

- **Lawyers already have to check.** US court rules (Rule 11) make a lawyer who signs a filing vouch that its legal arguments rest on real law.
- **Tools exist.** Legal databases and “citators” confirm whether a case exists. The firm’s research service had limited coverage of federal cases, the court found.
- **A simpler check would have caught it:** a search in a legal database. Labels add one thing: the unchecked references stand out before anyone signs.

---

## What the Labels Wouldn’t Have Caught

- **What came after.** Much of the court’s criticism was about how the lawyers responded once they were warned. Labels don’t make anyone own up.
- **Signing without reading.** A label only helps if the person signing reads it.
- **Limited research tools.** The firm turned to ChatGPT partly because its usual service didn’t cover the cases it needed.

---

## Further Reading

- [Mata v. Avianca, the court’s ruling on sanctions, 22 June 2023 (Justia)](https://law.justia.com/cases/federal/district-courts/new-york/nysdce/1:2022cv01461/575368/54/)
- [Bloomberg Law: phony ChatGPT brief leads to $5,000 fine](https://news.bloomberglaw.com/esg/chatgpt-phony-legal-filing-case-gets-lawyers-a-5-000-fine)

Also listed in the [AI Incident Database (#541)](https://incidentdatabase.ai/cite/541/).

**Next:** [All case studies](../../cases.html) · [How the labels work](../../labels.html)
