Alleged Fake Citations Undermine Expert Testimony in Minnesota Deepfake Law Case

A misinformation expert submitted an affidavit in a Minnesota legal case that included multiple non-existent academic citations. Attorneys for the plaintiffs argued these citations were AI-generated hallucinations from ChatGPT, which has cast doubt on the credibility of the expert's entire declaration and the legal proceedings.

In a legal case defending Minnesota’s deepfake election misinformation law, Stanford misinformation expert Professor Jeff Hancock's affidavit allegedly cited non-existent academic sources, potentially generated by ChatGPT. The reportedly fabricated citations appear to have undermined the credibility of his testimony.

Source: AI Incident Database

Risk classification

  • Primary risk domain: 3 Misinformation
  • Primary risk subdomain: 3.1 False or misleading information

The AI system generated fabricated academic citations, which constitutes false and misleading information that was subsequently introduced into a federal court proceeding.

Additional risk subdomains

  • 5.1 Overreliance and unsafe use: The expert relied on the AI-generated citations without verifying their authenticity, leading to severe reputational and legal consequences.
  • 7.3 Lack of capability or robustness: The AI system exhibited a lack of robustness by hallucinating plausible-sounding but entirely fabricated academic sources.

Causal factors

  • Entity: AI
  • Intent: Unintentional
  • Timing: Post-deployment

The generation of non-existent academic citations was an unintentional hallucination produced by the AI system during its post-deployment use.

EU AI Act risk tier

  • Risk tier: 3 Limited Risk

Risk Level 3: Limited Risk. The system is a chatbot or general-purpose AI used to generate text, which falls under limited risk requiring transparency obligations.

AI system and alleged parties

  • AI system: ChatGPT (OpenAI)
  • AI purpose: Writing Assistant; Question Answering
  • Behaviour type: Assistant
  • Alleged developer: OpenAI, ChatGPT
  • Alleged deployer: Jeff Hancock
  • Alleged harmed parties: Mary Franson, Keith Ellison, Jeff Hancock, Christopher Kohls, Chad Larson

Harm severity

Highest direct severity in any category: Minor. Severity is scored from Negligible to Catastrophic in each harm category, for harm the reports describe as caused directly or indirectly by the AI system.

  • Physical: direct Negligible, indirect Negligible
  • Infrastructure: direct Negligible, indirect Negligible
  • Property: direct Negligible, indirect Negligible
  • Financial: direct Negligible, indirect Negligible
  • Environmental: direct Negligible, indirect Negligible
  • Malicious content: direct Negligible, indirect Negligible
  • Differential treatment: direct Negligible, indirect Negligible
  • Civil rights: direct Negligible, indirect Negligible
  • Democracy: direct Negligible, indirect Negligible
  • Privacy: direct Negligible, indirect Negligible
  • Psychological: direct Negligible, indirect Negligible
  • Epistemic: direct Minor, indirect Minor
  • Child sexual exploitation and abuse: direct Negligible, indirect Negligible

Epistemic

Reported: The report explicitly describes epistemic harm through the fabrication of academic citations by an AI, which were then presented as true facts in a legal document.

Directly caused: The AI generated non-existent academic sources, such as 'The Influence of Deepfake Videos on Political Attitudes and Behavior', which were included in a formal legal declaration.

Indirectly caused: The inclusion of these fake citations misled the court and the public, casting doubt on the credibility of the expert's entire declaration and the research lab.

Inferred additional harm: N/A

People affected

  • Occurrences reported: 1
  • People reportedly harmed: 1
  • People reportedly exposed: 3

Potential causes

Management

  • Inadequate Delegation Oversight: Failure to supervise assistants or co-authors who may have used AI tools.

Technology

  • LLM Citation Hallucination: ChatGPT generated plausible-sounding but entirely fictional academic citations.

Human Factors

  • Failure to Verify References: The expert signed the declaration without verifying the cited academic papers.
  • Overreliance on AI Tools: Assuming AI-generated text is accurate without performing manual fact-checks.

Process and Methods

  • Inadequate Drafting Review: No formal process to cross-reference academic citations before legal submission.

Information quality

  • Classification confidence: High
  • Reason for confidence: The report clearly details the specific fake citations, the court case, the parties involved, and the consensus that they represent AI hallucinations. The facts are consistent across the reporting.
  • Ambiguities identified: It is not 100% confirmed which specific LLM was used (though ChatGPT is suspected) or who exactly ran the prompt (Hancock or an assistant).

A prominent misinformation expert submitted a federal court declaration containing fabricated academic citations generated by ChatGPT. While highly embarrassing and damaging to the credibility of the legal proceedings regarding deepfake regulations, the incident represents a minor domestic legal disruption with negligible direct national security impact.

  • Overall national security impact: Minor
  • Response level: Moderate
  • Scope: Single nation
  • Primary target: United States
  • Alleged perpetrator: Unknown

Threat characteristics

  • Imminence: Long-term. The incident represents an ongoing concern regarding LLM reliability in professional settings rather than an active national security crisis.
  • Autonomy: Human-controlled. The AI system was used as a writing assistant, and humans retained final control over the submission of the document.
  • Novelty: Established threat. Similar incidents of AI-generated hallucinations being submitted in legal proceedings have occurred previously.

Impact by dimension

  • Physical security: Negligible. No physical systems, kinetic threats, or critical infrastructure were affected by this incident.
  • Information security: Negligible. The incident involves unintentional AI hallucinations in a domestic legal filing, not a state-sponsored information warfare campaign or intelligence compromise.
  • Sovereignty: Minor. The submission of fabricated citations temporarily impacted the credibility of an expert witness in a federal court case concerning election deepfake laws.
  • Economic security: Negligible. No strategic technologies, financial systems, or economic security interests were compromised.
  • Societal stability: Negligible. The incident did not result in mass surveillance, societal instability, or systematic human rights violations.
Explore in the interactive Incident Tracker