Activity Bank
A student asks the librarian for the source behind a citation in an AI-written research brief.

All activities Student learning

Hallucination Hunt

Students trace every claim and citation in a polished AI research brief and discover which ones lead somewhere and which ones lead nowhere.

Print handouts

Overview

Generative AI tools can produce research summaries that look finished: confident prose, tidy numbers, a reference list in perfect format. Some of those references are real, some are real but misrepresented, and some do not exist at all, because the model predicts what a citation should look like rather than retrieving one. In this hunt, students audit a mock AI brief about a fictional town's heat problem against a "library shelf" of source cards, sort each claim by what the evidence shows, and then repair the brief. Every source in the packet is fictional, so the only way to win is to trace, not to recognize.

Objectives

  • Students will explain why a generative AI tool can produce citations that look real but cannot be found (prediction from patterns, not retrieval).
  • Students will trace each claim in an AI-generated brief to its cited source and classify it as supported, distorted, unsupported, or fabricated.
  • Students will revise the brief so every remaining claim is backed by a source they checked, and document what they changed and why.

Materials

On paper

  • Handout A: The AI Research Brief (1 per student, for annotating)
  • Handout B: The Library Shelf source cards (1 set per team, cut apart and kept at a "reference desk" at the front of the room)
  • Handout C: Claim Trace Log (1 per team)
  • Highlighters in three colors; sticky notes

On screen

  • One device per team with access to your school or college library databases and a search engine (Round 2 only)
  • A generative AI chatbot your district or institution approves, for students 13+; otherwise the teacher projects it

Before you start

  1. Print Handout B on card stock and cut it apart. Keep the full set at a reference desk; teams must come to the desk and ask for a specific source by its citation, the way they would at a library.
  2. Read the card backs (facilitator key) so you know which of the brief's seven claims are supported, distorted, unsupported, or fabricated.
  3. For the digital round, ask a generative AI tool your institution approves to write a short research summary with five citations on a topic your class is studying. Save the output so every team audits the same text.
  4. Optional: invite your school or college librarian to co-teach the Round 2 database search.

Step by step

  1. 10–5 min

    Hook

    Would you turn this in?

    Hand out Handout A and give students 90 seconds to skim it. Ask: "If a classmate turned this in, what grade would it get just on looks?" Take a quick show of fingers (1–5). Then ask: "How many of these seven references did you check?" Say: "Today you're a research auditor. The brief was written by an AI chatbot. Your job is to find out which parts are true, which parts are twisted, and which parts were never true at all."

    Facilitator noteExpect high scores. The format, the numbers, and the reference list all signal "done." That reaction is the thing you are teaching students to notice.

  2. 25–13 min

    Model

    Why AI invents citations

    Explain in plain terms: a chatbot builds text by predicting likely next words from patterns in its training data. It has seen thousands of reference lists, so it can produce one that looks correct: author, year, title, journal. Unless the tool is connected to a search or database and shows you the retrieved source, nothing guarantees the reference exists. Then model the trace on Claim 1: underline the claim, find its citation, walk to the reference desk, and ask for that source by author and year. Read the card aloud and think out loud: "Does the source exist? Does it say what the brief says it says? Is the number the same?"

    Facilitator noteName the four verdicts now and write them on the board: Supported, Distorted (source exists but says something different), Unsupported (no citation, or source doesn't address it), Fabricated (the source cannot be found).

  3. 313–28 min

    Explore

    Round 1: the reference desk hunt

    Teams trace Claims 2–7. Roles rotate each claim: the Tracer requests the source card at the desk, the Reader reads the card aloud, the Recorder logs the verdict and the evidence on Handout C. Teams may request only one card at a time and must name the citation exactly as the brief gives it. If they ask for a source that isn't on the shelf, hand them the "No record found" card and have them log what they searched for.

    Facilitator noteListen for teams that mark a claim "supported" because the source exists. Push: "The source is real, but does it say that?" Claims 3 and 5 are built to catch that shortcut.

  4. 428–35 min

    Debrief

    Score the brief

    Tally verdicts on the board. Ask: "How many claims survived? Which kind of error was hardest to catch, and why?" Most teams find fabricated citations easier to catch than distorted ones, because a missing source is obvious while a twisted one requires reading. Ask: "What single habit would have caught every problem?" Steer toward: open the source and compare the exact wording and numbers.

  5. 535–50 min

    Apply

    Round 2: hunt a live output

    Project the AI summary you generated during prep, on a topic your class is studying. Teams choose three of its citations and search for them in your library databases and a search engine. For each, they log: Did it exist? Did it say what the AI claimed? What better source did we find? Teams under a device limit share one device and rotate the Tracer role. Unplugged option: skip the live output and have teams repair Handout A instead: rewrite the brief's summary paragraph using only supported claims, with the citations corrected.

    Facilitator noteResults vary by tool and topic, and some tools now retrieve real sources. That is fine. The finding "this one checked out" is as valuable as "this one didn't," because the student verified it rather than assumed it.

  6. 650–60 min

    Reflect

    The auditor's note

    Each student writes a short auditor's note at the bottom of Handout C: "The claim I would have believed without checking was ___, and here's what the source actually showed: ___." Then: "When is it reasonable to use an AI tool for research, and what must you do before anything it says goes into your work?" Collect Handout C as the exit ticket.

    Facilitator noteLook for students who separate using AI to find leads from trusting AI as a source. That distinction is the goal.

Paper or screen

Unplugged

Run Round 1 exactly as written: the source cards at the reference desk are the unplugged version of opening a database. For Round 2, teams repair the brief on paper, keeping only supported claims, correcting distorted ones, and striking fabricated citations, then write one sentence per change explaining it. The trace-and-compare habit is fully preserved without a screen.

Digital

Post Handout A as a document students annotate with comments, one comment per claim with the verdict and the evidence. In Round 2, teams search library databases and the open web for the live AI output's citations and paste the direct link or database record into their log. For students 13+ at institutions that allow it, teams may ask an approved chatbot to "provide the source for claim X" and then check whether that answer holds up, which usually makes the lesson land harder. Students under 13 do not use this path; the teacher projects any AI tool.

Does it need a screen? The live round shows students what the paper round can only describe: a real tool, on their own topic, producing references they must actually find in a real database. The evidence of learning is the log of what they searched, what they found, and what they changed, which is something a finished essay never shows.

Evidence of learning

What you should be able to see or collect if it worked.

  • Handout C logs a specific piece of evidence for each verdict (a quoted phrase or number from the source card), not just a label.
  • Students distinguish a real-but-distorted source from a fabricated one and explain the difference.
  • In Round 2, teams record at least one citation they verified and one they could not, with what they searched for.
  • Auditor's notes name a claim the student would have believed and describe what changed after checking.

Adaptations

Higher Ed
Use a live AI output on a topic from your discipline and require teams to locate each source through the campus library, including checking whether a cited article is peer reviewed, a preprint, or a press release about either.
Grade 9 or emerging readers
Trace Claims 1–4 only and provide the verdict words on a strip so teams choose from them. Read the source cards aloud at the desk.
Emergent bilinguals
Pair the brief with a glossary of claim, citation, source, fabricated, and distorted, and let teams discuss verdicts in any language before logging in English.
Short on time (40 min)
Skip Round 2 and extend the debrief into the repaired-paragraph task.

Standards connections

Students trace each claim in an AI research brief to its source, catch fabricated and distorted citations, and revise the brief, core ELAR inquiry and research.

TEKS
inquiry and researchethics and laws (c)(9)computational thinkingcompositionTEKS sections: Technology Applications high school Technology Applications courses (19 TAC Chapter 126); ELAR §110.36–§110.39

See how all activities align

Reflect

  • How do I question and verify AI-generated information before I use it? (AI4.1, AI4.2)
  • How can I tell a trustworthy source from one that only looks trustworthy? (S2.2)
  • What limitation of AI tools did I see today, and how will it change the way I research? (AI1.3)

What comes next

For the next research assignment, students attach a short trace log to any work where they used an AI tool: each AI-suggested source, whether they found it, and what they used instead. At home, students can run the three-question check (Does it exist? Does it say that? Is the number the same?) on one claim they see in their feeds this week.

Pairs well with