The thought experiment lab

arqiv✳

Lesson · Philosophy & ethics

The thought experiment lab

Some questions can’t be looked up; they have to be reasoned through. Students work through classic dilemmas about AI on paper, discover their own principles, and then test them against an AI’s strongest counterargument.

Level 1 AI after thinkingGrades 7–1250 minutesSocial studies · ELA · EthicsPaper + 1 teacher screen

The big idea

When there’s no answer key, your reasoning is the answer.

Students will be able to, without AI:

  • State a moral principle in their own words
  • Notice when changing one detail changes their answer, and why
  • Steelman an opposing view
  • Judge whether an AI’s argument is actually strong

Lesson flow · the 5 P’s

Think first. Then the tool.

Most of the lesson happens on paper. AI appears for one short, structured phase, and the teacher can run it from a single projector screen. How the 5 P’s work →

1

Ponder

15 min○ Screens off · paper only

Read the dilemma aloud. Students write their answer and one reason before any discussion.

  1. The dilemma: A self-driving car’s brakes fail. It can stay in its lane and hit three people who stepped into the road, or swerve and hit one person on the sidewalk. What should it be programmed to do?
  2. Write your answer and your reason.
  3. Change one detail. What if the one person is a child? What if the three were jaywalking? What if swerving means crashing into a wall instead, and the person who gets hurt is you, riding in the car? Write whether your answer changes each time.
  4. Circle the detail that changed your mind the most. That’s a clue to your real principle.
Say to students“There’s no right answer to look up, and I’m not grading whether you agree with me. I’m grading how well you think.”
2

Plan

10 min○ Screens off · paper only

Students turn their gut reaction into a principle they can defend.

  1. Complete: “A self-driving car should ______ because ______.”
  2. Predict the two strongest objections someone could raise.
  3. Write one question you want to ask the AI to test your principle.
3

Partner

10 min● AI on · teacher screen or lab

On the teacher’s screen, the class asks AI to argue against the principles students wrote. AI is a sparring partner, not a judge.

  1. Run 3–4 student principles that take different positions (for example, “save the most lives,” “never swerve into someone who was safe” or “protect the people in the car”). Students copy the AI’s main point in their own words.
  2. Students whose principle wasn’t run use the counterargument to the principle closest to theirs.
  3. If time allows, ask the AI 1–2 of the students’ own questions from the Plan.
Here is a student’s principle: “[principle]”. Give the single strongest argument against it, in plain language, in under 120 words.
How would a philosopher who judges actions by their outcomes answer this dilemma? How would one who believes some actions are always wrong answer it?
4

Probe

10 min○ Screens off · paper only

Screens off. Students evaluate the AI like a debate judge.

  1. Was the AI’s argument actually strong, or did it just sound confident?
  2. Did it dodge the question or pretend there’s one right answer?
  3. Did it just agree with whatever it was told? (Try asking it to argue the opposite next time.)
  4. Real-world check: MIT’s Moral Machine experiment collected about 40 million decisions from people in 233 countries and territories, and found people’s choices varied by culture. Who should decide what real cars do?
5

Prove

5 min + optional seminar○ Screens off · paper only

Students defend their position without any screen.

  1. Rewrite your principle. Did you keep it, change it or add an exception? Explain why.
  2. Optional: hold a 15-minute Socratic circle where students must respond to the strongest opposing argument first.
  3. Exit ticket: one thing the other side gets right.

Printable

Student worksheet

Print one per student. Everything students hand in is handwritten, so their thinking is visible.

Thought experiment lab

Write your own thinking first. There is no answer key.

Name:Date:Class:
My first answer and reason
Change one detail: did my answer change?
☐ If the one person is a child
☐ If the three were jaywalking
☐ If swerving means crashing into a wall and I’m in the car
My principle: “A self-driving car should ____ because ____.”
Two objections I predict
My question for the AI
The AI’s strongest argument, in my own words
Was it actually strong? Why or why not?
My final principle (kept, changed or added an exception) and why
Exit ticket: one thing the other side gets right

Assessment

Grade the thinking, not the output

Every row scores something students did themselves. An AI can produce a product; it can’t produce a student’s reasons.

Criteria4 · Excellent3 · Solid2 · Developing1 · Beginning
ReasoningClear principle with reasons that hold across variationsClear principle with reasonsOpinion with weak reasonsNo reasoning given
CounterargumentsAnticipates strong objections and answers themNames objectionsNames a weak objectionNone
Evaluating the AIJudges strength, spots dodges or flattery, explains whyJudges the argument with a reasonAgrees or disagrees without reasonsCopies the AI
RevisionThoughtfully revises or defends with new insightRevises or defends with a reasonMinimal change, little explanationNo final position

Why this is hard to cheat

There’s nothing to copy. Each student’s first answer, variations and final principle are handwritten in class, and the AI only appears on the teacher’s screen to argue with them.

Adjust it

Younger students: Use a gentler dilemma: “Should a robot helper always tell the truth, even if it hurts someone’s feelings?”

Older students: Add John Searle’s Chinese Room thought experiment (1980). A person who knows no Chinese sits alone in a room with a rulebook written in English. Questions in Chinese are slipped under the door. By following the rules, the person sends back answers so good that people outside think a Chinese speaker is inside. Does the person understand Chinese? Does a computer that follows rules the same way? Connect to how transformers work.

No devices at all? The teacher prepares two AI counterarguments in advance on a handout: one against “save the most lives” and one against “never swerve into someone who was safe.” Students judge them on paper.

Discussion questions

  • Should companies, governments or voters decide how self-driving cars make choices?
  • Why might people in different countries answer differently?
  • Can an AI have a principle, or only predict what people would say?
  • What did you learn about your own values today?

Extensions & connections

Put curiosity to work

Read it. Try it. Question it.

Explore new AI research, or take a paper-first investigation into your classroom.