The fix isn’t a better detector — it’s work whose thinking a chatbot can’t do for the student.
If a task can be finished by pasting the prompt into a chatbot, it now measures the chatbot. So you redesign the task around two questions: can I see the student’s thinking, and could a bot alone pass this. Moving work in-class, adding an oral defense, and requiring visible drafts all raise the first and lower the second. And instead of policing with unreliable detectors, you set disclosure norms — allow AI where it helps, ask students to say what they used — so honest use surfaces instead of going underground.
Interactive · assessment redesign lab
You start with a take-home essay: easy to outsource, hard to see the thinking. Flip the design moves and watch both dials respond, then see how your classroom’s norms compare with a policing one.
ban AI, run detectors, catch offenders
Students hide it, detectors misfire, and honest kids get falsely accused. You’re in an arms race you can’t win.
right now: policing
Turn on “allow AI with disclosure” to move from policing to disclosure norms.
Run every task through two questions, then reach for the moves that fix a weak answer:
Q1 Can I SEE the student's thinking? drafts, notes, an oral defense
Q2 Could a chatbot alone PASS it? if yes, it grades the tool
Design moves, strongest first for "can't delegate":
ORAL DEFENSE explain your reasoning, answer two follow-ups
IN-CLASS / PROCESS do it here; show outline, sources, versions
DISCLOSURE NORMS allow AI, require a line on what you used & how
Where AI helps vs. takes the thinking away:
HELPS brainstorming, feedback, formatting, a rough translation,
scaffolding for a struggling student
TAKES when constructing the argument or doing the derivation
IS the skill you're assessing
The no-proof suspicion talk. Detectors are unreliable, so you often can’t prove it. Don’t accuse — ask the student to walk you through their drafts and explain a paragraph. If the assessment already required process and defense, you rarely need this talk at all: understanding was demonstrated, not litigated.
| Move | Trade-off to weigh |
|---|---|
| Oral defense — highest signal on real understanding | Time-intensive; hard to scale to 150 students; keep questions equitable |
| In-class / process-visible — kills real-time outsourcing | Limits research depth; plan for accommodations and test anxiety |
| Visible drafts — shows the thinking develop | Trails can be faked; adds grading load |
| Disclosure norms — builds honesty over a gotcha culture | Only works if you genuinely allow AI somewhere and teach the line |
In a curriculum-design interview: “Your take-home essay is now trivially AI-completable. Redesign the unit.” A strong answer doesn’t reach for a detector. It keeps the essay as a thinking artifact but requires an outline, annotated sources, and revision history alongside it; adds a five-minute oral defense on two of the student’s own paragraphs; permits AI for brainstorming and feedback with a one-line disclosure; and moves the final synthesis paragraph in-class. Watch the dials: evidence of thinking climbs into “clearly visible” and AI-fakeable drops toward zero — and because AI is allowed honestly, no student needs to hide it and no teacher needs a gotcha.
You strongly suspect a paper is AI-written, but a detector score is your only evidence. Best next step?
Which task best resists being delegated to a chatbot?