OpenAI
ChatGPT Edu at Stanford
Stanford’s ChatGPT service, with university sign-in. This is where the skills and the agent go.
AI Learning Hub › AI Skills › Teach this
Workshop packet · and a do-it-yourself path
A session you can run — for a class, a study group, a journal, a clinic seminar, or a library workshop — on what these writing skills do, where their limits are, and who is responsible for the result. Everything it needs is already on this site. Work through it alone and it becomes a structured way to learn the whole set instead of poking at it.
This is a session about reviewing writing, not about producing it. Nobody should bring work they intend to submit unless the course that will receive it permits AI review, and no part of this packet authorises anything. Use the fictional practice drafts as the shared material: they exist so that a room can look at the same writing without anyone handing their graded work to an AI tool, or to each other.
Before anything else
Sign in through Stanford, not with a personal account. The university versions are the ones covered by Stanford’s agreements and its guidance on what may be put into them; a personal free account is not, and what you paste into one is not governed by the same terms.
OpenAI
Stanford’s ChatGPT service, with university sign-in. This is where the skills and the agent go.
Anthropic
Stanford’s Claude service, with university sign-in. The same ten skill files install here.
Facilitators: confirm access a week out, not on the day. Getting an account provisioned is the single thing most likely to cost you the first fifteen minutes of a session, and it is the one thing nobody in the room can fix.
Do this first
The point of installing in both tools and as an agent is not thoroughness for its own sake. It is that the three set-ups behave differently, and the differences are most of what there is to discuss. Ask participants to finish all three before the session; budget twenty minutes at the start if you would rather do it together.
Everything below comes from one download: the SLS Writing Partner Set. Unzip that set once. Inside are ten skill files, each its own ZIP, and each of those stays zipped at every step that follows.
In ChatGPT: Plugins, switch from Plugins to Skills, click +, and drop in each ZIP. Ten uploads, once. Full walkthrough with a recording: install a skill in ChatGPT.
In Claude: Customize, then Add at the top right, and drag each ZIP into the upload box. Same ten files, unchanged. Walkthrough: install a skill in Claude.
An agent is a named, reusable assistant that carries the skills, a standing set of instructions, and any files you attach — so you stop re-uploading and re-explaining in every chat. Follow Build a Writing Partner agent, which includes the instructions to paste and a short recording of the upload.
Everyone is ready when
If you only have time for one: install the ten in whichever tool the group already uses, and demonstrate the other two from the front. The comparison activity needs two tools in the room, but it does not need two tools on every laptop — pairs work.
The shared material
Five fictional student drafts sit on the skills page under Practice drafts: a case brief, three memos of increasing length and citation density, and a timed exam answer. They were written as teaching samples, and each one has real weaknesses built into it — that is what makes them useful. Nobody has to volunteer their own writing to have something worth reviewing in front of the room.
Everything in them is invented — the parties, the facts, the drafting, and the student authors. No authority cited in them should be relied on, and none of them is a model answer. Say this out loud at the start; someone always asks whether Patel v. Green is real.
Which draft to hand out depends on what you want the room to find:
| Draft | What is weak in it | Pairs well with |
|---|---|---|
| Case brief Daniels, ~400 words |
The rule statement claims more than the case decides, and the takeaway drifts past the holding. | Genre Fit, Argument and Structure |
| Memo — false imprisonment Patel, ~700 words |
Conversational register, thin authority, and a brief answer more confident than the analysis under it. | Clarity and Precision, Audience and Reception |
| Memo — emotional distress Lee, ~1,000 words |
Elements handled one at a time, with the contested one carrying weight the reasoning does not. | Counterargument, Flow and Organization |
| Memo — private facts Morgan, ~1,200 words |
Six authorities and a quoted email — the richest target for source and citation work. | Claims and Source Traceability, Bluebook Audit |
| Torts exam answer The wet floor, ~500 words |
Conclusions arriving ahead of the analysis, and elements marched through in order regardless of weight. | Argument and Structure, Audience and Reception |
A word of warning worth passing on. A group that reviews the same draft with the same skill will get five different reviews. That is not a malfunction, and noticing it is half the lesson — see the discussion topics below.
Part two
Choose two for an hour, three for ninety minutes. They are independent: any order, any combination. Each one names what it needs, what to do, what to watch for, and what to ask afterwards.
25–30 minutes · no laptop needed
The question: what is a skill, and why does that matter to a lawyer?
You need: the case study, Reverse-engineering Anthropic’s AI governance legal skills. Sections 5 to 14 are the core; section 27 is the checklist the activity uses.
Watch for: the moment someone realises the file is just written instructions, not a separate intelligence. That realisation is what makes the rest of the session possible.
Debrief: If a skill is a written procedure, whose procedure should it be — Anthropic’s, a vendor’s, the library’s, or your professor’s? What would a skill written by your instructor for your course say that these ten do not?
25 minutes · both tools
The question: how much of what comes back is the skill, and how much is the model running it?
You need: one practice draft, the same skill installed in ChatGPT and in Claude, and an identical opening message in both. The Morgan memo with Claims and Source Traceability is the sharpest pairing; Bluebook Audit is a close second.
Watch for: confident disagreement. Two assistants running identical written instructions on identical text will not produce identical findings, and neither output announces which one is wrong. Watch also for the tempting conclusion — that the longer review is the better one.
Debrief: If two runs of the same skill disagree, what does that tell you about a single run you never compared? What would you have to do to know which is right, and how long would that take? Is the answer “use both every time” realistic?
30 minutes · the agent
The question: what does a full review actually give you, and what does accepting it uncritically cost?
You need: the ChatGPT agent from set-up C, and the Morgan or Lee memo.
Watch for: the pull towards acceptance. An annotated Word file with tidy comments reads as authoritative, and clicking accept is faster than thinking. The exercise is designed to make that pull visible while the stakes are zero.
Debrief: Whose writing is it after you accept twenty suggestions you did not evaluate? Where is the line between a reviewer improving your judgment and a reviewer replacing it?
15–20 minutes · the short one
The question: where is the boundary, and how firm is it?
Every one of these skills refuses the same thing: it will not draft, rewrite, or hand back replacement prose. This activity tests that claim rather than trusting it.
Watch for: the grey zone. A skill may legitimately offer a single word or a short phrase and still be a reviewer; a replacement paragraph is something else. Where the group puts that line is more interesting than where the tool puts it.
Debrief: If you can talk a tool past its own boundary in four turns, what is the boundary actually for — and what is left holding the line? (Answer: the Honor Code, the course rule, and the person typing.) Would you disclose the turn that worked?
Facilitator note. Frame this as a stress test of a tool, not as instruction in evasion. The finding you want is that a written boundary is a real constraint and a soft one, which is exactly why the rules that govern the student are not written in the file.
30 minutes · the build
The question: can you encode your own standard well enough that a machine could follow it?
You need: the case study’s teaching template (there is a copy button on it), and a real standard from a course, journal, or clinic — a rubric, a style rule, a citation convention, a professor’s standing complaint.
Watch for: how quickly “write well” turns out to be unwritable, and how much of legal writing instruction is tacit. Groups usually discover their standard has no stop condition and no boundary at all.
Debrief: What part of your standard could not be written down? Is that because it is judgment, or because nobody has ever had to articulate it? Which answer should worry a teacher more?
Part three
Five sets. Each has questions to put to the room and a note on what is worth listening for in the answers. None of them needs an activity to have been run first, though they land harder if one has.
Set one
Listen for: the slide from “it only diagnosed” to “so the revision is entirely mine.” Both halves can be true and the conclusion can still be too comfortable. Push on the case where the diagnosis was the insight.
Set two
Listen for: people treating “the assistant found the case” as the end of the process. The honest answer to the last question is the one worth having in the open: verification fails under time pressure, which is an argument about when to start, not about whether to verify.
Set three
Listen for: the assumption that a skill is a guarantee. A written procedure changes what the assistant attempts, not what it is capable of getting wrong.
Set four
Listen for: winner-takes-all framing. The case study’s stack diagram is the corrective: the interesting question is which layer you are choosing at, not which company wins.
Set five
Listen for: the difference between prohibited and unwise. Both matter, and only one of them is written down. The PAUSE Rule is the site’s workflow for the second.
Part four
One note per skill: what it does, and the thing worth saying about it out loud. Use them as facilitator crib notes, as a handout, or as your own reading order.
The ten
The reading
Reverse-engineering Anthropic’s AI governance legal skills takes apart a real, public, Apache-2.0 project — Anthropic’s Claude for Legal — and asks what a skill is made of. It is the reading that turns “I installed a thing” into “I know what I installed.”
There is also a 21-minute audio explainer covering the same ground. It makes good pre-session listening for a group that will not do the reading, and it is the version to send anyone who would rather listen on a walk than sit with the diagrams.
SKILL.md file contains, what YAML frontmatter is, and why the description is
what routes work to the right skill. Use it for: activity A, and for anyone who thinks a
skill is code.Attribution to repeat. The case study is an independent educational reading of someone else’s open-source project. It is not affiliated with or endorsed by Anthropic, and its code samples are teaching illustrations rather than reproductions.
Part five
Set-up happens before all three. If it cannot, add twenty minutes at the front and drop an activity.
Sixty minutes
Leaves with: the tools disagree, and the boundary is soft.
Ninety minutes
Leaves with: what it is, what it produces, and who is responsible.
On your own
Spread over a week, an hour at a time.
Make it yours
Nothing here is fixed. Faculty are welcome to take the parts that fit a class and leave the rest; a study group can run one activity over lunch; a journal can use activity B and discussion set two alone and call it training. Three ways it bends:
One thing not to adapt away: the course-policy question. Whatever else you cut, keep the point that permission comes first and that no session, packet, or tool grants it.
Part six
The words that come up in this packet, in plain English. Most of them sound more technical than they are.
SKILL.md.md means Markdown: ordinary text with a little formatting, readable in any text editor.# for a heading, - for a bullet, **bold** for bold.--- lines, holding its name and description. The label card on the outside of the folder.Everything this packet uses
The ten skills and the practice drafts are on the skills page; the agent guide, the case study, and the PAUSE Rule each have their own. Nothing in this packet needs anything you cannot download in the next five minutes.
Running this with a class? Email library@law.stanford.edu and we will help.
Robert Crown Law Library · last reviewed August 2026.