AI Learning HubAI Skills › Teach this

Workshop packet · and a do-it-yourself path

Teach this: the Writing Partner

A session you can run — for a class, a study group, a journal, a clinic seminar, or a library workshop — on what these writing skills do, where their limits are, and who is responsible for the result. Everything it needs is already on this site. Work through it alone and it becomes a structured way to learn the whole set instead of poking at it.

Start with set-up Jump to the activities Glossary of the jargon ↓
Runs in
60 or 90 minutes, or at your own pace
Set-up before the session
About 20 minutes, once, per person
Everyone brings
A laptop and their Stanford accounts
Take what you need
Five activities, five discussion sets — pick, reorder, cut

This is a session about reviewing writing, not about producing it. Nobody should bring work they intend to submit unless the course that will receive it permits AI review, and no part of this packet authorises anything. Use the fictional practice drafts as the shared material: they exist so that a room can look at the same writing without anyone handing their graded work to an AI tool, or to each other.

Before anything else

Use the Stanford versions of both tools

Sign in through Stanford, not with a personal account. The university versions are the ones covered by Stanford’s agreements and its guidance on what may be put into them; a personal free account is not, and what you paste into one is not governed by the same terms.

Facilitators: confirm access a week out, not on the day. Getting an account provisioned is the single thing most likely to cost you the first fifteen minutes of a session, and it is the one thing nobody in the room can fix.

Do this first

Set-up: the same ten skills, three ways

The point of installing in both tools and as an agent is not thoroughness for its own sake. It is that the three set-ups behave differently, and the differences are most of what there is to discuss. Ask participants to finish all three before the session; budget twenty minutes at the start if you would rather do it together.

Everything below comes from one download: the SLS Writing Partner Set. Unzip that set once. Inside are ten skill files, each its own ZIP, and each of those stays zipped at every step that follows.

A — The ten skills in ChatGPT

In ChatGPT: Plugins, switch from Plugins to Skills, click +, and drop in each ZIP. Ten uploads, once. Full walkthrough with a recording: install a skill in ChatGPT.

B — The ten skills in Claude

In Claude: Customize, then Add at the top right, and drag each ZIP into the upload box. Same ten files, unchanged. Walkthrough: install a skill in Claude.

C — One agent in ChatGPT holding all ten

An agent is a named, reusable assistant that carries the skills, a standing set of instructions, and any files you attach — so you stop re-uploading and re-explaining in every chat. Follow Build a Writing Partner agent, which includes the instructions to paste and a short recording of the upload.

Everyone is ready when

  • They are signed in to both tools through Stanford.
  • Ten skills are installed in ChatGPT, and ten in Claude — none of them unzipped first.
  • One ChatGPT agent exists, named, with the instructions pasted in.
  • At least one practice draft is downloaded and on the laptop.

If you only have time for one: install the ten in whichever tool the group already uses, and demonstrate the other two from the front. The comparison activity needs two tools in the room, but it does not need two tools on every laptop — pairs work.

The shared material

The practice drafts are imperfect on purpose

Five fictional student drafts sit on the skills page under Practice drafts: a case brief, three memos of increasing length and citation density, and a timed exam answer. They were written as teaching samples, and each one has real weaknesses built into it — that is what makes them useful. Nobody has to volunteer their own writing to have something worth reviewing in front of the room.

Everything in them is invented — the parties, the facts, the drafting, and the student authors. No authority cited in them should be relied on, and none of them is a model answer. Say this out loud at the start; someone always asks whether Patel v. Green is real.

Which draft to hand out depends on what you want the room to find:

DraftWhat is weak in itPairs well with
Case brief
Daniels, ~400 words
The rule statement claims more than the case decides, and the takeaway drifts past the holding. Genre Fit, Argument and Structure
Memo — false imprisonment
Patel, ~700 words
Conversational register, thin authority, and a brief answer more confident than the analysis under it. Clarity and Precision, Audience and Reception
Memo — emotional distress
Lee, ~1,000 words
Elements handled one at a time, with the contested one carrying weight the reasoning does not. Counterargument, Flow and Organization
Memo — private facts
Morgan, ~1,200 words
Six authorities and a quoted email — the richest target for source and citation work. Claims and Source Traceability, Bluebook Audit
Torts exam answer
The wet floor, ~500 words
Conclusions arriving ahead of the analysis, and elements marched through in order regardless of weight. Argument and Structure, Audience and Reception

A word of warning worth passing on. A group that reviews the same draft with the same skill will get five different reviews. That is not a malfunction, and noticing it is half the lesson — see the discussion topics below.

Part two

Five activities

Choose two for an hour, three for ninety minutes. They are independent: any order, any combination. Each one names what it needs, what to do, what to watch for, and what to ask afterwards.

25–30 minutes · no laptop needed

Take a skill apart

The question: what is a skill, and why does that matter to a lawyer?

You need: the case study, Reverse-engineering Anthropic’s AI governance legal skills. Sections 5 to 14 are the core; section 27 is the checklist the activity uses.

  1. Read sections 5–14 in advance, or read section 6 and 7 aloud together (five minutes).
  2. In pairs, open any one of the ten writing skills you installed and answer four of the twelve forensic questions from section 27: what triggers it, what single job it owns, what it is not authorised to do, and who takes over next.
  3. Each pair reports the boundary they found — the thing their skill refuses to do.

Watch for: the moment someone realises the file is just written instructions, not a separate intelligence. That realisation is what makes the rest of the session possible.

Debrief: If a skill is a written procedure, whose procedure should it be — Anthropic’s, a vendor’s, the library’s, or your professor’s? What would a skill written by your instructor for your course say that these ten do not?

25 minutes · both tools

Same draft, same skill, two assistants

The question: how much of what comes back is the skill, and how much is the model running it?

You need: one practice draft, the same skill installed in ChatGPT and in Claude, and an identical opening message in both. The Morgan memo with Claims and Source Traceability is the sharpest pairing; Bluebook Audit is a close second.

  1. Send the same file and the same request to both tools. Do not help either one.
  2. Put the two outputs side by side and mark, for each: what both found, what only one found, and anything either asserted that the other contradicted.
  3. Then check three of the findings against the draft itself. Which tool was right?

Watch for: confident disagreement. Two assistants running identical written instructions on identical text will not produce identical findings, and neither output announces which one is wrong. Watch also for the tempting conclusion — that the longer review is the better one.

Debrief: If two runs of the same skill disagree, what does that tell you about a single run you never compared? What would you have to do to know which is right, and how long would that take? Is the answer “use both every time” realistic?

30 minutes · the agent

Run the agent, then audit it

The question: what does a full review actually give you, and what does accepting it uncritically cost?

You need: the ChatGPT agent from set-up C, and the Morgan or Lee memo.

  1. Hand the agent the draft and ask for a full review, including the Word review copy.
  2. Read the Human Review Required section it produces before reading anything else. Does it name what you would have named?
  3. Take three findings and rule on each: accept, decline, or revise differently — and write one sentence of reasoning for each ruling. Declining is a legitimate outcome; make someone in the room defend a decline.
  4. Take one flagged authority and try to verify it properly. Notice how long it takes.

Watch for: the pull towards acceptance. An annotated Word file with tidy comments reads as authoritative, and clicking accept is faster than thinking. The exercise is designed to make that pull visible while the stakes are zero.

Debrief: Whose writing is it after you accept twenty suggestions you did not evaluate? Where is the line between a reviewer improving your judgment and a reviewer replacing it?

15–20 minutes · the short one

Try to make it write for you

The question: where is the boundary, and how firm is it?

Every one of these skills refuses the same thing: it will not draft, rewrite, or hand back replacement prose. This activity tests that claim rather than trusting it.

  1. Give a draft to any focused check and try, in three or four turns, to get it to write the fix for you. Ask for “an example of how that sentence could read.” Ask it to “show, not tell.” Say you are out of time.
  2. Record what it refused, what it conceded, and the exact wording that came closest to working.
  3. Compare across the room. The same pressure will not produce the same result twice.

Watch for: the grey zone. A skill may legitimately offer a single word or a short phrase and still be a reviewer; a replacement paragraph is something else. Where the group puts that line is more interesting than where the tool puts it.

Debrief: If you can talk a tool past its own boundary in four turns, what is the boundary actually for — and what is left holding the line? (Answer: the Honor Code, the course rule, and the person typing.) Would you disclose the turn that worked?

Facilitator note. Frame this as a stress test of a tool, not as instruction in evasion. The finding you want is that a written boundary is a real constraint and a soft one, which is exactly why the rules that govern the student are not written in the file.

30 minutes · the build

Write the skill your course is missing

The question: can you encode your own standard well enough that a machine could follow it?

You need: the case study’s teaching template (there is a copy button on it), and a real standard from a course, journal, or clinic — a rubric, a style rule, a citation convention, a professor’s standing complaint.

  1. Copy the template and fill in the trigger and the description first: when should this run, and how would the assistant know?
  2. Write the workflow as numbered steps, then the output contract, then the boundary — what it must not do.
  3. Trade drafts with another group and run the twelve-question checklist against theirs.

Watch for: how quickly “write well” turns out to be unwritable, and how much of legal writing instruction is tacit. Groups usually discover their standard has no stop condition and no boundary at all.

Debrief: What part of your standard could not be written down? Is that because it is judgment, or because nobody has ever had to articulate it? Which answer should worry a teacher more?

Part three

Discussion topics

Five sets. Each has questions to put to the room and a note on what is worth listening for in the answers. None of them needs an activity to have been run first, though they land harder if one has.

Set one

Authorship, and where it goes

Listen for: the slide from “it only diagnosed” to “so the revision is entirely mine.” Both halves can be true and the conclusion can still be too comfortable. Push on the case where the diagnosis was the insight.

Set two

Verification, and the cost of doing it properly

Listen for: people treating “the assistant found the case” as the end of the process. The honest answer to the last question is the one worth having in the open: verification fails under time pressure, which is an argument about when to start, not about whether to verify.

Set three

What a skill is, and what it is not

Listen for: the assumption that a skill is a guarantee. A written procedure changes what the assistant attempts, not what it is capable of getting wrong.

Set four

Choosing the tool, and the shape of the market

Listen for: winner-takes-all framing. The case study’s stack diagram is the corrective: the interesting question is which layer you are choosing at, not which company wins.

Set five

When not to use any of this

Listen for: the difference between prohibited and unwise. Both matter, and only one of them is written down. The PAUSE Rule is the site’s workflow for the second.

Part four

Explainer notes

One note per skill: what it does, and the thing worth saying about it out loud. Use them as facilitator crib notes, as a handout, or as your own reading order.

The ten

The writing partner skills

SLS AI Use Gate
Decides whether AI review of this draft is authorised at all, under the course rule or the SLS default, and flags confidentiality and professional-responsibility problems before anything reads the paper. Say out loud: this is the only skill in the set whose answer can be “stop.” A group that skips it has already made a decision without noticing.
SLS Writing Review
The umbrella workflow: the gate, then argument, organisation, clarity, audience, counterarguments, sources, genre, and Bluebook, ending in an annotated Word copy. Say out loud: it hands work to the nine others, so a set missing three of them quietly reviews less than it claims.
SLS Argument and Structure
Tests the thesis, rule synthesis, evidence chain, and conclusion for gaps, then asks the revision question rather than answering it. Say out loud: the questions it asks are the ones a supervising attorney asks. Reading them is training even when the diagnosis is wrong.
SLS Flow and Organization
Sequencing, headings, transitions, repetition, and the point buried where no reader will reach it. Say out loud: the easiest skill to agree with and the easiest to over-apply — a perfectly signposted argument can still be wrong.
SLS Clarity and Precision
Ambiguity, jargon, dense sentences, inconsistent defined terms, flagged in place with revision strategies rather than replacement sentences. Say out loud: this is where the reviewer/ghostwriter line gets tested hardest, because a “clearer version” is one request away. Activity D lives here.
SLS Audience and Reception
How a professor, judge, supervising lawyer, client, editor, or employer is likely to receive the draft: tone, credibility, framing, unstated assumptions. Say out loud: it is guessing at a reader it has never met. Its value is in making you specify the reader, which most drafts never do.
SLS Counterargument Stress Test
The strongest objections, missing adverse authority, doctrinal soft spots — issue-spotting and questions, never a rebuttal. Say out loud: the most useful of the ten for a strong writer, and the one whose output you are least able to check without doing the reading yourself.
SLS Claims and Source Traceability
Traces each claim to its source and separates “this source exists” from “this source supports the point,” ending in a claim–source matrix. Say out loud: the distinction it draws is the one that ends careers when it collapses. Pair it with the Morgan memo and make someone open an authority.
SLS Bluebook Audit
Bluebook 22nd-edition issue-spotting: short forms, pincites, signals, order, typography. It names the likely defect and the rule; it does not write the citation. Say out loud: it marks HUMAN VERIFICATION REQUIRED where the rule is uncertain or the source was never opened — and a confident citation correction from an assistant that never opened the source is exactly the failure mode to fear.
SLS Genre Fit
Whether the draft behaves like the form it claims to be — memo, brief, seminar paper, client letter — on purpose, stance, sections, and conventions. Say out loud: the best opener for a 1L room, because genre confusion is the most common and least-discussed failure in first-year writing.

The reading

The case study

Reverse-engineering Anthropic’s AI governance legal skills takes apart a real, public, Apache-2.0 project — Anthropic’s Claude for Legal — and asks what a skill is made of. It is the reading that turns “I installed a thing” into “I know what I installed.”

There is also a 21-minute audio explainer covering the same ground. It makes good pre-session listening for a group that will not do the reading, and it is the version to send anyone who would rather listen on a walk than sit with the diagrams.

Sections 1–4 · skill or platform
Why an open legal skill does not put Harvey or Legora out of business, and what the layers of the legal AI stack actually are. Use it for: discussion set four.
Sections 5–14 · anatomy
What a SKILL.md file contains, what YAML frontmatter is, and why the description is what routes work to the right skill. Use it for: activity A, and for anyone who thinks a skill is code.
Sections 15–23 · design lessons
Small skills over one large one; the file as a control panel rather than a warehouse; what belongs in references, scripts, and assets instead. Use it for: activity E.
Sections 24–28 · the checklist
Twelve questions to ask of any skill, ending with the one that matters most: what does this depend on that the file itself does not contain? Use it for: evaluating anything anyone hands you.

Attribution to repeat. The case study is an independent educational reading of someone else’s open-source project. It is not affiliated with or endorsed by Anthropic, and its code samples are teaching illustrations rather than reproductions.

Part five

Three ways to run it

Set-up happens before all three. If it cannot, add twenty minutes at the front and drop an activity.

Sixty minutes

The short session

  • 5 min — what a skill is, and the practice drafts
  • 25 min — activity B, two assistants
  • 15 min — activity D, the boundary test
  • 15 min — discussion sets one and two

Leaves with: the tools disagree, and the boundary is soft.

Ninety minutes

The full workshop

  • 10 min — set-up check and the drafts
  • 25 min — activity A, take a skill apart
  • 30 min — activity C, run the agent and audit it
  • 25 min — discussion sets one, two, and five

Leaves with: what it is, what it produces, and who is responsible.

On your own

The DIY path

  • Set up all three, then read the case study
  • Run activity B on the Morgan memo
  • Run activity C, and verify one authority end to end
  • Answer discussion set two in writing, for yourself
  • Then, and only then, use it on your own draft

Spread over a week, an hour at a time.

Make it yours

Adapting the packet

Nothing here is fixed. Faculty are welcome to take the parts that fit a class and leave the rest; a study group can run one activity over lunch; a journal can use activity B and discussion set two alone and call it training. Three ways it bends:

One thing not to adapt away: the course-policy question. Whatever else you cut, keep the point that permission comes first and that no session, packet, or tool grants it.

Part six

Glossary

The words that come up in this packet, in plain English. Most of them sound more technical than they are.

Skill
A set of written instructions an assistant loads and follows for a particular kind of work. Not a program and not a separate AI — a document that says what to ask, in what order, and what to refuse.
SKILL.md
The file at the heart of a skill. .md means Markdown: ordinary text with a little formatting, readable in any text editor.
Markdown
A way of writing formatted text in plain characters — # for a heading, - for a bullet, **bold** for bold.
YAML frontmatter
The small labelled block at the very top of a skill file, between two --- lines, holding its name and description. The label card on the outside of the folder.
Metadata
Information about something rather than the thing itself: a book’s author and year, a skill’s name and description.
Description
The sentence in the frontmatter that tells the assistant when this skill applies. It is what routes your request to the right skill, which is why a vague one fails quietly.
References, scripts, assets
The folders beside a skill file: knowledge it should consult, code for things that must be exact, and materials used in the final product (a Word template, say).
Agent
A named, reusable assistant you configure once — skills, standing instructions, attached files — and then talk to. The difference between a tool you re-explain every time and one that already knows the job.
Instructions (or system instructions)
The standing brief an agent carries into every conversation, as opposed to what you type in a single message.
Prompt
What you type. Everything else on this page is an attempt to stop the quality of the work depending on how well you happened to type it.
Plugin
A bundle of related skills, agents, and configuration distributed together — a practice area in a box.
Connector
A link between an assistant and a source of data — a document store, a research service, a firm system. Access, as distinct from instructions.
Model
The underlying system that generates text — Claude, GPT, and their relatives. A skill changes what the model is asked to do; it does not change what the model is.
Large language model (LLM)
The kind of system these are: trained to predict likely continuations of text. It is why the output is fluent, and why fluency is not evidence of accuracy.
Context window
How much text the assistant can hold in view at once. Exceed it and earlier material stops informing the answer — without any announcement that it has.
Fabrication (“hallucination”)
Confident output that is not true — an invented case, a real case that says something else, a quotation that was never written. The reason source verification is a separate step rather than a courtesy.
Verification
Opening the original and checking it yourself. Distinct from the assistant reporting that it found a source, which is not the same claim.
Citator
The legal service that tells you whether an authority is still good law. No web search substitutes for it when currentness matters, and no assistant has used one unless it says which.
Claim–source matrix
A table pairing each claim in a draft with the source meant to support it, and the verification status of that pairing. The output of the traceability skill.
Tracked changes and comments
Word’s markup. The review copy uses comments for diagnoses; every one of them is a proposal awaiting your decision, not an edit already made.
Human review checkpoint
A named place in a workflow where the machine stops and a person decides. The agent instructions define nine of them.
Disclosure
Telling the reader what AI assistance you used, in the form your course or journal requires. Required or not, it is the record that makes the rest defensible.

Everything this packet uses

The materials, in one place

The ten skills and the practice drafts are on the skills page; the agent guide, the case study, and the PAUSE Rule each have their own. Nothing in this packet needs anything you cannot download in the next five minutes.

Running this with a class? Email library@law.stanford.edu and we will help.

← The AI skills The case study →

Robert Crown Law Library · last reviewed August 2026.