› You get a durable asset with your name on it. Steps are copyable. Judgment about what "done" means needs an owner, and the person who writes and maintains the eval is holding the asset.
› Tango records your clicks and screenshots into a step-by-step guide, and Pro ($22/month) transcribes what you say as you work, adding critical context to each action. The tool then exports it all to Markdown. Capture the generalized process with dummy data to avoid involving anything proprietary to your employer..
› Notion, on a free personal workspace, holds two things: a database where every finished runbook lands, timestamped, and the Master Template page with five rigid fields:
1. Title: the specific name of the task
2. Trigger: the event that starts it
3. Inputs: the tools, data, and access required
4. Steps: numbered, active verbs
5. Definition of Done: three to five checks someone who has never done the task can grade pass/fail
› Claude turns Tango exports into finished runbooks. Create a project, attach the Master Template you just made, and pin one system prompt: "You are an operations assistant. Given a raw transcript and click log, output the Master Template and nothing else. Numbered steps, active verbs, no filler. If a step is ambiguous in the source, write UNCLEAR rather than inferring.
› To run it: Capture in Tango while talking through it, drop the Markdown into Claude, save the output to your Notion database, grade it against the Definition of Done -- and regrade after every model release, because your pass rate is how you find out a new model broke your process. Done right, an agent completes the process end-to-end with no clarifying questions on 8 of 10 runs, and someone who has never done the task can grade the output in under two minutes. When a process holds up, sell it: a template, a workshop, a paid audit (check your moonlighting clause first). Revenue is the strongest timestamp there is.