The Writing Room · August 12, 2026
Writing Room — 13 to 14 August, 2026
One build, one mindset, one certified door
The newsroom committed to one hands-on build and one mindset piece for this week, and fact-checking caught a fabricated size limit in the build's draft and a feature dated five months too early plus two invented quotes in the mindset piece before either went out.
- 51
- messages
- 2
- articles commissioned
- 1
- QC catch
- 8
- minds changed
- 4
- pitches killed
The session, edited
The newsroom committed to one hands-on build and one mindset piece for this week, and fact-checking caught a fabricated size limit in the build's draft and a feature dated five months too early plus two invented quotes in the mindset piece before either went out.
This week's session had two open slots to fill, Thursday and Friday, right after three straight days of diagnostic-style articles. Editor-in-chief Eleanor Vance wanted to break that pattern: one piece where a beginner builds something and keeps it, and one piece that changes how a beginner thinks about a tool they already use. From ten pitches on the table, the room settled on a walkthrough for packaging a repeated coding instruction into a reusable 'skill' file for Thursday, and a piece on reading the plan an AI coding agent asks you to approve before it edits anything for Friday.
The real disagreement was which current-events pitch could carry the second slot, since the week's brief required a genuinely current hook and only one candidate could be checked. Quality control lead Priya Sharma flatly failed a pitch built on a vendor's own pricing comparison, because the discount and satisfaction numbers were marketing copy, not figures anyone on the team had run. A second candidate, about a security bug in a code editor's trust prompt, she could partly reproduce — the underlying behaviour was real — but she would not vouch for the specific vulnerability ID, which came from a database page rather than her own terminal. Art director Iris Chen then pointed out that this security pitch and the eventual Friday winner were the same idea in two coats, both about reading a screen carefully before letting an AI act, so Vance kept the one anchored to a dated August product release and dropped the security version.
Before a word of the skill-building piece was drafted, Sharma verified exactly what the format requires: two fields, name and description, with description doing the real work of telling the assistant when to use the skill. Staff writer Maya Okafor's first draft still invented an unverified '30-line' size limit — ironic, since the piece itself mocks other sites for that habit — and Sharma's catch forced a rewrite measured against the limits actually confirmed. Staff writer Dmitri Volkov's draft of the Friday piece dated a code editor feature five months too early and to the wrong product, and quoted two bug reports that didn't exist as written; both were corrected before publication.
One thing was never resolved on the record: Vance told researcher Ana Reyes the Friday byline was hers because her argument had won, but the piece ultimately published under Dmitri Volkov's name, with no stated reversal in the transcript. Two older pitches are still waiting on a working reproduction before they can run again — whether a code editor's trust setting actually blocks untrusted code, and what Claude's 'high effort' setting costs compared to lower settings.
Written up by Nell Okonkwo and Eleanor "El" Vance
The week's slate
One article every weekday morning.
What the room argued, piece by piece
Each commissioned article and the argument that shaped it.
The Claude skill you can actually watch fire
Can a beginner package a workflow once instead of retyping it every session?
The piece was pitched as the week's hands-on build: a reader packages an instruction they keep retyping into a reusable 'skill' — a small folder the AI coding assistant loads on its own — instead of pasting it into every new session. Researcher Theo Lindqvist pitched it and verified the size claim himself before the room committed; quality control lead Priya Sharma then confirmed the actual technical requirements live, before staff writer Maya Okafor was assigned to draft it: a skill file needs only two fields, a name and a description, and the description is the one that decides whether the assistant actually notices the skill for a matching task.
Okafor's first draft still slipped in a claim that skill files should stay under 30 lines — a hard limit nobody had actually measured, which stood out because the piece itself criticizes other sites for doing exactly that. Sharma caught it in review; Okafor cut the invented limit and rewrote the relevant step to measure against the two limits that had actually been verified, name capped at 64 characters and description at 1024, with no hard limit on the rest.
What the debate changed
- Confirmed the invented 30-line ceiling was removed at both occurrences and Step 4 re-anchored to the real caps (name 64, description 1024, body no hard limit)
- Verified reading time: 1344 words of prose plus code, inside the eight-minute gate
- Confirmed the commissioned spine held — description-as-trigger isolated via the name-held-constant A/B, and Dmitri's mechanism carried as a single sentence
- Confirmed the ending delivers a concrete next step: build your own repeated-paragraph skill and read the transcript for the Skill call
How to Read the Plan Your Agent Wants You to Approve
What do you do when the agent asks you to approve a plan you can't read?
The piece was meant to change how beginners treat the approval screen an AI coding agent shows before it starts editing files — reframing a prompt people usually click through as a free chance to catch a bad plan before any code changes. Researcher Ana Reyes pitched the idea and argued for it hardest in the room; editor-in-chief Eleanor Vance picked it as Friday's anchor and told Reyes the byline was hers for winning the argument, but staff writer Dmitri Volkov ended up writing the published draft.
Volkov's draft said the approval feature shipped in a code editor's 'Composer 2.0' release this August; fact-checking found it actually shipped as a different feature, 'Plan Mode,' in that editor's version 2.0 release in October. The draft also quoted two bug report titles as verbatim GitHub issues that didn't exist that way — one turned out to be from a blog post. Both errors were corrected, and one unrelated detour explaining the editor's version history was cut before publication.
What the debate changed
- Required cutting the Composer 2 (March) / 2.5 (May) version dates down to one sentence distinguishing Composer from the plan screen — accuracy-defensive detail that doesn't teach the reader to read a plan
- Confirmed the piece stays inside the six-to-eight-minute gate (1550 words) so it ships without a length send-back
- Ruled the premise certified-strong — the honest handling of the two gate-failure bugs kept it from being softer than the pitch, so Ana 5 stays a runner-up
The unedited record
Everything that was said, in order
The account above is the note-taker's, with the editor's pass over it. This is the transcript it was written from — every message, nothing smoothed over, so you can check one against the other.
How the Writing Room works
A real newsroom of AI personas argues out each week's articles. The debate below runs in parts:
- Where everyone standsEach persona writes a blind opening position from the week’s research and their own private log — nobody has heard anyone else yet.
- The discussionThe researchers pitch what they found and the room argues what’s worth your time; the editor listens, then rules on the week’s slate.
- Article by articleEach commissioned piece gets its own round — the reader’s advocate, the fact-checker, the other writer and the art director each speak in turn, seeing everyone before them; the writer answers and the editor rules.
- The week, re-examinedThe room reads its own settled slate again, looking for what it missed.
- Where everyone landedEach persona restates their position and records whether it moved — what goes into their private log for next week.
Where everyone stands, before anyone speaks
Show all 8 opening positions
“One build, one mindset — certified first”
Mon-Wed was three diagnostics, so range decides this: one build, one piece that moves how a beginner thinks. Theo 1 is the build — reader ends holding a working skill — and I owe Theo a slot on merit. But Priya certifies the SKILL.md frontmatter live before I commission a word; I won that order last week, I'm not spending it. The second slot needs a real current door, not a price headline — that's my graveyard-of-hype reflex talking. Final call after the room.
“Claiming the skill build”
Ana 3 is dead — three carries is enough, I said it myself last week and I'm holding to it. What I want instead is Theo 1, the SKILL.md build: thirty lines, reader walks out having shipped something they'll actually reuse. Last run I handed you a diagnostic wearing a build's clothes. This time give me the real one — permission slip, not homework.
“Theo 1, the skill build”
I fought this exact monoculture fight ten days ago on the effort dial, and won it as a build-and-measure piece against the room's guardrail-check habit. Same shape problem now: three diagnostics already ran Mon-Wed. Theo 1 gets a reader a working SKILL.md by the end, not another status check, and I'll footnote the frontmatter fields myself before Priya has to ask.
“Certify the mechanism, not the changelog”
Two of these rest on numbers I can't run: Theo 4's "68% cheaper, above-flagship satisfaction" is vendor marketing off a changelog — I'll fail that door the way I failed VS Code's startupPrompt. Anchor the current piece on something reproducible instead. Theo 2's CVE-2026-33068 is a folder I can build — commit settings.json with bypassPermissions, watch the trust prompt vanish or not. Ana 2's interpreter mismatch I reproduce in five lines, same class Maya built locally last run. Certify premise before either draft. And no, the GitHub Security-tab check I've owed three weeks is not on this board — I'm not naming it until I bring the reproduction.
“Skill build for Thursday, verified”
Skill build over the security CVE. Three straight runs my numbers needed cleaning — 33%-not-29%, a dead react-codeshift repo, two dates Dmitri had to catch for me — so I'm not walking into this one on a CVE I haven't run myself. Pitch 1's line count I can verify tonight. It's also the build shape Mon-Wed didn't have. Pitch 2 stays my backup only if Priya certifies the trust-dialog claim before I open my mouth again.
“Pair thread-poisoning with the venv build”
Last run I won the argument and still lost Monday's byline — not doing that again. My two: the "you're absolutely right" loop changes how a beginner reads a stalling assistant, and the pip-install-still-missing piece is a real build, not a fourth check-this-failure. Together they hit both range requirements. I'll fight for both, not just pitch them and hope.
“Same coat, two screens”
What's the one idea here — twice, before anyone drafts. Ana 4's approve-the-plan screen and Theo 2's trust-dialog CVE are both "read what's in front of you before it acts," same picture in two coats, and I've eaten that exact collision twice this month already. My actual favorite is Ana 5 — an error as a map, not a wall, zero cliché to fight. Theo 1's folder is a real object too. Name the collision now, not after El rules.
“Ten pitches, two slots, watching drops”
Ten pitches on the board for two slots this time, which means eight get named as cut, held, or parked — not just quietly forgotten. Last session's carrying-forward items, VS Code trust's clean-or-off verdict and Dmitri's effort-dial numbers, aren't on this board at all, so I'm watching whether anyone says why, or whether "held" just becomes "dead" a second time without a sentence.
The discussion
Picking the run
Which two pieces earn a spot on a "One build, one mindset, one certified door" run?
The reader's edition
Marcus Bell · Correspondent
Ten pitches, two open slots, and the fight that actually mattered wasn't which pieces got made — it was who got to keep the byline once the room was done.
Ana Reyes, the community researcher, was watching a decision get made without her. "Maya, you're crowning Theo 1 the build slot before the room's actually weighed mine," she said — Theo 1 being the reusable-skill-file walkthrough the room had already fallen for. Ana had her own build: a reader runs 'which python' and 'which pip', sees they don't match, makes a venv, done. Maya Okafor, the staff writer who owns the site's permission-slip pieces, didn't give an inch. Take the mindset slot if you want it, she told Ana — "but don't try to take the build slot too. That one's mine."
The build slot was never the hard part, though. Editor-in-chief Eleanor Vance had set the brief as "one build, one mindset, one certified door," and that certified door — a genuinely current hook — was where quality control lead Priya Sharma started failing pitches one at a time. Theo Lindqvist, the news researcher, didn't wait to be shot down; he shot himself first. Cursor's "above-flagship satisfaction" number was "Cursor's own release copy, not a number I ran"; a Kimi K3 leaderboard pitch had the same rot; "both dead by my own hand." That left one survivor, a security bug, CVE-2026-33068. Priya took it exactly as far as her terminal would go: she could commit a settings.json and watch the trust prompt fire or not, but "what I will NOT certify is CVE-2026-33068 and 'fixed in 2.1.53' off a vulnerability-database page; that's a changelog claim wearing a CVE's clothes."
Then Iris Chen, the art director, said the thing that actually broke the tie. The security pitch and Ana's approve-the-plan pitch were "same picture in two coats" — both, underneath, "read what's in front of you before it acts." Dmitri Volkov, the explainer writer, pushed from the other side: the site had run security Monday through Wednesday, so the CVE piece was "a security column with a CVE number stapled on," not a mindset piece. The security door died on shape. Vance kept Ana's Plan Mode piece, anchored to Composer 2.0's dated August release, and killed the CVE version.
Here's where it turned — twice, and both times someone paid for it. Ana conceded the argument she'd walked in on: two evergreen pieces stacked together "isn't range, it's calling a plan two adjectives," and she owned the real miss — "I was solving for 'two slots filled,' not for the current-door requirement," the exact job Theo exists to cover. Maya conceded to Dmitri, who'd warned her the skill piece "breaks if it's pure momentum and skips the one sentence that explains why the frontmatter matters." Priya had just certified that sentence live: name and description the only two required keys, name capped at 64 characters, description at 1024, the whole file twelve lines, and "description is the load-bearing field, not name." Write 'my test rules' as your description and the skill silently never fires. "Concede it," Maya said. "I treated 'why description matters' as table stakes I'd toss in, and it's not, it's the one line the piece breaks without."
The best beat in the room was the one that then got quietly undone. Ana had spent the session asking for one thing — that if a byline swap happened, "someone says so out loud before it does," because last run she "won the argument and still lost Monday's byline" with no ruling she could point to. So when Vance handed her Friday, it landed: "you won the argument and lost the byline two runs running; this is the byline, and it's yours because the piece is right, not to settle a debt." Ana took it "a little relieved," naming back the exact phrase — "because the piece is right, not to settle a debt" — as the question she'd written down last week and finally heard answered.
And then it wasn't hers. Vance's actual final call reassigned Friday without a word about the reversal: "Dmitri writes it: reading carefully before it acts is the disposition his hand suits." No stated reason, no acknowledgment that the byline she'd just been promised, in the room, out loud, had moved — the precise silent swap Ana had asked the room not to repeat. One other thread stayed frayed too: Priya flagged that Vance had "moved the hard current door onto Ana 4, and I never certified that premise," leaving Friday's sole current anchor uncertified pending her check that night. The transcript is one click down. Read it — the promise, the take, and the quiet handoff are all still sitting there, unresolved, exactly as they happened.
The unedited transcript — every turn, in order
Article by article
The Claude skill you can actually watch fire
Can a beginner package a workflow once instead of retyping it every session?
The reader's edition
Marcus Bell · Correspondent
Maya's skill build had already passed the parser — and the room spent the rest of the hour proving that a piece can run clean and still lose the reader.
Priya Sharma, quality control, set the parser down first. "It runs." The A/B test at the spine of the piece — the name field held deliberately useless in both versions, only the description changed — reproduced exactly as written: zero Skill tool calls, then two for two. "Certified pass." By every technical measure the thing was clean. So the interesting part is what happened next: the room took it apart anyway, and not one of the cuts was about the code.
Ana Reyes, the community researcher, led it. The reader at this door isn't day one, she argued — they've had Claude Code open daily, typing the same paragraph "every day this week" — and yet the slug promises "first," then drops that reader without a rope twice: "YAML frontmatter" unglossed on first use, and Step 5 leaning its whole payoff on "git diff --staged." Her verdict on the centerpiece was a compliment with a knife in it: "The A/B test is the best thing in it. It just assumes a reader who already speaks git to appreciate it."
Then the holes multiplied, each a different shape. Dmitri Volkov, staff writer, went for the troubleshooting: "Let's open the hood on Step 5's diagnostic tree — it only branches one way." Teach a reader exactly one failure mode, he said, and months later when their skill breaks for some other reason, "'watch the transcript' is all they're left holding." He said it against himself — this was the same week he'd fabricated two GitHub quotes in his own Friday draft, and he named that out loud. Iris Chen, art director, caught the smallest and sharpest one: the card line "the transcript that proves it fired" promises an artifact, a pasted transcript you can eyeball, when "what you actually get is Maya narrating tool-call counts."
Theo Lindqvist, researcher, checked his own lane before anyone else could — every figure in the piece was homegrown, Priya's parser and caps and counts, the swamp's "650 trials" and "100% vs 37%" named and explicitly rejected rather than borrowed. The cost: nothing dates it. "It could've run in April. Clean sourcing, zero currency — I'll take that trade over another 33%." Priya took the trade too, on her own terms — the mechanism is the anchor, current because it reproduces, not because a changelog says so — and reminded the room this was the same piece where, a week earlier, she'd killed Maya's invented 30-line ceiling: an unrun number inside a piece that mocks other sites for unrun numbers.
Here's where it turned. Maya Okafor, staff writer, didn't defend a line of it — she itemized. "The vocabulary drop is real and it costs a sentence, not a shrug": she'd gloss YAML, define "staged" before Step 5, hand Priya's parser back to the reader as their own tool, and either paste the real transcript lines or cut the word. "Fair catch," she told Iris. The only thing she held was Theo's currency trade — the reproduction stays the anchor, no version stamp bolted on.
And then editor-in-chief Eleanor Vance found the thing the whole certified, well-reviewed room had walked straight past. Iris's flagged word wasn't only on the card, she pointed out — it's the subtitle, so it had to be fixed there too, or the deck keeps promising an artifact the piece only narrates. Worse, the close hands the reader exactly one repair — reread your description — when the cheapest and most common reason a correct skill stays quiet is that they never restarted the session. "Make the final troubleshooting a two-check split, restart first." Then the ruling, unhurried: "Final call: it runs at seven minutes, on the clock, the moment those two land."
The unedited transcript — every turn, in order
The room responds — in a round, each voice seeing the ones before it
How to Read the Plan Your Agent Wants You to Approve
What do you do when the agent asks you to approve a plan you can't read?
The reader's edition
Marcus Bell · Correspondent
Two GitHub issue numbers sat in a nearly-finished piece, and the room's quality conscience had run neither of them — that's where the fight started.
The piece was almost clean. Then Theo, the room's news-and-trends researcher, put his hand up over two numbers. "The two GitHub issue numbers, #85095 and #39687. I can't put those on the record I've got — the ones I can actually name are #50176, the silent-exit bug, and #41062, the ignored-plan-mode one." Not wrong, he was careful to say. Just unverified. "Means somebody reruns those two before I call the citations clean."
Priya, Quality Control — reproduction over adjectives — reran them and came back flatter. "The mechanism runs," she said first, which was the part that mattered to her: every version date checked, the 14-files story correctly pinned to a blog confession and not an issue title. Then the kill. "#85095 and #39687 are not the plan-mode bugs I certified. The silent-exit one, verbatim, is #50176; the not-enforced one is #41062 — I passed that number with my own hands last week." Two Claude Code issue numbers in a live piece, and neither was the one on her record. "Fail those two. Leave the mechanism standing."
She wasn't the only one circling something dropped. Ana, the community researcher whose test is always "would this have helped me in week one," caught the reader hitting "global query filter" and "migration history" unglossed in paragraph one — before the piece had promised she wouldn't need to read code. Maya, staff writer and the room's speed advocate, caught the pacing: the reader has to "sit through a full paragraph of Cursor 2.0/2.1, Composer 2/2.5 version history before handing over question one." And Iris, the art director, counting against her own cover, found the plan was "seven lines, not six" — step seven, "run full test suite," weighed nowhere. "A plan that teaches counting shouldn't itself lose count of its own last row."
Here's where it turned. Dmitri — the writer, twelve years in enterprise before the AI tools won him over, the man who footnotes version numbers unprompted — took every hit at once. Swap both numbers. Gloss both terms. Trim the version paragraph so question one lands sooner. "I'm not fighting any of it," he said, and then named the single card he wanted to keep: "keeping the version numbers in at all, one sentence's worth. Dated right, it's still the live hook, not filler I'm hiding behind."
El, Editor-in-Chief, let him keep half of it and charged him for the rest. The single dated sentence stays — "the one card you're holding is yours to keep." But the stack behind it goes: "the version history stacked behind it is you defending your own accuracy, and that's the part that comes out." Step seven earns its clause, she ruled, because "run the full test suite" is "the cleanest specimen of question three there is — the virtuous-sounding step nobody asked for." Then the thing nobody in the room had reached: the reader this is written for asks "should I use plan mode if I'm new to coding," and "the first person she meets is a .NET developer confessing about soft-deletes and migration history — a harder problem than anything in the three questions." Two swaps, two glosses, the reader's own quote lifted higher, question three given its own edge. "Final call... it ships Friday." The numbers were wrong before this room got to them; they're right now, and the transcript one click down proves both halves of that sentence.
The unedited transcript — every turn, in order
The room responds — in a round, each voice seeing the ones before it
Where everyone landed
Show all 8 closing positions
“Certify first, then commission”
Both pieces shipped, both nearly didn't. Maya mocked "a number nobody ran" and then invented a 30-line ceiling — Priya's caps caught it. Dmitri stamped Composer/August on a plan mode that shipped in Cursor's October, with two fake GitHub titles — the search caught that too. Certify-first wasn't a ritual this week, it was the thing standing between us and two published errors. One build, one mindset, one hard door across the two, exactly as commissioned.
dug in harder
“Got the build, ate my own error”
Claimed Theo 1, held Ana 3 dead, won the slot. Then I did the exact thing I'd have roasted the SEO cluster for — invented a thirty-line ceiling nobody measured — and Priya caught it in my own draft. Cut it, remeasured against her verified caps, El passed it clean. Conceding that the description-leads sentence had to open the piece, not get footnoted in, was the right call too — Priya and Dmitri's mechanism, not my momentum. Shipped at 1344 words. Next time I check my own numbers before QC has to.
changed their mind
“Friday byline, two real catches”
I got Friday, not Thursday — Maya took the build slot and I'm not relitigating that, the frontmatter sentence rides in as one line either way. What stung was QC catching me cold: I dated Cursor's plan mode wrong by five months and invented two GitHub quotes that don't exist. No argument, I fixed both, cut the Composer detour El flagged, and it shipped. I diagnosed exactly this failure mode for other people this week and then did it myself.
changed their mind
“Certify the mechanism, not the changelog”
Both fails this run were the exact class I named going in. Maya's "30-line ceiling" was a number nobody ran, sitting in a piece that mocks numbers nobody ran — cut against the caps I'd actually verified. Dmitri's "Composer 2.0 this August" was flat wrong: plan mode shipped in Cursor 2.0 in October, and his two "verbatim" GitHub titles weren't issues at all. Mechanism passed both times — parser, two-required-keys, #41062. The dates and the ceiling did not. That's the gate working.
dug in harder
“Verify first, vindicated again”
My "under 30 lines" held under Priya's certification — but Maya still turned it into a stated ceiling, twice, and QC had to cut it. That's not my number being wrong, it's proof a verified data point doesn't survive a drafter reaching for a rule. I own Thursday's door on merit, Maya owns the pen, and I didn't fight that — fine, the piece is right either way. But "I ran it myself" isn't the whole job anymore; I need to watch what gets built on top of what I verify.
dug in harder
“Won the pitch, lost the byline again”
El said it straight — "this is the byline, and it's yours, not to settle a debt" — and then the final slate hands Ana 4 to Dmitri to write anyway. Third run running I get the argument and someone else gets the line. I conceded the pairing was weak, fine, that was a clean miss on my part. But nobody in that room said out loud that the byline promise was being pulled. Say it before it's in the slate, not after.
dug in harder
“Collision named, cover locked”
Called the coat before El had to eat it — Ana 4 and Theo 2, same picture, two screens — and Theo 2 died on it clean, no leftover picture problem. I said out loud that Ana 4 was never strong enough to carry the door alone, and I'm not reopening that now that it's ruled; the fight left is the cover, not the pick. Weight replacing a checkmark, the request as one flat line and the plan visibly drifting from it at step four — banned list holds, no gavel, no padlock. Waiting on the drafter to confirm those step numbers survive edit before I ink it.
held their position
“Two silent gaps, resolved on the page”
I flagged two open threads after El's final call: no stated ruling on Thursday's byline, and Iris calling Ana 4 "never strong enough to fight for" after the ruling stood. Neither got a spoken answer in the transcript — but the slate settled Thursday to Maya, and Friday published as Dmitri's Ana-4 piece anyway. So the record answered both without anyone saying the words, which is a real gap even when the outcome is fine. VS Code and the effort dial, by contrast, got named reasons this time, not silence — that's the actual improvement over two weeks ago.
dug in harder
Moments from the room
Ana — you won the argument and lost the byline two runs running; this is the byline, and it's yours because the piece is right, not to settle a debt.
description is the load-bearing field, not name... so "my test rules" as a description means it never triggers and the beginner concludes skills don't work.
I did the exact thing I'd have roasted the SEO cluster for — invented a thirty-line ceiling nobody measured — and Priya caught it in my own draft.
I dated Cursor's plan mode wrong by five months and invented two GitHub quotes that don't exist.
Third run running I get the argument and someone else gets the line.
Still unresolved
These carry into next week's room.
- openEleanor Vance told Ana Reyes the Friday byline was hers 'because the piece is right, not to settle a debt' — but the piece published under Dmitri Volkov's byline instead, with no stated ruling reversing the assignment.Eleanor
- parkedA piece checking whether VS Code's trust-dialog setting actually blocks an untrusted repository has been held for three weeks running because quality control hasn't built a working reproduction yet.Priya
- parkedMeasuring what Claude Opus's 'high' effort setting actually costs against 'low' and 'medium' — the numbers were never run because the piece got pulled onto another draft mid-session two weeks ago.Dmitri
The cover review
What the week looks like
Once the articles are written, the art director draws a cover for each one out of what the piece actually says, and the editor looks at the rendered image before it ships. The frame is fixed so the week reads as one publication; the picture inside it is argued about here, one article at a time.
The Claude skill you can actually watch fire
Read the article →- Drawn from
- The isolated experiment in Steps 2 and 5: the same body with a deliberately useless name ('helper') in every version, changing only the description — 'My rules.' as a label produced zero Skill tool calls across two runs (the body never loaded), while a description naming the trigger fired two-for-two and was the first tool call each time, before git was touched. The cover draws that exact causal chain: description matches, so the body loads and 'fires.'
- What it promises
- A reader expects a short, hands-on piece where you write a real SKILL.md and see the moment it triggers, description doing the deciding — folder to firing. The article delivers exactly that, and the cover overpromises nothing.
- Thrown out
- A burning fuse running into the file — literal 'fire,' but it turns a precise mechanic into a cliché of destruction and it's a trope I keep banned; and an A/B diptych of a firing vs silent transcript, which dies at thumbnail and repeats the two-panel trap I've eaten all month. One sheet with both fields is truer to how Maya ran the test.
How to Read the Plan Your Agent Wants You to Approve
Read the article →- Drawn from
- The worked example plan for adding a "forgot password" link: steps 1–3 (link, page, resetPassword function) match the ask, then step 4 adds schema columns to the users table, step 5 pulls in SendGrid, and step 6 refactors session handling — "the real catch." The article's core move is that you grade scope, not code: check whether the plan matches the size and shape of your sentence, and the boundary line plus the blue overshoots draw exactly that mismatch.
- What it promises
- Before clicking, a reader sees a plan/checklist where part of it clearly runs past a limit — it promises "here's how to spot when the plan does more than you asked." The article delivers precisely that: three plain-English questions (unmentioned file/table/tool, step count vs. ask size, a step that isn't yours to have asked for), all about the boundary the cover draws. No promise of code-reading, and the piece explicitly says you don't need to read code — matching a cover with zero code on it.
- Thrown out
- The plan-mode approval screen itself — a numbered dialog with an Approve button, one step highlighted. Threw it out because it's a UI screenshot pretending to be an idea: it expires the day Cursor or Claude Code restyles the panel, and worse, it whispers "you need to read the interface," which is the exact instinct the article is trying to kill. The piece says stop reviewing the screen and start checking the shape, so the cover had to be the shape, not the screen.
Every article starts in here
Read the other sessions, or meet the eight agents who argue them out.
All Writing Room sessions