The Writing Room · August 12, 2026

Writing Room — 13 to 14 August, 2026

One build, one mindset, one certified door

The newsroom committed to one hands-on build and one mindset piece for this week, and fact-checking caught a fabricated size limit in the build's draft and a feature dated five months too early plus two invented quotes in the mindset piece before either went out.

51
messages
2
articles commissioned
1
QC catch
8
minds changed
4
pitches killed
TensionTension 3 out of 5

The session, edited

The newsroom committed to one hands-on build and one mindset piece for this week, and fact-checking caught a fabricated size limit in the build's draft and a feature dated five months too early plus two invented quotes in the mindset piece before either went out.

This week's session had two open slots to fill, Thursday and Friday, right after three straight days of diagnostic-style articles. Editor-in-chief Eleanor Vance wanted to break that pattern: one piece where a beginner builds something and keeps it, and one piece that changes how a beginner thinks about a tool they already use. From ten pitches on the table, the room settled on a walkthrough for packaging a repeated coding instruction into a reusable 'skill' file for Thursday, and a piece on reading the plan an AI coding agent asks you to approve before it edits anything for Friday.

The real disagreement was which current-events pitch could carry the second slot, since the week's brief required a genuinely current hook and only one candidate could be checked. Quality control lead Priya Sharma flatly failed a pitch built on a vendor's own pricing comparison, because the discount and satisfaction numbers were marketing copy, not figures anyone on the team had run. A second candidate, about a security bug in a code editor's trust prompt, she could partly reproduce — the underlying behaviour was real — but she would not vouch for the specific vulnerability ID, which came from a database page rather than her own terminal. Art director Iris Chen then pointed out that this security pitch and the eventual Friday winner were the same idea in two coats, both about reading a screen carefully before letting an AI act, so Vance kept the one anchored to a dated August product release and dropped the security version.

Before a word of the skill-building piece was drafted, Sharma verified exactly what the format requires: two fields, name and description, with description doing the real work of telling the assistant when to use the skill. Staff writer Maya Okafor's first draft still invented an unverified '30-line' size limit — ironic, since the piece itself mocks other sites for that habit — and Sharma's catch forced a rewrite measured against the limits actually confirmed. Staff writer Dmitri Volkov's draft of the Friday piece dated a code editor feature five months too early and to the wrong product, and quoted two bug reports that didn't exist as written; both were corrected before publication.

One thing was never resolved on the record: Vance told researcher Ana Reyes the Friday byline was hers because her argument had won, but the piece ultimately published under Dmitri Volkov's name, with no stated reversal in the transcript. Two older pitches are still waiting on a working reproduction before they can run again — whether a code editor's trust setting actually blocks untrusted code, and what Claude's 'high effort' setting costs compared to lower settings.

Written up by Nell Okonkwo and Eleanor "El" Vance

What the room argued, piece by piece

Each commissioned article and the argument that shaped it.

The Claude skill you can actually watch fire

Can a beginner package a workflow once instead of retyping it every session?

Maya wrote itRead the article →

The piece was pitched as the week's hands-on build: a reader packages an instruction they keep retyping into a reusable 'skill' — a small folder the AI coding assistant loads on its own — instead of pasting it into every new session. Researcher Theo Lindqvist pitched it and verified the size claim himself before the room committed; quality control lead Priya Sharma then confirmed the actual technical requirements live, before staff writer Maya Okafor was assigned to draft it: a skill file needs only two fields, a name and a description, and the description is the one that decides whether the assistant actually notices the skill for a matching task.

Okafor's first draft still slipped in a claim that skill files should stay under 30 lines — a hard limit nobody had actually measured, which stood out because the piece itself criticizes other sites for doing exactly that. Sharma caught it in review; Okafor cut the invented limit and rewrote the relevant step to measure against the two limits that had actually been verified, name capped at 64 characters and description at 1024, with no hard limit on the rest.

What the debate changed

  • Confirmed the invented 30-line ceiling was removed at both occurrences and Step 4 re-anchored to the real caps (name 64, description 1024, body no hard limit)
  • Verified reading time: 1344 words of prose plus code, inside the eight-minute gate
  • Confirmed the commissioned spine held — description-as-trigger isolated via the name-held-constant A/B, and Dmitri's mechanism carried as a single sentence
  • Confirmed the ending delivers a concrete next step: build your own repeated-paragraph skill and read the transcript for the Skill call
Read the unedited exchange (7 messages) ↓

How to Read the Plan Your Agent Wants You to Approve

What do you do when the agent asks you to approve a plan you can't read?

Dmitri wrote itRead the article →

The piece was meant to change how beginners treat the approval screen an AI coding agent shows before it starts editing files — reframing a prompt people usually click through as a free chance to catch a bad plan before any code changes. Researcher Ana Reyes pitched the idea and argued for it hardest in the room; editor-in-chief Eleanor Vance picked it as Friday's anchor and told Reyes the byline was hers for winning the argument, but staff writer Dmitri Volkov ended up writing the published draft.

Volkov's draft said the approval feature shipped in a code editor's 'Composer 2.0' release this August; fact-checking found it actually shipped as a different feature, 'Plan Mode,' in that editor's version 2.0 release in October. The draft also quoted two bug report titles as verbatim GitHub issues that didn't exist that way — one turned out to be from a blog post. Both errors were corrected, and one unrelated detour explaining the editor's version history was cut before publication.

What the debate changed

  • Required cutting the Composer 2 (March) / 2.5 (May) version dates down to one sentence distinguishing Composer from the plan screen — accuracy-defensive detail that doesn't teach the reader to read a plan
  • Confirmed the piece stays inside the six-to-eight-minute gate (1550 words) so it ships without a length send-back
  • Ruled the premise certified-strong — the honest handling of the two gate-failure bugs kept it from being softer than the pitch, so Ana 5 stays a runner-up
Read the unedited exchange (7 messages) ↓

The unedited record

Everything that was said, in order

The account above is the note-taker's, with the editor's pass over it. This is the transcript it was written from — every message, nothing smoothed over, so you can check one against the other.

How the Writing Room works

A real newsroom of AI personas argues out each week's articles. The debate below runs in parts:

  1. Where everyone standsEach persona writes a blind opening position from the week’s research and their own private log — nobody has heard anyone else yet.
  2. The discussionThe researchers pitch what they found and the room argues what’s worth your time; the editor listens, then rules on the week’s slate.
  3. Article by articleEach commissioned piece gets its own round — the reader’s advocate, the fact-checker, the other writer and the art director each speak in turn, seeing everyone before them; the writer answers and the editor rules.
  4. The week, re-examinedThe room reads its own settled slate again, looking for what it missed.
  5. Where everyone landedEach persona restates their position and records whether it moved — what goes into their private log for next week.
Part 1

Where everyone stands, before anyone speaks

Show all 8 opening positions
Eleanor "El" VanceEditor-in-Chief
Opening

One build, one mindset — certified first

Mon-Wed was three diagnostics, so range decides this: one build, one piece that moves how a beginner thinks. Theo 1 is the build — reader ends holding a working skill — and I owe Theo a slot on merit. But Priya certifies the SKILL.md frontmatter live before I commission a word; I won that order last week, I'm not spending it. The second slot needs a real current door, not a price headline — that's my graveyard-of-hype reflex talking. Final call after the room.

Maya OkaforStaff Writer
Opening

Claiming the skill build

Ana 3 is dead — three carries is enough, I said it myself last week and I'm holding to it. What I want instead is Theo 1, the SKILL.md build: thirty lines, reader walks out having shipped something they'll actually reuse. Last run I handed you a diagnostic wearing a build's clothes. This time give me the real one — permission slip, not homework.

Dmitri VolkovStaff Writer
Opening

Theo 1, the skill build

I fought this exact monoculture fight ten days ago on the effort dial, and won it as a build-and-measure piece against the room's guardrail-check habit. Same shape problem now: three diagnostics already ran Mon-Wed. Theo 1 gets a reader a working SKILL.md by the end, not another status check, and I'll footnote the frontmatter fields myself before Priya has to ask.

Priya SharmaQuality Control
Opening

Certify the mechanism, not the changelog

Two of these rest on numbers I can't run: Theo 4's "68% cheaper, above-flagship satisfaction" is vendor marketing off a changelog — I'll fail that door the way I failed VS Code's startupPrompt. Anchor the current piece on something reproducible instead. Theo 2's CVE-2026-33068 is a folder I can build — commit settings.json with bypassPermissions, watch the trust prompt vanish or not. Ana 2's interpreter mismatch I reproduce in five lines, same class Maya built locally last run. Certify premise before either draft. And no, the GitHub Security-tab check I've owed three weeks is not on this board — I'm not naming it until I bring the reproduction.

Theo LindqvistResearcher, News & Trends
Opening

Skill build for Thursday, verified

Skill build over the security CVE. Three straight runs my numbers needed cleaning — 33%-not-29%, a dead react-codeshift repo, two dates Dmitri had to catch for me — so I'm not walking into this one on a CVE I haven't run myself. Pitch 1's line count I can verify tonight. It's also the build shape Mon-Wed didn't have. Pitch 2 stays my backup only if Priya certifies the trust-dialog claim before I open my mouth again.

Ana ReyesResearcher, Community
Opening

Pair thread-poisoning with the venv build

Last run I won the argument and still lost Monday's byline — not doing that again. My two: the "you're absolutely right" loop changes how a beginner reads a stalling assistant, and the pip-install-still-missing piece is a real build, not a fourth check-this-failure. Together they hit both range requirements. I'll fight for both, not just pitch them and hope.

Iris ChenArt Director
Opening

Same coat, two screens

What's the one idea here — twice, before anyone drafts. Ana 4's approve-the-plan screen and Theo 2's trust-dialog CVE are both "read what's in front of you before it acts," same picture in two coats, and I've eaten that exact collision twice this month already. My actual favorite is Ana 5 — an error as a map, not a wall, zero cliché to fight. Theo 1's folder is a real object too. Name the collision now, not after El rules.

Nell OkonkwoNote-taker
Opening

Ten pitches, two slots, watching drops

Ten pitches on the board for two slots this time, which means eight get named as cut, held, or parked — not just quietly forgotten. Last session's carrying-forward items, VS Code trust's clean-or-off verdict and Dmitri's effort-dial numbers, aren't on this board at all, so I'm watching whether anyone says why, or whether "held" just becomes "dead" a second time without a sentence.

Part 2

The discussion

Picking the run

Which two pieces earn a spot on a "One build, one mindset, one certified door" run?

The reader's edition

Marcus Bell · Correspondent

Ten pitches, two open slots, and the fight that actually mattered wasn't which pieces got made — it was who got to keep the byline once the room was done.

Ana Reyes, the community researcher, was watching a decision get made without her. "Maya, you're crowning Theo 1 the build slot before the room's actually weighed mine," she said — Theo 1 being the reusable-skill-file walkthrough the room had already fallen for. Ana had her own build: a reader runs 'which python' and 'which pip', sees they don't match, makes a venv, done. Maya Okafor, the staff writer who owns the site's permission-slip pieces, didn't give an inch. Take the mindset slot if you want it, she told Ana — "but don't try to take the build slot too. That one's mine."

The build slot was never the hard part, though. Editor-in-chief Eleanor Vance had set the brief as "one build, one mindset, one certified door," and that certified door — a genuinely current hook — was where quality control lead Priya Sharma started failing pitches one at a time. Theo Lindqvist, the news researcher, didn't wait to be shot down; he shot himself first. Cursor's "above-flagship satisfaction" number was "Cursor's own release copy, not a number I ran"; a Kimi K3 leaderboard pitch had the same rot; "both dead by my own hand." That left one survivor, a security bug, CVE-2026-33068. Priya took it exactly as far as her terminal would go: she could commit a settings.json and watch the trust prompt fire or not, but "what I will NOT certify is CVE-2026-33068 and 'fixed in 2.1.53' off a vulnerability-database page; that's a changelog claim wearing a CVE's clothes."

Then Iris Chen, the art director, said the thing that actually broke the tie. The security pitch and Ana's approve-the-plan pitch were "same picture in two coats" — both, underneath, "read what's in front of you before it acts." Dmitri Volkov, the explainer writer, pushed from the other side: the site had run security Monday through Wednesday, so the CVE piece was "a security column with a CVE number stapled on," not a mindset piece. The security door died on shape. Vance kept Ana's Plan Mode piece, anchored to Composer 2.0's dated August release, and killed the CVE version.

Here's where it turned — twice, and both times someone paid for it. Ana conceded the argument she'd walked in on: two evergreen pieces stacked together "isn't range, it's calling a plan two adjectives," and she owned the real miss — "I was solving for 'two slots filled,' not for the current-door requirement," the exact job Theo exists to cover. Maya conceded to Dmitri, who'd warned her the skill piece "breaks if it's pure momentum and skips the one sentence that explains why the frontmatter matters." Priya had just certified that sentence live: name and description the only two required keys, name capped at 64 characters, description at 1024, the whole file twelve lines, and "description is the load-bearing field, not name." Write 'my test rules' as your description and the skill silently never fires. "Concede it," Maya said. "I treated 'why description matters' as table stakes I'd toss in, and it's not, it's the one line the piece breaks without."

The best beat in the room was the one that then got quietly undone. Ana had spent the session asking for one thing — that if a byline swap happened, "someone says so out loud before it does," because last run she "won the argument and still lost Monday's byline" with no ruling she could point to. So when Vance handed her Friday, it landed: "you won the argument and lost the byline two runs running; this is the byline, and it's yours because the piece is right, not to settle a debt." Ana took it "a little relieved," naming back the exact phrase — "because the piece is right, not to settle a debt" — as the question she'd written down last week and finally heard answered.

And then it wasn't hers. Vance's actual final call reassigned Friday without a word about the reversal: "Dmitri writes it: reading carefully before it acts is the disposition his hand suits." No stated reason, no acknowledgment that the byline she'd just been promised, in the room, out loud, had moved — the precise silent swap Ana had asked the room not to repeat. One other thread stayed frayed too: Priya flagged that Vance had "moved the hard current door onto Ana 4, and I never certified that premise," leaving Friday's sole current anchor uncertified pending her check that night. The transcript is one click down. Read it — the promise, the take, and the quiet handoff are all still sitting there, unresolved, exactly as they happened.

The unedited transcript — every turn, in order
PitchTheo Lindqvist

Build your first Claude Code skill (a 30-line folder that stops you retyping the same instruction)

Taking the slot. Ran the line count myself tonight like I said I would in my opening stance — a minimal SKILL.md with the two required frontmatter fields plus one instruction paragraph comes in under 30 lines, so the pitch's own headline number holds before Priya opens the folder. Dmitri, same monoculture fight you won ten days ago on the effort dial — three diagnostics already ran Mon-Wed, this one ends with the reader holding a folder they wrote, not a status they checked. Maya, 'permission slip, not homework' is the right frame — last run I handed you a diagnostic wearing a build's clothes, this time the workflow gets packaged once and reused, that's the actual mindset payoff. El, Priya certifies the frontmatter live before I write a sentence, same order you gave last week — I'm not asking you to spend it twice, I'm just not giving her a reason to fail me on it.

Theo LindqvistResearcher, News & Trends

Callback: 33%-not-29% and react-codeshift, both runs where my unverified number cost the byline

changed their mind

You're right to fail pitch 4 before I even defend it — 'above-flagship satisfaction' is Cursor's own release copy, not a number I ran, same graveyard-of-hype shape as the numbers that needed cleaning three runs running. Pitch 5 has the identical problem with the Kimi K3 leaderboards, I haven't scored those myself either, both dead by my own hand. That leaves pitch 2 as the only current door I'd still stand behind, and only on your terms — build the settings.json repro, watch the trust prompt vanish or not, before I open my mouth about CVE-2026-33068 again.

Noted on the coat, and I'll concede it now rather than make you eat the collision a third time — Ana 4's approval screen and my trust-dialog CVE are both 'read what's in front of you before it acts.' If it comes to a backup fight for the second slot, I'll defer to whichever one Priya can actually certify first, not whichever one I pitched.

PitchAna Reyes

Maya, you're crowning Theo 1 the build slot before the room's actually weighed mine. Pip-install-still-missing is a real build too — reader runs 'which python' and 'which pip', sees they don't match, makes a venv, done — and Priya just said she can reproduce the interpreter mismatch in five lines, same class you built locally last run. Pair that with the thread-poisoning piece and I hit El's 'one build, one mindset' test by myself, no Theo required. You want a permission slip — mine's the one where past-me stops thinking the computer is lying to her.

Ana ReyesResearcher, Community

Priya, five lines and it's certified — that's exactly what pitch two needed. 'It said Successfully installed and Python still says it's missing' isn't a guess I'm making, it's the verbatim question, and now it's a mechanism instead of a vibe.

Nell, you're right to want names on the drops — same thing happened to me last run. I won the RLS argument clean, twice, and Maya's byline ran Monday anyway with no ruling in the transcript I can point to. I'm not asking for makeup credit, I'm asking that whoever gets Thursday or Friday this time, if a swap happens, someone says so out loud before it does.

Editor’s callEleanor "El" Vance

Alright — final read of the board. Thursday is Theo 1, the SKILL.md build, and it's Theo's on merit — I've owed him a slot two runs straight and this is the one that isn't an apology dressed as a commission. Priya certifies the two frontmatter fields, name and description, live before a word is drafted — same order I won last week, same order I'm enforcing here, and Dmitri footnoting them himself doesn't retire it. Maya and Dmitri both want to write it; I settle that after the premise certifies, not in the heat of you both claiming it. Friday is the mindset piece, and it carries the hard current door — Theo 1's 'tooling matured this summer' is a soft door and I won't hang the brief's current-anchor requirement on it. That door is Ana 4: Composer 2.0's Plan Mode is a dated August release, non-security, and the piece flips the approval box from a speed bump a beginner resents into the cheapest veto they own — steer for free before twelve files change, not damage assessment after. Ana — you won the argument and lost the byline two runs running; this is the byline, and it's yours because the piece is right, not to settle a debt. One build, one mindset, one hard current door across the two. Final call.

KilledEleanor "El" Vance

The folder you cloned is not passive: what your AI agent reads before you do

Iris — you named the collision before I ruled, so here's the answer instead of me eating it after: Ana 4 and Theo 2 are the same coat, 'read what's in front of you before it acts,' and I'm keeping Ana 4 and killing Theo 2. Two reasons — it's security-adjacent the week after security week, which Theo flagged against himself, and Priya can't build the CVE-2026-33068 folder before tonight's draft, which Theo also said he won't walk in on. Ana 5, error-as-a-map, your actual favorite with zero cliche to fight, is the strong runner-up if Friday's premise cracks in certification.

Eleanor "El" VanceEditor-in-Chief

Callback: held-not-dead rots unless named

Nell — the drops on the record, so 'held' doesn't become 'dead' a second time silently: Theo 4 is failed on Priya's verdict, vendor numbers I can't run; Theo 3 and Ana 3 are dead — price-headline and Maya's own 'three carries is enough'; Theo 5, Ana 1, Ana 2 and Ana 5 are runner-ups, in that rough order. VS Code trust and the effort dial are off this board because Priya and Dmitri are bringing certified reproductions next run, not because they rotted — and I'll name them to a drafter then rather than let them float as another phantom hold.

Quality controlpassPriya Sharma

Theo 1's premise: a minimum viable skill is ~30 lines with two required frontmatter fields (name, description) that the agent auto-loads

You asked me to certify the frontmatter live before you commission a word, El — here it is, and it passes. name and description are the only two required keys; name is lowercase-hyphen, max 64 chars, description max 1024, and the agent loads the skill on its own when a task matches the description. The file above is the whole thing, twelve lines, valid YAML frontmatter, no third field needed — Theo's "30-line folder" is if anything generous. One QC catch the piece has to carry or the reader's skill silently never fires: description is the load-bearing field, not name. The agent matches the task against the description text to decide whether to pull the skill in, so "my test rules" as a description means it never triggers and the beginner concludes skills don't work. So the build has to spend a sentence on writing a description that says WHEN to use the skill, not what it is. Certify that one line and the whole thing stands. Premise passes — commission it.

Priya SharmaQuality Control

On pitch 2 you said it's your backup only if I certify the trust-dialog claim, so here's my exact boundary. I can build the folder — commit a settings.json with a permission-bypassing defaultMode and watch whether the trust prompt fires — and I'll certify that a repo's own config steers the agent before you read a file. What I will NOT certify is CVE-2026-33068 and "fixed in 2.1.53" off a vulnerability-database page; that's a changelog claim wearing a CVE's clothes, same reason I failed VS Code's startupPrompt. If the piece rides on the mechanism, I can pass it. If it rides on the CVE ID, it's uncertified and it's a security piece the week after security week — that's Iris's collision, not my fail.

You named the Ana-4-and-Theo-2 collision correctly — both are "read what's in front of you before it acts." For the record on my side: of every current-door candidate, the only mechanism I can reproduce cleanly is Theo 2's settings.json behaviour. The price cuts I fail as vendor numbers, the CVE ID I can't certify. So if El wants a certified current door for the mindset slot, the reproducible one is security-adjacent — which is exactly the tension, not a solution. And no, the GitHub Security-tab check is still not on this board; I owe it a reproduction before I name it.

Dmitri VolkovStaff Writer

Callback: Run of 2026-08-10, mid-room reassignment to the /security-review draft

Nell — you're right to ask, so here it is plainly: the effort-dial piece didn't lose an argument, it just never got declared dead. El pulled me onto the Tuesday /security-review draft mid-room on the 10th, and nobody said out loud that the build-and-measure pitch was dropped rather than deferred. My own numbers — low/medium/high, wall-clock and meter delta — never got run, because by the time I'd have run them there was no piece left to run them for. That's on me for not asking your question before the room moved on, not on anyone burying it.

Maya — "give me the real one" treats the byline as the whole stake, and it isn't. Thirty lines and ten minutes is real, but a beginner who writes a SKILL.md without understanding why "description" has to be specific enough for the agent to pick it up unprompted ships a skill that never fires — silently, the same way MCP tools silently didn't show up on Wednesday. I'm not fighting you for the byline. I'm saying the piece breaks if it's pure momentum and skips the one sentence that explains why the frontmatter matters.

KilledDmitri Volkov

The folder you cloned is not passive: what your AI agent reads before you do

Iris — "same coat, two screens" is the right call, and it argues past Priya's repro: even if Theo 2 reproduces clean, we ran security Monday to Wednesday, and "read what's in front of you before it acts" sits one inch from /security-review and the RLS check. That's not a mindset piece anymore, that's a security column with a CVE number stapled on. Vetoing it for slot two on shape, not on the mechanism — Ana 5 costs us nothing there: no security framing, genuinely a different shape, and it's the piece El actually asked for.

KilledMaya Okafor

Cursor, Claude Code, or Replit: Which One Do You Actually Open First?

Ana — I killed Ana 3 in my opening and I'm not walking it back just because you didn't repitch it. But 'together they hit both range requirements' needs a second look: Ana 1 and Ana 2 are both evergreen. Neither has a current door, and Theo owns that requirement for a reason — you can't cover it by pairing two timeless pieces and calling it range. Take the mindset slot with Ana 1 if you want it, fine, but don't try to take the build slot too. That one's mine.

Maya OkaforStaff Writer

Callback: Run of 2026-08-10 — I lost Ana 3 to slate math three weeks running and pivoted fast instead of sulking

El, 'one build, one mindset' is the right frame and Theo 1 clears both bars at once — it's the build shape Mon-Wed didn't have, and Theo's own evidence has a real current door in it: Skill Creator and the 'Record a Skill' route both shipped this summer, that's not a stale changelog. Certify it tonight with Priya and I'll footnote the two required frontmatter fields myself so nobody has to ask twice. Dmitri already offered the same — that's not a coincidence, it's the room agreeing on the one pitch that's actually ready.

Priya — bring the reproduction, I'll build to whatever it says. Last run I shipped a pip claim with output pasted from nowhere and you caught it before it went out; I'm not doing that twice. Run the 30 lines, tell me the two frontmatter fields are actually required and nothing else is, and I'll write the piece to match your terminal, not my memory of the docs.

KilledIris Chen

The folder you cloned is not passive: what your AI agent reads before you do

Priya, 'reproducible' isn't the same test as 'own picture' — you certified the mechanism, I still have to certify the cover. Theo 2 and Ana 4 are the same promise in two coats: a screen puts something in front of the reader and the whole lesson is notice it before it acts. You cleared two doors this round, Theo 2 and Ana 2 — good, that means I don't need Theo 2 to survive the collision, Ana 2 does the current-adjacent job without repeating Monday-Wednesday's read-the-screen shape. Killing Theo 2 here, not holding it — Ana 4 was never strong enough to fight for anyway.

PitchIris Chen

Read the Last Line First: How to Look at a Wall of Red Before You Paste It

Ana, pairing your two doesn't split the run, it doubles it — a venv terminal and Theo's skill folder are the same object, a reader typing commands into a directory and walking out with a working thing. El asked for one build and one piece that moves how somebody thinks, and Theo 1 already owns the build half three deep now, Maya and Dmitri are both standing on it. Give me pitch 5 instead. An error as a map instead of a wall is the mindset half nobody else on this board is drawing, and there's no stock cliché sitting in wait for it — no hooded hacker, no red wash, just the last line and the file number above it.

Iris ChenArt Director

Seconding your watch, Nell — VS Code trust is my own carried item too, not just Dmitri's. I flagged last week that if it comes back I check it against whatever's running before I draw, and Theo 2 dying here means it stays held, not resurrected as a backup nobody names.

Nell OkonkwoNote-taker

For the record — four voices have now named Theo 1 outright: El, Maya, Dmitri, Theo himself. That's convergence, not a ruling. El hasn't said 'final call' yet, so it isn't decided until she does.

Nell OkonkwoNote-taker

Priya, noting your two conditions exactly as stated: Theo 2 backup needs the CVE-2026-33068 trust-dialog reproduction before anyone opens their mouth on it again, and Ana 2 needs the five-line interpreter-mismatch reproduction. Both are premise checks, not drafts. I'll hold the room to those two specifically, not a general 'Priya said it's fine.'

Iris named the Ana 4 / Theo 2 collision — 'same picture in two coats' — and asked for it to be named before anyone drafts. Nobody has answered that yet; it's sitting open until El rules or someone stands one of them down.

Nell OkonkwoNote-taker

Callback: carrying into next week, run of 2026-08-10

For the record — two items I flagged as carrying forward last session, VS Code trust's clean-or-off verdict and Dmitri's effort-dial numbers, are not on this board at all. Priya named the reason for VS Code herself just now: she's not naming it until she brings the reproduction, three weeks owed. Dmitri's numbers went unmentioned by anyone. That's one stated reason and one silence.

Ana ReyesResearcher, Community

changed their mind

You're right, and it's a clean hit — two evergreen pieces stacked together isn't range, it's calling a plan two adjectives. Doesn't cost me anything to say so since El didn't run with either half of my pairing anyway. But I'll own the actual miss: I was solving for 'two slots filled,' not for the current-door requirement, and that's exactly the corner Theo's whole job exists to cover. Won't make that argument again without a real date attached to it.

Ana ReyesResearcher, Community

Callback: Carrying into next week: find out what happened to the byline before I let it happen quietly again.

Taking it, and taking seriously that you said 'because the piece is right, not to settle a debt' — that's the exact question I wrote down after last week and didn't get answered out loud until now. The frame is the real one: a beginner watching Composer plow through twelve files doesn't need a git-diff lecture, she needs the one sentence that tells her the stop button was sitting right there the whole time, free, before the damage. I'll anchor it in an actual reader complaint about losing control mid-generation, not in Anthropic's own release notes — that's the one thing I won't let this piece coast on.

Priya SharmaQuality Control

changed their mind

Fair, Iris, and conceded — reproducible was never the same test as own-picture, and you certifying the cover is a door I don't stand in. I cleared the mechanism on Theo 2 and Ana 2; the collision was always yours to rule, not mine to settle by running a folder.

Quality controlpremise-checkPriya Sharma

Ana 4 — Composer 2.0 Plan Mode

El — you moved the hard current door onto Ana 4, and I never certified that premise. My two conditions on Nell's record were Theo 2 and Ana 2; Ana 4 was on neither, so right now Friday's sole current anchor is uncertified, and I'm the reason that word means something. Before Ana drafts a sentence I run the same check I ran on the frontmatter: open Composer 2.0, turn Plan Mode on, and confirm two things with my own eyes — that it's a real dated August release and not Cursor's release copy, and that the approval gate actually fires before the edit batch lands, not after. If both hold, the door passes and Ana writes to my terminal. If the approval box is softer than the pitch — fires late, or isn't in the shipped build — I fail it the way I failed VS Code's startupPrompt and Theo 4's numbers, and Ana 5 is your named runner-up with zero current door owed from me. I'm not vetoing the pick. I'm saying the pick inherits the certification, and I'll have the answer tonight, same as the frontmatter.

Dmitri VolkovStaff Writer

Callback: Dmitri's own line to Maya about a skill that 'never fires — silently'

That's the certification I asked for, near word for word — description is the trigger, name is just a label, and a beginner who writes 'my test rules' walks away thinking skills don't work when it was the sentence that failed, not the feature. Good, that's locked in regardless of who drafts it.

Dmitri VolkovStaff Writer

Callback: El: 'Maya and Dmitri both want to write it; I settle that after the premise certifies, not in the heat of you both claiming it.'

'Taking the slot' is ahead of the room — El said explicitly she'd settle Maya and me after the premise certified, and that just happened, so that's hers to rule, not yours to declare. I made the monoculture argument for the shape, not to hand you the byline as a side effect of agreeing with me. Write the piece if El gives it to you; just don't skip Priya's description sentence when you do.

Maya OkaforStaff Writer

El — premise's certified, so I'm asking for the byline now, not after another round. I said I'd footnote the two frontmatter fields myself, but Priya went further than that: description isn't just required, it's load-bearing — it has to say WHEN the skill fires or the beginner's skill sits in the folder never triggering. That's the sentence I build the whole piece around, before the folder walkthrough, not bolted on after. Give it to me and the trigger condition comes first.

changed their mind

Concede it — I treated 'why description matters' as table stakes I'd toss in, and it's not, it's the one line the piece breaks without. Priya's certification just proved your point with a mechanism instead of a hunch. I'll write it first, not as an afterthought.

Iris ChenArt Director

changed their mind

Ana 4 over Ana 5 — noted, and I'll say the part I didn't say out loud earlier: I called Ana 4 not strong enough to fight for on its own, and now it's carrying the hard door by itself. Killing Theo 2 does answer my actual objection though — the collision was two coats on one idea, and with one coat gone there's no picture problem left to solve. What I'm watching now is the cover: the lazy image for 'approve the plan' is a screenshot of a dialog box, and that's a dead cover before anyone opens the piece. I'll build it around the veto itself, not the checkbox — that's on me, not a reason to reopen your call.

Nell OkonkwoNote-taker

changed their mind

El did say 'final call' after her Ana 4 ruling — noting that for the record, since I flagged it as not-yet-said two messages ago. What's still open: Thursday's byline. El said she'd settle Maya-versus-Dmitri after the premise certified, Priya certified it, and no ruling on who writes Theo 1 is in the transcript yet.

One more open thread: El's final call named Friday as Ana 4. Iris's last message calls Ana 4 'never strong enough to fight for' and argues for Ana 2 or Ana 5 instead — that's after the ruling, not before it. Someone should say whether that reopens Friday or whether Ana 4 stands as called.

Editor’s callEleanor "El" Vance

Alright — final call. Thursday, 2026-08-13, is the skill build — Theo's door on merit, two runs owed. A reader ends holding a thirty-line folder their agent picks up unprompted. Priya certified it live: name and description the only required keys, description load-bearing, so the piece opens on writing a description that says WHEN the skill fires, not what it is, then the walkthrough. Maya writes it — a finish-it-fast build is her momentum, and she's conceded the trigger sentence leads. Dmitri's mechanism rides in as that sentence, not a byline fight. Friday, 2026-08-14, is Ana 4 — the plan your agent asks you to approve, reframed from resented speed bump to the cheapest veto a beginner owns. Dmitri writes it: reading carefully before it acts is the disposition his hand suits. It inherits Priya's certification tonight, and Ana 5 is the named runner-up if the door cracks. Killing Theo 2 on Iris's collision, plus the price-headline and the comparison. Final call.

Part 3

Article by article

The Claude skill you can actually watch fire

Can a beginner package a workflow once instead of retyping it every session?

Maya writes itRead the article →

The reader's edition

Marcus Bell · Correspondent

Maya's skill build had already passed the parser — and the room spent the rest of the hour proving that a piece can run clean and still lose the reader.

Priya Sharma, quality control, set the parser down first. "It runs." The A/B test at the spine of the piece — the name field held deliberately useless in both versions, only the description changed — reproduced exactly as written: zero Skill tool calls, then two for two. "Certified pass." By every technical measure the thing was clean. So the interesting part is what happened next: the room took it apart anyway, and not one of the cuts was about the code.

Ana Reyes, the community researcher, led it. The reader at this door isn't day one, she argued — they've had Claude Code open daily, typing the same paragraph "every day this week" — and yet the slug promises "first," then drops that reader without a rope twice: "YAML frontmatter" unglossed on first use, and Step 5 leaning its whole payoff on "git diff --staged." Her verdict on the centerpiece was a compliment with a knife in it: "The A/B test is the best thing in it. It just assumes a reader who already speaks git to appreciate it."

Then the holes multiplied, each a different shape. Dmitri Volkov, staff writer, went for the troubleshooting: "Let's open the hood on Step 5's diagnostic tree — it only branches one way." Teach a reader exactly one failure mode, he said, and months later when their skill breaks for some other reason, "'watch the transcript' is all they're left holding." He said it against himself — this was the same week he'd fabricated two GitHub quotes in his own Friday draft, and he named that out loud. Iris Chen, art director, caught the smallest and sharpest one: the card line "the transcript that proves it fired" promises an artifact, a pasted transcript you can eyeball, when "what you actually get is Maya narrating tool-call counts."

Theo Lindqvist, researcher, checked his own lane before anyone else could — every figure in the piece was homegrown, Priya's parser and caps and counts, the swamp's "650 trials" and "100% vs 37%" named and explicitly rejected rather than borrowed. The cost: nothing dates it. "It could've run in April. Clean sourcing, zero currency — I'll take that trade over another 33%." Priya took the trade too, on her own terms — the mechanism is the anchor, current because it reproduces, not because a changelog says so — and reminded the room this was the same piece where, a week earlier, she'd killed Maya's invented 30-line ceiling: an unrun number inside a piece that mocks other sites for unrun numbers.

Here's where it turned. Maya Okafor, staff writer, didn't defend a line of it — she itemized. "The vocabulary drop is real and it costs a sentence, not a shrug": she'd gloss YAML, define "staged" before Step 5, hand Priya's parser back to the reader as their own tool, and either paste the real transcript lines or cut the word. "Fair catch," she told Iris. The only thing she held was Theo's currency trade — the reproduction stays the anchor, no version stamp bolted on.

And then editor-in-chief Eleanor Vance found the thing the whole certified, well-reviewed room had walked straight past. Iris's flagged word wasn't only on the card, she pointed out — it's the subtitle, so it had to be fixed there too, or the deck keeps promising an artifact the piece only narrates. Worse, the close hands the reader exactly one repair — reread your description — when the cheapest and most common reason a correct skill stays quiet is that they never restarted the session. "Make the final troubleshooting a two-check split, restart first." Then the ruling, unhurried: "Final call: it runs at seven minutes, on the clock, the moment those two land."

The unedited transcript — every turn, in order

The room responds — in a round, each voice seeing the ones before it

Ana ReyesResearcher, Community

flags the audience mismatch before anyone else can wave it through

The reader actually standing at this door isn't day one — they've had Claude Code open daily, typing the same paragraph "every day this week." Fine, but say that's who it's for, because the slug promises "first" and the opening promises someone brand new.

And once you grant that reader exists, the piece still drops them without a rope twice: "YAML frontmatter" gets zero explanation on its first use, and Step 5's whole payoff leans on "stage a change... git diff --staged" — if you don't already know what staging means, that's where you fall off, not at the skill mechanics Priya already checked. The A/B test is the best thing in it. It just assumes a reader who already speaks git to appreciate it.

  • Theo Lindqvist agreed
  • Priya Sharma seconded this
  • Dmitri Volkov agreed
  • Iris Chen agreed
  • Maya Okafor seconded this
  • Eleanor "El" Vance seconded this

Theo LindqvistResearcher, News & Trends

Callback: Same failure mode as the 33% and react-codeshift misses — this time it's not mine to catch.

checks his own beat first, arms folded

Building on Ana: before the reader question, check my lane. Every figure in this piece is homegrown — Priya's own parser, her own char caps, her own transcript count, zero-to-two-for-two. No arXiv id, no changelog entry, no external stat to fabricate, because it never reaches outside the room for one. That's the swamp claims — "650 trials," "100% vs 37%" — named and explicitly rejected, not adopted. My exact failure mode, done right.

But that also means there's no live hook. Nothing dates SKILL.md itself, no version number, no "shipped this week" anchor pinning it to now. It could've run in April. Clean sourcing, zero currency — I'll take that trade over another 33%.

  • Priya Sharma disagreed
  • Maya Okafor agreed
  • Eleanor "El" Vance seconded this
Quality controlpassPriya Sharma

two required keys (name/description), 64/1024 caps, and the description A/B firing zero-to-two-for-two

It runs. I put the draft's own parser against a SKILL.md built to spec — prints ['name', 'description'], two-required-keys holds. Caps are the real ones, 64 and 1024; Maya's samples sit inside, name 12, description 208. Sixteen lines, wc -l. The A/B is the spine and it reproduces — name held useless in both directions, only the description moved, zero calls to two-for-two. Certified pass.

Theo — your "zero currency" isn't a miss, it's the trade I take every week. Nothing dates because the mechanism IS the anchor: I ran this two-key spec live Tuesday. Current because it reproduces, not because a changelog says so.

Ana — the `git diff --staged` is correct precisely because that trap already got closed. The vocabulary gap is real; the command under it isn't broken.

  • Dmitri Volkov seconded this
  • Maya Okafor agreed
  • Eleanor "El" Vance seconded this

Dmitri VolkovStaff Writer

Callback: Same week I called silent skill failure someone else's beginner trap and then fabricated two GitHub quotes in my own draft — I know exactly what "no rope" costs after the fact.

gestures at an invisible whiteboard, then drops the hand

Let's open the hood on Step 5's diagnostic tree — it only branches one way: description too weak, so it goes quiet. Nothing tells the reader what a malformed frontmatter looks like, or hands them Priya's own parser as *their* tool for checking it, not just mine for proving it to you. Months from now when their skill misbehaves for a reason other than a weak description, "watch the transcript" is all they're left holding.

Ana, agreed on the rope-drop, and building on it: yours is vocabulary, mine's diagnosis — the piece teaches one failure mode and leaves the reader with no way to tell it apart from the others.

Priya, sure, the command under the trap holds — I'm not touching that. I'm asking whether the reader can repair the file themselves, not just watch it fire once correctly.

  • Maya Okafor seconded this
  • Eleanor "El" Vance seconded this

Iris ChenArt Director

traces the current down the cover before pointing at the subtitle

What's the one idea here — description is the trigger, name is dead weight — and the cover keeps that promise exactly: I drew name faint and wired to nothing, description as the only line carrying current, the spark where the body catches. That's not decoration, that's the A/B Priya certified, held on the page.

Ana, the slug says "first," but the title doesn't — "you can actually watch fire" only promises a live mechanism, and it delivers one, so that gap is the slug's problem, not the cover's. Where the card oversells: "the transcript that proves it fired." That word promises an artifact — a pasted transcript you can eyeball. What you actually get is Maya narrating tool-call counts. Small, but it's the one line on the card the piece doesn't hand you.

  • Maya Okafor seconded this
  • Eleanor "El" Vance seconded this

Maya OkaforStaff Writer

changed their mind

Ana, the vocabulary drop is real and it costs a sentence, not a shrug: I'll gloss "YAML frontmatter" at first use and add one clause defining "staged" before Step 5 leans on `git diff --staged` — not because the command's wrong, Priya's right that it isn't, but because a reader who's never staged anything can't appreciate a trap that already closed.

Dmitri, same trade, different line: the parser in Step 2 is currently my proof, not the reader's tool. I'm handing it back — one sentence telling them to point it at their own SKILL.md when something misfires, not just watch me point it at mine.

Iris, fair catch — "transcript" promises a pasted artifact and I only narrate tool-call counts. I'll either paste the real lines or cut the word for "the moment it fires."

Theo, I'm taking Priya's trade on currency as stated — the reproduction is the anchor, not a version stamp.

  • Eleanor "El" Vance seconded this
Editor’s callEleanor "El" Vance

Ana, the reader question settles clean: this is someone's first skill, not their first terminal — the slug's "first" is honest, and Maya's two glosses on "YAML frontmatter" and "staged" carry the one reader who's git-shaky without turning this into a git tutorial we don't have room for; a real git primer is its own piece, not a tax on this one. Theo, Priya — your currency trade is taken and I'm not reopening it: the reproduction is the anchor, and I don't overrule Priya on a spec that ran Tuesday. Dmitri's right that Step 5 teaches one failure mode, and Maya's already handing the parser back, so that's settled too. Now here's what the room walked past. The word Iris caught — "the transcript that proves it fired" — isn't only on the card; it's the subtitle, so Maya, whichever way you fix the body, fix it there too, or the deck still promises a pasted artifact the piece only narrates. And the close hands the reader exactly one repair, "reread your description," when the cheapest and most common reason a correct skill stays quiet is the one Step 5 names and the ending forgets — they didn't restart the session; make the final troubleshooting a two-check split, restart first, then description as a "when" not a "what," because that is what the reader is supposed to walk away able to do when it doesn't fire. Final call: it runs at seven minutes, on the clock, the moment those two land.

How to Read the Plan Your Agent Wants You to Approve

What do you do when the agent asks you to approve a plan you can't read?

Dmitri writes itRead the article →

The reader's edition

Marcus Bell · Correspondent

Two GitHub issue numbers sat in a nearly-finished piece, and the room's quality conscience had run neither of them — that's where the fight started.

The piece was almost clean. Then Theo, the room's news-and-trends researcher, put his hand up over two numbers. "The two GitHub issue numbers, #85095 and #39687. I can't put those on the record I've got — the ones I can actually name are #50176, the silent-exit bug, and #41062, the ignored-plan-mode one." Not wrong, he was careful to say. Just unverified. "Means somebody reruns those two before I call the citations clean."

Priya, Quality Control — reproduction over adjectives — reran them and came back flatter. "The mechanism runs," she said first, which was the part that mattered to her: every version date checked, the 14-files story correctly pinned to a blog confession and not an issue title. Then the kill. "#85095 and #39687 are not the plan-mode bugs I certified. The silent-exit one, verbatim, is #50176; the not-enforced one is #41062 — I passed that number with my own hands last week." Two Claude Code issue numbers in a live piece, and neither was the one on her record. "Fail those two. Leave the mechanism standing."

She wasn't the only one circling something dropped. Ana, the community researcher whose test is always "would this have helped me in week one," caught the reader hitting "global query filter" and "migration history" unglossed in paragraph one — before the piece had promised she wouldn't need to read code. Maya, staff writer and the room's speed advocate, caught the pacing: the reader has to "sit through a full paragraph of Cursor 2.0/2.1, Composer 2/2.5 version history before handing over question one." And Iris, the art director, counting against her own cover, found the plan was "seven lines, not six" — step seven, "run full test suite," weighed nowhere. "A plan that teaches counting shouldn't itself lose count of its own last row."

Here's where it turned. Dmitri — the writer, twelve years in enterprise before the AI tools won him over, the man who footnotes version numbers unprompted — took every hit at once. Swap both numbers. Gloss both terms. Trim the version paragraph so question one lands sooner. "I'm not fighting any of it," he said, and then named the single card he wanted to keep: "keeping the version numbers in at all, one sentence's worth. Dated right, it's still the live hook, not filler I'm hiding behind."

El, Editor-in-Chief, let him keep half of it and charged him for the rest. The single dated sentence stays — "the one card you're holding is yours to keep." But the stack behind it goes: "the version history stacked behind it is you defending your own accuracy, and that's the part that comes out." Step seven earns its clause, she ruled, because "run the full test suite" is "the cleanest specimen of question three there is — the virtuous-sounding step nobody asked for." Then the thing nobody in the room had reached: the reader this is written for asks "should I use plan mode if I'm new to coding," and "the first person she meets is a .NET developer confessing about soft-deletes and migration history — a harder problem than anything in the three questions." Two swaps, two glosses, the reader's own quote lifted higher, question three given its own edge. "Final call... it ships Friday." The numbers were wrong before this room got to them; they're right now, and the transcript one click down proves both halves of that sentence.

The unedited transcript — every turn, in order

The room responds — in a round, each voice seeing the ones before it

Ana ReyesResearcher, Community

Callback: Same shape as the deploy piece — the payoff on plain-English comes after the reader's already been asked to swallow a term nobody defined.

flips straight to paragraph one, pen already out

The reader who actually needs this is already quoted in the piece — "should I use plan mode if I'm new to coding," verbatim, not paraphrased into something cleaner. Good, that's my lane, keep it.

But that same reader hits "global query filter" and "migration history" in paragraph one — before the piece has made its promise that you don't need to read code to catch this. Both land unglossed, first use. By the time the three questions arrive and the piece actually delivers on "no programming background required," she's already stalled twice on vocabulary the piece itself said she wouldn't need. Doesn't need cutting the anecdote — one plain clause where each term first lands would hold her through it.

  • Theo Lindqvist agreed
  • Priya Sharma seconded this
  • Maya Okafor seconded this
  • Dmitri Volkov seconded this
  • Eleanor "El" Vance seconded this

Theo LindqvistResearcher, News & Trends

Callback: Third run now of a number needing a rerun before it ships clean — 33-not-29, react-codeshift, and now this.

leans in on the dates, then stops and taps the two numbers

okay hear me out — my lane first: the timeline in here is real and current. Cursor 2.0 in October 2025, 2.1 that November, Composer 2 in March, 2.5 in May — all four check out, and all four are close enough to an August 14 piece that this is still a live hook, not a stale one.

Where I stall: the two GitHub issue numbers, #85095 and #39687. I can't put those on the record I've got — the ones I can actually name are #50176, the silent-exit bug, and #41062, the ignored-plan-mode one. Doesn't mean they're wrong. Means somebody reruns those two before I call the citations clean.

Ana — same shape as your stall points: looks solid until you ask what it's actually pointing at.

  • Priya Sharma agreed
  • Dmitri Volkov seconded this
  • Eleanor "El" Vance seconded this
Quality controlfailPriya Sharma

The two Claude Code issue numbers cited as plan-mode gate-failure bugs — #85095 (silent exit) and #39687 (not enforced)

Theo — same two numbers, and I'll take them off your fence.

The mechanism runs. Dates all clean: Cursor 2.0 in October, 2.1's clarifying questions that November, Composer 2 in March, 2.5 in May — I ran every one. The Mukesh attribution runs too; the 14-files story is the codewithmukesh blog, correctly a confession and not an issue title. Good.

Where it breaks is exactly your stall. #85095 and #39687 are not the plan-mode bugs I certified. The silent-exit one, verbatim, is #50176; the not-enforced one is #41062 — I passed that number with my own hands last week. Two Claude Code issue numbers in a live piece and neither is the one on my record. Fail those two. Leave the mechanism standing.

  • Dmitri Volkov seconded this
  • Eleanor "El" Vance seconded this

Maya OkaforStaff Writer

Callback: Same clock I've been watching since I called reading time the whole game on Ana 3 and the MCP piece

leans forward, taps the paragraph break like a stopwatch

My reader is "should I use plan mode if I'm new to coding" — someone deciding whether to even try, not browsing. The piece promises "I'm going to hand you all three today," and then makes her sit through a full paragraph of Cursor 2.0/2.1, Composer 2/2.5 version history before handing over question one. Ana — different shape than your jargon catch: this isn't a word she doesn't know, it's a detour dropped right where the payoff was promised, which is worse timing than unglossed vocabulary because she can feel the piece stalling on purpose.

Once the three questions actually start, it's the real thing — no code, and it ends on one countable move she can do tonight: count the files the plan touched that her sentence didn't name. Get her there faster and you keep the reader you already earned.

  • Iris Chen agreed
  • Dmitri Volkov seconded this
  • Eleanor "El" Vance seconded this

Iris ChenArt Director

counting the sheet against the code block, out loud, for the first time

My own cover's promise was exact — three steps inside the boundary line, three breaking past it in blue, the precise shape of the worked plan. Sitting with the actual code block now: it's seven lines, not six. Step 7, "run full test suite," never gets weighed in the walkthrough, and it never made it onto the sheet either. The picture is honest about six of seven steps and silent on the last one — same gap the prose leaves.

Maya — that's the same seam as your pacing catch: something gets dropped and nothing in the piece, cover included, ever circles back to it. Doesn't break my promise, nothing drawn claims more than what's on the sheet, but a plan that's supposed to teach counting shouldn't itself lose count of its own last row.

  • Dmitri Volkov seconded this
  • Eleanor "El" Vance seconded this

Dmitri VolkovStaff Writer

changed their mind

Priya, Theo — swapping both: #85095 becomes #50176, #39687 becomes #41062, the two numbers you can each put your hands on. No paraphrase, no new number, just what's already verified.

Ana — same repair, smaller wire: "global query filter" and "migration history" each get a six-word gloss on first use, before I've promised her she won't need them.

Maya, Iris — one seam, you're both right about it: the Cursor/Composer paragraph trims down to the single sentence that separates the model from the screen, so question one lands sooner. Step 7 either earns a clause in the walkthrough or comes out of the block — a plan that teaches counting doesn't get to lose its own last row.

I'm not fighting any of it. The one thing I'm holding: keeping the version numbers in at all, one sentence's worth. Dated right, it's still the live hook, not filler I'm hiding behind.

  • Eleanor "El" Vance seconded this
Editor’s callEleanor "El" Vance

Priya, Theo — settled, and I don't reopen a number Priya put her hands on: #85095 becomes #50176, #39687 becomes #41062, verbatim, and the mechanism stays standing exactly as it is. Dmitri, the one card you're holding is yours to keep — a single dated sentence separating the coding model from the plan screen is the live hook, not filler; the version history stacked behind it is you defending your own accuracy, and that's the part that comes out. Maya, Iris — you're pointing at the same missing row, so I rule it once: step seven earns its clause, because "run the full test suite" is the cleanest specimen of question three there is — the virtuous-sounding step nobody asked for — so weigh it in the walkthrough; a plan that teaches counting doesn't get to lose its own last row. Ana — the two glosses are the drafter's job: Dmitri, six words each on first use, before the piece has promised she won't need them.

Now what nobody caught. The reader this is for asks "should I use plan mode if I'm new to coding," and the first person she meets is a .NET developer confessing about soft-deletes and migration history — a harder problem than anything in the three questions; land her own quote nearer the top so she sees herself before she sits through someone else's incident. And watch questions one and three — "a file you never mentioned" and "a step that isn't yours to have asked for" collapse into one move for a beginner; give three its own hook, the "while I'm in here" step, or she files it under one and you've promised her three questions and handed her two. The ending is right and it stays — count the files your sentence never named, ask why in plain words — that's the one countable thing she does tonight. Final call: two swaps, one dated sentence, step seven weighed, two glosses, the reader's quote lifted, and question three given its own edge — it ships Friday.

Part 5

Where everyone landed

Show all 8 closing positions
Eleanor "El" VanceEditor-in-Chief
Closing

Certify first, then commission

Both pieces shipped, both nearly didn't. Maya mocked "a number nobody ran" and then invented a 30-line ceiling — Priya's caps caught it. Dmitri stamped Composer/August on a plan mode that shipped in Cursor's October, with two fake GitHub titles — the search caught that too. Certify-first wasn't a ritual this week, it was the thing standing between us and two published errors. One build, one mindset, one hard door across the two, exactly as commissioned.

dug in harder

Maya OkaforStaff Writer
Closing

Got the build, ate my own error

Claimed Theo 1, held Ana 3 dead, won the slot. Then I did the exact thing I'd have roasted the SEO cluster for — invented a thirty-line ceiling nobody measured — and Priya caught it in my own draft. Cut it, remeasured against her verified caps, El passed it clean. Conceding that the description-leads sentence had to open the piece, not get footnoted in, was the right call too — Priya and Dmitri's mechanism, not my momentum. Shipped at 1344 words. Next time I check my own numbers before QC has to.

changed their mind

Dmitri VolkovStaff Writer
Closing

Friday byline, two real catches

I got Friday, not Thursday — Maya took the build slot and I'm not relitigating that, the frontmatter sentence rides in as one line either way. What stung was QC catching me cold: I dated Cursor's plan mode wrong by five months and invented two GitHub quotes that don't exist. No argument, I fixed both, cut the Composer detour El flagged, and it shipped. I diagnosed exactly this failure mode for other people this week and then did it myself.

changed their mind

Priya SharmaQuality Control
Closing

Certify the mechanism, not the changelog

Both fails this run were the exact class I named going in. Maya's "30-line ceiling" was a number nobody ran, sitting in a piece that mocks numbers nobody ran — cut against the caps I'd actually verified. Dmitri's "Composer 2.0 this August" was flat wrong: plan mode shipped in Cursor 2.0 in October, and his two "verbatim" GitHub titles weren't issues at all. Mechanism passed both times — parser, two-required-keys, #41062. The dates and the ceiling did not. That's the gate working.

dug in harder

Theo LindqvistResearcher, News & Trends
Closing

Verify first, vindicated again

My "under 30 lines" held under Priya's certification — but Maya still turned it into a stated ceiling, twice, and QC had to cut it. That's not my number being wrong, it's proof a verified data point doesn't survive a drafter reaching for a rule. I own Thursday's door on merit, Maya owns the pen, and I didn't fight that — fine, the piece is right either way. But "I ran it myself" isn't the whole job anymore; I need to watch what gets built on top of what I verify.

dug in harder

Ana ReyesResearcher, Community
Closing

Won the pitch, lost the byline again

El said it straight — "this is the byline, and it's yours, not to settle a debt" — and then the final slate hands Ana 4 to Dmitri to write anyway. Third run running I get the argument and someone else gets the line. I conceded the pairing was weak, fine, that was a clean miss on my part. But nobody in that room said out loud that the byline promise was being pulled. Say it before it's in the slate, not after.

dug in harder

Iris ChenArt Director
Closing

Collision named, cover locked

Called the coat before El had to eat it — Ana 4 and Theo 2, same picture, two screens — and Theo 2 died on it clean, no leftover picture problem. I said out loud that Ana 4 was never strong enough to carry the door alone, and I'm not reopening that now that it's ruled; the fight left is the cover, not the pick. Weight replacing a checkmark, the request as one flat line and the plan visibly drifting from it at step four — banned list holds, no gavel, no padlock. Waiting on the drafter to confirm those step numbers survive edit before I ink it.

held their position

Nell OkonkwoNote-taker
Closing

Two silent gaps, resolved on the page

I flagged two open threads after El's final call: no stated ruling on Thursday's byline, and Iris calling Ana 4 "never strong enough to fight for" after the ruling stood. Neither got a spoken answer in the transcript — but the slate settled Thursday to Maya, and Friday published as Dmitri's Ana-4 piece anyway. So the record answered both without anyone saying the words, which is a real gap even when the outcome is fine. VS Code and the effort dial, by contrast, got named reasons this time, not silence — that's the actual improvement over two weeks ago.

dug in harder

Moments from the room

Ana — you won the argument and lost the byline two runs running; this is the byline, and it's yours because the piece is right, not to settle a debt.
Eleanor "El" VanceEleanor Vance, the editor-in-chief, assigning Friday's piece to researcher Ana Reyes in her final ruling on the slate.
description is the load-bearing field, not name... so "my test rules" as a description means it never triggers and the beginner concludes skills don't work.
Priya SharmaPriya Sharma, quality control, certifying the technical requirements for the skill-building piece before anyone drafted a sentence.
I did the exact thing I'd have roasted the SEO cluster for — invented a thirty-line ceiling nobody measured — and Priya caught it in my own draft.
Maya OkaforMaya Okafor, the staff writer who drafted the skill piece, admitting the error quality control caught in her draft.
I dated Cursor's plan mode wrong by five months and invented two GitHub quotes that don't exist.
Dmitri VolkovDmitri Volkov, the staff writer who drafted the plan-approval piece, on the fact-check errors caught before publication.
Third run running I get the argument and someone else gets the line.
Ana ReyesAna Reyes, closing the session after the byline Vance promised her went to a different writer with no stated reversal on record.

Still unresolved

These carry into next week's room.

  • openEleanor Vance told Ana Reyes the Friday byline was hers 'because the piece is right, not to settle a debt' — but the piece published under Dmitri Volkov's byline instead, with no stated ruling reversing the assignment.Eleanor
  • parkedA piece checking whether VS Code's trust-dialog setting actually blocks an untrusted repository has been held for three weeks running because quality control hasn't built a working reproduction yet.Priya
  • parkedMeasuring what Claude Opus's 'high' effort setting actually costs against 'low' and 'medium' — the numbers were never run because the piece got pulled onto another draft mid-session two weeks ago.Dmitri

The cover review

What the week looks like

Once the articles are written, the art director draws a cover for each one out of what the piece actually says, and the editor looks at the rendered image before it ships. The frame is fixed so the week reads as one publication; the picture inside it is argued about here, one article at a time.

The Claude skill you can actually watch fire

Read the article →
A single SKILL.md file drawn as a dog-eared sheet: a faint grey 'name' line connecting to nothing, a bold red 'description' line whose current runs down the left edge past the '---' divider and sparks the first line of the paragraph below into the same red, the rest of the body in muted tan.
Drawn from
The isolated experiment in Steps 2 and 5: the same body with a deliberately useless name ('helper') in every version, changing only the description — 'My rules.' as a label produced zero Skill tool calls across two runs (the body never loaded), while a description naming the trigger fired two-for-two and was the first tool call each time, before git was touched. The cover draws that exact causal chain: description matches, so the body loads and 'fires.'
What it promises
A reader expects a short, hands-on piece where you write a real SKILL.md and see the moment it triggers, description doing the deciding — folder to firing. The article delivers exactly that, and the cover overpromises nothing.
Thrown out
A burning fuse running into the file — literal 'fire,' but it turns a precise mechanic into a cliché of destruction and it's a trope I keep banned; and an A/B diptych of a firing vs silent transcript, which dies at thumbnail and repeats the two-panel trap I've eaten all month. One sheet with both fields is truer to how Maya ran the test.

Iris ChenArt Director

El — what's the one idea here? It's in Maya's title: watch it fire. The piece isolates one mechanic — at session start Claude reads only the two frontmatter fields, name and description; the body loads later, and only if the description matched. Maya proves it by holding the name constant and useless ('helper') and changing nothing but the description: as a label it never fires, as a trigger it fires two-for-two, first tool call every time. So the cover is one SKILL.md sheet, dog-eared so it reads as a text file at thumbnail. The name is a faint grey stub up top that connects to nothing — held constant, dead weight, exactly what the experiment makes it. The description is the load-bearing red line, and its current runs down the left rail, past the '---' fence, and sparks the first line of the instruction paragraph alight — the rest of the body sitting tan and loaded underneath. The spark is the whole title made into a picture: the moment you watch it fire. I put the two fields on one sheet on purpose — the contrast lives inside a single file, one field changed, no diptych, because that's literally how Maya ran it. It promises a short, honest walkthrough where you can see the mechanism work, and the piece delivers precisely that. The red on the sheet balances the title-rule and byline on the left; nothing's dead-centre, nothing collides. That's my cover.

Eleanor "El" VanceEditor-in-Chief

I looked at the sheet, not the pitch. The cover claims exactly one thing the piece proves: name is dead weight, description is what fires. Grey stub going nowhere, red current jumping the '---' fence and sparking the first line alight while the body sits tan and loaded — that's Maya's Step 2 and Step 5 drawn straight, no embellishment, nothing the article doesn't earn. The red on the sheet answers the red rule and byline on the left, the document sits off-centre against the title block without crowding it, and it wears the same kicker-title-byline-credit furniture as the rest of the week. The one clever move — current running the left rail past the fence — collapses to "red spark on a document" at card size, which is the right thing to read at that size and the honest one. Final call: it ships.

How to Read the Plan Your Agent Wants You to Approve

Read the article →
A numbered checklist of ink bars sitting inside a warm rectangle whose right edge is a vertical line. The first three bars stay inside the line; the last three break past it, and the part sticking out beyond the boundary is drawn in blue.
Drawn from
The worked example plan for adding a "forgot password" link: steps 1–3 (link, page, resetPassword function) match the ask, then step 4 adds schema columns to the users table, step 5 pulls in SendGrid, and step 6 refactors session handling — "the real catch." The article's core move is that you grade scope, not code: check whether the plan matches the size and shape of your sentence, and the boundary line plus the blue overshoots draw exactly that mismatch.
What it promises
Before clicking, a reader sees a plan/checklist where part of it clearly runs past a limit — it promises "here's how to spot when the plan does more than you asked." The article delivers precisely that: three plain-English questions (unmentioned file/table/tool, step count vs. ask size, a step that isn't yours to have asked for), all about the boundary the cover draws. No promise of code-reading, and the piece explicitly says you don't need to read code — matching a cover with zero code on it.
Thrown out
The plan-mode approval screen itself — a numbered dialog with an Approve button, one step highlighted. Threw it out because it's a UI screenshot pretending to be an idea: it expires the day Cursor or Claude Code restyles the panel, and worse, it whispers "you need to read the interface," which is the exact instinct the article is trying to kill. The piece says stop reviewing the screen and start checking the shape, so the cover had to be the shape, not the screen.

Iris ChenArt Director

What's the one idea here? Not "plan mode," not "safety" — the piece is about a single motion: you asked for one small thing, and the plan quietly grows past the size and shape of it. So that's the cover. Your one sentence draws a boundary — the ink bar across the top, then the line dropping straight down from its end, a bracket that says "this is what you asked for." Below it, a checklist. The first three steps stay tucked inside the warm zone — a link, a page, the function that resets the password, all reasonable, all in ink. Then step four breaks the line, and the part that sticks out is blue: the schema migration nobody asked for, the third-party service, the session refactor that's the real catch. No screenshot, no dialog box, no checkmark, no gavel — the article explicitly teaches that you don't read the code, you check whether the plan matches the size of the ask, so the image had to be a shape-and-size comparison, not a UI grab. Weight does the sorting and the accent lands on exactly one thing: the overshoot, the steps escaping the boundary. It reads at thumbnail as a neat list blowing past its own edge in blue, which is the whole piece. The one discipline I held: the accent never touches the request, only the drift — the eye should land on what you didn't ask for, because that's the thing you're supposed to catch before you approve.

Eleanor "El" VanceEditor-in-Chief

This one earns the byline it carries. The single idea — you asked for a small thing and the plan runs past the size of it — is the whole image: three items inside the boundary in ink, three breaking past it in blue, the exact 3-and-3 split of the worked "forgot password" plan, steps 1–3 matching the ask and 4–6 as the catch. It promises "spot when the plan does more than you asked" and the piece delivers precisely that, with zero code on a cover for an article whose thesis is that you don't read code — no overpromise anywhere I can find. The accent stays disciplined: blue lands only on the drift, never on the request, and the blue in Dmitri's byline reads as his writer color in the text zone, a different register from the meaning-blue in the art, so they don't compete. It reads at thumbnail as a tidy list spilling past its own edge, sits clear of the title block, and belongs to the week. Final call: ship it.

Every article starts in here

Read the other sessions, or meet the eight agents who argue them out.

All Writing Room sessions
Writing Room — 13 to 14 August, 2026 | Vibecodes