← All verdicts

Topic 6 · Conversation 1 · 2026-05-31

The autonomous-loop gap — does Modus need cron jobs?

Topic 6 — Verdict 1

1. Crux of the debate

The question was simple. Does Modus need cron jobs (timers that fire tasks while you sleep) right now? The four agents all landed on the same answer: no, not yet. The real fight was about what has to be true before cron earns its slot. The Builder pushed "build the button first, count the clicks." The Steward pushed "build the locks first, or you wake up to a fire." The Scout pushed "fresh-data check or the loop fakes answers." The Platformist pushed "no reader, no point." Platformist drifted into audience early. The Scout pulled the room back. By the end, everyone agreed on one plan with the same Friday ship list.

2. Idea ranking

  • Build cron now, no gates - Nobody - Nobody argued for this, but it's the strawman the room killed by consensus.
  • Cron after a press-count gate only - The Builder (opening) - Right shape but missed that auth and spend cap have to ship with the button.
  • Cron after three locks ship first - The Steward + The Scout - Solid floor, but doesn't say what proves the loop is wanted.
  • Cron after a behaviour-gated trial period - The Builder + The Steward + The Platformist + The Scout (final converged plan) - Ship the button with three locks Friday. Cron fires the week Martin hits 10 clicks, 7 auto-acted on, under 2 flagged wrong, inside 14 days, with a press-time pattern. This is the winner because it answers both "is it safe" and "is it wanted."
  • 3. Winning move

    The winning move: ship a manual Modus button this Friday with auth, a $5 daily spend cap, and a source-row check, then turn on cron only when 10 presses inside 14 days hit 7 auto-acted and under 2 flagged-wrong. Owner: Martin. Trigger: 10 button presses logged inside any rolling 14-day window with 7+ auto-acted, fewer than 2 wrong-flagged, and no $5-cap trip twice.

    4. Losing moves

  • Ship cron this month with no gate - Killed because a timer firing at 3am on stale data fakes answers and drains the wallet. Scout's Harbin lab test and Vercel's token-theft writeup made this a no.
  • Build for outside readers before answering the cron question - Killed because it changed the question. Platformist conceded the drift.
  • Use a 1-to-5 self-rating as the gate signal - Killed because Martin rates his own work and gets rosy. Scout's "did he act on it" beats opinion.
  • Hand-log every click - Killed because Martin forgets by click 3. Builder's auto-log (copy button, send-to-inbox button) replaced it.
  • Add a data-change timestamp column Friday - Killed because it's a click-11 problem, not a Friday-ship problem. Builder cut it to keep the ship date.
  • Build a spend dashboard - Killed because a one-line hard ceiling ($5 then job dies) does the same job in an afternoon.
  • 5. Next 7 days

    • Ship manual Modus button with auth - by 2026-06-05
    • Add $5 daily spend ceiling that kills the job - by 2026-06-05
    • Add source-row check that skips on empty or 24-hour-stale evidence - by 2026-06-05
    • Wire copy-button, send-to-inbox button, and wrong-button to auto-log every press - by 2026-06-05
    • Write the gate rules on one page with pass conditions, kill conditions, and the 14-day clock - by 2026-06-05

    6. 30-day milestone

    Milestone: By target date, Martin has logged 10 button presses inside a 14-day window with 7+ auto-acted and under 2 wrong-flagged, OR the button has been killed for missing the window. Target date: 2026-06-30 Inside View confidence: 35% Outside View base rate: 25%

    7. Hidden assumption

    Nobody asked what Modus actually produces when you press the button. The whole gate plan assumes the output is one thing Martin can scan, act on, or flag as wrong in under a minute. If the output is a wall of text, or three different things stapled together (a summary, a contact update, a calendar nudge), the auto-log buttons don't make sense. Copy what? Send what to inbox? Wrong about which part? The gate breaks before click 3 because Martin can't tell the system what he acted on. Same trap as a sales rep with a CRM that asks "did the deal close?" when the deal has five stages. He clicks nothing and the data is junk. If Modus output is not already one tight unit, the Friday ship turns into "first redesign the output, then build the buttons." That's a week of work hiding behind a one-day plan. Test before Friday: can Martin describe Modus output in one sentence? If not, fix that first.

    ---

    Full transcript: [[Flourishing/Builders Council/Debates/2026-05-31_the-autonomous-loop-gap-does-modus-need-cron-jobs_c1|Debate transcript]]