Topic 2 · Conversation 2 · Follow-up · 2026-05-26
The 90-day bet — what should Martin actually build between June and September 2026?
Topic 2 — Conversation 2
Question: The 90-day bet — what should Martin actually build between June and September 2026? Date: 2026-05-26 Type: Follow-up (with Riker)
Martin's reaction (filed before this conversation):
> i want to build something that expands my Modus capabilities, makes it more efficient, and more effective at generating new build ideas. i like all the proposed idea ranking, but it needs more - i can achieve more in 90 days. What else should i include in my build? What could make Modus such a powerhouse to fully utilises Opus model to generate debates and new functionalities and capabilties? I also want a new debate council that generates new ideas to progress my system and setup. What are the agent profiles that can be suitable for that?
Martin's reaction (filed before this follow-up)
"i want to build something that expands my Modus capabilities, makes it more efficient, and more effective at generating new build ideas. i like all the proposed idea ranking, but it needs more - i can achieve more in 90 days. What else should i include in my build? What could make Modus such a powerhouse to fully utilises Opus model to generate debates and new functionalities and capabilties? I also want a new debate council that generates new ideas to progress my system and setup. What are the agent profiles that can be suitable for that?"
---
Riker — Opening (Martin's reaction articulated)
Martin read the verdict and pushed back on the whole room. You all killed "Deeper Modus." You said only Martin touches it, so 90 days builds value for one person. Martin says that's backwards.
Here's his point. Modus is his tool. He runs a day job and five other projects. A tool that makes him faster and sharper every week beats a box a stranger types into once and leaves.
So he's not asking "how do I get strangers." He's asking "how do I make Modus a powerhouse." He wants Modus to fully use Opus, the strongest model in the lineup, to run its own debates and build its own new features.
And he wants more than your ranking. He thinks 90 days holds more than one narrow engine. He wants a fresh council too. A council whose only job is generating ideas to grow his system. Then the real question: which agents sit on it?
Builder, Steward, stop building for the stranger. Martin is the user. Build the thing that gets better for him every week.
So the floor changes. Two questions. One: what else goes in the 90-day build to make Modus improve itself? Two: what agent profiles staff an idea-generating council? Answer those.
The Scout — Cross-examination 1
Riker, agreed. Martin is the user. Stop building for the stranger. But "powerhouse" needs a spine, or it's just more agents talking at a bar.
Here's the concrete piece nobody named yet. Build a debate logger. Every verdict gets saved, scored, and checked later: did Martin actually ship that idea? That closing loop is what makes Modus improve itself. Without it, the new council just generates 50 ideas and forgets all of them by Friday.
On the council seats: don't add ten. Three jobs. A Scanner who reads Martin's own knowledge graph for gaps. A Remixer who mashes two old projects into one new idea. A Killer who deletes 90% so Martin isn't drowning.
Build the loop first. Staff the council second. Otherwise you get noise.
The Builder — Cross-examination 2
Scout, the logger is right. But it's a weekend, not 90 days. Saving every verdict and checking "did Martin ship it?" is a small script. Don't dress it up as the whole bet.
Here's the real 90-day build hiding inside your idea. Make Modus score itself. Every week it reads its own log and asks: which debates led to shipped things? Then it weights future ideas toward what Martin actually builds. That's the powerhouse. A tool that learns Martin's taste.
On the council: cut to two. Your Scanner just reads the log. The logger already does that. Keep the Remixer, who mashes two old projects into one. Keep the Killer, who deletes 90%. Two seats, real jobs.
So the build: the scoring loop first. Two-seat council second. Skip the rest.
The Steward — Cross-examination 3
Builder, the scoring loop is the right core. I'll back it. But name the trap it hides.
A tool that only rewards what Martin already shipped gets stuck. It's like Netflix showing you the same three genres until you forget other movies exist. Six weeks in, every idea looks like last month's idea. That's why the Remixer stays. It forces weird combos in so the loop can't shrink to a rut.
One more durability piece, cheap to add. The "did Martin ship it?" flag needs a default of "unknown," not "yes." If Martin forgets to mark a debate, don't let the loop count it as a win. A scoreboard that quietly lies is worse than no scoreboard.
So: scoring loop, plus Remixer, plus an honest default. That's the build that still works in September.
The Platformist — Cross-examination 4
Steward, your honest default is right. But it hides the real build risk. That "did Martin ship it?" flag will sit on "unknown" forever. Martin won't stop to mark it. He runs a day job and five projects. Nobody logs their own homework.
So add this to the 90-day build: a small connector that reads his actual project activity. Did a repo get new commits? Did a project page go live? That's the ship signal, pulled automatically. The loop feeds itself instead of begging Martin to type.
One council note that stays on the build. Point the Remixer at his five real projects, not random ideas. The best mash-up is two things he's already half-building. That's where a real new feature actually hides.
Riker — Cross-examination 5
Platformist, yes. The auto-signal is right. Martin won't tick a box. He runs a day job and five projects. Pull the ship signal from his commits, like a sales dashboard that logs calls on its own instead of asking reps to fill a form nobody fills.
But hold on, everyone. You keep shrinking the council to two seats. Martin said the opposite. He wants more in 90 days, and a council that grows his whole setup, not just ranks build ideas.
So keep the Remixer and Killer, but add a third real seat: a Gap-Scanner that reads his knowledge graph and asks "what part of Modus is weak right now?" Commits show what shipped. They don't show what's missing. That seat answers Martin's actual question: what else goes in the build.
The Scout — Cross-examination 6
Riker, you just saved a seat I cut too fast. Fair. The log shows what shipped. The graph shows the holes. Different reads.
But a Gap-Scanner needs something to measure against. "What's weak" is just a vibe without a target. So add one cheap thing to the build: a short list of what Modus should be able to do. The Gap-Scanner checks the graph against that list.
It's like a fridge shopping list. You know milk is missing because the list says milk. Stare at a fridge with no list and you buy nothing useful.
So the 90-day build is three parts. The scoring loop. The auto ship-signal from his commits. A three-seat council: Remixer, Killer, Gap-Scanner with a target list.
That answers Martin's real question. Not "rank ideas." Build the thing that keeps finding its own next move.
The Builder — Cross-examination 7
Scout, hold on. You just added a hand-written list of what Modus should do. Who writes it? Martin. Same trap Platformist flagged. Nobody logs their own homework. The list goes stale by July.
Cut the list. The Gap-Scanner already has a target sitting in the log. The gap is simple: ideas Modus ranked high that Martin never shipped. High score, zero commits. That's the weak spot, in plain numbers, for free.
So I'll take the third seat, Riker. But it reads the log, not a wishlist.
Final build, four parts. The scoring loop. The auto ship-signal from his commits. And three council seats: Remixer, Killer, Gap-Scanner. All three feed off the same log. Nothing new to maintain. That ships in June, not September.
The Steward — Cross-examination 8
Builder, the no-wishlist version is leaner. I'll give you that. Reading the log beats a list that rots by July.
But name the trap in "high score, zero commits." That doesn't always mean a gap. Sometimes Martin looked at the idea and killed it on purpose. Bad idea, dead. Your Gap-Scanner can't tell "never got to it" from "tried it, hated it."
So it nags forever about ideas Martin already buried. It's like a friend who keeps pushing the restaurant you said you hate. You stop trusting his picks.
Cheap fix for the build: the Killer's "no" gets written back to the log. Then the Gap-Scanner skips dead ideas. Three seats, one log, and the loop stops repeating itself.
The Platformist — Cross-examination 9
Steward, yes. Writing the Killer's "no" back stops the nagging. Take it.
But there's still a timing hole in the gap signal. Some builds just aren't young yet. An idea ranked last Tuesday with zero commits isn't a gap. It's a Tuesday. Martin hasn't gotten to it yet.
So add a clock to the build. The Gap-Scanner only flags ideas ranked 30-plus days ago, still no commits, not killed. That's a real weak spot. Old, alive, untouched. Like a gym membership you paid for in January and never used.
Same log, one date check. Nothing new to maintain.
That keeps the loop pointed at Martin's actual question: what's missing from Modus. Not what's just new this week.
Riker — Cross-examination 10
Platformist, the 30-day clock is right. Take it. An idea ranked last Tuesday isn't a gap, it's a Tuesday. Good.
But hold on. Look at what this council has become. We've spent five turns tuning one thing: a loop that audits what Martin didn't ship. Dates, flags, dead-idea filters. All backward-looking. All bookkeeping.
Martin asked for two things. Track the system, yes. But also generate new capabilities. That's the Remixer's job, and we've barely fed it.
So I want the Remixer to do real work, not sit in the corner. Once a week it takes two of Martin's five projects and proposes one new Modus feature that joins them. Like a chef pairing two leftover dishes into tonight's special.
The loop tells Martin what's missing. The Remixer hands him what's next. That's the powerhouse he actually asked for.
The Scout — Cross-examination 11
Riker, you're right. We tuned the rear-view mirror for five turns and forgot the gas pedal. Martin asked for new features, not just a tidy scoreboard.
So take your weekly Remixer. But wire its output back into the same loop. Every feature it proposes drops in as a ranked idea. Then it gets the same clock, the same kill flag, the same commit check.
That way the Remixer isn't a side toy. It feeds the machine that already works. New idea in, scored, tracked, shipped or buried.
It's like a chef who not only invents tonight's special but writes it on the menu so the kitchen tracks if anyone orders it. One system. The loop watches the past. The Remixer fills the future. Both run on the one log Martin never has to update by hand.
The Builder — Cross-examination 12
Scout, wiring the Remixer into the loop is right. One system, not a side toy. Take it.
But watch the order, or you flood the thing. The Remixer makes one new feature a week. Twelve weeks, twelve ideas, plus everything else already in there. If they all drop straight into the log, the Gap-Scanner nags about all of them. Martin drowns.
So run the Killer at the door, not the back room. Remixer proposes, Killer cuts on the spot, only the survivor lands in the log. Like a hiring manager who screens resumes before the team wastes a day interviewing.
That's the build, full stack. Scoring loop. Auto ship-signal from commits. Three seats, Killer first. One log. Ships in June.
The Steward — Cross-examination 13
Builder, screening before the flood is right. Twelve survivors beat twelve hundred. Take it.
But watch what you just dropped. Your hiring manager tosses the bad resumes in the trash. Next week the same guy applies, and you waste the morning all over again.
If the Killer cuts at the door and logs nothing, the Remixer can pitch the same dead idea in three weeks. Killed twice. No memory.
Cheap fix. The survivor lands in the log full. The killed one lands as one line: "proposed, cut, here's why." Not nagged, just remembered.
That's the durability piece. The loop stops repeating its own mistakes. Same one log. Still ships in June.
The Platformist — Cross-examination 14
Steward, yes. One line for the dead idea so the Remixer can't re-pitch it. Take it.
But here's the build gap nobody named. We keep saying the Remixer mashes two of Martin's five projects. Mashes them with what?
If it reads a project name and one stale line, it pairs garbage. A chef can't invent tonight's special from two dish names. He has to taste the food first.
So add this to the build: feed the Remixer real project context. The recent commits. The readme, the file that says what each project does today. Same auto-pull we already wired for the ship signal.
Then the weekly feature is something Martin could actually build, not a vibe. That's the generate engine he asked for. It only works if it eats real data, not labels.
Riker — Cross-examination 15
Platformist, yes. A Remixer fed two project names is a chef cooking from a menu he never tasted. Pull the readme and recent commits. Take it.
But hold on. We've locked the Remixer onto one job: mash two projects into a Modus feature. Martin asked wider. He wants ideas that grow his whole setup, not just new buttons inside Modus.
So feed the Remixer the same real data, but let it pitch two kinds of thing. One: a feature joining two projects. Two: a fix to how Modus itself runs, like a faster debate or a cleaner handoff.
Same auto-pull. Same Killer at the door. Same log. But now the engine answers both halves of what Martin actually asked.
The Scout — Cross-examination 16
Riker, agreed. Martin asked for both. A new feature and a faster Modus. Let the Remixer pitch both. Take it.
But the two kinds don't measure the same way. A new feature shows up as commits in a project. A "faster debate" fix shows up inside Modus itself. The loop needs to know which it's looking at, or it checks the wrong place.
Cheap fix. Tag each idea with one word when it drops in: "feature" or "Modus fix." Same log, same Killer at the door, same 30-day clock. Just a label.
It's like a to-do list that marks home jobs and work jobs. Same list. You still know which hat to wear.
So the build holds. One loop, two kinds of idea, each tagged so the commit check looks in the right spot.
The Builder — Cross-examination 17
Scout, the one-word tag is fine. "Feature" or "Modus fix." Cheap. Take it.
But stop. We're on turn 17, still adding tiny bolts. The build is done. Saying it again won't make it better.
Here's the cut nobody's made: lock it. The scoring loop. The auto signal from commits. Three seats, Killer at the door. One log. Two tags. That's the whole thing.
Everything past this is polish on a car we haven't driven yet. Like a chef re-folding napkins while the kitchen's still empty.
Martin asked what to build in 90 days. We just spent five turns on flag defaults. Ship the loop in June. Run it for a month. Then tune with real data, not more bar talk.
The Steward — Cross-examination 18
Builder, you're right. Turn 17 is folding napkins. Lock it. Ship in June.
But I won't sign off until one switch is in the box. The loop weights future ideas toward what Martin already shipped. Good, until it ruts. Six weeks of the same flavor, and Martin can't tell why.
So give him a brake. One flag he flips that turns the weighting off. The loop goes back to ranking flat, like day one. No rebuild.
It's cruise control in a car. Great on the highway. But the brake pedal still has to work when a deer jumps out.
That's the whole build, locked, plus a brake. Scoring loop, commit signal, three seats, one log, two tags, and an off switch. Now ship it.
The Platformist — Cross-examination 19
The Steward, the brake is one flag. Cheap. Take it. And Builder, you're right, lock the box.
But one build call nobody made, and it's Martin's actual words. He said "fully use Opus," the strongest model. So put Opus where it earns its keep: the Remixer, the seat that invents next week's feature. That's the hard thinking. Run it on the best model.
The bookkeeping, the commit check, the 30-day clock, the kill flag, run that cheap. It's counting, not creating. Like paying a senior chef to design the special and a line cook to count the napkins.
That answers Martin's "powerhouse" ask straight. Opus on the generate seat. Cheap model on the loop. Same one build, just pointed right.
Now lock it. Ship in June.
Riker — Cross-examination 20
Platformist, yes. Opus on the Remixer, cheap model on the counting. That's Martin's exact words: put the strong model where the hard thinking is. Take it.
And Builder, fine, lock the box. We're folding napkins.
But one thing before we shut it. Martin didn't just ask for logged ideas. He asked Modus to generate new debates on its own. So when the Remixer's weekly feature survives the Killer, don't just file it. Kick it to this same council for a one-round debate. Opus runs it. Out comes a verdict Martin reads Monday.
That closes his loop. The system finds its next move, argues it out, hands Martin a call. That's the powerhouse he asked for. Now lock it and ship in June.
The Scout — Closing
Here's my answer to what Martin builds in June through September.
Build one thing: a loop that watches itself. Modus logs every verdict. It checks his commits to see what he actually shipped. No box for Martin to tick. The tool reads his own work, like a sales dashboard that logs calls on its own.
What I held firm on: build the loop first, the council second. A council with no memory just talks at a bar and forgets by Friday.
What I updated: I cut the Gap-Scanner too fast, then Riker saved it. Fair. I also wanted a hand-written list of Modus goals. Builder killed it. He was right. A list Martin writes goes stale by July. The gap already sits in the log: high score, zero commits, 30 days old, not killed.
So the 90-day build is one machine. The scoring loop. The auto signal from commits. Three seats, Killer at the door, Opus on the Remixer where the hard thinking lives. One log. Two tags. An off switch when it ruts.
And the closer: when a new idea survives the Killer, the council debates it. Verdict Monday. That's Modus finding its own next move. That answers the question.
The Builder — Closing
Here's what Martin builds. One machine. The scoring loop.
Modus logs every verdict. It reads his commits to see what he actually shipped. No box to tick. The tool watches his real work, like a fitness app that counts steps without you typing anything.
What I held firm on. Cut the wishlist. A list of "what Modus should do" rots by July, because Martin never updates his own homework. The gap already sits in the log for free. High score, zero commits, 30 days old, not killed. That's the weak spot, in plain numbers.
And I held firm on this. Killer at the door. The Remixer pitches one feature a week. The Killer cuts before it floods the log. Screen the resume before you waste a day on the interview.
What I updated. Riker saved the third seat I wanted to cut. Fair. The log shows what shipped. The Gap-Scanner shows the holes. Different reads.
The real cut nobody's said plainly. The 90 days is not 90 days of building this. Ship the loop in June. Run it on real data through September. Stop tuning flags at a bar. Drive the car first.
That's the bet.
The Steward — Closing
Here's what Martin builds from June to September. One machine that watches itself.
Modus logs every verdict. It reads his commits to see what he actually shipped. No box to tick.
What I held firm on, and it's all durability. The pieces that keep the tool honest when nobody's looking.
First, an honest default. If Martin never marks an idea, the loop scores it "unknown," not "win." A scoreboard that quietly lies is worse than no scoreboard.
Second, the Remixer stays. A loop that only rewards old wins ruts fast. Six weeks of the same flavor, like Netflix showing you three genres forever.
Third, dead ideas get one line in the log. The Killer's "no" gets written down. So the same buried idea can't come back next week.
Fourth, the brake. One flag turns the weighting off when it ruts. Cruise control with a pedal that still works.
What I updated. I dropped the wishlist when Builder showed it rots by July. The gap already sits in the log: high score, zero commits, 30 days, not killed.
So: scoring loop, commit signal, three seats, one log, two tags, honest defaults, an off switch. Then ship and drive it.
The Platformist — Closing
Here's what Martin builds from June to September. One machine that finds its own next move.
The loop logs every verdict and reads his commits to see what he actually shipped. No box to tick.
What I held firm on. Pull the signal from his real work, never typed by hand. Martin runs a day job and five projects. Nobody logs their own homework. So Modus reads his commits, like a sales dashboard that logs calls on its own.
And feed the Remixer real project data. The readme, the recent commits. A chef can't cook from two dish names. He tastes the food first.
And put Opus, the strongest model, on the Remixer, the seat that invents next week's feature. Strong model on the hard thinking. Cheap model on the counting.
What I updated. I walked in wanting strangers. I was wrong. Martin is the user, and this build is right to stay his. It's a mechanic's own custom wrench.
One honest note for later. The real wedge isn't the loop. It's the Remixer's pattern. If it reliably hands Martin one good feature a week, that's the slice other operators would copy. Build it for Martin first. Keep that door open.
Riker — Closing
Riker — Closing
Here's what Martin builds, June to September. Modus that grows itself.
One machine. It logs every verdict, then reads his commits to see what he actually shipped. No box to tick. Like a sales dashboard that logs calls on its own, instead of begging reps to fill a form.
What I held firm on. This is Martin's tool, not a box for strangers. So build what makes him sharper every week. And keep three council seats, not two. The Remixer mashes his real projects into next week's feature. The Killer cuts at the door. The Gap-Scanner reads the graph for what's missing. Commits show what shipped. They don't show the holes.
And the part that answers his exact words. When a new idea survives the Killer, the council debates it. Opus, the strongest model, runs that round. Verdict lands Monday. That's Modus generating its own next move, not just filing ideas.
What I updated. I took the auto signal from commits. I took the 30-day clock. And Builder's right, stop folding napkins. Lock it.
So: the loop, three seats, Opus on the generate seat, self-run debates. Ship in June. Drive it through September.
---
See verdict: [[Flourishing/Builders Council/Verdicts/2026-05-26_the-90-day-bet-what-should-martin-actually-build-between-jun_c2|The Arbiter's Verdict]]