Routing ← Home

Not every question deserves the full ensemble.

Triangulation is expensive in attention, and attention is the scarce resource in this method. Running three platforms on a question that needed one is not rigor — it is ceremony, and it burns the focus you'll want when a question genuinely warrants the session.

The core routing question

Before you open anything, ask one question: what does this task most require — breadth, depth, or grounding?

  • Breadth — options, variations, angles you haven't considered.
  • Depth — structured reasoning, tradeoff analysis, a coherent argument that holds together.
  • Grounding — facts, sources, verification, contact with the external record.

Most real tasks need more than one. But they usually have a dominant need, and naming it decides both who you call and whether you need everyone.

When the dominant need is grounding, naming it also decides the order. Brief whichever of the three you will use for retrieval first — before turning the others loose — so the generative work builds on a grounded foundation instead of being retrofitted to one afterward. Order matters more than people expect.

Depth

Claude first

When the question is ambiguous or poorly defined and needs to be made tractable before anything else can happen. When you need a structured document that will serve as the foundation for the rest of the work. When the stakes involve ethical or long-term dimensions. When you have raw material — notes, transcripts, a pile of sources — that needs organizing into an argument.

Not first whenYou need ten ideas fast. The hedging that protects you at high stakes will slow you down here.
Breadth

ChatGPT first

When you need a first draft on the page. When you want variations on a message or an approach. When you're in early-stage exploration and the goal is to widen the field rather than narrow it. When you need the same concept explained at three different levels of complexity.

Not first whenAccuracy matters more than fluency and you cannot easily verify the output yourself.
Friction

Grok first

When you suspect you're stuck in a frame and can't see its edges. When you have a tentative plan and want it stress-tested before you commit. When the other platforms have been agreeable and you don't trust that agreement. When you want the unconventional angle that a more cautious system won't volunteer.

Not first whenYou're locking a final decision. Improvisation belongs earlier in the process.
Grounding

Retrieval first

When the answer has to be connected to sources. When facts need checking before anything is built on them. When the question touches a domain where training-data limits could produce plausible but outdated information — pricing, regulations, current status, anything that moves. Brief one of the three for retrieval and citation explicitly; grounding is a mode you ask for, not a member you add.

AlsoGrok is usually the strongest starting point for live material, and ChatGPT's Deep Research for a cited report — but the assignment is yours to make, and worth re-auditioning as the platforms change.

When to convene everyone

Brief the full ensemble when at least one of these is true:

  • The stakes are real and the question is genuinely open.
  • You suspect your framing is wrong — which is exactly the thing you can't detect alone.
  • The domain crosses modes: it needs facts and judgment and range.
  • You haven't run this kind of question before, so you don't yet know what a good answer looks like.

Escalation

You will often start with one platform and discover mid-session that you needed three. Escalate when a single response surfaces complexity you didn't anticipate, when it contradicts an assumption you'd been treating as settled, or when the stakes changed while you were working.

Escalating is not an admission that you routed badly. It is the routing decision working — you learned something that changed the shape of the question.

Stress-test brief · friction pass
CONTEXT:
[Brief description of the plan, decision, or recommendation to be tested.]

TASK:
I have a tentative direction on the above. Before I commit, I want you
to stress-test it. Please:

· Identify the weakest assumptions in my reasoning
· Describe what could go wrong, and how I would know early
· Surface any angles I have not considered
· Tell me what you would need to see to change your assessment

Do not soften this. I need the friction, not reassurance. If the plan is
sound, say so — but only after you have genuinely tried to break it.

When the result disappoints

A weak answer does not always mean you asked badly. Before you rewrite the prompt for the fourth time, run the diagnosis: is the problem the prompt, the platform, the task, or the routing choice?

Sometimes the task is simply pressing against the platform's native architecture. Prompting can improve your access to a system's strengths; it cannot manufacture strengths the system doesn't have. Architecture sets the reachable range; prompting determines how much of that range you can actually use. If the limitation is architectural, stop polishing and change the instrument.

And be careful with reputation. A platform's standing in your memory reflects the version you used months ago, filtered through whatever you were doing at the time. Capability moves faster than reputation does, in both directions. Audition the platform in the task, not in your memory of its brand.

Field note

Knowing when not to convene is part of the skill. Once the question is framed well, a single platform can carry a low-stakes task without full ensemble review. Stepping back is not carelessness — it is part of the discipline.

The field guides

Eight guides, one method. Roughly in the order a reader meets them: how the field got here, how to work across models, then the tools themselves.