No. 08 · In design · AI & agents · ROUTE
The Model Portfolio
How to Route, Mix, and Govern LLMs Like a Strategic Asset Allocator
Stop picking models off the leaderboard.
Choosing one model off the leaderboard is a bet you keep re-losing. The Model Portfolio treats your LLMs like a governed asset mix — routed and managed by cost, latency, privacy, and capability — so model choice becomes a strategy, not a guess.
- pages
- 418
- chapters
- 14
- hours of reading
- ± 6
- editions
- EN · NL
- Design
- Drafting
- Manuscript
- Production
- Launched
The book
How to Route, Mix, and Govern LLMs Like a Strategic Asset Allocator
A new flagship ships, the benchmark chart looks decisive, and you migrate everything to it by Friday. Then latency creeps, the bill doubles on calls that never needed the smartest model, and a capability you relied on regresses in an update you cannot refuse. You run production on one vendor's roadmap and call it simplicity. It is exposure.
The standard move is to pick the best model and standardize on it. That fails because there is no best model, only one that is best for a task, a budget, and a failure tolerance, and those stop agreeing the moment traffic is real.
The Model Portfolio treats your models the way an investor treats capital: assets you allocate on purpose, not one bet you keep doubling. It introduces the PORTFOLIO framework (Profile, Options, Route, Test, Fallback, Iterate), the moves that turn a stack of API keys into a managed system. The framework is the structure; the model you favor this quarter is weather. When the leaderboard turns over, your architecture holds.
What you learn
What this book puts in your hands
- The ROUTE protocol: govern a living mix of models, not one pick
- Manage by cost, latency, privacy and capability
- Route work to the right model on purpose
The framework
ROUTE, step by step
govern a living mix of models, not one pick
Look inside
The strongest pages — frameworks, figures, and worksheets from the print edition.
The contents
Chapter by chapter
Every chapter of The Model Portfolio with its printed epigraph, what you can do afterwards, and the moment it is built for.
Chapter 0
Stop Worshipping the Leaderboard
The benchmark winner is always someone else's model on someone else's tasks: priced, measured, and celebrated in conditions that have nothing to do with your production environment.
Chapter 1
The Model Picking Problem
You can find a model selection process by its artifacts: a Slack message, a benchmark screenshot, and a decision that took four minutes. The evidence that it needs review often appears much later.
Chapter 2
ROUTE Overview
Five skills. One feedback loop. The difference between a model collection and a model portfolio is not the models you hold: it is the discipline you apply to them.
Chapter 3
Register the Portfolio
Chapter 4
Objective Typing
Chapter 5
Routing Policies
A policy written in prose can be debated. A policy written as a decision table can be tested. Only one of those options tells you whether your routing does what you intended.
Chapter 6
Cascades and Fallbacks
Chapter 7
Local vs Cloud vs Hybrid
Chapter 8
Privacy and Compliance Tiers
Chapter 9
Track Cost and Latency
The invoice is not the dashboard. By the time a billing statement surfaces a cost problem, the routing decision that caused it has been running for weeks.
Chapter 10
Quality Signals and Eval
Chapter 11
Evolve the Mix
Chapter 99
Conclusion: Allocate Models Like Capital
A portfolio that is not governed will still be managed: by whoever made the last ad-hoc decision. The only question is whether that management is deliberate.
Chapter 100
Evidence ledger
Who it is for
Who this book was written for
The result is concrete: you stop overpaying a frontier model for work a smaller one does well, you stop being hostage to one provider's release notes, and your system grows cheaper and more resilient over time.
If you run more than one model in production and refuse to confuse a default with a decision, this was written for the operator who treats capability as an allocation, not a habit.
The reader it was written for
The model-routing lead. Founder or tech lead choosing among open, closed, local, and frontier models across an agent stack — tired of ad-hoc picks driven by benchmarks and Twitter heroes.
Also a fit for
The cost-conscious operator. Owns API bills and latency SLOs; needs explicit trade-offs, not "use the latest."
What you will use it on
- Register models with capability, cost, and privacy class
- Type tasks and map them to routing requirements
- Implement routing policies, cascades, and fallbacks
- Track cost, latency, and quality signals per route
- Evolve the portfolio — upgrade, demote, retire
Probably not for you if
- Readers training models from scratch
- Readers implementing retrieval pipelines
- Pure ML governance without routing implementation
Editions
Editions and specifications
| Edition | Formats | Chapters | Pages | Reading time | ISBN (paperback) |
|---|---|---|---|---|---|
| English The Model Portfolio | In production | 14 | 418 | ± 6 hours | — |
| Dutch Het Model-Portfolio | In production | 14 | 454 | ± 5 hours | — |
Both editions are written natively. The Dutch text is not a machine translation of the English. · Trim size: 6x9″
Frequently asked
What readers usually want to know
What is The Model Portfolio about?
ROUTE turns model choice into portfolio management across cost, latency, privacy, reliability, capability, and governance. The subtitle is: How to Route, Mix, and Govern LLMs Like a Strategic Asset Allocator.
What is the ROUTE framework?
govern a living mix of models, not one pick
Is there a Dutch edition?
Yes. The Dutch edition is Het Model-Portfolio, written as a native edition rather than a machine translation. It moves through the same production line.
How long is The Model Portfolio?
This edition runs 14 chapters, 418 pages in print and roughly 6 hours of reading.
Who is The Model Portfolio for?
If you run more than one model in production and refuse to confuse a default with a decision, this was written for the operator who treats capability as an allocation, not a habit.
The production system
How this book was made
Every title moves through the same gated production line: sourced research, a claim-level evidence ledger, structural review, fact-checking, red-team critique, and a bilingual final edit. AI agents do specialist work inside those gates; judgment, voice, and accountability stay human.
- Claims enter an evidence ledger with a source and a confidence grade before they reach the page
- English and Dutch are two native editions, not a translation of one another
- Every chapter clears readability, rhythm, and style gates before it is typeset
The series