Dashboard
0%
1
Curious builder0 XP earned · 300 to level 2
0 daysFinish a lesson to begin
Badge collection0 of 6 unlocked
51 small wins to finish your pathNext question →

Q17IntermediateConcept

Explain shadow, canary, blue-green and A/B deployments for LLM changes. When do you use each?

30-second answerSay your answer out loud first, then reveal.
A router sends 100% of live traffic to the current version, a copy to a shadow version whose outputs are logged and judged but not served, and 5% to a canary; a comparison of errors, latency, cost, judge scores and feedback then either expands the rollout to 25%, 50% and 100% or rolls back.

When to use each

StrategyUser exposureBest forCost / caveat
ShadowNoneModel swaps, big prompt rewrites, risky retrieval changesDoubles inference cost for shadowed traffic; can't measure user reactions; avoid side-effecting tools in shadow
CanarySmall %Most releasesNeeds enough traffic for signal; automated analysis
Blue-greenAll at once (switchable)Infrastructure changes, inference-server upgradesDuplicate infrastructure (expensive for GPUs)
A/B testSplit for weeksMeasuring business impact of prompt/model changesRequires experiment design and statistical rigour

LLM-specific note: shadowing agents is tricky because tools have side effects. Run shadows with read-only or mocked write tools.

You understood something today that you didn't yesterday.