A software company's incident tool uses Claude to turn on-call engineers' notes into public status-page updates. Its prompt holds a written instruction and four worked examples, which were added a year ago because zero-shot updates varied widely in tone and length. Last month the communications team changed the house format from a single paragraph to three labelled lines (impact, current status, time of next update) and rewrote the instruction to match, but left the four examples unchanged. Of 300 updates since, 71 per cent still arrive as a single paragraph. The fix must ship this week, and the platform team will not add model calls per update or change the model tier. What should the architect recommend?
- ARewrite the four worked examples in the three-line format so they match the new instruction Correct
- BRestate the three-line rule in capitals at the end of the prompt so it outweighs the examples
- CRemove the four worked examples and let the rewritten instruction set the format by itself
- DAdd a second call that checks each update and regenerates any not written in three lines
Why A is correct: Examples are the strongest signal of output format in a few-shot prompt, and these four still demonstrate the retired paragraph shape. Rewriting them removes the conflict, keeps the tone and length control they were added for, and needs no extra calls or model change.
Why B is wrong: Emphasis and placement do make an instruction more salient, so this is a natural first move. It leaves four examples demonstrating the old paragraph format, so the prompt still sends two conflicting format signals and the examples, which show the exact output shape, keep pulling the model back to them.
Why C is wrong: Removing the conflicting examples would probably fix the format, which makes this attractive. The stem says the examples were added because zero-shot updates varied in tone and length, so deleting them trades away a control the team still needs when updating them keeps both.
Why D is wrong: A checking pass is a reasonable compensating control and would catch format failures. The stem rules out extra model calls per update, and a check-and-retry loop treats the symptom while the prompt keeps producing the wrong format most of the time.