My numbered prompt hit all 10 rules. It still dropped a warning.
The advice everyone repeats
Common advice says to number your instructions so the AI follows all of them. Anthropic's prompting guide recommends numbered lists "when the order or completeness of steps matters." So I tested it.
Source: Anthropic, Prompting best practices
Same notice, two prompt formats
The same fictional office-move notice, rewritten by GPT-5.5 (OpenAI API, Oct 2026) with 10 requirements written two ways. Each ran twice, plus a smaller 6-requirement notice.
Numbering didn't make it obey more
Every run met every listed requirement in both formats: 6 of 6 on the short notice, 10 of 10 on the move notice. In this test, numbering didn't add obedience.
The warning that went missing
Real output · GPT-5.5 (OpenAI API) · Oct 2026
Step 8 said only "Say that IT handles monitors and docking stations." In one numbered run, that's all it did, and the "Do NOT pack your monitor" warning was gone.
The model reads each line literally
OpenAI's guide says GPT-5.5 "interprets prompts in a literal and thorough manner." So a line that states part of a requirement can get just that part. The paragraph prompt used similar wording and kept the warning.
Source: OpenAI API docs, prompt guidance for GPT-5.5
Write the full requirement
Tested in GPT-5.5 (OpenAI API), Oct 2026
Rewriting step 8 as the whole requirement, with its reason, brought the warning back in both reruns in this test.
Number it, then tick it off
Number your requirements anyway, so you can tick each one off against the output. Write each line as the full requirement, reason included.
Try it: Number your next rewrite request, then check the result line by line.
Sources and assumptions
- Anthropic, Prompting best practices (Claude Platform Docs), section 'Be clear and direct': Verbatim: 'Provide instructions as sequential steps using numbered lists or bullet points when the order or completeness of steps matters.' Also: 'Be specific about the desired output format and constraints.' (checked 2026-10-05)
- OpenAI API docs, Prompt guidance (model: GPT-5.5): Verbatim: 'GPT-5.5 interprets prompts in a literal and thorough manner, enabling specific, descriptive instructions when the product requires them.' Also: 'Describe the expected outcome, success criteria, allowed side effects, evidence rules, and output shape. Avoid step-by-step process guidance unless the exact path matters.' (checked 2026-10-05)
Assumptions:
- All sample notices are fictional (Dana ext 214, Priya Shah ext 309 and Marcus are invented names); no real personal data.
- All runs used GPT-5.5 via the OpenAI Chat Completions API with default settings (no system message, default reasoning effort and temperature), on 2026-10-05 US time (ran_at timestamps 2026-10-06 UTC). The result is from the API, not the ChatGPT app.
- Claude Sonnet 5.5 runs were attempted (test ids paragraph-1 and numbered-1), but the local claude runner failed in this research session and produced no output. The post must not claim any Claude result.
- Small sample: 2 runs per prompt version (10 runs in total). Results can vary between runs and models, so the slides should say 'in this test', not 'always'.
- The 'numbered list' here is a list of requirements for the output (success criteria), not process steps. That fits both vendors' guidance.
- Scoring by hand: short notice (6 requirements: exact date/times, subject line, under 100 words, parking left out, 'What to do:' line, ends with contact + extension): all 4 runs met 6/6. Move notice (10 requirements: subject under 8 words, every date/time exact, under 120 words, no parking, no apology, exactly 3 pack bullets, 'Deadline:' line, IT handles monitors/docks, signed 'Facilities Team', ends with contact + extension): all 4 runs met 10/10. The longest output (move-paragraph-1) is about 94 words.
- The dropped warning: only move-numbered-2 left out the original 'Do NOT pack your monitor or docking station'. move-numbered-1 and both paragraph runs kept a 'Do not pack' line, even though the paragraph prompt also only said 'make sure people know IT handles monitors and docking stations'. So the cause shown is narrow wording read literally, not the numbered format.
The short version
- Numbering didn't make GPT-5.5 obey more here
- It does turn your prompt into a checklist
- Each line must state the full requirement
- Add the reason, like "because IT moves those"


