Force the model to quote your essay before it scores it
Marks your Task 2 essay in four locked steps — evidence first, descriptors second, band last — so it cannot work backwards from a flattering number.
Best on Claude · ChatGPT · Bands 6.0–8.0
Marking prompts that quote your essay before they score it.
Almost every IELTS writing prompt circulating online is one sentence long: "act as an IELTS examiner and score my essay." It returns a number, the number is usually flattering, and running the same essay twice returns two different numbers. That is not a marking tool — it is a random band generator with a friendly tone.
The prompts in this category are built around four mechanisms that fix it. First, the rubric is loaded into context rather than recalled from memory, because a model paraphrasing the descriptors is not applying them. Second, an output contract — a fixed table, a fixed field list — so two runs are comparable and inflation has nowhere to hide. Third, evidence before judgement: a model forced to quote your text before assigning a criterion band cannot work backwards from a number it already picked. Fourth, a calibration anchor where possible, so the standard being applied is one you can point at.
What none of them do is give you a trustworthy band on their own. In our own study of 1,200 essays, purpose-built AI grading matched the human examiner consensus within ±0.5 bands 94.2% of the time and exactly 78.3% of the time. A general-purpose chatbot with a pasted rubric is looser than that, and its number moves with the framing of the prompt — which is exactly what the "defend a lower band" prompt below is designed to expose. Use these prompts for the reasoning, not the digit.
The order that works: mark the essay with evidence first, then stress-test that mark by making the model argue for half a band lower, then take whichever specific weakness both passes agree on and fix that one thing. Coherence and cohesion is usually the criterion worth isolating on its own, because it is the one candidates self-assess worst — your own writing always feels logical to you.
Marks your Task 2 essay in four locked steps — evidence first, descriptors second, band last — so it cannot work backwards from a flattering number.
Best on Claude · ChatGPT · Bands 6.0–8.0
Marks a reference essay whose real band you already know, forces a recalibration, then marks yours against that same standard.
Best on ChatGPT · Claude · Bands 6.0–7.5
Asks the model to build the strongest case against the band it just gave you, then arbitrate honestly between the two.
Best on Claude · ChatGPT · Bands 6.0–7.5
Maps your controlling ideas, referencing chains and linker repeats — the criterion most learners cannot self-assess — with grammar and vocabulary explicitly off-limits.
Best on Claude · ChatGPT · Bands 6.0–7.5
Runs a yes/no scorecard over the two most formulaic paragraphs in Task 2, quotes the evidence, then rewrites both in your own voice.
Best on ChatGPT · Claude · Bands 5.5–7.0
For descriptor-anchored marking with long inputs, Claude and ChatGPT both handle the full rubric plus an essay reliably, and Gemini is strong when you need a chart or an image read as well. The tool matters far less than whether you pasted the actual descriptors and forced evidence before judgement.
It can give you a plausible one, which is not the same thing. Bands from general chatbots move with prompt framing — the same essay can come back 6.5 or 7.5 depending on how strict the persona is. Treat any AI band as a range and rely on the quoted evidence instead.
Yes, into every new chat. Models paraphrase the descriptors from memory and the paraphrase drifts, which defeats the entire point of descriptor-anchored marking. The official public descriptors are a free download from ielts.org.