The Definitive Guide to Evaluations — WBS12
PART 1: WHAT AN EVALUATION ACTUALLY IS AND WHY IT IS NOT WHAT MOST STUDENTS THINK
26 min read
The most common misunderstanding on WBS12 is what evaluation means. Most students treat evaluation as one of these three things:
STUDENT BELIEF 1:
"Evaluation means writing 'however' and then
a second point."
→ WRONG. A second point is analysis.
Evaluation is the assessment of whether
the first point holds, matters, or is
more significant than the alternative.
STUDENT BELIEF 2:
"Evaluation means writing a conclusion."
→ PARTIALLY RIGHT. A conclusion can contain
evaluation. But evaluation must appear
throughout the answer — not just at the end.
A body with no evaluation and a conclusion
is Level 3. Full bilateral evaluation across
the whole answer is Level 4.
STUDENT BELIEF 3:
"Evaluation means 'it depends'."
→ WRONG. "It depends" is the absence of
evaluation, not evaluation itself. It
identifies that evaluation should exist
without actually providing it. Zero marks.
Confirmed every series in examiner reports.
What evaluation actually is:
Evaluation is the act of assessing the validity, significance, or reliability of an argument — answering one of three questions about every analytical point you make:
QUESTION 1 — VALIDITY:
"Is this argument actually correct given
the specific circumstances of this business?"
QUESTION 2 — SIGNIFICANCE:
"Does this argument matter more or less
than the competing argument, and why?"
QUESTION 3 — CONDITIONS:
"Under what specific circumstances does
this argument hold? What would have to
be true for it to be correct? What would
change the conclusion?"
Every evaluation you write on WBS12 is answering one or more of these three questions about an argument. If your "evaluation" is not answering one of these three questions, it is not evaluation — it is more analysis dressed up with the word "however."
PART 2: WHY EVALUATION EXISTS IN THE MARKING SYSTEM
Evaluation exists as a separate AO (AO4) because Pearson is testing whether you can do something fundamentally different from analysis. Analysis says: this causes that. Evaluation says: this causes that, but only if, or unless, or compared to the alternative, it matters this much because.
The distinction is built into the level descriptors:
LEVEL 2 DESCRIPTOR (all extended questions):
"A generic or superficial assessment is presented."
→ Evaluation attempted but not executed.
"It depends" territory. States that conditions
exist without identifying them.
LEVEL 3 DESCRIPTOR:
"An attempt at assessment is presented...
though unlikely to show the significance
of competing arguments."
→ Evaluation present. Conditions identified.
But the weight and significance of each
argument relative to the other is not
explicitly compared. One side tends to
dominate without justification.
LEVEL 4 DESCRIPTOR:
"Assessment is balanced, wide ranging,
well contextualised...shows an awareness
of the SIGNIFICANCE of competing arguments."
→ Not just: evaluation is present.
But: the significance is explicitly shown.
The student weighs which argument matters
more and explains why, with extract evidence.
The word significance is the Level 4 differentiator. Level 3 has evaluation. Level 4 has evaluation of the significance of competing arguments. That extra word is worth 2–5 marks depending on the question type.
PART 3: THE THREE TYPES OF EVALUATION — COMPLETE BREAKDOWN
Every evaluation on WBS12 is one of three types. Knowing which type you are writing determines what it needs to contain.
TYPE 1 — THE LIMITING EVALUATION
What it does: Identifies the condition under which the main argument is less strong, weaker, or invalid.
When to use it: P2 of the 20-marker (evaluating your own main argument). Mini-evaluations within the body of 10-markers. Any time you need to show that your main argument is not unconditionally true.
Why it exists: The Level 4 descriptor says "full awareness of the validity and significance of competing arguments." If your main argument is presented as unconditionally true, you are not demonstrating awareness of its limits. The limiting evaluation is what shows the examiner you understand when your argument holds and when it does not.
Structure:
LIMITING EVALUATION STRUCTURE:
TRIGGER:
"However, the significance of [main argument]
is limited by..."
CONDITION:
"...whether/if [specific condition from extract]"
MECHANISM:
"Because if [condition is true], then [mechanism
of why main argument weakens]..."
CONSEQUENCE FOR THE ARGUMENT:
"...meaning [main argument] holds only in
circumstances where [condition is satisfied]."
OPTIONALLY — WHAT THIS MEANS FOR THE QUESTION:
"This suggests [main argument] is [more/less]
convincing as the primary explanation because
[specific reason related to this business]."
Model limiting evaluation — Q3 P2 of this paper:
"However, the significance of domestic competition as the primary external cause is limited by the fact that China's digital reading market grew 18% to $6bn in 2021 — meaning external market conditions were actually favourable, not hostile to e-reader sales. This matters because if the external environment had been the primary cause, Amazon would expect market contraction, not growth. The growth of 18% in one year demonstrates that demand for e-readers was strong — meaning Amazon's failure to capture that demand is more convincingly attributed to internal product and marketing failures than to unfavourable external conditions. The external competition argument holds only if Chinese competitors succeeded because of inherent structural advantages Amazon could not replicate — but since they succeeded through product adaptation and targeted marketing, both of which were available to Amazon, the external cause is weakened significantly."
Why this is Type 1 (limiting): It takes the main argument (external competition) and systematically reduces its strength by showing that the conditions required for it to be the primary cause are not fully met in this specific context.
TYPE 2 — THE COMPARATIVE EVALUATION
What it does: Explicitly weighs two competing arguments against each other and states which is more significant and why.
When to use it: P4 of the 20-marker (after developing the competing argument, before or within the conclusion). The conclusion of 10-mark Assess questions. Any time the question asks you to choose between two explanations, methods, or strategies.
Why it exists: Level 4 requires "awareness of the significance of competing arguments." Awareness means more than presenting both — it means explicitly comparing them. The comparative evaluation is the mechanism by which this comparison is made. Without it, you have two good arguments sitting next to each other without any judgement about which matters more. That is Level 3. The comparative evaluation is what converts Level 3 content into Level 4 marks.
Structure:
COMPARATIVE EVALUATION STRUCTURE:
ACKNOWLEDGE:
"While [Argument B] presents a genuine
[cause/risk/benefit]..."
LIMIT ARGUMENT B:
"...this is less significant than [Argument A]
because [specific reason]..."
EXTRACT ANCHOR:
"...as evidenced by [specific extract data]
which shows [why A outweighs B]..."
DECISIVE FACTOR:
"The decisive factor is [X] because [Y]..."
CONDITION:
"This comparison holds only if [condition]..."
COUNTER-CONDITION:
"However if [condition changes], then
[Argument B] would become more significant
because [mechanism]."
Model comparative evaluation — Q3 conclusion of this paper:
"On balance, internal causes were more significant than external ones in explaining Amazon's withdrawal. While domestic competition from Xiaomi, iFlytek and Huawei presented genuine external pressure, this competition operated in a market growing 18% to $6bn in 2021 — meaning external conditions were favourable. The decisive factor is that Amazon's internal failures were controllable: neglecting the Harry Potter library, failing to build brand awareness among 500 million Chinese digital readers, and not adapting the hardware to Chinese preferences were all decisions Amazon made, not circumstances it faced. Unlike Chinese government regulation — which would represent a genuinely insurmountable external barrier — competitor innovation was a challenge Amazon had the resources and capability to meet. This comparison holds only if Chinese regulation did not specifically restrict Amazon's operations; if it did, external causes would become the primary explanation."
Why this is Type 2 (comparative): It does not just present both arguments — it explicitly states which wins ("internal more significant"), explains why ("controllable decisions"), uses extract evidence ($6bn/18%, Harry Potter, 500 million users), names the decisive factor (controllability), states the condition (regulatory barrier), and acknowledges when the conclusion would change.
TYPE 3 — THE CONDITIONAL EVALUATION
What it does: States the specific condition under which the overall conclusion holds, and the counter-condition under which it would change. This is the "only if / provided that / unless" evaluation.
When to use it: Every conclusion of every 10-mark and 20-mark levels question. Also valuable at the end of P2 and P4 evaluations within the body.
Why it exists: Pearson's supported judgement rule is explicit: a conclusion without a condition is unconditional — which is Level 3 maximum. The conditional evaluation is what converts a good conclusion into a supported judgement. It demonstrates that you understand your conclusion is not universally true — it depends on specific, nameable factors that you can identify and justify.
Confirmed from examiner reports across every series: "Trigger phrases — only if / provided that / conditional on — are required for Level 4." Not suggested. Required. The absence of a condition in a conclusion is a confirmed Level 4 blocker.
Structure:
CONDITIONAL EVALUATION STRUCTURE:
OVERALL CONCLUSION:
"Overall, [position — commit to one side]
because [decisive reason]..."
EXTRACT EVIDENCE:
"...as evidenced by [specific extract data]."
CONDITION:
"This conclusion holds only if [specific
condition — named and extract-anchored]."
COUNTER-CONDITION:
"However, if [condition changes / alternative
condition is true], then [other argument]
would become the more significant factor
because [mechanism]."
OPTIONAL — FURTHER QUALIFICATION:
"In the short run [X], but in the long run
[Y] — meaning the conclusion may change
over time as [business circumstance] evolves."
Model conditional evaluation — Q2e of this paper:
"Overall, Grupo Tamazula will face moderate difficulty keeping waste to a minimum, primarily because the perishable nature of fresh chilli ingredients combined with the growing distance of export supply chains to the US and Canada creates structural logistical challenges that factory-based technology cannot fully resolve. This conclusion holds only if Grupo fails to invest in dedicated cold-chain export logistics — if such investment were made, the difficulty in managing waste during international transport would reduce significantly, making the modern equipment argument the more decisive factor. In the short run, Grupo's existing domestic operations may be well-managed; but in the long run, as export volumes continue to increase, waste minimisation becomes progressively harder without specific supply chain investment."
Why this is Type 3 (conditional): The conclusion commits to a position, uses extract evidence (perishable chillies, US/Canada exports), states a condition explicitly ("holds only if"), states a counter-condition (cold-chain investment), and adds a time-based qualification (short run vs long run). Every element of a supported judgement is present.
PART 4: WHAT MAKES AN EVALUATION GENUINE VS FAKE
This is the most practically important distinction in this guide. The examiner report flags "superficial evaluation" and "generic evaluation" in every series. Understanding the difference between genuine and fake evaluation is what separates a student who scores 6/8 from one who scores 8/8.
FAKE EVALUATION — THE FIVE PATTERNS:
FAKE PATTERN 1 — THE "IT DEPENDS" PLACEHOLDER:
"However, it depends on the situation."
"This may or may not be true in all cases."
"It depends on whether Arditi Tours makes
the right decisions."
WHY IT'S FAKE: It identifies that a condition
exists without naming it. Zero evaluation marks.
Confirmed every examiner report: these phrases
score nothing.
FAKE PATTERN 2 — THE GENERIC COUNTER-ARGUMENT:
"However, there are disadvantages to this approach."
"On the other hand, not all businesses benefit."
"This might not always work."
WHY IT'S FAKE: No specific mechanism. No
extract data. No named condition. Generic
counter-arguments are Level 2 analysis
dressed up as evaluation.
FAKE PATTERN 3 — THE RESTATEMENT:
"Overall, Arditi Tours should be concerned
about its margin of safety because the margin
of safety is only 8 passengers."
(Conclusion restates the analysis.)
WHY IT'S FAKE: A conclusion that simply
repeats what the body argued is not evaluation.
Evaluation adds something new — a weighing,
a condition, a significance assessment.
FAKE PATTERN 4 — THE ADVANTAGE/DISADVANTAGE LIST:
"There are advantages: [list]. There are
disadvantages: [list]. Therefore it depends."
WHY IT'S FAKE: Listing more arguments is
analysis, not evaluation. Evaluation is the
assessment of which argument wins and why —
not the production of more arguments.
FAKE PATTERN 5 — THE UNANCHORED CONDITION:
"This only works if the company has enough
money." / "This depends on the economic
climate."
WHY IT'S FAKE: The condition is not anchored
to the extract. It could apply to any business
in any situation. Genuine evaluation uses
specific circumstances from this extract to
justify the condition.
GENUINE EVALUATION — THE FIVE MARKERS:
GENUINE MARKER 1 — NAMED CONDITION:
The condition is specific and named.
Not "if circumstances are right" but
"only if Grupo Tamazula can secure multiple
reliable fresh chilli suppliers who can
deliver to specification across its growing
export supply chain."
GENUINE MARKER 2 — EXTRACT ANCHORED:
The condition uses specific data or facts
from the extract.
Not "if the market grows" but "given that
China's digital reading market grew 18%
to $6bn in 2021, suggesting the external
environment was favourable not hostile."
GENUINE MARKER 3 — MECHANISM PRESENT:
The evaluation explains WHY the condition
matters — not just stating it.
Not "this holds only if demand is high"
but "this holds only if demand is high —
because at below-quarter capacity as Extract A
describes, even a small cost increase would
erode the 8-passenger margin of safety
entirely, turning profitable average months
into loss-making ones."
GENUINE MARKER 4 — SIGNIFICANCE STATED:
The evaluation explicitly says which
argument is more or less significant and why.
Not "both arguments are valid" but "the
internal cause is more significant than
the external one because it was controllable
— Amazon could have adapted its product,
unlike Chinese government regulation which
represented a genuine external barrier."
GENUINE MARKER 5 — SOMETHING NEW ADDED:
The evaluation adds information or perspective
not present in the analysis. The conclusion
does not just restate the body — it weighs,
qualifies, or extends what the body argued.
PART 5: HOW TO WRITE EACH TYPE OF EVALUATION — THE FORMULA
These are the practical templates. Not optional frameworks — the exact structure the mark scheme rewards.
FORMULA FOR TYPE 1 — LIMITING EVALUATION (P2 / P4 of 20-marker)
Trigger phrase options:
- "However, the significance of [argument] is limited by..."
- "This argument holds only if..."
- "However, [argument] depends critically on whether..."
- "The strength of this argument is contingent on..."
Required components:
① Name the limiting condition (specific to extract)
② Explain the mechanism by which it limits
the argument
③ State what the argument becomes if the
condition is not met
④ Connect back to the question: what does
this mean for internal vs external /
best method / concern level?
Time budget: 3–4 minutes. Word count: 60–80 words. Never more.
Example template applied:
"However, [main argument] holds only if [specific extract-anchored condition]. If [condition is not met], then [mechanism by which argument weakens], meaning [business name] would face [different consequence] — suggesting [argument] is a less decisive cause/factor/method than [alternative] in this specific context."
FORMULA FOR TYPE 2 — COMPARATIVE EVALUATION (Conclusion body)
Trigger phrase options:
- "On balance, [A] outweighs [B] because..."
- "[A] is more significant than [B] because..."
- "The decisive factor is [X] rather than [Y] because..."
- "While [B] presents a genuine [argument], it is less significant because..."
Required components:
① Acknowledge the other argument genuinely
(not dismiss it)
② State which argument wins and name it
③ Give the specific reason it wins using
extract evidence
④ Name the decisive factor
⑤ State the condition under which the
comparison would change
Time budget: 4–5 minutes. Word count: 80–100 words.
Example template applied:
"While [Argument B] presents a genuine [cause/risk], [Argument A] is more significant because [specific reason], as evidenced by [extract data]. The decisive factor is [X] — unlike [B], which [limitation], [A] [advantage specific to this business]. This comparison holds only if [condition]; if [condition changes], [B] would become the more significant factor because [mechanism]."
FORMULA FOR TYPE 3 — CONDITIONAL EVALUATION (Every conclusion)
Trigger phrase options:
- "This conclusion holds only if..."
- "However, this holds only provided that..."
- "This recommendation is valid only if..."
- "This judgement is contingent on..."
Required components:
① State the conclusion/position clearly
② Name the condition — specific, extract-anchored
③ Explain what changes if condition is not met
(the counter-condition)
④ For 20-marker: add something new not in body
(time-based dimension, further recommendation,
secondary condition)
Time budget: 3–4 minutes. Word count: 60–80 words on condition alone.
Example template applied:
"Overall, [position] because [decisive reason], as [extract evidence] confirms. This conclusion holds only if [specific extract-anchored condition] — if [condition is false / changes], then [other argument] would become the more significant factor because [mechanism]. [Optional: In the short run X, but in the long run Y as this business's specific circumstances evolve]."
PART 6: EVALUATION ACROSS EVERY QUESTION TYPE
Different question types require different amounts and types of evaluation. Using the wrong type or the wrong amount is as costly as not evaluating at all.
6-MARK ANALYSE — EVALUATION:
REQUIRED: ZERO evaluation.
ZERO AO4 marks exist on Analyse questions.
Writing evaluation on a 6-marker wastes time
and earns nothing.
IF YOU WRITE IT: Zero marks. 45-60 seconds
wasted. That time belongs to your second
reason's chain.
RULE: Hard stop at Stage 4. No "however."
No "this depends." No conditions. Stop.
8-MARK DISCUSS — EVALUATION:
REQUIRED: Implicit evaluation through balance.
The balance requirement IS the evaluation.
By developing both sides with chains you are
implicitly evaluating. A formal conclusion
is not required.
WHAT COUNTS AS EVALUATION HERE:
— Two genuine competing arguments each with
full chains (not one chain + one sentence)
— If you write a conclusion: Type 3
conditional evaluation for 7–8/8
WHAT DOES NOT COUNT:
— "Overall there are advantages and
disadvantages" — zero evaluation marks
— One good argument + one assertion = Level 2
regardless of how strong the first argument is
MARK POSITIONS:
6/8: Both sides with chains. No conclusion.
7/8: Both sides with chains + weak conclusion.
8/8: Both sides with chains + Type 3 conditional
evaluation in conclusion.
10-MARK ASSESS — EVALUATION:
REQUIRED: Type 1 (limiting) in body + Type 3
(conditional) in conclusion.
MINIMUM FOR LEVEL 4 (8/10):
— At least one Type 1 evaluation challenging
your main argument in the body
— Type 3 conditional evaluation in conclusion
— Both using extract evidence
WHAT MOST STUDENTS DO (scores 6–7/10):
— Two chains with no body evaluation
— Conclusion present but no condition
— This is Level 3 because the descriptor
says "unlikely to show significance of
competing arguments" — which is exactly
what a body-only analysis with a weak
conclusion produces
WHAT LEVEL 4 REQUIRES:
The competing argument must not just exist —
it must REDUCE confidence in your main argument.
Not: "here is another argument."
But: "here is why my main argument is not
the whole story."
MARK POSITIONS:
5/10: Chains present. No evaluation anywhere.
6/10: Chains + one evaluation attempt (weak).
7/10: Chains + Type 1 in body (condition stated
but no mechanism) + conclusion present
but no condition.
8/10: Chains + Type 1 with mechanism +
Type 3 conclusion with condition.
9/10: Everything at 8 + counter-condition +
significance explicitly stated in body.
10/10: Everything at 9 + conclusion adds
something new beyond the body.
20-MARK EVALUATE — EVALUATION:
REQUIRED: Type 1 (P2) + Type 2 (P4 transition)
+ Type 3 (conclusion) + significance throughout.
MINIMUM FOR LEVEL 4 (16/20):
— P2: Type 1 limiting evaluation of P1
— P4: Type 1 limiting evaluation of P3
— Conclusion: Type 3 conditional + Type 2
comparative in same paragraph
— All using extract evidence
MARK POSITIONS:
12/20: Chains present. Evaluation slots empty
or labelled but not written (your v1).
14/20: Evaluation content present but
incomplete chains in P2/P4.
Conclusion present. (your v2/v3)
16/20: P2 has mechanism chain. P4 has content.
Conclusion with condition. Primary cause
declared. (your v3/v4)
17/20: P2 full evaluation chain with condition.
P4 evaluation chain present.
Conclusion: condition + counter-condition
+ recommendation. (your v4)
18/20: P2 and P4 both fully developed with
mechanism chains AND significance
statements. Conclusion adds something new.
19/20: Everything at 18 + P2/P4 significance
explicitly answers the question's central
argument (not just evaluates the point).
20/20: Every evaluation adds something beyond
the analysis. Nothing is asserted.
Everything is argued. Conclusion decisive,
conditional, specific, new.
PART 7: THE EVALUATION TRIGGER PHRASES — COMPLETE LIST
These are the signal words that tell the examiner evaluation is occurring. Without them, evaluation content can be present but invisible — the examiner may read an evaluative point as analysis because it is not signalled.
╔══════════════════════════════════════════════════════╗
║ TYPE 1 — LIMITING EVALUATION TRIGGERS ║
╠══════════════════════════════════════════════════════╣
║ "However, this argument holds only if..." ║
║ "The significance of this is limited by..." ║
║ "This depends critically on whether..." ║
║ "However, [argument] is contingent on..." ║
║ "The strength of this point is conditional on..." ║
║ "This analysis assumes [X], which may not hold ║
║ because [extract evidence]..." ║
║ "In the short run [X], but in the long run [Y]..." ║
╠══════════════════════════════════════════════════════╣
║ TYPE 2 — COMPARATIVE EVALUATION TRIGGERS ║
╠══════════════════════════════════════════════════════╣
║ "On balance, [A] is more significant than [B]..." ║
║ "[A] outweighs [B] because..." ║
║ "The decisive factor is [X] rather than [Y]..." ║
║ "While [B] presents a genuine argument, ║
║ [A] is ultimately more convincing because..." ║
║ "[A] matters more in this specific context ║
║ because [business-specific reason]..." ║
║ "The primary explanation is [A] not [B] because ║
║ unlike [B], [A] was [controllable/specific/ ║
║ evidenced by extract]..." ║
╠══════════════════════════════════════════════════════╣
║ TYPE 3 — CONDITIONAL EVALUATION TRIGGERS ║
╠══════════════════════════════════════════════════════╣
║ "This conclusion holds only if..." ║
║ "This judgement is valid only provided that..." ║
║ "However, if [condition changes], then..." ║
║ "This recommendation is contingent on..." ║
║ "This holds unless..." ║
║ "The conclusion would change if..." ║
║ "In the short run [X] — but if [condition], ║
║ then in the long run [Y]..." ║
╠══════════════════════════════════════════════════════╣
║ ZERO MARKS — THESE ARE NOT EVALUATION: ║
╠══════════════════════════════════════════════════════╣
║ ✗ "It depends on the situation." ║
║ ✗ "There are advantages and disadvantages." ║
║ ✗ "This may or may not work." ║
║ ✗ "Every business is different." ║
║ ✗ "Overall both arguments are valid." ║
║ ✗ "This could be good or bad." ║
║ ✗ "It is difficult to say." ║
╚══════════════════════════════════════════════════════╝
PART 8: THE EVALUATION QUALITY LADDER — FROM ZERO TO PERFECT
Every evaluation you write sits on this ladder. Knowing which rung you are on tells you exactly what to add to move up.
RUNG 0 — NO EVALUATION (Level 1/2):
No attempt at evaluation. Pure analysis
or assertion on both sides.
Example: Two chains. No "however." No
conditions. No weighing. Body stops
after Stage 4 on each side.
↑ ADD: Any evaluation trigger phrase
+ one condition (even without mechanism)
RUNG 1 — PLACEHOLDER EVALUATION (Level 2):
Evaluation signalled but not delivered.
"It depends." "However this may vary."
"This is not always the case."
Identifies that conditions exist without
naming them. Zero evaluation marks awarded.
↑ ADD: Name the specific condition
(what specifically does it depend on?)
RUNG 2 — STATED CONDITION (Low Level 3):
Condition named but not anchored to extract
and no mechanism.
"This holds only if demand is high enough."
"This depends on whether Amazon had
good products."
Right structure. Wrong content.
Condition is generic, not business-specific.
↑ ADD: Extract data to anchor condition
+ one sentence explaining why the
condition matters
RUNG 3 — ANCHORED CONDITION (Level 3):
Condition named AND anchored to extract.
"This holds only if diesel prices remain
at €1.58/litre — if they rise, the
break-even threshold increases."
Condition is specific. Extract data present.
Missing: mechanism chain + significance.
↑ ADD: What happens if condition
is not met (mechanism) +
what this means for which argument wins
RUNG 4 — ANCHORED CONDITION + MECHANISM (Top Level 3):
Condition named, anchored, with mechanism.
"This holds only if diesel remains at
€1.58/litre — if prices rise, variable
costs per journey increase, contribution
per passenger falls, break-even rises above
17 passengers, and the current 8-passenger
margin of safety is eroded."
Missing: significance statement
(why this matters more/less than other side)
↑ ADD: One sentence stating which
argument wins overall and why,
OR what condition would change
the conclusion entirely
RUNG 5 — FULL EVALUATION (Level 4 entry):
Condition + anchor + mechanism +
significance statement.
"This holds only if diesel remains at
€1.58/litre — if it rises, the margin of
safety erodes, making concern justified.
This is significant because diesel is a
variable cost Arditi cannot control,
meaning the external cost environment
is the decisive factor in whether the
current margin is sufficient."
All components present. Level 4 accessed.
↑ ADD: Counter-condition +
something new in conclusion
RUNG 6 — COMPLETE EVALUATION (Level 4 strong):
Everything at Rung 5 PLUS:
Counter-condition that would change
the conclusion entirely.
"However if diesel prices stabilise and
demand returns to above-average levels,
the 8-passenger margin becomes comfortable
and concern is not warranted."
Conclusion adds new perspective not in body.
RUNG 7 — PERFECT EVALUATION (20/20):
Everything at Rung 6 PLUS:
Time-based dimension (short run vs long run)
OR secondary recommendation
OR limitation of the conclusion itself.
Nothing asserted. Everything argued.
Every condition extract-anchored.
Every comparison explicitly justified.
PART 9: YOUR EVALUATION PATTERNS — THIS PAPER
Based on every question marked in this session, here is exactly where your evaluations sat on the ladder and what moved them up.
╔══════════════════════════════════════════════════════╗
║ YOUR EVALUATION PATTERNS — WBS12 OCT 2023 ║
╠══════════════════════════════════════════════════════╣
║ Q1d DISCUSS (6/8): ║
║ Evaluation type: Implicit through balance + ║
║ partial Type 3 in conclusion. ║
║ Rung reached: Rung 3 (anchored condition on ║
║ diesel, but mechanism chain incomplete). ║
║ What was missing: mechanism chain on the ║
║ cash flow argument + significance statement. ║
║ What would have scored 8/8: Complete the ║
║ diesel chain to break-even consequence + ║
║ state which argument wins overall. ║
║ ║
║ Q1e ASSESS (6/10): ║
║ Evaluation type: Type 3 attempted (condition ║
║ stated: "if demand is elastic/inelastic"). ║
║ Rung reached: Rung 2 (condition stated but ║
║ generic — not extract-anchored specifically). ║
║ What was missing: Explicit comparison of ║
║ price increase vs alternatives + specific ║
║ recommendation for Arditi Tours. ║
║ What would have scored 9/10: "This conclusion ║
║ holds only if Arditi's demand is inelastic — ║
║ given several competitors offer the same route, ║
║ demand is more likely elastic, meaning the ║
║ price increase would reduce revenue. Therefore ║
║ the best alternative is [specific extract-based ║
║ recommendation]." ║
║ ║
║ Q2e ASSESS (7/10 v3): ║
║ Evaluation type: Type 3 in conclusion (condition ║
║ on reliable suppliers). Type 1 partially in body. ║
║ Rung reached: Rung 4 (condition anchored, ║
║ mechanism present but condition qualifies ║
║ recommendation not judgement). ║
║ What was missing: Condition applied to the ║
║ overall judgement not just the recommendation + ║
║ explicit weighing of capability vs difficulty. ║
║ What would have scored 9/10: "This conclusion ║
║ holds only if Grupo cannot secure reliable ║
║ suppliers — if it can, JIT becomes viable ║
║ and difficulty reduces significantly. The ║
║ decisive factor is therefore supplier ║
║ reliability, not modern equipment capability." ║
║ ║
║ Q3 EVALUATE: ║
║ v1 (12/20): Evaluation slots empty. Rung 0. ║
║ v2 (14/20): Evaluation content present. ║
║ Rung 2 (conditions stated, generic). ║
║ v3 (16/20): P2 condition anchored. Rung 3–4. ║
║ Conclusion with condition. Rung 5. ║
║ v4 (17/20): P2 mechanism chain. Rung 5–6. ║
║ P4 evaluation chain. Rung 4. ║
║ Target v5 (18/20): P4 significance sentence. ║
║ Rung 6 throughout. ║
╠══════════════════════════════════════════════════════╣
║ OVERALL EVALUATION PATTERN: ║
║ Starting position: Rung 0–2 on all questions ║
║ End position: Rung 4–6 on extended questions ║
║ Primary gap: Rung 3→4 transition (adding ║
║ mechanism chain to stated conditions) ║
║ Secondary gap: Rung 5→6 transition (adding ║
║ counter-condition to conclusions) ║
╚══════════════════════════════════════════════════════╝
PART 10: THE EVALUATION CHECKLIST — EXAM DAY
╔══════════════════════════════════════════════════════╗
║ BEFORE WRITING ANY EVALUATION — CHECK: ║
╠══════════════════════════════════════════════════════╣
║ □ Am I on a 6-marker? If yes: WRITE NO EVALUATION. ║
║ Move to next reason immediately. ║
║ ║
║ □ Am I writing P2/P4 of a 20-marker or body ║
║ evaluation of a 10-marker? ║
║ → Use TYPE 1 (limiting evaluation). ║
║ → Must include: condition + extract anchor + ║
║ mechanism chain. ║
║ ║
║ □ Am I writing a conclusion? ║
║ → Use TYPE 3 (conditional) as minimum. ║
║ → Must include: position declared + ║
║ "only if..." condition + counter-condition. ║
║ → On 20-marker: also TYPE 2 (comparative). ║
║ ║
║ □ Have I named the specific condition? ║
║ → Not "if circumstances allow" ║
║ → But "only if [specific extract-anchored fact]" ║
║ ║
║ □ Have I built a mechanism chain from the ║
║ condition? (What happens if condition fails?) ║
║ ║
║ □ Have I stated significance? ║
║ → Which argument matters more and why? ║
║ → What is the decisive factor? ║
║ ║
║ □ Does my conclusion add something new? ║
║ → Not a restatement of the body. ║
║ → A specific recommendation or time-based ║
║ qualification or secondary condition. ║
║ ║
║ □ Have I avoided all zero-mark phrases? ║
║ → No "it depends." No "both are valid." ║
║ → No "advantages and disadvantages." ║
╚══════════════════════════════════════════════════════╝
PART 11: WHY THIS STRUCTURE AND NOT ANYTHING ELSE
Every component of evaluation technique maps to a specific Pearson marking requirement:
EVALUATION COMPONENT PEARSON REQUIREMENT
──────────────────────────────────────────────────────
Named condition "Supported judgement" —
Level 4 Assess/Evaluate
descriptor. Without a
named condition, the
judgement is unconditional.
Unconditional = Level 3 max.
Confirmed every series.
Extract-anchored "Well contextualised" and
condition "effective use of business
context throughout" — Level
3/4 descriptor. Generic
conditions that could apply
to any business do not
satisfy contextualisation.
Mechanism chain "Chains of reasoning showing
from condition cause(s) and/or effect(s)" —
AO3 component present in
all level descriptors. A
condition without a mechanism
is an assertion, not analysis.
Significance "Awareness of the significance
statement of competing arguments" —
the exact Level 4 differentiator
phrase in every extended
question descriptor. Without
explicit significance, Level 3
is the ceiling regardless of
how good everything else is.
Counter-condition "Effective conclusion" —
20-mark descriptor. "Leading
to a supported judgement" —
10-mark descriptor. The counter-
condition is what makes a
judgement supported rather than
merely conditional.
Something new "Effective conclusion...
in conclusion proposes a solution and/or
recommendations" — 20-mark
Level 4 descriptor. A conclusion
that only restates the body
does not add the "recommendation"
element the descriptor requires.
Remove any component and a Level 4 descriptor requirement goes unmet. Every component exists because a specific mark criterion demands it. Nothing is stylistic preference. Everything is marks.
Chains earn the analysis marks. Evaluations earn the Level 4 marks. Master both and you have mastered every mark available on this paper.
Up next
The Definitive Guide to Judgements — WBS12
PART 1: WHAT A JUDGEMENT ACTUALLY IS AND WHY IT IS THE RAREST SKILL ON THIS PAPER
36 min