Academy Writing guide

CELPIP Writing scoring: how your level is decided

How CELPIP Writing scoring works: the four rated dimensions, the four to six raters who score you, what happens when they disagree, and where marks are really lost.

HelloCelpip Team
12 min read

Your CELPIP Writing result is a single level from 0 to 12, decided by four to six trained raters who score both of your tasks against four rated dimensions. There is no percentage, no raw mark, and no separate score for Task 1 and Task 2. Both tasks feed one Writing level, and from level 3 upward that level equals the same CLB level.

That structure explains most of what confuses people about Writing results. You cannot work out your level by counting anything, one weak task cannot be hidden behind a strong one, and a single mistake will not sink you, because several people score you independently.

#How is CELPIP Writing scored?

By human raters, not by a computer. CELPIP describes the process in its published test-results material, and four details matter to you as a test taker.

The CELPIP Writing rating process, in four facts
Who scores it
Qualified raters trained to apply the same scoring rubrics, with ongoing training and regular monitoring.
How many
Each Writing performance is rated by four to six writing raters. Speaking uses three to five.
How they work
Tests are randomly assigned by an online system, your identity is not shown, and each rater works without seeing any other rater’s score.
What they score
Both tasks together, across four rated dimensions, each divided into five performance levels.

The phrase to hold on to is tangible evidence. Raters do not award a level for a general impression. They assign a level in each dimension by finding evidence in your response that matches the written descriptors for that level. That is why a vague, safe answer often scores lower than learners expect: there is less in it for a rater to point at.

#What are the four rated dimensions?

CELPIP publishes the four dimensions and the factors inside each one. These are the actual categories your response is judged on.

The four rated dimensions, and what sits inside each
Content / Coherence
Number of ideas, quality of ideas, organization of ideas, and examples and supporting details.
Vocabulary
Word choice, suitable use of words and phrases, range of words and phrases, and precision and accuracy.
Readability
Format and paragraphing, connectors and transitions, grammar and sentence structure, and spelling and punctuation.
Task Fulfillment
Relevance, completeness, tone, and word count.

Two things are worth noticing in that list.

Tone and word count sit under Task Fulfillment, not under style. Writing a warm, chatty email to a city official is not a small stylistic slip. It is a Task Fulfillment problem, in the same category as missing a required point.

Spelling and punctuation sit under Readability, alongside paragraphing and connectors. So Readability is not only about elegant sentences. It covers the small mechanical accuracy that runs through every line you write.

#Who actually scores your writing?

Real people, with published minimum qualifications. CELPIP lists them, and they are stricter than most test takers assume.

Requirement What CELPIP asks for
English proficiency A native speaker of English, or a non-native speaker at CLB 11 or 12
Education A minimum of an undergraduate degree
Teaching and assessment An ESL teaching certification recognised by TESL Canada, or graduate training in language education or linguistics, plus at least three years of relevant experience
Residency Resident in Canada at the time of scoring

CELPIP also uses rater agreement statistics to check rating quality: a rater is judged to agree with the others scoring the same test taker when their rating is close enough to reach consensus.

The practical takeaway is reassuring. Your response is not being read by one tired marker on a Friday afternoon. It is being read by several qualified raters who cannot see each other's decisions, and whose consistency is measured.

#What happens if the raters disagree?

A benchmark rater is brought in automatically. When the ratings for one performance are complete, they are inspected for agreement, and if they disagree the system assigns an additional rater. Benchmark raters are experienced raters who have shown consistent accuracy and reliability, and they assess the performance without knowing the ratings already given.

This is the part of CELPIP Writing scoring that almost nobody explains, and it should change how you think about a borderline answer. A response that sits between two levels does not get rounded down by whoever happened to open it first. Disagreement triggers more scrutiny, not less.

#Why is there no raw score or percentage for Writing?

Because a raw count would not mean the same thing on two different test forms. CELPIP does not report raw scores for any component, and explains why: test forms are built to the same guidelines but can still vary slightly in difficulty, so a raw score of 30 would not carry the same meaning across forms. Scores are corrected for those differences through score equating, and then reported as a level.

For Writing the point goes further, because there is nothing countable to begin with. There is no set of right answers to total. There is a performance, and four dimensions of judgement applied to it by several people.

Try a Task 1 email question

#Where do learners actually lose marks?

Readability, more than any other dimension. We looked at more than 70,600 anonymised, aggregated Writing responses from over 4,500 learners on HelloCelpip and grouped every issue our evaluator flagged by the dimension it belongs to. About 35% of everything flagged is a Readability issue, and about 27% is Vocabulary. The remaining two dimensions sit behind those, close enough together that ranking them would be reading more into the numbers than they support.

Readability about 35% Vocabulary about 27% Content / Coherence and Task Fulfillment the rest Share of issues flagged in Writing responses on HelloCelpip. Anonymised and aggregated.

So roughly six in every ten issues are about how cleanly you write and which words you choose, rather than about what you actually say.

And the split barely moves between the two tasks. Task 1 emails and Task 2 survey responses produce almost identical proportions, dimension for dimension. That is more useful than it first sounds: there is no separate scoring strategy for Task 2. The same four things cost marks in both, in the same order, so anything you fix carries across.

None of this means content does not matter. It means content failures are a different kind of problem, and they show up in a different place. Our Task 1 and Task 2 guides each cover the one requirement learners most often miss on that specific task, with the numbers behind it:

#How does your Writing level map to CLB?

Directly, from level 3 upward. A CELPIP Writing 9 is CLB 9, and a Writing 7 is CLB 7. This is the official chart, and it is the same scale across all four skills.

CELPIP Writing level CLB level CEFR
12 12 C2
11 11 C1
10 10 C1
9 9 B2
8 8 B2
7 7 B2
6 6 B1
5 5 B1
4 4 A2
3 3 A2
2 1 to 2 A1

Below that, CELPIP reports 1 and 0 as insufficient information to assess, and NA where a component was not taken. Immigration programmes set their requirement per ability rather than as an average, so your Writing level has to clear the bar on its own. The score to CLB calculator will show you which of your four abilities is holding you back.

#Can you get your Writing re-evaluated?

Yes, within six months of your test date, through your CELPIP account. A few conditions are worth knowing before you pay:

How CELPIP re-evaluation works
  • You can request a re-evaluation of some or all components, within six months of your test date.
  • You pay a fee at the time of the request, and it depends on which components you choose.
  • If your level changes for a component that was re-evaluated, the fee for that component is refunded.
  • There is a limit of one re-evaluation per component, and requests cannot be cancelled once submitted.
  • Results usually arrive about one to two weeks after you apply and pay.

Writing and Speaking are the components where a re-evaluation makes most sense, because they are human-rated. CELPIP notes that re-evaluating Listening or Reading is unlikely to change anything, since those are computer rated.

#How to use this when you practise

The useful move is to stop practising Writing as one general skill and start practising against the dimension that is costing you. If most of what goes wrong is grammar, punctuation, paragraphing and connectors, then writing more responses without checking them will not fix it.

A simple way to work: write a Task 1 email or a Task 2 survey response under time, read the report across the same four rated dimensions, and fix only the dimension with the most flagged issues before writing the next one. Then check Progress after a few attempts to see whether that dimension is actually moving.

Next Step
See which of the four dimensions is costing you
Write a timed Task 1 or Task 2 response on HelloCelpip and get a report across Content/Coherence, Vocabulary, Readability and Task Fulfillment, with the specific lines that were flagged and a corrected version.
Both Writing tasks feed the same Writing level, so it is worth practising the weaker one.
Frequently asked questions
How is CELPIP Writing scored?
By trained human raters against four rated dimensions: Content/Coherence, Vocabulary, Readability, and Task Fulfillment. Each dimension has five performance levels, and raters assign a level by finding evidence in your response that matches the descriptors. Each Writing performance is rated by four to six raters working independently.
Is Task 1 or Task 2 worth more?
CELPIP reports one Writing level for both tasks together rather than a score for each, so a weak task is not something a strong one can quietly cover for. Our own data also shows the two tasks lose marks in almost identical proportions across the four dimensions.
What is a good CELPIP Writing score?
It depends entirely on why you are taking the test. Express Entry programmes set a minimum per ability rather than an overall average, so the level you need for Writing is set by your programme. Check the current requirement on canada.ca before you plan around a number.
Why did I get a lower Writing level than I expected?
The most common reason is that marks are lost somewhere other than ideas. Across more than 70,600 Writing responses on HelloCelpip, about 35% of what our evaluator flags is a Readability issue and about 27% is Vocabulary, so grammar, punctuation, paragraphing and word choice account for roughly six in ten issues even in responses that answer the question fully.
Can I ask for my Writing to be re-marked?
Yes. You can request a re-evaluation of some or all components within six months of your test date through your CELPIP account. You pay a fee up front, and it is refunded for any component whose level changes. There is a limit of one re-evaluation per component.
How long does it take to get Writing results?
CELPIP scores are available in your CELPIP account in two to four business days after your test date, and you can download your Official Score Report as a PDF. Scores stay viewable in your account for two years from your test date.
Does spelling matter in CELPIP Writing?
Yes. Spelling and punctuation sit inside the Readability dimension, together with paragraphing, connectors, grammar and sentence structure. That is the dimension where our evaluator flags the largest share of issues, so it is worth proofreading rather than spending the last minutes adding another idea.
Official sources

Confirm the four rated dimensions, the rating procedure, rater qualifications and re-evaluation rules on the official CELPIP Results page, and the level to CLB mapping on the CELPIP Score Comparison Chart. Immigration requirements are set per ability and change, so check them on canada.ca. Your HelloCelpip level uses the same 0 to 12 CELPIP scale, so you can track real progress between now and test day. Your official score comes from Paragon Testing Enterprises on test day. HelloCelpip is an independent study resource and is not affiliated with CELPIP or Paragon Testing Enterprises.

Official sources:

Last updated: August 8, 2026