Skip to content
PTE

PTE mock test vs practice questions: they are not measuring the same thing

Aman Batth · 13 min read · · Updated

Both hand you a number, so both feel like preparation. One measures whether you can do a task type. The other measures whether you still can at minute seventy, under a block clock, with the previous task still in your head. Most plateaus are made of doing the cheap one and counting it.

Almost every page on this question answers it as a balance: do some of both, here is a four-phase plan, neither is enough on its own. That is true and it is not useful, because it never says what each one measures. Without that, "do both" is advice you cannot act on — you have no way to tell which one your next free hour should go to.

Here is the version that decides it.

A practice question measures whether you can do a task type. A full mock measures whether you do it under a block clock, seventy minutes in, having just done something else. Those are different properties of the same candidate, and a candidate can have the first and not the second. That gap is what a plateau usually is.

There is a second thing only a full paper can show you, and on PTE it is the more important one. We build the engine that scores speaking and writing answers on Hilingo, so the rest of this is written from the marking side.

Why the plateau happens: one of them is cheap

A single question returns a score in seconds. It fits in a commute. It tells you something encouraging most of the time, because you chose the task type and you were fresh.

A full PTE Academic paper costs you the better part of an afternoon. From Pearson's own published format pages, the three parts carry these limits:

PartContainsPublished time limit
1Speaking & Writing76–84 minutes
2Reading23–30 minutes
3Listening31–39 minutes

Add the ranges and a paper runs from about two hours ten to about two hours thirty-five, before the check-in and the microphone test. Pearson describes the test as around two hours; the published part limits add to a little more than that, which is worth knowing before you plan an evening around one.

So the cheap instrument runs dozens of times a week and the expensive one runs rarely, and study time flows to whichever one returns a number fastest. That is a perfectly rational response to a feedback loop, and it is why people arrive on test day fluent in every task type and short of their target anyway.

What a full paper measures that a question cannot, mechanically

Four things, and none of them is motivational.

1. The block clock. Reading is a single limit over the whole part — there is no per-question timer inside it. That means pacing is a property of the block, not the item. In practice mode there is no question twelve waiting while you sit on question three, so nothing can ever fail you for overspending. The failure mode does not exist in the instrument. Sitting a reading section is the only way to find out that Reorder Paragraph eats four minutes of a budget that had thirty in it.

2. Carry-over. Your Read Aloud at minute four and your Read Aloud at minute seventy are not the same performance, because Part 1 runs 76 to 84 minutes of continuous speaking and writing and you have to arrive at the end of it. Practice always samples you fresh. That is the definition of stamina, and it is not trainable by an instrument that cannot observe it.

3. One pass at the audio. Pearson's listening format page states it plainly: "You hear each audio or video clip once." In practice mode most people replay. A replayed clip is no longer a listening item — it is a reading-of-your-own-memory item, and it scores far better than the thing it is standing in for.

4. The four scores. This is the PTE-specific one, and neither of the pages that usually rank for this question mentions it at all.

The part that only exists on a whole paper

Most PTE questions are scored for two communicative skills at once. Pearson calls them integrated skills questions and states that the score on such a question contributes to both skills. It also states that the overall score is not an average of the four.

Follow that through and something falls out: a per-question score and a skill score are not the same kind of object. Summarize Written Text produces marks in reading and writing. Repeat Sentence produces marks in listening and speaking. Your listening score is therefore partly manufactured inside the speaking section, and there is no sequence of single questions that will ever show you that, because the thing you are looking at only exists once a whole paper has been scored together.

The practical version: a candidate whose listening is capping their profile can drill listening-section items for a month and not discover that Repeat Sentence, sitting in Part 1, is where the damage is. The full task-to-skill map — with the counts, and with the rows where PTE Academic and PTE Core disagree — is in PTE marks distribution. Read that before you decide what "practising your weak skill" even means.

Worth saying here too, because neither reference distinguishes them: PTE Academic and PTE Core do not have the same task list or the same task-to-skill map. Read Aloud feeds reading on Core and does not on Academic. Advice written for one can send a candidate for the other to the wrong question type entirely. If you are not certain which you are sitting, that decision comes first.

From the marking side: why a single answer is a weak estimate

Pearson does not publish how its engine works, and nobody outside Pearson can tell you. What we can describe is the problem any system marking recorded responses has to solve, because we run one.

A Describe Image response gives a scorer 25 seconds of your preparation and 40 seconds of audio. A speaking score on a full paper is computed from every speaking-feeding item in that paper. An estimate built on more evidence of the same kind is more stable than an estimate built on less — that is a property of estimation, not a claim about anyone's algorithm. So a per-item score is a good signal about that item and a poor signal about your speaking score, on our engine and on any other.

This is why we think practice should be read as trait feedback and a mock should be read as a profile, and why it is a mistake to treat a run of good practice scores as evidence you are ready. They are evidence you can do the task types.

What our engine returns, so you can hold it to it: per-criterion marks on writing and speaking rather than one number; a written reason on every objective answer with the evidence sentence quoted; and on PTE and TCF read-aloud, a breakdown down to individual sounds in phonetic notation. Those are our outputs against published descriptors. They are not Pearson's numbers and we do not present them as a prediction of one — the fuller version of that argument, including what official practice does better than we can, is in official PTE practice vs third-party.

The third instrument nobody mentions: section tests

The two references treat this as a binary, and it is not. A section test is a section of a full paper, offered on its own — the same material presented a second way, not extra papers. We count them separately because using one is a different exercise, and you should know what you are comparing when anyone quotes you a library size, ours included.

PTE AcademicPTE Core
Full-length mocks4530
Section tests, per section4530

A section test buys you the thing single-question practice structurally cannot give you — the block clock, the item order, the pacing decisions — for thirty minutes instead of two and a half hours. It does not give you carry-over across parts, and it does not give you the four skill scores. For most of a preparation period it is the highest-value hour available, and it is the one almost nobody schedules.

What makes a mock result meaningless

A mock is only a measurement if the conditions that produced the number match the conditions you are modelling. Break one and you have not sat a mock — you have done a long practice session and given it a misleading name.

The useful way to see this is by which number each shortcut destroys:

What you didWhat the score no longer tells you
Paused mid-sectionYour pacing. The block clock was the one thing this instrument had that practice does not.
Replayed a listening clipListening, and every integrated item feeding it — so the corruption spreads into a second column.
Retook a task you fluffedRecovery. The real cost of one bad Read Aloud is the next three items, not that item.
Used laptop speakers or no headsetSpeaking. Whatever scores the audio only has the audio — room noise and mic distance are in the file.
Looked a word upReading and vocabulary, and worse, it points the diagnosis at the wrong skill.
Split it across two eveningsCarry-over and stamina, which is most of the reason to sit a full paper at all.

An honest note about our own product, because it cuts against us: Hilingo lets you save an attempt and resume it. That is the right behaviour for a practice session you got interrupted in. It is the wrong behaviour for a paper whose score you intend to read as a measurement, and the software cannot tell the difference — only you can. If you paused, treat the result as practice and do not put it on your progress chart.

The same applies to the thing candidates do most: retaking a speaking task because the recording "did not capture how you really speak". It captured exactly what any scorer will get. That is the point of the exercise.

A sequence, and the reasoning for it

We do not publish success rates or average point gains, here or anywhere, because we could not show you how they were calculated. So this is a rule derived from what each instrument measures, not from a statistic.

StageWhat to sitWhy this and not the other
Before you plan anythingOne full mock, full conditionsYou cannot choose what to practise until you know which of the four scores is capping you, and self-assessment will not tell you. The score itself does not matter here.
The working phaseSingle-question practice on the task types that feed your capping skillThis is where a trait actually changes. Repetition is the mechanism, and it needs volume a full paper cannot supply.
Once a week, roughlyOne full mock, or a section test on the weak sectionRe-measure only when the practice could plausibly have changed something. Measuring more often just reprints the same ranked list.
Final two weeksTwo or three full mocks, full conditions, at your real test time of dayPacing and stamina are the last properties to arrive and the only ones that need rehearsal rather than study.

If you want a ratio: roughly one full paper to every four to six practice sessions, and the reasoning is the same in both directions. Mock more often than that and you spend hours re-confirming a weakness you have not yet worked on. Mock less often than that and you are steering on stale information — the leak you fixed three weeks ago may have been replaced by a pacing problem you cannot see from inside a practice queue.

Two rules that do not flex:

  • Never sit a mock to learn a task type. It is the most expensive possible way to find out you do not know one, and it wastes a measurement.
  • Sit at least one complete, uninterrupted, full-conditions paper before test day. The first time you meet an 80-minute Part 1 should not be the time it counts.

And if you are planning to lean on a memorised structure to survive the clock, read why PTE templates do not work first — there is a Content-zero gate in the scoring order that makes that strategy worse than doing nothing.

Start with the measurement

The argument above only pays off if the mock comes back with something you can act on. A number by itself puts you straight back into guessing which task type to drill.

Hilingo gives you one full scored mock free, with no card. Full length, real section timers, our own engine on every speaking and writing answer — per criterion rather than one figure, so you can see which dimension is holding a skill down. Objective questions come back with the reason the answer was wrong and the evidence sentence quoted. Read Aloud answers come back broken down to individual sounds in phonetic notation. The AI teacher works from your own result rather than a generic syllabus, the study centre keys itself to whichever skill came back weakest, and any question, passage or transcript can be shown in 50 languages if the English of the prompt is getting in the way of the English being tested.

Then do the volume on single-question practice, which sits alongside the mocks rather than behind a separate product.

Take a free scored mock

Frequently asked questions

What is the difference between a PTE mock test and practice questions?

A practice question measures whether you can do one task type, in isolation, while fresh. A full mock measures whether you can still do it under the section clock, in the exam's own order, after the tasks that came before it — and it is the only thing that produces the four communicative skill scores, because most PTE questions feed two skills at once and that map only resolves across a whole paper.

How many PTE mock tests should I take before the exam?

There is no researched number, and treat anyone quoting one with suspicion. A workable rule from mechanism: roughly one full paper for every four to six practice sessions, so that each mock measures a version of you that the practice could plausibly have changed. Then two or three full-conditions papers in the final fortnight, when pacing and stamina are the remaining variables. The floor is one complete uninterrupted paper before test day.

Are practice questions enough to pass PTE?

They are enough to learn the task types and not enough to predict your result. Practice cannot observe pacing across a timed block, cannot observe fatigue across an 80-minute Part 1, and cannot produce a skill score, because a PTE skill score is assembled from questions spread across the whole paper rather than from any single item.

Does it matter if I pause a PTE mock test?

Yes, and it removes the specific thing the mock was for. Pearson publishes one time limit per part — 76 to 84 minutes for Speaking and Writing, 23 to 30 for Reading, 31 to 39 for Listening — and pacing inside a block is exactly what single-question practice cannot teach you. A paused paper is a long practice session. It is still useful; it just should not go on your progress chart as a score.

Can I take a PTE mock test in two sittings?

You can, but the result stops being comparable to a real one. PTE is sat in a single session covering all four skills, so splitting a mock removes carry-over and fatigue — the two properties that separate a mock from practice in the first place. If your evening cannot hold a full paper, sit a section test instead: you keep the block clock, and you are not pretending the number means something it does not.

Do I need headphones for a PTE mock test?

For the result to mean anything, yes. Whatever system marks your speaking has only the audio file, so background noise, microphone distance and speaker bleed are not conditions around the answer — they are part of the answer. Practising on laptop speakers trains you for an input the test centre will not give you, and it makes your speaking score at home unrelated to your speaking score on the day.

Practise with the engine this article describes.
One full scored mock, free. No card.
Take a free mock

All articles