Internal prototype every figure here is example data. The marking numbers (90 / 63 / 29) are unverified against NAATI and must not be relied on.
Prototype NAATI Ready start at the dashboard — every screen is reachable by clicking
NAATI Ready Hi, Jince! JM
Exam · 21 Nov 2026FR-AC-003 On track 76 days
Started 8 Jul 44% through prep window
74%
READY
Almost exam-ready!

Based on last 12 sessions. The band reflects how much your recent scores vary — not a predicted NAATI result.FR-RS-001

Scoring criteriaFR-RS-006
NAATI topicsFR-RS-007 7 / 12 practised
View full report →
7-day streak — keep it going!
Healthcare Dialogue Session 6 of 6 0:42
Source dialogue
English → Hindi
Doctor

The patient has been experiencing persistent lower back pain for the past three weeks. It appears to be muscular in origin, though we need to rule out any spinal involvement.

Nurse

Should we schedule an MRI, or would X-rays be sufficient at this stage?

Doctor

Let's start with an X-ray. If the results are inconclusive, we'll escalate to an MRI. Please ensure the patient is aware of the preparation requirements.

1:14 1:52
Your turn — interpret the full dialogue into Hindi
Recording
0:42
of 3:00 max
Stop & Submit
or

Scoring: Accuracy · Fluency · Omissions · Additions · Delivery

Speak clearly and interpret the entire dialogue in sequence.

This flow does not match the real CCL test

The real test plays the dialogue in segments of 35 words or less, with a chime after each, and the candidate interprets segment by segment — not the whole dialogue in one 3-minute take. See Conflicts → C1.

/Segment practice
Healthcare · Dialogue 3 Segment 4 of 9
English → Mandarin · 31 words (cap 35)

The doctor says your blood pressure is quite high, so she'd like you to come back next Tuesday for another check before she prescribes anything.

Replays are unlimited in practice. The mock test restricts them.

0:00

Press to record your interpretation. Recording never starts on its own.FR-DP-005

This segment
TopicHealthcare
DirectionEN → ZH
Words31
Propositions6
Recording cap2:00
Mic is live and capturing

The level meter is the honest signal here. Silent capture is the worst failure in this loop, so a near-zero take is rejected before scoring rather than marked as a total omission.FR-DP-010

/Segment feedback
Healthcare · Dialogue 3 · Segment 4 How that segment scoredFR-DP-011
Source · English

The doctor says your blood pressure is quite high, so she'd like you to come back next Tuesday for another check before she prescribes anything.

Your rendition · transcribed

医生说您的血压很高,所以她想让您下周二再来检查一下,可能需要吃药

STT confidence 94% · Mandarin · Tier A
Reference rendition

医生说您的血压偏高,所以她希望您下周二再来复查一次,然后再决定是否开药。

One acceptable rendering of six. Matching is on meaning, not wording.FR-SC-007
Proposition breakdownFR-SC-004 Tap any row for the reasoning
Segment score
6.0/ 9

Four of six propositions conveyed. One degree distortion, one omission, one substantive addition.

By criterion
Accuracy (distortions)1 issue
Omissions1
Additions1
FluencyGood
DeliveryGood

Accuracy is scoped to distortions only, so it doesn't double-count omissions and additions.Q7

These five criteria are our practice rubric, not NAATI's marking scheme. NAATI deducts from a fixed total, weighted by how much an error impairs communication.BR-SC-005

Delivery · measured, not judgedFR-SC-010
Speech rate118 wpm
Longest pause2.3 s
Start latency1.8 s
Filler ratio4.2%
False starts1
Finished in capYes

Computed arithmetically from your audio and word timings — no model involved, so these are the most reproducible numbers on this page.

/Review queue
Housing · Dialogue 1 · Segment 2 We couldn't score this confidentlyFR-SC-011
Punjabi · Tier C
Score withheld, not zero

Speech recognition returned 41% mean confidence on this take. That is below the threshold for Punjabi, so we are not going to pretend a score derived from it is meaningful. Correct the transcript below and we'll score it properly.FR-SC-012

What we heard · low-confidence tokens marked

ਡਾਕਟਰ ਕਹਿੰਦਾ ਹੈ ਕਿ ਤੁਹਾਡਾ ਬਲਡ ਪਰੈਸ ਕੁਜ ਵਧਿਆ ਹੋਇਆ ਹੈ

Mean confidence 41% · 3 tokens below 50%
Correct it
Why this happens

Speech recognition for Punjabi is materially weaker than for Mandarin or Spanish. Rather than hide that behind a confident-looking number, Punjabi is marked Tier C: scores are advisory, and they don't move your readiness.FR-CP-006

Counts toward readinessNo
Flagged for content reviewYes
Threshold for this language65%
This take41%

A cluster of these in one topic usually means our content or language pack is wrong — not that you are.FR-SC-013

08:42 Dialogue 1 of 2 · Segment 5 of 11 Replays left BR-MT-004 Exam conditions
Mandarin → English · 28 words

我上个月交了押金,可是房东到现在还没有把收据给我,我该怎么办?

0:00

No feedback until the test ends. You cannot pause, re-record or go back.FR-MT-003

What's different here
Replay source audioLimited
FeedbackWithheld
Pause / resumeNo
Re-recordNo
Readiness weightHigher
Real test conditionsFR-MT-010
HeadphonesNot allowed
AudioSpeakers + mic
Start after chimeWithin 5 s
Segment length≤ 35 words

NAATI banned headphones of any kind in mid-2024. Practising on them rehearses acoustics you won't have on the day.

Reloading is safe

Completed segments survive a page reload and the timer keeps running. Abandoning records an incomplete attempt rather than discarding your work — but incomplete attempts don't count toward readiness.BR-MT-003

/Mock test result
Mock test · 6 Sep 2026 · 19:24 Your resultFR-MT-006
68/ 90

Above the practice pass mark of 63, and above the minimum on both dialogues. This is an estimate from our own marking model — not an official or predicted NAATI result.FR-MT-007

Thresholds
By criterion
3 of 22 segments scored with low confidence

Those segments carry up to ±4 marks of uncertainty, so treat 68 as roughly 64–72. We disclose this rather than present a clean total built on shaky parts.FR-MT-009

⚠ Every marking number here is UNKNOWN

A 49-agent research sweep established the CCL's structure — 2 dialogues, segments of 35 words or less, a 5-second start window — but verified no marking number at all, because naati.com.au is blocked by our egress proxy and no primary source was read. Total marks, pass mark and the per-dialogue minimum are uninvestigated, not disputed. The circulating 90 / 63 / 29 are shown so they can be checked, not because they are right. One config object holds them all. Do not ship scoring against these. Q1 remains blocking.

Readiness moved
71 74

A mock counts for more than practice, because practice conditions are easier.FR-RS-004

Calibration
Readiness before71
Implied mark64
Actual mark68
Error−4

Every mock is logged against the readiness that preceded it. That comparison is the only thing that keeps the headline number honest.FR-RS-011

Weakest segments
D2 · Segment 73.0 / 9
D1 · Segment 44.5 / 9
D2 · Segment 25.0 / 9

You dropped temporal qualifiers in 4 of 22 segments — the most repeated single pattern in this attempt.

/Design review
Design review Where the Figma and the verified test format disagree

Every screen in this prototype is built from your Figma. Six things in it conflicted with what the CCL research established, or with itself. C6 is now resolved — the design system follows the marketing site. None of the rest are hard to fix now; all are expensive to fix after build.

C1
The session flow interprets the whole dialogue at once

Session — Dialogue Recording gives a single 3:00 take for a three-turn dialogue. The real CCL plays segments of 35 words or less, sounds a chime, and the candidate interprets that segment before the next plays. This is the single biggest gap: it changes the memory load, the note-taking strategy, and what we can score. Practising one long take does not prepare anyone for segment-by-segment interpreting.

C2
"Accuracy — word-for-word faithfulness to the source"

From Detail — Mock Test. Word-for-word transfer is what a weak interpreter does; NAATI assesses whether meaning survives, and a literal rendering that mangles idiom is penalised. Teaching candidates to aim for word-for-word would actively lower their scores.

C3
Criteria weights 30 / 25 / 20 / 15 / 10

From Detail — Mock Test. No NAATI source publishes any such weighting — the research looked and found only a single unsourced prep site claiming a split. Presenting invented weights as "what's scored in the exam" is the exact failure BR-SC-005 exists to prevent.

C4
"45 min" mock test duration

Test duration came back UNKNOWN. The only duration-adjacent fact established is that long pauses may cost marks, particularly past a 20-minute total recording. 45 minutes has no basis found.

C5
Two different languages across the design

The session screens say English → Hindi; Vocabulary shows Persian/Dari (ترتیبی، مترجم شفاهی). Both are plausible CCL languages, but a single candidate has exactly one. Pick one for the prototype, and make sure RTL is tested — Persian, Dari, Arabic and Urdu all need it.

C6
Three different accent colours — resolved

The indigo modals were right and the green dashboard was the outlier. The marketing site is a slate/indigo system built on Geist — its stylesheet literally says “keep the slate/indigo system”. Green was never the brand: emerald appears there only as a success colour. This prototype is now rebuilt on the marketing tokens, and green has been demoted to what it should mean — passed, conveyed, improved.

Prototype only — no audio is captured, no data leaves this page. Every number is example content from the Figma. 11 Figma frames + 5 v1 scoring screens