STANAG 6001 · SLP Levels 2 & 3

One score is a guess.
Four is a profile.

An example profile

Gating level

Listening

2

Speaking

2

Reading

3

Writing

2

Never an average — the lowest digit, here Listening and Speaking at 2, is what gates the posting. SLP Command measures each skill on its own, the way a real sitting rates it.

So what you practise next follows from your own evidence, not from a single number that hides which skill is actually holding you back.

Free plan, no card required · iOS app coming to the App Store

Independent educational platform. SLP Command is not affiliated with NATO, any Ministry of Defence, or any official examining body. AI-generated feedback is indicative guidance — not an official SLP / STANAG 6001 assessment.

Four skills, one learner

Reading, Listening, Writing, Speaking

Not four identical exercise banks. Each skill trains against the specific criteria a rater actually applies to it — and each shows you a different kind of evidence.

Reading Comprehension · Evidence

Practice adapts to your estimated level and to specific weak sub-skills, not just overall difficulty. A spaced-review scheduler resurfaces what you got wrong until it sticks.

Listening One pass · Exam pace

One play, exam timing — the same condition the real sitting imposes. Catch the key data the first time, because that is the only time it is offered.

Writing Examiner · Precision

Rated on task achievement, content & organisation and language precision — separately. An improved version comes back beside your own, so the gap is visible, not just described.

Speaking Voice · Performance

Transcribed and rated across five criteria — fluency, grammar, vocabulary, coherence, task achievement — every time you speak.

Platform

Everything you need to reach SLP 2 or SLP 3

One platform, both target levels. Every exercise, every exam simulation and every AI evaluation is built to the level you are training for — SLP 2 or SLP 3 — with its own question formats, time limits and rating criteria, not a single difficulty curve with the numbers moved.

4
Skills — Reading, Listening, Writing, Speaking
2
Target levels — SLP 2 and SLP 3
11
Structured Academy topics per skill
5
AI Speaking dimensions scored per response
01

Today’s session

Open the app and the first thing you see is a plan, not a menu: which skills, in which order, for how many minutes, and the reason for each block. It also says what to skip today. Two productive skills are never placed back to back, and no single block takes more than half the time you have.

02

Confidence you can read

Every estimate carries how much the platform actually trusts it, shown as a four-step ladder rather than a word: where you are, what each step means, and what specifically buys the next one.

03

Your timeline

Not a chart — a chronology. Every point at which something changed, what caused it, and what happened while you were away, reconstructed from your measured history.

04

Practice Mode

Exercises adapt in real time to your estimated level and to weak sub-skills specifically — not just overall difficulty. A spaced-review scheduler resurfaces items you got wrong until they stick.

05

Exam Simulation

Timed, full-length mock exams for Reading, Listening, Speaking and Writing, built to the SLP 2 and SLP 3 formats separately — question count, time limits and scoring bands match the level you select.

06

Academy

11 structured topics per skill — grammar and strategy references, worked examples with step-by-step reasoning, and a short quiz per topic with immediate, explained feedback.

07

Intelligence

Per-skill weakness profiles, a readiness assessment against your target level, and the sub-skills your recent results actually separate you on. Every figure is traceable to the attempts that produced it — never a rolling average nobody can audit.

08

AI evaluation, with its reasoning

Speaking is transcribed and rated across five criteria — fluency, grammar, vocabulary, coherence and task achievement. Writing is rated on task achievement, content & organisation and language precision, with an improved version returned beside your own. Writing is also assessed on whether it answers the task that was set, judged separately from how well it is written — good English on the wrong subject is not credited as good English on the right one.

09

Mastery & progression

Accuracy trends over 7, 30 and 90 days per skill, with a mastered / developing / needs-work state. A decline is reported as a decline: skills decay when they are not used, and a platform that hides that is not measuring anything.

10

Adaptive Coach

A short, prioritised list of what to work on next, derived from your own weakness profile rather than from a syllabus. It changes when your results change, and it says what it is responding to.

11

Cross-device sync

Your account, progress, scores and Academy status sync automatically — pick up training on any device without losing history.

Method

Six rules the platform will not break

Most exam apps are content libraries with a score on top. The difference here is not the volume of material — it is what the product is willing to claim, and what it refuses to say when it has not measured it.

01

Nothing is asserted without measurement

If there is not enough evidence about a sub-skill, the product says so and shows you how much is missing. It does not fill the gap with a plausible number.

02

Every recommendation names its evidence

“Work on this next” always arrives with the results that produced it — how many observations, from which skills, and what changed.

03

Strengths do not cancel weaknesses

STANAG 6001 levels are not averages. The level you are shown is bounded by the evidence you have actually produced at that level, never by a mean across dimensions.

04

Confidence is reported, not hidden

Every estimate carries how certain it is and what would make it more certain. Uncertainty is information, not a flaw to be smoothed over.

05

A decline is a fact, not a judgement

Language skills fade when they are not used. Recovering something you had is faster than learning it, and it is worth doing before it slips further.

06

Answering the task is part of being right

If a response addresses something other than what was asked, the language quality cannot stand in for the missing task — the two are judged separately.

07

Grounded in the rating criteria

The constructs the platform measures are the ones the exam rates — the Speaking criteria, the Reading and Listening sub-skills, the Writing dimensions.

Intelligence

Four skills.
One learner.

The point of a single model is that a Speaking problem caused by a Writing gap can be named as such — which is impossible when each skill keeps its own opinion about the same person.

1

Observation

Every rated attempt — a reading question, a listening item, a written submission, a spoken response — becomes an observation carrying what was asked, at which level, and how it went. Nothing is recorded that was not measured.

2

Belief update

Observations update an estimate of your ability together with how uncertain that estimate is. Items near your current level carry the most information; an item far above or below it moves the estimate less, because it tells us less.

3

Ageing

Evidence loses weight over time rather than being deleted. A correct answer from eight months ago is not the same claim about you as one from last week, and the model treats it accordingly instead of averaging both as if they were.

4

Reportable level

The number you are shown is capped by what you have actually demonstrated at that level, and it is accompanied by its confidence and by what would raise it. Where the evidence does not reach, the product says the evidence does not reach.

Validation before promotion. A new assessment engine built to these rules runs in parallel with the one currently in service, observing the same activity and recording what it would have concluded, without affecting anything you see. It replaces the current estimator only when a published set of criteria is met — a minimum volume of real evidence, a minimum period of continuous observation, zero invariant violations when real learner histories are replayed through it, and a demonstrated ability to roll back. Passing those criteria makes it eligible for calibration against field data. It does not make it live. That decision is separate, staged, and reversible at every step.

Pricing

Free measures all four skills. Professional takes the cap off.

Same trainer. SLP 2 or SLP 3. You pay for volume, not a different exam.

Free

€0/month

Get started with real practice volume, every skill included.

  • Reading Practice — 10 sessions / week
  • Listening Practice — 10 sessions / week
  • Writing, AI-scored — 3 submissions / month
  • Speaking, AI-scored — 3 evaluations / month
  • Exam Simulation — 1 full exam / month
  • Academy — full access, all topics
  • Intelligence Dashboard — full access
Start free

No card required.

On the web you pay with a card (Stripe). The iOS app is not for sale on the App Store yet. Independent trainer — not an official STANAG 6001 assessment.

How it works

From first session to exam day

1

Create your account

Sign up with your email. Your progress syncs across devices automatically.

2

Find your starting position

Set your target level — SLP 2 or SLP 3 — and start training in any skill. The first sessions are as much measurement as practice: they establish where you actually are, which is the only thing that makes the rest of the plan meaningful.

3

Work the weakness, not the syllabus

Intelligence updates after every session and the Adaptive Coach names one thing to do next, with the results behind it. Practice adapts to the specific sub-skills you are losing marks on, and a spaced-review scheduler brings back what you got wrong until it holds.

4

Test it under exam conditions

Full-length, timed simulations in the format of your target level. Evidence from a timed exam counts for more than practice, because it was produced under the conditions the real assessment imposes.

5

Know whether you are ready

Readiness is reported against your target level with the confidence attached, and with what is still missing named explicitly. Walking into the exam knowing which two things are weakest is worth more than a number that says you are fine.

Roadmap

What has shipped, and what has not

Listed here because a roadmap that only appears once something ships is a marketing page, not a roadmap. This section is kept honest in both directions: when something ships it moves out of the list, and it says so.

Shipped since this list was written

One place to continue is now the session described under Today’s session. The story of how you got here is now your timeline. Both are in the app and included in what you pay for. They are described in Features rather than here, because that is where things that exist belong.

🔬

The certified assessment engine

The estimator described under Architecture. It is the source of truth for Listening today; the remaining skills continue on the previous method until each meets its published promotion criteria on real evidence. The rollout is per skill and reversible at any point, which is why it arrives one skill at a time rather than all at once.

🗺️

A roadmap that knows what blocks what

Not built yet. Sub-skills depend on each other: you cannot infer an argument from a text whose vocabulary you do not have. The intention is to draw those dependencies, and to never show anything as locked without naming what locks it and why working on that first is faster. Nothing of this is in the app today.

🎯

A probability of passing

Deliberately absent, and not planned. The platform can tell you what your evidence supports and how far it is from your target; it will not convert that into a percentage chance of passing a real exam, because no such number has been calibrated against real STANAG outcomes. People book exams on that figure. It is not one to estimate.

FAQ

Questions

Can I cancel anytime?

Yes. If you subscribed on the web, manage or cancel from your account. The iOS app is not for sale on the App Store yet.

Do I lose my progress if I cancel?

No. Your practice history and results are always saved to your account.

Is this billed monthly?

Yes, €9.99 billed monthly until cancelled.

Is SLP Command an official STANAG 6001 assessment?

No. SLP Command is an independent educational platform, not affiliated with NATO, any Ministry of Defence, or any official examining body. AI-generated feedback is indicative guidance to help you prepare — not an official SLP / STANAG 6001 assessment.

Which levels does it cover?

SLP Command trains SLP Level 2 and SLP Level 3 across all four skills — Reading, Listening, Writing and Speaking. Every exercise and every AI evaluation adapts to whichever level you select.

Does my progress sync across devices?

Yes. Your account, progress, scores and Academy status sync automatically, so you can pick up training on any device without losing history.

How is my estimated level calculated?

From your own rated attempts, nothing else. Each answer, submission and spoken response updates an estimate of your ability along with how uncertain that estimate is; items close to your current level carry more weight than ones far from it, and older evidence counts for less than recent evidence. The level shown is bounded by what you have actually demonstrated — it is never an average across skills, because STANAG 6001 levels are not averages.

Why does it sometimes refuse to give me a number?

Because it has not measured enough to give you an honest one. When there is too little evidence about a sub-skill, the platform says so and tells you what would resolve it, rather than showing a plausible figure derived from nothing. A number you cannot rely on is worse than no number, particularly when you are deciding whether to book an exam.

Why did my level go down?

Usually because recent results are below earlier ones, or because evidence ages. That is a measurement, not a verdict on you: language skills decay when they are not used. It is also the most actionable thing the platform can tell you, since recovering something you had is considerably faster than learning it for the first time.

There is one other reason, and it is not about you at all: we sometimes improve how the estimate is worked out. When that happens your number can move without anything about your English having changed. You are told when it does, and why — see My estimate changed and I did not do anything differently below.

My estimate changed and I did not do anything differently

Then the change is ours, not yours. An estimate is only as good as the method behind it, and when we find that a method credited progress too easily we correct it — which can move your number even though your English is exactly where it was.

The earlier method nudged your level up a little with every correct answer and down by less with every wrong one. Over enough questions that drifts upward on its own, so the figure could end up above what your answers actually supported. The current method asks a different question: how do you perform on questions at each level? That is the same question the real assessment asks.

We show you the new figure alongside the old one before anything changes, so it is never a surprise. Your history is kept as it was — we do not rewrite numbers you were previously shown — and results from before and after the change are drawn separately, because they are not directly comparable.

What does “Out of date” mean next to my level?

That we are no longer confident the estimate describes you today. It says nothing about how good you are — it is a statement about the evidence, which has simply aged.

Confidence answers one question: how far should we trust this estimate right now? So it has four settings, and each tells you something different:

  • Reliable — enough recent work to be confident.
  • Fairly reliable — supported, but a little more would make it firm.
  • Limited evidence — not enough yet. Do more.
  • Out of date — your work was solid, but it is old. One recent session is usually enough, and doing more is not what is needed.

Confidence starts to fade after about six weeks without practice in a skill, and each skill ages on its own — practising Reading every day does not keep your Speaking estimate current.

Why does the app decide what I train instead of letting me choose?

You can still open any skill you like at any time — nothing is locked. But the app opens on a plan rather than a menu, because choosing what to practise is a decision most people get wrong in the same direction: they practise what they are already good at, in long blocks of one skill.

The session is built the other way round. It interleaves rather than blocks, it puts recovery before new material when something has gone out of date, and it never places two productive skills — writing and speaking — back to back, because that is what makes a session unfinishable rather than difficult. Every block on the plan states the reason it is there, so you can overrule it knowing what you are overruling.

Why does it sometimes tell me not to train something?

Because half of what an instructor is for is saying what can wait. If a skill is current and well evidenced, another session on it today buys you very little, and the time is worth more elsewhere. The app names those skills explicitly instead of quietly leaving them off the plan.

It will never tell you to skip something that is also on today's plan. If that ever happens, it is a bug — report it.

Will you tell me my chance of passing the exam?

No, and this is deliberate. We can tell you what your evidence supports, how confident we are in it, and how far it is from your target level. Turning that into “72 % likely to pass” would require calibration against real STANAG outcomes that we do not have. People book exams and make career decisions on that number, so we do not produce one rather than produce one that looks credible.

My writing was good English but got a low level. Why?

Almost always because the response did not answer the task that was set. Writing assessments rate two things that are genuinely separate: how well you write, and whether you did what the prompt asked. A fluent, accurate, well-organised piece about something other than the question cannot be credited as though it answered the question — an examiner would not do that, and neither will we.

When this happens the feedback says so explicitly rather than leaving you to guess from the number: it tells you the language was not the problem, which part of the task was missed, and what to change. Usually nothing about your English needs to change at all.

Is the AI making up its feedback?

The evaluation is produced by AI models against fixed rating criteria, and the criteria — not the model's opinion — are what the score is expressed in. Every rating comes back with the reasoning that produced it, so you can disagree with it on specifics. It remains indicative guidance for preparation: it is not, and does not claim to be, an official STANAG 6001 rating.

How does this differ from a language academy or a course?

A course delivers a syllabus in a fixed order to everyone. This does not have a syllabus: it has a model of you, rebuilt after every session, and it uses that model to choose the next thing. It also does something a course structurally cannot — it tells you when it does not know, and it shows you the evidence behind every claim it makes about your English.

Ready when you are

Now train.

Free plan, no card required. Every skill measured on its own — so what you practise next comes from your own evidence.