Product requirements · Student side

IELTS Preparation Platform v2

A student app that carries one learner from Band 3.5 to Band 9.0: a plan built from their own mistakes, practice cut by skill and question type, speaking face to face with an AI examiner, and a mistake notebook that follows them into their inbox and their browser.

Surface
Student web + mobile web
Prototype
34 screens, clickable
Screens covered
34 · both layouts
Status
Proposal, pre-build
Written
11/09/2026
Draft for discussion — no scope has been committed
Where this document came from, and what it is not

On 06/09/2026 the brief for this product was given in full, verbatim: "Build for me an ielts preparation platform UI/UX and deploy to roadmap.flyer.vn/ielts-v2 — Personalization for students / Gamification, pride / Students can study by idivisual skill / Details onboarding / Knowledge band from 3.5, 4.5, to 9.0 / Ranking between students / Speaking facing with AI Teacher / Details questions filter / Mock Test, recently offical test / Mobile app UI/UX friendly / Email notification to remind to learn vocabulary mistake ( each mistake will be store in students notebook) / Teacher conner (will use Teado.ai) no worries at this stage / Training conner: IPA, learn pronuncation conner / Chrome extention connection (when reading new paper, can open extention to add vocabulary list)". On 11/09/2026 the follow-up was: "tạo thêm prd và cho vào luôn".

The clickable prototype was built first and this document was written after it, by an AI assistant in a single session, describing what the prototype already shows plus the rules needed to build it for real. No user research, no engineering estimate and no content licensing check sits behind it yet.

It is a proposal to argue with, not a decision. Numbers marked proposed have no baseline behind them and exist to be replaced with real ones.

01Why this, why now

IELTS preparation is bought by learners with a deadline and a number. They know the band they need and roughly when they sit the test, and almost every other decision — what to study today, which question type to drill, whether they are on track — is one they are unqualified to make alone. Most preparation products answer this with a content library and leave the sequencing to the student.

The bet in v2 is that sequencing is the product. Every screen in the prototype exists to answer one of three questions a student asks: where am I, what do I do next, and is it working. The content bank, the AI examiner and the gamification are inputs to those answers rather than features in their own right.

This is not greenfieldFLYER already ships an AI-first IELTS product, IELTS Guru at ieltsguru.ai, whose own PRD documents skill-by-skill mock tests with exam simulation, AI grading for Writing and Speaking, an AI Mentor for chat and voice, and a diamond plus subscription model. The first question this document raises is therefore not "should we build it" but what v2 reuses — content bank, grading pipeline, accounts — and what it genuinely replaces, which is the student experience: onboarding, the band ladder, the daily plan, the mistake loop and the ranking. That comparison has not been done yet and belongs in the next revision.
The second bet: mistakes are the assetA wrong answer in Reading, a mispronounced /θ/ in Speaking, a missing article in Writing and a word saved from a news article all become the same kind of object — a notebook entry with a due date. One capture mechanism feeds the daily plan, the review emails and the drills. If that loop works, retention follows it.

02Who it serves

LearnerSituationWhat they need firstWhere they break
Deadline student
primary
Test booked in 8–16 weeks, needs 6.5 or 7.0 for university admissionA plan that fits the date, and proof each week that the band is movingStudies whatever is easiest, avoids Writing and Speaking, discovers the gap in the last fortnight
Undated improverWants a stronger band, no test bookedA visible ladder and a daily habit small enough to keepLoses the streak in week two and never returns
Class studentStudying with a teacher who works in Teado.aiHomework in the same place as self-study, and one band estimate both sides trustTeacher tooling and student tooling disagree about level
Re-takerSat the test, missed by 0.5 in one skillTo spend nearly all time in the one skill that failedRestarts a general course instead of a targeted one

The prototype is drawn for the deadline student and degrades sensibly for the others: an undated learner skips the date step in onboarding and gets a plan by band rather than by week.

03What success looks like

All targets below are proposed. None has a measured baseline in this product yet, since nothing is built; the first job after launch is to replace this table with real numbers rather than to defend these.

QuestionMeasureTargetRead it as
Does onboarding land?Students finishing all 8 steps including placement70%Activation
Do they come back?Active on 4+ days in their first 1445%Habit formed
Does the band move?Median overall band change after 8 weeks of use+0.5The only outcome that matters
Do they face Speaking?Students with 3+ AI Teacher sessions in 30 days50%The avoided skill
Does the notebook loop close?Due items reviewed within 48 hours60%Mistake capture is worth the build
Do the emails earn their place?Review email open rate, and unsubscribe rate35% / <2%Reminder, not spam
Is the estimate honest?Gap between predicted band and real test result±0.5Trust in everything else
The metric that can embarrass usThe last row. A product that tells a student they are at 6.5 when they sit the test and get 5.5 has done them real harm, and they will say so publicly. Band estimates should be conservative and should show their working — which attempts, how recent, how many.

04Scope: brief to screen

Every line of the original brief, and where it is answered in the prototype. "Drawn" means the screen exists and is clickable with sample data; nothing in the prototype is wired to a real backend.

From the briefAnswered byState
Personalization for studentsOnboarding answers drive the Today plan; mock results and notebook mistakes re-weight it; profile shows performance over 7/30/90 days and ranked weaknessesDrawn
Gamification, prideStreak, XP, coins, leagues, 12 badges, earned titles beside the name, shareable result cardDrawn
Study by individual skillFour skill hubs with per-question-type mastery and a band path eachDrawn
Detailed onboarding8 steps: goal, target band, test date, level, placement, daily time, reminders, planDrawn
Knowledge band 3.5 → 9.012-rung ladder with locked, current and mastered rungs; every unit and question carries a band tagDrawn
Ranking between studentsLeague, class, target-band peers, friends, global, plus a public profileDrawn
Speaking facing with AI TeacherExaminer picker, live call with cue card and transcript, feedback by the four criteriaDrawn
Detailed questions filterSkill, part, question type, band, topic, source, status, timeDrawn
Mock test, recent official testOfficial recent tests by month, Cambridge full mocks, per-skill mocks, in-test screen, result reportDrawn · content source unresolved
Mobile app UI/UX friendlyBottom tab bar under 720px, filters as sheets, phone frame toggle in the prototype barDrawn
Email reminder for vocabulary mistakesNotebook with spaced repetition, plus a rendered review email with an inline quizDrawn
Teacher corner (Teado.ai)Placeholder plus class-code join; deliberately out of scope for nowPhase 2
Training corner: IPA, pronunciation44-sound chart, per-sound record and compare, minimal pairs, word stressDrawn
Chrome extension for vocabularyPopup over an article, saved lists, sync into notebook and emailsDrawn · extension itself not built

Out of scope for v2

05The band ladder, 3.5 to 9.0

The ladder is the spine of the product. It is twelve rungs — 3.5, 4.0, 4.5, 5.0, 5.5, 6.0, 6.5, 7.0, 7.5, 8.0, 8.5, 9.0 — and it does three jobs: it tags content, it gates content, and it reports progress.

Rules

What a rung contains

RungListeningReadingWritingSpeaking
4.5Short factual exchanges, form completionShort texts, short answerSentence-level accuracy, 150 wordsAnswers of one or two sentences
5.5Part 2 monologues, map labellingSummary completion, TFNG at 60%Paragraph structure, linking, 250 wordsFluent on familiar topics, short long-turn
6.5Part 3 discussion, paraphrase trackingMatching headings, YNNG reliablyClear position, developed body paragraphsTwo-minute long turn without long pauses
7.5Part 4 lectures, note completion at speedDense academic text, inferenceCollocation range, cohesion without over-linkingAbstract opinion, hedging, natural stress
Why 3.5 is the floorBelow 3.5 the useful intervention is general English, not IELTS technique. Students who place under 3.5 should be told that plainly and pointed at a foundation path rather than sold drills they cannot use.

06Personalisation and the daily plan

Onboarding collects five inputs, and each one changes something visible. If an answer changes nothing, the question should be cut.

InputCollected inWhat it changes
Goal (study abroad, work, school, self)Step 2Academic or General module; topic weighting; which Task 1 type is taught first
Target band, overall and per skillStep 3The top of the ladder; what counts as "on track"; difficulty ceiling in the question bank
Test dateStep 4Plan length, weekly intensity, when mocks are scheduled
Starting level (placement or self-report)Steps 5–6The starting rung per skill
Daily time and reminder settingsStep 7Size of the daily plan; push and email cadence

How the daily plan is assembled

The plan is rebuilt each morning. It holds two to four tasks sized to the student's declared minutes, chosen in this priority order:

  1. Notebook items due today, capped at one block so review never eats the whole session.
  2. The weakest skill against target, measured as target minus current estimate. The prototype student is 2.0 below in Writing, so Writing appears first.
  3. The weakest question type inside that skill, by accuracy over the last twenty attempts.
  4. A scheduled mock, if the test date makes one due this week.
  5. Something the student is good at, once a week, deliberately — a plan made only of weaknesses is a plan people quit.

Every task states its reason in one line: "Because your last essay scored 5.0 on Coherence". A plan whose reasoning is hidden is indistinguishable from a random content feed, and students treat it as one.

07Study by skill and the question bank

Each skill hub shows the current rung, the mastery of every question type inside that skill, and the band path. Mastery is accuracy across the last twenty attempts of that type, which keeps it responsive without being noisy.

Question type taxonomy

SkillPartsQuestion types the bank must tag
Listening4Form completion, multiple choice, map labelling, matching, sentence completion, note completion
Reading3 passagesTrue/False/Not Given, Yes/No/Not Given, matching headings, matching information, summary completion, multiple choice, short answer
Writing2 tasksTask 1 chart, Task 1 process or map, Task 2 opinion, discuss both views, problem/solution, advantages/disadvantages
Speaking3 partsPart 1 topics, Part 2 cue types (person, place, event, object), Part 3 abstract discussion, pronunciation focus

Filters the bank must support

The filter panel is the difference between a library and a practice tool. All eight dimensions combine, and the active set is shown as removable chips so a student can see why a list is short.

Identity
Skill · Part
Which of the four, and which part within it
Shape
Question type
The taxonomy above, multi-select
Difficulty
Band 5.0 → 7.5+
Tied to the ladder, not to a private scale
Subject
Topic
Environment, education, technology, health, work…
Provenance
Source
Cambridge 17–19, official recent tests, FLYER original
History
Status
New, done, wrong before, flagged
Budget
Time
Under 5, 15 or 40 minutes
Order
Sort
Recommended, newest official, hardest, wrong before
Default
Recommended
Opens pre-filtered to the student's weak types
Player behaviour that is not negotiableListening audio plays once, as in the real test. Reading keeps the passage and the questions on screen together on desktop and switches to tabs on mobile. Writing counts words against the task minimum. Every player has a flag-for-review action and a question palette showing answered, flagged and current.

08Mock tests and official recent tests

Three products sit under one roof, and they are not the same thing:

A completed mock produces a result report with an overall band, four skill bands, a breakdown of where marks were lost by question type, an updated plan, and a shareable card. The report is the moment the product earns trust or loses it, so it shows the student's own wrong answers next to the correct ones, with the passage or transcript that settles the argument.

Open risk: where does this content come from"Official recent test" is a category competitors publish, and the legal footing varies by how the material is reconstructed and worded. Before any of this ships, the content and legal position needs a decision: licensed material, original material written to the same specification, or a differently named category. Nothing in the prototype should be read as a commitment to reproduce exam papers.

09Speaking with the AI Teacher

Speaking is the skill students avoid, so the design goal is to make starting a session feel closer to a video call than to a recording exercise. The student picks an examiner — accent and manner differ — picks a part or a full test, and is then in a call: examiner on screen, their own camera in the corner, question asked aloud, live transcript running underneath.

Session flow

StageWhat happensRequirement
ChooseExaminer, part, topic; weak sounds offered as a warm-upFour examiner personas, at least three accents
InterviewExaminer asks, student answers, follow-ups adapt to the answerResponse latency low enough to feel conversational; barge-in handled
Long turnCue card, one minute preparation, two minutes speakingReal timer, notes allowed, no interruption during the turn
FeedbackBand by the four official criteria, with evidence per criterionEvery claim cites the moment it came from
CaptureErrors written into the notebookPronunciation, grammar and lexis, each typed correctly

Scoring

Feedback reports Fluency and coherence, Lexical resource, Grammatical range and accuracy, and Pronunciation, each as a band with a one-line justification pointing at real evidence — "/θ/ produced as /s/ three times", "no complex sentences in the two-minute turn". A band without evidence is not shown.

Cost and latency are product constraints, not implementation detailsA conversational speaking session is the most expensive minute in this product. The per-session cost ceiling, and what happens when a student exceeds a fair-use limit, are decisions that belong in this document once the numbers exist. They are not in it yet.

10Mistake notebook and email reminders

One notebook, four sources, three kinds of entry. Everything the student gets wrong lands here without them doing anything.

SourceCapturesEntry type
Reading and ListeningWrong answers, plus words tapped in a passageVocabulary
WritingSpelling, grammar and collocation slips from AI markingVocabulary, Grammar
SpeakingMispronounced sounds, wrong word stress, grammar errors in speechPronunciation, Grammar
Chrome extensionWords saved while reading anything on the webVocabulary

Each entry keeps the word or rule, its IPA, its meaning, what the student actually wrote or said, and the sentence it came from. The wrong version matters: it is what makes the entry theirs rather than a dictionary line.

Review schedule

Spaced repetition with four stages — 10 minutes, 1 day, 3 days, 7 days — then mastered. A failed review drops the entry back one stage rather than to the start.

Email rules

11Gamification, pride and ranking

Two different motivations are being served and they should not be confused. Gamification is private and keeps a daily habit alive. Pride is public and gives a reason to tell someone else.

MechanicEarned byServes
StreakCompleting the daily planHabit; the one thing a reminder can protect
XPEvery finished task, weighted by difficulty and by whether it was avoided workProgress inside a week
CoinsMilestonesReserved for a reward mechanic not yet designed
LeaguesWeekly XP, thirty students per league, top five promoteCompetition with peers of similar effort
BadgesTwelve defined achievementsCollection, and nudges toward avoided behaviour
TitlesSustained behaviour, shown beside the name in rankingsPublic identity
Share cardBand improvement or a mock resultPride, and the cheapest acquisition channel available

Ranking scopes

Five boards, because one global leaderboard tells a Band 5.0 student only that they are losing: league, class, students with the same target band, friends, and global. The peer board is the useful one — it shows what students who started where you are actually practise.

Anti-gamingXP must not be farmable by answering easy questions quickly. Weight by band difficulty, cap XP per question type per day, and exclude questions answered faster than a plausible reading time. If the league can be won by grinding Band 4.0 multiple choice, it will be.

12Training corner: IPA and pronunciation

A full chart of the 44 sounds — twelve monophthongs, eight diphthongs, twenty-four consonants — where the student's own weak sounds are marked from real speaking sessions rather than from a generic list for Vietnamese learners.

Each sound opens to: how the mouth makes it, a short mouth video, minimal pairs to hear the contrast, a record-and-compare with a score and a named failure ("/θ/ heard as /s/ in think"), and the words from the student's own notebook that contain it. Word stress and sentence rhythm are separate drills, because four-syllable stress errors and flat sentence rhythm cost Pronunciation marks independently of individual sounds.

The loop that makes this worth building: AI Teacher flags a sound, the sound appears in the training corner as weak, the drill uses the student's own vocabulary, and the next speaking session measures whether it moved.

13Chrome extension

Students who are ready for Band 7.0 vocabulary read English outside the product. The extension makes that reading count.

Build notesPermissions should be requested per-site on activation rather than as blanket access to all pages — a vocabulary tool that can read every page a student visits is a privacy conversation nobody wants to have later. Sync goes one way for now: extension to notebook.

14Teacher corner (Teado.ai)

Deliberately minimal, as the brief asked. The student side holds a class-code join and a placeholder explaining what appears once a teacher connects. Teachers themselves work in Teado.ai.

What phase 2 has to settle: assignments with due dates that auto-mark Listening and Reading and pre-mark Writing and Speaking for teacher review; a class ranking; a teacher override on any AI band, with the override visible to the student; and attendance. The hard part is not the UI. It is agreeing one band estimate that both the student app and the teacher app display, so that a student is never told two different things about their own level.

15Design system, mobile, accessibility

16Data the product must hold

EntityKey fieldsFeeds
StudentGoal, module, target band overall and per skill, test date, daily minutes, reminder settings, streak, XP, leaguePlan, reminders, ranking
Band estimateSkill, band, confidence, evidence count, computed at, stale flagLadder, plan, profile, report
QuestionSkill, part, type, band, topic, source, expected minutes, answer keyBank filters, mocks, mastery
AttemptQuestion, student, answer given, correct, seconds taken, flagged, context (practice or mock)Mastery, weakness ranking, anti-gaming
Notebook entryType, term, IPA, meaning, student's wrong version, source sentence, origin, stage, due dateReview, emails, drills, extension
Speaking sessionExaminer, part, audio, transcript, four criterion bands, evidence, errors extractedFeedback, notebook, training corner
Mock resultTest, four bands, overall, per-type loss breakdown, durationReport, band history, plan

Two fields carry more weight than their size suggests: the student's wrong version on a notebook entry, and evidence on a band estimate. Both are what make the product feel like it is paying attention rather than guessing.

17Delivery order and dependencies

Ordered so that each phase is usable on its own. A student could stop at the end of any phase and still have something worth opening.

PhaseContainsUsable becauseBlocked by
P0 doneClickable prototype, 34 screens, both layoutsScope is now arguable against something real
P1Onboarding, placement, band ladder, question bank with filters, skill hubs, daily planA student can practise deliberately and see their levelTagged content bank; placement item set
P2Mistake notebook, spaced repetition, review emails, Writing AI markingThe mistake loop closes and brings students backEmail infrastructure; marking quality bar
P3AI Teacher speaking, feedback by criterion, training cornerThe avoided skill becomes practisableVoice stack choice; cost ceiling; latency budget
P4Mocks and official recent tests, result report, band historyStudents can rehearse the real thingContent licensing decision
P5Leagues, badges, titles, ranking boards, share cardsRetention and word of mouthEnough students for a league to be non-empty
P6Chrome extension, Teado.ai teacher cornerStudy extends outside the app and into classExtension review; shared band estimate with Teado
The sequencing riskP5 depends on population. Leagues with four students in them are worse than no leagues. If the student base at launch is small, ranking should ship against class and target-band peers only, and global boards should wait.

18Decisions still open

#DecisionWhy it cannot waitOwner
1Where "official recent test" content comes fromBlocks P4 entirely, and the answer may change what the category is calledContent + legal
2Voice stack, per-session cost ceiling, fair-use limitDecides whether speaking is unlimited or metered, which changes the pricing pageProduct + engineering
3How v2 relates to IELTS Guru: successor, redesign of its student side, or separate productDecides authentication, student identity, whose content bank is used, and how Teado connects. Everything else in this document assumes an answerProduct
4English only, or English and Vietnamese at launchCheap to decide now, expensive to retrofit across 34 screensProduct
5The accuracy bar AI Writing marking must clear before students see a bandA wrong band is worse than no band; this bar gates P2Academic
6Who owns the tagged content bank, and how many questions P1 needsEvery filter, mastery score and band estimate depends on taggingContent
7Free and paid boundaryChanges which screens need an upgrade path drawnBusiness

Related: clickable prototype · IELTS Guru PRD (shipped product) · onboarding walkthrough · all 34 screens and the requirement map · Cloud theme · Teado teacher app PRD · all PRDs