riivo
AI-NATIVE PRODUCT BUILD · CUSTOM ENGINEERING
Proposal for Christy Kirkpatrick · Empanda

AI Assessment and Marking Platform

A concept-proving MVP that marks qualitative answers instantly, consistently and defensibly, with an accredited assessor overseeing every release.
Prepared by riivo for Christy Kirkpatrick · Empanda
Concept-proving MVP · Draft for discussion

01 Why do this now

Assessment is the broken slice of the learning and development value chain. Multiple choice became the default because it was the only format markable at scale, while everything richer still depends on scarce human markers who are slow, inconsistent and expensive. In South Africa the accredited assessor and moderator engine is visibly strained, and moderation backlogs stall accreditation outcomes for training providers.

AI marking of qualitative answers against a proper rubric, distributed as a simple link, does not exist as a product today. It will, and soon. This MVP proves it first, with real learners, for a fixed R150 000.

The rubric studio

CONSISTENT · DEFENSIBLE · YOURS

AI helps the assessment creator turn a question and a learning outcome into criteria, weightings and a model answer. The creator refines and approves; marking only ever runs against an approved, versioned rubric. That is what makes every mark consistent and explainable.

Rubrics drafted in minutes, not meetingsEvery mark traceable to a rubric version

The marking engine

INSTANT · EXPLAINABLE · MEASURED

Typed answers are marked the moment they arrive: a score and written feedback per rubric criterion, with a consistency tolerance the platform must meet. Agreement is measured against a human-marked benchmark set of at least 50 answers, so trust is evidence, not assertion.

Marking turnaround from weeks to secondsA published agreement rate to show sceptical buyers

The compliance unlock

CETA / QCTO ALIGNED · HUMAN IN THE LOOP

The frameworks do not prohibit AI-assisted marking; they require an accredited assessor and moderator overseeing outcomes. The MVP is built around exactly that: AI does the volume, low-confidence marks route to a moderation queue, and every override lands in an immutable audit trail.

Providers retire the marking bottleneckAn evidence pack an external verifier can read

Distribution without integration

ANY DEVICE · ANY LMS · NO LOGIN

Each assessment is a tokenised link. It opens on a flagship laptop or a low-end Android phone from a WhatsApp message, embeds in an LMS page, and needs no account, app or integration project. Buyers can adopt it without asking IT for anything.

Zero integration cost for the pilot clientLearners complete assessments where they already are

02 What we will build

Nine backlog items deliver the full concept-proving loop for one pilot organisation: author, AI-draft the rubric, deliver by link, mark by AI, moderate by exception, release and export.

Rubric studio: AI drafts, the assessor approves

Prototype: the creator builds choice and free-text questions; AI drafts the rubric criteria, weights and model answer for the free-text question, and nothing marks until the assessor approves it.

Moderation: AI marks, an accredited human signs off

Prototype: a low-confidence criterion routed to the moderator, reviewed side by side with the learner's answer, overridden with a reason, and recorded in the audit trail before release.

The learner's result: instant, explained, trusted

Prototype: a provisional result seconds after submission, with a score and written feedback per rubric criterion and a clear note where the assessor is confirming a mark.

03 How it works

04 What's in and what's out

The MVP is deliberately smaller than the full platform vision: enough to prove trustworthy AI marking with a real cohort and support a raise, at a fixed price. Everything in the right column is planned, priced indicatively, and deferred on purpose.

In the MVP · R150 000 fixed

  • One pilot organisation with creator, moderator and admin roles
  • Assessment builder: multiple choice and free-text questions
  • AI-drafted rubrics with model answers, versioned and assessor-approved
  • Tokenised delivery link on any device, autosave and resume, basic iframe embed
  • AI marking per rubric criterion with written feedback, benchmarked against your human-marked set of 50+ answers
  • Moderation queue with override, sign-off and an immutable audit trail
  • Results list with per-criterion detail and CSV export
  • Pilot go-live with checklist, smoke test and rollback plan

Deliberately deferred to the scale roadmap · R325 000 indicative

  • Multi-tenancy with per-organisation data isolation (the pilot data model is rebuilt for this)
  • Voice answers, transcription and multilingual capture (MVP marks typed answers, verified for English)
  • Proctoring: photo identity, timing, focus logging
  • Tiered model routing, confidence auto-routing and cost telemetry
  • Dashboards and standards-aligned accreditation reporting
  • Results API, Marking-as-a-Service API, rubric API and QTI import
  • Subscription billing with usage metering
  • Documentation, formal UAT cycles and a hypercare window

These deferrals are what hold the MVP at R150 000: each one adds scale-readiness, not proof. The roadmap is re-baselined against pilot actuals before contracting.

05 Effort & sizing

Effort is expressed with our Small / Medium / Large sizing and a build-point scale, giving a single comparable measure of build size across the project.

SizeEffortBuild points
S  SmallHalf a day1
M  MediumOne day2
L  LargeTwo days3

Foundation and access

What it doesSizePoints
Production environment with CI/CD, secrets vault, daily backups and error alerting.L3
Creator, moderator and admin roles for the pilot organisation; learners join by tokenised link, no accounts.M2
Sub-total5

Authoring and the rubric studio

What it doesSizePoints
Assessment and question authoring with locked publish versions; every attempt records the version it ran on.L3
Multiple choice with automatic scoring, single and multiple correct options.S1
AI drafts rubric criteria, weights, level descriptors and a model answer per free-text question.L3
Review, edit and approve flow with rubric versioning; unapproved rubrics cannot mark.M2
Sub-total9

Learner delivery

What it doesSizePoints
Mobile-first assessment player opened from a tokenised link, verified on the agreed lean device set.L3
Autosave with resume after a dropped connection, plus a basic iframe embed.S1
Sub-total4

AI marking and moderation

What it doesSizePoints
Per-criterion AI marking with written feedback and a consistency tolerance the platform must meet.L3
Marking queue with retries and full mark provenance: model, prompt version and rubric version.M2
Agreement measured against your human-marked benchmark set of at least 50 answers.S1
Moderation queue with side-by-side review and per-criterion override with a mandatory reason.M2
Immutable audit trail and a release policy gate on every result.S1
Sub-total9

Results and go-live

What it doesSizePoints
Results list with per-criterion detail per learner.S1
CSV export matching on-screen filters.S1
Pilot go-live: checklist, smoke test and rollback plan.S1
Sub-total3

Total build effort

WorkstreamSizesBuild points
Foundation and access1 × L, 1 × M5
Authoring and the rubric studio2 × L, 1 × S, 1 × M9
Learner delivery1 × L, 1 × S4
AI marking and moderation1 × L, 2 × M, 2 × S9
Results and go-live3 × S3
Total build points30

Project activities

Alongside the build, delivery includes:

06 Delivery plan

The MVP fits a focused six-week window from entry criteria to pilot go-live, followed by the pilot cohort running live. The scale roadmap is re-baselined against pilot actuals and typically funded from the raise the MVP enables.

Entry criteria
Before Sprint 1
  • Human-marked benchmark set of 50+ answers
  • Data residency and POPIA hosting decision
  • Product name and brand assets
Build sprints
Weeks 1 to 5
  • Foundation, authoring and rubric studio
  • Delivery link, marking engine and moderation
  • Benchmark evaluation and results
Pilot go-live
Week 6
  • Checklist, smoke test, rollback plan
  • Pilot cohort live with real learners
Scale roadmap
Post-raise
  • Multi-tenancy, dashboards, voice, proctoring
  • Results and Marking-as-a-Service APIs, billing

Indicative timeline

Entry criteria
Client
Foundation & authoring
Build
Marking & moderation
Build
Benchmark & results
Evaluate
Pilot go-live
Live
Week 0Week 2Week 4Week 6

07 Key assumptions

08 Investment

A single fixed price for the concept-proving MVP. The scale roadmap is indicative and contracted separately once pilot evidence is in.

AI Assessment and Marking Platform · concept-proving MVP
Fixed price for the complete MVP scope in this proposal · 30 build points
R150 000
excl. VAT
What is included
Authoring with the AI rubric studioTokenised delivery link on any deviceAI marking benchmarked against your human-marked setModeration queue with immutable audit trailResults, CSV export and pilot go-live

Billing follows riivo's effort-weighted milestones: 20 percent on story sign-off, 60 percent on development completion and 20 percent on client QC per sprint. The scale roadmap is indicative at R325 000 (65 build points, plus or minus 25 percent) and is re-baselined against pilot actuals.

09 Roles & approval

RoleWho
Driverriivo delivery team
ApproverChristy Kirkpatrick, Empanda
Responsibleriivo build team; Christy's accredited assessor for moderation
InformedPilot organisation stakeholders

Next step: confirm the MVP scope and entry criteria, and riivo schedules Sprint 1. The pilot can be live with real learners inside six weeks of the benchmark set arriving.