Education guide

Assessment

Assessment makes learning visible through completed work, explanation and reflection, with one fixed standard that works equally for a robot, a song and a community project.

Version 4.1.0 Published 27 August 2026 Updated 3 September 2026

A standard for completion

Creative and technical work rarely follows one neat sequence. Learners may begin with a problem, a material, a sound, a reference or an unexpected discovery. So what gets assessed is the evidence that exists when a project is complete, rather than whether one prescribed process was followed.

For assessed work from the Maker stage onward, six evidence states must be present. Their order may vary; a facilitator may teach a workflow, but nobody is assessed on having followed one. The six are evidence functions, not submission formats: how each is shown can be adapted to the learner, the subject and the project, as long as what it evidences is there.

Why not publish a cycle

A cycle that opens with “identify a problem” is engineering-shaped. Presenting it as universal tells every musician and artist in the room that they are working wrongly. Research on where knowledge construction actually happens in maker education locates it in the reflective work around the artefact, not in adherence to a procedure[4].

Six evidence states

All six present, any order Aim The intent is stated, with criteria to judge it by. Options More than one route was considered before committing. Made The thing exists, an object, a performance, a service. Evaluated It was judged against the aim’s own criteria. Shared There is a record, and an audience actually saw it. Refined At least one improvement came out of what was learned. There is no arrow on this diagram. Some makers start from an aim; others start from a material, a sound or an accident. The standard is what exists at the end, not the route taken to it.
The categories align making, engineering design, documentation and reflection. Compare engineering design practice in the science standards[22] and open portfolio work in maker settings[21].

The wording adapts to the subject. Made may be a robot, a digital service, a song, an artwork or a facilitated community activity. Evaluated means a technical test in one project and a critique or a rehearsal in another. The category is constant; the evidence that satisfies it is native to the subject.

Explorer work is exempt. It is unassessed by definition, run as open tinkering with learner-set goals[5]. The Definition of Done applies exactly where certification applies.

The same standard in every subject

Every structured programme publishes its own subject-specific reading of the standard, so the interpretation is settled in advance rather than improvised. A facilitator still explains it (in a learner’s own language, or in whatever form lands) and that is presentation, not a change to what the evidence has to show.

Evidence stateRoboticsMusicVisual arts
AimWhat problem, and for whomWhat piece, what feeling, what occasionWhat idea, and what response is wanted
OptionsSketch two or three mechanismsTry several melodic or rhythmic ideasThumbnails, references, several compositions
MadeBuilt and programmedComposed, arranged, recordedThe work produced
EvaluatedTested against spec, does it work?Rehearsed and listened back, as intended?The critique, does it read as intended?
SharedDemonstration plus build logPerformance plus session notesExhibition plus process journal
RefinedRedesigned from test dataArrangement revised, re-recordedReworked from the critique

In arts subjects the artist’s statement is the natural Shared record. Documented portfolios are a well-precedented way of making process and learning visible across settings, and around three-quarters of surveyed makerspaces already assess in some form, so this is established practice rather than an experiment[21].

Three sources of certificate evidence

1 The artifact The hard result. It exists, and it meets the six evidence states. Answers What was produced? 2 The talk-through A short structured conversation: how and why does this work? Answers Is it understood? 3 The reflection What did I learn, and what would I do differently next time? Answers What transfers? The second is the one that carries the most weight and costs the least, only a facilitator’s time. It works identically for a robot, a song and a painting, and it is the practical way to tell a learner who understands their project from a polished artifact produced with substantial help.
Each source answers a question the other two cannot, and all three are required for a certificate at any level. The three are functions, not fixed media: the explanation and the reflection may be spoken, written, signed, drawn, recorded or given through supported communication, provided they answer the same questions.

The reflection is not decoration. Experiential accounts of learning place reflection as the step that helps a learner connect a specific experience to the next situation[7], and, just as importantly here, it is what makes that connection visible to an assessor. Without it, the evidence shows only that this particular thing was built once.

The talk-through carries a lot of weight and costs the least. It needs only a facilitator’s time, and it works across subjects better than most instruments, which matters because digital work is far easier to assess automatically than craft work is.

It is also the cheapest way to tell genuine understanding from a polished artefact produced with substantial help. That has become more pressing, not less, since generative tools became available. An equivalent accessible format can replace the spoken conversation, as long as it answers the same questions.

Fair, legible assessment

  • Criteria are visible before work begins, written in language learners can actually use.
  • Comparable evidence is compared. A learner’s work is judged against the level rubric they enrolled at, never against the strongest person in the room.
  • Documentation makes progress reviewable, so a judgement can be revisited rather than remembered.
  • Assessment is separated from encouragement. Both matter; conflating them makes the first meaningless and the second suspect.
  • Prior learning is recognised on evidence, not provenance, against the same prerequisites or outcomes as taught learning, wherever it was acquired. The placement check does this for entry; completion is judged against the programme’s own standard.

Attendance creates an opportunity to learn. Demonstrated, explained work shows what was learned. Only the second belongs on a certificate.

Judging quality

The Definition of Done establishes that work is complete. How good it is, is a separate question, and the tools for judging that are rougher than a percentage mark would suggest.

Two published frameworks shape how the level rubrics describe rising quality without inventing precision. Structural accounts of learning outcomes separate work that handles one element, several unconnected elements, several related elements, and a generalisation beyond the task[20]. Skill-acquisition accounts describe the move from following rules, through situational judgement, to fluent practice[18].

Revised taxonomies of educational objectives treat creating as a complex form of learning rather than a decorative one[19]. That is what lets a maker programme put it in its advanced outcomes rather than at the end of everything else.

On creativity specifically

Creativity is not one thing at one scale: published work separates it into four, from a personally meaningful insight up to field-changing contribution[23]. Almost everything a learner produces sits in the first two, so a rubric that grades against professional novelty is measuring the wrong thing.

What this does not solve

An open limitation

Hard results are straightforward to verify: the artefact exists, it works, it was tested. Soft results (persistence, judgement, collaboration, confidence) are genuinely hard to verify, and maker educators report this consistently.

The talk-through and the reflection together are what the academy does about it. They are not a proof. That flag stays on the certification system until the academy's own delivery data says otherwise. The limitation is published rather than covered with a rubric that implies a precision nobody has.

Two further items are open: the exact exit criteria at each level in each subject are still being authored alongside the curricula themselves, and whether certificates should carry external accreditation or remain internally issued has not been decided. Neither is presented here as finished.