Assessment Rubrics for Timing and Rhythm

Five formative rubrics for judging a learner's timing growth with Chronkle: observation-led, low-stakes, and honest about their own limits.

What this is: a set of five classroom rubrics for noticing how a learner's timing and rhythm are developing while they play Chronkle. Each rubric covers one dimension — accuracy, stability, improvement, strategy use, and self-reflection — and describes observable behaviour at four levels rather than handing out a score. They are teaching tools, not tests.

Why formative, and why zero data

Rhythm is learnable, and like most learnable skills it improves fastest when assessment feeds back into teaching instead of sitting in a grade book. That is the core idea of formative assessment, or assessment for learning: assessment used to shape the next step rather than to sum up the last one, a practice classroom research associates with some of the larger gains available in ordinary teaching (Black and Wiliam, 1998). The rubrics below are written in that spirit: they help you notice what a learner is doing as they play, describe it plainly, and decide what to teach next.

A word on data. Chronkle keeps no account, no login, and no server-side record of anyone's play; scores live only on the device for the length of a session and vanish when the tab closes. That "zero data" design shapes how these rubrics work: there is no stored dashboard of results to mine, so the assessment evidence is simply what you and the learner observe together in the moment. That keeps the stakes low, keeps a child's timing data private, and keeps the focus where formative assessment wants it: on the conversation, not the record.

Read against the learner's own starting point

Timing varies naturally from person to person and from day to day, and much of that variation is ordinary motor noise rather than a difference in skill (see motor variability in timing). So these rubrics describe growth relative to a learner's own earlier work, not against an absolute standard or an age norm. A learner is "Extending" when they have moved well beyond where they themselves began, not because they have beaten a benchmark. None of this is a clinical or diagnostic instrument: Chronkle is a practice game, and a rubric level is a snapshot of behaviour in one setting, never a verdict on ability. For how the game frames timing, see the methodology note.

Rubric 1: Accuracy

Accuracy looks at how close a learner's responses land to the target beat or pattern: where the taps fall, not how they feel about them.

LevelWhat it looks like
EmergingResponses often land clearly early or late, with large errors in both directions within one round. The learner may not yet hear that a tap missed.
DevelopingResponses cluster nearer the target and bigger misses get noticed; the learner can usually say whether an attempt was early or late. A pulse is there but wanders.
SecureMost responses sit close to the target across a whole round; the learner reliably lands on the beat at comfortable tempos and holds it through a short pattern.
ExtendingAccuracy holds up under harder conditions (faster tempos, longer patterns, changes of subdivision), and the centre stays stable even when a few taps slip.

Rubric 2: Stability (consistency)

Stability looks at how even the timing is: whether spacing stays regular within a round and repeats across rounds. A learner can be fairly accurate on average yet still lurch between rushing and dragging (see rushing and dragging), which is a stability issue rather than an accuracy one.

LevelWhat it looks like
EmergingSpacing is uneven; strong and weak attempts sit side by side, and results swing widely from one round to the next.
DevelopingSome runs are noticeably steadier than others; a decent result is repeatable but not yet on demand, and consistency is best at slow tempos.
SecureSpacing is even within a round and repeatable across several rounds at a familiar tempo; a wobble is the exception rather than the rule.
ExtendingSteadiness survives distraction and tempo changes; after a slip the learner recovers within a beat or two instead of unravelling.

Rubric 3: Improvement (growth from own baseline)

Improvement looks at change over time measured against the learner's own earlier sessions, never against peers. Keeping a light record of where a learner started makes this visible; the note on recording and evaluating your timing suggests simple ways to do that.

LevelWhat it looks like
EmergingLittle clear change yet from an early baseline, or change that is hard to separate from ordinary day-to-day variation.
DevelopingA modest but repeated gain over several sessions compared with the learner's own starting point, though better on some game types than others.
SecureClear, sustained improvement from baseline that shows across most game types and still holds after a break from practice.
ExtendingThe learner carries gains over to new material or a different instrument, and sets their own next target from where they have been.

Rubric 4: Strategy use

Strategy use looks at whether a learner brings a deliberate approach to keeping time: counting in, subdividing internally, warming up, or pacing their breathing to steady the pulse.

LevelWhat it looks like
EmergingPlays reactively with no visible plan; does not yet use counting, subdivision, or a warm-up before a round.
DevelopingTries a strategy when prompted (counting in, subdividing, steady breathing) but applies it patchily and drops it under pressure.
SecureChooses and uses a fitting strategy without prompting; subdivides internally, counts in, or paces themselves to hold the pulse.
ExtendingSwitches strategy to suit the task, can explain why a given approach helps, and adapts it when the material gets harder.

Rubric 5: Self-reflection

Self-reflection looks at how well a learner can judge and talk about their own timing. In formative assessment it is the most valuable dimension of all: a learner who can hear their own errors no longer depends on you to spot them.

LevelWhat it looks like
EmergingCannot yet say whether an attempt was early, late, or steady, and relies on the game's feedback to know.
DevelopingDescribes a result in general terms ("that one felt rushed") and begins to connect it to something they did.
SecureJudges their own timing accurately before seeing any feedback, names what went wrong, and suggests a fix for the next try.
ExtendingReflects across sessions, spots recurring patterns in their own errors, and plans practice around them.

How to use these rubrics

These tools reward light hands. A few cautions keep them fair and stop them from claiming more than they can.

Watch out for device and input latency. Any browser game measures a chain that includes the keyboard or touchscreen, the operating system, and the display, and every link adds a little delay. Studies of web-based timing tasks find that measured response times carry device-dependent lag and jitter that can reach tens of milliseconds and differ between machines (Anwyl-Irvine and colleagues, 2021). So the absolute numbers a learner sees partly describe the tablet, not the child. Compare like device with like device where you can, and treat raw millisecond readings as rough, not exact.

A single measurement is not an ability level. One round is a noisy sample. Fatigue, mood, an unfamiliar device, or a lucky streak can move a result far more than a real change in skill, which is why measurement guidance stresses gathering enough evidence before drawing a conclusion (AERA, APA and NCME, 2014). Place a learner on a level only after watching across several sessions, and hold that placement loosely.

Read for growth, keep it a conversation. The point is not to rank a class but to find each learner's next useful step and talk it through. Share the rubric language so learners can locate themselves, invite them to set the next target, and revisit it together. Used that way the rubrics stay what they are meant to be: a support for learning, not a label.

Rubrics describe what you see; the game gives you something to see. Sit a learner down with a short round of Chronkle, watch one dimension at a time, and let the plain-language levels do the rest. Free, and it stores nothing.

Play Chronkle →

Frequently asked questions

Are these rubrics a diagnostic or clinical test?

No. They are formative teaching tools for describing observable behaviour during a practice game. They are not designed or validated to identify any condition, and they should never be used to diagnose one. If you have a clinical concern about a learner, refer to an appropriate professional rather than reading it from a rubric.

Why measure against a learner's own baseline instead of a class standard?

Because timing varies naturally between people and from day to day, an absolute cut-off would penalise ordinary variation and reward a fast starting point rather than genuine learning. Comparing a learner with their own earlier work keeps the assessment about growth.

How many sessions should I watch before placing a learner on a level?

More than one, and ideally several across different days. A single round is a noisy sample that a bad device, tiredness, or luck can distort. Look for a pattern that repeats before you settle on a level, and revise it freely as you see more.

Does device or browser lag affect the results?

Yes. Keyboards, touchscreens, and displays each add small delays, so the exact millisecond figures partly reflect the hardware rather than the learner. Compare results on the same kind of device where you can, and treat absolute numbers as approximate rather than precise.

Does Chronkle store any of this assessment data?

No. Chronkle keeps no account and no server-side record; scores stay on the device for the session only and vanish when the tab closes. The assessment evidence is what you and the learner observe together, not a stored profile.

Keep reading

Sources: Black, P. and Wiliam, D. (1998), "Assessment and Classroom Learning," Assessment in Education, 5(1), 7-74, doi:10.1080/0969595980050102; Assessment Reform Group (2002), Assessment for Learning: 10 Principles; Anwyl-Irvine, A., Dalmaijer, E. S., Hodges, N. and Evershed, J. K. (2021), "Realistic precision and accuracy of online experiment platforms, web browsers, and devices," Behavior Research Methods, 53, 1407-1425, doi:10.3758/s13428-020-01501-5; American Educational Research Association, American Psychological Association and National Council on Measurement in Education (2014), Standards for Educational and Psychological Testing.

Understand it, now train it. The daily games stay free forever. A Personal Pass adds progress tracking (trends + export); Teacher / School tools let you assign it and follow a whole class. For teachers & self-learners →

Updated · By Lajos Toldi — educator and researcher in adaptive intelligent tutoring systems at AI24EduLabs (part of AI24Labs). Scientific claims are checked against authoritative sources; see our editorial policy.