A helper says each word naturally, then models a tap for each syllable. The learner may tap, point or listen. Say the whole word naturally again afterward. The questions use stated models; different pronunciations are not mistakes. Do not count letters. No independent reading, timer or recording is needed. Questions check models, not speech or reading fluency.
Worked example
Moon has one spoken beat in our model, even though it has four letters. Robot has two beats. Each circle stands for one beat.
Listen and speak first
Say moon, robot and banana with your helper. Try their beats before checking: moon has one, robot two, banana three in our models. Ask your helper to listen and model another try. This spoken practice is not automatically scored.