Tasks

Set up a custom controlled TTS process to pre-generate audio for letters, syllables, and words instead of relying on unpredictable browser SpeechSynthesis.

Recognize short spoken answers from a child using the Web Speech API in "What is shown in the picture?" tasks.

Check whether the mechanism helps the child recover a previously unseen word, rather than just memorizing studied examples.

Implement a working interactive prototype that goes through a single complete learning cycle without unnecessary platform infrastructure.

Determine which connections the system considers mastered and when it can hide the image, sound, or spelling.

Put together a compact learning set to test the transfer of letter correspondences to new words.

Establish one short interactive cycle from a familiar image and sound to the recognition of a new spelling.