Tasks
Theoretically investigate what properties of PCM and the speech signal we can reliably measure in time and frequency before designing a new generator.
An interactive multitrack editor for manual experimentation with synthesizer parameters over time, cyclic listening, and discovering interesting sounds.
Build an experimental map of how control parameter changes over time manifest in PCM and acoustic features, taking into account DSP lags and memory.
Verify whether generator scenarios can be automatically reconstructed from recorded audio, first using our own synthetic examples with a known ground truth, and then on human speech.
Speed up the manual search for quality sounds: edit scenarios visually, play them back instantly, and be able to send reference and synthetic logs to the AI agent to return modified JSON.
Find and study browser-based projects and DSP approaches that programmatically reproduce at least a few human syllables or vocal sounds with high quality, without generative TTS and without tying the core to letters/syllables.