New side project: Samuel, a model I trained to mimic speech using Pink Trombone, a (pre-existing) silly vocal tract…
Summary
A machine learning researcher created Samuel, a speech autoencoder that uses Pink Trombone (a vocal tract simulator) as its decoder to generate intelligible speech. The project is non-trivial because Pink Trombone is non-differentiable and sample-by-sample, requiring careful optimization to maintain intelligibility rather than just matching waveforms.
Summarized by ThreadOut AI from the full thread. May miss nuance — read the thread below.
- #1
- #2
- #3