Decode, move and speak!
Self-supervised learning of speech units, gestures and sounds relationships using vocal imitation.
Read the paper
| Original sound | Model reconstruction |
|---|---|
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |
Self-supervised learning of speech units, gestures and sounds relationships using vocal imitation.
Read the paper
| Original sound | Model reconstruction |
|---|---|
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |
![]() | ![]() |