I built a browser-based digital human that listens to a user and responds with generated, lip-synced video. I worked on this project 2 years back, I would like share the learnings with the community.
I’d appreciate feedback from anyone who has worked on real-time avatars, streaming TTS, or interruption handling.
2 comments of 3
[ 330 ms ] story [ 315 ms ] threadI’d appreciate feedback from anyone who has worked on real-time avatars, streaming TTS, or interruption handling.