websitehot

Sparrow-2

Noise cancellation isn't designed for conversational AI

sparrow2.tavuslabs.org

About Sparrow-2

Have a real conversation with a Tavus PAL and watch Sparrow-2 read it live: who holds the floor, who is acknowledging, and who just cut in.

In the maker’s words

Hey there, I’m Brian. I've been shipping conversational models at Tavus for the past two years. I want to tell you about our latest audio-understanding/turn-taking model: Sparrow-2! It’s a new category of model and a unique new approach to conversational flow understanding. Earlier this year we launched Sparrow-1, (at the time) our SoTA turn-taking model. Since our Sparrow-1 launch, I’ve spent a lot of time listening to humans talking and trying to really understand how people know when to talk, when to listen, and when to wait. I’ve also been hunting down failure modes of the current SoTA models. And while Sparrow-1 is great, there are some failure patterns I see. We tried solving the problems with existing approaches, but solving one problem only created another. Current turn-taking models (like Sparrow-1) use noise cancellation to strip background sound, leaving only basic prosody and…
code_brian, launching on Hacker News

Where people found it

More sites like Sparrow-2