Sparrow-2
Noise cancellation isn't designed for conversational AI
About Sparrow-2
Have a real conversation with a Tavus PAL and watch Sparrow-2 read it live: who holds the floor, who is acknowledging, and who just cut in.
In the maker’s words
Hey there, I’m Brian. I've been shipping conversational models at Tavus for the past two years. I want to tell you about our latest audio-understanding/turn-taking model: Sparrow-2! It’s a new category of model and a unique new approach to conversational flow understanding. Earlier this year we launched Sparrow-1, (at the time) our SoTA turn-taking model. Since our Sparrow-1 launch, I’ve spent a lot of time listening to humans talking and trying to really understand how people know when to talk, when to listen, and when to wait. I’ve also been hunting down failure modes of the current SoTA models. And while Sparrow-1 is great, there are some failure patterns I see. We tried solving the problems with existing approaches, but solving one problem only created another. Current turn-taking models (like Sparrow-1) use noise cancellation to strip background sound, leaving only basic prosody and…
Where people found it
- Hacker NewsShow HN: Sparrow-2 – Noise cancellation isn't designed for conversational AI11 points5 comments25 days ago
More sites like Sparrow-2
- Giving Opus 5.5 a simulated paint canvasstillwet.art
- AIHOT一个自己找热点、自己写日报的网站框架。把信源和精选标准换成你的,它就是你的行业热点站。
- Offrunmanage every coding agent from one workspace
- OpenDotsYour always-on AI coworkers that move between text, calls, and Slack.
- Pi podRun your pi coding agent in sandboxes on your own server
- Ledge.shRunnable Markdown Notes