Viewing profile — bpanahij
bpanahij
HN member- Joined
- Fri, Apr 07, 2023, 6:59 AM UTC
- HN karma
- 14
- Public activity
- 20 items
- HN profile
- View on Hacker News ↗
About bpanahij
No profile information was provided.
Recent public activity
-
comment
Comment #46635725
That’s unfortunate and certainly not what I spend my time dreaming about. My favorite use case for the elderly is as a sort of companion for sharing their story for future generati…
-
comment
Comment #46635684
The response timing in the chart in the blog post shows that even with perfect precision/recall Sparrow-1 also has the fastest true positive response times. The turn taking models …
-
comment
Comment #46635545
You can try Sparrow-1 with any of our PALs, or by signing up for a developer account.
-
comment
Comment #46635535
Try out the PALs: they all use Sparrow-1. You can try Charlie on Tavus.io on the homepage in one of the retro retro-styled windows there.
-
comment
Comment #46635503
This is a very good idea. We currently have a model in our perception system (Raven-1) that performs this partially. It uses audio to understand tone and augment the transcription …
-
comment
Comment #46635444
You should be skeptical, and try it out. I selected 28 long conversations for our evaluation set, all unseen audio. Every turn taking model makes tradeoffs, and I tried to make the…
-
comment
Comment #46635362
That’s great! I also built Sparrow-0, and Sparrow-1 was designed to address Sparrow-0’s shortcomings. 1 is a much better model, both in terms of responsiveness and patience.
-
comment
Comment #46635328
I haven’t tried that one yet, I’ll check it out.
-
comment
Comment #46635303
Maybe infiniband is a bit more than we can handle. That technology is incredible! You are right though, we have been willing to build things we needed that didn’t exist yet, or wer…
-
comment
Comment #46635231
As a dev myself, I see a couple of modes of operation: - push to talk - long form conversation - short form conversation In both conversational approaches the AI can respond with s…
-
comment
Comment #41713359
Thanks for checking it out!
-
comment
Comment #41713316
https://www.tavus.io/pricing Scroll down the page to find our pricing.
-
comment
Comment #41713306
You could hack this together now with OBS and Tavus.
-
comment
Comment #41713006
Thanks for these thoughts and compliments. I love the idea of preventing landfill with this tech. Our team is awesome and we really love our customers and all the jobs that can be …
-
comment
Comment #41712383
We're partnering with GPU infrastructure providers like Replicate. In addition, we have done some engineering to bring down our stack's cold and warm boot times. With sufficient ca…
-
comment
Comment #41712349
Scroll down the page and the per minute pricing is there: https://www.tavus.io/pricing We bill in 6 second increments, so you only pay for what you use in 6 second bins.
-
comment
Comment #41712282
We have the ability to send phonetic pronunciations as guidance, and this could be a great addition to our LLM/response generation stack! Adding a check for names and then adding i…
-
comment
Comment #41711681
Thanks for that insight. Brian here, one of the engineers for CVI. I've spoken with CVI so much, and as it has become more natural, I've found myself becoming more comfortable with…
-
comment
Comment #41478225
Checkout Tavus.io for realtime. They have a great API for realtime conversational replicas. You can configure the CVI to do just about anything you want to do with a realtime strea…
-
comment
Comment #41478061
Tavus.io already does this. They have realtime conversational replicas: with a < 1 second response time. Hyper realistic too.