Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
bpanahij
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
bpanahij
8mo ago
That’s unfortunate and certainly not what I spend my time dreaming about. My favorite use case for the elderly is as a sort of companion for sharing their story for future generations. One of our partners uses our technology to help elderly
2.
▲
by
bpanahij
8mo ago
The response timing in the chart in the blog post shows that even with perfect precision/recall Sparrow-1 also has the fastest true positive response times. The turn taking models were evaluated in a controlled environment with no addi
3.
▲
by
bpanahij
8mo ago
You can try Sparrow-1 with any of our PALs, or by signing up for a developer account.
4.
▲
by
bpanahij
8mo ago
Try out the PALs: they all use Sparrow-1. You can try Charlie on Tavus.io on the homepage in one of the retro retro-styled windows there.
5.
▲
by
bpanahij
8mo ago
This is a very good idea. We currently have a model in our perception system (Raven-1) that performs this partially. It uses audio to understand tone and augment the transcription we send to the conversational LLM. That seems to have an imp
6.
▲
by
bpanahij
8mo ago
You should be skeptical, and try it out. I selected 28 long conversations for our evaluation set, all unseen audio. Every turn taking model makes tradeoffs, and I tried to make the best tradeoffs for each model by adjusting and tuning the i
7.
▲
by
bpanahij
8mo ago
That’s great! I also built Sparrow-0, and Sparrow-1 was designed to address Sparrow-0’s shortcomings. 1 is a much better model, both in terms of responsiveness and patience.
8.
▲
by
bpanahij
8mo ago
I haven’t tried that one yet, I’ll check it out.
9.
▲
by
bpanahij
8mo ago
Maybe infiniband is a bit more than we can handle. That technology is incredible! You are right though, we have been willing to build things we needed that didn’t exist yet, or were not fast enough or natural enough. Sparrow-1, Raven-1, and
10.
▲
by
bpanahij
8mo ago
As a dev myself, I see a couple of modes of operation: - push to talk - long form conversation - short form conversation In both conversational approaches the AI can respond with simple acknowledgements. When prompted by the user the AI cou
11.
▲
by
bpanahij
2y ago
Thanks for checking it out!
12.
▲
by
bpanahij
2y ago
https://www.tavus.io/pricing Scroll down the page to find our pricing.
13.
▲
by
bpanahij
2y ago
You could hack this together now with OBS and Tavus.
14.
▲
by
bpanahij
2y ago
Thanks for these thoughts and compliments. I love the idea of preventing landfill with this tech. Our team is awesome and we really love our customers and all the jobs that can be done with this kind of tech!
15.
▲
by
bpanahij
2y ago
We're partnering with GPU infrastructure providers like Replicate. In addition, we have done some engineering to bring down our stack's cold and warm boot times. With sufficient caches on disk, and potentially a running process&#x
16.
▲
by
bpanahij
2y ago
Scroll down the page and the per minute pricing is there: https://www.tavus.io/pricing We bill in 6 second increments, so you only pay for what you use in 6 second bins.
17.
▲
by
bpanahij
2y ago
We have the ability to send phonetic pronunciations as guidance, and this could be a great addition to our LLM/response generation stack! Adding a check for names and then adding in the phoneme.
18.
▲
by
bpanahij
2y ago
Thanks for that insight. Brian here, one of the engineers for CVI. I've spoken with CVI so much, and as it has become more natural, I've found myself becoming more comfortable with a conversational style of interaction with the va
19.
▲
by
bpanahij
2y ago
Checkout Tavus.io for realtime. They have a great API for realtime conversational replicas. You can configure the CVI to do just about anything you want to do with a realtime streaming replica.
20.
▲
by
bpanahij
2y ago
Tavus.io already does this. They have realtime conversational replicas: with a < 1 second response time. Hyper realistic too.