Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mahmoudfelfel
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
mahmoudfelfel
2y ago
The current deployed model is English only, we are rolling out a multilingual version later this week!
2.
▲
by
mahmoudfelfel
2y ago
PlayAI (fma PlayHT) founder here, this is a native multiturn voice model that is built for conversations like real-time agents or podcasts. Try it through our playground ( https://play.ai/playground ) or API ( https:/&#x
3.
▲
PlayHT (YC W23) Is Hiring Fullstack Engineers
(jobs.ashbyhq.com)
1 points
by
mahmoudfelfel
2y ago
4.
▲
PlayAI (YC W23) Is Hiring Product Engineers
(ycombinator.com)
1 points
by
mahmoudfelfel
2y ago
5.
▲
PlayHT (YC W23) Is Hiring a Founding LLM Engineer and Researcher
(ycombinator.com)
1 points
by
mahmoudfelfel
3y ago
6.
▲
AI Voice API Benchmark
(twitter.com)
2 points
by
mahmoudfelfel
3y ago
|
0 comments
7.
▲
PlayHT (YC W23) Is Hiring Senior ML Engineers (LLMs, Generative AI)
(ycombinator.com)
1 points
by
mahmoudfelfel
3y ago
8.
▲
PlayHT (YC W23) Is Hiring Senior ML Engineer (Large Language Models)
(ycombinator.com)
1 points
by
mahmoudfelfel
3y ago
9.
▲
PlayHT (YC W23) Is Hiring ML Engineers (LLMs, Generative AI)
(ycombinator.com)
1 points
by
mahmoudfelfel
3y ago
10.
▲
PlayHT (YC W23) Is Hiring a Senior ML Engineer (LLMs/Generative Models)
(ycombinator.com)
1 points
by
mahmoudfelfel
3y ago
11.
▲
by
mahmoudfelfel
3y ago
Very soon.
12.
▲
by
mahmoudfelfel
3y ago
The free accounts have 5k words and then you can upgrade to an API-plan from here https://play.ht/app/api-plans
13.
▲
by
mahmoudfelfel
3y ago
Society definitely needs to adapt to this new norm; we are trying to roll this out as safely as possible, but others are not as careful, and this technology will just become more ubiquitous over time.
14.
▲
by
mahmoudfelfel
3y ago
We have seen use cases in audiobooks, podcasts, marketing videos, explainer videos, Commercials, and Gaming, among others.
15.
▲
by
mahmoudfelfel
3y ago
Whisper is Speech to Text; we are building Text to Speech LLMs.
16.
▲
by
mahmoudfelfel
3y ago
Yes, we just released that for the UltraRealistic TTS ( https://docs.play.ht/reference/api-getting-started ), and it will soon be added to our Standard voices as well.
17.
▲
by
mahmoudfelfel
3y ago
What is in this demo is a very rate-limited, early version of our new model. We have many mitigations in place to increase the safety of our main product (Play.ht); I mentioned some of that here https://news.ycombinator.com/
18.
▲
by
mahmoudfelfel
3y ago
We have been seeing some of these genuine use cases: youtube creators, audiobooks, elearning videos, podcasts, commercials, dubbing, and gaming.
19.
▲
by
mahmoudfelfel
3y ago
We have many mitigations in place to increase the safety of this service, I mentioned some of that here https://news.ycombinator.com/item?id=35331310
20.
▲
by
mahmoudfelfel
3y ago
You can try the high-fidelity voice cloning here https://play.ht/voice-cloning/
21.
▲
by
mahmoudfelfel
3y ago
Cofounder here, What you see in the above demo is a very rate-limited demo of our upcoming model. We realize how dangerous this technology can be and have built a lot of mitigations on our main product (Play.ht) to reduce possible abuse: -
22.
▲
by
mahmoudfelfel
4y ago
Everything, the content and the voices are generated by AI. check my other comments for how we did that.
23.
▲
by
mahmoudfelfel
4y ago
Yes, a finetuned GPT3 on Jobs' biography
24.
▲
by
mahmoudfelfel
4y ago
All AI generated.
25.
▲
by
mahmoudfelfel
4y ago
Thank you :) that is the main point of it, to show people what is possible and inspire them to create.
26.
▲
by
mahmoudfelfel
4y ago
No, All was GPT3 generated, check the other comment about how we prompted GPT3 to start the conversation
27.
▲
by
mahmoudfelfel
4y ago
We will never use any cloned voice in any commercial way without consent and compensation, we only wanted to show the community what is possible and what generative AI models can do.
28.
▲
by
mahmoudfelfel
4y ago
The original model ( https://play.ht/blog/introducing-truly-realistic-text-to-spe... ) was trained on 50k hours of audio, the above voices were just finetuned on the model, only 4-6 hours each. We just finetuned another
29.
▲
by
mahmoudfelfel
4y ago
This was the prompt: " Podcast.AI Great people, great interviews with our host Joe Rogan. Episode 1 - Steve Jobs Summary: " Then GPT3 generated the summary, then we added: " Transcript: Joe: " That is all.
30.
▲
by
mahmoudfelfel
4y ago
It was really hard to find good quality of Steve Jobs' voice, most of his speeches are keynotes on stage, I honestly was surprised that we managed to get to that quality from such poor quality recordings.
More ›