Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
zan2434
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
zan2434
5mo ago
I am unfortunately just paying for this out of pocket! Didn't really expect it to blow up like this.
2.
▲
by
zan2434
1y ago
interesting, so you think the issue with the above approach is the graph structure being too rigid / lossy (in terms of losing semantics)? And embeddings are also too lossy (in terms of losing context and structure)? But you guys are w
3.
▲
by
zan2434
2y ago
Buried, but on Page 24 they reveal to me the most surprising massive capability leap - that o3-mini is way better at conning gpt-4o for money (79% win rate for o3-mini vs 27% for full o1!). It isn't surprising to me that "reasonin
4.
▲
by
zan2434
2y ago
I was running into some scaling issues, but should be all working now!
5.
▲
by
zan2434
2y ago
The voice is just OpenAI’s default tts voice. I agree that Veritasium video is an incredible work and the ai version is absurd by comparison! This is mostly a proof of concept that this is possible at all, and as LLMs get smarter it’ll be
6.
▲
by
zan2434
2y ago
Hmm the initial version of the app only took me about a day to get something working, but that version took minutes to generate a single video and even then only worked a third of the time. It took a solid 2 weeks from there to add all the
7.
▲
by
zan2434
2y ago
There is a job queue on the backend with statuses, just not worth breaking the streaming experience to ask the LLM rewrite broken manim segments out of order
8.
▲
by
zan2434
2y ago
totally fair! I like the XKCD comic as well because it hints at a potential solution - even if you can't always be correct, how you respond to critical questions can really help. I'm working on a feature for users to ask follow up
9.
▲
by
zan2434
2y ago
These are amazing examples! Thanks for all the feedback, detailed info, and persistence in trying! HN hug of death means I'm running into Gemini rate limits unfortunately :( will def make that more clear when it happens in the UI and t
10.
▲
by
zan2434
2y ago
sad looks like I already hit the Gemini rate limit :( Switching to Claude!
11.
▲
by
zan2434
2y ago
thanks! Streaming was actually pretty hard to get working, but it goes roughly like this as a streaming pipeline: - The LLM is prompted to generate an explainer video as sequence of small Manim scene segments with corresponding voiceovers -
12.
▲
Show HN: AI that generates 3blue1brown-style explainer videos
(tma.live)
93 points
by
zan2434
2y ago
|
46 comments
13.
▲
by
zan2434
2y ago
This actually makes a lot of sense! Sounds like finding dangerous chemicals is easy and is not the actual limitation at all.
14.
▲
by
zan2434
2y ago
This is a textbook bad faith comment / attacking the person but not the subject of the argument. I’m just asking about others’ assessment of the benefits and risks. What do you think? Or do you think it’s just not worth considering?
15.
▲
by
zan2434
2y ago
Clear snark aside, content piracy has pretty bounded risks so isn’t a reasonable comparison
16.
▲
by
zan2434
2y ago
This is both awesome and feels very dangerous to release publicly, no? Can’t this be used to discover novel bioweapons as easily as it can be used to discover new medicines? Genuinely curious, would love to learn if that isn’t true / o
17.
▲
by
zan2434
3y ago
Does this ruling make IVR systems illegal, too? I applaud the effort because this really could curb a lot of spam, but I am curious because AI generated voices in phone calls are already ubiquitous and have been for decades. Do they have a
18.
▲
Show HN: 2-way interruptible voice AI
(twitter.com)
2 points
by
zan2434
3y ago
|
0 comments
19.
▲
by
zan2434
3y ago
I agree. Have been working on a 2 way interruptions system + streaming like this. It's not robust yet, but when it works it does feel magical.
20.
▲
by
zan2434
3y ago
Hey! Awesome work. It seems like in theory this encoding scheme should enable the a model like this to generate images as well, by outputting image tokens, is that right?
21.
▲
by
zan2434
3y ago
This was an inspiring read! Reminds me of Simon Willison's analogy of LLMs to "calculators for words" but this author takes the idea even further. I agree the analogy points to foundation model companies like OpenAI and Anthr
22.
▲
by
zan2434
3y ago
Anyone wanna convert this to GGML so we can run it with LLaMa.cpp?
23.
▲
by
zan2434
4y ago
I don't think that's even necessary! You could use this acoustic technique to shape ordinary light-curing resin and then flash with UV light to harden it.
24.
▲
by
zan2434
4y ago
This should replace the OP link
25.
▲
by
zan2434
4y ago
I think programmer salaries have been so high because FAANG companies were willing and able to pay through the nose to hoard talent. This is changing.
26.
▲
by
zan2434
4y ago
:D Thanks for mentioning this! I hadn't noticed it and it is an amazing touch.
27.
▲
by
zan2434
5y ago
This is fascinating, groundbreaking research I think! Never heard of anything like this. Any neuroscientists here who can speak to implications?
28.
▲
by
zan2434
5y ago
Correct me if I’m wrong, but don’t transparent pigments only change light color by allowing certain wavelengths of light to pass through them? That would mean the glaze in question here is only changing the color of the light by blocking ou
29.
▲
by
zan2434
5y ago
Fascinating! Can you explain this a bit more and share some examples?
30.
▲
by
zan2434
6y ago
Heh, I thought this was going to be an article about Engelbart, et. al.'s mothers!
More ›