Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
darknoon
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
14 ms
·
1.
▲
by
darknoon
29d ago
they really buried the lede on which models are supported, then the link is broken as of now to the list ( https://app.router.com/supported-models )
2.
▲
by
darknoon
2mo ago
Really want to use this, but for web page design you really need vision!
3.
▲
by
darknoon
3mo ago
I frequently run into scenarios where it won't let me generate the email within 1password on a website, and I have to go to Fastmail and then manually do it. Is this something you have bene able to work around?
4.
▲
by
darknoon
4mo ago
the `claude` binary is essentially a packed copy of bun + the js code, so this will replace the native runtime part of claude code.
5.
▲
by
darknoon
5mo ago
This relies on knowledge of the distribution, just querying in the middle of A = [1, 2, 4, 8, 16, ..., 2^(n-1)] is slower than binary search
6.
▲
by
darknoon
8mo ago
I haven't seen one that worked properly—can you list a couple examples? Some of the ones that say they're "AI" are just VTracer / Potrace and don't give nice control points.
7.
▲
by
darknoon
8mo ago
I think you'd find that it's far from "any human" who can do this without looking anything up. I have 15y of dev exp and couldn't do this from memory on the cli. Maybe in c, but less helpful to getting stuff done!
8.
▲
by
darknoon
8mo ago
really weird graph where they're comparing to 3x H100 PCI-E which is a config I don't think anyone is using. they're trying to compare at iso-power? I just want to see their box vs a box of 8 h100s b/c that's what p
9.
▲
by
darknoon
9mo ago
Here's the problem, you're still going to get scraped and the LLM will understand it anyway. Maybe at best you'll get filtered out of the dataset b/c it's high perplexity text?
10.
▲
by
darknoon
10mo ago
The developers also gave a talk about Helion on GPU Mode: https://www.youtube.com/watch?v=1zKvCLuvUYc
11.
▲
by
darknoon
1y ago
Here's the thing, they've completely given up and started making their (inferior to AMD) CPUs on TSMC. For example, Arrow Lake is on TSMC N3B. So it's not getting amortized over anything at all and their valuation is going to
12.
▲
by
darknoon
1y ago
It's ok, somewhere between a qwen 2.5 VL and the frontier models (o3 / opus 4) on visual reasoning
13.
▲
by
darknoon
1y ago
In ML, often it does work to a degree even if it's not 100% correct. So getting it working at all is all about hacking b/c most ideas are bad and don't work. Then you'll find wins by incrementally correcting issues with
14.
▲
by
darknoon
1y ago
If you were doing a lot of scraping, you could just solve this on a GPU in 1/10 or less of the time it takes a human's phone to do it. Generally you need a decent computer to render a webpage while scraping it these days, so I don
15.
▲
by
darknoon
1y ago
anyone know why they mix in the 3 previous tokens? could have just as easily done 5 or 2 right?
16.
▲
by
darknoon
2y ago
One vote for image inputs here. I would love a fine-tuned qwen-2-vl-72b on demand, but most of the solutions are "talk to us" level expensive. I'm assuming you beat the price or convenience of a replicate / modal solutio
17.
▲
by
darknoon
2y ago
I think just reading the code wouldn't make you a good programmer, you'd need to "read" the anti-code, ie what doesn't work, by trial and error. Models overconfidence that their code will work often leads them to fa
18.
▲
by
darknoon
2y ago
this is somewhat similar, but diffusion transformers typically use a pre-trained text model as the text conditioning whereas, in this case it's integrated and trained together multimodally.
19.
▲
by
darknoon
2y ago
I tried this w/ AM5, but realized that despite there theoretically being enough lanes for dual x16 PCI-e 4.0 GPUs, I couldn't find any motherboards that are actually configured this way, since dual-GPU is dead in consumer for gami
20.
▲
by
darknoon
2y ago
Relevant video about some of the history of superconducting computers: https://www.youtube.com/watch?v=14r2oMsAaE8
21.
▲
by
darknoon
2y ago
Why does this webpage have auto-playing audio?
22.
▲
by
darknoon
2y ago
No, it's connected to AirTrain, which is slow and unpredictable, which is then connected to either the A or the LIRR.
23.
▲
by
darknoon
3y ago
I think Vision Pro isn't really a product, it's for developers / early adopters to make apps that will then be available once a consumer version is ready.
24.
▲
by
darknoon
3y ago
in the video, it seems like you were designing something more complicated than the logos I was expecting-more of a vector illustration. It seems roughly in line with stable diffusion, and other models which struggle to make the clean, symbo
25.
▲
by
darknoon
3y ago
It's striking vs the France chart, though I wonder how this reflects electricity import / exports https://www.energy-charts.info/charts/energy/chart.htm?l=en&...
26.
▲
by
darknoon
3y ago
Interesting context I would like to know
27.
▲
by
darknoon
3y ago
Would be more interesting if Pytorch with MPS backend was also included.
28.
▲
by
darknoon
3y ago
Please reword this clickbait headline. It's just disk offloading.
29.
▲
by
darknoon
3y ago
Whisper v3, just a couple weeks ago https://huggingface.co/openai/whisper-large-v3
30.
▲
by
darknoon
3y ago
Stable Diffusion + fine-tuning is quite powerful for creating specific art styles that aren't easily described to DALL-E / MJ
More ›