Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
pllbnk
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
pllbnk
3d ago
No. We can let them edit spreadsheets, write code, summarize content, control robots even, etc. But people must be held accountable for the actions of their computers.
2.
▲
by
pllbnk
3d ago
Asking the model to follow the safety guidelines is much like asking it to “make no mistakes”. The way these hacks which they are bragging about happened is not by loading an LLM into a GPU with ethernet cable plugged out and providing a pr
3.
▲
by
pllbnk
4d ago
I have collected three reasonable versions from today’s comments: 1. They hit a wall from the technical perspective 2. Inference costs are getting out of hand and newer models require significantly more resources for marginal gains, meanin
4.
▲
by
pllbnk
4d ago
I think they will all conveniently "decide" to slow down because models are becoming crazy expensive, both per token and how much tokens they need to do anything meaningful. The difference between Sol and Astra is 150% pricing inc
5.
▲
by
pllbnk
5d ago
I have listened quite a few interviews with Tao and I see him being very careful about criticizing AI. He very often emphasizes the usefulness of it. Where he is critical has a lot of merit. One of the points I clearly remember him saying t
6.
▲
by
pllbnk
5d ago
LLMs are the fast food for the brain. I don't know how they can be used correctly.
7.
▲
by
pllbnk
6d ago
AI will cure all the people and kill them afterwards.
8.
▲
by
pllbnk
6d ago
Let's just say it's better not to risk it if there's anything you might not want them to see because they see everything. There are local models which are very capable and can be run on cloud if running on own hardware is not
9.
▲
by
pllbnk
6d ago
Companies are driven by people oriented at the next quarter's goals to maximize their stock portfolio's value. Nobody cares what will happen, everybody just creates narratives.
10.
▲
by
pllbnk
6d ago
It's funny how out of touch they are. To be fair, he also said: > The reason, he said, was that the best way to achieve reach in a chronological feed was to post more often, and businesses could afford to post more often than ordina
11.
▲
by
pllbnk
6d ago
I think the main limit of their model is that it was prompted and with LLMs you get what you prompt for. It's as truthful as any complex models. Could be good, could be bad, could be meh.
12.
▲
by
pllbnk
7d ago
I don't remember where I saw it but there was a woman in one conference who very eloquently put it that since frontier AI companies took humanity's work to train their models [without explicit permission of every single person who
13.
▲
by
pllbnk
8d ago
The guy on the left in that same photo only has three fingers (and a thumb, I suppose). I thought image generation has already outlived that. Edit: I feel stupid I didn't see the original OP already mentioning three finger issue. I
14.
▲
by
pllbnk
8d ago
Do all Silicon Valley corporations use the same jingle for their product promotional videos? The creator must be really rich by now.
15.
▲
by
pllbnk
10d ago
> Can you do murder? Only if it kills many people over the long time.
16.
▲
by
pllbnk
10d ago
Long time ago I had one particular hardware issue, which Opus 4.6 found a workaround to fix. I don't remember the workaround and being careless (I thought I could ask an LLM again if I needed) I lost that solution. Some time passed and
17.
▲
by
pllbnk
12d ago
Shouldn't this new reduction in thinking output make the model cheaper to operate? Could it be that the model is the same old LLM, with a bit newer architecture but still doing the same things, including huge amounts of thinking, just
18.
▲
by
pllbnk
12d ago
You still need the models to be able to perform web searches, don't you? In which case the data goes in and out of your machine and there is risk for prompt injection attacks. I think it's needed at least for documentation purpose
19.
▲
by
pllbnk
12d ago
The repository has many forks, suggesting that folks are trying to (vibe) code support for different GPUs. Might be worth a shot.
20.
▲
by
pllbnk
12d ago
Why would they sound the alarm if they were not trained (reinforced) to do that? I hope we don't expect sudden emersion of moral values from statistical models.
21.
▲
by
pllbnk
13d ago
Even without ninfer I would get over 80 on LM studio with default settings, so it should be noticeably more on 6000. You might want to try different a different inference engine or settings.
22.
▲
by
pllbnk
13d ago
Just a couple days ago I learned about ninfer ( https://github.com/Neroued/ninfer ) and on RTX 5090 I can now get ~200 tok/s and over 400 tok/s on concurrent requests which is plenty fast for a local model of t
23.
▲
by
pllbnk
13d ago
It's not that LLMs (I think that's what we are talking about when talking about AI) are not that useful. They are. But the most straightforward way to use them is to generate walls of text, which contain a lot of BS. In order to c
24.
▲
by
pllbnk
14d ago
If everybody is rich, then no one is rich. In other words, if everyone has a lot of money, then the prices will be high enough to suck this money out of everyone. America, and now much of the Western world, runs on debt and taking on debt i
25.
▲
by
pllbnk
16d ago
Last update from Anthropic was that they only allow third party harnesses via extra usage credits or API key: https://www.reddit.com/r/ClaudeAI/s/cNb55vmWqW
26.
▲
by
pllbnk
16d ago
AI industry does everything on vibes :)
27.
▲
by
pllbnk
16d ago
I think that technologies will continue to commoditize and arrive at the global mean (median?), so the need for accountants to have the vocabulary just disappear. They will say what they want and it will be coded, deployed, with CDNs and fi
28.
▲
by
pllbnk
16d ago
Well, to be fair, you cut the 6 important words that show the context: > We’re working on exciting changes that will make it feel like you’re getting more from Claude On the other hand, I think the biggest improvement they could do is al
29.
▲
by
pllbnk
18d ago
Just tried it, really cool and works as advertised.
30.
▲
by
pllbnk
19d ago
I run similar workflows as the author on my 5090. It's really good and reasonably fast at ~90 tok/sec on LM Studio. I haven't tried ninfer yet. The only problem is having to be mindful about the context size. I am jealous of
More ›