Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
lnenad
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
lnenad
4d ago
What does saying "working tech" do for you? If I have a working Samsung CRT from 25 years ago do I ping them about smart TV support? Nowadays it's a shitty situation with planned obsolescence; but 6 years for an open source p
2.
▲
by
lnenad
6d ago
I'm well aware but it just reaffirms that it's a cluster fuck from a product perspective and makes no sense to create such a weird segmentation. You didn't even mention Jules lol, who knows where it fits in.
3.
▲
by
lnenad
6d ago
What is happening at Google? It seems like they've got multiple teams building the same thing and competing for love from the higherups. Depending on who's in the lead the chosen package gets into the spotlight. It's been Jul
4.
▲
XBOW's Native Team Has Claimed the First Chrome Full Chain Exploit Bonus
(twitter.com)
1 points
by
lnenad
14d ago
|
0 comments
5.
▲
by
lnenad
16d ago
The agent wrote the code that has a mechanism that triggers a file as a side effect. That file started the separate process, as it could have started any other binary.
6.
▲
by
lnenad
17d ago
But you're not actually hijacking the agent if you start a new process.
7.
▲
by
lnenad
17d ago
Yeah, I agree, this is a different vector. Still scary though and very related to AI.
8.
▲
by
lnenad
17d ago
But there is a cool blog post about it though.
9.
▲
by
lnenad
19d ago
48c 7643. I'm getting about 10tps @Q3kxl with 2x3090s.
10.
▲
by
lnenad
19d ago
I'm getting about 10tps @Q3kxl with 2x3090s.
11.
▲
by
lnenad
19d ago
What model are you interested in? DS Flash 0731@Q4KXL I'm about 25-30tps. Same as the new Qwen3.8 Flash Next. The new GLM 5.3Q3KXL at 10tps. I've got 2x3090s which I didn't mention in the original message.
12.
▲
by
lnenad
19d ago
About 5k with RAM and GPUs bought used. Eastern Europe.
13.
▲
by
lnenad
19d ago
I am getting 10t/s on unsloth's Q3kxl with 2x3090s@250w. It's enough for me for now. I will probably upgrade the GPUs down the line. DDR5 would have made the price of the machine double and I just wasn't prepared to pay
14.
▲
by
lnenad
19d ago
I have just built an Epyc with 512gb DDR4 3200 RAM for a "reasonable" price and I'm hoping to have a setup with GLM as the architect and Qwen 27b/Next Flash as the implementer. This is 1/5 of the price of the Mac, b
15.
▲
by
lnenad
20d ago
But there isn't? What does it factually mean to be conscious? How can we claim other living beings aren't are conscious?
16.
▲
by
lnenad
21d ago
I've got a 48c Epyc with 2x3090s and 512gb ddr4 3200. It's good enough for 25+ tps with deepseek so I'm hoping for similar performance with less overthinking.
17.
▲
by
lnenad
21d ago
Yeah I understand, it's my assumption that the actually/wait/but have a point. It doesn't reduce the fact that it increases the time for tasks substantially.
18.
▲
by
lnenad
21d ago
Especially on practical tasks. One shot prompts work better at Q6_K_XL for me. It loads a file, then analyses then second guesses itself then again then again then it tries to come up with a solution then second guess rinse and repeat. 122b
19.
▲
by
lnenad
21d ago
Yeah 122B is the sweet spot for me as well. Even deepseek flash overthinks on stuff way too much. I think they fully rely on large reasoning turns to achieve better quality. The result of course means we wait a long time to get results even
20.
▲
by
lnenad
22d ago
Adding to my homelab stack, hopefully it doesn't overthink like the little model. Actually, hoping it thinks a bit less. Wait actually I'm really praying it reasons a bit more directly. But wait, I'm really sure that it must
21.
▲
by
lnenad
23d ago
I think as with most human undertakings, building isn't too much of a problem. Maintaining is. Even with what is still a relatively tame number of chargers you get a large number of them that are broken.
22.
▲
by
lnenad
24d ago
I'm assuming it's definitely part of the equation, but considering that I'm getting more tps but still waiting a lot more time for code to come out I'd assume it's not a 1:1 comparison. Plus I'm running quants,
23.
▲
by
lnenad
24d ago
As a small background, I have a local server and I've been trying out different models with different inference engines, quants, configurations etc... I'm also using Opus and Sol at work consistently. I've used AI since the f
24.
▲
by
lnenad
25d ago
I LOVE the concept. I will play around with the execution, if it works as described this is a great product.
25.
▲
by
lnenad
27d ago
> If your needs are met otherwise stick to that and move on. Weird to post such a thought in a forum where OP has posted their project for people to look at. I never mentioned any needs, I am saying the readme holds very little value for
26.
▲
by
lnenad
27d ago
I think as many things that are posted here lately there is no *why* attached to the readme. Why would one use this, what is the benefit of this approach? Am I really gonna need my model to build exotic tools around it; or is exec/web_
27.
▲
by
lnenad
1mo ago
You've built a great piece of software, thank you!
28.
▲
by
lnenad
1mo ago
I'm running gitea successfully with very little resources.
29.
▲
by
lnenad
1mo ago
Deepseek is the enemy, the implication is being on hn you should know that and not work for them. /s
30.
▲
by
lnenad
1mo ago
They've probably got companies lined up to give them jobs that will pay multiples of those. The runner up is even more interesting, with a slightly larger reason to feel like they should have won it. They did ~3nm (out of 4 iirc) of 4.
More ›