Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
_ache_
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
_ache_
4d ago
Ok, so OpenAI is going to compensate RubyGem for the cyberattack on its servers?
2.
▲
by
_ache_
14d ago
https://ache.one/gpt6_now_down.png Big claims, expensive and not release to the public yet.
3.
▲
by
_ache_
14d ago
It's up then down again. https://openai.com/index/gpt-6-astra/ What a bunch of amateurs. Here is it anyway : https://ache.one/gpt6_now_down.png The claims: https://share-md.com
4.
▲
by
_ache_
16d ago
In computer science, that is technically a language. A formal language if you want to look it up on Wikipedia.
5.
▲
by
_ache_
19d ago
I will rephrase it. Will it run on any consumer hardware?
6.
▲
by
_ache_
19d ago
Reduce limits or usage? Twitter is blocked. Can you do more or less?
7.
▲
by
_ache_
21d ago
Is it linked to the last month COLT problem?
8.
▲
Qwen3.8-Flash-Next: A New Architecture, Towards Ultimate Cost-Efficiency
(qwen.ai)
12 points
by
_ache_
22d ago
|
0 comments
9.
▲
by
_ache_
23d ago
I'm hearing Tencent, Zhipu and Baidu shaking from here. It's fair to assume BATX / 6 Tigers don't sleep very well either.
10.
▲
by
_ache_
23d ago
I think a reasonable expectation of MAX requirement to claim "runable on consumer hardware" is to 32G VRAM and 128GB RAM and it run at +10tps.
11.
▲
by
_ache_
23d ago
What will be the requirement, like 128G of RAM and 12G of VRAM ?
12.
▲
Show HN: A visual ping utility that is pretty
(source.tube)
4 points
by
_ache_
1mo ago
|
0 comments
13.
▲
by
_ache_
1mo ago
They need to train a new model every month to keep at the top of most benchmarks. They don't own any DC, the price is insane. Most of people are aiming at smaller models because Claude one's are too expansive. Evolution of intelli
14.
▲
by
_ache_
1mo ago
From your benchmark, Qwen3.8 is nearer than Opus 4.8 than Qwen3.6. 0.1pp but still. Also, a lot of people don't really care about german language capacity, maybe people programming in DDP idk. PS: You benchmark seems saturated. Most va
15.
▲
by
_ache_
1mo ago
I actually expect them to explain me how they will manage to not go bankrupt soon.
16.
▲
by
_ache_
1mo ago
It's crazy how Anthropic talks so much about their "AGI risk" and not enough about the risk of bankruptcy.
17.
▲
by
_ache_
1mo ago
It is already. You can buy it online. There is not a lot of places where hackers sell that kind of stuffs. 1k lines are already shared, seems legit. Most of them is just <30k€ people. Only a handful of millionaires (8 >10M if I rememb
18.
▲
by
_ache_
1mo ago
No yet finished! Still waiting for tonight Qwen3.8-27B and the unsloth Q5_K_M/S quantification. Hopping for an AgentWorld variant from Qwen but I guess, I have too high expectations.
19.
▲
by
_ache_
1mo ago
What can you do with "only" 64G of VRAM that a 32G can't? Also, the R9700 are so loud!
20.
▲
Ask HN: What's the story Behind BSD-3-Clause-No-Nuclear-Warranty
1 points
by
_ache_
1mo ago
|
1 comments
21.
▲
by
_ache_
1mo ago
I planed to do exactly this, with podman instead of docker, volume support. Like: $ podman run -it --rm -v .:/workspace local-dev-ia /usr/bin/oc Configured with a .env file. Hope to do it hopefully before the end of th
22.
▲
by
_ache_
1mo ago
It is interesting but it does look like a careful distillation of (Spark and) biggers open-weight models. The progress compared to Qwen3.6 27B is good, not that impressive, it's a 4 months old model. (kuto to them to compare to 27B den
23.
▲
by
_ache_
2mo ago
I think, it's the end of Anthropic. They can't compete, they have bills. By the end of the year, if they can't react, it's game over. Maybe US clients could be a little patriotic here, but money is money. They won'
24.
▲
by
_ache_
2mo ago
Note that HAWK-256 is also a simpler version of the proposed HAWK standard. But yeah, à priori it may apply to HAWK-512 too, so looks like a big found. https://hawk-sign.info/ PS: Never heard of LEA, looks like a Korean c
25.
▲
by
_ache_
2mo ago
And them, ... actually not an error. Server costs went up with AI needs, you know. The AWS incident is about inaccuracy, not "absurd values". You gotta catch-up or die.
26.
▲
by
_ache_
3mo ago
I used to do that. Them move to a NodeJS script. Because bash is maybe worst than C for this task.
27.
▲
by
_ache_
3mo ago
Yeah, exactly. My girlfriend did it too, using information theory. In my solution, I did chose a word that minimized the size of the larger set of possible guesses. It's not strictly information theory, but it's a good approximati
28.
▲
by
_ache_
3mo ago
And new Github profile too. https://github.com/TensorOne
29.
▲
by
_ache_
3mo ago
Note that Qwen from Alibaba choose to align the model with the PCC. It's not a same as DeepSeek who ensure it at the "service" level.
30.
▲
by
_ache_
3mo ago
Yet.
More ›