Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
djsjajah
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
djsjajah
22d ago
funny, I am not surprised that it is now much harder to run an automated twitter account for legitimate purposes. Running one for illegitimate purposes on the other hand....
2.
▲
by
djsjajah
1mo ago
It’s been renamed recently. Maybe a few times. I think it’s rocq now. [1] [1] https://rocq-prover.org/docs
3.
▲
by
djsjajah
1mo ago
You have to power it all the time, but the amount of power it uses while it’s on will change by up to a few orders of magnitude depending on the gpu. It’s not uncommon for a gpu to be pulling just a couple of watts at idles and several hund
4.
▲
by
djsjajah
2mo ago
but what is the point? A ban is supposed to make a certain thing less likely to occur. Does a ban of open source models do that? Presumably, the behavior you are trying to limit is the miss-use of these models but I don't know how many
5.
▲
by
djsjajah
2mo ago
I think they were making a joke. In the future, you might consider the advice you are giving as well as giving it.
6.
▲
by
djsjajah
3mo ago
Yes. Wait a day
7.
▲
by
djsjajah
3mo ago
Not with 800 examples. If you are going to consider an ngram model, I think you are better off getting a frontier llm to write you an absurd regex.
8.
▲
by
djsjajah
3mo ago
Except, that won’t help. By the time a new fab is up and running, we will probably have a massive surplus.
9.
▲
by
djsjajah
3mo ago
You need to think this thought through all the way to the end. What it has said also influences what it will say. If it has consistently made combative responses, then the most likely thing to do is to continue to be combative. I don't
10.
▲
by
djsjajah
3mo ago
It’s amusing that a lot of the agents have worked out that sampling doesn’t change ppl.
11.
▲
by
djsjajah
3mo ago
I think what they mean by “now” is the stuff announced today.
12.
▲
by
djsjajah
5mo ago
I don't follow. Can you explain how your comment is relevant to mine? It might help if you also explain how you interpreted my comment.
13.
▲
by
djsjajah
5mo ago
You just failed the Turing test.
14.
▲
by
djsjajah
5mo ago
I have 2 of them. I would advise against if you want to run things like vllm. I have had the cards for months and I still have not been able to create a uv env with trl and vllm. For vllm, it’s works fine in docker for some models. With one
15.
▲
by
djsjajah
5mo ago
> or by the community Hmmm
16.
▲
by
djsjajah
6mo ago
yes, but the difference between one model and one 4x larger is usually a lot more than that. It is not a question of do a run Qwen 8b at bf16 or a quantized version. It more of a question of do I run Qwen 8b at full precision or do I run a
17.
▲
by
djsjajah
6mo ago
trl. give me a uv command to get that working. But even in the amd stack things (like ck and aiter) consumer cards are not even second class citizens. They are a distance third at best. If you just want to run vllm with the latest model, if
18.
▲
by
djsjajah
6mo ago
No. It seems to me that the comment is objectively incorrect. The original comment was talking about inference and from what I can tell, it is strictly going to run slower than the model trained to the same loss without this approach (it ha
19.
▲
by
djsjajah
7mo ago
That’s kind of a moot point. Even if none of those overheads existed you would still be getting a a fractions of the mfu. Models are fundamental limited by memory bandwidth even with best case scenarios of sft or prefill. And what are you d
20.
▲
by
djsjajah
8mo ago
> including all previous experiments How far back do you go? What about experiments into architecture features that didn’t make the cut? What about pre-transformer attention? But more generally, why are you so sure that they team that bu
21.
▲
by
djsjajah
8mo ago
Not only can it be streamed, but lz4 will probably make things quicker.
22.
▲
by
djsjajah
8mo ago
You just ruined my day. The post makes it sound like gel is now dead. The post by Vercel does not give me much hope either [1]. Last commit on the gel repo was two weeks ago. [1] https://vercel.com/blog/investing-in-the
23.
▲
by
djsjajah
9mo ago
> Do you really though? Yes. It stays in on the hbm but it need to get shuffled to the place where it can actually do the computation. It’s a lot like a normal cpu. The cpu can’t do anything with data in the system memory, it has to be l
24.
▲
by
djsjajah
9mo ago
GPUs might not be bandwidth starved most of the time, but they absolutely are when generating text from an llm. It’s the whole reason why low precision floating point numbers are being pushed by nvidia.
25.
▲
by
djsjajah
10mo ago
I can't tell if you are making a joke or not. They are not even remotely equivalent. tinygrad is a toy. If you are serious, I would be interested to hear how you see tinygrad replacing CUDA. I could see a tiny grad zealot arguing that
26.
▲
by
djsjajah
10mo ago
I went to check how many services are being impacted on down detector, but it was down.
27.
▲
by
djsjajah
1y ago
If we didn't have people like that, then they would be right.
28.
▲
by
djsjajah
1y ago
I few people have mentioned dagster and I took a look at that for some machine learning things I was playing with but I found dvc (data version control [1]) and I think it is fantastic. I think it also has more applications than just machin
29.
▲
by
djsjajah
2y ago
No. Kids would need to memorize the private key of their parents id card.
30.
▲
by
djsjajah
2y ago
You could have also self-hosted the GitHub Actions runner which might have been easier as long as you had something to run the runner on.
More ›