Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
fzimmermann89
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
88 ms
·
1.
▲
Kernel Fusion in Nvidia CUDA: Optimizing Memory Traffic and Launch Overhead
(developer.nvidia.com)
2 points
by
fzimmermann89
1mo ago
|
0 comments
2.
▲
by
fzimmermann89
1y ago
The switch by Artificial Analysis from per-token-cost to per-benchmark-cost shows some effect! Its nice that labs are now trying to optimize what I actually have to pay to get an answer - It always annoys me to have to pay for all the sens
3.
▲
by
fzimmermann89
1y ago
..and for complex valued tensors, you need to conjugate.
4.
▲
by
fzimmermann89
1y ago
Also, for an auto complete I think a small llm trained from scratch should already work well. Have you tried on if the tinystories(also only 3gb..)/nanogpt speed runs without any fancy loss terms etc as a baseline?
5.
▲
by
fzimmermann89
1y ago
How foreign is the language - was it likely included in pre training to some degree? Does it use grammar, syllables, and logic similiar to one of the large languages? Your approach assumes there is an easy to learn mapping between context i
6.
▲
by
fzimmermann89
1y ago
Contacting support obviously interfered with Apple services. Duh.
7.
▲
Ukraine says Russia drone hits Chernobyl nuclear plant, radiation levels normal
(cnn.com)
26 points
by
fzimmermann89
2y ago
|
14 comments
8.
▲
by
fzimmermann89
2y ago
If I am not mistaken, this is done by modulation in Fourier space. We have already been using this in optical setups for ages - at the speed of light. The interesting part imo is the implementation of this idea in their work and the efficie
9.
▲
by
fzimmermann89
2y ago
Sad to hear that they removed most of their content as well.
10.
▲
by
fzimmermann89
2y ago
If only they would fix the memory leak and freeze on resume from hibernate that has been an issue for the last year at least...
11.
▲
by
fzimmermann89
3y ago
I thought the same, until I noticed a really annoying WSL2 bug: On two machines I own, waking up from hibernate or standby causes a wsl related process (vmmem) to consume 100% CPU, with wsl becoming completely unresponsive (including wsl te
12.
▲
by
fzimmermann89
3y ago
*shoes.
13.
▲
by
fzimmermann89
4y ago
I got access by providing an academic email adress without mentioning any relevant publications etc.. Took maybe 2-3 days..
14.
▲
by
fzimmermann89
4y ago
There is a typo in the in app purchase page: "unlimmited" I clicked on one story, did not like it, so i canceled it. I wanted to give the app a second chance. But alas, only one story is allowed per day.. Maybe increase this at le
15.
▲
by
fzimmermann89
5y ago
So the code looks (apart from pointers) similiar to numba which feels much closer to numpy/pytorch high level code. Are there huge advantages in the triton model compared to numba that I don't see? Or is there a big performance ga
16.
▲
by
fzimmermann89
5y ago
or the text went through latex at some point, which can easily lead to a"->ä. letter+quotation mark is a common way to write umlauts...