Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
bigdict
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
1.
▲
by
bigdict
10d ago
I would! If you are thinking to write one, do it!
2.
▲
by
bigdict
4mo ago
But... they are equivalent?
3.
▲
by
bigdict
8mo ago
Was it a controlled study or just correlation?
4.
▲
by
bigdict
8mo ago
Huh? Triton inference server and Triton the kernel language are two distinct, very different things… Is this AI-generated?
5.
▲
by
bigdict
8mo ago
> What actually happens: random Internet users spend two seconds skimming, then click their favorite. > They're not reading carefully. They're not fact-checking, or even trying. Uhhh, how was that established?
6.
▲
Artificial General Cleverness
(mathstodon.xyz)
4 points
by
bigdict
9mo ago
|
0 comments
7.
▲
by
bigdict
10mo ago
cuz area and power
8.
▲
by
bigdict
11mo ago
There's antimony, arsenic, aluminum, selenium…
9.
▲
by
bigdict
1y ago
What's the point of the relu in the loss function? Its inputs are nonnegative anyway.
10.
▲
by
bigdict
1y ago
Exactly, it's a natural way to write a college essay . I've never not cringed reading an article/blog post that is structured that way, it comes across very contrived. I've also noticed that LLMs tend to prefer it, and
11.
▲
by
bigdict
1y ago
"delves" dashes an explicit "conclusion" section at the end
12.
▲
by
bigdict
1y ago
Lisp In Small Pieces
13.
▲
by
bigdict
1y ago
Pairs great with the tale of the Death Valley Germans: https://www.otherhand.org/home-page/search-and-rescue/the-hu... .
14.
▲
by
bigdict
1y ago
Once you realize that elisp is a better shell programming language than bash, it ceases to appear insane.
15.
▲
by
bigdict
1y ago
Has this been used widely since?
16.
▲
by
bigdict
1y ago
Thank you so much for continuing to support Gemma 3 with these updates.
17.
▲
by
bigdict
1y ago
Amazing, I've been wishing for this! Do you have any estimates on how much accuracy is first lost then recovered compared to the original bf16 and the naively quantized models?
18.
▲
by
bigdict
1y ago
Sure, you can get better model performance by throwing more compute at the problem in different places. Does is it improve perf on an isoflop basis?
19.
▲
by
bigdict
1y ago
Gemma 3 is.
20.
▲
by
bigdict
2y ago
* the code is at https://github.com/google-deepmind/gemma * you download the weights at https://www.kaggle.com/models/google/gemma-3/
21.
▲
by
bigdict
2y ago
"undo their disgusting propaganda, apply our beautiful correct opinions" This is so cringe.
22.
▲
by
bigdict
2y ago
Why is it timely?
23.
▲
by
bigdict
2y ago
Have you seen the RFK confirmation hearing?
24.
▲
by
bigdict
2y ago
You are quoting Wikipedia, not US foreign policy :)
25.
▲
by
bigdict
2y ago
Does it have an entry for what we don't know we don't know?
26.
▲
by
bigdict
2y ago
> What gave you the impression I was assuming bad faith? You said "I would guess you're not asking a serious question here"
27.
▲
by
bigdict
2y ago
Why are you assuming bad faith?
28.
▲
by
bigdict
2y ago
I thought lexical scoping is standard now?
29.
▲
2024 Lebanon Pager Explosions
(en.wikipedia.org)
14 points
by
bigdict
2y ago
|
2 comments
30.
▲
by
bigdict
2y ago
nah, this is about codebases that are not themselves primarily lisp implementations
More ›