Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
tbalsam
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
tbalsam
27d ago
There was a story once about a boy with a wheelchair who needed a ramp to get into school, and the school made him use the loading dock ramp used for garbage and other things at the back. The school argued that it was an appropriate accommo
2.
▲
by
tbalsam
11mo ago
This is the common belief but not quite correct! The Muon update was proposed by Bernstein as the result of a theoretical paper suggesting concrete realizations of the theory, and Keller implemented it and added practical things to get it t
3.
▲
by
tbalsam
1y ago
shocked quack
4.
▲
by
tbalsam
1y ago
I'm not entirely sure, to be honest. If you look at the linked video, they state that it's oftentimes not in the best interest of the private equity group's moneymaking capabilities to announce that a channel has been sold ou
5.
▲
by
tbalsam
1y ago
They unfortunately recently (last few years) sold out to private equity (which tends to glaze over fundamentals and tries to pump out massive content using previous brand quality to give it credence), so beware of quality in more recent vid
6.
▲
by
tbalsam
1y ago
There are versions of this kind of benchmark with a higher threshold, however, it only seems to adjust the timetables by a linear amount, so you're only buying 1-2 years or so depending on what you want that % success rate to be.
7.
▲
by
tbalsam
1y ago
For those curious: https://en.m.wikipedia.org/wiki/Zombo.com
8.
▲
by
tbalsam
1y ago
The only limit is yourself Source: One of the most classic internet websites, zombo.com (sound on)
9.
▲
by
tbalsam
1y ago
No! This is not good. Iteration speed trumps all in research, most of what Python does is launch GPU operations, if you're having slowdowns from Pythonland then you're doing something terribly wrong. Python is an excellent (and ye
10.
▲
by
tbalsam
1y ago
This is (and was) the dream of Cerebras and I am very glad to see it embraced if even in small part on a GPU. Wild to see how much performance is left on the table for these things, it's crazy to think how much can be done by a few bol
11.
▲
by
tbalsam
1y ago
Not bad frustrations at all. That said -- IoU is how the final box scores are calculated, that doesn't change how you do feature aggregation, this will happen in basically any technique you use. Modern SSD/YOLO-style detectors use
12.
▲
by
tbalsam
1y ago
> The MSE here is not intended to be a training loss, but as a means to demonstrate that both approaches lead to almost the same result except for some rounding error. Ah, gotcha > I don't think that max pooling the last feature
13.
▲
by
tbalsam
1y ago
As someone who's done a fair bit of architecture work -- both are important! Making it either or is a very silly thing, both are the limiting factor for the other and there are no two ways about it. Also, for classification, MaxPooling
14.
▲
by
tbalsam
1y ago
As someone who has worked in computer vision ML for nearly a decade, this sounds like a terrible idea. You don't need RL remotely for this usecase. Image resolution pyramids are pretty normal tho and handling them well/efficiently
15.
▲
by
tbalsam
1y ago
McCormick is a popular brand of seasonings hahaha https://i5.walmartimages.com/seo/McCormick-Pure-Ground-Black...
16.
▲
by
tbalsam
1y ago
Yes, it's a synthesizer -- you may know it inside and out, but having demo videos showing what it can do will help people with no context get that quick "ahhh, that makes sense" moment from things. :)
17.
▲
by
tbalsam
1y ago
If you would like an original link non-heisted through the gwern domain, I'd encourage you to read it from the original UPenn link (University the professor who wrote this works at): https://web.english.upenn.edu/~cavit
18.
▲
by
tbalsam
1y ago
I did my part and manually reloaded the page about once a second for 5 minutes so that Andrew could get their dev validation beep quota in for the day (unless it's not naive hits, and unique user based, in which case this has a been a
19.
▲
by
tbalsam
1y ago
I don't think they really did? A single scratch to cause it to break down doesn't seem like it would really be a scalable solution for any kind of mass produced material like this. Would cause chaos if any individual container wen
20.
▲
by
tbalsam
2y ago
I'm a speedrunner, and I'm pretty sure this is well known -- and accepted as standard in some categories! It's a pretty well accepted standard (to the point of the headline being almost a mild offense!). In the gaming world,
21.
▲
by
tbalsam
2y ago
Yes, this is the principle that violates the WFC algorithm and makes it no longer WFC It is now just a procedural algorithm, which is faster than but loses some of the magic of what makes WFC _so good_. You can tell by looking at the render
22.
▲
by
tbalsam
2y ago
This is neat! However, it is now longer the WFC algorithm. The idea of the WFC algorithm is that it is limited to one-at-a-time iteration, this mathematically means that every element in the input communicates its full "state" to
23.
▲
Visualizing 6D Mesh Parallelism
(main-horse.github.io)
1 points
by
tbalsam
2y ago
|
0 comments
24.
▲
by
tbalsam
2y ago
>> The classification of several microstates into the same macrostate, is this not a distinctly observer-centred function? > It seems that way if we consider only our neat models, but it fails to explain why experimental measuremen
25.
▲
Muon: An optimizer for hidden layers in neural networks
(kellerjordan.github.io)
4 points
by
tbalsam
2y ago
|
0 comments
26.
▲
by
tbalsam
2y ago
That is an understandable statement, and probably fair as well I feel. Much of this comes in reference to statements from fchollet w.r.t. replacing deep learning -- around the time of the initial prize, with a lot of the much more hype mark
27.
▲
by
tbalsam
2y ago
> Now that the AI research field is coming around to the idea that something beyond deep learning is needed, I have not heard this from anyone that I work with! It would be a curious violation of info theory were this to be the case. Cer
28.
▲
by
tbalsam
2y ago
I feel rather consternated that this response effectively boils down to "yes, we know we overhyped this to get people's attention, and now that we have it we can be more honest about it". Fighting for place in the attention e
29.
▲
by
tbalsam
2y ago
As a rather experienced ML researcher, ARC is a great benchmark on its own, but is punching below its weight in terms of claiming that it is a gate (or in terms of this post -- a "steward") towards AGI, and in my perspective and t
30.
▲
by
tbalsam
2y ago
I have chronic fatigue issues and after trying thousands of dollars' worth of supplements over the years what worked best for me were blood glutamate scavengers (N-acetyl cysteine for me, I take an unnaturally high dose of up to 8g a d
More ›