Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
nil-sec
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
nil-sec
3y ago
Sorry maybe I should have added more explanation. One way to think about attention, which is the main distinguishing element in a transformer, is as an adaptable matrix. A feedforward layer is a matrix with static entries that do not change
2.
▲
by
nil-sec
3y ago
I guess I was more thinking about self attention, so yes. The more general case is covered by your notation!
3.
▲
by
nil-sec
3y ago
Feedforward: y=Wx Attention: y=W(x)x W is Matrix, x & y Are vectors. In the second case, W is a function of the input.
4.
▲
by
nil-sec
4y ago
I had the same experience a couple of years back when people started discussing AI. I try to keep this in mind but somehow I keep forgetting. It’s genuinely difficult to filter good from bad takes without expert level understanding of a top
5.
▲
by
nil-sec
4y ago
It’s quite nice to have something irreversible though. It gives you time. Also nobody really thinks QM is the end of it, assuming semi classical physics under the hood is just odd to me. There is something below QM that we don’t understand
6.
▲
by
nil-sec
4y ago
It’s funny because, for me, this was one of the major confusions when I moved from Europe to the US. In Europe, there are a lot of prominent subcultures, particularly in college. This was totally absent in the US in my experience. Even at p
7.
▲
by
nil-sec
4y ago
This isn’t true, the quality of images generated by DALL-E are really good, but they are an incremental improvement and based on a long chain of prior work. See e.g. https://github.com/CompVis/latent-diffusion
8.
▲
by
nil-sec
5y ago
Insane take, the west vs Russia in open confrontation ends in a nuclear winter.
9.
▲
by
nil-sec
5y ago
I have no background in control theory but this sounds very similar to the identifiability problem in nonlinear ICA. Are those equivalent?
10.
▲
by
nil-sec
5y ago
This is a good point which I think stems from wrongly equating human level intelligence to AGI in the popular literature. It’s not at all clear what a general intelligence should be and it’s much less clear that humans have general intellig
11.
▲
by
nil-sec
5y ago
While I agree with the general point of this paper I don’t think it’s quite right to compare the current situation with the last AI spring. It’s not AGI but it’s very good narrow AI that has real commercial value right now. The systems back
12.
▲
by
nil-sec
6y ago
For a given causal model it is (1) in my understanding.
13.
▲
by
nil-sec
6y ago
For one it lets you avoid controlling for the wrong variables and causing e.g. spurious correlations by doing so. In fact this is one of the best examples of why a causal model is necessary, because without one you can easily end up with a
14.
▲
by
nil-sec
6y ago
It is incomprehensible to you, because you just simply do not understand what your parent is talking about. You are the ignorant one here and indeed quite rude. Doesn't matter that genetics is not natural language. The point is we can
15.
▲
by
nil-sec
6y ago
Agreed, I trained a 3D version of b0-b2 on a classification task I worked on and besides being very slow to train they did not outperform a simple baseline VGG architecture. Interestingly training time was much improved by setting the cudnn
16.
▲
by
nil-sec
6y ago
Location: Zürich Remote: Yes Willing to relocate: Yes, within Europe. Technologies: Pytorch, Tensorflow, Python, C++, JS, HTML/CSS, Gurobi, MongoDB CV: https://nilsec.github.io Email: nils [dot] eckstein [at] googlemail [do
17.
▲
by
nil-sec
6y ago
They could have used all negative samples for testing (and even training if they would have done it better), yes. But once your test set is large enough, whatever that means, its not that relevant anymore. They are anyway "under sampli
18.
▲
by
nil-sec
6y ago
1. Isn't an issue. They make inference on a sample by sample basis. The network has no memory so it won't expect a 50/50 distribution on the test set just because its trained like that. Having a balanced distribution is the e
19.
▲
by
nil-sec
6y ago
Would love to play around with this data but there are no seeds for this torrent. Anyone here who can provide this dataset?
20.
▲
by
nil-sec
6y ago
Let me be very concrete: There is currently a lot of research being done in so called “contrastive learning”. This is an unsupervised technique in which you train a network to build good representations of its input data. Not by explicitly
21.
▲
by
nil-sec
6y ago
I happen to work in AI research and what you are saying isn’t true. There is theoretical machine learning and applications of it. They are distinct. The former is largely task agnostic and deals with fundamental issues such as, how to train
22.
▲
by
nil-sec
6y ago
In industry, yes. I had the impression this article was aimed at academia and research in AI.
23.
▲
by
nil-sec
6y ago
The fraction of people in AI working on problems that need to consider diversity/fairness/etc. is rather small. Yes, those people, these specific applications, should be designed with care, should be overseen properly and in the b
24.
▲
by
nil-sec
6y ago
This is a very one-sided view of what this article is talking about. That there is systemic racism in the US is not up for debate. It's a fact. Even if your back of the envelope calculation there were correct, you have to ask yourself
25.
▲
America's identity politics went from inclusion to division (2018)
(theguardian.com)
12 points
by
nil-sec
6y ago
|
4 comments
26.
▲
by
nil-sec
6y ago
It’s an interesting description of the space of computation and I quite enjoyed this article even though, as many, I have issues with the way wolfram participates in science. Certainly what he is describing can be a way to view the universe
27.
▲
by
nil-sec
6y ago
Turing completeness isn’t necessarily an interesting thing to have in common. Many (very simple) models of computation are Turing complete but have vastly different properties. Take for example a cellular automata, a Turing machine, Wang ti
28.
▲
by
nil-sec
6y ago
Why exactly shouldn’t scientists be politicians? It’s not like most of the people who are politicians have a formal education in political science. In an ideal world the ministers should come from exactly the background they are tasked with
29.
▲
by
nil-sec
6y ago
This is the point. There are many people at these protests. This is a heated, complex situation and at any large protest the people there are not homogeneous. The majority wants to peacefully protest, some loot, some want violence, the poli
30.
▲
by
nil-sec
6y ago
It is utterly ridiculous to try to blame either, far right or far left wing organizations for what is happening in the US right now. None of these groups have the power to cause that much damage. I’d be surprised if the number of antifa
More ›