Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sls56
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
A new deep learning blog
(samtalksml.net)
2 points
by
sls56
9y ago
|
0 comments
2.
▲
by
sls56
9y ago
yes tempay precisely, because we use English as our reference, there is bias in favour of languages similar to English. Additionally the bigger a language's wikipedia, the higher quality the word vectors, which also tends to improve pe
3.
▲
by
sls56
9y ago
Cool rspeer, we will check NumberBatch out! Yes, as you say, our goal here is to show you can get something for nothing from pre-trained embeddings (I can learn the 78 matrices on my macbook in about ten minutes...). The alignment procedure
4.
▲
by
sls56
9y ago
People have actually tried to do things like this; though I suspect a linear transformation wouldn't work well. Chris Olah talks about it in his awesome blog: https://colah.github.io/posts/2014-07-NLP-RNNs-Represen
5.
▲
by
sls56
9y ago
Yes definitely! We didn't want to complicate the repository, but from a few in-house experiments we already know that it is possible to improve the rotation matrices by: 1. First aligning to a reference language (English) 2. Then defin
6.
▲
by
sls56
9y ago
We haven't done this in a rigorous way yet no, but it would be really interesting! In particular, I wonder if it's possible to reconstruct the linguistic tree of languages evolving over time; without any prior knowledge, solely fr