Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
alcinos
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
Meta Segment Anything Model 3
(ai.meta.com)
178 points
by
alcinos
10mo ago
|
48 comments
2.
▲
by
alcinos
1y ago
> We've just only started RL training LLMs That's just factually wrong. Even the original chatGPT model (based on gpt3.5, released in 2022) was trained with RL (specifically RLHF).
3.
▲
by
alcinos
7y ago
It is already possible to know if a particular image has been used in training (see eg. https://arxiv.org/abs/1809.06396 by the same authors), but this new work also provides a p-value to give you a confidence on the r
4.
▲
Inside Kdenlive: How to fuzz a complex GUI application?
(kdenlive.org)
4 points
by
alcinos
8y ago
|
1 comments
5.
▲
by
alcinos
8y ago
Well it's deterministic once you know the random seed, which is stored in the replay file. An agent doesn't know the seed, hence cannot predict the exact outcome of its actions, only a probability over the outcomes. So, from the a
6.
▲
by
alcinos
11y ago
I completely agree that serious researchers should review definitions and results they are basing their findings upon. But then, I'd find much more reliable an editable reference paper corrected and improved by a significant number of
7.
▲
by
alcinos
11y ago
To me, the main problem with papers in their current shape is that they are required to be more or less self-contained. When one wants to state a result that improves a little bit the knowledge in a well established field, he has to waste t