Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
fred123123
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
fred123123
9d ago
That sound cool!, do you have hints on the cash prizes the website said something like 2 M ???
2.
▲
by
fred123123
9d ago
Hey! Any indication on what area of mathematics theses questions are from?
3.
▲
ArXiv
5 points
by
fred123123
1mo ago
|
1 comments
4.
▲
by
fred123123
2mo ago
Classic God in a Box type argument, yes you can then argue anything
5.
▲
by
fred123123
2mo ago
ok but why no inference time RL? also i think that it is reasoning exactly like us humans do. Indistinguishable
6.
▲
by
fred123123
2mo ago
Yes maybe you could scheme it, like think of 10 things you could try, spawn subagents, try them, perhaps one would succseed and yes that think that worked could be placed in a memory system. But there is also value in things that dont work
7.
▲
by
fred123123
2mo ago
The reason i am asking is about research mathematics. In this application what is really important is continual intuition and theory building, this takes time and trial and error. However this is very fruitful in solving problem. Currently
8.
▲
by
fred123123
2mo ago
I mean it gains the same level of expertise as it does on stuff in training data.
9.
▲
by
fred123123
2mo ago
Yes i agree, but somehow knowledge got put into the LLMs head during train time, but why does it not work during inference time? Too little examples? Do llms know stuff with exactly one occurence in the training data?
10.
▲
What is the status on continual learning for LLMs?
5 points
by
fred123123
2mo ago
|
14 comments
11.
▲
by
fred123123
3mo ago
How do you handle things like scrolling quickly in a video?