Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
lumpilumpi
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
lumpilumpi
7mo ago
My experience is that the first iteration output from a single agent is not what I want to be on the hook for. What squares it for me with "not writing code anymore" is the iterative process to improve outputs: 1) Having review lo
2.
▲
by
lumpilumpi
7mo ago
I get the justification but I found it hard to understand how the actual evaluation at each step is carried out. For example, is there any calibration to some human gold standard involved or is the AI evaluating the AI without calibration&#
3.
▲
Show HN: MCGrad – Fix ML Calibration in Subgroups (Open Source from Meta)
(github.com)
4 points
by
lumpilumpi
7mo ago
|
0 comments