Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
james-mxtech
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
10 ms
·
1.
▲
by
james-mxtech
2mo ago
Doubt this is really a language problem instead of a model-capability one. A better prompt might get you most of the way there.
2.
▲
by
james-mxtech
2mo ago
Reserving judgment until the actual feature is set public. The announcement post itself doesn't tell you much.
3.
▲
by
james-mxtech
3mo ago
the rename from checkpoint to system boundary is how you lose traceability. a multi-model black box that benchmarks well is great until it fails on your workload and theres no trace to debug why.
4.
▲
by
james-mxtech
3mo ago
speed is half of why i've gotten sloppier with agents lately. at 200 t/s i actually read each diff, at 1000 i just accept them and catch the breakage later.
5.
▲
by
james-mxtech
3mo ago
the eval agent you're running 8-10 wide is the part i'd want more detail on. silent chapter drops and dupes were the only extraction failures that ever bit me, and grading those automatically needs a ground-truth toc to diff again
6.
▲
by
james-mxtech
3mo ago
can't wait for the Xbox Mortgage Edition